Model Quant Measured Estimated RAM used App Source Date
Phi-4 Mini Phi-4 Mini Q5 42 tok/s 15 tok/s +180% 2.9 GB Ollama editorial 2026-03-20
Mistral 7B Mistral 7B Q4 28 tok/s 8 tok/s +250% 4.3 GB LM Studio editorial 2026-03-20