Model Quant Measured Estimated RAM used App Source Date
Mistral 7B Mistral 7B Q4 186 tok/s 54 tok/s +244% 4.3 GB LM Studio editorial 2026-03-20
Phi-4 Mini Phi-4 Mini Q8 174 tok/s 103 tok/s +69% 4.6 GB Ollama editorial 2026-03-20