Phi-4 Mini Instruct
3.8BMicrosoft Phi
Latest Phi mini with improved math and code. Strong 4B-class performer.
Consumer GPUMac / Apple SiliconCPU / VPS
131K
Max Context
3
Quant Variants
GGUF Q8_0
Best Quality
99.8%
Accuracy Retained
Quantization Variants
Per-quant VRAM, quality loss, and inference speed on RTX 4090
Measured = site benchmarks · Estimated = formula · Community = public reports
Similar models
Compare with Phi-3.5 Mini3.8B
Phi-3.5 Mini Instruct
Microsoft Phi
Consumer GPUMac / Apple Silicon
2.5 GBmin VRAM·99.8%accuracy
Microsoft's tiny powerhouse. Best 4B model for on-device deployment.
14B
Phi-3 Medium 14B Instruct
Microsoft Phi
Consumer GPUMac / Apple Silicon
8.8 GBmin VRAM·99.2%accuracy
Microsoft's mid-size Phi-3. Excellent quality-per-GB on 16GB cards.
14B
Phi-4 14B
Microsoft Phi
Consumer GPUMac / Apple Silicon
8.2 GBmin VRAM·98.7%accuracy
Microsoft Phi-4 dense 14B — strong reasoning for size. Q4 ~9GB fits 12GB cards with moderate context.
7B
Qwen2.5 7B Instruct
Alibaba Qwen2.5
Consumer GPUMac / Apple Silicon
4.8 GBmin VRAM·99.3%accuracy
Alibaba's highly optimized 7B. Punches well above its weight, especially in coding.