Phi-4 14B

14B

Microsoft Phi

Microsoft Phi-4 dense 14B — strong reasoning for size. Q4 ~9GB fits 12GB cards with moderate context.

10.8K HF downloads62 likesbartowski/phi-4-GGUF· stats from 7/27/2026
Consumer GPUMac / Apple Silicon

16K

Max Context

3

Quant Variants

GGUF Q5_K_M

Best Quality

98.7%

Accuracy Retained

Quantization Variants

Per-quant VRAM, quality loss, and inference speed on RTX 4090

Measured = site benchmarks · Estimated = formula · Community = public reports

FormatLevelBPWVRAMPPL LossSpeedSourceActions
GGUFQ4_K_M4.859.1 GB2.7%88 tok/sCommunity
CalcHF
GGUFQ5_K_M5.6810.5 GB1.3%78 tok/sEstimated
CalcHF
AWQINT448.2 GB3.6%112 tok/sEstimated
CalcHF