GPT-OSS 120B

117B MoE

OpenAI GPT-OSS

The big GPT-OSS (117B total / 5.1B active). Native MXFP4 checkpoint is ~61GB — fits one 80GB card or a 128GB unified-memory Mac. Partial offload on 24GB consumer cards is slow but works.

4.1M HF downloads5087 likesopenai/gpt-oss-120b· stats from 8/8/2026
Pro GPUMac / Apple Silicon

131K

Max Context

2

Quant Variants

GGUF MXFP4

Best Quality

100.0%

Accuracy Retained

Quantization Variants

Per-quant VRAM, quality loss, and inference speed on RTX 4090

Measured = site benchmarks · Estimated = formula · Community = public reports

FormatLevelBPWVRAMPPL LossSpeedSourceActions
GGUFMXFP44.2561.0 GB0.0%22 tok/sCommunity
CalcHF
GGUFQ4_K_M4.158.5 GB1.6%24 tok/sEstimated
CalcHF