gpt-oss-120b

Try Try

gpt-oss-120b has 117B parameters (5.13B active per token) and takes 65.2 GB on disk. See where its parameters live, layer by layer, and what hardware it needs.

At a glance

openai/gpt-oss-120b
Repo createdAug 4, 2025updated Aug 26, 2025
Model typeMixture of expertsFull attention (GQA)
InputsText
Total parameters117B
Active per token5.13B4.4% of the model
Experts4 of 128 active
Max context (from config)128K tokens
Layers36
On disk65.2 GB15 files
Precision4-bit MXFP4 (88%), BF16 (6.6%), scales (5.5%)
Quantization4-bit MXFP4blocks of 32 · 98% of parameters · embeddings & output head and attention kept in BF16
Fewest GPUs1× NVIDIA H100a single 80 GB card · weights only
Licenseapache-2.0
GitHubopenai/gpt-oss