Kimi-K2-Instruct

Try Try

Kimi-K2-Instruct has 1.03T parameters (31.7B active per token) and takes 1.03 TB on disk. See where its parameters live, layer by layer, and what hardware it needs.

At a glance

moonshotai/Kimi-K2-Instruct
Repo createdJul 11, 2025updated Apr 23, 2026
Model typeMixture of expertsFull attention (MLA)
InputsText
Total parameters1.03T
Active per token31.7B3.1% of the model
Experts8 of 384 activeplus 1 shared, always on
Max context (from config)128K tokens
Layers61
On disk1.03 TB61 files
PrecisionFP8 E4M3 (99%), other (0.51%)
QuantizationFP8 E4M3blocks of 128×128 · 100% of parameters
Fewest GPUs7× NVIDIA B200fits in one 8-GPU server · weights only
Licensemodified-mit (custom)
GitHubMoonshotAI/Kimi-K2