Qwen3.8-Flash-Next vs MiniCPM5-2B

Try Try

Qwen3.8-Flash-Next has 180B parameters and takes 360 GB on disk; MiniCPM5-2B has 2.52B (2.25B active) and takes 5.03 GB. Compare them layer by layer.

At a glance

Qwen/Qwen3.8-Flash-Nextopenbmb/MiniCPM5-2B
Repo createdAug 24, 2026updated Aug 27, 2026Sep 6, 2026updated Sep 12, 2026
Model typeNot shown: breakdown incompleteDenseFull attention (GQA)
InputsText + images449M vision encoderText
Total parameters180B72× more2.52B72× less
Active per tokenNot shown: breakdown incomplete2.25B89% of the model
ExpertsNot shown: breakdown incompleteNone (dense)
Max context (from config)256K tokens2.0× more128K tokens2.0× less
Layers481.1× more36 linear + 12 full attention421.1× less
On disk360 GB72× more131 files5.03 GB72× less1 file
PrecisionBF16 (100%)BF16 (100%)
QuantizationNoneoriginal precision (BF16)Noneoriginal precision (BF16)
Fewest GPUs5× NVIDIA H100fits in one 8-GPU server · weights only1× RTX 4060a single 8 GB card · weights only
Licenseqwen-community-1.0 (custom)apache-2.0
GitHubQwenLM/Qwen3.8-Flash-NextOpenBMB/MiniCPM