Qwen3.8-Flash-Next vs jina-ocr-v1

Try Try

Qwen3.8-Flash-Next has 180B parameters and takes 360 GB on disk; jina-ocr-v1 has 3.37B (708M active) and takes 6.74 GB. Compare them layer by layer.

At a glance

Qwen/Qwen3.8-Flash-Nextjinaai/jina-ocr-v1
Repo createdAug 24, 2026updated Aug 27, 2026Sep 1, 2026updated Sep 18, 2026
Model typeNot shown: breakdown incompleteMixture of expertsFull attention (MHA)
InputsText + images449M vision encoderText + images303M vision encoder
Total parameters180B53× more3.37B53× less
Active per tokenNot shown: breakdown incomplete708M21% of the model
ExpertsNot shown: breakdown incomplete6 of 64 activeplus 2 shared, always on
Max context (from config)256K tokens8.0× more32K tokens8.0× less
Layers484.0× more36 linear + 12 full attention124.0× less
On disk360 GB53× more131 files6.74 GB53× less2 files
PrecisionBF16 (100%)BF16 (100%)
QuantizationNoneoriginal precision (BF16)Noneoriginal precision (BF16)
Fewest GPUs5× NVIDIA H100fits in one 8-GPU server · weights only1× RTX 4060a single 8 GB card · weights only
Licenseqwen-community-1.0 (custom)cc-by-nc-4.0
GitHubQwenLM/Qwen3.8-Flash-NextNot linked