Qwen1.5-MoE-A2.7B

Try Try

Qwen1.5-MoE-A2.7B has 14.3B parameters (2.38B active per token) and takes 28.6 GB on disk. See where its parameters live, layer by layer, and what hardware it needs.

At a glance

Qwen/Qwen1.5-MoE-A2.7B
Repo createdFeb 29, 2024updated Apr 18, 2024
Model typeMixture of expertsFull attention (MHA)
InputsText
Total parameters14.3B
Active per token2.38B17% of the model
Experts4 of 60 activeplus 1 shared, always on
Max context (from config)8K tokens
Layers24
On disk28.6 GB8 files
PrecisionBF16 (100%)
QuantizationNoneoriginal precision (BF16)
Fewest GPUs1× RTX 5090a single 32 GB card · weights only
Licensetongyi-qianwen (custom)
GitHubQwenLM/Qwen1.5