DeepSeek-Coder-V2-Lite-Instruct

Try Try

DeepSeek-Coder-V2-Lite-Instruct has 15.7B parameters (2.45B active per token) and takes 31.4 GB on disk. See where its parameters live, layer by layer, and what hardware it needs.

At a glance

deepseek-ai/DeepSeek-Coder-V2-Lite-Instruct
Repo createdJun 14, 2024updated Jul 3, 2024
Model typeMixture of expertsFull attention (MLA)
InputsText
Total parameters15.7B
Active per token2.45B16% of the model
Experts6 of 64 activeplus 2 shared, always on
Max context (from config)160K tokens
Layers27
On disk31.4 GB4 files
PrecisionBF16 (100%)
QuantizationNoneoriginal precision (BF16)
Fewest GPUs1× RTX 6000 Adaa single 48 GB card · weights only
Licensedeepseek-license (custom)
GitHubdeepseek-ai/DeepSeek-Coder-V2