Xing4.0-29B-A4B-GGUF

Try Try

Xing4.0-29B-A4B-GGUF has 31.2B parameters (3.93B active per token) and takes 20.1 GB on disk. See where its parameters live, layer by layer, and what hardware it needs.

At a glance

XingChen-AGI/Xing4.0-29B-A4B-GGUF
Repo createdSep 16, 2026updated Sep 20, 2026
Model typeMixture of expertsFull attention (MLA)
InputsText
Total parameters31.2B
Active per token3.93B13% of the model
Experts4 of 64 activeplus 1 shared, always on
Max context (from config)256K tokens
Layers40plus 1 extra prediction layer
On disk20.1 GB1 file
PrecisionIQ4_NL (81%), BF16 (17%), Q6_K (1.9%), other (0.19%)
QuantizationIQ4_NL (GGUF)mix of IQ4_NL, Q6_K · blocks of 32 / blocks of 256 · 95% of parameters · attention kept in BF16
Fewest GPUs1× RTX 4090a single 24 GB card · weights only
Licenseapache-2.0
GitHubXingChen-AGI/Xing4.0-29B-A4B