Qwen3.8-Flash-Next-GSQ-RCO-GGUF

Try Try

Qwen3.8-Flash-Next-GSQ-RCO-GGUF has 51.8B parameters and takes 29.5 GB on disk. See where its parameters live, layer by layer, and what hardware it needs.

At a glance

pfeifferj/Qwen3.8-Flash-Next-GSQ-RCO-GGUF
Repo createdSep 13, 2026updated Sep 19, 2026
Model typeNot shown: breakdown incomplete
InputsText + imagesvision encoder in a separate mmproj file, not counted here
Total parameters51.8B
Active per tokenNot shown: breakdown incomplete
ExpertsNot shown: breakdown incomplete
Max context (from config)256K tokens
Layers48the files show 0, which doesn't match; layer details hidden
On disk29.5 GB1 file
PrecisionQ4_0 (98%), Q8_0 (2.3%)
QuantizationQ4_0 (GGUF)mix of Q4_0, Q8_0 · blocks of 32 · 100% of parameters
Fewest GPUs1× RTX 6000 Adaa single 48 GB card · weights only
Licenseqwen-community-1.0 (custom)
GitHubNot linked