Can I run Qwen3.6-35B-A3B-Uncensored-HauhauCS-Aggressive on an NVIDIA H100?
Yes
Its Q8_K file (43.6 GB) fits in the 72.0 GB an NVIDIA H100 can use, with room for ≈ 256K tokens of context.
Full analysis of Qwen3.6-35B-A3B-Uncensored-HauhauCS-Aggressive
Everything that runs on NVIDIA H100
Get the Q8_K file on Hugging Face
ollama run hf.co/HauhauCS/Qwen3.6-35B-A3B-Uncensored-HauhauCS-Aggressive:Q8_KFiles and context on NVIDIA H100
Each file of Qwen3.6-35B-A3B-Uncensored-HauhauCS-Aggressive (and of GGUF versions the site has analyzed), and how much context fits next to it, with the context cache in 16-bit or 8-bit.
| File | Size | On NVIDIA H100 | Context (16-bit) | Context (8-bit) |
|---|---|---|---|---|
| Q8_K | 43.6 GB | Fits | 256K | 256K |
| Q6_K | 30.6 GB | Fits | 256K | 256K |
| Q5_K | 28.0 GB | Fits | 256K | 256K |
| Q4_K | 23.4 GB | Fits | 256K | 256K |
| Q4_K_M | 21.2 GB | Fits | 256K | 256K |
| IQ4_NL | 19.8 GB | Fits | 256K | 256K |
| Q3_K | 19.0 GB | Fits | 256K | 256K |
| IQ4_XS | 18.7 GB | Fits | 256K | 256K |
| IQ3_M | 15.4 GB | Fits | 256K | 256K |
| Q2_K | 15.0 GB | Fits | 256K | 256K |
| IQ2_M | 11.7 GB | Fits | 256K | 256K |