Can I run Qwen3-Coder-480B-A35B-Instruct-FP8 on an NVIDIA H100?
No
Even its smallest file (4-bit (estimate), 270 GB) is more than an NVIDIA H100 can hold (72.0 GB usable): it needs 4 of them.
Files and context on NVIDIA H100
Each file of Qwen3-Coder-480B-A35B-Instruct-FP8 (and of GGUF versions the site has analyzed), and how much context fits next to it, with the context cache in 16-bit or 8-bit.
| File | Size | On NVIDIA H100 | Context (16-bit) | Context (8-bit) |
|---|---|---|---|---|
| As published | 482 GB | Doesn't fit | — | — |
| 4-bit (estimate) | 270 GB | Doesn't fit | — | — |