Can I run gpt2 on an RTX 5080?
Yes
Its published weights (548 MB) fit in the 14.4 GB an RTX 5080 can use, with room for ≈ 1K tokens of context.
hf download openai-community/gpt2Files and context on RTX 5080
Each file of gpt2 (and of GGUF versions the site has analyzed), and how much context fits next to it, with the context cache in 16-bit or 8-bit.
| File | Size | On RTX 5080 | Context (16-bit) | Context (8-bit) |
|---|---|---|---|---|
| As published | 548 MB | Fits | 1K | 1K |
| 4-bit (estimate) | 77.1 MB | Fits | 1K | 1K |