Can I run gpt2 on an RTX 3090?
Yes
Its published weights (548 MB) fit in the 21.6 GB an RTX 3090 can use, with room for ≈ 1K tokens of context.
hf download openai-community/gpt2Files and context on RTX 3090
Each file of gpt2 (and of GGUF versions the site has analyzed), and how much context fits next to it, with the context cache in 16-bit or 8-bit.
| File | Size | On RTX 3090 | Context (16-bit) | Context (8-bit) |
|---|---|---|---|---|
| As published | 548 MB | Fits | 1K | 1K |
| 4-bit (estimate) | 77.1 MB | Fits | 1K | 1K |