Can I run gpt2 on an RTX 4060?
Yes
Its published weights (548 MB) fit in the 7.20 GB an RTX 4060 can use, with room for ≈ 1K tokens of context.
hf download openai-community/gpt2Files and context on RTX 4060
Each file of gpt2 (and of GGUF versions the site has analyzed), and how much context fits next to it, with the context cache in 16-bit or 8-bit.
| File | Size | On RTX 4060 | Context (16-bit) | Context (8-bit) |
|---|---|---|---|---|
| As published | 548 MB | Fits | 1K | 1K |
| 4-bit (estimate) | 77.1 MB | Fits | 1K | 1K |