Can I run Gemma-4-E4B-Uncensored-HauhauCS-Aggressive on an RTX 4060?
Yes, just
Its Q6_K file (6.25 GB) fits in the 7.20 GB an RTX 4060 can use, with room for ≈ 30K tokens of context.
Full analysis of Gemma-4-E4B-Uncensored-HauhauCS-Aggressive
Everything that runs on RTX 4060
Get the Q6_K file on Hugging Face
ollama run hf.co/HauhauCS/Gemma-4-E4B-Uncensored-HauhauCS-Aggressive:Q6_KFiles and context on RTX 4060
Each file of Gemma-4-E4B-Uncensored-HauhauCS-Aggressive (and of GGUF versions the site has analyzed), and how much context fits next to it, with the context cache in 16-bit or 8-bit.
| File | Size | On RTX 4060 | Context (16-bit) | Context (8-bit) |
|---|---|---|---|---|
| Q8_K | 8.13 GB | Doesn't fit | — | — |
| Q6_K | 6.25 GB | Tight | 30K | 62K |
| Q5_K | 5.81 GB | Tight | 45K | 92K |
| Q5_K_M | 5.76 GB | Tight | 46K | 95K |
| Q4_K | 5.37 GB | Fits | 60K | 122K |
| Q4_K_M | 5.34 GB | Fits | 61K | 125K |
| IQ4_XS | 5.07 GB | Fits | 70K | 128K |
| Q3_K | 4.88 GB | Fits | 76K | 128K |
| Q3_K_M | 4.85 GB | Fits | 78K | 128K |
| IQ3_M | 4.71 GB | Fits | 82K | 128K |
| Q2_K | 4.43 GB | Fits | 92K | 128K |