← Back to Models
Googletransformer
Gemma 3 4B
4B parameters • 2.4GB VRAM (Q4 estimate) • 128,000 context
Estimate · Q4_K_M · 4k context · batch 1. For dense 7B–70B this sits ~5–8% above the Q4_K_M GGUF file. Not peak runtime VRAM and not a measured bench.
Specifications
Parameters
4B
VRAM (Q4 est.)
2.4 GB
VRAM (FP16)
8 GB
Context Window
128,000
Architecture
transformer
License
Gemma
Finding cloud alternatives...
Run Gemma 3 4B on GTX 1060 3GB
~2.4GB VRAM needed at Q4. GTX 1060 3GB has 3GB — buy hardware or rent cloud.
We may earn a commission from cloud and hardware partners at no extra cost to you.
Find the cheapest GPU that can run Gemma 3 4B
Compatibility Lab — check every GPU × quantization combination