← Back to Models
Googletransformer
Gemini 1.5 Flash 8B
8B parameters • 4.9GB VRAM (Q4 estimate) • 1,000,000 context
Estimate · Q4_K_M · 4k context · batch 1. For dense 7B–70B this sits ~5–8% above the Q4_K_M GGUF file. Not peak runtime VRAM and not a measured bench.
Specifications
Parameters
8B
VRAM (Q4 est.)
4.9 GB
VRAM (FP16)
16 GB
Context Window
1,000,000
Architecture
transformer
License
Google TOS
Finding cloud alternatives...
Run Gemini 1.5 Flash 8B on Tesla K20c
~4.9GB VRAM needed at Q4. Tesla K20c has 5GB — buy hardware or rent cloud.
We may earn a commission from cloud and hardware partners at no extra cost to you.
Find the cheapest GPU that can run Gemini 1.5 Flash 8B
Compatibility Lab — check every GPU × quantization combination