← Back to Models
OpenAItransformer
Whisper Large v3 Turbo
0.8B parameters • 0.5GB VRAM (Q4 estimate) • 4,096 context
Estimate · Q4_K_M · 4k context · batch 1. For dense 7B–70B this sits ~5–8% above the Q4_K_M GGUF file. Not peak runtime VRAM and not a measured bench.
Specifications
Parameters
0.8B
VRAM (Q4 est.)
0.5 GB
VRAM (FP16)
1.6 GB
Context Window
4,096
Architecture
transformer
License
MIT
Finding cloud alternatives...
Run Whisper Large v3 Turbo on GTX 1050
~0.5GB VRAM needed at Q4. GTX 1050 has 2GB — buy hardware or rent cloud.
We may earn a commission from cloud and hardware partners at no extra cost to you.
Find the cheapest GPU that can run Whisper Large v3 Turbo
Compatibility Lab — check every GPU × quantization combination