← Back to Models
Qwentransformer
Qwen2 VL 72B Instruct
72B parameters • 43.8GB VRAM (Q4 estimate) • 8,192 context
Estimate · Q4_K_M · 4k context · batch 1. For dense 7B–70B this sits ~5–8% above the Q4_K_M GGUF file. Not peak runtime VRAM and not a measured bench.
Specifications
Parameters
72B
VRAM (Q4 est.)
43.8 GB
VRAM (FP16)
144 GB
Context Window
8,192
Architecture
transformer
License
Apache-2.0
Finding cloud alternatives...
Run Qwen2 VL 72B Instruct on Radeon PRO W7900
~43.8GB VRAM needed at Q4. Radeon PRO W7900 has 48GB — buy hardware or rent cloud.
We may earn a commission from cloud and hardware partners at no extra cost to you.
Find the cheapest GPU that can run Qwen2 VL 72B Instruct
Compatibility Lab — check every GPU × quantization combination