← Back to Models
Mistral AImoe
Mixtral 8x7B
46.7B total / 12.9B active · load estimate uses total • 28.4GB VRAM (Q4 estimate) • 32,768 context
Estimate · Q4_K_M · 4k context · batch 1. For dense 7B–70B this sits ~5–8% above the Q4_K_M GGUF file. Not peak runtime VRAM and not a measured bench.
Specifications
Parameters
46.7B total / 12.9B active · load estimate uses total
VRAM (Q4 est.)
28.4 GB
VRAM (FP16)
93.4 GB
Context Window
32,768
Architecture
moe
License
Apache-2.0
Finding cloud alternatives...
Run Mixtral 8x7B on MI100
~28.4GB VRAM needed at Q4. MI100 has 32GB — buy hardware or rent cloud.
We may earn a commission from cloud and hardware partners at no extra cost to you.
Find the cheapest GPU that can run Mixtral 8x7B
Compatibility Lab — check every GPU × quantization combination