← Back to Models
meta-llamatransformer
Llama 3.3 70B Instruct
70B parameters • 38.5GB VRAM (Q4) • 8,192 context
Specifications
Parameters
70B
VRAM (Q4)
38.5 GB
VRAM (FP16)
140 GB
Context Window
8,192
Architecture
transformer
License
Apache-2.0
Finding cloud alternatives...
Run Llama 3.3 70B Instruct on NVIDIA A100 40GB PCIe
~38.5GB VRAM needed at Q4. NVIDIA A100 40GB PCIe has 40GB — buy hardware or rent cloud.
We may earn a commission from cloud and hardware partners at no extra cost to you.
Find the cheapest GPU that can run Llama 3.3 70B Instruct
Compatibility Lab — check every GPU × quantization combination