← Back to Models
Alibabatransformer
GTE-Qwen2 1.5B Instruct
1.5B parameters • 0.9GB VRAM (Q4 estimate) • 131,072 context
Estimate · Q4_K_M · 4k context · batch 1. For dense 7B–70B this sits ~5–8% above the Q4_K_M GGUF file. Not peak runtime VRAM and not a measured bench.
Specifications
Parameters
1.5B
VRAM (Q4 est.)
0.9 GB
VRAM (FP16)
3 GB
Context Window
131,072
Architecture
transformer
License
MIT
Finding cloud alternatives...
Run GTE-Qwen2 1.5B Instruct on GTX 1050
~0.9GB VRAM needed at Q4. GTX 1050 has 2GB — buy hardware or rent cloud.
We may earn a commission from cloud and hardware partners at no extra cost to you.
Find the cheapest GPU that can run GTE-Qwen2 1.5B Instruct
Compatibility Lab — check every GPU × quantization combination