Qwen
Qwen2.5 7B
Total params
7.6B
Active params
7.6B
Layers
28
Hidden size
3584
Attention heads
28
KV heads (GQA)
4
Vocab size
152,064
Native context window
131,072
Native precision
BF16
Size this model
Opens the sizing calculator pre-filled with this model at BF16 weights / FP16 KV cache and a typical workload — sign in to run it and see GPU/cloud recommendations.
Size this model →