Qwen

Qwen2.5 32B

Total params
32.5B
Active params
32.5B
Layers
64
Hidden size
5120
Attention heads
40
KV heads (GQA)
8
Vocab size
152,064
Native context window
131,072
Native precision
BF16
Size this model

Opens the sizing calculator pre-filled with this model at BF16 weights / FP16 KV cache and a typical workload — sign in to run it and see GPU/cloud recommendations.

Size this model →