Gemma
Gemma 2 9B
Total params
9.24B
Active params
9.24B
Layers
42
Hidden size
3584
Attention heads
16
KV heads (GQA)
8
Vocab size
256,128
Native context window
8,192
Native precision
BF16
Size this model
Opens the sizing calculator pre-filled with this model at BF16 weights / FP16 KV cache and a typical workload — sign in to run it and see GPU/cloud recommendations.
Size this model →