OLMo 2 32B Instruct March 2025 is post-trained variant of the OLMo-2 32B March 2025 model, which has undergone supervised finetuning on an OLMo-specific variant of the Tülu 3 dataset, further DPO training on this dataset, and final RLVR training o...
chat
32.2B
Parameters
4K
Context length
15
Benchmarks
14
Quantizations
7K
HF downloads
Architecture
Dense
Released
2025-06-15
Layers
64
KV Heads
8
Head Dim
128
Family
olmo
Quantization Options
Quant
Bits
VRAM @ 4K
Quality
IQ3_XXS
3.25
13.6 GB
low
IQ3_XS
3.5
14.6 GB
low
Q3_K_S
3.64
15.1 GB
low
IQ3_M
3.76
15.6 GB
low
Q3_K_M
4
16.6 GB
low
Q3_K_L
4.3
17.8 GB
moderate
IQ4_XS
4.46
18.4 GB
moderate
Q4_K_S
4.67
19.3 GB
moderate
Q4_K_M
4.89
20.2 GB
good
Q5_K_S
5.57
22.9 GB
good
Q5_K_M
5.7
23.4 GB
good
Q6_K
6.56
26.9 GB
excellent
Q8_0
8.5
34.7 GB
lossless
FP16
16
64.9 GB
lossless
Select your GPU above to see speed estimates and compatibility for each quantization.
Detects your GPU, recommends the best model, downloads it, and starts chatting — zero config. Benchmarks your speed and contributes anonymous data to improve predictions.