Model record
OLMo 3.1 32B Think model overview
OLMo 3.1 32B Think is a open weights model from Ai2. It was released on December 12, 2025. Its strongest case is open weights, while buyers should account for context unavailable.
OLMo 3.1 32B Think benchmark snapshot
| Evaluation | Reported score | What it probes |
|---|---|---|
| MMLU-Pro | Not reported | Broad knowledge and multi-step reasoning |
| GPQA Diamond | Not reported | Graduate-level science reasoning |
| SWE-bench Verified | Not reported | Verified real-repository issue resolution |
| LiveCodeBench | Not reported | Contamination-aware competitive programming |
| SWE-Bench Pro | Not reported | Longer, harder professional software tasks |
| Artificial Analysis Intelligence Index | Not reported | Composite third-party capability index; version matters |
Scores are percentages reported by model creators or benchmark maintainers under varying settings. A blank is preferable to an inferred result. See our methodology.
OLMo 3.1 32B Think strengths
- Open weights
- Transparent release
OLMo 3.1 32B Think limitations
- Context unavailable
- Pricing unavailable
Editorial take: Benchmark rank should narrow a shortlist, not close a purchase. Run a private evaluation with representative prompts, failure cases, latency targets, and total token costs.
Source record
Specifications and scores are linked to the best source located during review. Provider-reported results are not presented as third-party lab reproductions.
Source limitation: Supplemental official model card: https://huggingface.co/allenai/Olmo-3.1-32B-Think
Read Ai2 OLMo 3 announcement ↗