Ai2

OLMo 3.1 32B Think

An open reasoning checkpoint for coding, math, and multi-step tasks.

open-weightsreasoningcodingmath

Model record

OLMo 3.1 32B Think model overview

OLMo 3.1 32B Think is a open weights model from Ai2. It was released on December 12, 2025. Its strongest case is open weights, while buyers should account for context unavailable.

OLMo 3.1 32B Think benchmark snapshot

EvaluationReported scoreWhat it probes
MMLU-ProNot reportedBroad knowledge and multi-step reasoning
GPQA DiamondNot reportedGraduate-level science reasoning
SWE-bench VerifiedNot reportedVerified real-repository issue resolution
LiveCodeBenchNot reportedContamination-aware competitive programming
SWE-Bench ProNot reportedLonger, harder professional software tasks
Artificial Analysis Intelligence IndexNot reportedComposite third-party capability index; version matters

Scores are percentages reported by model creators or benchmark maintainers under varying settings. A blank is preferable to an inferred result. See our methodology.

OLMo 3.1 32B Think strengths

  • Open weights
  • Transparent release

OLMo 3.1 32B Think limitations

  • Context unavailable
  • Pricing unavailable
Editorial take: Benchmark rank should narrow a shortlist, not close a purchase. Run a private evaluation with representative prompts, failure cases, latency targets, and total token costs.

Source record

Specifications and scores are linked to the best source located during review. Provider-reported results are not presented as third-party lab reproductions.

Source limitation: Supplemental official model card: https://huggingface.co/allenai/Olmo-3.1-32B-Think

Read Ai2 OLMo 3 announcement