Alibaba

Qwen2.5-Max

A large mixture-of-experts model trained on a broad multilingual corpus. It targets general reasoning, coding, and instruction-following workloads through Alibaba Cloud.

codingmultilingualreasoning

Model record

Qwen2.5-Max model overview

Qwen2.5-Max is a api model from Alibaba. It was released on January 29, 2025. Its strongest case is strong multilingual coverage, while buyers should account for core weights are not available.

Qwen2.5-Max benchmark snapshot

EvaluationReported scoreWhat it probes
MMLU-Pro76.1Broad knowledge and multi-step reasoning
GPQA Diamond60.1Graduate-level science reasoning
SWE-bench Verified38.8Verified real-repository issue resolution
LiveCodeBench38.7Contamination-aware competitive programming
SWE-Bench ProNot reportedLonger, harder professional software tasks
Artificial Analysis Intelligence IndexNot reportedComposite third-party capability index; version matters

Scores are percentages reported by model creators or benchmark maintainers under varying settings. A blank is preferable to an inferred result. See our methodology.

Qwen2.5-Max strengths

  • Strong multilingual coverage
  • Competitive general reasoning
  • MoE architecture

Qwen2.5-Max limitations

  • Core weights are not available
  • Shorter context than some peers
Editorial take: Benchmark rank should narrow a shortlist, not close a purchase. Run a private evaluation with representative prompts, failure cases, latency targets, and total token costs.

Compare Qwen2.5-Max

Source record

Specifications and scores are linked to the best source located during review. Provider-reported results are not presented as third-party lab reproductions.

Read Qwen model announcement