Model record
Kimi K3 model overview
Kimi K3 is a api + app model from Moonshot AI. It was released on July 16, 2026. Its strongest case is one-million-token context, while buyers should account for full weights were announced for a later date and were not yet available when checked.
Kimi K3 benchmark snapshot
| Evaluation | Reported score | What it probes |
|---|---|---|
| MMLU-Pro | Not reported | Broad knowledge and multi-step reasoning |
| GPQA Diamond | Not reported | Graduate-level science reasoning |
| SWE-bench Verified | Not reported | Verified real-repository issue resolution |
| LiveCodeBench | Not reported | Contamination-aware competitive programming |
| SWE-Bench Pro | Not reported | Longer, harder professional software tasks |
| Artificial Analysis Intelligence Index | Not reported | Composite third-party capability index; version matters |
Scores are percentages reported by model creators or benchmark maintainers under varying settings. A blank is preferable to an inferred result. See our methodology.
Kimi K3 strengths
- One-million-token context
- Native vision and long-horizon coding
- First-party app, coding-agent, and API access
Kimi K3 limitations
- Full weights were announced for a later date and were not yet available when checked
- Requires preserved thinking history and may act too proactively without explicit constraints
- Provider reports a remaining UX gap versus its cited proprietary frontier peers
Compare Kimi K3
- Kimi K3 vs GPT-5.6 Sol — review fit, trade-offs, listed prices, and an evaluation plan.
Source record
Specifications and scores are linked to the best source located during review. Provider-reported results are not presented as third-party lab reproductions.
Source limitation: First-party Kimi source. Full weights were announced for release by 2026-07-27 and were not treated as available on 2026-07-19. Prices are official Kimi API USD per MTok; inputPrice records cache-miss input because the schema has no cached-input field.
Read Kimi K3 technical blog ↗