Model intelligence
Find the right model.
Skip the hype.
A curated reference to leading LLMs—reported benchmarks, listed pricing, licensing, and practical trade-offs in one place.
Latest model releases
Top 5 latest
models.
The five most recent models with verified release dates, ordered newest first. Unavailable dates and unverified fields are never inferred.
Gemini 3.6 Flash
Google's current stable Flash model for fast multimodal reasoning, coding, and agent workflows through Gemini apps and API.
Laguna S 2.1
Poolside's sparse open-weight coding and agent model, combining a million-token context window with a comparatively small active parameter count.
Kimi K3
Moonshot AI's native-vision, 2.8T-parameter MoE for long-horizon coding, knowledge work, and reasoning, available through Kimi apps and API with a one-million-token context window.
Inkling
A multimodal reasoning model for text, images, audio, coding, and tools.
Mercury 2
A diffusion language model for reasoning, coding, editing, tools, and real-time agents.
Start with the job
What are you
building?
Purpose-built shortlists account for deployment, cost, and licensing—not just a leaderboard position.
Best coding models
Code generation & maintenance
↗02Best reasoning models
Science & multi-step reasoning
↗03Cheapest LLM API models
Provider-listed token prices
↗04Fastest LLMs: speed watchlist
Output speed & latency
↗05Best large-context-window models
Large documents & repositories
↗06Best open-weight models
Downloadable & self-hosted
↗07Best multilingual LLMs
Languages & regional coverage
↗08Best LLMs for RAG
Search-grounded answers
↗09Best LLMs to run locally
Private & controlled deployment
↗Current benchmark snapshot
One table.
More context.
Scores are included only when a source reports the exact evaluation. Blank cells are deliberate, not an invitation to estimate.
| Model | Provider | GPQA Diamond | SWE-Bench Pro | AA Intelligence Index |
|---|---|---|---|---|
| OpenAI | 94.6 | 64.6 | 58.9 | |
| OpenAI | 92.9 | 63.4 | 55 | |
| OpenAI | 92.3 | 62.7 | 51.2 | |
| Moonshot AI | — | — | 57 | |
| Anthropic | 92.6 | 80 | 59.9 |
Provider-reported and cited comparative results can use different settings. Missing benchmark values remain blank rather than estimated. Data checked July 19, 2026.