GPT-5.6
OpenAI
Sol · Terra · Luna · default in ChatGPT since 9 Jul
The volume leader. Sol is the high-compute flagship; Terra is the everyday workhorse; Luna was cut 80% on 30 Jul to $0.20 / $1.20 per million tokens. Best general assistant most people can actually open. Sol leads several hard math and reasoning boards.
Claude Opus 5
Anthropic
Released 24 Jul · Fable 5 still the long-horizon specialist
Current intelligence-index leader (Artificial Analysis v4.1: 61). Took Arena’s coding and agent boards. Same $5 / $25 price as Opus 4.8. Fable 5 / Mythos-class sits beside it for the hardest writing and long jobs. Enterprise revenue now outruns OpenAI’s last reported run-rate.
Gemini 3.1 Pro
Google DeepMind
Plus 3.5 / 3.6 Flash and 3 Deep Think
The accuracy and multimodal pick, wired through Search, Android, and Workspace. 3.1 Pro has led GPQA Diamond and cheap ARC-AGI-1 work. Google has been shipping Flash-tier speed this summer rather than a new closed flagship. Gemini Spark is the always-on agent from I/O.
Grok 4.5
xAI / SpaceX
Launched 8 Jul · also Grok Build
Real-time X context, fewer content limits, and a 2M-token class window on the prior 4.20 line. AA index around 54 — a step behind Opus 5 and Sol, cheaper than both. Distribution is the X graph: 117M feature users, not a ChatGPT-scale destination app.
DeepSeek V4
DeepSeek
V4-Pro · V4-Flash 0731 · MIT license
The price-performance wrecking ball. Flash 0731 is about $0.14 / $0.28 per million tokens with a mid-tier intelligence score. Pro is a 1.6T MoE (49B active) that sits within a few GPQA points of the Western flagships. Open weights, 1M context, China-heavy users.
Llama 4
Meta
Open weights · Meta AI / Muse Spark in-app
The default open stack for everyone who will not pay OpenAI or Anthropic. Meta AI is the most-touched assistant on Earth if you count WhatsApp and Instagram taps, and almost invisible as a destination site. Llama remains the ecosystem play, not the LMSYS crown.
Copilot
Microsoft
Routes GPT-5.6, Claude, and Microsoft’s own MAI models
Not a single frontier model so much as the workplace surface. GitHub Copilot now auto-routes by task. Microsoft is also shipping cheaper in-house models to reduce OpenAI dependence. Tiny standalone web share, enormous Office and Windows footprint.
Sonar
Perplexity
Sonar / Pro / Reasoning · also routes Kimi K2.5
The cited-answer engine. Losing consumer web share as ChatGPT and Gemini absorbed search-like answers, still the cleanest product if you want sources first. $22.6B company on roughly half a billion of annualized revenue.
Kimi K3
Moonshot
AA index 57 · between Sol and Opus 4.8
The surprise on the July board. K3 sits ahead of last-gen Western flagships on Artificial Analysis while remaining far cheaper. Evidence that the second tier of Chinese labs is now a frontier participant, not a clone factory.
Qwen 3.7
Alibaba
Plus / Max · China + global open weights
Alibaba’s open line remains the other default for self-hosting beside Llama and DeepSeek. Strong coding variants, aggressive price, and a large domestic chat surface that never appears in English-language Similarweb pies.
Mistral
Mistral AI
Magistral · Medium · Le Chat
Europe’s lab. Not winning the intelligence index, still the sovereign and on-prem option for buyers who will not send tokens to US or Chinese hosts. Le Chat is a small slice of the “Other 4.2%” in US chatbot share.