The Capitals.
The key players — the dominant labs and flagship models.
Claude Code
Anthropic
Powered by Opus 4.8 — SWE-bench Pro 69.2 %, Verified 88.6 %. Dynamic workflows with parallel subagents; fast mode 3x cheaper.
Claude Cowork
Anthropic
Desktop agent now with Claude for Small Business: 15 agentic workflows across finance, ops, sales, marketing, HR and CS. Opus 4.8 upgrades flow through.
Claude Opus 4.8
Anthropic
SWE-bench Pro 69.2 %, Verified 88.6 %; dynamic workflows enable hundreds of parallel subagents. Anthropic filed a confidential S-1 at ~$965 B valuation.
Gemini 3.1 Pro
Leads most published reasoning benchmarks and has the cheapest output among the majors; paired with Gemini Spark for long cloud tasks.
HappyHorse
Alibaba
Currently tops the Artificial Analysis leaderboard.
Nano Banana Pro
Gemini 3 Pro Image; leads the image-arena leaderboard, with the best multilingual text rendering, free in the Gemini app.
Seedance 2.0
ByteDance
Currently tops the Artificial Analysis leaderboard.