Models that run tools well: the benchmark band leads with Terminal-Bench and the agent-loop demos show real multi-step work where captures exist.Last reviewed 2026-07-17.
Anthropic's fifth-generation Sonnet — near-Opus intelligence at Sonnet speed and price.
xAI's agentic build model for autonomous software engineering.
MoonshotAI's 2.8T-parameter flagship for complex coding, knowledge work, and long-horizon agent runs.
No shared demos for these models yet.
No captured outputs for these models yet.
Grok Build 0.1 has the lowest blended list price. Per 1M tokens: Claude Sonnet 5 at $2.00 in / $10.00 out; Grok Build 0.1 at $1.00 in / $2.00 out; Kimi K3 at $3.00 in / $15.00 out.
Kimi K3 leads with 1M tokens. The other two: Claude Sonnet 5 at 1M, Grok Build 0.1 at 256K.
Yes, all three accept images.
This idapt comparison runs all three in aligned columns: the same prompt, plus benchmarks, specs, and prices per model. Open any model in chat from its column to keep working with it.