The frontier tier: the newest reasoning-first flagships from Anthropic, OpenAI, and xAI. Benchmarks tell part of the story; the shared output band shows how differently they write and reason on the same prompt.Last reviewed 2026-09-09.
Anthropic's upgraded Mythos-class flagship, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work.
OpenAI's new flagship for demanding end-to-end work: advanced analysis, software engineering, deep research, scientific work, and document creation.
xAI's smartest model, with frontier performance on coding, knowledge work, and STEM.
No shared demos for these models yet.
No captured outputs for these models yet.
Grok 4.6 has the lowest blended list price. Per 1M tokens: Claude Fable 5.1 at $10.00 in / $50.00 out; GPT 6 Astra at $10.00 in / $50.00 out; Grok 4.6 at $2.00 in / $6.00 out.
GPT 6 Astra leads with 1.1M tokens. The other two: Claude Fable 5.1 at 1M, Grok 4.6 at 500K.
Yes, all three accept images.
This idapt comparison runs all three in aligned columns: the same prompt, plus benchmarks, specs, and prices per model. Open any model in chat from its column to keep working with it.