Fast models carry most day-to-day work. These three trade a little peak capability for speed and price; the columns show exactly how much of each.Last reviewed 2026-09-09.
OpenAI's fastest, most affordable GPT-5.6 tier for high-volume, latency-sensitive work.
Google's most intelligent Flash model, with significant gains over 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.
Fast and efficient with near-frontier intelligence.
No shared demos for these models yet.
No captured outputs for these models yet.
Gemini 3.8 Flash has the lowest blended list price. Per 1M tokens: GPT 5.6 Luna at $1.00 in / $6.00 out; Gemini 3.8 Flash at $0.75 in / $3.75 out; Claude Haiku 4.5 at $1.00 in / $5.00 out.
GPT 5.6 Luna leads with 1.1M tokens. The other two: Gemini 3.8 Flash at 1M, Claude Haiku 4.5 at 200K.
Yes, all three accept images.
This idapt comparison runs all three in aligned columns: the same prompt, plus benchmarks, specs, and prices per model. Open any model in chat from its column to keep working with it.