Skip to main content
idapt
HomeCodeAI ModelsPricing
Sign inStart free
  • Home
  • Pricing
  • AI Models
  • Image models
  • Voice models
  • Video models
  • Rankings
  • New models
  • Model status
  • Multi-Model Chat
  • Agents
  • Computers
  • Drive
  • Automations
  • AI Gateway
  • All features →
  • LLM cost calculator
  • Token counter
  • All free tools →
  • Blog
  • Use cases
  • Comparisons
  • Best of
  • Benchmarks
  • Demos
  • Changelog
  • Help center
  • FAQ
  • Privacy
  • Compare all models
  • Support
  • idapt Code
  • Developers
  • Quickstarts
  • API reference
  • API pricing
  • CLI
  • MCP
  • Downloads
  • Desktop
  • Badges and embeds
© idapt[email protected]TermsPrivacy PolicyLegal noticeReport content
X (Twitter)

DeepSeek V4 Pro (0813) vs Qwen3.8 2.4T A95B vs Nemotron 3 Ultra

The strongest open-weight lines in one view. All three publish weights; the columns show what each costs to run hosted and how they score on the named benchmarks.Last reviewed 2026-09-09.

DeepSeek V4 Pro (0813) vs Qwen3.8 2.4T A95B vs Nemotron 3 Ultra

At a glance

The GA release of DeepSeek's open 1.6T-parameter MoE flagship for advanced reasoning and long-horizon agents.

1.0M ctx$0.66/$1.98 per M

The open-weight variant of Qwen3.8 Max, with 95 billion active parameters out of 2.4 trillion total.

1M ctx$2.00/$6.00 per M

NVIDIA's open frontier reasoning model for orchestration, coding, and deep research.

512K ctx$0.60/$3.60 per M
Demos
All demos

No shared demos for these models yet.

AI Output
All AI outputs

No captured outputs for these models yet.

Benchmarks
See rankings
The paired significance test (McNemar, in "Measured on Idapt") appears only when exactly two models are pinned.
Reasoning
CodingInsufficient data
AgenticInsufficient data
ReasoningInsufficient data
MathInsufficient data
Sources:OpenRouter
Cheaper27%
Reasoning
CodingInsufficient data
AgenticInsufficient data
ReasoningInsufficient data
MathInsufficient data
Sources:OpenRouter
Reasoning
CodingInsufficient data
AgenticInsufficient data
ReasoningInsufficient data
MathInsufficient data
Sources:OpenRouter
Specs
1.0M
Context
256K
Max output
Aug 2026
Released
1M
Context
131K
Max output
Aug 2026
Released
512K
Context
16K
Max output
Jun 2026
Released
Capabilities
Vision
Audio
Reasoning
Vision
Audio
Reasoning
Vision
Audio
Reasoning
Pricing
Live pricing, shown before you run
Input$0.66/M tokens
Output45%$1.98/M tokens
Cache read90%$0.02/M tokens
Live pricing, shown before you run
Input$2.00/M tokens
Output$6.00/M tokens
Cache read$0.25/M tokens
Live pricing, shown before you run
Input9%$0.60/M tokens
Output$3.60/M tokens
Cache read$0.20/M tokens
Chat with DeepSeek V4 Pro (0813)Go to model
Chat with Qwen3.8 2.4T A95BGo to model
Chat with Nemotron 3 UltraGo to model

Frequently asked

Which of DeepSeek V4 Pro (0813), Qwen3.8 2.4T A95B, and Nemotron 3 Ultra is the cheapest?

DeepSeek V4 Pro (0813) has the lowest blended list price. Per 1M tokens: DeepSeek V4 Pro (0813) at $0.66 in / $1.98 out; Qwen3.8 2.4T A95B at $2.00 in / $6.00 out; Nemotron 3 Ultra at $0.60 in / $3.60 out.

Which of DeepSeek V4 Pro (0813), Qwen3.8 2.4T A95B, and Nemotron 3 Ultra has the largest context window?

DeepSeek V4 Pro (0813) leads with 1M tokens. The other two: Qwen3.8 2.4T A95B at 1M, Nemotron 3 Ultra at 512K.

Do DeepSeek V4 Pro (0813), Qwen3.8 2.4T A95B, and Nemotron 3 Ultra support image input?

No, none of the three accepts image input.

Where can I run DeepSeek V4 Pro (0813), Qwen3.8 2.4T A95B, and Nemotron 3 Ultra side by side?

This idapt comparison runs all three in aligned columns: the same prompt, plus benchmarks, specs, and prices per model. Open any model in chat from its column to keep working with it.

More three-way comparisons

GPT 6 Astra vs Claude Opus 5 vs Gemini 3.1 ProClaude Fable 5.1 vs GPT 6 Astra vs Grok 4.6GPT 5.6 Luna vs Gemini 3.8 Flash vs Claude Haiku 4.5Claude Sonnet 5 vs Kimi K2.7 Code vs Qwen3 Coder NextGLM 5.3 Flash vs MiniMax M3 vs DeepSeek V4 FlashClaude Sonnet 5 vs Grok Build 0.1 vs Kimi K3Muse Spark 1.3 vs Kimi K3 vs Grok 4.6DeepSeek V4 Flash vs Gemini 3.8 Flash vs Qwen3.7 PlusGLM 5.3 vs MiniMax M3 vs DeepSeek V4 Pro (0813)