Skip to main content
idapt
HomeCodeAI ModelsPricing
Sign inStart free trial
  • Home
  • Pricing
  • AI Models
  • Image models
  • Voice models
  • Video models
  • Rankings
  • New models
  • Model status
  • Multi-Model Chat
  • Agents
  • Computers
  • Drive
  • Automations
  • AI Gateway
  • All features →
  • LLM cost calculator
  • Token counter
  • All free tools →
  • Blog
  • Use cases
  • Comparisons
  • Best of
  • Benchmarks
  • Demos
  • Changelog
  • Help center
  • FAQ
  • Privacy
  • Compare all models
  • Support
  • idapt Code
  • Developers
  • Quickstarts
  • API reference
  • API pricing
  • CLI
  • MCP
  • Downloads
  • Desktop
  • Badges and embeds
© idapt[email protected]TermsPrivacy PolicyLegal noticeReport content
X (Twitter)

DeepSeek V4 Pro vs Qwen3.7 Max vs Nemotron 3 Ultra

The strongest open-weight lines in one view. All three publish weights; the columns show what each costs to run hosted and how they score on the named benchmarks.Last reviewed 2026-07-17.

DeepSeek V4 Pro vs Qwen3.7 Max vs Nemotron 3 Ultra

At a glance

Open-source 1.6T-parameter MoE with 49B active, built for advanced reasoning and long-horizon agents.

821.0M ctx$0.43/$0.87 per M#21/92

Alibaba's flagship Qwen3.7 model, built for agent-centric workloads.

851M ctx$1.48/$4.42 per M#15/92

NVIDIA's open frontier reasoning model for orchestration, coding, and deep research.

512K ctx$0.60/$3.60 per M
Demos
All demos

No shared demos for these models yet.

AI Output
All AI outputs

No captured outputs for these models yet.

Benchmarks
See rankings
The paired significance test (McNemar, in "Measured on Idapt") appears only when exactly two models are pinned.
Reasoning
AgenticInsufficient data
Sources:Epoch ECI·Epoch AI·OpenRouter· as of 2026-07-10
Cheaper60%
Capability
CapabilityECI81.6
Reasoning & Knowledge
Graduate Science89.6%Expert-Level—Factual Recall57.0%
Math
Competition Math2%96.7%
Coding
Real-World Coding77.6%Scientific Code—
Agentic
Terminal Tasks—
Reasoning
Sources:Epoch ECI·Epoch AI·Provider·OpenRouter· as of 2026-07-10
Smarter4%
Capability
CapabilityECI4%84.7
Reasoning & Knowledge
Graduate Science2%91.6%Expert-Level38.1%Factual Recall3%58.5%
Math
Competition Math95.0%
Coding
Real-World Coding77.3%Scientific Code53.5%
Agentic
Terminal Tasks69.7%
Reasoning
CodingInsufficient data
AgenticInsufficient data
ReasoningInsufficient data
MathInsufficient data
Sources:OpenRouter
Capability
Capability—
Reasoning & Knowledge
Graduate Science—Expert-Level—Factual Recall—
Math
Competition Math—
Coding
Real-World Coding—Scientific Code—
Agentic
Terminal Tasks—
Specs
1.0M
Context
256K
Max output
Apr 2026
Released
1M
Context
66K
Max output
May 2026
Released
512K
Context
16K
Max output
Jun 2026
Released
Capabilities
Vision
Audio
Reasoning
Vision
Audio
Reasoning
Vision
Audio
Reasoning
Pricing
Live pricing, shown before you run
Input27%$0.43/M tokens
Output76%$0.87/M tokens
Cache read98%$0.004/M tokens
Live pricing, shown before you run
Input$1.48/M tokens
Output$4.42/M tokens
Cache read$0.29/M tokens
Cache write$1.84/M tokens
Live pricing, shown before you run
Input$0.60/M tokens
Output$3.60/M tokens
Cache read$0.20/M tokens
Chat with DeepSeek V4 ProGo to model
Chat with Qwen3.7 MaxGo to model
Chat with Nemotron 3 UltraGo to model

Frequently asked

Which of DeepSeek V4 Pro, Qwen3.7 Max, and Nemotron 3 Ultra is the cheapest?

DeepSeek V4 Pro has the lowest blended list price. Per 1M tokens: DeepSeek V4 Pro at $0.43 in / $0.87 out; Qwen3.7 Max at $1.48 in / $4.42 out; Nemotron 3 Ultra at $0.60 in / $3.60 out.

Which of DeepSeek V4 Pro, Qwen3.7 Max, and Nemotron 3 Ultra has the largest context window?

DeepSeek V4 Pro leads with 1M tokens. The other two: Qwen3.7 Max at 1M, Nemotron 3 Ultra at 512K.

Do DeepSeek V4 Pro, Qwen3.7 Max, and Nemotron 3 Ultra support image input?

No, none of the three accepts image input.

Where can I run DeepSeek V4 Pro, Qwen3.7 Max, and Nemotron 3 Ultra side by side?

This idapt comparison runs all three in aligned columns: the same prompt, plus benchmarks, specs, and prices per model. Open any model in chat from its column to keep working with it.

More three-way comparisons

GPT 5.6 Sol vs Claude Opus 4.8 vs Gemini 3.1 ProClaude Fable 5 vs GPT 5.6 Sol vs Grok 4.5GPT 5.6 Luna vs Gemini 3.5 Flash vs Claude Haiku 4.5Claude Sonnet 5 vs Kimi K2.7 Code vs Qwen3 Coder NextMiniMax M3 vs DeepSeek V4 Flash vs Gemini 3.1 Flash LiteClaude Sonnet 5 vs Grok Build 0.1 vs Kimi K3Muse Spark 1.1 vs Kimi K3 vs Grok 4.5DeepSeek V4 Flash vs Gemini 3.5 Flash vs Qwen3.7 PlusGLM 5.2 vs MiniMax M3 vs DeepSeek V4 Pro