Skip to main content
idapt
HomeCodeAI ModelsPricing
Sign inStart free
Qwen

Qwen2.5 VL 72B

$$Standard
Qwen#65 of 92
CompareChat

Large 72B vision-language model from Qwen 2.5 generation. Strong at visual understanding, document parsing, and image QA. Good multimodal model for complex visual reasoning tasks.

▼

Qwen2.5 VL 72B by Qwen has a Capability Index of 62.8, making it below the field average in capability. It ranks #65 out of 92 models on the idapt capability leaderboard. Qwen2.5 VL 72B is somewhat expensive at $0.80/M tokens input and $1.00/M tokens output. Cached input reads are priced at $0.40/M tokens, making repeated context cheaper. The model supports vision/image input, a 128K context window.

Best forFrontier reasoningVision tasksCost-effective

Benchmarks

CodingInsufficient data
AgenticInsufficient data
ReasoningInsufficient data
MathInsufficient data
Sources:Epoch ECI·OpenRouter· as of 2026-07-10

Category Rankings

Capability
#65of 92

Specifications

Vision
Audio
Reasoning
Dense
Architecture
72B
Parameters
128K
Context
64K
Max output

Performance

Providers

Misc

February 1, 2025
Released
Same family
QwenQwen2.5 VL 32BQwenQwen VL MaxQwenQwen VL PlusQwenQwen2.5 72BQwenQwen2.5 7BQwenQwen2.5 Coder 32BQwenQwen2.5 Coder 7BQwenQwen2.5 VL 7B
CompareChat with Qwen2.5 VL 72B

Compare with

OpenAIGPT 5.6 Solvs Qwen2.5 VL 72BOpenAIGPT 5.6 Lunavs Qwen2.5 VL 72BAnthropicClaude Fable 5vs Qwen2.5 VL 72BAnthropicClaude Opus 4.8vs Qwen2.5 VL 72BAnthropicClaude Sonnet 5vs Qwen2.5 VL 72BAnthropicClaude Haiku 4.5vs Qwen2.5 VL 72BGoogleGemini 3.6 Flashvs Qwen2.5 VL 72BGoogleGemini 3.5 Flashvs Qwen2.5 VL 72BGoogleGemini 3.5 Flash-Litevs Qwen2.5 VL 72BGoogleGemma 4 31Bvs Qwen2.5 VL 72BxAIGrok 4.5vs Qwen2.5 VL 72BxAIGrok Build 0.1vs Qwen2.5 VL 72BxAIGrok 4.1 Fastvs Qwen2.5 VL 72BMetaMuse Spark 1.1vs Qwen2.5 VL 72BDeepSeekDeepSeek V4 Provs Qwen2.5 VL 72BDeepSeekDeepSeek V4 Flashvs Qwen2.5 VL 72BMiniMaxMiniMax M3vs Qwen2.5 VL 72BMoonshotAIKimi K3vs Qwen2.5 VL 72BMoonshotAIKimi K2.7 Codevs Qwen2.5 VL 72BZhipu AIGLM 5.2vs Qwen2.5 VL 72BQwenQwen3.7 Maxvs Qwen2.5 VL 72BQwenQwen3.7 Plusvs Qwen2.5 VL 72BQwenQwen3 Coder Nextvs Qwen2.5 VL 72BNVIDIANemotron 3 Ultravs Qwen2.5 VL 72BMistral AIMistral Small 4vs Qwen2.5 VL 72B
Frequently Asked Questions about Qwen2.5 VL 72B▼

When was Qwen2.5 VL 72B released?

▼

Qwen2.5 VL 72B was released on February 1, 2025.

Who created Qwen2.5 VL 72B?

▼

Qwen2.5 VL 72B was created by Qwen.

How capable is Qwen2.5 VL 72B?

▼

Qwen2.5 VL 72B has a Capability Index of 62.8/100, making it below the field average in capability. It ranks #65 out of 92 models.

How much does Qwen2.5 VL 72B cost?

▼

Qwen2.5 VL 72B costs $0.80/M input tokens and $1.00/M output tokens.

What is Qwen2.5 VL 72B API pricing?

▼

The Qwen2.5 VL 72B API is priced at $0.80 per million input tokens and $1.00 per million output tokens. You can access Qwen2.5 VL 72B through idapt.app alongside 200+ other AI models in one workspace.

Is Qwen2.5 VL 72B a reasoning model?

▼

No, Qwen2.5 VL 72B is not a reasoning model. It responds directly without an extended internal thinking step. For tasks requiring deep chain-of-thought reasoning, consider models like Claude, o-series from OpenAI, or Gemini Thinking.

Does Qwen2.5 VL 72B support image or vision input?

▼

Yes, Qwen2.5 VL 72B supports vision/image input. You can share images in your conversation and the model will analyze and respond based on their content.

Does Qwen2.5 VL 72B support audio?

▼

No, Qwen2.5 VL 72B does not support audio input.

What is the context window of Qwen2.5 VL 72B?

▼

Qwen2.5 VL 72B has a context window of 128K tokens, meaning it can process approximately 96,000 words of context at once.

Is Qwen2.5 VL 72B open source?

▼

Qwen2.5 VL 72B is a proprietary model developed by Qwen. The model weights and training data are not publicly available.

What is Qwen2.5 VL 72B good at?

▼

Based on benchmark data, Qwen2.5 VL 72B excels at: visual analysis and image understanding. It is Qwen's capable model suitable for a wide range of tasks.

Where can I use Qwen2.5 VL 72B?

▼

You can chat with Qwen2.5 VL 72B on idapt.app — no separate API key required. idapt is an AI workspace with 200+ models, agents, computers, and tasks. You can also access Qwen2.5 VL 72B through Qwen's own API or platform.

  • Home
  • Pricing
  • AI Models
  • Image models
  • Voice models
  • Video models
  • Rankings
  • New models
  • Model status
  • Multi-Model Chat
  • Voice Mode
  • Voice HUD
  • Web Search
  • Image Generation
  • Video Generation
  • Audio Generation
  • Transcription
  • Drive
  • Credentials
  • Sharing
  • Workspaces
  • Tasks
  • Memory
  • Agents
  • Subagents
  • Automations
  • Skills
  • idapt Code
  • Code Execution
  • Computers
  • Computer Use
  • Computer Assist · Soon
  • Containers · Soon
  • Cloud Computers
  • Local AI
  • AI Gateway
  • API & SDK
  • CLI
  • MCP
  • Tunnels
  • All features →
  • LLM cost calculator
  • Token counter
  • Context window checker
  • Can I run it
  • Model picker quiz
  • Savings finder
  • Video cost estimator
  • Text to speech cost
  • Transcription cost
  • API endpoint tester
  • All free tools →
  • Blog
  • Use cases
  • Comparisons
  • Best of
  • Skills
  • Learn
  • Changelog
  • Help center
  • FAQ
  • Privacy
  • Compare all models
  • Support
  • idapt Code
  • Developers
  • Quickstarts
  • API reference
  • API pricing
  • CLI
  • MCP
  • Downloads
  • Desktop
  • Badges and embeds
© idapt[email protected]TermsPrivacy PolicyLegal noticeReport content
X (Twitter)