45 text
Alibaba's cost-efficient Qwen3.7 model for long-context multimodal work.
Alibaba's flagship Qwen3.7 model, built for agent-centric workloads.
Open-source 1.6T-parameter MoE with 49B active, built for advanced reasoning and long-horizon agents.
Efficiency-optimized DeepSeek V4 at 284B total / 13B active for fast, high-throughput inference.
Qwen's latest hybrid model combining linear attention with sparse MoE routing.
Compact 9B vision-language model from the Qwen3.5 family.
Efficient vision-language MoE with linear attention and 3B active parameters.
Dense 27B vision-language model with linear attention for fast response times.
Qwen's second-strongest model — text capabilities exceeding Qwen3 235B, vision surpassing Qwen3 VL 235B.
Ultra-fast Qwen3.5 variant with 1M-token context at rock-bottom pricing.
Qwen3.5 Plus variant from February 2025 — optimized for balanced performance.
Qwen's largest MoE model with 397B total parameters, 17B active.
Flagship reasoning model for complex multi-step tasks.
Efficient coding agent with 80B parameters, only 3B active per token.
Open-source model with GPT-5 class reasoning performance.
32B dense vision-language model for high-quality multimodal tasks.
Compact vision-language model with always-on reasoning for visual tasks.
Efficient 30B MoE vision-language model with 3B active parameters.
Qwen's largest vision-language model with 235B MoE architecture.
Enhanced coding model with superior performance on complex engineering tasks.
Fast variant of Qwen3 Coder for speed-prioritized coding tasks.
Next-generation 80B MoE model with always-on reasoning mode.
Qwen Plus with always-on extended thinking for July 2025.
July 2025 Qwen3 30B MoE model with always-on extended thinking.
Efficient 30B MoE coding model with only 3B active parameters.
Qwen's primary code generation model for software development workflows.
July 2025 update to Qwen3 235B A22B with improved capabilities.
Efficient 30B MoE language model with 3B active parameters per token.
Compact Qwen3 model for fast, efficient inference.
Mid-size dense Qwen3 model balancing capability and efficiency.
Dense 32.8B parameter model optimized for reasoning and dialogue.
Qwen's flagship 235B MoE model with 22B active parameters.
Large 72B vision-language model from Qwen 2.5 generation.
Qwen's mid-tier commercial model for balanced performance.
Second-generation Qwen coding model at 32B scale.
Compact 7B model from Qwen 2.5 for efficient deployment.
Large 72B model from the Qwen 2.5 generation.
Qwen's dedicated reasoning model at 32B scale.
Qwen's top commercial model for the most demanding tasks.
Fast and cost-efficient Qwen model for high-throughput applications.
32B vision-language model from Qwen 2.5 with strong visual understanding.
Compact 7B code generation model from Qwen 2.5.
Earlier Qwen vision-language model with multimodal capabilities.
Top-tier earlier Qwen vision-language model.
Compact 7B vision-language model from Qwen 2.5.