Skip to main content
idapt
HomeCodeAI ModelsPricing
Sign inStart free
Qwen

Qwen3 8B

$AffordableRun locally
Qwen
CompareChat

Compact Qwen3 model for fast, efficient inference. Small yet capable with reasoning mode support. Ideal for edge deployment and high-volume, cost-sensitive applications.

▼

Qwen3 8B is somewhat expensive at $0.12/M tokens input and $0.45/M tokens output. The model supports extended reasoning, a 131K context window.

Best forCost-effective
ReplacesQwenQwen2.5 7B

Benchmarks

No benchmarks available

This model hasn't been evaluated by Artificial Analysis yet. Benchmark scores will appear once they're published.

Specifications

Vision
Audio
Reasoning
Dense
Architecture
8B
Parameters
131K
Context
8K
Max output
5.2 GB-8.7 GB
Local download
8.0 GB-12 GB
Local RAM

Performance

Providers

Run locally

Run Qwen3 8B on your own computer with the Local AI Engine for free, with no rate limits and no third-party AI provider involved. idapt installs it and routes chat to your hardware automatically.

Download
5.2 GB-8.7 GB
Recommended RAM
8.0 GB-12 GB

ollama pull qwen3:8b-q4_K_M

Misc

April 28, 2025
Released
Same family
QwenQwen3 14BQwenQwen3 235B A22BQwenQwen3 235B A22B (2507)QwenQwen3 30B A3BQwenQwen3 30B A3B Thinking (2507)QwenQwen3 32BQwenQwen3 Max ThinkingQwenQwen3 Next 80B A3B ThinkingQwenQwen3 VL 235B A22BQwenQwen3 VL 30B A3BQwenQwen3 VL 32BQwenQwen3 VL 8B Thinking
CompareChat with Qwen3 8B

Compare with

OpenAIGPT 5.6 Solvs Qwen3 8BOpenAIGPT 5.6 Lunavs Qwen3 8BAnthropicClaude Fable 5vs Qwen3 8BAnthropicClaude Opus 4.8vs Qwen3 8BAnthropicClaude Sonnet 5vs Qwen3 8BAnthropicClaude Haiku 4.5vs Qwen3 8BGoogleGemini 3.6 Flashvs Qwen3 8BGoogleGemini 3.5 Flashvs Qwen3 8BGoogleGemini 3.5 Flash-Litevs Qwen3 8BGoogleGemma 4 31Bvs Qwen3 8BxAIGrok 4.5vs Qwen3 8BxAIGrok Build 0.1vs Qwen3 8BxAIGrok 4.1 Fastvs Qwen3 8BMetaMuse Spark 1.1vs Qwen3 8BDeepSeekDeepSeek V4 Provs Qwen3 8BDeepSeekDeepSeek V4 Flashvs Qwen3 8BMiniMaxMiniMax M3vs Qwen3 8BMoonshotAIKimi K3vs Qwen3 8BMoonshotAIKimi K2.7 Codevs Qwen3 8BZhipu AIGLM 5.2vs Qwen3 8BQwenQwen3.7 Maxvs Qwen3 8BQwenQwen3.7 Plusvs Qwen3 8BQwenQwen3 Coder Nextvs Qwen3 8BNVIDIANemotron 3 Ultravs Qwen3 8BMistral AIMistral Small 4vs Qwen3 8B
Frequently Asked Questions about Qwen3 8B▼

When was Qwen3 8B released?

▼

Qwen3 8B was released on April 28, 2025.

Who created Qwen3 8B?

▼

Qwen3 8B was created by Qwen.

How much does Qwen3 8B cost?

▼

Qwen3 8B costs $0.12/M input tokens and $0.45/M output tokens.

What is Qwen3 8B API pricing?

▼

The Qwen3 8B API is priced at $0.12 per million input tokens and $0.45 per million output tokens. You can access Qwen3 8B through idapt.app alongside 200+ other AI models in one workspace.

Is Qwen3 8B a reasoning model?

▼

Yes, Qwen3 8B is a reasoning model. It can think step-by-step through complex problems before providing an answer, often yielding better results on difficult tasks like math, coding, and multi-step logic.

Does Qwen3 8B support image or vision input?

▼

No, Qwen3 8B does not currently support image or vision input. It is a text-only model.

Does Qwen3 8B support audio?

▼

No, Qwen3 8B does not support audio input.

What is the context window of Qwen3 8B?

▼

Qwen3 8B has a context window of 131K tokens, meaning it can process approximately 98,304 words of context at once.

Is Qwen3 8B open source?

▼

Qwen3 8B is a proprietary model developed by Qwen. The model weights and training data are not publicly available.

Where can I use Qwen3 8B?

▼

You can chat with Qwen3 8B on idapt.app — no separate API key required. idapt is an AI workspace with 200+ models, agents, computers, and tasks. You can also access Qwen3 8B through Qwen's own API or platform.

  • Home
  • Pricing
  • AI Models
  • Image models
  • Voice models
  • Video models
  • Rankings
  • New models
  • Model status
  • Multi-Model Chat
  • Voice Mode
  • Voice HUD
  • Web Search
  • Image Generation
  • Video Generation
  • Audio Generation
  • Transcription
  • Drive
  • Credentials
  • Sharing
  • Workspaces
  • Tasks
  • Memory
  • Agents
  • Subagents
  • Automations
  • Skills
  • idapt Code
  • Code Execution
  • Computers
  • Computer Use
  • Computer Assist · Soon
  • Containers · Soon
  • Cloud Computers
  • Local AI
  • AI Gateway
  • API & SDK
  • CLI
  • MCP
  • Tunnels
  • All features →
  • LLM cost calculator
  • Token counter
  • Context window checker
  • Can I run it
  • Model picker quiz
  • Savings finder
  • Video cost estimator
  • Text to speech cost
  • Transcription cost
  • API endpoint tester
  • All free tools →
  • Blog
  • Use cases
  • Comparisons
  • Best of
  • Skills
  • Learn
  • Changelog
  • Help center
  • FAQ
  • Privacy
  • Compare all models
  • Support
  • idapt Code
  • Developers
  • Quickstarts
  • API reference
  • API pricing
  • CLI
  • MCP
  • Downloads
  • Desktop
  • Badges and embeds
© idapt[email protected]TermsPrivacy PolicyLegal noticeReport content
X (Twitter)