Ling 3.0 Flash VL
StandardInclusionAI's hybrid vision-language model, layering native visual perception on the Ling 3.0 Flash MoE (124B total, 5.5B active).
131K-token context with image and video input, a hybrid instant/reasoning mode, and tool calling. Handles visual agent work: document and screen understanding, grounded tool use, and multimodal automation.
▼
Ling 3.0 Flash VL is somewhat expensive at $0.02/M tokens input and $0.06/M tokens output. Cached input reads are priced at $0.004/M tokens, making repeated context cheaper. The model supports vision/image input, extended reasoning, a 262K context window.
Benchmarks
Specifications
Performance
Providers
Misc
Compare with
Frequently Asked Questions about Ling 3.0 Flash VL▼
When was Ling 3.0 Flash VL released?
▼
Ling 3.0 Flash VL was released on September 10, 2026.
Who created Ling 3.0 Flash VL?
▼
Ling 3.0 Flash VL was created by InclusionAI.
How much does Ling 3.0 Flash VL cost?
▼
Ling 3.0 Flash VL costs $0.02/M input tokens and $0.06/M output tokens.
What is Ling 3.0 Flash VL API pricing?
▼
The Ling 3.0 Flash VL API is priced at $0.02 per million input tokens and $0.06 per million output tokens. You can access Ling 3.0 Flash VL through idapt.app alongside 200+ other AI models in one workspace.
Is Ling 3.0 Flash VL a reasoning model?
▼
Yes, Ling 3.0 Flash VL is a reasoning model. It can think step-by-step through complex problems before providing an answer, often yielding better results on difficult tasks like math, coding, and multi-step logic.
Does Ling 3.0 Flash VL support image or vision input?
▼
Yes, Ling 3.0 Flash VL supports vision/image input. You can share images in your conversation and the model will analyze and respond based on their content.
Does Ling 3.0 Flash VL support audio?
▼
No, Ling 3.0 Flash VL does not support audio input.
What is the context window of Ling 3.0 Flash VL?
▼
Ling 3.0 Flash VL has a context window of 262K tokens, meaning it can process approximately 196,608 words of context at once.
Is Ling 3.0 Flash VL open source?
▼
Ling 3.0 Flash VL is a proprietary model developed by InclusionAI. The model weights and training data are not publicly available.
What is Ling 3.0 Flash VL good at?
▼
Based on benchmark data, Ling 3.0 Flash VL excels at: visual analysis and image understanding. It is InclusionAI's capable model suitable for a wide range of tasks.
Where can I use Ling 3.0 Flash VL?
▼
You can chat with Ling 3.0 Flash VL on idapt.app — no separate API key required. idapt is an AI workspace with 200+ models, agents, computers, and tasks. You can also access Ling 3.0 Flash VL through InclusionAI's own API or platform.