Ling 3.0 Flash VL
Newinclusionai/ling-3.0-flash-vl-20260910
inclusionAI: Ling 3.0 Flash VL is a model that processes text, image, and video inputs to generate text outputs. It has a context length of 131,072 tokens, supports tool use and reasoning, and is available as open weights. It was released on September 9, 2026.
Tracked, but not yet independently ranked
This model is in the catalog, but we do not yet have enough public benchmark data to place it confidently in quality-focused recommendations. It may still appear in price-based recommendations.
- tool use
- structured outputs
- reasoning
- open-weight
At a glance
Specifications
What the API gives you
- Context window
- 131K tokens
- Max output
- 33K tokens
- Input price
- $0.060 / 1M tokens
- Output price
- $0.18 / 1M tokens
- Cache read
- $0.012 / 1M tokens
- Modalities
- text, image, video → text
- Knowledge cutoff
- —
- Released
- Sep 9, 2026
Benchmarks
No independent benchmark results are linked to this model yet. We keep it visible while withholding quality claims until usable results arrive.
Open-weight model on Hugging Face.
Task evidence
How broad is the support?
Each status is exact to this model version and the selected task. Official release claims remain provisional.
No exact-version task evidence is linked yet. Official release claims will appear provisionally; independent results are required for established status.
Recommendation coverage
Where this model appears
Not currently included in a published recommendation. The task evidence panel above says exactly what is confirmed and what is still missing.
Catalog data via OpenRouter and models.dev. Last checked within the past 6 hours.