Ling 3.0 Flash VL

New

inclusionai/ling-3.0-flash-vl-20260910

inclusionAI: Ling 3.0 Flash VL is a model that processes text, image, and video inputs to generate text outputs. It has a context length of 131,072 tokens, supports tool use and reasoning, and is available as open weights. It was released on September 9, 2026.

Tracked, but not yet independently ranked

This model is in the catalog, but we do not yet have enough public benchmark data to place it confidently in quality-focused recommendations. It may still appear in price-based recommendations.

  • tool use
  • structured outputs
  • reasoning
  • open-weight

At a glance

providerinclusionai
releasedSep 9, 2026
context131K
input / M$0.060

Specifications

What the API gives you

Context window
131K tokens
Max output
33K tokens
Input price
$0.060 / 1M tokens
Output price
$0.18 / 1M tokens
Cache read
$0.012 / 1M tokens
Modalities
text, image, video → text
Knowledge cutoff
Released
Sep 9, 2026

Benchmarks

No independent benchmark results are linked to this model yet. We keep it visible while withholding quality claims until usable results arrive.

Open-weight model on Hugging Face.

Task evidence

How broad is the support?

Each status is exact to this model version and the selected task. Official release claims remain provisional.

No exact-version task evidence is linked yet. Official release claims will appear provisionally; independent results are required for established status.

Recommendation coverage

Where this model appears

Not currently included in a published recommendation. The task evidence panel above says exactly what is confirmed and what is still missing.

Catalog data via OpenRouter and models.dev. Last checked within the past 6 hours.