Gemini 3.1 Flash Lite
google/gemini-3.1-flash-lite-20260507
Google's Gemini 3.1 Flash Lite is a multimodal model that accepts text, image, video, file, and audio inputs and generates text output. It provides a context length of 1,048,576 tokens, supports tool usage and reasoning capabilities, and is priced at $0.25 per million input tokens and $1.5 per million output tokens. The model is not open weights and was released on May 7 2026.
- tool use
- structured outputs
- reasoning
At a glance
Specifications
What the API gives you
- Context window
- 1.0M tokens
- Max output
- 66K tokens
- Input price
- $0.25 / 1M tokens
- Output price
- $1.50 / 1M tokens
- Cache read
- $0.025 / 1M tokens
- Modalities
- text, image, video, file, audio → text
- Knowledge cutoff
- —
- Released
- May 7, 2026
Recommendation coverage
Where this model appears
Not currently included in a published recommendation. This means the available task-specific evidence did not support a ranking yet—not that the model was reviewed and rejected.
Benchmarks
No independent benchmark results are linked to this model yet. We keep it visible while withholding quality claims until usable results arrive.
Catalog data via OpenRouter and models.dev. Last checked within the past 6 hours.