GPT Audio Mini

openai/gpt-audio-mini

OpenAI's GPT Audio Mini, released on 2026-01-20, processes both text and audio inputs and produces text and audio outputs. It supports tool usage, does not support reasoning, and has a context length of 128 000 tokens. Pricing is $0.6 per million input tokens and $2.4 per million output tokens; the model’s weights are not open.

  • tool use
  • structured outputs

At a glance

provideropenai
releasedJan 20, 2026
context128K
input / M$0.60

Specifications

What the API gives you

Context window
128K tokens
Max output
16K tokens
Input price
$0.60 / 1M tokens
Output price
$2.40 / 1M tokens
Cache read
Modalities
text, audio → text, audio
Knowledge cutoff
Released
Jan 20, 2026

Recommendation coverage

Where this model appears

Not currently included in a published recommendation. This means the available task-specific evidence did not support a ranking yet—not that the model was reviewed and rejected.

Benchmarks

No independent benchmark results are linked to this model yet. We keep it visible while withholding quality claims until usable results arrive.

Catalog data via OpenRouter and models.dev. Last checked within the past 6 hours.