GPT Audioon Krater.ai

The gpt-audio model is OpenAI's first generally available audio model. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Audio is priced

Released January 19, 2026

Specifications

GPT Audio at a glance

Context

128K tokens

Max output

16.4K tokens

Input

Text, Audio

Output

Text, Audio

Input price

$2.50 / 1M

Output price

$10.00 / 1M

Tool callingStructured outputsProvider moderated

Is ChatGPT down?

ChatGPT not loading? Krater checks OpenAI every 5 minutes and shows what to switch to while it is down.

Live ChatGPT status

Capabilities

What it is good at

Catalog capabilities

  • Generates text responses
  • Supports tool calling
  • Supports structured outputs

Good for

Long documentsUse with toolsStructured responses

FAQ

Questions about GPT Audio

Clear answers based on the model catalog and Krater access.

GPT Audio supports a context window of 128,000 tokens.

Ready to use GPT Audio?

Access leading models in one focused workspace.