GPT Audioon Krater.ai
The gpt-audio model is OpenAI's first generally available audio model. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Audio is priced
Released January 19, 2026
Specifications
GPT Audio at a glance
Context
128K tokens
Max output
16.4K tokens
Input
Text, Audio
Output
Text, Audio
Input price
$2.50 / 1M
Output price
$10.00 / 1M
Capabilities
What it is good at
Catalog capabilities
- Generates text responses
- Supports tool calling
- Supports structured outputs
Good for
Compare
Compare GPT Audio
GPT Audio vs GPT Audio Mini
A cost-efficient version of GPT Audio. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency.
Open comparisonGPT Audio vs GPT-5 Pro
GPT-5 Pro is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience.
Open comparisonGPT Audio vs GPT-3.5 Turbo
GPT-3.5 Turbo is OpenAI's fastest model. It can understand and generate natural language or code, and is optimized for chat and traditional completion tasks.
Open comparisonFAQ
Questions about GPT Audio
Clear answers based on the model catalog and Krater access.
Use GPT Audio in your API
Connect with Krater's unified API.
