O

GPT Audio

Audioby OpenAI·Model page

OpenAI's audio-native model for real-time speech input and output with a 128K-token context.

Max output

Most tokens the model can return in a single response.

16Ktokens
Share:

Pricing

per 1M tokens
Input$2.5 /1M
Output$10 /1M
Audio$32 /1M

Model Description

The gpt-audio model is OpenAI's first generally available audio model. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Audio is priced...

Author
O
OpenAI
Organization · ✓
openai
Details
Downloads
Likes
AccessClosed Source
Context128K tokens
Input price$2.5 /1M
Output price$10 /1M
CreatedJan 19, 2026
Updated
View on Hugging Face
Get the full context.

Sign up to read complete case studies, access detailed metrics, and unlock all use cases.

GPT Audio — AI Model Details | Applied