← All models

openai

OpenAI: GPT Audio

The gpt-audio model is OpenAI's first generally available audio model. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Audio is priced at $32 per million input tokens and $64 per million output tokens.

128,000 context
Modalities:text, audio->text, audio
Released:1/19/2026

The gpt-audio model is OpenAI's first generally available audio model. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Audio is priced at $32 per million input tokens and $64 per million output tokens.

Weekly tokens

23.9M

Tokens generated this week (network-wide)

Usage by period

Today5.3M tokens
This week37.4M tokens
This month138.2M tokens
Trending37.4M tokens