The gpt-audio model is OpenAI's first generally available audio model. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Audio is priced...
Specifications
| Provider | OpenAI |
|---|---|
| Context window | 128,000 tokens (128K) |
| Input price | $2.5 / 1M tokens |
| Output price | $10 / 1M tokens |
| Input modalities | text, audio |
| Output modalities | text, audio |
| Model ID | openai/gpt-audio |
Pricing is per 1M tokens, indicative pricing via OpenRouter (updated 2026-08-22); a provider's own list price may differ.
Frequently asked questions
How much does GPT Audio cost?
Indicative pricing (via OpenRouter, updated 2026-08-22) is $2.5 per 1M input tokens and $10 per 1M output tokens. A provider's own list price may differ. On Vincony, GPT Audio is available on credit-based pricing alongside 800+ other models.
What is GPT Audio's context window?
GPT Audio supports a context window of up to 128,000 tokens (128K), which is how much text it can consider at once.
Who makes GPT Audio?
GPT Audio is built by OpenAI. You can access it — and compare it against other providers' models — through Vincony's multi-model platform.
How do I use GPT Audio?
You can use GPT Audio directly through OpenAI, or via Vincony, which aggregates 800+ models (including GPT Audio) into one interface with side-by-side comparison and credit-based pricing.