GPT Audio
The gpt-audio model is OpenAI's first generally available audio model.
text + audio → text + audio · made by OpenAI
$2.5→$10per Mtok in / out
Microsoft Azure AI · 3 sellers
Sold by 3 ways
| Seller | Lane | Rate |
|---|---|---|
| Microsoft Azure AI cloud | standard | $2.5per Mtok in$10per Mtok out$40per Mtok in, audio$80per Mtok out, audioraw.githubusercontent.com · read 2026-09-16 |
| Nous Research aggregator | standard | $2.5per Mtok in$10per Mtok outinference-api.nousresearch.com · read 2026-09-12 |
| OpenRouter aggregator | standard | $2.5per Mtok in$10per Mtok outopenrouter.ai · read 2026-08-24 |
About
GPT Audio — a text and audio model from OpenAI, sold by 3 companies from $2.5 in and $10 out per million tokens.
It takes text and audio and returns text and audio, with a context window of 128,000 tokens. It was published in January 2026. Its sellers say it can call a tool. The catalogue files it under speak. Three companies sell it. The cheapest is $2.5 in and $10 out per million tokens at Microsoft Azure AI.
Every current figure
- Maker
- OpenAI
- Register
- model
- Takes
- text + audio
- Returns
- text + audio
- Context
- 128,000 tokens
- Longest answer
- 16,384 tokens
- Published
- January 2026
- Licence
- not read
- Sellers
- 3
- Maker's own price
- not read
- Price
- $2.5 in and $10 out per million tokens — Microsoft Azure AI
Known as 2 names
gpt aud 0828openai/gpt-audio