OpenAI: GPT Audio
0.0 / 10
128K context
Proprietary
About this model
The gpt-audio model is OpenAI's first generally available audio model. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Audio is priced...
Technical Specs
- Architecture: text+audio->text+audio
- Context Window: 128,000 tokens
- Input Modalities: text
Hardware Requirements
- API-only (no local hardware needed)
Pricing
| Input | Output | Currency |
|---|---|---|
| 2.5 / 1M tokens | 10.0 / 1M tokens | USD |