Official provider results
GPT Audio 1.5 model documentation
OpenAI’s model page documents the modalities, limits, and pricing of gpt-audio-1.5; it publishes no benchmark scores for the model.
Source snapshots refreshed
Five measurement lanesVOICE / S2S
Reasoning, task completion, conversation dynamics, experience, and latency stay separate.
Benchmark profile
- Scope
- Modalities, context and output limits, endpoint support, and token pricing from the official model page
- Primary metric
- Documented capabilities (no scored benchmark)
- Owner
- OpenAI
- Available evidence
- Specifications only; OpenAI publishes no evaluation numbers for this model
What it measures
Voice experience
Is the exchange natural, robust, and responsive?
Latency
How long do responses, tool calls, and full tasks take?
Available results
- Modalities
- Audio + text in, audio + text out
- Context window
- 128,000 tokens
- Audio pricing
- $32 in / $64 out per 1M tokens
Documented
Chat Completions is the only supported endpoint.
Documented
Maximum output is 16,384 tokens.
Documented
Text tokens are $2.50 in and $10.00 out per million.
Interpretation limit
This entry records what OpenAI documents for gpt-audio-1.5 rather than a benchmark result. BenchLM assigns no score.