# Artificial Analysis Speech-to-Speech Index: Voice Benchmark Profile

> A source-owned index for native audio models, with separate results for spoken reasoning, customer-service task completion, live preference, conversation flow, latency, and audio cost.

- Scope: Native audio-input and audio-output models with all four index components; the source table also lists partial-result models
- Measurement lanes: Spoken reasoning, Task completion, Conversation dynamics, Voice experience, Latency
- Primary metric: Equal-weighted source index (higher is better); component results stay separate
- Available evidence: Published source table and repeatable page snapshot
- Owner: Artificial Analysis
- Source snapshots refreshed: 2026-09-18

## Interpretation limit

The index combines Big Bench Audio reasoning, Artificial Analysis’s τ-Voice implementation, frozen Arena preference, and Arena task success at 25% each. Some τ-Voice rows use fewer than three trials, and live Arena Elo can differ from the frozen value used in the index. The source-owned voice result does not enter weighted text-model scores; audio prices here are source observations, not canonical provider pricing.

## Primary sources

- [Dataset](https://huggingface.co/datasets/ArtificialAnalysis/big_bench_audio)
- [Owner page](https://artificialanalysis.ai/speech-to-speech)

Canonical page: https://benchlm.ai/voice-benchmarks/artificial-analysis-speech-to-speech-index
