# Parakeet TDT 0.6B v3 Benchmark Scores & Performance

> Parakeet TDT 0.6B v3 has Fermion speech evaluation results. Voice evidence is display-only and does not produce a text-model score.

Canonical page: https://benchlm.ai/models/parakeet-tdt-0-6b-v3

Last updated: September 30, 2026

## Model Details

| Property | Value |
|----------|-------|
| Creator | NVIDIA |
| Source Type | Open Weight |
| Reasoning Type | Non-Reasoning |
| Context Window | N/A |
| Official model card | [NVIDIA model documentation](https://huggingface.co/nvidia/parakeet-tdt-0.6b-v3) |
| Overall Score | Not scored (voice evidence only) |
| Overall Rank | Unranked |

## Family & Coverage

- Family: Parakeet TDT 0.6B v3
- Variant: base
- Benchmarks covered: 0 of 486
- Coverage note: Speech and voice results appear below; no weighted text-model benchmark rows are stored.

## Speech and voice evidence

[Phonon-2 launch evaluation](/voice-benchmarks/phonon-2-launch-evaluation)

Fermion ran Phonon-2, Phonon-1, and Parakeet Redux with the Open ASR Leaderboard code on full test sets. The other accuracy rows are leaderboard figures quoted by Fermion, without independent re-verification here. Some model cards report different runs; those values are not mixed into this comparison. Voxtral's download size is an estimate at 16 bits. Throughput includes log-mel processing and excludes model load; batch-128 GPU results are separate from single-stream results. The Mac runtime comparison uses 20 dictations totaling 797 seconds at each runtime's defaults. Download size does not measure runtime memory. These results do not enter text-model or voice-agent rankings.

| Metric | Result |
| --- | --- |
| Seven-set mean WER | 4.96% |
| LibriSpeech clean WER | 1.52% |
| LibriSpeech other WER | 3.13% |
| AMI WER | 9.42% |
| Earnings-22 WER | 5.85% |
| GigaSpeech WER | 7.99% |
| SPGISpeech WER | 3.63% |
| VoxPopuli WER | 3.19% |
| Download | 2,508 MB |
| FluidAudio, Parakeet TDT 0.6B v3 (Core ML) | 104.9x realtime on M5 MacBook Air |
| sherpa-onnx, Parakeet TDT 0.6B v3 (int8) | 16.5x realtime on M5 MacBook Air |

Leaderboard row quoted by Fermion; Fermion launch table retrieved September 30, 2026. Lower WER is better; throughput excludes model load.

- [Evaluation source](https://www.fermionresearch.com/research/phonon-2/)

Open multilingual speech recognition weights under CC-BY-4.0, with NeMo and Transformers inference documented on the official model card. This is the full-precision teacher of Phonon-2.

## Other NVIDIA Models

- [Nemotron 3 Ultra](/models/nemotron-3-ultra) - Score: 43.01
- [Nemotron 3 Nano Omni 30B A3B](/models/nemotron-3-nano-omni-30b-a3b) - Score: 30.71
- [Nemotron 3.5 Lightning 30B A3B NVFP4](/models/nemotron-3-5-lightning-30b-a3b-nvfp4) - Score: 18.94
- [Nemotron-4 15B](/models/nemotron-4-15b) - Score: 11.45
- [Nemotron 3 Nano 30B](/models/nemotron-3-nano-30b) - Score: not computed
- [Nemotron 3 Super 120B A12B](/models/nemotron-3-super-120b-a12b) - Score: not computed
- [Nemotron 3 Super 100B](/models/nemotron-3-super-100b) - Score: not computed
- [Nemotron Ultra 253B](/models/nemotron-ultra-253b) - Score: not computed
- [Cosmos3-Edge](/models/cosmos3-edge) - Score: not computed
- [Audio Flamingo 3 7B](/models/audio-flamingo-3-7b) - Score: not computed
- [Canary 180M Flash](/models/canary-180m-flash) - Score: not computed
- [Nemotron 3.5 ASR Streaming 0.6B](/models/nemotron-3-5-asr-streaming-0-6b) - Score: not computed
- [Parakeet CTC 1.1B](/models/parakeet-ctc-1-1b) - Score: not computed
