Best speech-to-text models

Transcription models by word error rate, lowest first.

ranked by word error rate · 57 models · updated Aug 30 · 51–57

#ModelWord error rate
51Gradium Speech-to-Text
0.10
52Nova-3, Deepgram
0.10
53Qwen3.5 Omni Flash
0.10
54Chirp 2, Google
0.10
55Rev AI
0.10
56Qwen3 ASR Flash, Alibaba
0.10
57Cloud Speech-To-Text (Chirp), Google
0.30

what word error rate means

The share of words a model gets wrong when transcribing. Lower is better, so this board ranks upward from zero.