AssemblyAI Universal 3.5 Pro shows 1.93% word error rate in a 15-model test
Tarush Agarwal reported a speech-to-text benchmark of 15 models on the same audio and pipeline. On 1,000 read-aloud clips from Pipecat FLEURS, AssemblyAI Universal 3.5 Pro had a 1.93% word error rate and a 489 ms median time to first text, both the best results in that test.