Research

Benchmarks

Evaluations of frontier speech models on two-person conversation, measured against the public audio those models are usually ranked on.

Work with us

Word error rate · lower is better

Show 10 more
Audio

Converse-STT

Conversational Speech-to-Text Benchmark

In collaboration with Cekura, we scored 15 speech-to-text models on Ocular AI Real World Conversational Data and on public Pipecat audio. Twelve had a higher error rate on the conversational recordings, and the leading model changed with the dataset.

Category

Speech-to-text

Languages

English
In collaboration withCekura
Ready to bring AI into the real world?