FLEURS
FLEURS Streaming - English
WER on the English subset of FLEURS across providers in streaming mode. Lower is better.
A note on audio amplitude normalization
Audio amplitude normalization materially changes WER on FLEURS. Most competitors benchmark on raw FLEURS - which has variable, often low amplitude - without normalizing peak audio to −10 dBFS. This makes some models look much better than they actually are. Pulse is stable across all amplitude regimes.
Pre-recorded - FLEURS
Google’s multilingual speech dataset covering 102 languages, built on the FLoRes-101 translation benchmark. Contains ~12 hours of read speech per language and is the standard benchmark for evaluating multilingual ASR, including low-resource languages.
Evaluated on the FLEURS dataset (non-streaming / batch mode).
Sources: Deepgram internal benchmarks; Smallest AI internal evaluation.
Streaming - FLEURS
Evaluated on the FLEURS dataset (streaming mode).
Sources: Deepgram internal benchmarks; Smallest AI internal evaluation.