Skip to navigation

WildASR robustness

Pulse accuracy on noisy, accented, real-world audio.
View as Markdown

ASR Robustness - WildASR Dataset (Streaming)

An open-source robustness benchmark designed to stress-test STT under real-world degraded conditions: clipping, far-field capture, background noise, phone codec compression, reverberation, and accented speech. Lower WER is better. n/a = not supported by that provider.

Evaluated on the open-source WildASR dataset. Numbers from internal evaluation.

DatasetSmallest PulseAssembly Universal 3 ProAWS TranscribeAzureDeepgram Nova 3Sarvam Saras V3ElevenLabs Scribe V2
Clean5.313.337.0111.1111.627.024.24
Clipping10.316.5942.104.3547.3528.7411.20
Far-field8.9926.0738.76n/a62.9921.277.38
Noise Gap6.914.049.77n/a15.049.746.30
Phone Codec6.643.458.70n/a9.1310.644.98
Reverberation8.0623.5014.83n/a27.274.356.48
Accent5.742.804.45n/a7.31n/a4.01
Aggregate7.6012.5218.358.8228.1717.756.47