Skip to navigation

Open ASR Leaderboard

Pulse Pro against the leaderboard top three, FLEURS English and throughput.
View as Markdown

Pulse Pro: Open ASR Leaderboard

Pulse Pro is tied for #2 on the public Open ASR Leaderboard at 5.42% average WER across eight ESB datasets. Whisper EnglishTextNormalizer applied, normalized WER. Lower is better.

Head-to-head vs leaderboard top-3

DatasetPulse ProGranite 4.1 2BCohere Transcribe
AMI (meetings)7.328.098.13
Earnings229.048.3710.86
GigaSpeech9.529.809.34
LibriSpeech clean1.731.331.25
LibriSpeech other3.742.502.37
SPGISpeech (financial)2.043.783.08
TED-LIUM3.683.072.49
VoxPopuli6.325.705.87
Average (8 datasets)5.425.335.42
Open ASR rank🥈 #2 (tied)🥇 #1🥈 #2 (tied)

Pulse Pro leads on conversational (AMI) and financial (SPGISpeech) workloads. Cohere edges ahead on read speech (LibriSpeech, TED-LIUM).

Position on the public leaderboard

Sorted by ESB average WER. Lower is better. Commercial APIs in our accuracy band:

RankModelESB Avg WER ↓
1IBM Granite Speech 4.1 2B5.33
2Pulse Pro5.42
2Cohere Labs Transcribe (tied)5.42
3Zoom Scribe v15.47
5NVIDIA Canary Qwen 2.5B5.63
8ElevenLabs Scribe v25.83
12AssemblyAI Universal-3 Pro6.21
18Speechmatics Enhanced6.91
23OpenAI Whisper Large v37.44

FLEURS English

MetricPulse Pro
WER (FLEURS en_us)3.92%
CER (FLEURS en_us)1.73%

Throughput

Measured on 1× NVIDIA L40S (48 GB), long-form audio.

ModeThroughput (RTFx)
No word timestamps250–300×
With word timestamps~200×

L4 is the recommended production GPU and runs at lower throughput than the L40S reference. See Cloud deployment for sizing.