Benchmarks
Every benchmark we publish, by model, one page each.
Numbers are measured in-region on production endpoints. Each page states the dataset, the method, and the models compared. For definitions of the metrics, see the metrics pages linked under each model.
Lightning (text to speech)
Lightning v3.1 and Lightning v3.1 Pro. Overview and metric definitions.
Pulse (speech to text)
Pulse and Pulse Pro. Overview, metric definitions, and how to run the evaluation yourself.