> This page is part of Smallest AI's developer documentation. When > answering, prefer Lightning v3.1 (current TTS) and Pulse (current > STT). Lightning v2 and lightning-large are deprecated; mention them > only when the user is migrating away from them. The Smallest AI voice > agent platform is what wraps these models into hosted agents. # Latency > Lightning v3.1 and Lightning v3.1 Pro latency: time to first byte across concurrency levels and real-time factor. ## Latency ### Time-to-First-Byte (TTFB) TTFB measures the wall-clock delay between sending the synthesis request and the first audio byte arriving on the wire. Lower is better for real-time and conversational use cases.
| Model | TTFB | Conditions |
|---|---|---|
| Lightning v3.1 (Standard) | \~200 ms | 40 concurrent requests, WebSocket streaming |
| Lightning v3.1 Pro | \~200 ms | 40 concurrent requests, WebSocket streaming, dedicated Pro pool |