Separate from TTFT (LLM). A fast model does not guarantee low TTFA — turn detection, TTS warmup, and transport all sit on the path.
Even “fast” end-to-end stacks often land around ~1–2s median. Budget UX accordingly.
Connections
No outgoing connections.