Separate from TTFT (LLM). A fast model does not guarantee low TTFA — turn detection, TTS warmup, and transport all sit on the path.

Even “fast” end-to-end stacks often land around ~1–2s median. Budget UX accordingly.

Connections

No outgoing connections.