Temas / perceived latency thresholds

Punto de vista parafraseado

600–800 ms perceived willingness drop

Perceived willingness to continue the interaction begins to decline after 600 ms and drops significantly between 700 and 800 ms.

Detrás de la opinión

Las traducciones son para facilitar la lectura; los extractos originales siguen siendo la source evidence.

Real-Time Voice AI Stack for Agents: Architecture Guide

Extracto original

Perceived willingness begins to drop after 600 ms and steps down significantly from 700 to 800 ms.
Contexto

Transport, recognition, endpointing, the LLM's first token, TTS first byte, and the return network hop each take a slice of that budget. This guide breaks down what each layer costs and where you can cut to keep your voice AI stack inside it.