AI GlossaryㄹTechnical words in the news
Latency
The delay between asking and getting an answer. For voice conversations to feel natural, this has to come down to human reaction speed.
In plain words
Latency is the delay between sending a request and getting a response started. It's a number that shapes how good an AI product feels to use — voice conversations in particular only feel natural when latency approaches human conversational reaction time (0.2-0.3 seconds).
"Smarter" and "faster" pull in opposite directions — reasoning models get smarter by thinking longer, but that makes them slower. That's why companies release fast, small models alongside deep, large ones, and new chip announcements usually boast about how much they've cut latency.
See also
Stories using this term
- LTX achieves 'stutter-free avatars' with frame-by-frame real-time streamingAI · 2026.08.15
- Claude Desktop cuts background boot time by up to 2.3xAI · 2026.08.19
- Claude Code reveals six ways to steer hours-long tasksAI · 2026.09.04
- Claude web and desktop now stream long answers 4x smootherAI · 2026.08.25
- Claude Code makes "auto mode" the default, running without approvalAI · 2026.08.11
- Prime Intellect unveils self-improving agent harness 'Prime Agent'AI · 2026.08.09
