The delay between sending a request to an AI model and receiving a response — lower latency feels faster and more responsive.
Larger, more capable models are often slower (higher latency) than smaller ones, which is part of why companies offer both 'fast/lite' and 'powerful' model tiers.
One of 60 free AI glossary terms
Plain-language definitions for the AI jargon you'll actually run into — no email needed, ever, for this section.
Browse the Full Glossary →