Latency

AI Observability

The time delay between when a request is sent to an AI system and when a response is received. Low latency is crucial for real-time applications, while acceptable thresholds vary by use case (interactive chat vs. batch processing).

← Back to glossary