YFarmX

Latency

AI

Latency: The delay between sending a request and receiving the result; in language models it covers the wait for the first token and for the full response.

Used in these stories

Related terms

Browse the full glossary →