YFarmX

Inter-token latency

AI

Inter-token latency: The average gap between successive output tokens during generation; low values give smooth, fast-flowing streamed text.

Related terms

Browse the full glossary →