Skip to content

APIs & Integration

Latency

The time between making a request and receiving a response or result.

Example

The team measures 800 milliseconds from sending a prompt to receiving the first output.

Why people use it

Knowing how “Latency” works helps teams connect systems reliably and troubleshoot integrations.

What you'll hear

“How are we using Latency to connect the model to the rest of the product?”

Related terms