Skip to content

Model Serving & Economics

Token cost

The price charged for the amount of text or tokens processed or generated by a model.

Example

A long prompt costs more because the provider charges for each processed token.

Why people use it

Understanding “Token cost” helps teams balance model quality, reliability, speed, and cost.

What you'll hear

“How does Token cost affect latency, quality, or cost at scale?”

Related terms