Generative AI & LLMs
Max tokens
A limit on how many tokens a model is allowed to generate in a response.
Example
A response stops after 300 generated tokens because the application set that maximum.
Why people use it
Understanding “Max tokens” helps teams choose and configure generative systems more effectively.
What you'll hear
“How does Max tokens affect the answer the model gives the user?”