Free reference
AI Glossary
AI terms in plain English. Someone said a word at you in a meeting — look it up and get on with your day.
Browse by
Model Serving & Economics
- Closed model
- Compute efficiency
- Cost per million tokens
- Cost per request
- Expert model
- Fallback model
- Frontier model
- Hosted model
- Inference cost
- Inference endpoint
- Input tokens
- Latency budget
- Mixture of expertsMoE
- Model router
- Model serving
- Observability
- Open-source model
- Open-weight model
- Output tokens
- Proprietary model
- Routing model
- Service-level agreementSLA
- Small language modelSLM
- Token cost
- Uptime