Inference is the phase in which a trained and finalized model processes new input and, based on that input, generates an output—such as a response to a prompt—unlike training, during which the model is still learning from the data. The speed and cost of inference are key to the real-world commercial deployment of AI products, since inference occurs every time an end user uses the model.
Glossary
Inference
Inference
Similar articles
View all articles
Academy Article
13. 9. 2026
2026 Retraining Grant: How to Get a Free Course Through the Employment Office (Complete Guide)
14 min read
Academy Article
13. 9. 2026
Where Companies Will Save the Most with AI Automation in 2026 (and Where They’re Wasting Money)
8 min read
Academy Article
13. 9. 2026