Home / AI glossary / What is inference in AI?

What is inference in AI?

In short

Inference is the stage where a trained AI model is put to work, taking new input and producing an output such as a prediction, classification, or generated text. It is distinct from training, which is the learning phase. Every time you ask a chatbot a question or run a photo through a recognition model, you are running inference. Its speed and cost are key concerns when deploying AI.

Go deeper with Crux Digits

Want this applied in your business? See how we take it to production:

← All AI terms

From concept to working tool?

We build this AI in production — at fixed prices, with one named expert. Start with a free consultation.

Book a free consultation →