AI & DataBeginner
Inference
PronunciationIN-fuh-runss
Definition
Inference is the process of using a trained machine learning model to make predictions or generate outputs based on new, unseen input data. It is the stage where the model is "live" and performing its intended task rather than being trained.
Where you hear it
In discussions about model deployment, performance optimization, and API usage for AI services.
Examples
The model performs inference in milliseconds when a user sends a prompt.
We need to optimize our infrastructure to handle high-volume inference requests.
Common mistake
Confusing inference with training; training is the process of teaching the model, while inference is the process of using the model to get results.