QamoosTech
AI & DataBeginner

Inference

PronunciationIN-fuh-runss

Definition

Inference is the process of using a trained machine learning model to make predictions or generate outputs based on new, unseen input data. It is the stage where the model is "live" and performing its intended task rather than being trained.

Where you hear it

In discussions about model deployment, performance optimization, and API usage for AI services.

Examples

  • The model performs inference in milliseconds when a user sends a prompt.
  • We need to optimize our infrastructure to handle high-volume inference requests.

Common mistake

Confusing inference with training; training is the process of teaching the model, while inference is the process of using the model to get results.