AI product & model operations · Core

Inference

Also called: model inference

ELI5

The trained model produces an answer for new data.

Used in conversation

“Inference latency must stay below two seconds.”

Definition

Use of a trained model to generate a prediction or output from new input.

Here “Inference” means: Use of a trained model to generate a prediction or output from new input.

Pitch context

You will see “Inference” in product strategy decks, roadmaps, experiments and architecture reviews when the discussion reaches ai product & model operations.

Why it matters: In ai product & model operations, the scope can change what data may be used, who may access it and which controls are required.

Sources & evidence · 2

Direct term-level sources and supporting source families.