Zum Inhalt springen
Zurück zum Glossar
Bereitstellung

Inference

Definition

The process of running a trained model on new data to produce predictions or generated outputs. Inference cost and latency are the dominant operational concerns in production AI, particularly for large generative models that can cost cents per request at scale.

Hilfe beim Verständnis von KI Benötigt?

Buchen Sie ein Physical-AI-Eignungsgespräch, um zu besprechen, wie diese KI-Konzepte auf Ihre Branche und Ihre Herausforderungen anwendbar sind.

Inference | KI-Glossar — LLM, RAG, Fine-Tuning & Kernbegriffe