コンテンツへスキップ
用語集に戻る
デプロイメント

Inference

定義

The process of running a trained model on new data to produce predictions or generated outputs. Inference cost and latency are the dominant operational concerns in production AI, particularly for large generative models that can cost cents per request at scale.

AIの理解にお困りですか?

Physical AI 適合性コールをご予約いただき、これらのAI概念が貴社の業界や課題にどう適用されるかをご相談ください。

Inference | AI用語集 — LLM、RAG、ファインチューニング&主要概念