コンテンツへスキップ
用語集に戻る
デプロイメント

Inference Cost

定義

The compute and financial cost of running a model to produce a single prediction or generated response. Inference cost is often the dominant AI operational expenditure at scale and is managed through model compression, caching, quantization, and batching strategies.

関連サービス

AIの理解にお困りですか?

Physical AI 適合性コールをご予約いただき、これらのAI概念が貴社の業界や課題にどう適用されるかをご相談ください。

Inference Cost | AI用語集 — LLM、RAG、ファインチューニング&主要概念