跳至正文
返回术语表
部署

Inference Cost

定义

The compute and financial cost of running a model to produce a single prediction or generated response. Inference cost is often the dominant AI operational expenditure at scale and is managed through model compression, caching, quantization, and batching strategies.

了解术语只是第一步,将其落地应用才是第二步。

预约一次 Physical AI 适配性沟通,探讨这些 AI 概念如何转化到您所在的具体行业与业务挑战中。

Inference Cost | AI 术语表 — 解析 LLM、RAG 及 20+ 关键术语