跳至正文
返回术语表
部署

Triton Inference Server

定义

NVIDIA's open-source inference serving software that supports multiple frameworks (TensorRT, ONNX, PyTorch, TensorFlow) on GPU infrastructure. Triton is widely used in enterprise deployments requiring maximum throughput from GPU hardware.

了解术语只是第一步,将其落地应用才是第二步。

预约一次 Physical AI 适配性沟通,探讨这些 AI 概念如何转化到您所在的具体行业与业务挑战中。

Triton Inference Server | AI 术语表 — 解析 LLM、RAG 及 20+ 关键术语