Ga naar inhoud
Terug naar Woordenlijst
Deployment

Triton Inference Server

Definitie

NVIDIA's open-source inference serving software that supports multiple frameworks (TensorRT, ONNX, PyTorch, TensorFlow) on GPU infrastructure. Triton is widely used in enterprise deployments requiring maximum throughput from GPU hardware.

Hulp Nodig bij het Begrijpen van AI?

Boek een Physical AI-kennismaking om te bespreken hoe deze AI-concepten op uw branche en uitdagingen van toepassing zijn.

Triton Inference Server | AI-Woordenlijst — LLM, RAG, Fine-Tuning & Kernbegrippen