定義
A set of techniques—including quantization, distillation, pruning, and low-rank factorisation—that reduce model size and computational requirements while preserving performance. Model compression is essential for deploying powerful models on edge hardware or within cost budgets.
関連用語
AIの理解にお困りですか?
Physical AI 適合性コールをご予約いただき、これらのAI概念が貴社の業界や課題にどう適用されるかをご相談ください。