intel
★ 2,707neural-compressor
SOTA low-bit LLM quantization (INT8/FP8/MXFP8/INT4/MXFP4/NVFP4) & sparsity; leading model compression techniques on PyTorch, TensorFlow, and ONNX Runtime
HYSEN LABS ディレクトリ
intel組織の検証済みリポジトリ、分類、分析。
SOTA low-bit LLM quantization (INT8/FP8/MXFP8/INT4/MXFP4/NVFP4) & sparsity; leading model compression techniques on PyTorch, TensorFlow, and ONNX Runtime
Intel AutoRound は、AI 導入の精度を維持しながら推論コストを削減する、量子化ワークフロー用のモデル最適化ツールキットです。