intel
★ 2.707neural-compressor
SOTA low-bit LLM quantization (INT8/FP8/MXFP8/INT4/MXFP4/NVFP4) & sparsity; leading model compression techniques on PyTorch, TensorFlow, and ONNX Runtime
HYSEN LABS VERZEICHNIS
Geprüfte Repositories, Klassifikationen und Analysen der Organisation intel.
SOTA low-bit LLM quantization (INT8/FP8/MXFP8/INT4/MXFP4/NVFP4) & sparsity; leading model compression techniques on PyTorch, TensorFlow, and ONNX Runtime
Intel AutoRound ist ein Modelloptimierungs-Toolkit für Quantisierungsworkflows, das die Inferenzkosten senkt und gleichzeitig die Genauigkeit für die KI-Bereitstellung beibehält.