HYSEN LABS 项目目录

NVIDIA

来自 NVIDIA 组织的可信仓库、分类与深度解析。

28 个精选开源项目
14,627

TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way.

Python2,745 个 Fork
NVIDIA

TensorRT

13,348

NVIDIA® TensorRT™ is an SDK for high-performance deep learning inference on NVIDIA GPUs. This repository contains the open source components of TensorRT.

C++2,407 个 Fork
NVIDIA

cutlass

10,437

CUDA Templates and Python DSLs for High-Performance Linear Algebra

C++2,084 个 Fork
NVIDIA

garak

9,258

the LLM vulnerability scanner

Python1,282 个 Fork
NVIDIA

OpenShell

8,615

OpenShell is the safe, private runtime for autonomous AI agents.

Rust1,252 个 Fork
NVIDIA

DALI

5,759

A GPU-accelerated library containing highly optimized building blocks and an execution engine for data processing to accelerate deep learning training and inference applications.

C++676 个 Fork
NVIDIA

cuml

5,277

NVIDIA cuML: GPU-Accelerated Machine Learning

Python674 个 Fork
NVIDIA

nccl

5,092

Optimized primitives for collective multi-GPU communication

C++1,415 个 Fork
4,181

Generative AI reference workflows optimized for accelerated infrastructure and microservice architecture.

Jupyter Notebook1,098 个 Fork
3,810

量化、蒸馏、剪枝、神经架构搜索、推测解码等 SOTA 模型优化技术的统一库。它为 TensorRT-LLM、TensorRT、vLLM 等下游部署框架压缩深度学习模型,以优化推理速度。

PythonAI 与机器学习599 个 Fork
3,533

A library for accelerating Transformer models on NVIDIA GPUs, including using 8-bit and 4-bit floating point (FP8 and FP4) precision on Hopper, Ada and Blackwell GPUs, to provide better performance with lower memory utilization in both training and inference.

Python823 个 Fork
NVIDIA

skills

3,301

项目速览:适用于 NVIDIA 产品的代理技能,安装到 Claude Code、Codex 和其他编码代理中,以端到端运行物理 AI、机器人、模拟、CUDA 和 RAG 工作流程。技能目录产品|描述 |技能 |爱智商 | NVIDIA AI-Q 蓝图 - 部署本地 AI-Q 服务并作为代理技能运行浅层或深层研究工作流程。

Python394 个 Fork
3,253

Open-source deep-learning framework for building, training, and fine-tuning deep learning models using state-of-the-art Physics-ML methods

Python783 个 Fork
2,632

The NVIDIA NeMo Agent toolkit is an open-source library for efficiently connecting and optimizing teams of AI agents.

Python755 个 Fork
NVIDIA

kvpress

1,209

LLM KV cache compression made easy

Python179 个 Fork
NVIDIA

raft

1,042

RAFT contains fundamental widely-used algorithms and primitives for machine learning and information retrieval. The algorithms are CUDA-accelerated and form building blocks for more easily writing high performance applications.

Cuda250 个 Fork