HYSEN LABS DIRECTORY

NVIDIA

Verified repositories, classifications and analysis from the NVIDIA organization.

28 curated open-source projects
NVIDIA

NemoClaw

22,463

Run agents like Hermes, LangChain Deep Agents, and OpenClaw more securely inside NVIDIA OpenShell with managed inference.

TypeScriptAI & Machine Learning3,085 forks
14,627

TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way.

Python2,745 forks
NVIDIA

TensorRT

13,348

NVIDIA® TensorRT™ is an SDK for high-performance deep learning inference on NVIDIA GPUs. This repository contains the open source components of TensorRT.

C++2,407 forks
NVIDIA

cutlass

10,437

CUDA Templates and Python DSLs for High-Performance Linear Algebra

C++2,084 forks
NVIDIA

garak

9,258

the LLM vulnerability scanner

Python1,282 forks
NVIDIA

OpenShell

8,615

OpenShell is the safe, private runtime for autonomous AI agents.

Rust1,252 forks
NVIDIA

DALI

5,759

A GPU-accelerated library containing highly optimized building blocks and an execution engine for data processing to accelerate deep learning training and inference applications.

C++676 forks
NVIDIA

cuml

5,277

NVIDIA cuML: GPU-Accelerated Machine Learning

Python674 forks
NVIDIA

nccl

5,092

Optimized primitives for collective multi-GPU communication

C++1,415 forks
4,181

Generative AI reference workflows optimized for accelerated infrastructure and microservice architecture.

Jupyter Notebook1,098 forks
3,810

A unified library of SOTA model optimization techniques like quantization, distillation, pruning, neural architecture search, speculative decoding, etc. It compresses deep learning models for downstream deployment frameworks like TensorRT-LLM, TensorRT, vLLM, etc. to optimize inference speed.

PythonAI & Machine Learning599 forks
3,533

A library for accelerating Transformer models on NVIDIA GPUs, including using 8-bit and 4-bit floating point (FP8 and FP4) precision on Hopper, Ada and Blackwell GPUs, to provide better performance with lower memory utilization in both training and inference.

Python823 forks
NVIDIA

skills

3,301

Project brief: Agent Skills for NVIDIA products, install into Claude Code, Codex, and other coding agents to run Physical AI, robotics, simulation, CUDA, and RAG workflows end to end. Skill Catalog Product | Description | Skills | AIQ | NVIDIA AI-Q Blueprint - deploy local AI-Q services and run shallow or deep research workflows as agent skills.

Python394 forks
3,253

Open-source deep-learning framework for building, training, and fine-tuning deep learning models using state-of-the-art Physics-ML methods

Python783 forks
2,632

The NVIDIA NeMo Agent toolkit is an open-source library for efficiently connecting and optimizing teams of AI agents.

Python755 forks
NVIDIA

kvpress

1,209

LLM KV cache compression made easy

Python179 forks
NVIDIA

raft

1,042

RAFT contains fundamental widely-used algorithms and primitives for machine learning and information retrieval. The algorithms are CUDA-accelerated and form building blocks for more easily writing high performance applications.

Cuda250 forks