NemoClaw
Run agents like Hermes, LangChain Deep Agents, and OpenClaw more securely inside NVIDIA OpenShell with managed inference.
HYSEN LABS DIRECTORY
Verified repositories, classifications and analysis from the NVIDIA organization.
Run agents like Hermes, LangChain Deep Agents, and OpenClaw more securely inside NVIDIA OpenShell with managed inference.
Security scanner for AI agent skills. Detect vulnerabilities, malicious patterns, and security risks.
NVIDIA Linux open GPU kernel module source
TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way.
NVIDIA® TensorRT™ is an SDK for high-performance deep learning inference on NVIDIA GPUs. This repository contains the open source components of TensorRT.
NVIDIA Cosmos is an open platform of world models, datasets, and tools that enables developers to build Physical AI for robots, autonomous vehicles, smart infrastructure, and more.
CUDA Templates and Python DSLs for High-Performance Linear Algebra
OpenShell is the safe, private runtime for autonomous AI agents.
Samples for CUDA Developers which demonstrates features in CUDA Toolkit
the LLM vulnerability scanner
A PyTorch Extension: Tools for easy mixed precision and distributed training in Pytorch
NVIDIA Isaac GR00T N1.7 - A Foundation Model for Generalist Robots.
A Python framework for GPU-accelerated simulation, robotics, and machine learning.
A GPU-accelerated library containing highly optimized building blocks and an execution engine for data processing to accelerate deep learning training and inference applications.
NVIDIA cuML: GPU-Accelerated Machine Learning
Optimized primitives for collective multi-GPU communication
A unified library of SOTA model optimization techniques like quantization, distillation, pruning, neural architecture search, speculative decoding, etc. It compresses deep learning models for downstream deployment frameworks like TensorRT-LLM, TensorRT, vLLM, etc. to optimize inference speed.
Build and run containers leveraging NVIDIA GPUs
Generative AI reference workflows optimized for accelerated infrastructure and microservice architecture.
NVIDIA device plugin for Kubernetes