NemoClaw
通过托管推理,在 NVIDIA OpenShell 内更安全地运行 Hermes、LangChain Deep Agents 和 OpenClaw 等代理。
HYSEN LABS 项目目录
来自 NVIDIA 组织的可信仓库、分类与深度解析。
通过托管推理,在 NVIDIA OpenShell 内更安全地运行 Hermes、LangChain Deep Agents 和 OpenClaw 等代理。
AI 代理技能的安全扫描器。检测漏洞、恶意模式和安全风险。
TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way.
NVIDIA® TensorRT™ is an SDK for high-performance deep learning inference on NVIDIA GPUs. This repository contains the open source components of TensorRT.
CUDA Templates and Python DSLs for High-Performance Linear Algebra
the LLM vulnerability scanner
OpenShell is the safe, private runtime for autonomous AI agents.
用于 GPU 加速模拟、机器人和机器学习的 Python 框架。
A GPU-accelerated library containing highly optimized building blocks and an execution engine for data processing to accelerate deep learning training and inference applications.
NVIDIA cuML: GPU-Accelerated Machine Learning
Optimized primitives for collective multi-GPU communication
Generative AI reference workflows optimized for accelerated infrastructure and microservice architecture.
量化、蒸馏、剪枝、神经架构搜索、推测解码等 SOTA 模型优化技术的统一库。它为 TensorRT-LLM、TensorRT、vLLM 等下游部署框架压缩深度学习模型,以优化推理速度。
A library for accelerating Transformer models on NVIDIA GPUs, including using 8-bit and 4-bit floating point (FP8 and FP4) precision on Hopper, Ada and Blackwell GPUs, to provide better performance with lower memory utilization in both training and inference.
项目速览:适用于 NVIDIA 产品的代理技能,安装到 Claude Code、Codex 和其他编码代理中,以端到端运行物理 AI、机器人、模拟、CUDA 和 RAG 工作流程。技能目录产品|描述 |技能 |爱智商 | NVIDIA AI-Q 蓝图 - 部署本地 AI-Q 服务并作为代理技能运行浅层或深层研究工作流程。
Open-source deep-learning framework for building, training, and fine-tuning deep learning models using state-of-the-art Physics-ML methods
The NVIDIA NeMo Agent toolkit is an open-source library for efficiently connecting and optimizing teams of AI agents.
LLM KV cache compression made easy
用于探索、构建和部署人工智能天气/气候工作流程的开源深度学习框架。
RAFT contains fundamental widely-used algorithms and primitives for machine learning and information retrieval. The algorithms are CUDA-accelerated and form building blocks for more easily writing high performance applications.