Hysen Labs
オープンソースプロジェクト
kekzl/imp avatar
kekzl

imp

From-scratch C++23/CUDA inference engine for the NVIDIA RTX 5090 (sm_120a). The best single-GPU backend for agentic AI: tool calling, long-context loops, reasoning and concurrent sub-agents. Decode beats llama.cpp b9976 by 42-48% on dense GGUF (measured 2026-07-12), at-or-ahead of vLLM on NVFP4. 100% written by Claude Code.

スター 36フォーク 2CudaMIT
GitHub
情報の鮮度

分類

From-scratch C++23/CUDA inference engine for the NVIDIA RTX 5090 (sm_120a). The best single-GPU backend for agentic AI: tool calling, long-context loops, reasoning and concurrent sub-agents. Decode beats llama.cpp b9976 by 42-48% on dense GGUF (measured 2026-07-12), at-or-ahead of vLLM on NVFP4. 100% written by Claude Code.

このページのプロジェクト情報と編集内容は無料で読めます。元の GitHub リポジトリが最終的な情報源であり、保存や議論への参加時だけログインが必要です。

コミュニティノート

コミュニティノート