vllm-project
★ 86.907vllm
Eine speichereffiziente Inferenz- und Serving-Engine mit hohem Durchsatz für LLMs.
HYSEN LABS VERZEICHNIS
Geprüfte Repositories, Klassifikationen und Analysen der Organisation vllm-project.
Eine speichereffiziente Inferenz- und Serving-Engine mit hohem Durchsatz für LLMs.
Ein Framework für effiziente Modellinferenz mit Omnimodalitätsmodellen.
Kosteneffiziente und steckbare Infrastrukturkomponenten für GenAI-Inferenz.