ggml-org
★ 121,303llama.cpp
llama.cpp runs LLM inference in plain C/C++, serving models locally through a REST API and web UI with multimodal support in its llama-server.
HYSEN LABS DIRECTORY
Verified repositories, classifications and analysis from the ggml-org organization.
llama.cpp runs LLM inference in plain C/C++, serving models locally through a REST API and web UI with multimodal support in its llama-server.
whisper.cpp is a dependency-free C/C++ port of OpenAI's Whisper speech-recognition model, optimized for Apple Silicon and supporting CPU-only inference on many platforms.