omlx
LLM inference server with continuous batching & SSD caching for Apple Silicon, managed from the macOS menu bar.
What it solves
LLM inference server with continuous batching & SSD caching for Apple Silicon, managed from the macOS menu bar.
The project context on this page is free to read. The original GitHub repository remains the source of truth; sign in only when you want to save or join the discussion.
Best fit
Teams evaluating AI & Machine Learning and Developer Tools
Community notes