1 open source tools found
Run local LLMs on Mac effortlessly from the menu bar, with smart caching and multi-model support.
Local LLM inference server optimized for Mac with continuous batching and tiered hotspot/cold KV cache. Manage everything from the menu bar.