Local LLM inference server optimized for Apple Silicon, featuring continuous batching and memory-plus-SSD tiered caching with convenient menu-bar management for Mac users.