An edge-native MoE inference engine that runs massive open-weight models on consumer GPUs with fast execution and OpenAI-compatible APIs.