Distributed LLM grid that pools crowdsourced compute and VRAM into unified serving, automatically routing or splitting very large model inference to power agents and chat.