Provides open-source speech-to-text, intent detection, and text-to-speech models and toolkits optimized for ultra-low latency on-device inference, enabling developers to build real-time voice agents across multiple platforms.