Compact single-model release unifying speech recognition, text-to-speech synthesis, and voice conversion, enabling multiple audio tasks with one tiny general-purpose model.