Releases a 20-billion-parameter end-to-end continuous speech synthesis foundation model with pretrained and distilled weights plus inference and fine-tuning code for researchers.