Community-maintained open-source text-to-speech model for expressive, long-form multi-speaker conversational speech, offering multiple weight sizes plus fine-tuning and inference code for researchers.