An open-source model family for long-form speech, dialogue synthesis, voice design, sound effects, and real-time streaming TTS