Realtime TTS-2 adds natural-language voice direction, conversational awareness (the model hears prior-turn audio, not just transcripts), crosslingual identity preservation across 200+ languages, and Advanced Voice Design with three stability modes. That's where 'dub that sounds like acting' comes from — performance-level output rather than translated reading. Realtime TTS-2 is available now in research preview;
contact sales for production access.