MiniMax just updated their voice model to 2.6. Anyone doing dubbing or audio content should keep an eye on this—voice synthesis has come a long way in the last couple years, from that obvious robot sound early on to actually pulling off emotion and pauses pretty well now. MiniMax has been putting work into their voice line for a while, and 2.6 seems like another step forward in naturalness and expressiveness.
I make short videos and constantly need voiceover dubbing—hiring someone is expensive and slow, so a decent TTS saves a ton of hassle. But the voice libraries, multilingual support, and long-text stability vary a lot between models—you really have to run your script through it to see if it clicks. Anyone doing podcasts or audiobooks, come chat about how the voices in 2.6 are.