Yo, MiniMax Speech 2.6 just dropped a new voice model update.

MiniMax just updated their voice model to 2.6. Anyone doing dubbing or audio content should keep an eye on this—voice synthesis has come a long way in the last couple years, from that obvious robot sound early on to actually pulling off emotion and pauses pretty well now. MiniMax has been putting work into their voice line for a while, and 2.6 seems like another step forward in naturalness and expressiveness.

I make short videos and constantly need voiceover dubbing—hiring someone is expensive and slow, so a decent TTS saves a ton of hassle. But the voice libraries, multilingual support, and long-text stability vary a lot between models—you really have to run your script through it to see if it clicks. Anyone doing podcasts or audiobooks, come chat about how the voices in 2.6 are.

Having a big enough sound library is key—getting tired of hearing the same few sounds on repeat.

Does long text mess up the coherence as you read? That’s the biggest dealbreaker for the experience.

mark, gonna try this with my voiceover script later.

Long texts that lose their flow halfway through are the worst for the reading experience—short sentences are fine, but once they get long, the flaws start showing.

Long texts that lose their flow halfway through are the worst for the experience—if the vibe’s off, it’s all for nothing.