So apparently ZhiXiang Future’s image gen model hit #2 globally on some leaderboard, only behind OpenAI. Honestly kinda surprised a domestic model managed to squeeze into that spot.
But rankings have never really been my thing—they didn’t even specify which leaderboard this is or what dimensions they’re scoring on. The real test is just throwing a few images in and seeing how it performs. Rankings alone don’t mean much. Anyone here already tried it out? How’s the actual output quality? Especially curious about the two classic pain points: Chinese semantic understanding and hands.
Chinese typography has always been a weak spot for domestic models, so making it into the top two means they actually put some serious work into it. Just looking at the numbers, though, it does feel a bit hollow.