GPT Image 2 just hit #1 globally for text-to-image, straight up crushed Nano Banana 2.

GPT Image 2 just hit #1 globally for text-to-image, beating out the old champ Nano Banana 2. I couldn’t find which benchmark or scoring dimensions they used, so I’m not about to guess the numbers.

As someone who works on algorithms, I take these rankings with a grain of salt. Text-to-image benchmarks are super sensitive to the prompt set and evaluation criteria—swap the test method and the same model can drop several spots. Being #1 on some leaderboard doesn’t mean it crushes every scenario, especially since Chinese and English prompts often give totally different results.

What really matters is blind testing with your own stuff: run the same batch of prompts across a few models and compare. If you’ve got the hardware, just do a round yourself—way more useful than staring at rankings. So between these two, which one’s giving you better outputs?

Which ranking list are you even talking about? If you don’t name the list, it’s like saying nothing at all.

The whole text-to-image benchmark scene is a total mess. Swap out your prompt and the rankings flip completely.

Honestly, after testing it out, I still think it really depends on the subject matter. Portraits and landscapes each have their own strengths.