
Tech • AI • Robotics • Game
Google’s Nano Banana 2.1 delivered mixed results in side-by-side image tests, showing gains in editing and legible infographic text but not consistently beating Nano Banana Pro, Nano Banana 2, or Chat GBT Image 2.5 on photorealism, instruction following, or text rendering.
Nano Banana 2.1 has been launched as Google’s newest image model, with claims that it outperforms earlier releases including Nano Banana 2 and Nano Banana Pro while costing four times less. Google also says the model improves instruction following, in-image text rendering, search grounding, and character consistency. It is available through Google AI Studio, the Gemini API, and the Gemini app.
The model was compared directly with Nano Banana Pro, Nano Banana 2, and Chat GBT Image 2.5 using the same prompts. Tests focused on realistic scenes, camera-angle control, motion, infographics, and identity-preserving image edits. The comparisons highlighted not just visual appeal but also whether models followed detailed instructions such as hand placement, object orientation, perspective, and readable text.
In a busy street-level café prompt with multiple text elements, Nano Banana Pro was judged the strongest result, with legible signage, readable cup text, strong color balance, and convincing people. Nano Banana 2 also performed well but appeared overly saturated. Nano Banana 2.1 was a weak starter, producing a flatter, desaturated scene with repeated or broken text, while Chat GBT Image 2.5 also showed awkward repeated signage and hand issues.
A detailed gas-station prompt tested low-angle composition, hand assignments, sedan orientation, and object placement. None of the four models got everything right. Nano Banana Pro and Chat GBT Image 2.5 produced cinematic framing, while Nano Banana 2 was the only one to place the umbrella and cup in the correct hands, though it missed the car orientation and felt less cinematic. Nano Banana 2.1 improved on hand quality versus some rivals but still put the umbrella in the wrong hand and missed parts of the perspective setup.
In a trampoline action scene, Nano Banana 2.1 produced a believable subject with decent hair movement and a functional safety net, but Nano Banana 2 was preferred overall for more convincing motion in the shirt and hair. Nano Banana Pro struggled, generating a visibly strange body tilt and off-balance objects. Chat GBT Image 2.5 pushed the jump unrealistically high and introduced odd background texture in trees.
A loose prompt about the bends and decompression sickness tested whether models could assemble a coherent infographic with accurate text. All four outputs were considered factually accurate, but visual quality varied. Nano Banana 2 was judged the strongest overall for design, charts, and readability. Nano Banana 2.1 came second, with clean, legible text and no obvious garbling. Chat GBT Image 2.5 suffered from smudged or malformed lettering, and Nano Banana Pro was accurate but visually bland.
On most image prompts, Chat GBT Image 2.5 generated first, followed by Nano Banana 2.1, then Nano Banana 2, with Nano Banana Pro last. The infographic test was an exception: Nano Banana 2 and 2.1 finished first, ahead of Chat GBT Image 2.5, while Nano Banana Pro remained the slowest.
In a thumbnail-style edit featuring a boxer banana punching a person in the face, Nano Banana 2.1 produced the best overall result, while Nano Banana Pro also looked strong. Chat GBT Image 2.5 was criticized for adding a strange facial effect and weaker identity preservation. Nano Banana 2 was judged too exaggerated for the prompt.
In a rooftop-in-Tokyo identity edit with clothing changes and a New York Yankees cap, the strongest results were seen from Nano Banana 2 and Nano Banana 2.1. Chat GBT Image 2.5 again produced a recognizable but less flattering facial treatment. The result suggested that Nano Banana 2.1 may be more competitive in image editing than in fresh scene generation.
Nano Banana 2.1 appears to be an incremental upgrade rather than a clear across-the-board leader. It performs well on editing and clean text layouts, but older Google models and Chat GBT Image 2.5 still outperform it in several realistic scene and composition tests.
Ask a question