Beating AI Brief: Google has officially released Nano Banana 2.1, an upgraded version of Nano Banana 2 that continues to position itself as a highly efficient image generation and editing model. The main improvements are focused on image quality, instruction following, text generation, and multi-turn character consistency. Gemini, AI Mode, AI Studio, Flow, Stitch, Google Ads, and other products have already begun rolling it out, with the API model name gemini-nano-banana-2.1.
Image editing is the focus this time. Version 2.1 supports 1K, 2K, and 4K, can use up to 14 reference images simultaneously, and maintains consistency for up to 4 characters and 10 objects. Google has also strengthened mask editing, meaning only specified areas are changed, and improved infographic layout, text in images, and ultra-wide aspect ratio generation. The model can now also choose among three thinking levels: minimal, medium, and high.
In Google's own model card, 2.1 with thinking enabled has higher point estimates than Nano Banana 2 and Nano Banana Pro across all 10 evaluations, including text-to-image, general editing, multi-person consistency, and mask editing. The independent Arena leaderboard also confirms a clear improvement: 2.1 currently ranks 5th in text-to-image and 6th in single-image editing, the highest among Google models, but it still trails GPT Image 2.5 and GPT Image 2.
The price has actually dropped. Standard API image output for 1K, 2K, and 4K is $0.0336, $0.0504, and $0.0756 respectively, roughly only half that of Nano Banana 2. However, the input price has risen from $0.50 per million tokens to $1.50, and text and thinking output has also increased from $3 to $7.50. The more reference images and the heavier the thinking, the smaller the cost advantage brought by the lower image prices.

