Nano Banana 2 Lite

The first example of generating home interiors fills me with indescribable hatred. Recently real estate agents have taken to running every dilapidated unsellable apartment through these AI filters, and you have to scroll through a dozen of these Ikea-chic images of what the apartment presumably could look like, before you are allowed to see the horrors they are trying to peddle at insane prices.

I received early access to test this model. (through work — Google still does not like me personally lol)

It works as advertised here, and it does behave like a distilled Nano Banana 2 with respect to certain elements such as good text rendering, which Nano Banana 1 does much worse with. It is definitely not at the level of the base Nano Banana 2 of course particularly with highly-nuanced prompts. My main criticism is that you cannot programmatically force aspect ratios with NB2L but you can with NB2.

That said, the price of $0.034/image is higher than expected since price is generally correlated with generation time, and it takes half the time to generate than a Nano Banana 1 image which costs $0.039/image. Google's assertion that you can directly replace NB1 pipelines with NB2L is fair.

Yesterday, Google announced that the Gemini app will allow free image generations (https://blog.google/innovation-and-ai/products/gemini-app/pe...) but did not specify which model would be used: I suspect it's the main motivation for Nano Banana 2 Lite.

It's sort of amazing that Grok's image model beats Nano Banana on nearly every one of the metrics they chose to highlight.

They didnt include chatgpt in the comparison chart. That tells a lot

The speed is definitely impressive. I'm seeing under 5 seconds per image vs ~30 seconds for base NB2.

I built an app for my kids that generates illustrated stories for them with them as the characters. I wanted to prioritize likeness while still stylizing the illustrations. I tested a bunch of models but none seem to come close to maintaining likeness when stylized. I find the others generate generic looking characters.

I'm excited to incorporate this into the onboarding of my app since I want the users to experience the aha moment as soon as possible and waiting half a minute+ isn't ideal. I'll still be using the main NB2 for the actual illustrations as this lite version still has slight issues with nuance and consistency as others have pointed out.

I'm way behind in imagegen - only using it occasionally for roleplaying tokens, goofing around, and random personal assets. To me, this is nuts. It's able to create images in like 2 seconds... before with chatgpt it would take 30s-1m for the same quality image. I don't get the negative comments here

Expensive and Google doesn't even have enough resources to decently deploy a model like that. Creating 10 images in parallel gives me RESOURCE_EXHAUSTED error, which is a painfully common error when using Google AI products.

NB2 is an impressive tool. Camera File -> Heavy Changes in NB2 -> Final Tweaks in Photoshop -> Production Image

Wow, that's a pretty massive decrease in latency, which should unlock some use cases, but the linked web page doesn't exactly make it straightforward to understand the differences between the models.

However, based off my personal experiences with general images models, Google in my opinion is the best for my workflows. Granted, I haven't tried far-east providers yet.

What does everyone else think?

It seems to respond to edits much better than the current production image model, which often stubbornly locks on to prior iterations of the images.

gemini is so far behind. starting to wonder if their strategy is launching the low cost alternative to image/text models. last release was 3.5 flash

It's only half the full model price, $30/m output: https://cloud.google.com/gemini-enterprise-agent-platform/ge...

Nano Banana is head and shoulders above the rest, but still too steep for personal use, and half off doesn't really mean much for enterprise if the results are worse. Hopefully this drives the rest to catch up at least.

It seems to respond to edits much better than the current production image model, which often stubbornly locks on to prior iterations of the images.

NB2 is an impressive tool. Camera File -> Heavy Changes in NB2 -> Final Tweaks in Photoshop -> Production Image

Wow, that's a pretty massive decrease in latency, which should unlock some use cases, but the linked web page doesn't exactly make it straightforward to understand the differences between the models.

However, based off my personal experiences with general images models, Google in my opinion is the best for my workflows. Granted, I haven't tried far-east providers yet.

What does everyone else think?

It's only half the full model price, $30/m output: https://cloud.google.com/gemini-enterprise-agent-platform/ge...

> Google doesn't even have enough resources to decently deploy a model like that

Probably not for free but tbf, Google did scale "AI Mode" globally to its billion+ users, with its Gemini 3 series. Pretty much broke my habit of searching the web with pplx & Chat.

They didnt include chatgpt in the comparison chart. That tells a lot

That is fair to point out. For those who don't know, ChatGPT Image 2 has an absurd ELO of 1387; compared to the #2 model at 1273, it's over 100 points higher (https://arena.ai/leaderboard/text-to-image). The tradeoff is latency, and ChatGPT Image 2 at High is...slow (~2 minutes at 1024x1024). In both cases it would have skewed the charts here to uselessness.

I want to do a writeup on ChatGPT Image 2 but at this point I don't think people care about nuanced image generation anymore...even though ChatGPT Image 2 crushes all my existing tests.

It's sort of amazing that Grok's image model beats Nano Banana on nearly every one of the metrics they chose to highlight.

The speed is definitely impressive. I'm seeing under 5 seconds per image vs ~30 seconds for base NB2.

I received early access to test this model. (through work — Google still does not like me personally lol)

... does it? Are you seeing something I'm not seeing? Number one is that this just doesn't appear to be true (non-lite versions beat it across the board it seems), number two is that this specifically is a low-cost bulk model and not a SOTA frontier model, of course the benchmarks are lower.

I think that should be illegal and misrepresenting. Lots of gray area with AI usage.

And it's borderline fraud, I think I saw an apartment on Streeteasy where they were able to 'fit' an entire desk, drawers and a queen size bed, obviously these image models just scale these down to proportions that just don't exist in real life.

the actual bedroom could only fit queen size bed ;(

Where I live (NYC) putting altered images like that has been the norm for more than a decade.

It’s just used to be more expensive to hire someone to do it for you.

The altered images always e free stirs the same bright walls and grey magazine style furniture.

AI is just making it cheaper, but this was bound to happen.

(Images altered this way do have a small watermark stating so)

Just having a good photographer is amazing. When my friend was selling their place I was amazed and how good the house looked in the listing. How big it looked, when I know it was not big. This was before AI filters were available. So not a new issue but certainly made worse and cheaper to do.

In a sane world, this would be a clear cut case of false advertisement, and the real estate agents would be held liable for fraud. Sadly, we don't live in a sane world.

I tried something like that but an error told me it couldn't do anything with children. Did that change?

You can set aspect ratios with NB2 Lite programmatically through Vertex [1]. I updated the program I use to help create all the images for GenAI Showdown, set the model ID to `gemini-3.1-flash-lite-image`, and was able to use aspect ratios like 16:9, 4:3, and others.

[1] - https://cloud.google.com/developers/vertex-ai

but until this comes to edge gallery i won't care

What kind of work are you doing that requires automated image generation at scale?

> Google still does not like me personally lol

Please elaborate.

gemini is so far behind. starting to wonder if their strategy is launching the low cost alternative to image/text models. last release was 3.5 flash

Imagine saying Gemini is so far behind when Llama has unreleased max models from last year that are currently beat by quantized Qwen2.5 models.

ChatGPT detail is a lot better though. You can do stuff like complex 6-panel comics that Nano Banana can't match.

Also a lot of the negative comments are from people who hate the very idea of AI art and want it to fail.

Different use cases.

People making images, where the image is the focal point, want to spend more per image.

Where images are parts of reports or throwaways or demos, cheap is the better approach.

> Google doesn't even have enough resources to decently deploy a model like that

Probably not for free but tbf, Google did scale "AI Mode" globally to its billion+ users, with its Gemini 3 series. Pretty much broke my habit of searching the web with pplx & Chat.

Imagine saying Gemini is so far behind when Llama has unreleased max models from last year that are currently beat by quantized Qwen2.5 models.

ChatGPT detail is a lot better though. You can do stuff like complex 6-panel comics that Nano Banana can't match.

Also a lot of the negative comments are from people who hate the very idea of AI art and want it to fail.

Different use cases.

People making images, where the image is the focal point, want to spend more per image.

Where images are parts of reports or throwaways or demos, cheap is the better approach.

I want to do a writeup on ChatGPT Image 2 but at this point I don't think people care about nuanced image generation anymore...even though ChatGPT Image 2 crushes all my existing tests.

That arena leaderboard has some questionable results. Anyone who's used these models would know that ranking HiDream above Krea2 is a pretty hot take.

Many of these ELO comparative tests (ArtificialAnalysis is guilty as hell on this as well) also have other problems such as a considerable number of "amateur judges" tending to prioritize aesthetics over actual instruction-following given the prompt.

Also (less a critique of Arena.AI necessarily), but the MAI models are so incredibly locked down (e.g. censored) as to be functionally useless. I have a sneaking suspicion its fallout from Tay.

https://en.wikipedia.org/wiki/Tay_(chatbot)

I definitely appreciated your post about Nano Banana Pro. It's also a genuinely useful time-capsule for how these systems evolve and where they fall short. I've preferred the output of ChatGPT Image 2. I think a post would be very helpful for folks to see what they're missing.

Where I live (NYC) putting altered images like that has been the norm for more than a decade.

It’s just used to be more expensive to hire someone to do it for you.

The altered images always e free stirs the same bright walls and grey magazine style furniture.

AI is just making it cheaper, but this was bound to happen.

(Images altered this way do have a small watermark stating so)

In a sane world, this would be a clear cut case of false advertisement, and the real estate agents would be held liable for fraud. Sadly, we don't live in a sane world.

I think that should be illegal and misrepresenting. Lots of gray area with AI usage.

Why should that be illegal? It’s multiplying the productivity of our economy, instead of someone having to waste time and money making the apartment actually look like that, you can just generate an image of it, that’s massive productivity boost with no harm done to the final product, unless the tenant cares about the slippage between a generated image of an apartment that looks nice and an apartment that’s actually nice.

And plus thats time the real estate agent could have spent prompting claude to cure cancer so its a double win

the actual bedroom could only fit queen size bed ;(

but until this comes to edge gallery i won't care

I tried something like that but an error told me it couldn't do anything with children. Did that change?

[1] - https://cloud.google.com/developers/vertex-ai

What kind of work are you doing that requires automated image generation at scale?

> Google still does not like me personally lol

Please elaborate.

How is that borderline, that’s just actual fraud.

That arena leaderboard has some questionable results. Anyone who's used these models would know that ranking HiDream above Krea2 is a pretty hot take.

Also (less a critique of Arena.AI necessarily), but the MAI models are so incredibly locked down (e.g. censored) as to be functionally useless. I have a sneaking suspicion its fallout from Tay.

https://en.wikipedia.org/wiki/Tay_(chatbot)

And plus thats time the real estate agent could have spent prompting claude to cure cancer so its a double win

What, after all, is a bit of light fraud, if it saves an estate agent some time?

I'm not sure if you're being serious but it should be illegal because they're producing images that are often not physically possible. At least if an agent stages an apartment with real furniture they are doing something a tenant really could do. But these AI images tend to change the physical dimensions of a room, use images of furniture that don't make sense dimensionally, shift the "natural" light of the room in a way that the sun will never provide and sometimes even change the view through the windows of the room.

How is that borderline, that’s just actual fraud.

What, after all, is a bit of light fraud, if it saves an estate agent some time?

I think their last sentence is a pretty clear indicator that they were not being serious.

As with bitcoin fans before them, Poe's Law is in full effect with the AI boosters.

Create more for less with Nano Banana 2 Lite. Generate and edit images faster and more efficiently than ever before.

Capabilities
Hands-on
Comparisons
Showcase
Performance
Safety
Try Nano Banana 2 Lite

Lightning-fast latency

Explore, iterate, and keep your workflow moving with dramatically reduced latency.

Cost-efficient at scale

Generate thousands of images at a fraction of the cost of heavier production models.

No compromise on quality

The control and accuracy you expect from Nano Banana, accelerated. Maintain character consistency, edit visuals with precision, and lean on real-world knowledge.

Slide 1 of 4

Space Lift

Space Lift is an interior design app that lets you instantly reimagine any room. Upload an image of your space, and watch the app generate a variety of fully realized concepts, from Mid-Century Modern to Bohemian Chic. Swipe through each custom card to find the perfect design scheme for your home.

Gridscape

Gridscape’s infinite canvas lets you explore and learn about any topic. When you ask a question, an informational "node" maps out ideas using text and images generated with Nano Banana 2 Lite and Gemini 3.1 Flash Lite. Dive deeper using clickable pathways that explore related concepts.

Peek-A-Word

Transforming passive reading into an interactive learning journey, Peek-A-Word turns selected text into AI-generated visuals. Concise definitions and contextual imagery are generated in one space, without any distracting tab-switching to disturb flow of learning with Nano Banana 2 Lite and Gemini 3.1 Flash Lite.

Anywhere

Be transported to dream destinations across the world with Anywhere, an interactive 3D globe created with Nano Banana 2 Lite. Attach an image to generate a series of personalized postcards at iconic global landmarks. Spin the globe, click on any photo, and uncover fascinating travel facts about your virtual vacation spots.

Slide 1 of 5

“Nano Banana 2 Lite is fast and reliable, helping designers explore more ideas to craft unique images on Figma Weave's node-based canvas. It's ideal for rapid iteration while staying in the creative flow.”

Itay Schiff
Co-founder & Creative Director, Figma Weave

“We have been testing Nano Banana 2 Lite to power real-time image generation within Manus’s autonomous workflows—from slide decks to web pages. Its speed suits these scenarios well, allowing our AI Agent to iterate on visuals quickly and deliver results in seconds. The image quality is also impressive, coming close to the full Nano Banana 2. We look forward to continuing our partnership and building better experiences together.”

Tao Zhang
Co-Founder & Chief Product Officer, Manus AI

“Speed is no longer a limitation. When generation is faster than imagination, creators can stay inside the idea instead of waiting on the tool. Nano Banana 2 Lite brings that feeling into the creative process, letting thoughts move into visuals almost instantly. For Artlist’s users, it means less time staring at a progress bar and more time creating, iterating, personalizing, and moving at the speed of culture.”

Idan Yonas
Director of AI Content & Innovation, Artlist

"For our voice-controlled TV game Wit's End, [instant-ramen] delivers consistent, high-quality 1k images ~2.7× faster than Gemini 3.1 Flash Image with incredibly tight latency variance. Its ability to handle text-to-image, edits, and multi-image composition in one drop-in API gives us Flash-Lite speed and cost with Nano-Banana quality. It’s what makes real-time generative play viable at scale."

Max Child
CEO, Weekend

“Our engine generates the world as players explore it, so image speed generation is essential. Instant-ramen is a huge upgrade, enabling accurate visuals, while doing it quick enough to keep up with the player's experience. Instant-ramen's speed and fidelity make on-the-fly art generation something we can use to power living visual worlds.”

Nick Walton
CEO & Co-Founder, Latitude

Slide 1 of 4

Image Editing

Image editing Elo scores against competitors per lmarena.ai

Image Generation

Image generation Elo scores against competitors per lmarena.ai

Price per 1k resolution image

Creating your prompts

Use detailed prompts to take more control over the images you generate. Think about what you want to see – the characters, the setting, and the overall feel. The more detail you add, the closer the image will be to what you’ve imagined.

Visual and text fidelity

Not every image Gemini generates will be perfect – it can still struggle with small faces, accurate spelling, and fine details in images.

Data and factual accuracy

The model's real-world knowledge is extensive but not infallible. When generating infographics, annotating diagrams, or representing complex data, it may misinterpret information or produce factually incorrect results. Always verify data-driven outputs.

Translation and localization

The model is capable of generating and translating text in many languages, but it may struggle with grammar, spelling, cultural nuances, or idiomatic phrases.

Complex edits and image blending

Advanced features like masked editing, major lighting changes (like day to night), or blending multiple images may sometimes produce unnatural results, visual artifacts, or disjointed scenes.

Character features

The model excels at character consistency, but it may not always get it right. We're working to make this consistency even more reliable.

Gemini

Supercharge your creativity and productivity 

Google AI Studio

The fastest path from prompt to production 

Gemini API

Get started building with cutting-edge AI models 

Gemini Enterprise Agent Platform

Build, scale, and govern agents

Large Language Models (LLMs), such as Gemini 3.1 Flash-Lite Image, may sometimes provide inaccurate or offensive content that doesn’t represent Google’s views.

Use discretion before relying on, publishing, or otherwise using content provided by LLMs.

Don’t rely on LLMs for medical, legal, financial, or other professional advice. Any content regarding those topics is provided for informational purposes only and is not a substitute for advice from a qualified professional.

Hacker Times