Your Favorite Models — All in One Place, with Unlimited Generations.

¡Aprovecha hasta un 27 % de descuento por tiempo limitado! ›
Guías de imagen

El manual de Nano Banana 2

O
OpenArt Team
Mar 27, 2026 · 7 minutes read
The Nano Banana 2 Handbook

Gemini 2.0 Flash Image de Google, conocido por aquí como Nano Banana 2, es más rápido, más barato y, según nuestro colaborador Brian de Litany of Ignition, "significativamente mejor con las estrías". Le pedimos a Brian que lo pusiera a prueba en OpenArt y documentara lo que encontró. El resultado es un tutorial de 12 minutos repleto de consejos, prompts y un sorprendente número de visitas a Applebee's. Aquí tienes el desglose completo.

NB2 vs. Nano Banana Pro: The honest comparison

→ 1:44 en el vídeo

nb2-graphic1-comparison.png

Behind the banana, there is now a dynamic thinking engine capable of reasoning through complex prompts, pulling real information from the web rather than hallucinating it, and analyzing image inputs to maintain character consistency across generations. That is what makes NB2 genuinely different from its predecessor — not just a speed and cost improvement, but a new category of capability.

Tip 1: Don't edit inside the editing interface

→ 0:57 in the video

Brian's first tip is also his most immediately actionable: skip the editing interface entirely. Instead, drag your base image directly into the visual reference window and structure your prompt like this:

"Edit image one based on the prompt. [Your edit here]."

On OpenArt, this approach gives NB2 direct access to the image and consistently produces cleaner, more precise edits. Brian also recommends bumping your resolution up to 4K before generating. It is not the same as upscaling, but it gives the model significantly more pixel space to work with on fine details like text, name tags, or small props. This becomes especially important when you are prompting anything with intricate detail work.

Consejo 2: El prompt de la hoja de personaje

→ 2:45 en el vídeo

This was the most requested item in the comments section. Brian pinned the full prompt after the video went up, and here it is:

nb2-graphic2-prompt.png

Two things to keep in mind when using this prompt. First, always manually tag your reference image using @image1 — the prompt will not work without it. Second, NB2 can accept up to five character sheets tagged into the same scene, which Brian stress-tests later in the video with some genuinely impressive results. It is worth noting that both NB2 and Nano Banana Pro handle this prompt well, though NB2 has the advantage of supporting more simultaneous character inputs.

→ 3:46 in the video

This is where NB2 pulls most clearly ahead of its predecessor. The model can perform a web search when it determines it does not have enough information to complete a task, which opens up a genuinely new category of prompting. Brian tested this by photographing grocery store ingredients and asking NB2 to generate recipe infographics without telling it what the recipe should be. The model had to identify the ingredients from the photos, look up a relevant recipe, and render everything as a visual layout.

For the lasagna test, he used 12 separate reference images. OpenArt accepts up to 14.

Los resultados fueron en gran parte impresionantes, aunque Brian tenía sus observaciones:

"Many of the outcomes were a bit light on instructions as to the ingredient amounts, more or less implying that I should just dump entire boxes or jars of ingredients in the bowls, layer everything, and enjoy my yolo lasagna."

The practical takeaway is that web-grounded prompting works well for reference-heavy tasks where you want the model to synthesize information across multiple inputs. It is meaningfully better than NB Pro at this kind of task, even if the outputs occasionally require some interpretation.

Tip 4: Multilingual text generation

→ 7:10 en el vídeo

NB2 llega con una renderización de texto multilingüe mejorada, que Brian probó pidiéndole que tradujera su infografía de tarta de queso al japonés. La traducción se renderizó correctamente, aunque Brian es sincero sobre los límites de su capacidad para verificar la exactitud. Ha invitado a hispanohablantes de japonés a opinar en los comentarios. El veredicto sigue pendiente.

Tip 5: Spatial reasoning in complex scene edits

→ 7:58 en el vídeo

Brian brought NB2 to Times Square, photographically speaking, and asked it to remove all the LED billboards and signage from a rainy night shot. What makes this test genuinely interesting is the difficulty of what the model has to do: it must remove the signs, extrapolate what the buildings look like underneath them, and recalculate the lighting, since most of the illumination in a wet Times Square shot is reflected bounce light from those very signs. That is a significant amount of spatial and physical reasoning happening in a single generation.

The results are impressive, with some expected imperfection toward the center of the frame. Brian followed that test with a sequence of additional edits, removing pedestrians, switching to a sunny day, and then converting the scene to a blizzard. That last one worked well enough that he declared he can now generate winter B-roll without leaving the house.

Tip 6: Multi-character scenes with up to five consistent characters

→ 9:32 in the video

La prueba más ambiciosa del vídeo. Brian generó cinco hojas de personaje distintas y las etiquetó todas en una misma escena de comedor de Applebee's, pidiendo ángulos de cámara variados mientras los personajes comían costillas. La mayoría de los personajes se mantuvieron coherentes entre generaciones. El ama de casa retro fue la más complicada, algo que Brian atribuye a la mayor sensibilidad que tiene nuestro cerebro ante ligeras incoherencias en rostros humanos de aspecto realista.

A few unexpected things surfaced along the way: the robot had a persistent desire to resemble Bender from Futurama despite no such prompt, and some outputs showed floating character sheet artifacts. But the unambiguous highlight came when NB2 invented names for all five characters without being asked: Blonde Ambition, Steel Resolve, The Gentle Toad, Mew Pawsome, and Jinete de la Calavera. El veredicto de Brian: «Retiro todas mis críticas. Es el mejor modelo de imagen de la historia.»

Quick reference: which model for what

nb2-graphic3-models.png

The bottom line

NB2 is a meaningful upgrade over Nano Banana Pro, particularly for image editing, web-grounded prompts, and multi-character consistency. The thinking engine underneath makes it capable of things its predecessor simply could not do, and it does them faster and at a lower cost. It is not a perfect tool, but it is a genuinely powerful one, and it is available on OpenArt right now.

Brian's tutorial is worth watching in full. One thing the comments consistently praised was his willingness to show where the model produces unexpected or imperfect results, not just the polished outputs. That honesty gives you a much more realistic picture of what to expect when you sit down to use it yourself.

👉 Pruébalo en OpenArt

Crea sin límites

Join millions of creators using OpenArt to generate images, videos, characters, and stories - all in one platform.

Empieza gratis →