02 - Contextual models and multimodal editing
How to change a fragment of a photo in AI and keep the rest?
In this group, an existing image can be the starting point. You pass it to the model and describe the change: different lighting, background, clothing, prop, detail or composition. The remaining elements of the image are to be kept.
Contextual models, such as Nano Banana, FLUX Kontext or Seedream, allow you to change a shirt to a different one, insert a product into a lifestyle scene or recreate the same character in different shots. Previously, this required Photoshop, manual graphic design work or a photo shoot. Now all you need is a reference photo and a prompt.
From practice: one sentence that saves the frame. Sometimes you have to explicitly ask the model to keep elements from the reference. If I am changing something in a photo, I type: keep the entire structure of the photo, and replace the given character.
From practice: brand colour. A HEX code in a prompt gives a very close colour, but rarely a perfect one. It is much better to provide the model with a graphic containing that colour.
How to keep the same character or product in multiple AI graphics?
Nano Banana has changed the way we work with consistency. You can provide a photo of a product or a character, and then develop subsequent images while preserving their most important features.
With a product, you keep the object and change the surroundings. With a character, you keep the face and silhouette, and change the outfit, framing or the entire scene. Prepare a character sheet: a portrait from the front, profile, three-quarters and a full silhouette. You return to it with each subsequent shot, and the model holds the character's appearance more precisely.
The simpler the product shape, the better the result. Fine details, labels with small text or complex shapes require more iterations, and sometimes manual correction. Midjourney would recreate the product from scratch, resulting in slight variations. Contextual models have an advantage here and allow you to create additional formats for social media.
How to make an AI graphic with text?
For product creatives, I recommend Nano Banana: it maps the product best and handles text on the image well. To get the text right, keep it below 400 words in the entire graphic: more text means more errors and distortions. Nano Banana 2 writes very well, although occasional errors still happen.
GPT Image 2 and Seedream 5 also offer great control during generation and editing.