For the complete documentation index, see llms.txt. This page is also available as Markdown.

Image-to-Image

Edit and modify existing images by uploading an image and describing the changes you want. Perfect for style transfer, adding elements, changing colors, or transforming images.

Image-to-Image edits an existing image with a prompt. Use it when the image is already close, but you want to change the style, background, colors, mood, composition, objects, or visual direction.

This is different from Text-to-Image. Text-to-Image starts from nothing. Image-to-Image starts from a source image and asks the model to transform it.

Think of Image-to-Image as prompt-guided revision. The source image gives the model the starting point. Your prompt tells it what to change and what to preserve.


What This Tool Is For

Use Image-to-Image when you want to:

  • Restyle an image.

  • Change the lighting or mood.

  • Replace or adjust a background.

  • Add a visual element.

  • Remove a simple distraction.

  • Create a variation of an existing design.

  • Turn a rough concept into a more polished image.

  • Prepare a still for later video generation.

Choose a more specific workflow when the job needs stronger control:

You want to...
Use instead

Build a cinematic still from scratch

Studio Cinematic Lab

Make a controlled mask/layer edit

Canvas Editor

Remove the entire background

Background Removal

Increase resolution

Image Upscaling

Create a new image without a source

Text-to-Image

Animate the image into video

Image-to-Video or Studio Motion Director


How To Use It

  1. Enable Generate Media in the composer.

  2. Attach the image you want to edit.

  3. Choose an image model, or let Chat Video Pro switch to the matching edit version.

  4. Describe what should change.

  5. Choose aspect ratio, resolution, quality, or other available settings.

  6. Generate and compare the result with the source image.

When you attach an image, Chat Video Pro can route supported image models into their edit versions. For example, a text-to-image model may switch to its Image-to-Image version once the source image is attached.


When Image-to-Image Works Best

Image-to-Image is strongest when the edit is clear and focused.

Good fit
Example

Style change

Make this look like a cinematic 35mm film still with warm grain.

Lighting change

Make the scene feel like soft golden hour, preserving the same subject and pose.

Background change

Replace the plain studio background with a dark modern office.

Object addition

Add a small vintage camera on the table, keeping the rest of the scene unchanged.

Product variation

Change the bottle color to matte black while keeping the label readable.

Mood variation

Make this more premium and dramatic, with deeper shadows and subtle rim light.

It is weaker when the prompt asks for too many unrelated changes at once. If you need precise placement, masking, layered composition, or several separate edits, use Canvas Editor.


Choosing A Model

You do not need to memorize every edit model. Start with the type of edit.

If the edit needs...
Try...

A strong default for most image edits

Nano Banana 2

Text, labels, signs, packaging, or complex composition

GPT Image 2 Edit or Ideogram V4 Edit

High realism or polished final quality

Flux 2 Max Edit

Region-precise edits that keep the rest of the frame

Seedream 5.0 Pro Edit

Fast creative variations

Grok Edit or Ideogram V4 Edit

More exact placement, masks, or layered control

Canvas Editor

Use Supported Image Models for the broader chooser guide.


Writing Better Edit Prompts

A good edit prompt has three parts:

Part
What to say

Change

What should be different.

Preserve

What should stay the same.

Direction

The style, mood, lighting, or format you want.

Useful structure:

You do not need this exact wording every time, but it helps avoid accidental changes.


Prompt Examples

Style Transfer

Why it works:

  • It names the style.

  • It tells the model what not to change.

  • It avoids asking for a new scene.

Background Replacement

Why it works:

  • It changes only the background.

  • It protects the subject identity and pose.

  • It gives the replacement a clear visual direction.

Product Variation

Why it works:

  • It is specific about the product change.

  • It protects brand and composition details.

  • It is useful for testing design variants.

Social Version

Why it works:

  • It explains the destination format.

  • It asks for room for text.

  • It protects the subject.

Video-Ready Still

Why it works:

  • It improves the source frame for later animation.

  • It specifies light direction and composition.

  • It avoids changing the core image.


What To Preserve

If something matters, say so. Models may reinterpret parts of the image unless you protect them.

Common things to preserve:

  • Subject identity.

  • Face, pose, body shape, or wardrobe.

  • Product shape, logo, label, or color.

  • Camera angle and composition.

  • Background layout.

  • Text or UI elements.

  • Lighting direction.

  • Aspect ratio.

Example:


Aspect Ratio And Framing

For simple edits, keep the same aspect ratio as the source image. This usually preserves composition and avoids unexpected cropping.

Change aspect ratio when you are intentionally reformatting:

Destination
Common ratio

YouTube or video frame

16:9

Vertical short or Reel

9:16

Feed post

1:1 or 4:5

Cinematic banner

21:9

When changing aspect ratio, tell the model how to reframe the image. For example: Convert this to 9:16 with the subject centered and extra space above the head for title text.


Image-to-Image vs. Canvas Editor

Use Image-to-Image when a prompt can describe the edit clearly.

Use Canvas Editor when the edit depends on where something happens.

Use Image-to-Image when...
Use Canvas Editor when...

The whole image needs a style or mood change.

You need to paint, mask, or mark a specific area.

You want a quick variation.

You need controlled placement.

The edit is easy to explain in one prompt.

The edit has multiple layers or regions.

You are exploring directions.

You already know exactly what should change.

If the model changes the wrong part of the image twice, switch to Canvas Editor.


Best Practices

Make One Main Change At A Time

Image-to-Image works better when each generation has a clear job. If you need a new background, new outfit, new lighting, and new format, do them in stages.

Use Strong Source Images

The source image sets the ceiling. A blurry face, broken hand, bad product shape, or messy composition can carry into the edit.

Tell The Model What To Keep

Do not only say what to change. Add what should remain stable.

Save Good Variations

If an edit is close, save it before trying another direction. A later edit can drift away from the best version.

Prepare Video Source Frames Carefully

If the edited image will become video, keep the composition clean, avoid tiny text, leave room for motion, and use the same aspect ratio as the final video when possible.


Troubleshooting

The Edit Is Too Subtle

Use stronger language and make the change more specific. Try make the sky a dramatic stormy sunset instead of make it nicer.

The Image Changed Too Much

Add preservation language: Keep the same subject, pose, camera angle, and composition. You can also make a smaller edit first.

The Wrong Object Changed

Use location words like foreground, background, left side, on the table, or behind the subject. If the edit still lands in the wrong place, use Canvas Editor.

The Text Became Wrong

Use GPT Image 2 Edit when text matters, and keep text short. For final thumbnails or title graphics, it may be better to add text manually after generation.

The Aspect Ratio Cropped Important Details

Return to the source ratio or prompt the reframing directly: keep the full body visible, leave space above the head, or do not crop the product.



Next: If the edit needs exact placement or masking, open Canvas Editor. If the image is meant to become a cinematic video frame, use Studio Cinematic Lab.

Last updated