> For the complete documentation index, see [llms.txt](https://docs.chatvideopro.com/llms.txt). Markdown versions of documentation pages are available by appending `.md` to page URLs; this page is available as [Markdown](https://docs.chatvideopro.com/features/image-generation/text-to-image.md).

# Text-to-Image

Text-to-Image creates a still image from a written prompt. Use it when you want a new image, concept, thumbnail idea, product visual, style frame, background, or reference asset and you do not already have a source image to edit.

Chat Video Pro gives you two normal Text-to-Image paths: a fast chat path for quick ideas, and a full-control Generate Media path when you need model, aspect ratio, resolution, and quality settings. For cinematic stills and video-ready source frames, use Studio Cinematic Lab.

***

### When To Use Text-to-Image

Use Text-to-Image when you want to create a still from scratch:

* A quick concept image.
* A product or brand visual.
* A thumbnail background.
* A mood board image.
* A social post image.
* A background plate.
* A first draft before image editing.
* A source image for later video generation.

Choose another workflow when:

| You want to...                          | Use instead                              |
| --------------------------------------- | ---------------------------------------- |
| Create a cinematic frame for video work | Studio Cinematic Lab                     |
| Edit an existing image                  | Image-to-Image                           |
| Make a controlled layered edit          | Canvas Editor                            |
| Remove a background                     | Background Removal                       |
| Increase resolution                     | Image Upscaling                          |
| Build a YouTube thumbnail               | Thumbnail Mode                           |
| Animate the image into video            | Image-to-Video or Studio Motion Director |

***

### Quick Path vs. Full Control

#### Quick Path

Use the quick path when you want a fast image during a conversation.

1. Set your default image model in the top-right model selector.
2. Type a natural language request in chat.
3. Example: `Generate an image of a cozy cabin in the snow at night.`
4. Chat Video Pro generates the image using default settings.

Best for:

* Fast ideas.
* Casual requests.
* Conversation flow.
* Images where exact aspect ratio and resolution are not critical.

Limitations:

* Less control.
* Uses default settings.
* Not ideal for production-specific format requirements.

#### Full Control

Use Generate Media mode when the image needs specific settings.

1. Enable **Generate Media** in the composer.
2. Choose an image model.
3. Choose aspect ratio and resolution/quality.
4. Enter a more detailed prompt.
5. Generate and review.

Best for:

* Production work.
* Specific aspect ratios.
* Higher-quality outputs.
* Model comparisons.
* Images that may become thumbnails, key art, or video source frames.

***

### Text-to-Image vs. Cinematic Lab

Use normal Text-to-Image when you want direct model control or a quick image.

Use Cinematic Lab when the still needs to feel like a real production frame.

| Use Text-to-Image when...                          | Use Studio Cinematic Lab when...                                                     |
| -------------------------------------------------- | ------------------------------------------------------------------------------------ |
| You need a quick image or concept.                 | You need a cinematic still with camera/lens control.                                 |
| You know which image model you want.               | You want the workflow to guide the visual look.                                      |
| The image is a draft or simple asset.              | The image may become a Motion Director, Multi-Cam, AI Transition, or Relight source. |
| You want to manually write the whole image prompt. | You want camera body, lens, focal length, aperture, references, and model controls.  |

The practical rule: use Text-to-Image for general image generation, and use Cinematic Lab when the still is part of a video or production workflow.

***

### What A Good Prompt Includes

A strong image prompt usually describes:

<table><thead><tr><th width="185">Prompt element</th><th>What to include</th></tr></thead><tbody><tr><td>Subject</td><td>The main person, object, place, or idea.</td></tr><tr><td>Setting</td><td>Where the image takes place.</td></tr><tr><td>Composition</td><td>Close-up, wide shot, centered product, low angle, overhead, etc.</td></tr><tr><td>Lighting</td><td>Golden hour, studio lighting, neon, soft window light, moody shadows.</td></tr><tr><td>Style</td><td>Photoreal, editorial, cinematic, product photo, illustration, graphic design.</td></tr><tr><td>Details</td><td>Materials, textures, colors, wardrobe, props, background elements.</td></tr><tr><td>Mood</td><td>Calm, energetic, premium, mysterious, cozy, futuristic.</td></tr></tbody></table>

Useful structure:

{% code overflow="wrap" %}

```
[Subject] in [setting], [composition], [lighting], [style], [color palette], [specific details], [mood].
```

{% endcode %}

You do not need to include everything every time. Add the details that actually matter for the image.

***

### Prompt Examples

#### Cinematic Concept Image

{% code overflow="wrap" %}

```
A detective standing in a rain-soaked alley at night, trench coat dripping, neon signs reflected in puddles, low-angle composition, moody cinematic lighting, shallow depth of field, tense noir atmosphere.
```

{% endcode %}

Why it works:

* Clear subject and setting.
* Lighting and mood are specific.
* Composition gives the model a shot shape.

#### Product Visual

{% code overflow="wrap" %}

```
Premium product photo of a matte black wireless headphone case on a dark stone surface, soft studio key light, subtle rim light, shallow depth of field, clean luxury commercial style, detailed texture.
```

{% endcode %}

Why it works:

* Product, material, surface, lighting, and style are all clear.
* The image has a usable commercial direction.

#### Social Background

{% code overflow="wrap" %}

```
Vibrant abstract background for a vertical social post, electric blue and magenta gradient, soft grain texture, subtle light streaks, clean center area for text overlay, modern energetic style.
```

{% endcode %}

Why it works:

* It tells the model the final use.
* It reserves space for text.
* It avoids overloading the image with detail.

#### Thumbnail Concept

{% code overflow="wrap" %}

```
High-contrast YouTube thumbnail background, dramatic studio lighting, shocked creator silhouette on the left, glowing laptop screen on the right, bold red and yellow color palette, empty space at top for title text.
```

{% endcode %}

Why it works:

* It includes layout.
* It leaves room for text.
* It uses thumbnail-specific visual language.

***

### Weak Prompts To Avoid

Too vague:

```
A picture.
```

Better:

{% code overflow="wrap" %}

```
Photoreal image of a cozy mountain cabin at night, warm light glowing from windows, snow falling, pine trees around the cabin, cinematic winter atmosphere.
```

{% endcode %}

Missing composition:

```
Coffee shop.
```

Better:

{% code overflow="wrap" %}

```
Wide interior photo of a cozy coffee shop, wooden tables, plants near large windows, warm morning light, soft depth of field, inviting editorial lifestyle style.
```

{% endcode %}

No use case:

```
Abstract background.
```

Better:

{% code overflow="wrap" %}

```
Abstract 16:9 background for a tech presentation, dark navy gradient, subtle circuit-like light patterns, clean empty center space, premium modern look.
```

{% endcode %}

***

### Choosing A Model

Use the model based on the hardest part of the image.

<table><thead><tr><th width="434">Need</th><th>Good starting point</th></tr></thead><tbody><tr><td>Strong general image generation</td><td>Nano Banana 2</td></tr><tr><td>Harder prompt or reference reasoning</td><td>Nano Banana Pro</td></tr><tr><td>Readable text in the image</td><td>GPT Image 2</td></tr><tr><td>Final realism and detail</td><td>Flux 2 Max</td></tr><tr><td>Creative, affordable drafts</td><td>Seedream v5</td></tr><tr><td>Fast mobile/social formats</td><td>Grok</td></tr><tr><td>Quick low-cost preview</td><td>Z-Image Turbo</td></tr><tr><td>Cinematic frame for video</td><td>Studio Cinematic Lab</td></tr></tbody></table>

For a deeper chooser guide, see Supported Image Models.

***

### Aspect Ratio Guide

Choose aspect ratio based on where the image will be used.

<table><thead><tr><th width="155">Aspect ratio</th><th>Best for</th></tr></thead><tbody><tr><td>16:9</td><td>Video frames, YouTube thumbnails, website headers, landscape images.</td></tr><tr><td>9:16</td><td>Shorts, Reels, TikTok, vertical stories, phone screens.</td></tr><tr><td>1:1</td><td>Square social posts, profile-style images, balanced compositions.</td></tr><tr><td>4:5</td><td>Instagram feed portraits and social graphics.</td></tr><tr><td>21:9</td><td>Cinematic widescreen frames and banners.</td></tr><tr><td>4:3 or 3:4</td><td>Editorial, vintage, portrait, or alternate framing.</td></tr></tbody></table>

Pick the final deliverable shape before generating. Cropping after generation can cut off important subjects, text areas, or composition lines.

***

### Best Practices

#### Describe The Image You Need, Not Just The Topic

"A watch" gives the model a topic. "Premium product photo of a black watch on a dark reflective surface with rim light" gives it an image.

#### Include Composition Early

Composition controls whether the image is usable. Say:

* Close-up portrait.
* Wide establishing image.
* Centered product shot.
* Low-angle hero shot.
* Overhead flat lay.
* Empty space on the right for text.

#### Save Text For The Right Workflow

If the image needs readable words, use GPT Image 2 or add the final text manually after generation. For thumbnails and graphics, generating the background first and adding final text yourself often gives the cleanest result.

#### Use References For Consistency

If a character, product, or brand look must stay consistent, attach reference images in a workflow that supports them or use Cinematic Lab references.

#### Generate A Few Directions Before Polishing

Do not over-optimize the first result. Generate a few directions, pick what is working, then refine.

#### Upscale Last

Do not use upscaling to fix a bad image. Choose the image first, then use Image Upscaling if it needs more resolution.

***

### Example Workflows

#### Quick Concept Image

1. Use the quick path in chat.
2. Ask for a simple visual idea.
3. If the direction works, regenerate with full control or move into Cinematic Lab.

#### Production Still

1. Use Generate Media or Cinematic Lab.
2. Choose the final aspect ratio.
3. Write a detailed prompt with composition and lighting.
4. Generate several directions.
5. Use the best image as a source for Motion Director, Multi-Cam, or AI Transitions.

#### Thumbnail Background

1. Choose 16:9.
2. Prompt for strong contrast and empty title space.
3. Avoid asking the model to create final text unless using GPT Image 2.
4. Add final thumbnail text manually for control.

#### Product Visual

1. Describe the product, material, surface, and lighting.
2. Use a clean composition.
3. Keep the prompt focused on the product.
4. Use Image-to-Image or Canvas Editor for refinements.

***

### Troubleshooting

#### The image feels generic

Add specific composition, lighting, material, and mood. Avoid one-word prompts. If you want a production frame, try Cinematic Lab.

#### The image has bad text

Use GPT Image 2 or add the text manually after generation. Keep generated text short.

#### The image is the wrong shape

Set the aspect ratio before generating. If the final platform is vertical, generate vertical from the start.

#### The model ignored an important detail

Move the detail earlier in the prompt and remove competing instructions. If the detail is a person, product, logo, or brand look, use references.

#### The output looks too AI-generated

Add concrete physical details: material, texture, imperfect surfaces, realistic lighting, lens feel, and environment. Cinematic Lab can help if you want a more grounded production-frame look.

#### The result is close but not final

Use the right follow-up workflow:

| Problem                        | Better next step                  |
| ------------------------------ | --------------------------------- |
| Need to edit part of the image | Image-to-Image or Canvas Editor   |
| Need a transparent cutout      | Background Removal                |
| Need higher resolution         | Image Upscaling                   |
| Need motion                    | Image-to-Video or Motion Director |
| Need a different angle         | Multi-Cam                         |

***

### Related Pages

* [Supported Image Models](/features/image-generation/supported-image-models.md) - Choose the right model.
* [Image-to-Image](/features/image-generation/image-to-image.md) - Edit an existing image.
* [Canvas Editor](/features/image-generation/canvas-editor.md) - Make controlled image edits.
* [Thumbnail Mode](/features/image-generation/thumbnail-mode.md) - Create thumbnail-focused images.
* [Cinematic Lab](/features/studio/cinematic-lab.md) - Generate cinematic stills for production and video workflows.
* [Image-to-Video](/features/video-generation/image-to-video.md) - Animate a still image.

***

**Next:** If the still needs to become a video shot, use Image-to-Video or Studio Motion Director.
