How do you write AI image prompts that work?
Describe the picture in sentences: its purpose, subject, light and exact words. To edit your own photo, name one change and list what stays.
A good AI image prompt describes a scene in sentences, not a pile of keywords: what the picture is for, who is in it, where they are, the light, and the exact words to print. To edit your own photo, say what changes and list everything that stays. For a video, say what moves.
We wrote the same idea four times, each version adding one kind of decision, and drew every version with GPT image. The table below shows what each step adds and how long the prompt got, for the lemonade poster and for a manga cover.
| Step | What the step adds | Lemonade poster | Manga cover |
|---|---|---|---|
| 1. Minimal | only the product and the person | 9 words | 3 words |
| 2. Structured | the brand name, the action and the headline | 37 words | 15 words |
| 3. Pro | a strict palette, the layout and typography rules | 163 words | 71 words |
| 4. Polished | the camera, exact counts and every line of text | 268 words | 236 words |
To see what each part of a prompt does, we wrote two ideas, a lemonade poster and a manga cover, at four levels each and drew every step on our own engines on 8 October 2026. The shortest prompt was 3 words, the longest 268. Short prompts do not fail: the model fills every gap on its own, so you get a picture, just not yours. Each step after that takes one decision back from the model. The full ladders, with every picture, are on our AI image prompts page.
One prompt was refused by moderation in the Studio and went through after we rewrote it. One generation failed for a transient reason and succeeded on a retry. Every picture in this guide comes from this run.
What should a good AI prompt include?
| Part | Example | What goes wrong without it |
|---|---|---|
| What it is for | A finished vertical advertising poster | The model guesses the format, often a plain photo with no layout |
| Subject and action | A girl in a white bucket hat sips through a striped straw | A generic person in a generic pose |
| Setting | Against a flat bubblegum-pink wall | A busy background that fights the subject |
| Light and camera | Hard on-camera flash, shot from slightly below on a 35mm lens | Flat, even light that reads as stock |
| Style and palette | Strict palette: bubblegum pink, lemon yellow, mint green, white | Colours wander, and every version looks different |
| Exact text | The words "SOUR SUN" in quotes, once, across the top | Invented words or misspelled lettering |
| What stays (edits) | Keep her face, skin tone, hair, lighting and framing unchanged | The face changes along with the edit |
| What moves (video) | The can tips toward the camera, slow push in | Random motion, or none at all |
How do you write a prompt for an AI image?
The order matters less than the habit of deciding. Our lemonade poster went from one line to a brief. The second step named the brand, the girl and the headline. The third fixed a strict palette, placed the photograph in the upper two-thirds and the slogan on a yellow panel, and added typography rules. The fourth set a lens and a camera angle, gave exact counts, and added two stickers and a button. Google gives the same advice for Nano Banana in plain terms: stop writing tag lists like "cool car, neon, city, night" and describe the scene instead.

A finished vertical 4:5 advertising poster for a fictional sparkling pink lemonade brand. Loud, glossy, maximum saturation, hard-flash look. STRICT PALETTE: bubblegum pink, lemon yellow, mint green, white, black outlines. No other colors. PHOTOGRAPH in the upper two-thirds: a laughing girl in a white bucket hat and a mint tee against a flat bubblegum-pink wall, one eye squeezed shut from the sour taste, lips puckered around a pink-and-white striped straw. She holds a tall slim can toward a 35mm lens shot from slightly below; cold drops run down the can, three lemon slices hang in the air beside it. Hard direct flash: hot highlights on cheeks, a sharp black shadow on the wall behind her shoulder. GRAPHICS: "SOUR SUN" set enormous across the top on one line in a chunky extruded 3D rounded display face: lemon-yellow faces, mint-green extrusion, thick black outline, white gloss streak. The lettering overlaps the brim of her hat. TWO glossy sticker badges with soft drop shadows: a mint starburst reading "NEW" in white caps, and a yellow lemon-shaped badge reading "zero sugar" in black lowercase. Exactly two stickers. LOWER THIRD on a solid lemon-yellow panel: "PUCKER UP, SUMMER" in black bold caps, centered; below it one line of small black text: "Real lemons. Pink fizz. Zero chill."; below that a black pill button with yellow text "Find a cooler". Typography rules: English, spelled exactly as given (the brand is S-O-U-R S-U-N), at most two typefaces, straight baselines, no doubled glyphs, no gibberish, "SOUR SUN" appears exactly once. Exactly one can, one straw, two hands, two stickers, one button. No watermark, no extra logos.

A vertical 4:5 travel poster filling the frame edge to edge, with a thin bone-cream inner border. Style: mid-century screenprint, flat layered color separations, clean graphic shapes, light stipple texture, visible paper grain as if printed on aged uncoated paper and photographed flat. Layout: the name LOFOTEN at upper left in a large high-contrast serif, black ink, one line. Under it NORWAY in small widely spaced black caps. A solid black band across the bottom carries the tagline WHERE THE LIGHT STAYS UP in bone-cream spaced caps, centered. Illustration in the lower three quarters: sharp green-black peaks rising straight out of a still fjord, a row of red wooden fishing cabins on stilts along the shore, drying racks of fish beside them, a midnight sun low over the water in a banded sky from deep teal through coral to pale gold, a single small rowing boat leaving a white wake. Palette held to teal, coral, pale gold, red ochre, slate green, bone cream and black. Typography: exactly two typefaces, straight baselines, correct spelling, Latin characters only. One place name, one country line, one tagline, each once. No people, no watermark.
How do you write a prompt to edit your own photo?
Edits tend to fail the same way: you ask for a new jacket and get a new person. Two habits prevent most of it. First, write the keep list every time, even on a later turn, because each turn can drift a little further from the original. Second, ask for one change, check it, then ask for the next. Google's own advice for editing is to be direct and specific, as in "change the man's tie to green". Use your own photo or a generated one, never a picture of someone who has not agreed to it.

Add two pale pink rose petals resting on her cheekbone and small clear water droplets beaded across the petals and her skin, one droplet leaving a short wet trail. Brush a soft rose blush high on the cheekbone. Keep her exact facial features, skin tone, closed eyes, hair, lighting and framing unchanged.

Change the woman's faded denim jacket to a fitted mustard-yellow corduroy jacket with brown buttons, and change her white tee to a thin cream turtleneck. Keep the same pose, the same mint-green bicycle, the same street and lighting, and keep the woman with short platinum hair exactly as she is, while keeping the same facial features, freckles and expression.
How do you write a prompt for an AI video?
Video prompts are short for a reason: the later an instruction sits in a long prompt, the less reliably it is followed, so the subject and its action go first and the camera move second. One Seedance guide puts the budget at 60 to 100 words. To try one, write it on our Seedance video generator page or turn a still into a clip with image to video.
How do prompts differ between Nano Banana, GPT image, Flux, Grok and Seedance?
| Model | Best for | How to write for it | Edits and references |
|---|---|---|---|
| Nano Banana, the image model in the Gemini app | Edits of your own photo, seasonal portraits, posters | Full sentences: subject, composition, action, location, style | Blends up to 14 images, depending on where you use it |
| GPT image | Posters, covers and layouts with a lot of text | The intended use first, then the scene, the subject, the details and the constraints | Refer to each input image by number and say how they combine |
| Flux Kontext | One precise change to an existing photo | An instruction: change, replace, add, keep | Name the person by how they look and list what stays |
| Grok Imagine | Fast stylised stills and quick animation | Short and direct: the subject plus a style | An uploaded picture becomes the first frame; ratios from 1:1 to 9:16 |
| Seedance | Video from text or from a picture | Subject, action, scene and one camera move, in 60 to 100 words | Start from a still when the look must stay fixed |
From the model makers' own guides and tutorials published in the last twelve months, and our own run on 8 October 2026.
How do you get AI to spell text correctly?
Our manga cover ladder is the hard case: Japanese lettering, a Latin subtitle and a starburst on one cover. The last step named each piece of text, where it sits and how many times it appears, and asked for everything else on the cover to be too small to read, so the model has no reason to invent words to fill the space. Small text is still the weak spot of every model, and Google says as much about its own.

A finished weekly manga magazine cover, portrait 4:5. Printed look: visible halftone dots in the midtones, slight ink misregistration on the reds, faint newsprint paper texture over everything. STRICT PALETTE: signal red, acid yellow, black, cream newsprint. No other colors. ILLUSTRATION fills the cover: a girl with a short black bob and a torn acid-yellow scarf leans hard into a turn on a red sport motorbike, knee slider nearly touching the road, grinning with her teeth showing, eyes locked on the reader. Hard cel shading, crisp black outlines, heavy spot blacks. Dense black speed lines radiate from a vanishing point behind her helmet-free head. Sparks spray from the footpeg at the lower right in red and yellow dots. Low angle, we look slightly up at her. LAYOUT: MASTHEAD: the two kanji 疾風 rendered once, correctly formed, huge, top left, in black brush lettering with a red offset shadow, partly overlapped by her scarf. Under it, in small clean red Latin capitals: GALE RIDER. Top right: a yellow starburst with a black outline reading NEW SERIES in black capitals, two words exactly. Bottom edge: a solid black band with tiny cream decorative type that is too small to read. TYPOGRAPHY RULES: at most two typefaces, Latin text on straight baselines, exact spellings (疾風 once, GALE RIDER, NEW SERIES). No speech bubbles, no panel borders, no other readable text. Exactly two arms, one bike, two wheels. No watermark.
| Mistake | Kind | Fix |
|---|---|---|
| A list of tags: neon, city, night, ultra detailed | Image | Write the scene as sentences |
| No purpose stated | Image | Open with the finished thing: a poster, a cover, a product shot |
| Long text in a small space | Image | Fewer words, bigger type, each line quoted and placed |
| Several changes in one edit | Photo edit | One change per turn, check it, then the next |
| No keep list | Photo edit | Name what stays (face, skin tone, hair, pose, light, framing) and repeat it every turn |
| "She" and "him" instead of a description | Photo edit | Describe the person by look: the woman with short platinum hair |
| Several camera moves in one clip | Video | One move per clip, and a new clip for the next move |
| Describing the still again in an image to video prompt | Video | The still is the first frame: write only what moves |
More prompts to copy and adapt, each with the picture it made, are on our AI image prompts page. To try one now, open the GPT image generator or the Nano Banana image generator.
Questions and answers
Do negative prompts work?
Partly. A short exclusion list such as "no watermark, no extra text" works on many models, and OpenAI's own prompting guide uses them. For anything bigger, describe the clean state you want instead: "an empty street" rather than "no cars".
Can you use several reference images in one prompt?
Yes, on most image models. Give each image a job and refer to it by its number: "the person from image 1, the coat from image 2". Google says Nano Banana Pro can blend up to 14 images, depending on where you use it.
What aspect ratio should you use for short videos and social posts?
Vertical 9:16 for short video feeds and stories, 3:4 or 1:1 for feed posts, 16:9 for widescreen. Write the ratio into the prompt or pick it in the settings. When you turn a still into a clip, draw the still in the same ratio, because it becomes the first frame.
Is it free to try these prompts on N33 AI?
Writing and polishing a prompt in the chat is free and needs no sign-in. Drawing pictures and clips in the Studio uses energy, and a free account comes with some to start.