10 Tips for Better AI Image Generation in 2026

Practical prompt engineering for the modern image-model lineup — GPT-Image-2, Gemini 3 Pro Image, Flux-1 Krea. What still works, what's outdated, and how to iterate faster.

M

Media Moana Team

May 13, 2026

5 min read


The model lineup has changed. GPT-Image-2, Gemini 3 Pro Image, and Flux-1 Krea respond very differently from the Stable Diffusion era — most no longer benefit from "trending on artstation" tag stuffing or negative prompts. What still works is clear, specific, well-structured natural language plus knowing which model to point at which job.

Here are ten tips that hold up across today's best image models.

1. Describe the Scene, Not Just the Subject

Modern models reward natural-language descriptions of an entire scene over keyword soup.

Avoid: a dog

Better: a golden retriever puppy sitting in a field of wildflowers at sunset, soft warm light filtering through the grass, shallow depth of field, photorealistic

The more you write like a stage direction, the better today's models respond.

2. Set the Medium Explicitly

State the medium up front so the model commits to a visual register:

  • photorealistic photo
  • oil painting
  • watercolor illustration
  • 3D render
  • pencil sketch
  • cinematic still

Mixing mediums mid-prompt confuses the model. Pick one and stick with it.

3. Use Camera and Lighting Vocabulary

Photography terminology produces noticeably more polished output across GPT-Image-2, Gemini 3 Pro Image, and Flux:

cityscape at golden hour, shot on a 35mm lens at f/2.8,
shallow depth of field, warm tones, slight haze

Useful vocabulary to keep in your back pocket: focal length, aperture, lens type, time of day, color temperature, depth of field.

4. Control Composition

Spell out how elements should be arranged:

  • centered composition
  • rule of thirds
  • wide-angle establishing shot
  • close-up portrait
  • overhead / top-down view
  • low-angle hero shot

5. Set Aspect Ratio as a Parameter, Not in the Prompt

Aspect ratio is a generation parameter in every modern model — putting it in the prompt is unreliable. Media Moana's AI Generate dialog has a dedicated Aspect Ratio control. Common picks:

  • 16:9 — landscapes, cinematic scenes
  • 9:16 — portraits, mobile wallpapers, Reels/Shorts
  • 1:1 — social feed, product shots
  • 4:5 — Instagram posts

6. Pick the Right Model for the Job

Different models excel at different tasks, and a prompt that disappoints on one often shines on another:

  • GPT-Image-2 — strong at rendering legible text, follows complex instructions reliably
  • Gemini 3 Pro Image — high photorealism, excellent for editorial and product imagery
  • Flux-1 Krea — fast, flexible, very strong on artistic styles
  • Gemini 2.5 Flash Image — best when you're editing an existing image rather than generating from scratch

Media Moana's multi-model marketplace lets you swap models without changing your prompt — cost per generation is shown up front so you can A/B cheaply.

7. Reference Genres and Eras, Not Specific Artists

Most modern providers filter out named-artist mimicry, and prompts like "in the style of [living artist]" increasingly produce generic results. Reach for genres, movements, and eras instead:

landscape painting in the style of late-19th-century French
impressionism, vibrant complementary colors, loose brushwork

This is more durable across model updates — and more defensible for client work.

8. Iterate, Don't Restart

Modern models support conversational editing — Media Moana's AI Chat workspace and Gemini 2.5 Flash Image are built for it. Instead of rewriting from scratch:

  1. Generate a base image
  2. Ask for one targeted change ("warmer lighting", "remove the cup on the left", "shift the camera lower")
  3. Repeat

This converges on what you actually want much faster than re-rolling fresh generations.

9. Save Winners as Presets

When a prompt + settings combo nails the look you need, save it as a preset so you (and your team) can run it again with one click. Media Moana ships with 20+ built-in AI presets across generation, editing, upscaling, background removal, colorization, and face restoration — and you can build your own on top.

10. Know When to Stop Generating and Start Editing

Some fixes are easier in a manual editor than in another generation round — cropping, minor color tweaks, removing a stray pixel. Media Moana's built-in Photo Editor opens any generated image with Resize, Crop, and color adjustments, then saves back to the library. Treat generation and editing as two halves of one workflow.


A Worked Example

Attempt 1

a cat

Generic, low-effort, no surprises.

Attempt 2

a fluffy persian cat, studio photography

Better, but no scene or mood.

Attempt 3

a fluffy white persian cat with copper eyes, sitting upright
on a deep green velvet cushion in a Victorian sitting room,
soft afternoon light through tall windows, shallow depth of field,
shot on a medium-format camera, photorealistic, detailed fur texture,
warm muted color grading, centered composition

This is the prompt you'd save as a preset — it commits to medium, lighting, composition, and palette, and any model in the catalog will produce something usable from it.


Try It in Media Moana

Same prompt across GPT-Image-2, Gemini 3 Pro Image, and Flux-1 Krea, all in one place — and you'll find the model that fits your style faster than rewriting prompts. Cost per generation is shown before you commit.

The best prompt is usually the one you ran twice and refined the second time.