Skip to main content
WorkCrafter logoWorkCrafter.online
Tutorials

The AI Image Tool: Choosing the Right Engine

A guide to WorkCrafter's AI image models — FLUX, ZImage, Seedream, GPT Image, Gemini and more — when to use each, prompt structure, aspect ratios, credit costs and tips.

7 min read

By the WorkCrafter team · how we write these guides

A luminous artist palette dissolving into swirling generative pixels
Image generated with WorkCrafter AI

WorkCrafter's image tool offers more engines than any other tool, because different jobs want different looks. This guide explains what each engine is for so you stop guessing, plus the prompt habits that make any of them perform.

The models, and when to use them

You pick a model per generation — the name tells you the trade-off between speed, quality and style:

  • Automatic (recommended): a solid all-round model chosen for you. Start here.
  • FLUX.1 Schnell: the everyday workhorse — fast and versatile for most images.
  • ZImage Turbo: ultra-fast instant drafts when you're exploring ideas.
  • FLUX.2 Klein: a step up in detail and coherence for images you'll actually use.
  • ZAnime: stylised, non-photographic and anime art.
  • Gen-4 Image, Gemini Image 3.1 Flash, Gemini Image 3 Pro: premium models with richer detail and finish for hero images.
  • Seedream 5 Pro / Seedream 5 Lite: photoreal — when you need something that reads as a photograph.
  • GPT Image 2: follows the brief closely when the exact composition matters more than flair.
  • Gemini 2.5 Flash: quick, capable drafts.
  • Soul: cinematic and characterful, with a strong sense of style.
  • Reve: another capable all-rounder — worth trying when FLUX and Seedream both miss what you were after.
Draft on a fast model, finalise on a premium one. You'll waste far fewer credits than polishing every throwaway idea at top quality.WorkCrafter

A prompt structure that works on every engine

The model sets the ceiling; the prompt decides whether you reach it. Cover five elements in roughly this order:

  1. Subject: what is in the image.
  2. Context or action: where it is or what it's doing.
  3. Style: photorealistic, watercolor, 3D render, flat illustration, cinematic.
  4. Lighting and mood: golden hour, soft studio light, moody, high contrast.
  5. Technical: aspect ratio, lens, colour palette.
A ceramic coffee cup on a windowsill in morning light, photorealistic, soft warm backlight, shallow depth of field, 3:2 aspect ratio

What image engines are still bad at

  • Text in the image — it reads as nonsense. Add real text afterwards in a design tool.
  • Hands, teeth and precise anatomy, especially small or in motion.
  • Exact counts — "five people" is a hint, not an instruction.
  • The same character twice — each generation is independent, so consistency across images is the hardest ask.

Design around these rather than fighting them: crop out hands, keep on-image text in your editor, prefer compositions where an exact count doesn't matter.

Tips that raise your hit rate

  • Name a lighting style explicitly — it does more for mood than any other word.
  • Reuse the same style keywords across a set so the images look like a family.
  • If an unwanted element keeps appearing, name it to remove it rather than describing harder around it.
  • Change one word at a time so you learn what each does.
  • Keep a note of prompts that work — your prompt library is a real asset.

Lighting is the word that changes everything

If you can only add one thing to a prompt, add the lighting. It does more to separate a professional-looking image from an obviously generated one than resolution, detail words or style names — because lighting is what your eye reads as "a real photograph of a real place".

  • Soft diffused light — flattering, commercial, safe. Good default for products and people.
  • Golden hour backlight — warm, cinematic, generous with atmosphere.
  • Hard directional light — dramatic, strong shadows, high contrast.
  • Overcast — even and neutral, good when the subject must stay readable.
  • Practical light (lamps, screens, neon) — grounds a night scene and adds colour for free.

Consistency across a set

Getting one good image is easy. Getting six that look like they belong together is the real skill, and it comes down to holding things constant. Write your prompt as a stable block plus one variable — same style, same lighting, same palette, same lens; change only the subject.

[SUBJECT], flat vector illustration, bold simple shapes,
muted teal and sand palette, subtle grain, no outlines, 3:2

Keep that block in a note and swap only the first line. Fixing the seed as well makes the family resemblance stronger still.

Pick the ratio before the engine

Every engine listed above frames for the canvas it is handed, so the ratio is part of the brief rather than an export setting — and it is worth deciding before you decide which model to use, because a tall canvas suits a character portrait while a wide one suits the cinematic engines. Match it to where the image will actually live:

The same subject framed differently for a banner, a square and a story
Generate at the final ratio — the engine frames for the canvas you ask for.
  • 16:9 — video thumbnails, website heroes, presentation slides.
  • 1:1 — social posts, avatars, product tiles.
  • 9:16 — stories, reels, phone wallpapers.
  • 3:2 or 4:5 — editorial photography and print-like framing.

Reading a bad result

When an image misses, the fault is usually diagnosable rather than random.

  • Looks like stock photography — the prompt was generic. Add an unusual angle, specific light, a named style.
  • Cluttered and confused — too many competing instructions. Cut, do not add.
  • Right idea, wrong mood — change the lighting line before anything else.
  • Subject cropped oddly — you generated in the wrong ratio for the composition you described.
  • Garbled text in the image — expected. Add real text afterwards in a design tool.

Start from the job, not the model list

Fourteen names is too many to hold in your head, and picking by reputation is how people end up paying premium rates for a draft. Work backwards from the job instead — the answer is usually one of six:

  • **A product shot that has to look photographed** — Seedream 5 Pro, or GPT Image 2 when the composition is fixed and must be obeyed exactly.
  • **A hero image for a page or thumbnail** — Gen-4 Image or a Gemini Image 3 model. This is the one place the premium rate is worth it, because the image carries the click.
  • **A set of blog or social illustrations** — FLUX.1 Schnell with a fixed style block. Consistency across the set matters more than peak quality on any one of them.
  • **Something stylised or character-led** — ZAnime for anime, Soul when you want atmosphere and a sense of authorship.
  • **Twenty ideas in five minutes** — ZImage Turbo. You are looking for a direction, not a deliverable.
  • **You genuinely do not know** — Automatic, then move up only if the result is close but not sharp enough.

When to stop iterating and switch engines

The most common way to waste credits is loyalty to a model that is not going to get there. Two or three attempts on the same prompt tell you which situation you are in.

  • The composition is right and the finish is soft — the prompt works, the engine is the ceiling. Re-run the same prompt on a premium model.
  • Every attempt misreads the same word — that is a prompt problem, and a different engine will misread it too. Rewrite before you re-run.
  • Results swing wildly between attempts — the prompt is under-specified. Add lighting and style before changing anything else.
  • It is close but never quite right after four tries — stop. Generate the nearest thing that works and fix the rest in an image editor.

What it costs

An image generation costs 10 credits — about $0.20 at the $5 starter rate, and less on larger packs. Failed generations are refunded automatically, so trying a fast engine before committing to a premium one is essentially free.

Frequently asked questions

Which model gives the most realistic photos?

The Seedream models (and GPT Image 2) are tuned for photorealism. Pair them with photographic prompt language — lens, depth of field, natural lighting — for the strongest results.

My images look generic. Why?

A generic prompt lands on the average of the training data. Specific lighting, an unusual angle and a named style pull the output away from that centre.

Can I use the images commercially?

The platform's terms grant you the output, so for most commercial use the answer is yes. Copyright on purely machine-generated work is separate and unsettled, which only matters if you need to stop others reusing it.

Start creating

Open the image tool, draft on a fast engine, then finalise your favourite idea on a premium one at the final aspect ratio. That two-step habit gives portfolio-quality images without burning credits.

A crystalline abstract sculpture emerging from chaotic noise particles
Image generated with WorkCrafter AI
#AIimagegeneratorguide#whichAIimageengine#AIimageprompttips#WorkCrafterimagetool#texttoimageengine#photorealAIimages

Keep reading

Get the next guide by email

New guides on prompting, generating and what it actually costs — a few times a month, never more. One click to unsubscribe, and we never share your address.