AI Image Prompt Cheat Sheet: Stable Diffusion and Midjourney Basics

AI image prompts work best as a structured stack, subject, style, lighting, composition, and technical modifiers, rather than one long descriptive sentence. Stable Diffusion and Midjourney read prompts differently under the hood, but the same basic building blocks improve results on both: be specific about what you want, and layer the details rather than dumping them all into one run-on phrase.

Tarquin - Prompt Operations Manager AI Skill
Every prompt as a versioned artifact
Tarquin - Prompt Operations Manager AI Skill
Drop this into Claude and get prompts treated like production assets: versioned, tested, and tracked instead of reinvented from scratch every time.
$29
Get the skill →

The basic prompt stack for image generation

Layer Example
Subject "A red fox sitting in tall grass"
Style "Watercolor illustration" or "photorealistic"
Lighting "Golden hour lighting" or "soft studio lighting"
Composition "Close-up shot" or "wide angle, rule of thirds"
Technical modifiers Aspect ratio, quality tags, negative prompts (excluding unwanted elements)

Stable Diffusion vs Midjourney: syntax basics

  1. Stable Diffusion favors comma-separated tags. Keyword-style phrases work well, and negative prompts are a separate field for excluding unwanted elements.
  2. Midjourney favors natural language with parameters. Write closer to a sentence, then append parameters like aspect ratio or stylization strength at the end.
  3. Both benefit from specific style references. Naming an art style, medium, or era narrows the output more reliably than vague adjectives like "beautiful" or "nice."
  4. Both reward iteration over a single perfect prompt. Generate, adjust one variable at a time, regenerate.

Common mistakes in image prompts

  • Overloading one prompt with contradictory styles. "Photorealistic cartoon watercolor" confuses the model rather than blending the styles.
  • Vague quality descriptors instead of specifics. "High quality" does less than naming an actual lighting setup or composition style.
  • Skipping negative prompts. Explicitly excluding common artifacts (extra limbs, blurry edges) often improves results more than adding positive detail.

Related reading: prompt engineering 101. See also best AI chatbot alternatives to ChatGPT.

Frequently asked questions

Do Stable Diffusion and Midjourney use the same prompt syntax?

No, Stable Diffusion favors comma-separated tags while Midjourney favors natural language with appended parameters, but the same content principles apply to both.

What's a negative prompt and why does it matter?

A negative prompt tells the model what to exclude, like extra limbs or blurry edges. It's often as important as the positive prompt for cleaning up results.

Should I write image prompts as full sentences or keyword lists?

Depends on the tool: Midjourney handles natural sentences well, Stable Diffusion often responds better to structured keyword tags.

How specific should style references be?

Very. Naming an actual art style, medium, or photographic technique narrows results far more reliably than generic words like "beautiful" or "nice."

Why does my image prompt keep producing inconsistent results?

Often too many competing style or subject instructions in one prompt. Simplify to one clear style and subject, then adjust iteratively.

The bottom line

Good AI image prompts layer subject, style, lighting, composition, and technical modifiers rather than cramming everything into one sentence. Learn the syntax quirks of your specific tool, use negative prompts to clean up results, and iterate one variable at a time.

Browse all AI agents and LLM ops skills in the Claude skills collection at KissMySkills.

~/get-started

役立つSkills。無駄なし。

ストア内のすべてのスキル、promptパック、agentを閲覧する。

すべてのスキルを見る →または無料ツールを試す