How to Write Midjourney Prompts That Actually Work

6-Layer Midjourney Prompt Framework Concept Subject Colors Composition Lighting Camera

What you will learn

  • Most prompts fail for 3 reasons - missing layers beyond the subject, vague adjectives, and skipped visual language
  • The 6-layer framework - concept, subject, colors, composition, lighting, camera - turns vague prompts into specific ones
  • Replace vague words like "blue dress" with specific language like "deep cerulean silk midi dress with cool purple undertones"
  • See a direct before-and-after comparison of a prompt with and without the framework applied
  • Learn which Midjourney parameters matter most and the common mistakes to avoid

Quick Answer

Use the 6-layer framework - describe the content type, subject, colors and materials, composition, lighting, and camera settings. Each layer adds specificity that AI image generators respond to. More detail equals more accurate results. This guide is current for Midjourney V8.2, the default model as of July 24, 2026.

Updated August 2026 - Midjourney V8.2 is now the default model as of July 24, 2026

All --v examples in this guide have been updated to --v 8.2, the current default. See Midjourney's official documentation for the latest on what changed.

In this guide

  1. Why most prompts fail
  2. The 6-layer prompt framework
  3. Putting it all together
  4. Without and with the framework
  5. The quick test
  6. The Midjourney parameters that matter
  7. Common mistakes to avoid
  8. Try it now

This guide is part of our complete Midjourney prompt guide.

Most people write Midjourney prompts like search queries.

'a woman in a red dress outside'

Then they wonder why the result looks generic, flat, and nothing like what they imagined.

The problem is not Midjourney. The problem is that AI image generators are not search engines. They are creative translators. And like any translator, they can only work with what you give them.

Give them vague input, you get vague output. Give them specific, structured input, you get images that look like they came from a professional photoshoot.

Give them vague input, you get vague output.

This guide teaches you the exact framework that separates generic prompts from great ones.

How we tested this

We tested over 200 prompts across 6 content categories in Midjourney V8.2 to build and verify the 6-layer framework. Each layer was tested in isolation and in combination to measure its individual impact on output quality.

Why most prompts fail

There are three reasons most Midjourney prompts produce disappointing results:

  • They describe the subject but nothing else. No lighting, no composition, no camera, no mood.
  • They use vague adjectives. 'Nice', 'beautiful', 'good lighting' mean nothing to an AI.
  • They skip the visual language. 'Blue dress' tells AI almost nothing. 'Deep cerulean silk midi dress with cool purple undertones' tells it everything.

The fix is a framework called the 6-layer prompt structure.

The 6-layer prompt framework

Every great Midjourney prompt is built from 6 layers, written in this exact order: concept (what kind of image this is), subject (the person, product, or scene in full detail), colors and materials (specific descriptive language, not hex codes), composition (subject position and framing), lighting (source, direction, and quality), and camera and lens (focal length and aperture).

Concept frames everything that follows. Subject and colors give Midjourney the specific detail it needs instead of inventing details randomly. Composition and lighting are the two layers most prompts skip entirely, which is why so many AI images look flat and centered. Camera and lens is what makes the final image look like it was shot by a professional rather than generated by software.

We cover each layer in full depth, with bad-example/good-example pairs for every one, in The 6-Layer AI Prompt Framework Explained - the platform-agnostic version of this same methodology. What follows here is how those 6 layers apply specifically inside Midjourney, including the parameters that refine them.

Putting it all together

Here is the same basic idea written without the framework and with it:

Without and with the framework

Without the framework

'a skincare product on a surface with nice lighting'

With the 6-layer framework

'Product still life of a dark amber glass serum bottle with gold dropper, centered on a white marble surface with warm gray veining, soft diffused natural window light from the left with warm golden undertones, shallow depth of field, shot on 100mm macro lens at f/4.0, clean minimal background with generous negative space --ar 4:5 --v 8.2 --stylize 200'

The second prompt will produce a result that looks like it belongs in a premium skincare campaign. The first will produce something forgettable.

The quick test

Before you run any prompt, read it aloud. Ask yourself: could someone with no visual context build this scene in their head from your words alone?

If the answer is no - add more detail. Specificity is not optional in AI prompting. It is the entire job.

Specificity is not optional in AI prompting. It is the entire job.

The Midjourney parameters that matter

Once your prompt text is strong, these parameters refine the output:

  • --ar sets the aspect ratio. Use 4:5 for Instagram, 1:1 for square, 16:9 for landscape, 2:3 for Pinterest.
  • --v 8.2 uses the current default Midjourney model - check Midjourney's official documentation for the latest version, since this changes periodically.
  • --stylize controls how much creative interpretation Midjourney applies. Lower values (50-100) follow your prompt literally. Higher values (400-800) produce more artistic results that drift from the prompt.
  • --raw disables Midjourney's aesthetic filter and follows your prompt more literally. Use this when you want photorealistic results without artistic enhancement.

Common mistakes to avoid

  • Saying 'realistic' instead of specifying a camera and lens. The word realistic means nothing. A Canon lens specification means everything.
  • Using 'beautiful' or 'stunning' as descriptors. These are opinions, not visual instructions. Replace them with specific visual details.
  • Skipping lighting entirely. Midjourney defaults to flat even lighting when you give it no direction. This is why AI images often look shadowless and artificial.
  • Writing the prompt as a sentence. AI prompts work better as structured descriptive phrases, not grammatically correct sentences.
  • Stacking every parameter without knowing what it does. Adding --stylize 1000 --chaos 80 --weird 500 "just in case" produces unpredictable, often unusable results. Before: a prompt buried under five untested parameters. After: --ar 4:5 --v 8.2 --stylize 200 - only the parameters you can explain the effect of.
  • Repeating the same descriptor for two different things. Using "clean" for both the background and the lighting forces Midjourney to guess which one you mean. Before: 'clean product, clean lighting, clean background'. After: 'matte white background, soft diffused overhead lighting, glass bottle with a polished finish' - one distinct word per element.

Already have a prompt? Optimize it for your specific platform in one click.

The Platform Optimizer rewrites any prompt for Instagram, Pinterest, TikTok, Amazon, and more - correct aspect ratio, composition adjustments, and platform-native visual language, with your subject and brand descriptors kept intact.

Optimize my prompt

Try it now

The fastest way to apply this framework is to use a tool that builds the 6 layers for you automatically.

Our free Midjourney Prompt Builder lets you select each layer from dropdowns and watch the prompt assemble in real time. When you are done, hit Enhance with AI and get 3 variations - subtle, standard, and cinematic - all built on the 6-layer framework.

Or if you already have a prompt that is not working, paste it into our Prompt Fixer. It diagnoses which layers are missing and rewrites the prompt completely.

Try the Prompt Fixer →

Frequently asked questions about writing Midjourney prompts

It is a structure of six layers written in order - content type, subject, colors and materials, composition, lighting, and camera and lens. Each layer adds specificity, and together they turn a vague prompt into one that produces professional-looking results.

Generic results usually come from vague adjectives like "nice" or "beautiful", missing lighting direction, and prompts that describe only the subject with no composition or camera details. Midjourney defaults to flat, centered, evenly-lit images when it isn't given more to work with.

--ar for aspect ratio, --v 8.2 for the current model version, --stylize to control how much creative interpretation is applied, and --raw to disable the aesthetic filter for more literal, photorealistic output.

Enough that someone with no visual context could build the scene in their head from your words alone. That's the quick test - read your prompt aloud, and if it leaves gaps, add more specific detail rather than relying on general descriptors.

No. Structured descriptive phrases separated by commas work better than grammatically correct sentences. Midjourney parses phrases as discrete instructions, which gives you more precise control over each layer.