This guide is part of our complete Midjourney prompt guide.
Most people write Midjourney prompts like search queries.
What you will learn
Quick Answer
Use the 6-layer framework - describe the content type, subject, colors and materials, composition, lighting, and camera settings. Each layer adds specificity that AI image generators respond to. More detail equals more accurate results. This guide is current for Midjourney V8.2, the default model as of July 24, 2026.
Updated August 2026 - Midjourney V8.2 is now the default model as of July 24, 2026
All --v examples in this guide have been updated to --v 8.2, the current default. See Midjourney's official documentation for the latest on what changed.
In this guide
This guide is part of our complete Midjourney prompt guide.
Most people write Midjourney prompts like search queries.
'a woman in a red dress outside'
Then they wonder why the result looks generic, flat, and nothing like what they imagined.
The problem is not Midjourney. The problem is that AI image generators are not search engines. They are creative translators. And like any translator, they can only work with what you give them.
Give them vague input, you get vague output. Give them specific, structured input, you get images that look like they came from a professional photoshoot.
Give them vague input, you get vague output.
This guide teaches you the exact framework that separates generic prompts from great ones.
We tested over 200 prompts across 6 content categories in Midjourney V8.2 to build and verify the 6-layer framework. Each layer was tested in isolation and in combination to measure its individual impact on output quality.
There are three reasons most Midjourney prompts produce disappointing results:
The fix is a framework called the 6-layer prompt structure.
Every great Midjourney prompt is built from 6 layers, written in this exact order: concept (what kind of image this is), subject (the person, product, or scene in full detail), colors and materials (specific descriptive language, not hex codes), composition (subject position and framing), lighting (source, direction, and quality), and camera and lens (focal length and aperture).
Concept frames everything that follows. Subject and colors give Midjourney the specific detail it needs instead of inventing details randomly. Composition and lighting are the two layers most prompts skip entirely, which is why so many AI images look flat and centered. Camera and lens is what makes the final image look like it was shot by a professional rather than generated by software.
We cover each layer in full depth, with bad-example/good-example pairs for every one, in The 6-Layer AI Prompt Framework Explained - the platform-agnostic version of this same methodology. What follows here is how those 6 layers apply specifically inside Midjourney, including the parameters that refine them.
Here is the same basic idea written without the framework and with it:
Without the framework
'a skincare product on a surface with nice lighting'
With the 6-layer framework
'Product still life of a dark amber glass serum bottle with gold dropper, centered on a white marble surface with warm gray veining, soft diffused natural window light from the left with warm golden undertones, shallow depth of field, shot on 100mm macro lens at f/4.0, clean minimal background with generous negative space --ar 4:5 --v 8.2 --stylize 200'
The second prompt will produce a result that looks like it belongs in a premium skincare campaign. The first will produce something forgettable.
Before you run any prompt, read it aloud. Ask yourself: could someone with no visual context build this scene in their head from your words alone?
If the answer is no - add more detail. Specificity is not optional in AI prompting. It is the entire job.
Specificity is not optional in AI prompting. It is the entire job.
Once your prompt text is strong, these parameters refine the output:
--ar sets the aspect ratio. Use 4:5 for Instagram, 1:1 for square, 16:9 for landscape, 2:3 for Pinterest.--v 8.2 uses the current default Midjourney model - check Midjourney's official documentation for the latest version, since this changes periodically.--stylize controls how much creative interpretation Midjourney applies. Lower values (50-100) follow your prompt literally. Higher values (400-800) produce more artistic results that drift from the prompt.--raw disables Midjourney's aesthetic filter and follows your prompt more literally. Use this when you want photorealistic results without artistic enhancement.--stylize 1000 --chaos 80 --weird 500 "just in case" produces unpredictable, often unusable results. Before: a prompt buried under five untested parameters. After: --ar 4:5 --v 8.2 --stylize 200 - only the parameters you can explain the effect of.Already have a prompt? Optimize it for your specific platform in one click.
The Platform Optimizer rewrites any prompt for Instagram, Pinterest, TikTok, Amazon, and more - correct aspect ratio, composition adjustments, and platform-native visual language, with your subject and brand descriptors kept intact.
Optimize my promptThe fastest way to apply this framework is to use a tool that builds the 6 layers for you automatically.
Our free Midjourney Prompt Builder lets you select each layer from dropdowns and watch the prompt assemble in real time. When you are done, hit Enhance with AI and get 3 variations - subtle, standard, and cinematic - all built on the 6-layer framework.
Or if you already have a prompt that is not working, paste it into our Prompt Fixer. It diagnoses which layers are missing and rewrites the prompt completely.
It is a structure of six layers written in order - content type, subject, colors and materials, composition, lighting, and camera and lens. Each layer adds specificity, and together they turn a vague prompt into one that produces professional-looking results.
Generic results usually come from vague adjectives like "nice" or "beautiful", missing lighting direction, and prompts that describe only the subject with no composition or camera details. Midjourney defaults to flat, centered, evenly-lit images when it isn't given more to work with.
--ar for aspect ratio, --v 8.2 for the current model version, --stylize to control how much creative interpretation is applied, and --raw to disable the aesthetic filter for more literal, photorealistic output.
Enough that someone with no visual context could build the scene in their head from your words alone. That's the quick test - read your prompt aloud, and if it leaves gaps, add more specific detail rather than relying on general descriptors.
No. Structured descriptive phrases separated by commas work better than grammatically correct sentences. Midjourney parses phrases as discrete instructions, which gives you more precise control over each layer.