Midjourney guide: how to create cinematic key art, style systems, and campaign visuals in Phygital+

When I use Midjourney seriously, I am rarely looking for a “correct” image first.

I am looking for taste.

Midjourney is still one of the strongest models when the brief needs cinematic atmosphere, stylized campaign key art, or a mood that feels art-directed out of the box. That is different from models optimized mainly for product cleanup, vector export, or literal prompt obedience.

According to Midjourney’s own docs, the current default is V8.2 (default since July 24, 2026), with an update focused on aesthetics, image quality, and Personalization. Earlier, V8.1 and V7 shaped the modern workflow with HD output, Draft Mode, and Omni Reference.

In this guide I want to focus on four practical questions:

  • which Midjourney version to use right now
  • what Midjourney is actually good at
  • how to prompt it like an art director
  • how to use it inside a real Phygital+ workflow
Midjourney AI image generation hero visual

Midjourney versions compared

Midjourney makes more sense once you stop treating every version as a small quality bump.

Version Best use Why I would use it
V8.2 default production generation current default; aesthetics, image quality, Personalization focus
V8.1 HD / recent production baseline strong recent default before V8.2; useful when a workflow was tuned on it
V7 Draft Mode + Omni Reference workflows introduced Draft Mode and Omni Reference; still relevant for reference-heavy series
Niji anime / illustration direction use when the brief is explicitly anime-styled rather than general cinematic realism

My working rule

  • start on the current default for most campaign and concept work
  • use Draft Mode early when I need breadth before quality
  • lock style with --sref / moodboards when the series must stay consistent
  • use Omni Reference when the same character or object must survive across frames

What Midjourney is actually good at

I would not describe Midjourney as the best model for every production cleanup task.

I would describe it as a taste-first image model.

1. Aesthetic quality out of the box

What: cinematic composition, lighting, and stylization without over-prompting.
When to use: campaign key art, moodboards, concept frames, editorial looks.
Why it matters: the first useful image often arrives faster than with more literal models.

2. Style systems that persist

What: Style Reference, moodboards, and Personalization profiles.
When to use: multi-image campaigns that need one visual language.
Why it matters: one beautiful frame is easy; a consistent series is the real job.

3. Character / object anchoring

What: Omni Reference for likeness and form continuity.
When to use: recurring characters, product heroes, cast consistency.
Why it matters: prompt wording alone usually fails across a sequence.

4. Fast iteration via Draft Mode

What: cheaper/faster drafts before full-quality enhance.
When to use: early exploration, director-style branching.
Why it matters: it changes the economics of searching for the right direction.

How I think about prompting Midjourney

If Recraft feels like briefing a graphic designer, Midjourney feels more like briefing an art director.

I start with:

  1. subject and role in the frame
  2. camera and composition
  3. lighting and material mood
  4. style constraints
  5. what must stay out

A weak prompt sounds like:

beautiful cinematic portrait, highly detailed, 8k

A stronger prompt sounds like:

Stronger brief
medium shot of a product designer in a quiet concrete studio, holding a matte black prototype speaker, soft north-window daylight, shallow depth of field, muted graphite and warm oak palette, editorial campaign still, restrained color grade, no logo text, no cluttered desk props

Prompt structure table

Subject Action / pose Scene Camera Lighting Style lock
who/what what they do environment shot size / lens feel time + quality of light sref / mood / constraints

Example prompts

Campaign portrait

Close-up portrait of a runner at dawn on wet asphalt, breath visible in cold air, telephoto compression, soft rim light from sunrise, muted teal and amber grade, athletic campaign still, natural skin texture, no text overlays

Product key art

Hero still of a ceramic pour-over kettle on raw linen, top-three-quarter angle, soft diffused daylight, quiet luxury product photography, desaturated earth tones, generous negative space, no prop clutter

Fashion editorial

Full-body fashion editorial of a model in structured charcoal coat walking through foggy harbor docks, handheld documentary feel, cool overcast light, filmic grain, long stride, restrained color, no logo text

Concept world

Wide establishing shot of a floating greenhouse city above a desert at dusk, glass and copper architecture, volumetric haze, cinematic sci-fi concept art, grounded materials, no UI overlays

Brand moodboard

Flat lay moodboard of stone samples, linen swatches, and a single unfinished clay vase, soft overhead daylight, quiet interior design aesthetic, neutral palette, precise arrangement, no typography

Character continuity

Same young architect with short curly hair and round glasses, standing beside a blueprint table in a sunlit loft, consistent facial features, medium shot, warm afternoon light, architectural editorial style

A practical Midjourney workflow inside Phygital+

  1. Generate broad directions quickly, then lock the winner.
  2. Reuse style references so the campaign stays coherent.
  3. If character identity matters, keep the same reference path across frames.
  4. Move approved stills into cleanup, upscaling, or video nodes on the same canvas.
  5. Only then expand into motion or layout systems.

Try it from the Midjourney model page or jump into the workspace.

When I would not use Midjourney

  • I need native editable SVG / logo geometry
  • I need the most literal product-photo obedience at the lowest cost
  • the job is heavy multi-reference commercial editing with many product SKUs
  • I need dense, perfectly typeset long-form text inside the image

FAQ

What is Midjourney?

Midjourney is a text-to-image AI system from Midjourney, Inc. known for cinematic, stylized, and photorealistic images from natural-language prompts, with style/character consistency tools and image-to-video workflows.

Which Midjourney version should I use?

As of July 24, 2026, Midjourney’s docs list V8.2 as the default, focused on aesthetics, image quality, and Personalization. Use V7 when you specifically need Omni Reference / Draft Mode workflows that still rely on that version’s feature set.

What is Draft Mode good for?

Draft Mode is for fast branching. Official Midjourney materials describe it as much faster and cheaper than full-quality generation, so it is best for exploring directions before enhancing a winner.

What are Style Reference and Omni Reference?

Style Reference (–sref) transfers look and feel from reference images. Omni Reference (–oref), introduced around V7, anchors a person, object, or form across generations more reliably than prompt wording alone.

Can I use Midjourney commercially in Phygital+?

On paid Phygital+ plans, outputs are covered by commercial-use licensing under platform terms. Always check the current model card and plan terms for provider-specific limits.

Why run Midjourney inside Phygital+?

Because it sits on the same canvas as 30+ other models. You can generate key art, then continue into editing, upscaling, or video without rebuilding the workflow in a separate tool stack.


Explore more