DND Character Portrait Prompt Guide

Guide · last updated June 2026

A character portrait is the most common prompt type for DND players and dungeon masters. This guide covers when a portrait is the right choice, how to compose one that actually produces usable art, and what mistakes turn a promising idea into a blurry mess.

When to use character portrait prompts

A character portrait prompt is the right tool when you need a face-forward image of a specific character — a player character avatar, an NPC who appears repeatedly in your campaign, or a villain you want to show your table. Portraits emphasize the character's identity: facial features, expression, costume details, and the lighting that sets their mood.

Portraits are the wrong choice when you need to show the character in action or in context. If you want your rogue mid-backstab, a full-body or scene prompt will serve you better. If you need a token for a virtual tabletop, the token prompt type produces top-down silhouettes that portraits cannot replicate.

The sweet spot for a portrait is any situation where recognizing the character matters more than what they are doing. Session-zero character sheets, recurring NPC handouts, campaign wiki avatars — these are all portrait territory. One-shots where the NPC appears once and never again? A brief description at the table may be more efficient than generating art.

Composition decisions and trade-offs

The most important decision in a portrait prompt is the cropping. Head-and-shoulders framing keeps the focus on the face and expression but sacrifices armor and weapon detail. Three-quarter framing shows more of the body but reduces the emotional punch of the face. Full-body shots under a "portrait" prompt type often produce stiff, neutral poses because the model tries to fit everything in — that is what the full-body type is for.

Trade-off rule: If you find yourself adding "holding a sword" and "wearing plate armor" to a portrait prompt, consider switching to full-body. The portrait type optimizes for expression and facial detail; the full-body type optimizes for pose and silhouette.

Lighting direction is the second key decision. Dramatic side lighting (Rembrandt lighting) adds mood but can hide part of the face. Even, soft lighting shows detail but may look flat. A slight warm key light from above-left with a cool fill from the right is one reliable starting point for a fantasy torchlight mood. Rim light and backlight can work as accents, but pair them with enough front fill to keep the face readable.

Background complexity is the third axis. A simple dark background isolates the character and keeps the model focused on costume and face. An environment background adds story context but invites the model to spend tokens on scenery instead of the character. If you include a background, keep it to one or two nouns: "stone wall," "forest clearing," "throne room." Avoid full scene descriptions inside a portrait prompt.

Before and after examples

Before

a tiefling warlock with red skin and horns, wearing dark robes, holding a staff, in a dark place, fantasy art

After

head-and-shoulders portrait of a tiefling warlock, crimson skin, curved horns sweeping back from the brow, amber eyes catching firelight, wearing tattered velvet robes over black leather, a gnarled ironwood staff rests against the shoulder, warm torchlight from upper left, dark stone crypt wall behind, painterly fantasy illustration, expressive face, detailed skin texture

The "before" prompt is vague about framing, lighting, and background. The model has to guess what "dark place" means, and it will often guess wrong — producing a dark, muddy image. The "after" prompt specifies the cropping, the exact lighting setup, the background as a single element, and adds texture keywords ("detailed skin texture," "painterly") that push the model toward higher-quality output.

Before

elf ranger, green cloak, bow, forest background, detailed, high quality

After

three-quarter portrait of a wood elf ranger, sun-weathered copper skin, pointed ears emerging from braided auburn hair, sharp green eyes scanning the treeline, wearing a hooded moss-green cloak over studded leather armor, a longbow slung across the back, dappled forest light filtering through canopy, simple backdrop of ancient oak bark, digital fantasy painting, warm natural lighting, focused expression

The "before" prompt is a keyword list. Models treat keyword lists as a buffet — they sample from each term without prioritizing any. The "after" prompt is a sentence that flows from the character outward, giving the model a clear visual hierarchy: face first, then costume, then environment, then style.

Common failure patterns and corrections

Failure 1: The generic fantasy face. When a prompt only says "elf wizard" or "dwarf fighter," the model has no reason to make this character look different from the thousand other elf wizards it has seen. The result is a bland, average-of-all-elves face with no distinguishing features.
Correction: Add at least one unique physical trait per prompt. "A scar across the left cheek," "one gold eye and one silver eye," "facial tattoos in druidic script." These details anchor the character in the model's output and prevent generic blending.
Failure 2: Too many background elements. Adding "standing in a tavern with patrons, mugs on the table, a fireplace, a barmaid, a cat on the bar" to a portrait prompt causes the model to spend most of its rendering budget on the environment, shrinking the character to a small figure in a crowd.
Correction: Limit the background to one or two elements. If the scene is important, use the scene prompt type instead. Portraits work best with "simple backdrop" or a single background noun.
Failure 3: Contradictory lighting. Specifying both "dramatic chiaroscuro" and "bright cheerful lighting" forces the model to compromise, producing flat mid-tone images that satisfy neither instruction.
Correction: Pick one lighting direction and stick with it. If you want mood, go with "low-key lighting, strong shadows." If you want clarity, go with "soft diffused light, even illumination." Never mix.

Model-specific guidance

Midjourney

Midjourney v6+ responds well to natural-language descriptions. Add --ar 3:4 for portrait aspect ratio — the default 1:1 square crops heads or wastes space. Use --style raw if you want less of the Midjourney "painterly gloss" and more direct adherence to your description. Avoid comma-separated keyword dumps; Midjourney v6 prefers complete sentences.

ChatGPT / DALL-E

ChatGPT image generation and DALL-E generally respond well to clear natural-language layout instructions. For a head-and-shoulders result, specify the crop and explain which body parts should remain outside the frame. Content handling can change over time, so describe the visual purpose and context plainly rather than trying to work around a particular filter.

Stable Diffusion (SDXL / SD3)

Stable Diffusion responds strongly to comma-separated tags, which is the opposite of Midjourney's preference. Lead with the strongest visual tags: 1girl, tiefling, portrait, crimson skin, curved horns, dark robes, torchlit. Add quality tags like masterpiece, best quality, highly detailed early in the prompt. Use a negative prompt to exclude common SD artifacts: lowres, bad anatomy, extra fingers, blurry, watermark.

Reviewed prompt template

[framing] portrait of a [race] [class], [skin tone] skin, [distinctive facial feature], [hair description], [eye description], wearing [costume layers from outer to inner], [weapon or held object] [resting position], [lighting direction and quality], [background element], [art style], [expression keyword], [texture keyword]

Fill the relevant brackets with concrete details and remove any bracket you do not need. Models vary in how they weight prompt order, but placing framing and identity early creates a clearer hierarchy and makes the prompt easier to revise.

Frequently asked questions

Should I include the character's name in the prompt?
No, unless the name is a real-world historical or mythological figure the model knows. AI models do not know your homebrew character "Kaelthas Sunblade" (unless it matches a Warcraft character by coincidence). Use physical descriptions instead of names. A name wastes prompt tokens on something the model cannot render.
Why does my portrait look like a painting when I asked for a photo?
Most image models are trained heavily on painted and illustrated fantasy art. The word "photo" sometimes triggers photorealism, but it can also produce uncanny-valley results when combined with fantasy races like tieflings or dragonborn. If you want a realistic look, use "cinematic still, hyperrealistic" instead of "photo." If you want an illustrated look, be explicit: "oil painting," "digital illustration," or "watercolor."
Can I use the same portrait prompt for different characters?
Only if you change the distinguishing physical features. Swapping "elf" for "dwarf" while keeping everything else the same will produce two characters in the same outfit and lighting, which may or may not be what you want. The template above is designed to force you to specify what makes each character unique.
How long should a portrait prompt be?
A useful starting range is 40-80 words. Very short prompts leave more decisions to the model, while very long prompts can make priorities harder to read. The template above usually produces roughly 50-70 words when each selected bracket contains one detail.

Related guides