Skip to content
Muzea

How to Create a Consistent AI Character: The Full Method

Profile, reference face, LoRA, prompts — a step-by-step method to keep an AI character recognizable across hundreds of images.

The Muzea team · · 3 min read

Generating one great image of a character is easy. Generating a hundred where you always recognize the same person is another story. Without a method, the face drifts: the jaw softens, the eye color shifts, the freckles vanish from one photo to the next. Your audience notices right away.

Here is the method we use at Muzea, in four layers.

1. Start with a profile, not an image

The most common mistake is to open an image generator and iterate until you "land" on a face you like. The problem: you don't know why you like it, so you can't reproduce it.

Write the character first:

  • Identity: first name, age (adult), city, job.
  • Precise looks: hair color and cut, eye color, skin tone, build, distinctive marks (a mole, a scar, a piercing).
  • Style: three words that sum up their wardrobe.
  • Personality and voice: how they talk, what they love, what annoys them.

Distinctive marks are your best allies: a mole under the left eye is a detail the model can reproduce — and one your audience remembers.

Tip: in Muzea, a single sentence is enough to get this profile. You then refine it instead of starting from a blank page.

2. A locked appearance prompt

Turn the profile into an appearance prompt that describes only the character (never the scene):

photo of a 27 year old scandinavian woman, long wavy platinum blonde hair,
bright blue eyes, light freckles on the nose, small mole under the left eye,
athletic build, natural makeup

This block never changes again. Each scene is written next to it: outfit, setting, lighting, framing. If you rewrite the appearance for every image, you bring variation back in.

3. A reference kit

A prompt isn't enough: two generations of the same text give you two cousins, not the same person. You need a visual reference.

The reference kit is a set of 8 to 15 images of the character:

View Why
Front, three-quarter, profile The model understands the volume of the face
Neutral, smiling, laughing Expressions no longer distort the identity
Head-and-shoulders and full body The figure stays stable
Soft light and hard light The face holds up when lighting changes

Then pick the best 3 to 5 — the ones where the character looks most like "themselves". They serve as the reference for every new scene (InstantID- or PuLID-style techniques).

4. The LoRA: the final lock

To post every day, a visual reference quickly hits its limits: as soon as the face is small in the frame or seen from behind, the likeness drops.

The fix is to train a LoRA, a small add-on model learned from the reference kit. It "engraves" the character into the generator. Best practices:

  1. 15 to 30 varied images (angles, outfits, settings), all consistent with each other.
  2. Precise captions that describe what changes (outfit, place) and a unique trigger word for what doesn't (the character).
  3. Moderate strength at inference: too strong, and the LoRA freezes expressions and gives waxy skin.

Mistakes that break consistency

  • Switching base models midway: every model interprets the prompt its own way.
  • Overlong scene prompts that drown out the appearance.
  • Forgetting distinctive marks in the appearance prompt.
  • Approving "almost" matching images: they contaminate the kit or the LoRA.

In short

Consistency doesn't come from one good prompt, but from a chain: profile → appearance prompt → reference kit → LoRA. Each layer reduces the variation the previous one left behind.

That's exactly the chain Muzea automates. You can try the first step for free: describe your character in one sentence, and the profile and first portrait arrive in under a minute.