The expensive way to catch avatar drift is to render twenty images, look at them side by side, and notice the third one has different hair than the first one. By that point you've burned a good chunk of your render budget.
The cheap way is a thirty-second experiment you run before the batch starts.
The experiment
Take your avatar prompt — the locked persona, the appearance block, the whole canonical thing. Open a fresh chat. Paste the prompt. Then ask: "Describe the character to me as if you were going to draw them. Be specific about every visible detail."
ChatGPT is the prompt anvil — open it and try the skeleton in this post.
The model writes back its interpretation. Read it carefully. Anywhere the model "fills in" a detail you didn't specify is a place your prompt is leaving room for drift.
If you wrote "brown hair" and the model says "medium-length wavy chestnut brown," you just learned that "brown hair" is going to render as four different things across a batch — because the model is filling in the unspecified parts and there's no guarantee it'll fill them in the same way each render.
What to do with the result
For each detail the model added on its own, decide: do I want this nailed down? If yes, copy that exact phrasing back into your bio. If no (you genuinely don't care if the hair is wavy or straight), leave it.
The point isn't to have the longest possible bio. It's to lock in the things you do care about and consciously release the things you don't. Right now, most avatar bios aren't doing either — they're vague on details the creator does care about and verbose on details that don't matter. The experiment surfaces both.
Why it's cheap
The experiment is one prompt. No image generation. No render credits. Just text. It runs in the time it takes to type the question. The savings are huge: you'll catch drift sources before any pixels render.
I run this every time I make a meaningful change to an avatar bio. Add a new accessory? Run the experiment. Change a color? Run it. Want to test a new pose direction? Run it. Thirty seconds, every time. That habit alone has cut my "this batch looks off" reshoots roughly in half.
The bonus result
If the model's description matches what you actually want the avatar to look like, you don't need to render a single test image. You can go straight to the batch. The text-based pre-check is a faithful proxy for what the model will draw, because the same model is doing both — interpreting your bio is interpreting your bio, regardless of whether it ends in "now describe" or "now render."
One prompt. Thirty seconds. Most of your drift, gone.
— Jeff
ChatGPT is the prompt anvil — open it and try the skeleton in this post.