In the crowded market scene, the character stays highly recognisable, preserving face shape, smile, eyebrows, and overall facial structure.
What was measured
Identity preservation
How closely the generated face matches the reference image across scenes.
decisive for this rankingtransformation
This is the core of the task: the tool must keep the same character looking like the reference image across scenes. (3 of 3 judges)
What was given, what came back
Test input: Three-quarter face portrait · image
Input — what we sent

Input not captured
This run recorded no prompt or input file for the test, so we cannot show you what produced the result below. Capture gaps are tracked, not hidden.
Three-quarter face reference image with medium-dark skin, tight curly hair in an updo, bindi, and a floral dress. The partial angle and softer lighting make it a harder reference than the frontal portrait and are meant to stress identity consistency.
Why this input is hard
- · Identity preservation with partial face angle
- · Hair texture and updo retention
- · Skin tone fidelity
- · Reference-image difficulty under softer lighting
Output — unretouched

Also checked on this input — same tool, 4 other criteria
Accessory & detail retention✓ WorkedA busy market scene can still preserve fine clothing and hair details, keeping the sari, blouse, hairstyle, and curly hair texture recognizable.Expression accuracy✗ FailedThe tool can miss the requested emotion entirely, producing a neutral and emotionless face instead of the prompted angry and guarded look.Scene compliance✓ WorkedThe crowded market prompt is followed strongly, with a busy outdoor background, realistic crowd density, market stalls, and a natural walking pose with tote bag.Scene compliance✓ WorkedThe interrogation-room prompt is followed well, including the table, plain wall, overhead light, formal clothing, and clinical lighting.
Provenance
- Observation
- 1a0bf38a-fcde-4dc7-977f-cf458af631b5
- Evidence run
- 1dfb8fa4-f7f0-47dc-911b-7db3af467e9a
- Study
- Generate Consistent AI Characters Across Different Scenes and Poses
- Research task
- 86b96df11
- Tested at
- not recorded
- Source
- first-party
- Evidence state
- verified
- Proof shown
- input + output shown
- Cost / latency
- not captured
- Repeat run
- not captured
- Tester
- not captured
The last three rows are honest blanks, not placeholders — our capture has no field for them yet.
Query this
get_evidence({
tool: "gemini"
})MCP · mcp.aidemos.com/api/mcp
Free with attribution.
Same input, same check — 4 other tools
measured on Identity preservation
ChatGPT✗ FailedBecomes very hard to verify once the face turns too far, and the report says the bindi disappears entirely, leaving the identity lock extremely weak.ImagineArt✓ WorkedThe interrogation-room output stays close to the reference in face shape, skin tone, expression, posture, and overall appearance.Leonardo AI◐ MixedOn a harder 3/4 reference, it can keep face shape, skin tone and eyebrows close, but the hair geometry drifts enough that the result is only partly faithful.Scenario⚠ StruggledIdentity drifts in a crowded full-body scene, with a narrower jawline and a thinner body making the result read as a similar-looking person rather than the same subject.
This evidence is published in
Real inputs and real outputs, no retouching · every cell queryable via API & MCP · aidemos.com