It can stage a lively crowded market with a natural walking pose, sari styling and believable bag placement.
What was measured
Scene compliance
How accurately the output follows the prompt — environment, attire, pose, props.
decisive for this rankingtransformation
The ranking is about using the character in varied scenes and poses, so following the requested environment, attire, pose, and props is central to success. (3 of 3 judges)
What was given, what came back
Test input: Three-quarter face portrait · image
Input — what we sent

Three-quarter face portrait
Three-quarter face reference image with medium-dark skin, tight curly hair in an updo, bindi, and a floral dress. The partial angle and softer lighting make it a harder reference than the frontal portrait and are meant to stress identity consistency.
Why this input is hard
- · Identity preservation with partial face angle
- · Hair texture and updo retention
- · Skin tone fidelity
- · Reference-image difficulty under softer lighting
Output — unretouched

Also checked on this input — same tool, 5 other criteria
Accessory & detail retention⚠ StruggledDense natural curls are not retained; the tool flattens them into straight, oily-looking hair.Accessory & detail retention✓ WorkedIt preserves curl volume and hair texture better in the market scene than it does in the interrogation scene.Expression accuracy✗ FailedThe same neutral-expression failure repeats on a second reference, so the tool still misses the prompt's angry, guarded tone.Identity preservation◐ MixedOn a harder 3/4 reference, it can keep face shape, skin tone and eyebrows close, but the hair geometry drifts enough that the result is only partly faithful.Identity preservation◐ MixedWhen the face is turned partly away in a busy scene, the output still feels like the same character but cannot be fully verified.
Provenance
- Observation
- 20c6ad99-1b5e-4ad0-9a15-b1c8e13515d7
- Evidence run
- 1dfb8fa4-f7f0-47dc-911b-7db3af467e9a
- Study
- Generate Consistent AI Characters Across Different Scenes and Poses
- Research task
- 86b96df11
- Tested at
- not recorded
- Source
- first-party
- Evidence state
- verified
- Proof shown
- input + output shown
- Cost / latency
- not captured
- Repeat run
- not captured
- Tester
- not captured
The last three rows are honest blanks, not placeholders — our capture has no field for them yet.
Query this
get_evidence({
tool: "leonardo-ai"
})MCP · mcp.aidemos.com/api/mcp
Free with attribution.
Same input, same check — 3 other tools
measured on Scene compliance
ChatGPT✓ WorkedCleansly follows the interrogation-room prompt with the navy shirt, hands flat on the table, and bare room composition all in place.Gemini✓ WorkedThe crowded market prompt is followed strongly, with a busy outdoor background, realistic crowd density, market stalls, and a natural walking pose with tote bag.ImagineArt✓ WorkedThe interrogation-room output follows the clothing, lighting, pose, and environment accurately.
This evidence is published in
Real inputs and real outputs, no retouching · every cell queryable via API & MCP · aidemos.com