Captures the cold, guarded, unsmiling expression well, with a tight jaw and sharp eyes; the report marks this as the best expression accuracy of the five scenes.
What was measured
Expression accuracy
Whether the emotional tone and facial expression match what was explicitly prompted.
decisive for this rankingtransformation
If the character’s intended emotion or facial expression does not match the prompt, the generated character is not being controlled reliably across scenes. (3 of 3 judges)
What was given, what came back
Test input: Full frontal portrait · image
Input — what we sent
Input not captured
This run recorded no prompt or input file for the test, so we cannot show you what produced the result below. Capture gaps are tracked, not hidden.
Full frontal portrait reference image with fair skin, curly dark hair, bindi, gold jhumka earrings, and a green stone necklace. All features are clearly visible in good natural lighting, making it the easiest identity anchor for the tools.
Why this input is hard
- · Baseline identity preservation
- · Accessory retention
- · Best-case frontal face matching
- · Consistent character reuse across varied scenes
Output — unretouched

Also checked on this input — same tool, 6 other criteria
Accessory & detail retention✓ WorkedPreserves all three named accessories in a frontal portrait with near-100% retention: the bindi, gold jhumka earrings, and green stone necklace stay visible, and the hair colour/texture remains consistent; the main trade-off is slightly over-smoothed skin texture.Identity preservation⚠ StruggledIdentity drops to only a 50–70% match in an action scene, with significant facial-structure drift away from the reference even though the body styling remains plausible.Identity preservation✓ WorkedKeeps facial structure close to the reference in a frontal composition and is described as the best output from this input, with identity held strongly overall.Scene compliance◐ MixedOnly partially follows the cafe scene: the subject placement is correct, but the background stays weak and the coffee cup and plate are barely visible, so the environment lacks atmosphere.Scene compliance✓ WorkedRecreates the requested desert action scene accurately, including the environment, scene-appropriate attire, and hairstyle, with no visible distortion or artifacts.Scene compliance✓ WorkedFollows the full-frontal interrogation setup precisely: the navy shirt, hands on the metal table, and plain background are all present as prompted.
Provenance
- Observation
- dff2415a-48ec-4868-bc7e-970100f82c02
- Evidence run
- 1dfb8fa4-f7f0-47dc-911b-7db3af467e9a
- Study
- Generate Consistent AI Characters Across Different Scenes and Poses
- Research task
- 86b96df11
- Tested at
- not recorded
- Source
- first-party
- Evidence state
- verified
- Proof shown
- output only
- Cost / latency
- not captured
- Repeat run
- not captured
- Tester
- not captured
The last three rows are honest blanks, not placeholders — our capture has no field for them yet.
Query this
get_evidence({
tool: "chatgpt"
})MCP · mcp.aidemos.com/api/mcp
Free with attribution.
Same input, same check — 4 other tools
measured on Expression accuracy
Gemini✓ WorkedThe tool can capture a prompted angry and guarded mood, with direct eye contact and an intense expression in the interrogation shot.ImagineArt✓ WorkedThe interrogation-room output matches the requested serious, guarded expression.Leonardo AI✗ FailedIt does not translate an explicitly angry or guarded prompt into facial expression, defaulting to calm neutrality.Scenario✗ FailedIt fails to map an angry or guarded prompt onto the face; the output stays neutral or calm and even reads with a slight smile.
This evidence is published in
Real inputs and real outputs, no retouching · every cell queryable via API & MCP · aidemos.com