It fails to map an angry or guarded prompt onto the face; the output stays neutral or calm and even reads with a slight smile.
What was measured
Expression accuracy
Whether the emotional tone and facial expression match what was explicitly prompted.
decisive for this rankingtransformation
If the character’s intended emotion or facial expression does not match the prompt, the generated character is not being controlled reliably across scenes. (3 of 3 judges)
What was given, what came back
Test input: Full frontal portrait · image
Input — what we sent
Input not captured
This run recorded no prompt or input file for the test, so we cannot show you what produced the result below. Capture gaps are tracked, not hidden.
Full frontal portrait reference image with fair skin, curly dark hair, bindi, gold jhumka earrings, and a green stone necklace. All features are clearly visible in good natural lighting, making it the easiest identity anchor for the tools.
Why this input is hard
- · Baseline identity preservation
- · Accessory retention
- · Best-case frontal face matching
- · Consistent character reuse across varied scenes
Output — unretouched


Also checked on this input — same tool, 2 other criteria
Identity preservation✓ WorkedThe tool can keep a subject recognisably the same person in an action scene, with face shape, eyes, and overall look staying close to the reference.Scene compliance◐ MixedIt can reproduce the interrogation-room setup and wardrobe, but the requested harsh mood is softened because the lighting and facial affect stay gentle.
Provenance
- Observation
- 6a1796e6-76aa-4a88-84d2-ba870e97737e
- Evidence run
- 1dfb8fa4-f7f0-47dc-911b-7db3af467e9a
- Study
- Generate Consistent AI Characters Across Different Scenes and Poses
- Research task
- 86b96df11
- Tested at
- not recorded
- Source
- first-party
- Evidence state
- verified
- Proof shown
- output only
- Cost / latency
- not captured
- Repeat run
- not captured
- Tester
- not captured
The last three rows are honest blanks, not placeholders — our capture has no field for them yet.
Query this
get_evidence({
tool: "scenario"
})MCP · mcp.aidemos.com/api/mcp
Free with attribution.
Same input, same check — 4 other tools
measured on Expression accuracy
ChatGPT✗ FailedMisses the prompted brave/determined emotional tone and instead outputs a soft neutral expression, leaving the scene with essentially no emotional alignment.Gemini✓ WorkedThe tool can capture a prompted angry and guarded mood, with direct eye contact and an intense expression in the interrogation shot.ImagineArt✓ WorkedThe interrogation-room output matches the requested serious, guarded expression.Leonardo AI✗ FailedIt does not translate an explicitly angry or guarded prompt into facial expression, defaulting to calm neutrality.
This evidence is published in
Real inputs and real outputs, no retouching · every cell queryable via API & MCP · aidemos.com