Executes the market scene strongly: the crowd feels authentic, the mustard sari and red blouse match the prompt, and the jute bag with vegetables is present.
What was measured
Scene compliance
How accurately the output follows the prompt — environment, attire, pose, props.
decisive for this rankingtransformation
The ranking is about using the character in varied scenes and poses, so following the requested environment, attire, pose, and props is central to success. (3 of 3 judges)
What was given, what came back
Test input: Three-quarter face portrait · image
Input — what we sent
Input not captured
This run recorded no prompt or input file for the test, so we cannot show you what produced the result below. Capture gaps are tracked, not hidden.
Three-quarter face reference image with medium-dark skin, tight curly hair in an updo, bindi, and a floral dress. The partial angle and softer lighting make it a harder reference than the frontal portrait and are meant to stress identity consistency.
Why this input is hard
- · Identity preservation with partial face angle
- · Hair texture and updo retention
- · Skin tone fidelity
- · Reference-image difficulty under softer lighting
Output — unretouched

Also checked on this input — same tool, 4 other criteria
Accessory & detail retention✓ WorkedRetains the bindi and strong brows in the frontal output, while only slightly darkening skin tone relative to the reference.Expression accuracy✓ WorkedDelivers the requested cold, guarded, unsmiling expression with the exact vibe the prompt asked for.Identity preservation✗ FailedBecomes very hard to verify once the face turns too far, and the report says the bindi disappears entirely, leaving the identity lock extremely weak.Identity preservation✓ WorkedPreserves face shape and structure well in a frontal composition, though the report notes a slight darkening and a marginally wider, rounder face than the reference.
Provenance
- Observation
- c3fba2ff-eb7e-4ef1-9658-7af7cce5d52c
- Evidence run
- 1dfb8fa4-f7f0-47dc-911b-7db3af467e9a
- Study
- Generate Consistent AI Characters Across Different Scenes and Poses
- Research task
- 86b96df11
- Tested at
- not recorded
- Source
- first-party
- Evidence state
- verified
- Proof shown
- output only
- Cost / latency
- not captured
- Repeat run
- not captured
- Tester
- not captured
The last three rows are honest blanks, not placeholders — our capture has no field for them yet.
Query this
get_evidence({
tool: "chatgpt"
})MCP · mcp.aidemos.com/api/mcp
Free with attribution.
Same input, same check — 3 other tools
measured on Scene compliance
Gemini✓ WorkedThe interrogation-room prompt is followed well, including the table, plain wall, overhead light, formal clothing, and clinical lighting.ImagineArt✓ WorkedThe market output renders the market environment realistically and follows the prompt well.Leonardo AI◐ MixedIt reproduces the formal outfit and room, but the lighting becomes too bright and clean for an interrogation scene.
This evidence is published in
Real inputs and real outputs, no retouching · every cell queryable via API & MCP · aidemos.com