{"observation":{"id":"6a1796e6-76aa-4a88-84d2-ba870e97737e","tool":"scenario","tool_name":"Scenario","criterion":"expression-accuracy","criterion_name":"Expression accuracy","criterion_definition":"Whether the emotional tone and facial expression match what was explicitly prompted.","criterion_evidence_type":"transformation","criterion_rank_role":"decisive","criterion_rank_role_reason":"If the character’s intended emotion or facial expression does not match the prompt, the generated character is not being controlled reliably across scenes. (3 of 3 judges)","scenario":"full-frontal-portrait","scenario_name":"Full frontal portrait","group_tag":null,"scenario_description":"Full frontal portrait reference image with fair skin, curly dark hair, bindi, gold jhumka earrings, and a green stone necklace. All features are clearly visible in good natural lighting, making it the easiest identity anchor for the tools.","modality":"image","input_text":null,"input_artifact_refs":[],"stresses":["Baseline identity preservation","Accessory retention","Best-case frontal face matching","Consistent character reuse across varied scenes"],"verdict":"failed","score":null,"score_total":null,"note":"It fails to map an angry or guarded prompt onto the face; the output stays neutral or calm and even reads with a slight smile.","evidence_state":"verified","source":null,"artifacts":[{"url":"https://d3epheqghktydj.cloudfront.net/scenario-scenario-input1-interrogation-8a626f0b95c8.png","role":"output","alt":null},{"url":"https://d3epheqghktydj.cloudfront.net/scenario-scenario-input1-interrogation-failure-ex-fdb8037ae402.png","role":"output","alt":null}],"run_id":"1dfb8fa4-f7f0-47dc-911b-7db3af467e9a","study_title":"Generate Consistent AI Characters Across Different Scenes and Poses","study_kind":"generation","research_task":"86b96df11","tested_at":null,"completeness":"output-only","input":{"state":"not-captured","text":null,"files":[],"modality":"image","stresses":["Baseline identity preservation","Accessory retention","Best-case frontal face matching","Consistent character reuse across varied scenes"]},"tool_page_slug":"scenario","tool_url":"https://aidemos.com/tools/scenario","permalink":"https://aidemos.com/evidence/6a1796e6-76aa-4a88-84d2-ba870e97737e","api_url":"https://ai.aidemos.com/v1/observations/6a1796e6-76aa-4a88-84d2-ba870e97737e"},"peers":[{"id":"6636ced6-0dd3-4d55-aed4-228102bc03f6","tool":"chatgpt","tool_name":"ChatGPT","verdict":"failed","score":null,"score_total":null,"note":"Misses the prompted brave/determined emotional tone and instead outputs a soft neutral expression, leaving the scene with essentially no emotional alignment.","artifact_count":1,"thumbnail":"https://d3epheqghktydj.cloudfront.net/chatgpt-chatgpt-input1-horseride-dbbcd9f8517e.png","evidence_url":"https://aidemos.com/evidence/6636ced6-0dd3-4d55-aed4-228102bc03f6"},{"id":"d08a8783-fc20-4cf5-8b89-47e25f9f73b7","tool":"gemini","tool_name":"Gemini","verdict":"worked","score":null,"score_total":null,"note":"The tool can capture a prompted angry and guarded mood, with direct eye contact and an intense expression in the interrogation shot.","artifact_count":2,"thumbnail":"https://d3epheqghktydj.cloudfront.net/gemini-input-1-abf743cbaaf0.png","evidence_url":"https://aidemos.com/evidence/d08a8783-fc20-4cf5-8b89-47e25f9f73b7"},{"id":"a463c854-4e75-436e-bef8-df02f6bcfcff","tool":"imagineart","tool_name":"ImagineArt","verdict":"worked","score":null,"score_total":null,"note":"The interrogation-room output matches the requested serious, guarded expression.","artifact_count":1,"thumbnail":"https://d3epheqghktydj.cloudfront.net/imagineart-imagineart-input1-interrogation-cfe1542ae993.jpg","evidence_url":"https://aidemos.com/evidence/a463c854-4e75-436e-bef8-df02f6bcfcff"},{"id":"b57aa471-7b05-4433-a4f2-f1a621c5aaf3","tool":"leonardo-ai","tool_name":"Leonardo AI","verdict":"failed","score":null,"score_total":null,"note":"It does not translate an explicitly angry or guarded prompt into facial expression, defaulting to calm neutrality.","artifact_count":1,"thumbnail":"https://d3epheqghktydj.cloudfront.net/leonardo-ai-leonardo-input1-interrogation-77ba6593d6a0.jpg","evidence_url":"https://aidemos.com/evidence/b57aa471-7b05-4433-a4f2-f1a621c5aaf3"}],"other_criteria":[{"id":"ed66c473-a880-4d06-8d75-0990ec076849","criterion":"identity-preservation","criterion_name":"Identity preservation","rank_role":"decisive","verdict":"worked","score":null,"score_total":null,"note":"The tool can keep a subject recognisably the same person in an action scene, with face shape, eyes, and overall look staying close to the reference.","artifact_count":1,"evidence_url":"https://aidemos.com/evidence/ed66c473-a880-4d06-8d75-0990ec076849"},{"id":"bc2d761a-8541-43ff-83a0-014ce45ad8d2","criterion":"scene-compliance","criterion_name":"Scene compliance","rank_role":"decisive","verdict":"mixed","score":null,"score_total":null,"note":"It can reproduce the interrogation-room setup and wardrobe, but the requested harsh mood is softened because the lighting and facial affect stay gentle.","artifact_count":1,"evidence_url":"https://aidemos.com/evidence/bc2d761a-8541-43ff-83a0-014ce45ad8d2"}],"appears_in":[{"page_type":"ranking","slug":"consistent-ai-characters","title":"Best AI Tools for Consistent AI Characters Across Scenes and Poses","url":"https://aidemos.com/best/consistent-ai-characters","binding":"run"}],"same_scenario":[{"id":"6636ced6-0dd3-4d55-aed4-228102bc03f6","tool":"chatgpt","tool_name":"ChatGPT","verdict":"failed","score":null,"score_total":null,"note":"Misses the prompted brave/determined emotional tone and instead outputs a soft neutral expression, leaving the scene with essentially no emotional alignment."},{"id":"d08a8783-fc20-4cf5-8b89-47e25f9f73b7","tool":"gemini","tool_name":"Gemini","verdict":"worked","score":null,"score_total":null,"note":"The tool can capture a prompted angry and guarded mood, with direct eye contact and an intense expression in the interrogation shot."},{"id":"a463c854-4e75-436e-bef8-df02f6bcfcff","tool":"imagineart","tool_name":"ImagineArt","verdict":"worked","score":null,"score_total":null,"note":"The interrogation-room output matches the requested serious, guarded expression."},{"id":"b57aa471-7b05-4433-a4f2-f1a621c5aaf3","tool":"leonardo-ai","tool_name":"Leonardo AI","verdict":"failed","score":null,"score_total":null,"note":"It does not translate an explicitly angry or guarded prompt into facial expression, defaulting to calm neutrality."}]}