{"observation":{"id":"1429eee0-2569-4924-ade6-75a5272f1acf","tool":"pixverse-ai","tool_name":"PixVerse AI","criterion":"prompt-accuracy","criterion_name":"Prompt Accuracy","criterion_definition":"How faithfully the video follows the motion, scene, and detail instructions given in the prompt.","criterion_evidence_type":"transformation","criterion_rank_role":"decisive","criterion_rank_role_reason":"For image-to-video, following the requested motion and scene changes is central to judging whether the tool produced the intended video. (3 of 3 judges)","scenario":"3d-rendered-street-scene-image-with-forward-dolly-prompt","scenario_name":"3D rendered street scene image with forward dolly prompt","group_tag":"image-to-cinematic-video","scenario_description":"A rendered 3D street scene with multiple characters and environment detail, tested using a forward-dolly cinematic prompt with crowd motion, cart movement, atmospheric haze, and sunset lighting.","modality":"image","input_text":null,"input_artifact_refs":[{"alt":null,"url":"https://d3epheqghktydj.cloudfront.net/luma-ai-dream-machine-3d-image-input-2-8bec799b14bc.png","role":"input"}],"stresses":["Complex multi-subject scene animation","Character consistency in crowd motion","Camera movement stability during forward dolly","Environmental motion such as clouds, trees, and birds","Prompt adherence across foreground and background elements"],"verdict":"mixed","score":null,"score_total":null,"note":"The 3D run delivers the requested camera motion and sound effects, but the report also says clarity is reduced and characters distort, so the scene instructions are only partially preserved.","evidence_state":"verified","source":null,"artifacts":[{"url":"https://d3epheqghktydj.cloudfront.net/pixverse-ai-pixverse-ai-3d-image-output-1fc102a5d3b2.mp4","role":"output","alt":null}],"run_id":"b8778c3c-4fb1-48b2-bf35-4e08a3b82d8e","study_title":"Generate a cinematic AI video from a single image","study_kind":"generation","research_task":"86b94urgr","tested_at":null,"completeness":"input-and-output","input":{"state":"files","text":null,"files":[{"url":"https://d3epheqghktydj.cloudfront.net/luma-ai-dream-machine-3d-image-input-2-8bec799b14bc.png","filename":"luma-ai-dream-machine-3d-image-input-2-8bec799b14bc.png","alt":"3D rendered street scene image with forward dolly prompt","role":"input"}],"modality":"image","stresses":["Complex multi-subject scene animation","Character consistency in crowd motion","Camera movement stability during forward dolly","Environmental motion such as clouds, trees, and birds","Prompt adherence across foreground and background elements"]},"tool_page_slug":null,"tool_url":null,"permalink":"https://aidemos.com/evidence/1429eee0-2569-4924-ade6-75a5272f1acf","api_url":"https://ai.aidemos.com/v1/observations/1429eee0-2569-4924-ade6-75a5272f1acf"},"peers":[{"id":"8ddce692-ccba-4012-8f3d-714091bed33e","tool":"google-flow","tool_name":"Google Flow","verdict":"mixed","score":null,"score_total":null,"note":"It can reproduce the core movement language of a forward-dolly street prompt, especially natural walking and interactions, but the report does not verify every requested environmental detail.","artifact_count":1,"thumbnail":"https://d3epheqghktydj.cloudfront.net/google-flow-3d-output-478f5388b5c0.mp4","evidence_url":"https://aidemos.com/evidence/8ddce692-ccba-4012-8f3d-714091bed33e"},{"id":"293512d4-0fe0-47d3-a867-42336caaf9a4","tool":"leonardo-ai","tool_name":"Leonardo AI","verdict":"struggled","score":null,"score_total":null,"note":"The scene animates, but the report says the requested prompt details are not properly reflected.","artifact_count":1,"thumbnail":"https://d3epheqghktydj.cloudfront.net/leonardo-ai-3d-output-b445b62c2610.mp4","evidence_url":"https://aidemos.com/evidence/293512d4-0fe0-47d3-a867-42336caaf9a4"}],"other_criteria":[{"id":"0eebc77a-94e0-4e97-96e7-44fd48cbd60c","criterion":"visual-consistency-no-distortion","criterion_name":"Visual Consistency (No Distortion)","rank_role":"decisive","verdict":"struggled","score":null,"score_total":null,"note":"The 3D output shows noticeable facial distortion in characters and reduced clarity compared with the input, so structural fidelity is only partial.","artifact_count":1,"evidence_url":"https://aidemos.com/evidence/0eebc77a-94e0-4e97-96e7-44fd48cbd60c"}],"appears_in":[{"page_type":"ranking","slug":"image-to-video-generators","title":"Best AI Tools for Generating Cinematic Video from a Single Image","url":"https://aidemos.com/best/image-to-video-generators","binding":"run"},{"page_type":"use-case","slug":"generate-ai-video-from-image","title":"Generate a Cinematic AI Video from a Single Image","url":"https://aidemos.com/use-cases/generate-ai-video-from-image","binding":"run"}],"same_scenario":[{"id":"8ddce692-ccba-4012-8f3d-714091bed33e","tool":"google-flow","tool_name":"Google Flow","verdict":"mixed","score":null,"score_total":null,"note":"It can reproduce the core movement language of a forward-dolly street prompt, especially natural walking and interactions, but the report does not verify every requested environmental detail."},{"id":"293512d4-0fe0-47d3-a867-42336caaf9a4","tool":"leonardo-ai","tool_name":"Leonardo AI","verdict":"struggled","score":null,"score_total":null,"note":"The scene animates, but the report says the requested prompt details are not properly reflected."}]}