{"observation":{"id":"8ddce692-ccba-4012-8f3d-714091bed33e","tool":"google-flow","tool_name":"Google Flow","criterion":"prompt-accuracy","criterion_name":"Prompt Accuracy","criterion_definition":"How faithfully the video follows the motion, scene, and detail instructions given in the prompt.","criterion_evidence_type":"transformation","criterion_rank_role":"decisive","criterion_rank_role_reason":"For image-to-video, following the requested motion and scene changes is central to judging whether the tool produced the intended video. (3 of 3 judges)","scenario":"3d-rendered-street-scene-image-with-forward-dolly-prompt","scenario_name":"3D rendered street scene image with forward dolly prompt","group_tag":"image-to-cinematic-video","scenario_description":"A rendered 3D street scene with multiple characters and environment detail, tested using a forward-dolly cinematic prompt with crowd motion, cart movement, atmospheric haze, and sunset lighting.","modality":"image","input_text":null,"input_artifact_refs":[{"alt":null,"url":"https://d3epheqghktydj.cloudfront.net/luma-ai-dream-machine-3d-image-input-2-8bec799b14bc.png","role":"input"}],"stresses":["Complex multi-subject scene animation","Character consistency in crowd motion","Camera movement stability during forward dolly","Environmental motion such as clouds, trees, and birds","Prompt adherence across foreground and background elements"],"verdict":"mixed","score":null,"score_total":null,"note":"It can reproduce the core movement language of a forward-dolly street prompt, especially natural walking and interactions, but the report does not verify every requested environmental detail.","evidence_state":"verified","source":null,"artifacts":[{"url":"https://d3epheqghktydj.cloudfront.net/google-flow-3d-output-478f5388b5c0.mp4","role":"output","alt":null}],"run_id":"b8778c3c-4fb1-48b2-bf35-4e08a3b82d8e","study_title":"Generate a cinematic AI video from a single image","study_kind":"generation","research_task":"86b94urgr","tested_at":null,"completeness":"input-and-output","input":{"state":"files","text":null,"files":[{"url":"https://d3epheqghktydj.cloudfront.net/luma-ai-dream-machine-3d-image-input-2-8bec799b14bc.png","filename":"luma-ai-dream-machine-3d-image-input-2-8bec799b14bc.png","alt":"3D rendered street scene image with forward dolly prompt","role":"input"}],"modality":"image","stresses":["Complex multi-subject scene animation","Character consistency in crowd motion","Camera movement stability during forward dolly","Environmental motion such as clouds, trees, and birds","Prompt adherence across foreground and background elements"]},"tool_page_slug":"google-flow","tool_url":"https://aidemos.com/tools/google-flow","permalink":"https://aidemos.com/evidence/8ddce692-ccba-4012-8f3d-714091bed33e","api_url":"https://ai.aidemos.com/v1/observations/8ddce692-ccba-4012-8f3d-714091bed33e"},"peers":[{"id":"293512d4-0fe0-47d3-a867-42336caaf9a4","tool":"leonardo-ai","tool_name":"Leonardo AI","verdict":"struggled","score":null,"score_total":null,"note":"The scene animates, but the report says the requested prompt details are not properly reflected.","artifact_count":1,"thumbnail":"https://d3epheqghktydj.cloudfront.net/leonardo-ai-3d-output-b445b62c2610.mp4","evidence_url":"https://aidemos.com/evidence/293512d4-0fe0-47d3-a867-42336caaf9a4"},{"id":"1429eee0-2569-4924-ade6-75a5272f1acf","tool":"pixverse-ai","tool_name":"PixVerse AI","verdict":"mixed","score":null,"score_total":null,"note":"The 3D run delivers the requested camera motion and sound effects, but the report also says clarity is reduced and characters distort, so the scene instructions are only partially preserved.","artifact_count":1,"thumbnail":"https://d3epheqghktydj.cloudfront.net/pixverse-ai-pixverse-ai-3d-image-output-1fc102a5d3b2.mp4","evidence_url":"https://aidemos.com/evidence/1429eee0-2569-4924-ade6-75a5272f1acf"}],"other_criteria":[{"id":"420f9401-38a1-4590-b8ed-0b87ab42029c","criterion":"cinematic-enhancement","criterion_name":"Cinematic Enhancement (camera, environment, effects)","rank_role":"decisive","verdict":"worked","score":null,"score_total":null,"note":"It can sustain strong cinematic depth and warm golden-hour atmosphere through the clip.","artifact_count":1,"evidence_url":"https://aidemos.com/evidence/420f9401-38a1-4590-b8ed-0b87ab42029c"},{"id":"fc63fae5-55a9-498e-a014-7d89d9efa932","criterion":"motion-quality-realism","criterion_name":"Motion Quality & Realism","rank_role":"decisive","verdict":"worked","score":null,"score_total":null,"note":"The tool can animate a rendered street scene with smooth motion, natural walking, and scene-level interactions.","artifact_count":1,"evidence_url":"https://aidemos.com/evidence/fc63fae5-55a9-498e-a014-7d89d9efa932"},{"id":"04a73b4c-336e-4b43-9962-371922367429","criterion":"visual-consistency-no-distortion","criterion_name":"Visual Consistency (No Distortion)","rank_role":"decisive","verdict":"worked","score":null,"score_total":null,"note":"The 3D output preserved character structure cleanly enough that the report says there was no visible distortion in characters.","artifact_count":1,"evidence_url":"https://aidemos.com/evidence/04a73b4c-336e-4b43-9962-371922367429"}],"appears_in":[{"page_type":"ranking","slug":"image-to-video-generators","title":"Best AI Tools for Generating Cinematic Video from a Single Image","url":"https://aidemos.com/best/image-to-video-generators","binding":"run"},{"page_type":"use-case","slug":"generate-ai-video-from-image","title":"Generate a Cinematic AI Video from a Single Image","url":"https://aidemos.com/use-cases/generate-ai-video-from-image","binding":"run"}],"same_scenario":[{"id":"293512d4-0fe0-47d3-a867-42336caaf9a4","tool":"leonardo-ai","tool_name":"Leonardo AI","verdict":"struggled","score":null,"score_total":null,"note":"The scene animates, but the report says the requested prompt details are not properly reflected."},{"id":"1429eee0-2569-4924-ade6-75a5272f1acf","tool":"pixverse-ai","tool_name":"PixVerse AI","verdict":"mixed","score":null,"score_total":null,"note":"The 3D run delivers the requested camera motion and sound effects, but the report also says clarity is reduced and characters distort, so the scene instructions are only partially preserved."}]}