{"observation":{"id":"b482399a-b91b-4aaa-b6f1-a56363792548","tool":"meetgeek","tool_name":"MeetGeek","criterion":"summary-quality","criterion_name":"Summary Quality","criterion_definition":"Captures decisions and key points, is structured and skimmable, and drops nothing important.","criterion_evidence_type":"transformation","criterion_rank_role":"decisive","criterion_rank_role_reason":"The product is being judged on whether it produces a useful meeting summary that preserves key decisions and points without missing important content. (3 of 3 judges)","scenario":"ai-demos-daily-standup-31-july-2026","scenario_name":"AI Demos Daily Standup — 31 July 2026","group_tag":"ai-meeting-notetaker","scenario_description":"A real 25-minute technical engineering daily standup with 14 attendees and about 10 active speakers, used as the single parallel-capture meeting for evaluating AI meeting notetakers on transcription, diarization, summaries, action items, search/chat, and collaboration features.","modality":"image","input_text":null,"input_artifact_refs":[{"alt":null,"url":"https://cdn.futuresmart.ai/public/aidemos/547dd6f13e4a420fa8ad7bf2c88c7612.png?v=1","role":"input","filename":"31-july-meeting-screenshot.png"}],"stresses":["Transcription accuracy for real names, tool names, numbers, and technical jargon","Speaker diarization across multiple active speakers","Robustness to overlapping speech, crosstalk, and rapid turn-taking","Join reliability for bot-based and botless capture","Summary quality on identical source material","Action-item extraction with correct owners and commitments","Topic segmentation of standup updates","Search and chat grounded in the meeting content","Sharing, API, MCP, integrations, plan limits, languages, and privacy feature coverage"],"verdict":"worked","score":null,"score_total":null,"note":"It produces a clear, skimmable meeting summary with topic organization and a Next Steps section, and the report says it preserved the major decisions and discussion points.","evidence_state":"verified","source":null,"artifacts":[{"url":"https://cdn.futuresmart.ai/public/aidemos/348b3c4459f34fe8b0214debd8e0998b.png?v=1","role":"output","alt":null},{"url":"https://cdn.futuresmart.ai/public/aidemos/4074b4cecf0145948228bc248ac2d0be.png?v=1","role":"output","alt":null}],"run_id":"ace58582-3d1e-48ee-996c-9b3cd03f27a2","study_title":"AI Meeting Notetakers — Capture Accurate Transcripts, Summaries & Action Items From Live Calls","study_kind":"generation","research_task":"86baxegnv","tested_at":null,"completeness":"input-and-output","input":{"state":"files","text":null,"files":[{"url":"https://cdn.futuresmart.ai/public/aidemos/547dd6f13e4a420fa8ad7bf2c88c7612.png?v=1","filename":"31-july-meeting-screenshot.png","alt":"AI Demos Daily Standup — 31 July 2026","role":"input"}],"modality":"image","stresses":["Transcription accuracy for real names, tool names, numbers, and technical jargon","Speaker diarization across multiple active speakers","Robustness to overlapping speech, crosstalk, and rapid turn-taking","Join reliability for bot-based and botless capture","Summary quality on identical source material","Action-item extraction with correct owners and commitments","Topic segmentation of standup updates","Search and chat grounded in the meeting content","Sharing, API, MCP, integrations, plan limits, languages, and privacy feature coverage"]},"tool_page_slug":"meetgeek","tool_url":"https://aidemos.com/tools/meetgeek","permalink":"https://aidemos.com/evidence/b482399a-b91b-4aaa-b6f1-a56363792548","api_url":"https://ai.aidemos.com/v1/observations/b482399a-b91b-4aaa-b6f1-a56363792548"},"peers":[{"id":"53731c87-42c6-49f3-b19d-6a3ab5fe169e","tool":"fathom","tool_name":"Fathom","verdict":"worked","score":null,"score_total":null,"note":"Fathom produces a skimmable written recap with named sections such as Meeting Purpose, Key Takeaways, and Topics; the report describes the summary as structured and concise.","artifact_count":3,"thumbnail":"https://cdn.futuresmart.ai/public/aidemos/7a6b7c6bede341b1a0f8d941ba783549.png?v=1","evidence_url":"https://aidemos.com/evidence/53731c87-42c6-49f3-b19d-6a3ab5fe169e"},{"id":"d89239cd-ae18-4967-90d8-e968276badf1","tool":"fellow","tool_name":"Fellow","verdict":"worked","score":null,"score_total":null,"note":"The meeting recap was reported as clearly structured and complete, with the key decisions and discussion points preserved and nothing important dropped.","artifact_count":1,"thumbnail":"https://cdn.futuresmart.ai/public/aidemos/e4f4726c9f8b4d8e9ec8e58a8faf8a3a.png?v=1","evidence_url":"https://aidemos.com/evidence/d89239cd-ae18-4967-90d8-e968276badf1"},{"id":"2811037e-c564-48ad-a195-593c4e2ed1ea","tool":"fireflies-ai","tool_name":"Fireflies.ai","verdict":"worked","score":null,"score_total":null,"note":"Produced a structured notes summary with a named header ('Task Status and Issue Resolution') rather than a blob, and the report says the full summary was multi-section and did not drop important points.","artifact_count":1,"thumbnail":"https://cdn.futuresmart.ai/public/aidemos/791c17668df346a1aac4a7f11d716715.png?v=1","evidence_url":"https://aidemos.com/evidence/2811037e-c564-48ad-a195-593c4e2ed1ea"},{"id":"d9b8082d-37a5-4e94-8412-e2a6d59d4aa3","tool":"granola","tool_name":"Granola","verdict":"worked","score":null,"score_total":null,"note":"Granola produces skimmable summaries with named sections; the meeting output is organized into at least three top-level sections, including Use Case Status and Review Progress, Tool Research and Publishing, and Access Tracker Updates.","artifact_count":4,"thumbnail":"https://cdn.futuresmart.ai/public/aidemos/eed3e4d591e047e2990d9327d214923e.png?v=1","evidence_url":"https://aidemos.com/evidence/d9b8082d-37a5-4e94-8412-e2a6d59d4aa3"},{"id":"ab6fc227-dfb2-4524-a6ab-83adb117660f","tool":"happyscribe","tool_name":"HappyScribe","verdict":"mixed","score":null,"score_total":null,"note":"The summary can hallucinate a person name: the report says it substituted 'Nadine' for 'Mahreen' in a summary bullet, creating a false team-member attribution.","artifact_count":2,"thumbnail":"https://cdn.futuresmart.ai/public/aidemos/ac13349047eb4335af7ffc2f1f798d7e.png?v=1","evidence_url":"https://aidemos.com/evidence/ab6fc227-dfb2-4524-a6ab-83adb117660f"},{"id":"aa85bfe2-42ee-4df2-b89f-831056fd4fc3","tool":"notta","tool_name":"Notta","verdict":"worked","score":null,"score_total":null,"note":"The generated meeting summary was comprehensive and skimmable, with structured sections such as Task & Issue Management and a mindmap-style organization that reflected the meeting flow.","artifact_count":2,"thumbnail":"https://cdn.futuresmart.ai/public/aidemos/6868c5d787734044ae365ce951e12c06.png?v=1","evidence_url":"https://aidemos.com/evidence/aa85bfe2-42ee-4df2-b89f-831056fd4fc3"},{"id":"85b52298-f079-4410-b519-3fc1d34593a0","tool":"otter-ai","tool_name":"Otter.ai","verdict":"worked","score":null,"score_total":null,"note":"Otter produced a clear, structured meeting summary that the report says covered the key decisions and discussion points without dropping anything important, and the summary page loaded with organized sections like Overview and Action Items.","artifact_count":3,"thumbnail":"https://cdn.futuresmart.ai/public/aidemos/fb9982daf5bf4988b132cda637bf3053.png?v=1","evidence_url":"https://aidemos.com/evidence/85b52298-f079-4410-b519-3fc1d34593a0"}],"other_criteria":[{"id":"dfeacb85-baa0-4c8b-a44e-e4a7b005d54f","criterion":"action-item-extraction","criterion_name":"Action-Item Extraction","rank_role":"decisive","verdict":"worked","score":null,"score_total":null,"note":"It extracts real commitments as action items rather than noise; the report says all extracted items had correct ownership and timing, and the visible note includes an owned action item with timestamp 19:51.","artifact_count":1,"evidence_url":"https://aidemos.com/evidence/dfeacb85-baa0-4c8b-a44e-e4a7b005d54f"},{"id":"75aaf0b0-c9d5-4376-8b81-5b6f2574c0fb","criterion":"chat-with-notes-ask-questions","criterion_name":"Chat with Notes / Ask Questions","rank_role":"context","verdict":"mixed","score":null,"score_total":null,"note":"It answers direct grounded questions correctly, but the report records an incorrect answer on a speaker-dependent scheduling question, so chat is reliable for simple queries but weaker when attribution/context matters.","artifact_count":2,"evidence_url":"https://aidemos.com/evidence/75aaf0b0-c9d5-4376-8b81-5b6f2574c0fb"},{"id":"0a57537c-4ad4-4b2f-8115-4281f671bd30","criterion":"editability","criterion_name":"Editability","rank_role":"context","verdict":"worked","score":null,"score_total":null,"note":"Users can edit generated outputs inline before sharing; the report says summary, action items, and the full transcript are all editable, and the UI shows editable summary text.","artifact_count":1,"evidence_url":"https://aidemos.com/evidence/0a57537c-4ad4-4b2f-8115-4281f671bd30"},{"id":"9ebe6fbe-6f09-440e-afde-66593a710483","criterion":"join-method-reliability","criterion_name":"Join Method & Reliability","rank_role":"decisive","verdict":"worked","score":null,"score_total":null,"note":"The bot successfully joined a Google Meet call and the report says it captured the full ~30-minute meeting with zero disconnections or data loss.","artifact_count":3,"evidence_url":"https://aidemos.com/evidence/9ebe6fbe-6f09-440e-afde-66593a710483"},{"id":"e8929c90-b25d-47cf-ab27-f340ed881ef4","criterion":"search-across-notes","criterion_name":"Search Across Notes","rank_role":"context","verdict":"worked","score":null,"score_total":null,"note":"It supports transcript search with precise retrieval: searching for \"api\" surfaces the matching text in context and the report says timestamps are returned to within a few seconds.","artifact_count":1,"evidence_url":"https://aidemos.com/evidence/e8929c90-b25d-47cf-ab27-f340ed881ef4"},{"id":"4f7e5b21-ff4f-4846-9a07-3219c0659681","criterion":"speaker-diarization","criterion_name":"Speaker Diarization","rank_role":"decisive","verdict":"mixed","score":null,"score_total":null,"note":"It identifies most speakers in a multi-speaker standup, but leaves at least one utterance as \"Unknown speaker\" and misattributes some lines to the wrong speaker, so attribution is not fully reliable.","artifact_count":2,"evidence_url":"https://aidemos.com/evidence/4f7e5b21-ff4f-4846-9a07-3219c0659681"},{"id":"b45a857b-9306-4a29-a568-dcdf9f79fdf2","criterion":"topic-segmentation","criterion_name":"Topic Segmentation","rank_role":"context","verdict":"worked","score":null,"score_total":null,"note":"It breaks the meeting into useful numbered topic sections instead of one blob, with a visible hierarchy under \"Topics & Highlights\" and the report also noting an Insights tab alongside the segmentation.","artifact_count":1,"evidence_url":"https://aidemos.com/evidence/b45a857b-9306-4a29-a568-dcdf9f79fdf2"},{"id":"5a191ed7-02a9-4979-931d-9219e0b75fce","criterion":"transcription-accuracy","criterion_name":"Transcription Accuracy","rank_role":"decisive","verdict":"worked","score":null,"score_total":null,"note":"It transcribes a normal ~25-minute, ~10-active-speaker engineering standup mostly accurately, with only minor proper-noun/term drift noted in the report; one example given is \"Madin\" being misheard for \"Mahreen\".","artifact_count":1,"evidence_url":"https://aidemos.com/evidence/5a191ed7-02a9-4979-931d-9219e0b75fce"}],"appears_in":[{"page_type":"ranking","slug":"ai-meeting-notetakers","title":"Best AI Meeting Notetakers for Accurate Transcripts, Summaries, and Action Items","url":"https://aidemos.com/best/ai-meeting-notetakers","binding":"run"}],"same_scenario":[{"id":"53731c87-42c6-49f3-b19d-6a3ab5fe169e","tool":"fathom","tool_name":"Fathom","verdict":"worked","score":null,"score_total":null,"note":"Fathom produces a skimmable written recap with named sections such as Meeting Purpose, Key Takeaways, and Topics; the report describes the summary as structured and concise."},{"id":"d89239cd-ae18-4967-90d8-e968276badf1","tool":"fellow","tool_name":"Fellow","verdict":"worked","score":null,"score_total":null,"note":"The meeting recap was reported as clearly structured and complete, with the key decisions and discussion points preserved and nothing important dropped."},{"id":"2811037e-c564-48ad-a195-593c4e2ed1ea","tool":"fireflies-ai","tool_name":"Fireflies.ai","verdict":"worked","score":null,"score_total":null,"note":"Produced a structured notes summary with a named header ('Task Status and Issue Resolution') rather than a blob, and the report says the full summary was multi-section and did not drop important points."},{"id":"d9b8082d-37a5-4e94-8412-e2a6d59d4aa3","tool":"granola","tool_name":"Granola","verdict":"worked","score":null,"score_total":null,"note":"Granola produces skimmable summaries with named sections; the meeting output is organized into at least three top-level sections, including Use Case Status and Review Progress, Tool Research and Publishing, and Access Tracker Updates."},{"id":"ab6fc227-dfb2-4524-a6ab-83adb117660f","tool":"happyscribe","tool_name":"HappyScribe","verdict":"mixed","score":null,"score_total":null,"note":"The summary can hallucinate a person name: the report says it substituted 'Nadine' for 'Mahreen' in a summary bullet, creating a false team-member attribution."},{"id":"aa85bfe2-42ee-4df2-b89f-831056fd4fc3","tool":"notta","tool_name":"Notta","verdict":"worked","score":null,"score_total":null,"note":"The generated meeting summary was comprehensive and skimmable, with structured sections such as Task & Issue Management and a mindmap-style organization that reflected the meeting flow."},{"id":"85b52298-f079-4410-b519-3fc1d34593a0","tool":"otter-ai","tool_name":"Otter.ai","verdict":"worked","score":null,"score_total":null,"note":"Otter produced a clear, structured meeting summary that the report says covered the key decisions and discussion points without dropping anything important, and the summary page loaded with organized sections like Overview and Action Items."}]}