{"observation":{"id":"e10b1e9f-289c-4e45-bb24-24d7b3063e1e","tool":"granola","tool_name":"Granola","criterion":"topic-segmentation","criterion_name":"Topic Segmentation","criterion_definition":"Breaks long multi-topic meetings into useful sections instead of one blob.","criterion_evidence_type":"transformation","criterion_rank_role":"context","criterion_rank_role_reason":"Breaking long meetings into sections makes notes easier to use, but a tool can still succeed at core note-taking without perfect segmentation. (3 of 3 judges)","scenario":"ai-demos-daily-standup-31-july-2026","scenario_name":"AI Demos Daily Standup — 31 July 2026","group_tag":"ai-meeting-notetaker","scenario_description":"A real 25-minute technical engineering daily standup with 14 attendees and about 10 active speakers, used as the single parallel-capture meeting for evaluating AI meeting notetakers on transcription, diarization, summaries, action items, search/chat, and collaboration features.","modality":"image","input_text":null,"input_artifact_refs":[{"alt":null,"url":"https://cdn.futuresmart.ai/public/aidemos/547dd6f13e4a420fa8ad7bf2c88c7612.png?v=1","role":"input","filename":"31-july-meeting-screenshot.png"}],"stresses":["Transcription accuracy for real names, tool names, numbers, and technical jargon","Speaker diarization across multiple active speakers","Robustness to overlapping speech, crosstalk, and rapid turn-taking","Join reliability for bot-based and botless capture","Summary quality on identical source material","Action-item extraction with correct owners and commitments","Topic segmentation of standup updates","Search and chat grounded in the meeting content","Sharing, API, MCP, integrations, plan limits, languages, and privacy feature coverage"],"verdict":"worked","score":null,"score_total":null,"note":"The notes are split into named topic sections instead of one long blob; the published section header 'Diagram Animation and Other Use Cases' and the report’s multi-section summary structure show logical breakpoints for navigation.","evidence_state":"verified","source":null,"artifacts":[{"url":"https://cdn.futuresmart.ai/public/aidemos/5343a488a2ad46a48618165dac495d7c.png?v=1","role":"output","alt":null},{"url":"https://cdn.futuresmart.ai/public/aidemos/eed3e4d591e047e2990d9327d214923e.png?v=1","role":"output","alt":null}],"run_id":"ace58582-3d1e-48ee-996c-9b3cd03f27a2","study_title":"AI Meeting Notetakers — Capture Accurate Transcripts, Summaries & Action Items From Live Calls","study_kind":"generation","research_task":"86baxegnv","tested_at":null,"completeness":"input-and-output","input":{"state":"files","text":null,"files":[{"url":"https://cdn.futuresmart.ai/public/aidemos/547dd6f13e4a420fa8ad7bf2c88c7612.png?v=1","filename":"31-july-meeting-screenshot.png","alt":"AI Demos Daily Standup — 31 July 2026","role":"input"}],"modality":"image","stresses":["Transcription accuracy for real names, tool names, numbers, and technical jargon","Speaker diarization across multiple active speakers","Robustness to overlapping speech, crosstalk, and rapid turn-taking","Join reliability for bot-based and botless capture","Summary quality on identical source material","Action-item extraction with correct owners and commitments","Topic segmentation of standup updates","Search and chat grounded in the meeting content","Sharing, API, MCP, integrations, plan limits, languages, and privacy feature coverage"]},"tool_page_slug":"granola","tool_url":"https://aidemos.com/tools/granola","permalink":"https://aidemos.com/evidence/e10b1e9f-289c-4e45-bb24-24d7b3063e1e","api_url":"https://ai.aidemos.com/v1/observations/e10b1e9f-289c-4e45-bb24-24d7b3063e1e"},"peers":[{"id":"46b5c695-1c28-461a-ad98-573516adf29f","tool":"fathom","tool_name":"Fathom","verdict":"worked","score":null,"score_total":null,"note":"Fathom breaks the standup into named topical sections instead of one blob, including headers like \"Process & System Blockers\" and \"Content Quality & Review Process.\"","artifact_count":2,"thumbnail":"https://cdn.futuresmart.ai/public/aidemos/1d727e487573413ab4068646f67ff0af.png?v=1","evidence_url":"https://aidemos.com/evidence/46b5c695-1c28-461a-ad98-573516adf29f"},{"id":"b69458f2-5b1d-4915-bf53-39ab59f059d2","tool":"fellow","tool_name":"Fellow","verdict":"worked","score":null,"score_total":null,"note":"The tool broke the standup into logical topic sections with clear headers and separated discussion points, rather than leaving the meeting as one undifferentiated blob.","artifact_count":1,"thumbnail":"https://cdn.futuresmart.ai/public/aidemos/e9cfe1cbf2b543a3a9ce3c67d83d21ff.png?v=1","evidence_url":"https://aidemos.com/evidence/b69458f2-5b1d-4915-bf53-39ab59f059d2"},{"id":"cc75e1cb-8718-41a2-9065-0ad7df27d332","tool":"fireflies-ai","tool_name":"Fireflies.ai","verdict":"worked","score":null,"score_total":null,"note":"Broke the meeting notes into named sections with descriptive headers and short recap paragraphs, making the output skimmable instead of one undifferentiated block.","artifact_count":1,"thumbnail":"https://cdn.futuresmart.ai/public/aidemos/3146d83f0f2d4913ba1acc44c5a24530.png?v=1","evidence_url":"https://aidemos.com/evidence/cc75e1cb-8718-41a2-9065-0ad7df27d332"},{"id":"35b58400-b0b9-4e04-851c-9c0c51b07e1b","tool":"happyscribe","tool_name":"HappyScribe","verdict":"worked","score":null,"score_total":null,"note":"It breaks the standup into logical topic sections that reflect meeting flow, instead of presenting the notes as a single undifferentiated block.","artifact_count":1,"thumbnail":"https://cdn.futuresmart.ai/public/aidemos/bf864a8106ca4c3db1b8e79103b8e4f2.png?v=1","evidence_url":"https://aidemos.com/evidence/35b58400-b0b9-4e04-851c-9c0c51b07e1b"},{"id":"b45a857b-9306-4a29-a568-dcdf9f79fdf2","tool":"meetgeek","tool_name":"MeetGeek","verdict":"worked","score":null,"score_total":null,"note":"It breaks the meeting into useful numbered topic sections instead of one blob, with a visible hierarchy under \"Topics & Highlights\" and the report also noting an Insights tab alongside the segmentation.","artifact_count":1,"thumbnail":"https://cdn.futuresmart.ai/public/aidemos/802e2202df02459b9fe81b754b08212f.png?v=1","evidence_url":"https://aidemos.com/evidence/b45a857b-9306-4a29-a568-dcdf9f79fdf2"},{"id":"dacc4361-09c8-4791-8895-8a8095e21fe8","tool":"notta","tool_name":"Notta","verdict":"worked","score":null,"score_total":null,"note":"The meeting was segmented into useful topic blocks rather than one blob; the report names three sections, including Task & Issue Management, Individual Progress Updates, and API Benchmarking Task.","artifact_count":1,"thumbnail":"https://cdn.futuresmart.ai/public/aidemos/84691bb3ad6d4ac19dec1c5f25f3a0a9.png?v=1","evidence_url":"https://aidemos.com/evidence/dacc4361-09c8-4791-8895-8a8095e21fe8"},{"id":"eedc69a7-0929-4cbd-a001-6122eed42379","tool":"otter-ai","tool_name":"Otter.ai","verdict":"worked","score":null,"score_total":null,"note":"Otter broke the standup into useful topic sections rather than one blob, with named headings such as Issue Task Assignments and Status Updates and Error Resolution and Task Link Sharing, making the summary skimmable.","artifact_count":3,"thumbnail":"https://cdn.futuresmart.ai/public/aidemos/22a79c0f848e403c82d0a767e9e00914.png?v=1","evidence_url":"https://aidemos.com/evidence/eedc69a7-0929-4cbd-a001-6122eed42379"}],"other_criteria":[{"id":"9af81f24-1ac8-4a6a-b543-1959eafc22b0","criterion":"action-item-extraction","criterion_name":"Action-Item Extraction","rank_role":"decisive","verdict":"worked","score":null,"score_total":null,"note":"Granola extracts concrete next steps with ownership: the visible action item says to create a subtask and add details, names Mahreen Fathima as owner, and marks the item for same-day follow-up; the report says the extracted list contained no false positives.","artifact_count":2,"evidence_url":"https://aidemos.com/evidence/9af81f24-1ac8-4a6a-b543-1959eafc22b0"},{"id":"633e805b-8efa-4b9f-8667-7b79bb93072c","criterion":"chat-with-notes-ask-questions","criterion_name":"Chat with Notes / Ask Questions","rank_role":"context","verdict":"worked","score":null,"score_total":null,"note":"The chat/Q&A surface gives grounded answers from the meeting record: on the tool-access question it says access was confirmed that day, cites both the notes and transcript, and identifies rerunning testing as the next step.","artifact_count":2,"evidence_url":"https://aidemos.com/evidence/633e805b-8efa-4b9f-8667-7b79bb93072c"},{"id":"e181aa66-75ab-4b14-a804-0bb3f6481466","criterion":"editability","criterion_name":"Editability","rank_role":"context","verdict":"mixed","score":null,"score_total":null,"note":"Granola supports editing the summary layer, but transcript text is not editable in-app; the report describes transcript correction as locked by design, so users can fix notes and action items but not the underlying transcript.","artifact_count":2,"evidence_url":"https://aidemos.com/evidence/e181aa66-75ab-4b14-a804-0bb3f6481466"},{"id":"c83c6126-1c92-4f50-a102-cbdfe3713f0f","criterion":"join-method-reliability","criterion_name":"Join Method & Reliability","rank_role":"decisive","verdict":"worked","score":null,"score_total":null,"note":"The botless desktop capture recorded the full ~25-minute meeting end-to-end without visible dropouts; the transcript reaches the call’s closing lines, indicating uninterrupted capture rather than a mid-call failure.","artifact_count":3,"evidence_url":"https://aidemos.com/evidence/c83c6126-1c92-4f50-a102-cbdfe3713f0f"},{"id":"a5c429f5-458e-4d15-807a-a871c71e38be","criterion":"search-across-notes","criterion_name":"Search Across Notes","rank_role":"context","verdict":"worked","score":null,"score_total":null,"note":"Granola’s note search returns exact-match results with navigation: a query for 'api' produced a 1/1 hit and highlighted the matched word in the transcript, so keyword lookup works directly inside the meeting note.","artifact_count":2,"evidence_url":"https://aidemos.com/evidence/a5c429f5-458e-4d15-807a-a871c71e38be"},{"id":"73b827a9-4822-4d1e-a8ad-b6853c5eab4d","criterion":"speaker-diarization","criterion_name":"Speaker Diarization","rank_role":"decisive","verdict":"failed","score":null,"score_total":null,"note":"Granola’s default capture does not attribute speakers: the settings panel shows Speaker tags switched off, and the transcript excerpt is a plain text wall with no speaker labels, so diarization is absent unless the user manually enables it.","artifact_count":2,"evidence_url":"https://aidemos.com/evidence/73b827a9-4822-4d1e-a8ad-b6853c5eab4d"},{"id":"d9b8082d-37a5-4e94-8412-e2a6d59d4aa3","criterion":"summary-quality","criterion_name":"Summary Quality","rank_role":"decisive","verdict":"worked","score":null,"score_total":null,"note":"Granola produces skimmable summaries with named sections; the meeting output is organized into at least three top-level sections, including Use Case Status and Review Progress, Tool Research and Publishing, and Access Tracker Updates.","artifact_count":4,"evidence_url":"https://aidemos.com/evidence/d9b8082d-37a5-4e94-8412-e2a6d59d4aa3"},{"id":"87d23b55-bc59-4d3e-8547-9b80fc1107f6","criterion":"transcription-accuracy","criterion_name":"Transcription Accuracy","rank_role":"decisive","verdict":"failed","score":null,"score_total":null,"note":"On this 25-minute, multi-speaker standup, Granola’s transcript quality is unreliable: the published excerpt shows garbled phrasing and mistranscribed wording, and the report says the mishearing pattern recurs across early, middle, and late sections rather than being isolated to one moment.","artifact_count":2,"evidence_url":"https://aidemos.com/evidence/87d23b55-bc59-4d3e-8547-9b80fc1107f6"}],"appears_in":[{"page_type":"ranking","slug":"ai-meeting-notetakers","title":"Best AI Meeting Notetakers for Accurate Transcripts, Summaries, and Action Items","url":"https://aidemos.com/best/ai-meeting-notetakers","binding":"run"}],"same_scenario":[{"id":"46b5c695-1c28-461a-ad98-573516adf29f","tool":"fathom","tool_name":"Fathom","verdict":"worked","score":null,"score_total":null,"note":"Fathom breaks the standup into named topical sections instead of one blob, including headers like \"Process & System Blockers\" and \"Content Quality & Review Process.\""},{"id":"b69458f2-5b1d-4915-bf53-39ab59f059d2","tool":"fellow","tool_name":"Fellow","verdict":"worked","score":null,"score_total":null,"note":"The tool broke the standup into logical topic sections with clear headers and separated discussion points, rather than leaving the meeting as one undifferentiated blob."},{"id":"cc75e1cb-8718-41a2-9065-0ad7df27d332","tool":"fireflies-ai","tool_name":"Fireflies.ai","verdict":"worked","score":null,"score_total":null,"note":"Broke the meeting notes into named sections with descriptive headers and short recap paragraphs, making the output skimmable instead of one undifferentiated block."},{"id":"35b58400-b0b9-4e04-851c-9c0c51b07e1b","tool":"happyscribe","tool_name":"HappyScribe","verdict":"worked","score":null,"score_total":null,"note":"It breaks the standup into logical topic sections that reflect meeting flow, instead of presenting the notes as a single undifferentiated block."},{"id":"b45a857b-9306-4a29-a568-dcdf9f79fdf2","tool":"meetgeek","tool_name":"MeetGeek","verdict":"worked","score":null,"score_total":null,"note":"It breaks the meeting into useful numbered topic sections instead of one blob, with a visible hierarchy under \"Topics & Highlights\" and the report also noting an Insights tab alongside the segmentation."},{"id":"dacc4361-09c8-4791-8895-8a8095e21fe8","tool":"notta","tool_name":"Notta","verdict":"worked","score":null,"score_total":null,"note":"The meeting was segmented into useful topic blocks rather than one blob; the report names three sections, including Task & Issue Management, Individual Progress Updates, and API Benchmarking Task."},{"id":"eedc69a7-0929-4cbd-a001-6122eed42379","tool":"otter-ai","tool_name":"Otter.ai","verdict":"worked","score":null,"score_total":null,"note":"Otter broke the standup into useful topic sections rather than one blob, with named headings such as Issue Task Assignments and Status Updates and Error Resolution and Task Link Sharing, making the summary skimmable."}]}