Evidence · first-party tested/Best AI Meeting Notetakers for Accurate Transcripts, Summaries, and Action Items
It answers direct grounded questions correctly, but the report records an incorrect answer on a speaker-dependent scheduling question, so chat is reliable for simple queries but weaker when attribution/context matters.
What was measured
Chat with Notes / Ask Questions
Gives grounded answers with the cited moment and admits when unknown.
context, not decisivetransformation
Q&A over notes is useful, but it is an add-on to the capture and summarization job rather than a core measure of it. (3 of 3 judges)
What was given, what came back
Test input: AI Demos Daily Standup — 31 July 2026 · image · group: ai-meeting-notetaker
Input — what we sent

AI Demos Daily Standup — 31 July 2026
A real 25-minute technical engineering daily standup with 14 attendees and about 10 active speakers, used as the single parallel-capture meeting for evaluating AI meeting notetakers on transcription, diarization, summaries, action items, search/chat, and collaboration features.
Why this input is hard
- · Transcription accuracy for real names, tool names, numbers, and technical jargon
- · Speaker diarization across multiple active speakers
- · Robustness to overlapping speech, crosstalk, and rapid turn-taking
- · Join reliability for bot-based and botless capture
- · Summary quality on identical source material
- · Action-item extraction with correct owners and commitments
- · Topic segmentation of standup updates
- · Search and chat grounded in the meeting content
- · Sharing, API, MCP, integrations, plan limits, languages, and privacy feature coverage
Output — unretouched


Also checked on this input — same tool, 8 other criteria
Action-Item Extraction✓ WorkedIt extracts real commitments as action items rather than noise; the report says all extracted items had correct ownership and timing, and the visible note includes an owned action item with timestamp 19:51.Editability✓ WorkedUsers can edit generated outputs inline before sharing; the report says summary, action items, and the full transcript are all editable, and the UI shows editable summary text.Join Method & Reliability✓ WorkedThe bot successfully joined a Google Meet call and the report says it captured the full ~30-minute meeting with zero disconnections or data loss.Search Across Notes✓ WorkedIt supports transcript search with precise retrieval: searching for "api" surfaces the matching text in context and the report says timestamps are returned to within a few seconds.Speaker Diarization◐ MixedIt identifies most speakers in a multi-speaker standup, but leaves at least one utterance as "Unknown speaker" and misattributes some lines to the wrong speaker, so attribution is not fully reliable.Summary Quality✓ WorkedIt produces a clear, skimmable meeting summary with topic organization and a Next Steps section, and the report says it preserved the major decisions and discussion points.Topic Segmentation✓ WorkedIt breaks the meeting into useful numbered topic sections instead of one blob, with a visible hierarchy under "Topics & Highlights" and the report also noting an Insights tab alongside the segmentation.Transcription Accuracy✓ WorkedIt transcribes a normal ~25-minute, ~10-active-speaker engineering standup mostly accurately, with only minor proper-noun/term drift noted in the report; one example given is "Madin" being misheard for "Mahreen".
Provenance
- Observation
- 75aaf0b0-c9d5-4376-8b81-5b6f2574c0fb
- Evidence run
- ace58582-3d1e-48ee-996c-9b3cd03f27a2
- Study
- AI Meeting Notetakers — Capture Accurate Transcripts, Summaries & Action Items From Live Calls
- Research task
- 86baxegnv
- Tested at
- not recorded
- Source
- first-party
- Evidence state
- verified
- Proof shown
- input + output shown
- Cost / latency
- not captured
- Repeat run
- not captured
- Tester
- not captured
The last three rows are honest blanks, not placeholders — our capture has no field for them yet.
Query this
get_evidence({
tool: "meetgeek",
scenario: "ai-meeting-notetaker"
})MCP · mcp.aidemos.com/api/mcp
Free with attribution.
Same input, same check — 7 other tools
measured on Chat with Notes / Ask Questions
Fathom✓ WorkedAsk Fathom answers direct factual questions from the meeting notes with grounded references; for one query it answered that a call was scheduled for 6th August and linked the supporting transcript mention.Fellow✓ WorkedAsk Fellow returned grounded answers to natural-language questions against the meeting notes, and the tested query produced a cited response rather than an unsupported hallucination.Fireflies.ai✓ WorkedAskFred answered a natural-language question with a specific grounded response ('August 6th') and relevant context, with no hallucination reported in the tested query.Granola✓ WorkedThe chat/Q&A surface gives grounded answers from the meeting record: on the tool-access question it says access was confirmed that day, cites both the notes and transcript, and identifies rerunning testing as the next step.HappyScribe✓ WorkedAI chat answers a meeting question with a grounded transcript-backed response, returning that the call was scheduled for '6th August' and explicitly indicating it is reading the transcription.Notta✓ WorkedThe Q&A interface answered a natural-language question with a grounded response from the meeting record, including the specific date "6th August," and the report observed no hallucinations.Otter.ai✓ WorkedOtter’s AI Chat answered meeting questions with a grounded response and a specific timestamp, and the report says the answers were cited and free of hallucinations in the tested queries.
This evidence is published in
Real inputs and real outputs, no retouching · every cell queryable via API & MCP · aidemos.com