AI Demos Research — the structured-intelligence platform. Every verdict on these pages opens to the execution behind it.
BlazeSQL
The operator needs to find the questions the agent is failing on
untested
GROUP B — Observability. Whether failures and unanswered questions are surfaced AS failures, rather than filed as ordinary answered questions.
reading this cell
As ai-database-agents v1 reads it — only grades under the rubric that version pins.
reading policy v1 · fingerprint 4f53cda18c2baa0c…
What we measured
Nothing was measured numerically for this assessment.
Every assessment
0 counting · 0 from captured bytes · 0 transcribed
| Ref | Verdict | Rubric | Provenance | Graded by | Counts here |
|---|
Nobody has assessed this tool on this question. `untested` is what the page says, and it is a value rather than a blank.