Introduces a stray OCR-style space into every Ad-ID, so the identifier text is not preserved exactly as printed.

⚠ Struggled🧾 artifact-verifiedinput + output shownTest date not recordedUnstract
What was measured
Extraction Accuracy

Are field values correct, complete, and free of OCR or parsing errors? Includes numerical precision on financial fields.

decisive for this rankingtransformation

The whole point is to pull the right values from documents; wrong or incomplete field values mean the tool failed at the job. (3 of 3 judges)

What was given, what came back

Test input: Invoice PDF · pdf · group: financial-document-extraction
Input — what we sent
image
Input not captured
This run recorded no prompt or input file for the test, so we cannot show you what produced the result below. Capture gaps are tracked, not hidden.

A 2-page broadcast advertising invoice PDF with 8 line items, complex time/day fields, large dollar amounts, and compliance text, used to test hierarchical line-item extraction and financial validation.

Why this input is hard
  • · Nested line-item hierarchy extraction
  • · Multi-page line-item continuity across a page break
  • · Large dollar amount precision and total validation
  • · Parsing time slots, day patterns, and air dates
  • · Extraction of alphanumeric ad IDs and reference codes
  • · Structured metadata mapping for advertiser, station, billing, and remit sections
  • · Political advertising and FCC compliance text recognition
Output — unretouched
unstract-unstract-invoice-output-e53f3477db06.json
Loading file...
Provenance
Observation
668dd494-5404-4e0d-b091-134c04dde238
Evidence run
ec4d736d-95f9-4c88-884c-e280435f7b7b
Study
Extract and query structured data from documents using natural language
Research task
86b9y25e5
Tested at
not recorded
Source
first-party
Evidence state
verified
Proof shown
input + output shown
Cost / latency
not captured
Repeat run
not captured
Tester
not captured

The last three rows are honest blanks, not placeholders — our capture has no field for them yet.

Query this
get_evidence({
  tool: "unstract",
  scenario: "financial-document-extraction"
})
MCP · mcp.aidemos.com/api/mcp
Free with attribution.
Same input, same check — 7 other tools
measured on Extraction Accuracy
Real inputs and real outputs, no retouching · every cell queryable via API & MCP · aidemos.com