Extracts all 8 advertising line items as separate structured records without duplication or omission.
What was measured
Table & Record Completeness
Are all tabular rows and line items extracted without merging, duplication, omission, or phantom records?
decisive for this rankingtransformation
Missing, duplicated, or merged rows and line items directly corrupt the structured dataset and break query results. (3 of 3 judges)
What was given, what came back
Test input: Invoice PDF · pdf · group: business-document-extraction
Input — what we sent
A two-page broadcast advertising invoice PDF with nested header metadata, billing and remit addresses, and eight line items spanning a page break. It was used to stress hierarchical line-item extraction, amount precision, time/day parsing, code extraction, and summary validation.
Why this input is hard
- · Nested line-item hierarchy extraction
- · Multi-page continuity across a page break
- · Precision on large dollar amounts and totals
- · Parsing complex time slots and day patterns
- · Extraction of Ad IDs and reconciliation codes
- · Mapping structured metadata sections correctly
- · Financial summary validation
- · Handling political advertising compliance text
Output — unretouched


Also checked on this input — same tool, 6 other criteria
Extraction Accuracy◐ MixedRetains the source label in payment_terms, returning 'Payment Terms 30 Days' instead of only the requested value, so the field is not fully normalized.Extraction Accuracy✓ WorkedExtracts the invoice summary totals correctly, including agency_commission 4462.5, aired_spots 8, gross_total 29750, and net_amount_due 25287.5.Extraction Accuracy✓ WorkedExtracts invoice metadata with machine-readable values such as invoice_number 4064621-1, invoice_date 2012-10-28, invoice_period_start 2012-10-01, and advertiser_code 155.Schema Adherence✓ WorkedReconstructs the invoice schema into nested invoice_metadata, advertiser, station, account_details, billing_address, remit_address, flight_dates, line_items, and summary sections rather than returning a flat document dump.Semantic Field Enrichment✓ WorkedDerives scheduling fields from the line item, including day_of_week 'Su', days_pattern '------S', air_time '09:38:00', and time_slot '9a-10a'.Structural Clean Output⚠ StruggledPreserves valid JSON but does not keep the schema's original key order, so downstream consumers that expect schema-consistent ordering need an extra formatting step.
Provenance
- Observation
- dcbeb5fd-b152-4212-82e3-048c6a20b6d6
- Evidence run
- a061b9e7-a9c5-443d-a171-b296aaf51b8c
- Study
- Extract and query structured data from documents using natural language
- Research task
- 86b9y25e5
- Tested at
- not recorded
- Source
- first-party
- Evidence state
- verified
- Proof shown
- input + output shown
- Cost / latency
- not captured
- Repeat run
- not captured
- Tester
- not captured
The last three rows are honest blanks, not placeholders — our capture has no field for them yet.
Query this
get_evidence({
tool: "retab",
scenario: "business-document-extraction"
})MCP · mcp.aidemos.com/api/mcp
Free with attribution.
Same input, same check — 8 other tools
measured on Table & Record Completeness
Datalab✓ WorkedExtracts all eight invoice line items as separate records without merging, including rows that span the page break.Docsumo✓ WorkedThe line-item table keeps all 8 rows, including the row that spans the page break, without introducing a duplicate line item.Extend AI✓ WorkedIt extracts all 8 advertising line items as separate records, preserving item boundaries across the page break without merging adjacent rows.Landing AI✓ WorkedExtracts all 8 advertising line items as separate records, preserving the page-break continuation instead of merging or omitting rows.LlamaParse✗ FailedThe tool extracts invoice line items as structured rows, but it introduces an incorrect 8th line item with line_number 9, so the table record set is not fully faithful to the source invoice’s eight advertising records.Nanonets✓ WorkedKeeps all 8 advertising line items as separate records with no row merging or duplication.Reducto✓ WorkedExtracts all eight advertising line items as separate records and preserves the page-spanning table without merging adjacent rows.Unstract✓ WorkedExtracts the line-item table without losing the row that crosses the page break; line 8 is returned in full with no duplication or data loss.
This evidence is published in
Real inputs and real outputs, no retouching · every cell queryable via API & MCP · aidemos.com