Handles a table-heavy 18-page report with numerous tables while keeping the extracted sections aligned to the document flow.
What was measured
Complex Document Handling
Maintains quality across long, multi-section, and mixed-content documents without degradation.
decisive for this rankingtransformation
The subject explicitly says complex PDF, so sustained quality across long, mixed-content documents is central to the tool’s ability to do the job. (3 of 3 judges)
What was given, what came back
Test input: Financial Report - Table Heavy · pdf · group: financial-report-table-heavy
Input — what we sent
A table-heavy corporate financial report used to test extraction of dense, hierarchical financial statements with grouped columns, multi-row headers, segment-reporting tables, and narrative disclosures.
Why this input is hard
- · Multi-page financial statement extraction
- · Hierarchical table reconstruction
- · Grouped columns and multi-row headers
- · Reading order in a report with mixed narrative and tables
- · Document structure retention
- · Markdown usability
Output — unretouched
Also checked on this input — same tool, 4 other criteria
Markdown Quality✓ WorkedReturns the extraction as structured markdown in the Document Markdown workflow, making the output usable rather than a raw text dump.Reading Order & Structure✓ WorkedKeeps section ordering and narrative flow in an 18-page table-heavy financial report, preserving the operating-performance section before the tabular disclosures.Table Preservation✓ WorkedReconstructs a multilevel segment-results table with grouped previous/present quarter headers, year-over-year change columns, six segment rows, and a total row.Table Preservation✗ FailedBreaks on a more complex multi-header table by losing header hierarchy and omitting at least one header label, producing an incomplete and structurally incorrect reconstruction.
Provenance
- Observation
- 79747d1f-4b06-48ae-b43f-512e9c2d1465
- Evidence run
- 6e3160de-fe46-4b45-b071-72560b5c5d0e
- Study
- Convert a Complex PDF into Clean Markdown with an API
- Research task
- 86b9h7t37
- Tested at
- not recorded
- Source
- first-party
- Evidence state
- verified
- Proof shown
- input + output shown
- Cost / latency
- not captured
- Repeat run
- not captured
- Tester
- not captured
The last three rows are honest blanks, not placeholders — our capture has no field for them yet.
Query this
get_evidence({
tool: "tensorlake",
scenario: "financial-report-table-heavy"
})MCP · mcp.aidemos.com/api/mcp
Free with attribution.
Same input, same check — 6 other tools
measured on Complex Document Handling
Adobe API✓ WorkedProcesses a multi-section, table-heavy financial report in one automated run and returns structured Markdown output.Extend AI✓ WorkedProcesses the 18-page table-heavy financial report end-to-end and returns usable markdown without manual correction.LlamaParse✓ WorkedProcesses an 18-page table-heavy report end-to-end and reaches SUCCESS in the Results view without manual correction.Mistral AI✓ WorkedThe tool processes an 18-page table-heavy financial report end-to-end and returns both page-level and consolidated markdown outputs.Nutrient.io✓ WorkedHandles an 18-page table-heavy financial report and returns a parsed markdown output plus preview, showing end-to-end extraction on a dense corporate document.Reducto◐ MixedProcesses the 18-page report in 9.7 seconds with no truncation; the 9-column segment table at page 17 is flawless even though the first table on page 2 is broken.
This evidence is published in
Real inputs and real outputs, no retouching · every cell queryable via API & MCP · aidemos.com

