The output is organized as a downloadable ZIP containing a consolidated markdown document plus individual page-level files for localized inspection.
What was measured
Markdown Quality
Produces clean, well-structured, usable markdown rather than a flat text dump.
decisive for this rankingtransformation
The ranking is specifically about producing clean Markdown, so the quality and usability of the Markdown output directly measures success. (3 of 3 judges)
What was given, what came back
Test input: Financial Report - Table Heavy · pdf · group: financial-report-table-heavy
Input — what we sent
b12d5cd772a647818ad1833789b090d6.pdf
Financial Report - Table Heavy
A table-heavy corporate financial report used to test extraction of dense, hierarchical financial statements with grouped columns, multi-row headers, segment-reporting tables, and narrative disclosures.
Why this input is hard
- · Multi-page financial statement extraction
- · Hierarchical table reconstruction
- · Grouped columns and multi-row headers
- · Reading order in a report with mixed narrative and tables
- · Document structure retention
- · Markdown usability
Also checked on this input — same tool, 5 other criteria
Complex Document Handling✓ WorkedThe tool processes an 18-page table-heavy financial report end-to-end and returns both page-level and consolidated markdown outputs.Reading Order & Structure✓ WorkedDocument hierarchy and reading flow are preserved across the financial report, with headings staying connected to their explanatory text in the reconstructed output.Reading Order & Structure✗ FailedThe table of contents is flattened into raw text, so the original navigational hierarchy is not preserved.Table Preservation✓ WorkedThe multilevel segment table is preserved with its hierarchical headers intact, maintaining the relationships between the grouped columns and their values.Table Preservation✗ FailedDistinct header levels are collapsed into a single cell in the multilevel table, weakening the parent-child relationships that define the column semantics.
Provenance
- Observation
- 9c7a0e04-472d-43b7-a08c-8cba8763be40
- Evidence run
- 6e3160de-fe46-4b45-b071-72560b5c5d0e
- Study
- Convert a Complex PDF into Clean Markdown with an API
- Research task
- 86b9h7t37
- Tested at
- not recorded
- Source
- first-party
- Evidence state
- verified
- Proof shown
- input + output shown
- Cost / latency
- not captured
- Repeat run
- not captured
- Tester
- not captured
The last three rows are honest blanks, not placeholders — our capture has no field for them yet.
Query this
get_evidence({
tool: "mistral-ai",
scenario: "financial-report-table-heavy"
})MCP · mcp.aidemos.com/api/mcp
Free with attribution.
Same input, same check — 4 other tools
measured on Markdown Quality
LlamaParse◐ MixedRenders the table of contents as sequential text with page numbers instead of a nested TOC structure, so the markdown is usable but flattened.PDF.ai✗ FailedMarkdown export failed on an 18-page table-heavy financial report; the report records "Output MD: Output failed."Reducto✓ WorkedKeeps the syntax clean: no <signature>, <empty>, <b>, <i>, or <u> tags appear anywhere in the output, and the table-of-contents extract renders as a valid pipe table.Tensorlake✓ WorkedReturns the extraction as structured markdown in the Document Markdown workflow, making the output usable rather than a raw text dump.
This evidence is published in
Real inputs and real outputs, no retouching · every cell queryable via API & MCP · aidemos.com
