Visual assets are extracted into page-specific folders, keeping charts, signatures, and other document elements tied to their source pages instead of collapsing them into one consolidated output.
What was measured
Visual Content Retention
Retains charts, figures, diagrams, and images in the output and places them in the correct reading position.
decisive for this rankingtransformation
For complex PDFs, figures, charts, and images are part of the document content, so preserving them in the right place is part of the core job. (3 of 3 judges)
What was given, what came back
Test input: Hybrid Earnings Report · pdf · group: hybrid-earnings-report
Input — what we sent
07829e6dcb414445886e618cd047efb2.pdf
Hybrid Earnings Report
A long, real-world hybrid annual report used to benchmark end-to-end PDF-to-markdown conversion. It combines native text, financial tables, charts/graphics, and scanned signature/stamp regions, stressing preservation of document structure, reading order, and embedded visual content across a multi-section report.
Why this input is hard
- · Native digital text extraction
- · Complex financial table preservation
- · Chart and graphic retention
- · Image-based signatures and stamps
- · Reading order across a long multi-section report
- · Markdown quality and consistency
Output — unretouched


Also checked on this input — same tool, 5 other criteria
Complex Document Handling✓ WorkedThe tool handles an 84-page mixed-content annual report end-to-end, including tables, charts, and scanned signature/stamp regions, and returns usable markdown output without manual correction.Markdown Quality✓ WorkedThe export is usable markdown rather than a flat text dump: it includes an overall markdown file plus page-wise markdown files delivered in a downloadable ZIP.Reading Order & Structure✓ WorkedThe parser preserves document hierarchy across most sections of the long report, keeping headings and supporting text structurally aligned through the majority of the output.Reading Order & Structure◐ MixedHeading hierarchy is reconstructed inconsistently, so some parts of the document are flattened even though the underlying content is still recovered.Table Preservation✓ WorkedThe financial table is reconstructed into a usable structured layout without losing the relationships between headers, rows, and corresponding values.
Provenance
- Observation
- 457e867a-acde-40fc-ad65-8f9b5f51378d
- Evidence run
- 6e3160de-fe46-4b45-b071-72560b5c5d0e
- Study
- Convert a Complex PDF into Clean Markdown with an API
- Research task
- 86b9h7t37
- Tested at
- not recorded
- Source
- first-party
- Evidence state
- verified
- Proof shown
- input + output shown
- Cost / latency
- not captured
- Repeat run
- not captured
- Tester
- not captured
The last three rows are honest blanks, not placeholders — our capture has no field for them yet.
Query this
get_evidence({
tool: "mistral-ai",
scenario: "hybrid-earnings-report"
})MCP · mcp.aidemos.com/api/mcp
Free with attribution.
Same input, same check — 8 other tools
measured on Visual Content Retention
Adobe API✗ FailedDrops handwritten signature images from a scanned signatures page while retaining the surrounding printed legal text, so the signature content is missing from the extracted output.Extend AI◐ MixedRetains charts and logos as standalone figure blocks with captions, but they are extracted out of page position rather than embedded inline.Landing AI✗ FailedReplaces a waterfall chart with text, preserving the SG&A rate values and year-to-year change labels but not the chart as a visual asset in the output.LlamaParse◐ MixedConverts a waterfall chart into a text/table representation, preserving the SG&A rate sequence and deltas but not the chart as a visual object.Nutrient.io✗ FailedLoses the visual semantics of the SG&A waterfall chart; the 2013–2015 rate values are present, but the chart type, axes, and legend relationships are not preserved.Reducto✓ WorkedRetains visual content as real image outputs when requested: the portrait, donut chart, and bar chart are all returned as preserved figure/image crops rather than being dropped.Tensorlake✓ WorkedRetains scanned signature-page visuals as figure blocks plus signer text, including named executives and dates, rather than dropping the signature regions from the output.Upstage AI✗ FailedDoes not retain handwritten signature content as visual output on the signatures page, leaving only partial surrounding legal text and no clearly identifiable signatures.
This evidence is published in
Real inputs and real outputs, no retouching · every cell queryable via API & MCP · aidemos.com