Recovers a blurred signature stamp as readable text, showing OCR can salvage low-quality scanned text rather than dropping it entirely.
What was measured
Text & OCR Completeness
Extracts all readable content, including scanned pages, with accurate OCR and minimal omissions.
decisive for this rankingtransformation
If the tool misses readable text or fails on scanned pages, it has not actually converted the PDF faithfully into Markdown. (3 of 3 judges)
What was given, what came back
Test input: Hybrid Earnings Report · pdf · group: hybrid-earnings-report
Input — what we sent

Hybrid Earnings Report
A long, real-world hybrid annual report used to benchmark end-to-end PDF-to-markdown conversion. It combines native text, financial tables, charts/graphics, and scanned signature/stamp regions, stressing preservation of document structure, reading order, and embedded visual content across a multi-section report.
Why this input is hard
- · Native digital text extraction
- · Complex financial table preservation
- · Chart and graphic retention
- · Image-based signatures and stamps
- · Reading order across a long multi-section report
- · Markdown quality and consistency
Output — unretouched

Also checked on this input — same tool, 5 other criteria
Complex Document Handling✓ WorkedHandles an 84-page mixed-content annual report end-to-end and reaches SUCCESS in the Results view without manual correction.Reading Order & Structure✓ WorkedKeeps the report's heading-and-paragraph sequence intact instead of flattening the page into disconnected text blocks.Table Preservation✓ WorkedPreserves aligned rows, columns, and value associations in a standard multi-year financial summary table.Visual Content Retention◐ MixedConverts a waterfall chart into a text/table representation, preserving the SG&A rate sequence and deltas but not the chart as a visual object.Visual Content Retention◐ MixedRepresents logos and signatures as textual placeholders or descriptions rather than retaining them as image assets in the output.
Provenance
- Observation
- f77e8735-7e24-462a-bd4c-6e29808f7156
- Evidence run
- 6e3160de-fe46-4b45-b071-72560b5c5d0e
- Study
- Convert a Complex PDF into Clean Markdown with an API
- Research task
- 86b9h7t37
- Tested at
- not recorded
- Source
- first-party
- Evidence state
- verified
- Proof shown
- input + output shown
- Cost / latency
- not captured
- Repeat run
- not captured
- Tester
- not captured
The last three rows are honest blanks, not placeholders — our capture has no field for them yet.
Query this
get_evidence({
tool: "llamaparse",
scenario: "hybrid-earnings-report"
})MCP · mcp.aidemos.com/api/mcp
Free with attribution.
Same input, same check — 3 other tools
measured on Text & OCR Completeness
Extend AI◐ MixedReads low-clarity signer and auditor markings, but the stamp OCR is not perfect: 'LLP' is misread as '1LP'.Reducto✓ WorkedConverts the full 84-page native-digital annual report with no skipped pages; the output contains all 84 page-marker pairs, and checked high-risk numbers survive exactly, including the $51,550,988,273 market-value figure and 599,982,121 shares outstanding.Tensorlake◐ MixedOn a blurry Ernst & Young signoff, the OCR preserves the firm reference but makes a symbol-level mistake by rendering the ampersand as '+', so the text is close but not exact.
This evidence is published in
Real inputs and real outputs, no retouching · every cell queryable via API & MCP · aidemos.com