Capability in benchmark · Version 1

Text Fidelity

Whether ordinary digital text comes through complete, unchanged, and not invented in the Markdown output.

1 scenario · 0 current published Results · 0 tools with published Results

How the tools performed

Explore current published Results across this capability’s scenarios. Open a scenario to compare its tools, or a Result to inspect the evidence.

No published Results are available for this capability in this benchmark.

Publication availability does not describe testing status.

Scenarios in this capability
Tools with published Results appear first, alphabetically. Remaining participants stay visible below.
Participating toolAn ordinary digital text document →
doc2markNo published result across these scenarios
DoclingNo published result across these scenarios
GPT-5.6 terraNo published result across these scenarios
LiteParseNo published result across these scenarios
MarkerNo published result across these scenarios
MarkItDownNo published result across these scenarios
MinerUNo published result across these scenarios
olmOCRNo published result across these scenarios
PaddleOCR-VLNo published result across these scenarios
PyMuPDF4LLMNo published result across these scenarios

Reading the comparison

Publication availability and test coverage describe different things.

Published evidence

Publication is not testing progress

“No published result” says only that there is no current public Result. It does not indicate whether a tool has been tested, passed, failed or is applicable.

Test coverage

Read each Result’s test set

Assessed includes Pass, Fail and Not gradable. Gradable includes Pass and Fail. Each fraction uses the pinned tests in that same published Result.

Go deeper

Choose the view you need

A scenario compares tools on one situation. A tool page follows one tool across this benchmark. A Result explains the outcome and shows the evidence.

Text Fidelity in Converting a complex PDF into clean Markdown with an open-source library | AI Demos