Benchmark · Version 1

Converting a complex PDF into clean Markdown with a hosted API

This benchmark helps buyers choose a hosted API for PDF-to-Markdown conversion.

Benchmark overview

8Capabilities
12Scenarios
13Participating tools
What is included and excluded

Evaluates whether a hosted API turns complex PDFs into faithful, usable Markdown without dropping, reordering, or inventing content.

A reader learns whether the API preserves the document's text, order, structure, tables, figures, scanned pages, equations, and code, and whether it handles failures honestly when the source cannot be represented cleanly.

In scope

  • Conversion of complex, real-world PDFs into faithful Markdown.
  • Preservation of text, reading order, headings, section structure, tables, figures, charts, scanned text, equations, and code.
  • Honest fallback when a source element cannot be expressed cleanly in Markdown.

Out of scope

  • Structured field extraction.
  • Web-page conversion.
  • Form filling.
  • Handwriting.
  • Document classification and routing.
  • PDF creation or editing.

Participating tools

Tools in this benchmark’s public roster. Publication availability is not a performance ranking.

ToolPublished ResultsExplore
Adobe PDF Extract API1 scenario with a published ResultView tool in this benchmark →
Extend AI1 scenario with a published ResultView tool in this benchmark →
Mistral AI9 scenarios with published ResultsView tool in this benchmark →
Reducto5 scenarios with published ResultsView tool in this benchmark →
FirecrawlNo published resultView tool in this benchmark →
GPT-5.6 terraNo published resultView tool in this benchmark →
Landing AINo published resultView tool in this benchmark →
LlamaParseNo published resultView tool in this benchmark →
Nutrient.ioNo published resultView tool in this benchmark →
PDF VectorNo published resultView tool in this benchmark →
TensorlakeNo published resultView tool in this benchmark →
UnstructuredNo published resultView tool in this benchmark →
Upstage AINo published resultView tool in this benchmark →

Capabilities & scenarios

12 scenarios grouped by 8 capabilities. Open a group to explore its scenarios in this benchmark.

Text Fidelity1 scenario

Ordinary digital text survives completely and correctly in the Markdown output.

Capability in this benchmark → · Global definition →

  1. An ordinary digital text document1 tool with a published Result
Reading Order & Layout1 scenario

The source reading sequence survives the spatial-to-linear conversion without the columns being interleaved or reordered.

Capability in this benchmark → · Global definition →

  1. A page laid out in multiple columns1 tool with a published Result
Heading & Section Structure2 scenarios

Document structure survives as real Markdown headings and sections, with nesting and footnotes kept in the right relationship.

Capability in this benchmark → · Global definition →

  1. A document with styled headings and subheadings1 tool with a published Result
  2. Footnotes at the bottom of the page1 tool with a published Result
Table Extraction2 scenarios

Rows, columns, headers, and values survive as a table with the right relationships intact.

Capability in this benchmark → · Global definition →

  1. A simple, clearly formatted table1 tool with a published Result
  2. A table that continues across a page break1 tool with a published Result
Figures & Charts2 scenarios

Visual content survives with its position, caption or title, and a text-reachable trace for later use.

Capability in this benchmark → · Global definition →

  1. A document with figures and captions1 tool with a published Result
  2. A document with a data chartNo published results
Scanned Document OCR2 scenarios

Text that exists only as pixels is recovered into usable output instead of being silently lost.

Capability in this benchmark → · Global definition →

  1. A cleanly scanned document3 tools with published Results
  2. A document mixing digital and scanned pages2 tools with published Results
Equations & Mathematical Notation1 scenario

Mathematical notation survives as math markup or an honest fallback, without being turned into misleading prose.

Capability in this benchmark → · Global definition →

  1. A document containing mathematical equations2 tools with published Results
Code Extraction1 scenario

Code survives as a preformatted fenced block with its whitespace and line structure intact.

Capability in this benchmark → · Global definition →

  1. A document containing a code block2 tools with published Results

Results overview

Current published evidence in this benchmark.

16Current published Results
11Scenarios with published Results
4Tools with published Results

Publication availability is separate from test coverage and unpublished research progress.

Explore scenarios →

How the benchmark works

A public summary of the evaluation method. The same defined scope and evidence standard apply to every tool assessed under this version.

Evaluation unit

One test case, one tool

One test case is scored for one tool at a time.

Outcomes

Pass, fail, or partial

Each run ends pass, fail, or partial. A missing capability scores 0 and stays in the arithmetic — a product must not rank higher by having less product. not_applicable, not_measured and blocked (with the reason, never a 0) are distinct states.

Consistency

Versioned scenarios and resources

Use versioned scenarios and registered resource versions so runs stay comparable.

Evidence standard

Recorded stimulus and complete output

Record the input stimulus, the complete output, and the relevant registered references.

Full benchmark methodology →

Resources and fixtures

The registered material and systems that create a consistent test environment for this benchmark.

Converting a complex PDF into clean Markdown with a hosted API — Benchmark definition | AI Demos