The output#1
Extend AI
It got the scanned paper through with the main structure intact, including the two-column prose, tables, and chart transcription. But handwriting was only partly read, some between-column notes were lost, and the multirow table headers did not survive cleanly.
d8ec36e3083b4aec937c3063a2eda6da.png
The output#2MDresearch-media-llamaparse-scanned-pdf-output-8ff4647f2bcd.mdopen raw ↗ LlamaParse
It handled the scanned paper’s text flow and several tables well, but the harder grouped-header table and chart were only partly reconstructed, with visuals turned into a table-like form.
research-media-llamaparse-scanned-pdf-output-8ff4647f2bcd.md
The output#3ZIPresearch-media-mistral-ai-scanned-pdf-output-zip-c2a22ad51169.zipopen raw ↗ Mistral AI
It got through the scanned research paper end-to-end, recovered the text, tables, and chart assets, and returned page-wise markdown. The opening page hierarchy and the tougher tables were not kept faithfully enough for a higher score.
research-media-mistral-ai-scanned-pdf-output-zip-c2a22ad51169.zip
The output#4
Tensorlake
On the scanned research paper, it kept the reading order and pulled chart data into structured form, but the hierarchical tables were unreliable. That makes the result useful, though uneven on dense scanned pages.
4e52cca1198a46a2b5390735aa8882f5.png
The output#5
Reducto
It read the scanned pages, charts, and logo well enough to stay useful, but it hallucinated 'USA' into the title, displaced the byline, and badly corrupted the densest table.
7da94747d02a4142be818fe0b54106fb.png
The output#6MDresearch-media-nutrient-scannedpdf-output-04bb5fcf9fd2.mdopen raw ↗ Nutrient
It successfully parsed the scanned paper into markdown, preserved a multi-column section and a grouped table, but it misordered the first page, flattened the chart, and broke the denser table.
research-media-nutrient-scannedpdf-output-04bb5fcf9fd2.md
The output#7
Landing AI
It did a respectable job on the scanned paper, keeping the two-column section and the main diameter-class table readable, but the opening page hierarchy was off, the chart was text-only, and OCR picked up some noisy tokens.
baf588cebc964bc4b11f12f4627b8cf9.png
The output#8
Upstage AI
It read the scanned article text well and even pulled figure values into a chart summary, but the two-column flow and grouped table headers were still shaky.
2e845e61d6c94006bb897c7bf8b2720e.png
No output file capturedThe written finding remains available in Evidence.
PDFVector
It successfully read the 12-page scanned paper and produced a long extracted-text preview in about 20.6 seconds.
Written result only
The output#10MDresearch-media-scanned-research-pdf-pages-1-to-6-output-851ae2ab3972.mdopen raw ↗ Adobe API
After the file had to be split to fit the upload limit, it produced Markdown for both halves and kept the main OCR text, tables, and chart, but it lost section boundaries and broke one table when intervening text appeared.
research-media-scanned-research-pdf-pages-1-to-6-output-851ae2ab3972.md
Result not recorded per promptPDF.ai was tested, but its results were written up across all 3 prompts together rather than prompt by prompt.
PDF.ai
We did not run a recorded scanned research paper case here, so there is no basis to judge how it handles that input.
Covered run-wide
Comments (0)