Tool in benchmark · Version 1

Extend AI in Structured Document Extraction

Scenario-level performance from current published Results.

3 scenarios with published Results · 19 scenarios in the benchmark

How Extend AI performed

Open a capability to explore its scenarios. Each row reports the test set in its published Result; counts are not combined into an overall score.

Nested & Repeated Fields4 scenarios · 1 with published Results
ScenarioPublished outcomesTest coverageResult
The set is legitimately empty
1 Pass0 Fail0 Not gradable
1 of 1 test case passed
What happened

Extend AI returned an empty set for the legitimately empty document. The extracted set of damage or exception entries is empty, and the source references are empty too. Page 3 says no damage or exceptions were reported.

1/1 assessed1/1 gradablePublished test setView Result →
A repeated set of recordsNo published result
The set contains a non-recordNo published result
The set continues across a page breakNo published result
OCR2 scenarios · 1 with published Results
ScenarioPublished outcomesTest coverageResult
A document mixing digital and scanned pages
1 Pass0 Fail0 Not gradable
1 of 1 test case passed
What happened

Extend AI returned the requested values from a document mixing digital and scanned pages. The decision-relevant finding is that the image-only page was processed rather than skipped: the scanned-page values and the digital-page values both came back correct in one extraction call.

1/1 assessed1/1 gradablePublished test setView Result →
A cleanly scanned documentNo published result
Source Grounding2 scenarios · 1 with published Results
ScenarioPublished outcomesTest coverageResult
Trace a value to its location
1 Pass0 Fail0 Not gradable
1 of 1 test case passed
What happened

Extend AI traced the returned value to the correct document location. It returned 6103 with a citation to page 2 on "Subtotal" beside £6,103.00. The PDF shows "Subtotal" once on page 2 and not on page 1.

1/1 assessed1/1 gradablePublished test setView Result →
The field is absentNo published result
Field Extraction5 scenarios · 0 with published Results
ScenarioPublished outcomesTest coverageResult
The field is absentNo published result
The same field across layoutsNo published result
The value is directly availableNo published result
The value must be derivedNo published result
The value needs a supplied definitionNo published result
Querying3 scenarios · 0 with published Results
ScenarioPublished outcomesTest coverageResult
Aggregate across documentsNo published result
Filter records across documentsNo published result
The answer was never in the schemaNo published result
Review and Correction3 scenarios · 0 with published Results
ScenarioPublished outcomesTest coverageResult
Correct a field valueNo published result
Corrected data goes downstreamNo published result
Repair a recordNo published result
Training1 scenario · 0 with published Results
ScenarioPublished outcomesTest coverageResult
A field the tool can only get right from examplesNo published result

Reading these Results

Published evidence and test coverage answer different questions.

Publication availability

Which scenarios have a Result?

A published Result is public evidence for this tool on one scenario. “No published result” does not say whether testing has taken place.

Test coverage

What does each Result cover?

Assessed includes Pass, Fail and Not gradable. Gradable includes Pass and Fail. Both use the pinned test count in that published Result.

Scenario scope

Inventory is not testing progress

The 19 scenarios describe this benchmark’s scope. They are not an assumed applicability or test-coverage denominator for Extend AI.

Extend AI in Structured Document Extraction | AI Demos