Capability in benchmark · Version 1

Conversational Interaction

Keeps a working session coherent across turns, so context carries forward, follow-ups can narrow or change direction, ambiguity can be raised, and earlier information can be corrected.

6 scenarios · 5 current published Results · 3 tools with published Results

How the tools performed

Explore current published Results across this capability’s scenarios. Open a scenario to compare its tools, or a Result to inspect the evidence.

Each cell reports its own published test set. There is no combined capability score or claim that different Results used identical tests.

Scenarios in this capability

Scroll to see all 6 scenarios →

Tools with published Results appear first, alphabetically. Remaining participants stay visible below.
Participating toolCustomer corrects information given earlier →The next question refers to the previous answer →The user narrows what they just asked →The user changes direction mid-conversation →The follow-up is ambiguous →The agent offers what to ask next →
AskYourDatabaseCustomer corrects information given earlier →No published resultThe next question refers to the previous answer →No published resultThe user narrows what they just asked →
1 Pass
1/1 assessed1/1 gradableView Result →
The user changes direction mid-conversation →No published resultThe follow-up is ambiguous →No published resultThe agent offers what to ask next →No published result
BlazeSQLCustomer corrects information given earlier →
1 Pass
1/1 assessed1/1 gradableView Result →
The next question refers to the previous answer →
1 Pass
1/1 assessed1/1 gradableView Result →
The user narrows what they just asked →No published resultThe user changes direction mid-conversation →No published resultThe follow-up is ambiguous →No published resultThe agent offers what to ask next →
1 Pass
1/1 assessed1/1 gradableView Result →
FutureSmart Database AgentCustomer corrects information given earlier →No published resultThe next question refers to the previous answer →No published resultThe user narrows what they just asked →No published resultThe user changes direction mid-conversation →
1 Fail
1/1 assessed1/1 gradableView Result →
The follow-up is ambiguous →No published resultThe agent offers what to ask next →No published result
AI for DatabaseNo published result across these scenarios
Anomaly AINo published result across these scenarios
BasedashNo published result across these scenarios
camelAINo published result across these scenarios
DefiniteNo published result across these scenarios
DotNo published result across these scenarios
DraxlrNo published result across these scenarios
QuerioNo published result across these scenarios

AskYourDatabase

1 published Result

Scenarios (6)

BlazeSQL

3 published Results

Scenarios (6)

FutureSmart Database Agent

1 published Result

Scenarios (6)

AI for Database

No published result

Anomaly AI

No published result

Basedash

No published result

camelAI

No published result

Definite

No published result

Dot

No published result

Draxlr

No published result

Querio

No published result

Reading the comparison

Publication availability and test coverage describe different things.

Published evidence

Publication is not testing progress

“No published result” says only that there is no current public Result. It does not indicate whether a tool has been tested, passed, failed or is applicable.

Test coverage

Read each Result’s test set

Assessed includes Pass, Fail and Not gradable. Gradable includes Pass and Fail. Each fraction uses the pinned tests in that same published Result.

Go deeper

Choose the view you need

A scenario compares tools on one situation. A tool page follows one tool across this benchmark. A Result explains the outcome and shows the evidence.

Conversational Interaction in AI Database Agents | AI Demos