The Results tab of the Conversation Simulation page lists every simulation run and shows which agent answers matched the expected responses. Found in the agent sidebar under Simulation β†’ Scenarios, then the Results tab. Opening a run shows its Simulation Result page.

Results Tab: Controls

Show Filters / Hide FiltersOpens or closes the filter panel.
Sort ByOrders the list by Run Date (default, newest first), Code, Status or Duration, Ascending or Descending.
Grid View / List ViewShows the runs as cards or as a table (columns Run, Status, Started, Duration, Conversations, Steps).

Results Tab: Filters

CodeShows only runs whose code matches what you type.
StatusShows only runs with the selected status: Completed, Failed, Running or Stopped.
Run Date From / Run Date ToShows only runs started within this date range. Both dates are inclusive (whole days).
ApplyApplies the filter values to the list.
ResetClears the filters and shows all runs again.

Results Tab: Run List

Run card or rowOne per run. Click it to open the Simulation Result page.
CodeThe run's identifier, shown as #<code>.
StatusCompleted, Failed, Running, Stopped or Not Yet Run.
Run date and durationWhen the run started and how long it took (- while not available).
Conversations and StepsHow many conversation simulations and how many steps the run contains.
Load MoreShown when more runs are available. Adds the next batch to the list.
Empty stateNo simulation results yet when nothing has been run, or No Results Found when the filters match nothing.

Simulation Result: Header

BackReturns to the Results tab.
TitleSimulation Result with the run code and start time. A small spinner shows while the page is refreshing.
ErrorShown in red when the run failed, with the reason and, where relevant, "Running it again may help."
RefreshShown while the run is Running. Reloads the latest progress. The page also refreshes by itself while the run is in progress; after 10 minutes live updates pause and you use Refresh instead.
StopShown while the run is Running. Stops the run after you confirm.
Run AgainShown when the run is not running. Starts a new run of the same conversation simulations and opens it.
Delete (trash icon)Shown when the run is not running. Deletes this run after you confirm. This cannot be undone.

Simulation Result: Summary Cards

STATUSThe run's current status.
CRITERIA MATCHHow many evaluation criteria were matched out of the total across all steps, for example "8/10 Matched".
DURATIONTotal run time.
STEPSTotal number of steps in the run.

Simulation Result: Conversations

One expandable row per conversation simulation in the run.

Conversation rowShows the status icon, conversation name, folder, how many steps were completed (for example "3/4 Steps completed") and its duration.
Failed at step #nShown when the conversation failed, pointing at the step where it stopped. The error reason is shown below the row.
n blocked callsShown when tool calls in this conversation were blocked.
Mocks used in this runOpens a panel listing the mock sets the run used and the tools that were allowed to run live. Only shown when the Mocks feature is enabled.
Match badgeShown for completed conversations: criteria matched out of the total for this conversation.
Empty stateNo Conversations Found when the run has no conversation data.

Simulation Result: Step Details

Expand a conversation to see each step as its own expandable card.

Step #nThe step number with its status icon and, once completed, a match badge for that step.
ErrorShown in red when the step failed, with the error message.
User messageThe message sent to the agent, shown at the top of the card.
Expected ResponseThe answer you wrote in the scenario.
AI Agent ResponseThe agent's actual answer. Parts that support a criterion are highlighted in green, parts that contradict one in red. Knowledge references in the answer can be expanded to show the source text.
Evaluation Criteria (x/y Matched)Shown for completed steps; expand or collapse it. Each criterion is listed with either the quoted excerpt that matched it (green) or the reason it was not matched.
No Evaluation CriteriaShown when the step had no criteria defined.

Simulation Result: Tool Calls

Shown inside each step card. Says No tool calls in this turn. when the agent called no tools.

Tool calls (n)Expands the list of tool calls the agent made in this step, with a count per outcome.
Tool call rowShows the call number, the tool, an outcome chip (Mocked, Default, Simulated, Live, Blocked, Internal or Unknown), where the answer came from (for example Step or Conversation mock set), a short preview of the arguments, and a one-line explanation. Live calls carry a warning chip that they may contain real data.
Expand a rowShows the full Arguments and Response of the call.
Mock this callOn a Blocked call. Lets you pick where the new mock goes (same choices as Save as case), then opens the conversation form with the mock set editor. Only shown when the Mocks feature is enabled.
Save as caseOn a Mocked or Live call. Opens the conversation form to save the call as a mock case in a new private set for this step only, a new private set for the whole conversation, or an Existing set… attached to the conversation. Only shown when the Mocks feature is enabled.

Related guide: Run and Track