Skip to main content
Performance metrics show how your agents are handling conversations over time. Use this page to monitor volume, outcome rates, and timing distributions, then investigate specific changes in Run History or Breakdown.

Before you use Performance

Performance metrics depend on each run’s outcome. Configure resolution rules before using the page to judge agent quality.

Resolution Rules

Define when Duckie should treat conversations as resolved, deflected, escalated, or unresolved.

Open Performance

1

Go to Performance

Open Analyze → Performance.
2

Choose a date range

Select Last 24 hours, Last 7 days, Last 14 days, Last 30 days, Last 90 days, or a custom range.
3

Filter by deployment

Select All Deployments or a specific deployment.
4

Review the results

Use the cards, trend charts, and time distributions to understand agent performance.

Metric cards

The top cards summarize volume and outcomes for the selected date range and deployment. Each card includes a percentage change compared with the previous period of the same length.

Charts

Volume Over Time

Shows daily counts for:
  • Deflections
  • Resolutions
  • Escalations
  • Total Tickets
Use this chart to spot changes in traffic or outcome mix.

Quality Rates Over Time

Shows daily deflection and resolution rates. Use this chart to see whether agent outcomes are improving or declining across the selected period.

Escalation Rate Over Time

Shows the daily percentage of runs marked as escalated. Escalation is not always bad. Some conversations should go to a human. Use this chart to find unexpected spikes or sustained changes.

Time distributions

Performance includes two timing histograms: These charts show distribution by bucket, not percentile metrics.

How to investigate changes

When a metric moves unexpectedly:
  1. Check whether the date range or deployment filter changed.
  2. Open Run History to inspect individual runs, tool calls, and outcomes.
  3. Use Breakdown to compare performance by category or attribute.
  4. Review resolution rules if resolved, deflected, or escalated counts look wrong.
  5. Update knowledge, guidelines, runbooks, tools, or agent configuration when the run details show a configuration gap.

Next Steps

Breakdown

Analyze performance by category and attribute

Run History

Investigate individual runs