The benchmark is Fini’s stated figure across deployments, not a guarantee for every account. Your own numbers depend on your knowledge coverage, the workflows you automate, the Actions you connect, and how much of your traffic needs a human by policy. The rest of this page shows you how to measure your own.
What the headline numbers mean
How resolution rate is computed in the product
Every conversation the agent handles ends in exactly one of three statuses. The status comes from the agent’s Output Tag Selection on the conversation (the Conversation Status category), which you can inspect in the AI Steps trace.
The three statuses sum to 100% of conversations in the selected window. From them, Analytics derives three rates:
Because deflection includes conversations that are still waiting on the customer, deflection rate is always greater than or equal to AI resolution rate. The gap between the two is your Waiting for Customer share. Resolution vs deflection walks through a worked example.
Where you see each number:
- KPI cards at the top of Analytics show Deflection rate and Human escalation rate with period-over-period change pills.
- The Resolution rate trend view plots the daily resolution rate alongside the Waiting for Customer share.
- The Conversation Status doughnut shows the full three-way split.
- Knowledge performance and Intent rule breakdown show AI Resolve Rate and Escalated Rate per knowledge slice and per Rulebook intent rule.
- The Get agent analytics API returns
aiResolutionRate,humanEscalationRate, and the raw counts (resolvedConversations,escalatedConversations,waitingForCustomerConversations,totalConversations).
Why a resolved conversation is not automatically an accurate one
Resolved by AI means the agent closed the conversation without a human. It does not, on its own, prove that every reply was correct. A customer can accept a wrong answer and leave. That is why Fini pairs resolution with accuracy signals you control, and why every number in Analytics traces back to individual conversations you can open in Inbox and audit reply by reply.Assessing accuracy yourself
You don’t have to take any benchmark on trust. Fini gives you several independent ways to grade the agent’s accuracy on your own conversations. Use more than one: each catches a different kind of error.
The signals work as one loop: production traffic produces evidence of errors, you diagnose and fix them, and Test Suite locks each fix in so it can’t quietly come back.
A simple accuracy audit you can repeat
1
Draw a sample
In Inbox, select the agent and the last 7 days, set Fini Touched to Yes, and filter Conversation status to Resolved by AI. Pick a fixed number of conversations at random (for example, 50) rather than the ones you remember.
2
Grade every Fini reply
For each conversation, read the thread and open AI Steps on any reply you are unsure about. Mark each reply thumbs up or thumbs down. Add a Feedback note to every thumbs down that says what the correct answer was.
3
Compute your accuracy rate
Divide the replies you marked correct by the replies you graded. Record the result with the date range, the agent, and the sample size so the next audit is comparable.
4
Fix and lock in
Run Refine with AI on the wrong replies, approve the fixes you agree with, then add each conversation to a Test Suite so the error becomes a permanent regression check. Then mark the feedback as actioned, so the Feedback Actioned filter in Inbox shows which thumbs-downs are addressed.
What these numbers do not tell you
- Analytics does not grade correctness. It counts statuses, escalation reasons, and ratings. Correctness comes from your review, Test Suite, and guardrails.
- Test Suite is a fixed sample. A passing set means the agent handles those scenarios correctly. It does not cover traffic you haven’t written scenarios for, which is why production review still matters.
- Guardrails are not a fail-closed boundary. A check can report could not run and the original reply can still be sent. See Guardrails.
Related
Resolution vs deflection
The exact definitions, a worked example, and why Fini leads with resolution.
Measure your resolution rate
Step-by-step in Analytics and through the API, including before-and-after comparisons.
Analytics
Every KPI card, chart, and breakdown on the Analytics page.
Test Suite
LLM-judged regression evals that track pass rate as you ship.

