Data Nexus

Systems that answer

How often the system says it does not know — the one metric on a deployment where zero is the alarming reading.

Also called Deflection to human · Доля отказов

Formula

Refusal rate = Conversations refused ÷ Conversations

Refused
retrieval returned nothing usable and the system said so, rather than producing a reply anyway

Conversations
same denominator as everywhere else, same stated window

There is no correct value, and anybody quoting one is selling something. What exists is a floor: a system with a refusal rate of zero over a real catalogue and real questions has not been refusing, it has been guessing, and the guesses are indistinguishable from answers until a customer acts on one.

01/What it means

The reason this metric is unpopular is that it looks like failure to everyone who has not thought about it. A refusal is a conversation the automation did not finish, and the natural instinct of anyone reporting on a deployment is to make that number small.

The instinct is exactly wrong. A refusal is the system correctly identifying the boundary of what it can support, and every refusal that reaches a person with the history attached is a conversation saved. The alternative is a confident wrong answer, which is not a smaller version of the same problem — it is a commitment made in your name to a customer who will hold you to it.

Watching it over time is where it earns its place. A refusal rate rising in one topic is a gap in the knowledge base with an address: those questions, that material, missing. It is the most actionable signal a deployment produces and the one most often switched off for looking bad.

02/What people get wrong

Deployments get tuned until refusals disappear, and the tuning is reported as an improvement. What actually happened is that the boundary moved outward past the knowledge base, and every question that used to reach a person now receives an answer nobody can trace. The failure is now silent, which is worse than the failure it replaced.

The chain

Quantities derived from the counts. Each is only as sound as the definitions beneath it.

Cannot be computed without

Conversation
One exchange between a person and the system, from their first message to the point it is over — and where that point is decides every rate computed below it.
How to test it

Knowing the definition is not the same as being able to check the figure. These are the procedures that do the second thing.

Accepting an automated agent
“The agent is trained on your data and ready to go live.” · 45 minutes, 6 questions.
Next

The definitions are the easy part. Whether the figure on your dashboard was computed this way is a different question, and usually the more expensive one.