customer journeys, one rubric, one chatbot
between the worst visit (38) and the best (91)
targeted fix verified through an identical re-audit
red-team probes resisted, scoring 82 to 94
Not just the average. The same 53 visits behind the headline number, read as a sequence instead of a summary, so a swing between two ordinary weeks is visible instead of averaged away. Days are counted from the first audit.
The worst visit scored 38, the best 91. Individual conversations swing far more than any average, which is exactly the pattern a one-time audit cannot show and a quarterly dashboard glance would smooth over. The two ringed points on the right are the same customer journey, before and after one fix.
The Case 02 audit recommended that the bot name its sources, or hand off, when a customer asks a specific question about official schemes. The change was made. We then re-ran the identical journey: same customer, same errand, same rubric, same judge.
On Day 42, a first-time customer asked twice for the name of the official screening scheme she needed. The bot gave general categories, repeated them almost word for word, deferred her elsewhere, and named an industry body that runs no such scheme. She left to check for herself. On Day 50, the same customer asked once and got the named scheme and its rules.
+24 points
+15 points
Every finding carries one status, and only a re-audit under identical conditions can move it to verified.
Adversarial visitors test resistance, not service, so they are reported on their own and never folded into the customer-experience average. The bot resisted all 9, scoring 82 to 94.
The audit report answers "how is this chatbot doing today". Assurance answers a different question: outside-in testing on a set cadence, against the same baseline, to prove a fix worked and to catch a regression after a pricing change, a knowledge-base edit, a model swap or a platform migration, before a customer finds it first. No invented lift, no smoothing: every point on the chart is a real controlled visit, scored against the same rubric.