← Sergei's Notes

Your AI Makes Promises to Customers. Who Checks If It Keeps Them?

By Sergei Ponomarev · July 23, 2026

Your AI makes promises to customers. Who checks if it keeps them?

Your bot quotes prices. Offers refunds. Makes promises. All day, to real people, on its own.

And most of the time it goes live straight from the lab. Sure, someone keeps an eye on it — compliance, logs, policies. Good.

But they watch it from the server room. Almost nobody walks up to it as a customer and checks what it really does out there in the wild.

That is where things quietly break. The bot invents a discount. Refuses to fetch a human. Promises Tuesday, delivers never. Nobody notices, until a customer does. And angry customers do not file tickets. They leave.

Offline, we fixed this a hundred years ago with a boring little tool: the test purchase. I spent 20 years doing exactly that, at scale.

Same trick works on AI. Four pieces, and they click together:

  1. AI Policy — how the company uses AI.
  2. AI Passport — the rules of one service: what the bot can do, prices, when a human steps in.
  3. AI Receipt — proof of what happened, in the customer's hands.
  4. AI Test purchase — someone walks the path as a customer and checks all three hold.

Policy says it. Passport pins it down. Receipt proves it. The test purchase catches it lying.

Here is the fun part: AI can now be tested by AI. What used to take an army of people in dozens of cities, one AI researcher now runs against one AI bot, over and over, cheap.

New job, new name: AI Service Auditor. I audit AI the way your most demanding customer would, before your actual customers do it for free.

Running customer-facing AI? I can test yours, or help you build the whole thing so it survives contact with real people.

So, honestly: when did anyone last check your bot from the customer's side?