"Better than a human."
Which human, measured how?
Every AI feature was approved on some version of that sentence. Almost nobody who
said it had measured the human first.
There's rarely a baseline for the process the AI replaced: no error rate for the
people who used to do it, no count of how often the old way got it wrong. Without
that number, "better" is just a mood. So the first thing we measure on any
engagement is the humans. Sometimes the AI genuinely beats them, and then you can
finally prove it. That proof is the difference between "we think it's fine" and
"we can show it's fine", and only one of those survives a board meeting.
The sentence also moves responsibility, quietly. A person's mistake is one
mistake, caught by the colleague who checks their work and covered by your
professional indemnity policy. The model's mistake repeats on every matching
request until someone notices, and it's a defect in your product rather than an
employee error. A tribunal has already held an airline to what its chatbot
promised a customer. Worth asking your insurer whether AI output is covered;
the pause before the answer is informative.
A person gets it wrong
The model gets it wrong
01
One mistake, one case
The same mistake, every matching request
02
A colleague catches it
A customer reports it
03
Judgment you hired, vetted and insured
Output your name is on
04
An awkward conversation
A liability conversation
If the company carries your name, read that right-hand column again. The work your
team ships with AI is signed by you, whether or not you've seen it.