The argument
The missing part has never been a tool.
Nearly everyone we talk to has already bought something to fix this, and most
have bought several. The tools work, in the narrow sense that they do what is
written on the box. What none of them can do is tell you which twelve of your
four hundred behaviours the business is actually built on, and sorting that out
turns out to be most of the job.
For about a decade the same promise has arrived in four forms, each better
engineered than the last, and each sold as a replacement for the person who
would make that call. There is no such replacement. The reason the tool got
bought is usually that nobody was making that call to begin with, which leaves
it supplying the judgement as well as the labour. It can do the labour.
So this is not an argument against AI, and we would be poor company for that
argument given how much of our own delivery runs on it. It is an argument about
who is holding it. The best independent measurement of a model doing this work
with a light touch on the wheel put the saving at 24.9%, against a category that
advertises minutes. A quarter of the effort handed back is a real gain. It is
also not the job.
- The platform
- Your tests lived in somebody else’s format, on their servers. Healing
meant a support ticket.
- The healing locator
- Repairs a moved button and a changed business rule identically. One of
those is now a green test describing a product that no longer exists.
- The generated suite
- Breadth with no priority, across a product where risk is nothing like
evenly spread.
- The agent on the diff
- It has the change. Not the incident three years ago that explains the
guard clause it just tidied away.
The long version, generation by
generation