August 1, 2026
The word 'agent' stopped meaning much this year
Most products marketed as AI agents this year turned out to be automation with a new label.
A word can do a lot of damage once everyone starts using it for something slightly different. That happened to 'agent' this year. Industry analysts estimate that of the thousands of products now marketed as AI agents, only a small fraction, something like one in twenty, actually pursue a goal across multiple steps using real tools without a human approving each one. The rest are chatbots, scripted workflows, or ordinary automation wearing a new label.
The label matters because the buying decision changed underneath it. A year ago the question was whether a vendor had any AI at all. Now the question is whether what they built can finish a multi-step task on its own, and a rebranded script cannot do that, no matter how the demo is staged. The cost of getting that distinction wrong lands on whoever bought the thing.
The test practitioners converged on is simple enough to ask in a vendor meeting: give it a real, multi-step goal, hand it the actual tools it would use in production, and see how far it gets before a human has to step in. Anything that needs a person to approve every individual step is automation with a new name. Most of that automation is genuinely useful. The ask is just to call it what it is.
This is also the harder half of judging your own pilots, beyond a vendor demo. A model version that scores well on average can still be quietly failing the one case that matters, the same way a scripted tool can look agentic in a slide and fold the moment the task branches. Judging either one honestly takes the same discipline: specific cases, specific stakes, checked one at a time.