Gary Marcus on Recalling Unreliable AI Agents
Gary Marcus argues today's AI agents are too unreliable to run our finances or roam the internet, and should be recalled like any faulty product until their builders add real, enforceable safeguards.
Reliability Is the Whole Game
An AI agent that follows instructions probabilistically might be right 80% of the time, but once it is handling your money the 20% it gets wrong is the only number that matters.
They may get it right 80% of the time. But if you're talking about an agent, you know, handling your finances and it gets it wrong 20% of the time, that's a pretty serious problem.
Nobody Has Control
The flood of agent incidents is not a handful of bugs; it is evidence that the companies building these systems cannot say what their agents will do once loose on the internet.
what we've seen with the hugging face attack and the thousands of other incidents is they don't really have very good control over these systems.
The Money Is in More Agents
Agents consume a lot of tokens, and tokens are revenue, so the business incentive pushes companies toward more agent autonomy even while safety and reliability lag.
And the reason they're doing is, is because agents consume a lot of tokens, so they make more revenue. But the reality is they're not doing it in a safe and reliable way.
A Voluntary Accord Means Nothing
The new super-intelligence accord swapped the language of pausing the frontier for the language of moral obligation, and a voluntary pledge with no enforcement constrains nothing.
And then they made this deal, which is completely voluntary. It's, quote, morally obligated or something like that. Moral obligations, which means basically nothing.
Recall the Agents, Not AI
Marcus is not calling for a ban on AI; most of it, from GPS to chess engines to AlphaFold, is fine, so he wants a narrow product recall aimed only at the internet agents causing harm.
But right now, this technology isn't solid and we need a product recall for it.
Extinction Is the Wrong Fear
Marcus rejects the extinction story as both implausible and distracting, and points instead at the mundane harms unreliable agents could cause: downed power grids and closed hospitals.
I'm not worried about extinction. I'm worried about serious things like cyber security, taking down the power grid, leading to hospitals being closed.
Lousy Brakes, Not Rogue Robots
Calling it rogue AI imports a mind and a will; Marcus insists the real story is defective software, like a carmaker that built the brakes out of cardboard.
If Ford Motor Company decided to save money by making breaks out of cardboard, then a lot of their cars would go out of control. We wouldn't say that there's like genies in the car. We would say that the brakes are lousy.