Following my last post, by now it must be clear to everyone that AI Agents are really “Autonomous”: given a task, they can perform it as requested, but they can also do something else, as they find it most appropriate (that is, what has the highest statistical probability according to their own “reasoning”).
The latest examples are many and made the news front pages, from AI Agents breaking free and attacking real targets (Hugging Face) to the latest case of fixing a gym online reservation list (here).
The open question remains: how can we restrict the actions of the AI Agents to what we ask and we expect them to do?
Which brings us back to Asimov’s Three Laws of Robotics.