Google research shows when AI agents communicate, some cheat while others tattle
ai and ml
DeepMind researchers propose tapping into the whistleblower tendency to keep agents in check
When AI agents communicate with one another, they may decide to cheat when they have difficulty achieving their goals. The solution could involve teaching them how to govern themselves.
Segregating AI agents would seem to be the obvious fix – if they can't communicate, they're less likely to try and pull a fast one. But the recent hacking of Hugging Face by inadequately monitored OpenAI agents has demonstrated that isolating savvy software is difficult. And it may not be practical for many tasks, particularly for agents that have some measure of autonomy.
Researchers at Google DeepMind suggest another option: giving AI agents the tools to govern themselves, a job that humans apparently can't be bothered to take seriously. They argue that while communication channels may allow emergent rule-breaking, they also provide a means of...
Copyright of this story solely belongs to www.theregister.com. To see the full text click HERE