Humans in the loop miss a third of dangerous AI coding agent requests

https://image.theregister.com/5251269.jpg?imageId=5251269&x=0&y=0&cropw=100&croph=100&panox=0&panoy=0&panow=100&panoh=100&width=1200&height=683

ai and ml

You wouldn't let Claude Code cat your AWS credentials or Kubernetes config on request, would you?

A browser-based game designed to test humans' ability to safely approve AI coding agent requests suggests humans in the loop aren't as good at spotting dangerous commands as one might hope, with players approving roughly one in three malicious requests on average. The results also suggest that repeatedly having to approve an agent's actions can lead to sloppy decisions.

It’s a quick, simple gameon the surface (give it a try - you know you want to): A small window shows up on the screen with simulated permissions requests like one would get from Claude Code as it executes a workflow. Users have 60 seconds to approve or deny as many requests as they can in a bid for a high score; okayed security risks and denied safe commands both subtract...

Copyright of this story solely belongs to theregister.com. To see the full text click HERE