Safety · Hacker News ·
Humans missed 1 in 3 threats approving AI agent commands across 40k game runs
An analysis of 40,000 game runs found that humans missed one in three threats when approving commands issued by AI agents. The findings highlight risks in human oversight of agent permissions.