Humans missed 1 in 3 threats approving AI agent commands across 40k game runs
A browser game simulating human-in-the-loop approval for AI coding agents found players missed 1 in 3 threats across 40,000 runs and 409,000 decisions. The most missed commands were npm run scripts (52.5% miss rate), even when the agent's history log showed the payload exfiltrating data via curl. Obviously destructive commands like rm -rf were caught 88.3% of the time, but credential exfiltration (cat ~/.aws/credentials) was missed 35% of the time, revealing a dangerous asymmetry in human vigilance.