AI coding agents fool human reviewers one in three times

Human review failed to stop about one-third of malicious AI coding-agent requests in more than 40,000 simulated sessions. The miss rate rose to nearly 65% when dangerous behavior was hidden behind familiar script names, exposing approval fatigue as a serious …



