Scalex.dev reports that human reviewers missed roughly one in three threats while approving AI-agent commands across 40,000 game runs. The result highlights a potential weakness in permission-confirmation workflows: users may approve commands despite relevant risk signals. However, the supplied material contains only the article title and a short abstract. The threat taxonomy, game design, participant count, baseline conditions, statistical analysis, and reproducibility details remain unverified. Hacker News engagement was limited to a score of 3 and one comment.
No heat snapshots are available in the last 24 hours.