跳到正文
·
Archived topic · 归档话题,来源已停止追踪

Humans missed 1 in 3 threats approving AI agent commands across 40k game runs

AI 摘要

A browser game simulating a human-in-the-loop for an AI coding agent revealed that players missed one in three threats when approving or denying commands. Across 40,000 plays and 409,000 decisions, even with prior warnings about threats, human oversight proved fallible. This highlights potential challenges in human supervision of AI agents, even in a gamified context.

时间与来源

时间显示为 UTC

显示时区:UTC

本地时区尚不可用,暂时显示 UTC。

收录当时偏移:UTC+02026年8月6日 15:00 UTC

收录
2026年8月6日 15:00
来源类型
未分类