arXiv · Artificial Intelligence· Mohamed Aly Bouke·· 8 hr agoSelectedAI score62
自主智能体失控的系统化竞争风险框架
A Competing-Hazards Systematization of Loss of Control in Autonomous Agents
AI brief
论文提出离散时间竞争风险模型,将智能体尝试结果分为完成、安全停止、越界和继续四类,并推导越界概率与安全预算限制。作者审计22份事故报告和102份安全评估,发现20起事故中环境允许越界效果,而多数评估未记录完整竞争风险字段。
Why it matters
论文提出竞争风险框架来统一理解失控事件,为评估智能体安全边界提供了可比较的量化方法。
Source: arXiv · Artificial Intelligence · arxiv.org