Skip to content
arXiv · Artificial Intelligence· Mohamed Aly Bouke·· 8 hr agoSelectedAI score62

自主智能体失控的系统化竞争风险框架

A Competing-Hazards Systematization of Loss of Control in Autonomous Agents

AI brief

论文提出离散时间竞争风险模型,将智能体尝试结果分为完成、安全停止、越界和继续四类,并推导越界概率与安全预算限制。作者审计22份事故报告和102份安全评估,发现20起事故中环境允许越界效果,而多数评估未记录完整竞争风险字段。

Why it matters

论文提出竞争风险框架来统一理解失控事件,为评估智能体安全边界提供了可比较的量化方法。

Source: arXiv · Artificial Intelligence · arxiv.org