聚合到达-规避机会约束下有限MDP智能体群体策略合成
Get the story
2026年10月9日,arXiv Multiagent Systems 发布论文,提出面向有限规模智能体群体的策略合成方法,处理聚合到达-规避机会约束:保证至少比例 α_r 的智能体在时刻 t* 到达目标区域,同时不安全群体比例低于 β_u。方法沿均值场轨迹,通过离散时间 Lyapunov 递推传播经验密度的二阶矩,并利用 Cantelli 不等式将机会约束转化为可处理的确定性矩条件。该研究为有限种群场景下的群体策略合成提供了理论框架,目前未见后续验证或应用报道。
Generated from reports · updated 3 hr ago
Timeline
Follow the coverage from different angles.
- arXiv · Multiagent Systems聚合到达-规避机会约束下有限 MDP 智能体群体的策略合成
论文提出针对有限规模智能体群体的策略合成方法,在聚合到达-规避机会约束下保证至少比例 α_r 的智能体在 t* 时刻到达目标区域且不安全群体比例低于 β_u。方法沿均值场轨迹通过离散时间 Lyapunov 递推传播经验密度的二阶矩,并用 Cantelli 不等式将机会约束转化为可处理的确定性矩条件。
Heat trend
Current heat 9·Comparable peak 10(Oct 9)·Comparable change over 24 hours –
The trend compares only the same participants observed continuously; its range may be smaller than the current heat count. Move or click on the chart to inspect hourly heat; use the left and right arrow keys to switch.