Skip to content
Hot eventLive

XGenAct论文:几何增强世界动作模型

1 reports1 sources16 hr ago updated

Get the story

AI overview

2026年10月5日,研究者在arXiv(Robotics)发表论文,提出世界动作模型XGenAct。该模型将RGB观测、机器人动作、度量深度、表面法线和功能角色分割统一表示为RGB视频,用单一视频diffusion transformer和单一目标完成跨空间时序预测,无需模态专用头部。在RLBench留出任务上,其五任务外部比较成功率达52%,强于最强基线的26%,并更准确预测未来深度与分割。目前报道仅此一篇,尚无后续验证或矛盾信息。

Generated from reports · updated 16 hr ago

Timeline

Follow the coverage from different angles.

Oct 5, 2026
  1. arXiv · Robotics
    XGenAct:通过跨任务生成实现几何增强的世界动作模型

    研究者提出世界动作模型 XGenAct,将 RGB 观测、机器人动作、度量深度、表面法线和功能角色分割统一表示为 RGB 视频,用单一视频 diffusion transformer 和单一目标完成跨空间时序预测,无需模态专用头部。在 RLBench 留出任务上,其五任务外部比较成功率达 52%,强于最强基线的 26%,并更准确预测未来深度与分割。

Heat trend

Current heat 6·Comparable peak 10(Oct 5)·Comparable change over 24 hours –

02.557.510Oct5Oct5Oct5Oct6

The trend compares only the same participants observed continuously; its range may be smaller than the current heat count. Move or click on the chart to inspect hourly heat; use the left and right arrow keys to switch.