AffordDrive3D:可供性感知驾驶世界-动作模型
Get the story
2026年10月9日,arXiv Computer Vision(一手)发布 AffordDrive3D,一个可供性与几何感知的世界-动作模型。该模型基于 VLM 骨干网络,联合学习未来动作相关区域与空间结构,预测可行驶区域与碰撞关键区域,并从 RGB 世界模型潜变量预测未来几何结构。在 NAVSIM 基准上,AffordDrive3D 取得 91.3 PDMS 与 89.9 EPDMS 的 state-of-the-art 性能。目前尚无其他报道或更早版本可对照。
Generated from reports · updated 3 hr ago
Timeline
Follow the coverage from different angles.
- arXiv · Computer VisionAffordDrive3D:具备空间理解的可供性感知世界-动作模型
AffordDrive3D 是一个可供性与几何感知的世界-动作模型,联合学习未来动作相关区域与空间结构。该模型基于 VLM 骨干网络,预测可行驶区域与碰撞关键区域,并从 RGB 世界模型潜变量预测未来几何结构。在 NAVSIM 上,AffordDrive3D 取得 91.3 PDMS 与 89.9 EPDMS 的 state-of-the-art 性能。
Heat trend
Current heat 9·Comparable peak 10(Oct 9)·Comparable change over 24 hours –
The trend compares only the same participants observed continuously; its range may be smaller than the current heat count. Move or click on the chart to inspect hourly heat; use the left and right arrow keys to switch.