Skip to content
arXiv · Computer Vision· Zhaoyang Liu, Kun Jiang, Ziying Song, Diange Yang·· 4 hr agoAI score26

VGGTWorld-VLA:面向自动驾驶的意图条件化 3D 世界演化

VGGTWorld-VLA: Intent-Conditioned 3D World Evolution for Autonomous Driving

AI brief

研究者提出 VGGTWorld-VLA,将 VGGT-World 扩展为意图条件化的可控 3D 世界演化模型,用于自动驾驶场景。该方法引入 action–semantic conditioning 机制,把驾驶语义与自车运动表征注入 future-token 流,并通过 geometry–language–action bridge 联合条件化未来几何预测。

Source: arXiv · Computer Vision · arxiv.org