VGGTWorld-VLA:意图条件化的自动驾驶3D世界演化
Get the story
研究者提出 VGGTWorld-VLA,将 VGGT-World 扩展为意图条件化的可控 3D 世界演化模型,用于自动驾驶场景。该方法引入 action–semantic conditioning 机制,把驾驶语义与自车运动表征注入 future-token 流,并通过 geometry–language–action bridge 联合条件化未来几何预测。该成果发布于 arXiv(Computer Vision 分类),目前公开信息仅涉及模型框架与条件化机制,未见实验数据或第三方验证报道。
Generated from reports · updated 3 hr ago
Timeline
Follow the coverage from different angles.
- arXiv · Computer VisionVGGTWorld-VLA:面向自动驾驶的意图条件化 3D 世界演化
研究者提出 VGGTWorld-VLA,将 VGGT-World 扩展为意图条件化的可控 3D 世界演化模型,用于自动驾驶场景。该方法引入 action–semantic conditioning 机制,把驾驶语义与自车运动表征注入 future-token 流,并通过 geometry–language–action bridge 联合条件化未来几何预测。
Heat trend
Current heat 9·Comparable peak 10(Oct 9)·Comparable change over 24 hours –
The trend compares only the same participants observed continuously; its range may be smaller than the current heat count. Move or click on the chart to inspect hourly heat; use the left and right arrow keys to switch.