Hot eventLive
eRLT:基于动作相关Token路由的高效VLA强化学习
1 reports1 sources3 hr ago updated
Get the story
AI overview
2026-10-02,arXiv·Robotics 报道 eRLT 方法,旨在提升 VLA 强化学习的样本效率。其核心是在冻结的 VLA tokens 与层之间路由任务相关动作信息以构建有效状态表示;路由模块由专家演示初始化,再通过在线交互的 critic 反馈进行优化。
Generated from reports · updated 2 hr ago
Timeline
Follow the coverage from different angles.
Oct 2, 2026
- arXiv · RoboticseRLT:通过动作相关 Token 路由实现高效 VLA 强化学习
eRLT 通过在冻结 VLA 的 tokens 和层之间路由任务相关动作信息,构建有效状态表示,提升机器人操作强化学习样本效率。路由模块由专家演示初始化,再通过在线交互的 critic 反馈优化。
Heat trend
There is not enough continuous observation data to draw a trend yet.