Skip to content
Hot eventLive

SPLIT-RL:分阶段感知-语言推理训练方法

1 reports1 sources3 hr ago updated

Get the story

AI overview

2026年10月9日,arXiv Computer Vision 频道发布一手报道,介绍名为 SPLIT-RL 的分阶段后训练方法。该方法将视觉推理(VR)与语言推理(LR)拆分到不同阶段分别训练,并在 VR 阶段引入声明级优势(CLA-GRPO),对原子视觉声明提供细粒度优势信号。报道未涉及更早的相关工作或后续实验数据,也未与其他方法对比。目前公开信息仅限于该方法的基本设计思路,尚无训练规模、数据集或性能指标等细节。

Generated from reports · updated 3 hr ago

Timeline

Follow the coverage from different angles.

Oct 9, 2026
  1. arXiv · Computer Vision
    SPLIT-RL:分阶段感知-语言推理训练与声明级优势

    SPLIT-RL 是一种分阶段后训练方法,将视觉推理(VR)与语言推理(LR)拆分到不同阶段训练,并用声明级优势(CLA-GRPO)在 VR 阶段对原子视觉声明提供细粒度优势。

Heat trend

Current heat 9·Comparable peak 10(Oct 9)·Comparable change over 24 hours –

02.557.510Oct9Oct9Oct9Oct9

The trend compares only the same participants observed continuously; its range may be smaller than the current heat count. Move or click on the chart to inspect hourly heat; use the left and right arrow keys to switch.