Hot eventLive
Seg-OPD:分段在线策略蒸馏提升推理修订
1 reports1 sources16 hr ago updated
Get the story
AI overview
研究者提出分段式在线蒸馏方法 Seg-OPD,用于学习推理修订。该方法基于不确定性指标选取学生模型的推理片段,获取教师模型对应的重写版本,训练学生偏好教师重写片段,同时保留 token 级密集监督。在数学推理与竞赛编程任务上,Seg-OPD 的修订成功率高于基线,推理准确率平均相对提升 5.22%,代码已开源。
Generated from reports · updated 16 hr ago
Timeline
Follow the coverage from different angles.
Oct 5, 2026
- arXiv · Artificial IntelligenceSeg-OPD:用分段式在线蒸馏学习推理修订
研究者提出分段式在线蒸馏方法 Seg-OPD,用于学习推理修订。该方法基于不确定性指标选取学生模型的推理片段,获取教师模型对应的重写版本,训练学生偏好教师重写片段,同时保留 token 级密集监督。在数学推理与竞赛编程任务上,Seg-OPD 的修订成功率高于基线,推理准确率平均相对提升 5.22%,代码已开源。
Heat trend
Current heat 7·Comparable peak 10(Oct 5)·Comparable change over 24 hours –
The trend compares only the same participants observed continuously; its range may be smaller than the current heat count. Move or click on the chart to inspect hourly heat; use the left and right arrow keys to switch.