Hot eventLive
预算约束多源反事实标注离策略评估研究
1 reports1 sources3 hr ago updated
Get the story
AI overview
2026年10月9日,arXiv Machine Learning Theory(一手)发表研究,提出在预算约束下为 contextual-bandit 离策略评估(OPE)获取反事实标注的方法。该方法按各标注源的成本与误差特征,在上下文-动作对之间做整数分配,以降低估计方差中依赖标注计划的部分。目前公开信息仅涉及该方法框架,未见实验数据或后续验证报道。
Generated from reports · updated 3 hr ago
Timeline
Follow the coverage from different angles.
Oct 9, 2026
- arXiv · Machine Learning Theory预算约束下的多源反事实标注用于离策略评估
该研究提出在预算约束下为 contextual-bandit 离策略评估(OPE)获取反事实标注的方法,按各标注源的成本与误差特征在上下文-动作对间做整数分配,以降低估计方差中依赖标注计划的部分。
Heat trend
Current heat 9·Comparable peak 10(Oct 9)·Comparable change over 24 hours –
The trend compares only the same participants observed continuously; its range may be smaller than the current heat count. Move or click on the chart to inspect hourly heat; use the left and right arrow keys to switch.