面向分布内数据的保形数据污染检验研究
Get the story
2026年10月9日,arXiv 统计学机器学习栏目发布论文,提出一种面向分布内数据获取的无分布假设、污染感知数据获取框架。该框架仅检查少量数据,即可识别对模型个性化最有价值的外部数据代理。研究引入基于保形异常检测的两样本检验,声称在任意污染水平下保持有效;同时提出新型 Storey 型检验,通过 Benjamini-Hochberg 过程实现有限样本错误发现率控制。论文称实验在多种协作学习场景中验证了方法的稳健性与有效性。目前报道仅涉及该论文内容,尚无后续验证或独立评论。
Generated from reports · updated 3 hr ago
Timeline
Follow the coverage from different angles.
- arXiv · Statistics Machine Learning面向分布内数据获取的保形数据污染检验
论文提出一种无分布假设、污染感知的数据获取框架,仅检查少量数据即可识别对模型个性化最有价值的外部数据代理。研究引入基于保形异常检测的两样本检验,在任意污染水平下保持有效,新型 Storey 型检验通过 Benjamini-Hochberg 过程实现有限样本错误发现率控制。实验在多种协作学习场景中验证了方法的稳健性与有效性。
Heat trend
Current heat 9·Comparable peak 10(Oct 9)·Comparable change over 24 hours –
The trend compares only the same participants observed continuously; its range may be smaller than the current heat count. Move or click on the chart to inspect hourly heat; use the left and right arrow keys to switch.