心理健康污名自动评测基准发布
Get the story
2026年10月5日,研究者在arXiv(Computation and Language)发布一个基于理论的基准,用于自动评测在线交流中的心理健康污名。该基准涵盖六种心理健康状况的自然新闻与社交媒体文本;标注框架包含二元污名检测任务,以及覆盖污名模式、领域和具体成分的多级分类体系。实验显示,情感、毒性与仇恨言论分类器难以捕捉心理健康污名,LLM在缺少明确操作规则时常过度预测污名。目前公开部分基准、标注、示例与代码已发布。
Generated from reports · updated 16 hr ago
Timeline
Follow the coverage from different angles.
- arXiv · Computation and Language在线交流中心理健康污名的自动评测
研究者提出一个基于理论的基准,用于自动评测在线交流中的心理健康污名,涵盖六种心理健康状况的自然新闻与社交媒体文本。标注框架包含二元污名检测任务,以及覆盖污名模式、领域和具体成分的多级分类体系。实验显示,情感、毒性与仇恨言论分类器难以捕捉心理健康污名,LLM 在缺少明确操作规则时常过度预测污名;公开部分基准、标注、示例与代码已发布。
Heat trend
Current heat 6·Comparable peak 10(Oct 5)·Comparable change over 24 hours –
The trend compares only the same participants observed continuously; its range may be smaller than the current heat count. Move or click on the chart to inspect hourly heat; use the left and right arrow keys to switch.