Hot eventLive
DyaFDB:全双工对话模型的双体评估框架
1 reports1 sources3 hr ago updated
Get the story
AI overview
2026年10月7日,arXiv Audio and Speech 频道发布论文,提出 DyaFDB 评估框架,将对话视为双体问题。该框架让两个全双工对话模型在指定角色下直接对话,并由外部裁判离线为双方打分。框架包含 4 个任务、140 个场景,共记录 7,560 段对话,覆盖 6 种自我对弈与交叉对弈组合。实验显示,模型行为会持续重塑对话伙伴。作者表示将开源场景、角色提示词与双模型录制协议,但不包括预录音频。此前无其他报道,本报道为事件首发。
Generated from reports · updated 3 hr ago
Timeline
Follow the coverage from different angles.
Oct 7, 2026
- arXiv · Audio and Speech对话是一个双体问题:全双工对话模型的二元评估
论文提出 DyaFDB 评估框架,让两个全双工对话模型在指定角色下直接对话,并以外部裁判离线为双方打分。框架包含 4 个任务、140 个场景,记录 7,560 段对话,覆盖 6 种自我对弈与交叉对弈组合。实验显示模型行为会持续重塑对话伙伴,作者将开源场景、角色提示词与双模型录制协议,不含预录音频。
Heat trend
Current heat 9·Comparable peak 10(Oct 7)·Comparable change over 24 hours –
The trend compares only the same participants observed continuously; its range may be smaller than the current heat count. Move or click on the chart to inspect hourly heat; use the left and right arrow keys to switch.