Skip to content
Hot eventLive

DyaFDB:全双工对话模型的双体评估框架

1 reports1 sources3 hr ago updated

Get the story

AI overview

2026年10月7日,arXiv Audio and Speech 频道发布论文,提出 DyaFDB 评估框架,将对话视为双体问题。该框架让两个全双工对话模型在指定角色下直接对话,并由外部裁判离线为双方打分。框架包含 4 个任务、140 个场景,共记录 7,560 段对话,覆盖 6 种自我对弈与交叉对弈组合。实验显示,模型行为会持续重塑对话伙伴。作者表示将开源场景、角色提示词与双模型录制协议,但不包括预录音频。此前无其他报道,本报道为事件首发。

Generated from reports · updated 3 hr ago

Timeline

Follow the coverage from different angles.

Oct 7, 2026
  1. arXiv · Audio and Speech
    对话是一个双体问题:全双工对话模型的二元评估

    论文提出 DyaFDB 评估框架,让两个全双工对话模型在指定角色下直接对话,并以外部裁判离线为双方打分。框架包含 4 个任务、140 个场景,记录 7,560 段对话,覆盖 6 种自我对弈与交叉对弈组合。实验显示模型行为会持续重塑对话伙伴,作者将开源场景、角色提示词与双模型录制协议,不含预录音频。

Heat trend

Current heat 9·Comparable peak 10(Oct 7)·Comparable change over 24 hours –

02.557.510Oct7Oct7Oct7Oct7

The trend compares only the same participants observed continuously; its range may be smaller than the current heat count. Move or click on the chart to inspect hourly heat; use the left and right arrow keys to switch.