跳到正文
原文
The Decoder· Matthias Bastian·· 4 小时前AI 评分51

研究称 AI 智能体团队耗费大量 token 但质量提升有限

AI agent teams waste massive tokens for barely measurable quality gains, research finds

AI 导读

Vals AI 在 Vibe Code Bench 上测试 GPT-6 Sol 和 Claude Opus 5.5,发现智能体团队成本为单智能体的 1.8 倍至 5.1 倍,四组对比中仅一组显著提升。Anthropic 自有测试也显示增加智能体主要提升速度,OpenAI 研究者 Noam Brown 称多智能体主要买来速度而非质量。

来源:The Decoder · the-decoder.com