The Decoder· Matthias Bastian·· 2 小时前精选AI 评分66
Vals AI 研究:AI 智能体团队成本高数倍,质量提升却微乎其微
AI agent teams waste massive tokens for barely measurable quality gains, research finds
AI 导读
评测公司 Vals AI 在 Vibe Code Bench 上测试 GPT-6 Sol 和 Claude Opus 5.5,分别以单智能体和团队形式运行,并设置中等与最高两档推理强度。
推荐理由
评测数据显示多智能体团队成本高出数倍却几乎不提升成绩,可帮助判断何时不值得堆叠智能体。
来源:The Decoder · the-decoder.com