跳到正文
原文
The Decoder· Matthias Bastian·· 4 小时前AI 评分71

研究显示 AI 智能体团队消耗大量 token 却几乎不提升质量

AI agent teams waste massive tokens for barely measurable quality gains, research finds

AI 导读

评测公司 Vals AI 在 Vibe Code Bench 上测试 GPT-6 Sol 和 Claude Opus 5.5,分别以单智能体和团队形式运行,并设置中等与最高两档推理强度。

来源:The Decoder · the-decoder.com