跳到正文
The Decoder· Matthias Bastian·· 3 小时前AI 评分62

研究发现 AI 智能体团队成本高出数倍但质量提升有限

AI agent teams waste massive tokens for barely measurable quality gains, research finds

AI 导读

评测公司 Vals AI 在 Vibe Code Bench 上测试 GPT-6 Sol 和 Claude Opus 5.5 的单智能体与团队模式,团队成本高出 1.8x 至 5.1x,但四组对比中仅 GPT-6 Sol 在 medium 推理档下团队高出 7.3 分,maximum 推理档下团队无明显优势。

来源:The Decoder · the-decoder.com