跳到正文
原文
The Decoder· Matthias Bastian·· 4 小时前AI 评分62

研究显示 AI 智能体团队成本高但质量提升有限

AI agent teams waste massive tokens for barely measurable quality gains, research finds

AI 导读

评测公司 Vals AI 在 Vibe Code Bench 上测试 GPT-6 Sol 和 Claude Opus 5.5,分别以单体智能体和智能体团队形式运行,并设置中等与最高两档推理强度。

来源:The Decoder · the-decoder.com