亚马逊标记AI项目,一项Claude Sonnet计划超支860%

亚马逊标记AI项目,一项Claude Sonnet计划超支860%

内部报告显示,多个AI项目产生了大幅超出预期的成本,随着外界愈发质疑大规模使用AI是否能够带来价值,亚马逊正通过设置护栏和支出上限来收紧管控。

事实核查
The claim is corroborated by a Financial Times report whose headline directly confirms Amazon found AI causing runaway spending, and by multiple independent secondary outlets (Tom's Hardware, Unite.AI, AI Weekly, Daily.dev) that all cite the FT and consistently report the same key facts: a Claude Sonnet author-mapping project cost $1.8M — 860% over budget — and went undetected for five months, alongside smaller overruns of $541K and $134K. All sources agree Amazon is tightening controls with automated guardrails and spending caps and scrapped an internal usage leaderboard. The headline claim of '860% over budget' and 'guardrails and spending caps' is precisely matched. Only slight uncertainty remains because the underlying FT primary text is paywalled and full verification relies on consistent secondary reporting.
摘要

亚马逊内部将多个AI项目标记为严重超支,其中一项使用 Anthropic 的 Claude Sonnet 处理商品列表的项目,成本约为 $1.8 million,且超出预算 860%,管理层直到近五个月后才发现问题。内部报告还显示,一款财务审计工具产生了约 $541,000 的意外支出,一项物流优化项目则增加了约 $134,000 的成本,且两周多时间里都未被察觉。员工称,一些AI故障之所以“代价极其高昂”,是因为在错误未被严格限制时,自主代理会持续消耗 tokens。亚马逊还关闭了一个内部AI使用排行榜,因为员工会通过提高 token 消耗来“刷榜”;在一起导致 AWS 中断 13 小时的事件后,公司也收紧了其 Kiro 编码助手的权限。亚马逊表示,公司仍在试验并改进AI的使用方式,个别案例并不代表全公司的日常使用情况。这一事件凸显出更广泛的行业争论:尽管对亚马逊这样体量的公司而言,相关金额仍可控,但不断上升的AI使用量和 token 消耗是否真正转化为生产率提升,正受到越来越多质疑。

术语与概念
  • Claude Sonnet: Anthropic 推出的一款AI模型,被企业用于生成式AI任务。
  • tokens: 文本处理单位,用于衡量AI模型消耗的使用量及相应成本。
  • guardrails: 用于限制或监控AI系统运行方式的控制措施,以降低风险并防止支出失控。