公开观点

Gabriel Grinberg

Anthropic 产品发布中引用的客户 · Base44

1 份资料 · 1 条观点 · 1 个话题

内容更新于:

Gabriel Grinberg 关于应用构建迭代的观点。 按话题阅读 1 条观点,核对 1 个来源中的证据。

探索知识关联

按话题查看观点

按来源发布日期整理的个人观点,仅反映这些材料中的表达,不代表其全部立场。

译文仅辅助阅读;核查观点请以原始摘录为准。

应用构建迭代

查看话题
应用构建迭代

Base44 报告称,在 118 个应用构建任务中所需迭代次数更少

在 Anthropic 发布的一则客户证言中,Base44 公司的加布里埃尔·格林伯格(Gabriel Grinberg)表示,Claude Sonnet 5.5 在 118 个应用构建任务中与 Opus 5 的评分持平,平均每个构建仅需 3.6 轮迭代,而 Opus 5 则需 7.7 轮。

支持这项说法

Anthropic 关于 Claude Sonnet 5.5 的客户证言

“在 118 次真实应用构建中,Claude Sonnet 5.5 生成的应用评分与 Opus 5 持平;平均每次构建仅需 3.6 轮迭代,而 Opus 5 需要 7.7 轮。在我们比较的所有模型中,它的工具调用失败次数最少;也很少在构建中途暂停并询问用户,因此因等待人工响应而停滞的构建更少。”

原始摘录
“Across 118 real app builds, Claude Sonnet 5.5 produced apps that scored level with Opus 5. It got there in 3.6 iterations per build on average, where Opus 5 took 7.7. It had the fewest failed tool calls of any model we compared. It also rarely stopped mid-build to ask the user a question, so fewer builds stall waiting on someone to answer.”

按来源日期阅读1

按原始来源的发布日期排序;措辞不同不代表立场发生变化。