公开观点

Simon Willison

Source author ·

1 场访谈 · 3 条观点 · 3 个话题

内容更新于:

探索知识关联 ↗

译文仅辅助阅读;核查观点请以原始摘录为准。

按话题查看观点

按访谈日期排序的归属观点。这是相关对话的快照,而非对某人信念的最终定论。

AI model performance and cost

查看话题

Reported Sonnet 5.5 speed and cost improvements

Willison quotes Anthropic’s claim that Sonnet 5.5 "runs 30%+ faster, and costs up to 30% less for most work". He writes that it is priced the same as Sonnet 5, "but appears to beat it on every benchmark, and should be cheaper to run as well".

支持这项说法

原始摘录

Claude Sonnet 5.5 . New Sonnet model from Anthropic today. They say it "runs 30%+ faster, and costs up to 30% less for most work" - it's priced the same as Sonnet 5 but appears to beat it on every benchmark, and should be cheaper to run as well.
Claude Sonnet 5.5

此片段未附更多上下文,请阅读原始访谈。

位置来自原始文本;请结合原文语境核对措辞与归属。

AI model reliability and token limits

查看话题

Token exhaustion in a pelican-SVG test

In Willison’s pelican-SVG test, Sonnet 5.5 at “max” thinking effort used 128,000 tokens, costing $1.28, before running out of tokens without producing an SVG.

支持这项说法

原始摘录

the "max" thinking effort pelican thought for 128,000 tokens (at a cost of $1.28) before running out of tokens and failing to produce an SVG.
Claude Sonnet 5.5
上下文

Here are some pelicans riding bicycles . Sonnet 5.5 suffered from the same bug as Opus 5.5 :

位置来自原始文本;请结合原文语境核对措辞与归属。

AI model capability comparison

查看话题

Sonnet 5.5 near-parity with Opus 5.5 on coding tasks

Willison says Sonnet 5.5 appears almost as good as Opus 5.5 on some coding tasks, including viral 3D animation tricks.

支持这项说法

原始摘录

Sonnet 5.5 appears to be almost as good as Opus 5.5 on some coding tasks, including various viral 3D animation tricks .
Claude Sonnet 5.5

此片段未附更多上下文,请阅读原始访谈。

位置来自原始文本;请结合原文语境核对措辞与归属。

按访谈日期查看表达记录1

按原始访谈发布日期排列;表述不同不代表立场发生变化。