WTF Is Going On(Weights, Tools & Frameworks)

English
SiliconANGLE AI 2개 매체

Anthropic debuts Claude Sonnet 5.5 running 30% faster than the previous-generation AI model

원문 읽기 — SiliconANGLE AI →

어디가 다뤘나

어떤 얘기가 오갔나

816포인트 · 550댓글
Probably a first world problem, but with Opus 5.5's efficiency, the limits on the 5x plan are simply sufficient for my everyday work, even when running 2-3 sessions at a time. So I wonder when I would use Sonnet 5.5. More concurrency than …
Paying $200 a month and part of their Cyber Verification Program but can't use Opus 5.5 or Sonnet 5.5 for any authorized bounty work. Immediately get flagged for `Cyber`. This is bollocks. Their safeguards are shit.
Pelicans. Sonnet 5.5 has the same problem as Opus 5.5: on "max" thinking effort it burned through 128,000 thinking tokens (taking 15 minutes to do that) and ran out before it had produced the final SVG. https://tools.simonwillison.net/mark…
It does vey well at one shotting a PacMan clone, pretty much perfect. https://jonclegg.github.io/pacman-bakeoff/entries/claude-son... 2nd only to Opus 5.5, which is perfect. https://jonclegg.github.io/pacman-bakeoff/entries/claude-opu... U…
Sonnet 5.5 scoring higher (70.6) than Opus 5.5 (66.4) in Terminal-Bench is interesting. I looked into this, because it felt strange. Turns out that Opus had 10% of its trials answered by a fallback model due to safeguards; versus only 1.5%…
Unless you’re using frontier models like Astra, Sol, Fable, or Opus, I think you’re often better off using Chinese models for a fraction of the price. I’m not sure people realises just how competitive they’ve become. GLM and DeepSeek are g…
It costs 20x more than the Chinese models I use. I just don’t need them anymore. Sure I’d use them if forced to for a job, but I don’t pay them outside of that anymore. And my job won’t even pay for Claude now because it’s so ruinously exp…
"Sonnet 5.5’s cyber capabilities are a large improvement over Sonnet 5’s, so we’re deploying it with safeguards similar to those on Opus 5.5. Users can still find and fix bugs in their code as part of routine software development, but high…
For the folks asking what is the point of Sonnet 5.5 when Opus 5.5 is better in every way, Sonnet is now the new default in the free tier and now people using Claude Web in the free tier get access to an almost close to the frontier model.
We're still on Sonnet 4.6 as we found Sonnet 5 to perform worse across all of our evals, especially against time.

상황판의 다른 소식