hn.today

Claude Opus 5.5 Should Raise Your Ambitions

thezvi.substack.com9 points5 comments
Screenshot of Claude Opus 5.5 Should Raise Your Ambitions

Claude Opus 5.5 is a notable model release that delivers Fable‑5.1-level performance at substantially lower cost and improved usability: Anthropic claims roughly 40% lower running cost than Opus 5, pricing tiers at $4/$20 (standard) and $8/$40 (fast) with a cheap cache ($0.20), and speed gains around 30%. It shines on agentic coding, communication, and multi‑modal vision/3D tasks, often matching or exceeding Fable and approaching GPT‑6 Astra on many benchmarks. External testing places Opus 5.5 highly (Artificial Analysis intelligence 58), highlights a big ProgramBench jump (to 18.5%), and credits it with leading Omniscience scores by hallucinating less. Early reactions are uniformly positive: it’s pleasant to converse with, writes cleanly, handles long sessions, and produces compelling video and audio outputs.

Technically meaningful caveats matter: classifiers still block some prompts (but more sensibly and with better recovery), occasional nonsensical sentences appear, and a few domain‑specific grading benchmarks still favor Fable 5.1 where pure correctness trumps presentation. Zero data retention is offered despite claimed strong cyber capabilities, which raises principled concerns. Overall, Opus 5.5 is presented as the go‑to model for most tasks - cheap, fast, and capable - while continued caution about frontier model development and deployment remains important.

Read on thezvi.substack.com5 comments on Hacker News

Summary generated by AI from the linked article. hn.today is not affiliated with Hacker News or Y Combinator.

More in AI

The daily digest

Today's best Hacker News stories, summarized and screenshotted, one email a day.