ZDNET’s key takeaways
- Claude Opus 5.5 may very well be a giant win for energy customers.
- Builders may even see quicker coding with fewer steps.
- Anthropic says the improve is safer, cheaper, and fewer wordy.
Lower than two months after the discharge of Claude Opus 5, Anthropic is again with Opus 5.5. The large pitch for Opus 5 was that it had “close to Fable” efficiency at half the worth. This time, the headline is that the workhorse AI Opus 5.5 delivers Fable 5.1 efficiency for many work and prices about 40% much less to run.
“Our clients use Field AI on huge quantities of content material, so velocity and price are a prime precedence,” Yashodha Bhavnani, VP of AI Merchandise at Field, reviews. “In our evaluations, Claude Opus 5.5 used a 3rd of the tokens Opus 5 did, and its solutions have been 40% much less verbose with out dropping accuracy. We anticipate that to matter lots for groups working brokers throughout their content material in areas like monetary companies and the general public sector.”
Anthropic says that, “Over the approaching weeks, we’ll even be launching Claude Sonnet 5.5 and Haiku 5.5.” Opus 5.5 is accessible right this moment.
I’ve been utilizing the heck out of Claude Code with Opus 5, so I’m notably hopeful that the corporate’s efficiency claims are correct. The corporate says it “generates output greater than 30% quicker than Opus 5.”
Additionally: Anthropic merges Claude chat and Cowork into one
Token costs and subscription plans are addressed in right this moment’s announcement. Tokens are priced at 20% lower than when used with Opus 5. Opus 5.5 additionally reportedly “wants fewer tokens for larger high quality work.”
For subscription customers like me, Anthropic is elevating its five-hour utilization limits by 20%. That’s mainly a 20% larger fuel tank for a way a lot AI chomping you should utilize throughout 5 hours. Whereas my Max plan doesn’t get reset typically, it does get reset. For these on $20/month plans, this may very well be a substantial win. On prime of that, the corporate says that each 5-hour and weekly utilization limits go additional as a result of Opus 5.5 prices lower than Opus 5.
These appear to be additive. There’s a 20% bigger bucket coupled with a 25% slower burn, that means that efficient utilization appears to be a few 50% higher run capability with Opus 5.5. That’s not an inconsiderable quality-of-life enchancment.
Anthropic additionally says that Opus 5.5 communicates extra naturally than prior fashions. If it’s even just a bit much less obsequious, I’d be comfortable. Generally, Opus could be a whole suck-up, notably when it’s finished one thing improper.
Additionally: The AI fashions that cheat essentially the most, in keeping with new CAIS benchmark
“Verbose, hard-to-follow output has been my largest frustration with frontier fashions, and Claude Opus 5.5 fixes it,” says John Ruelas, employees software program engineer at Ramp.
Opus 5.5, he says, “writes like colleague and follows our writing guidelines. A design spec got here out usable with very minimal edits, and when it rewrote considered one of our prompts, I most popular its model to my very own. When it optimized our check suite, I may observe its reasoning simply and shipped the change with confidence.”
Pacing the frontier
Talking of doing one thing improper, the second half of Anthropic’s announcement is all about Opus 5.5 being a better-behaved AI citizen.
Citing CEO Dario Amodei’s weblog put up about moderating the velocity of AI functionality advances, Anthropic is hitting huge on a collection of Opus 5.5 finest practices, together with “intensive alignment testing, pre-release analysis by outdoors organizations, and safeguards for high-risk areas like cybersecurity and biology.”
Additionally: Why the DOJ’s OpenAI copyright stance is the true menace
Alignment is the AI time period that helps measure how a lot an AI appears inclined to run rogue. Anthropic says Opus 5.5 is “the strongest performing mannequin we’ve examined so far, with explicit enhancements on a number of of the behaviors that contributed to current cybersecurity incidents (e.g., biased reasoning, trying to flee a sandbox, and others).”
The corporate says they used exterior testing suppliers, together with a “comparable class of safeguards to Fable 5.1 on cybersecurity, biology, and frontier LLM improvement.”
If safeguards hearth, requests to the AI fall again from the Opus 5.5 stage to Opus 4.8. In apply, most cybersecurity duties will likely be rerouted to Opus 4.8, and people requests associated to biology and LLM improvement will likely be despatched to Opus 5.
Additionally: AI simply broke your profession ladder – 6 new methods to the highest
Some “vetted organizations” can now apply to Anthropic’s Life Sciences Verification Program to realize extra highly effective entry for organic analysis. These accepted for cybersecurity work via Anthropic’s Cyber Verification Program will be capable to begin utilizing Opus 5.5 in just a few weeks.
Disclosure: I’ve been personally accepted into the Cyber Verification Program as a part of work I do outdoors of ZDNET on nationwide infrastructure safety.
Buyer utilization experiences
Mario Rodriguez, GitHub’s chief product officer, has his tackle the brand new launch. He says, “Builders need brokers that may tackle actual software program work and end it. In our testing throughout GitHub Copilot CLI and VS Code, Claude Opus 5.5 used among the many fewest tokens and steps we measured. In VS Code, it solved extra terminal duties than Opus 5 in lower than half the steps. Greater than making particular person duties extra environment friendly, it’s making builders’ larger tasks extra achievable.”
Carl Bennett is CIO at Huge 4 accounting agency Deloitte Consulting LLP. He reviews, “Even at its lowest effort setting, Claude Opus 5.5 caught 72% of identified bugs in our code evaluations to Opus 5’s 56% at excessive effort, with fewer false alarms and a fraction of the output. On US consulting evaluation, low-thinking effort matched its higher-thinking settings on half the output and handed our high quality checks. When extra decrease considering efforts are deployed in manufacturing, that’s client-ready work delivered effectively.”
Additionally: This CIO doesn’t ‘rent engineers to put in writing code’
So there you go. Extra energy, extra security, much less price, and fewer rambling. That’s plenty of enchancment simply shy of two months after the final main launch.
What do you suppose? Are you planning on stepping up from Opus 5 to Opus 5.5 as quickly because it’s out there? Tell us within the feedback beneath.
You’ll be able to observe my day-to-day challenge updates on social media. You should definitely subscribe to my weekly replace publication, and observe me on Twitter/X at @DavidGewirtz, on Fb at Fb.com/DavidGewirtz, on Instagram at Instagram.com/DavidGewirtz, on Bluesky at @DavidGewirtz.com, and on YouTube at YouTube.com/DavidGewirtzTV.
