Elyse Betters Picaro/ZDNETZDNET’s key takeaways
- Claude Opus 5.5 may very well be a giant win for energy customers.
- Builders might even see sooner coding with fewer steps.
- Anthropic says the improve is safer, cheaper, and fewer wordy.
Lower than two months after the discharge of Claude Opus 5, Anthropic is again with Opus 5.5. The massive pitch for Opus 5 was that it had “close to Fable” efficiency at half the value. This time, the headline is that the workhorse AI Opus 5.5 delivers Fable 5.1 efficiency for many work and prices about 40% much less to run.
“Our clients use Field AI on huge quantities of content material, so velocity and value are a prime precedence,” Yashodha Bhavnani, VP of AI Merchandise at Field, reviews. “In our evaluations, Claude Opus 5.5 used a 3rd of the tokens Opus 5 did, and its solutions had been 40% much less verbose with out shedding accuracy. We count on that to matter so much for groups working brokers throughout their content material in areas like monetary companies and the general public sector.”
Anthropic says that, “Over the approaching weeks, we’ll even be launching Claude Sonnet 5.5 and Haiku 5.5.” Opus 5.5 is accessible right this moment.
I’ve been utilizing the heck out of Claude Code with Opus 5, so I’m notably hopeful that the corporate’s efficiency claims are correct. The corporate says it “generates output greater than 30% sooner than Opus 5.”
Additionally: Anthropic merges Claude chat and Cowork into one
Token costs and subscription plans are addressed in today’s announcement. Tokens are priced at 20% lower than when used with Opus 5. Opus 5.5 additionally reportedly “wants fewer tokens for larger high quality work.”
For subscription customers like me, Anthropic is elevating its five-hour utilization limits by 20%. That’s mainly a 20% larger gasoline tank for the way a lot AI chomping you need to use throughout 5 hours. Whereas my Max plan doesn’t get reset usually, it does get reset. For these on $20/month plans, this may very well be a substantial win. On prime of that, the corporate says that each 5-hour and weekly utilization limits go additional as a result of Opus 5.5 prices lower than Opus 5.
These appear to be additive. There’s a 20% bigger bucket coupled with a 25% slower burn, that means that efficient utilization appears to be a few 50% higher run capability with Opus 5.5. That’s not an inconsiderable quality-of-life enchancment.
Anthropic additionally says that Opus 5.5 communicates extra naturally than prior fashions. If it’s even just a bit much less obsequious, I’d be pleased. Generally, Opus is usually a whole suck-up, notably when it’s completed one thing unsuitable.
Additionally: The AI models that cheat the most, according to new CAIS benchmark
“Verbose, hard-to-follow output has been my greatest frustration with frontier fashions, and Claude Opus 5.5 fixes it,” says John Ruelas, workers software program engineer at Ramp.
Opus 5.5, he says, “writes like a superb colleague and follows our writing guidelines. A design spec got here out usable with very minimal edits, and when it rewrote certainly one of our prompts, I most popular its model to my very own. When it optimized our check suite, I might comply with its reasoning simply and shipped the change with confidence.”
Pacing the frontier
Talking of doing one thing unsuitable, the second half of Anthropic’s announcement is all about Opus 5.5 being a better-behaved AI citizen.
Citing CEO Dario Amodei’s blog post about moderating the velocity of AI functionality advances, Anthropic is hitting huge on a sequence of Opus 5.5 finest practices, together with “in depth alignment testing, pre-release analysis by exterior organizations, and safeguards for high-risk areas like cybersecurity and biology.”
Additionally: Why the DOJ’s OpenAI copyright stance is the real threat
Alignment is the AI time period that helps measure how a lot an AI appears inclined to run rogue. Anthropic says Opus 5.5 is “the strongest performing mannequin we’ve examined thus far, with explicit enhancements on a number of of the behaviors that contributed to current cybersecurity incidents (e.g., biased reasoning, trying to flee a sandbox, and others).”
The corporate says they used exterior testing suppliers, together with a “related class of safeguards to Fable 5.1 on cybersecurity, biology, and frontier LLM improvement.”
If safeguards hearth, requests to the AI fall again from the Opus 5.5 stage to Opus 4.8. In observe, most cybersecurity duties might be rerouted to Opus 4.8, and people requests associated to biology and LLM improvement might be despatched to Opus 5.
Additionally: AI just broke your career ladder – 6 new ways to the top
Some “vetted organizations” can now apply to Anthropic’s Life Sciences Verification Program to achieve extra highly effective entry for organic analysis. These accredited for cybersecurity work via Anthropic’s Cyber Verification Program will be capable of begin utilizing Opus 5.5 in a number of weeks.
Disclosure: I’ve been personally accredited into the Cyber Verification Program as a part of work I do exterior of ZDNET on nationwide infrastructure safety.
Buyer utilization experiences
Mario Rodriguez, GitHub’s chief product officer, has his tackle the brand new launch. He says, “Builders need brokers that may tackle actual software program work and end it. In our testing throughout GitHub Copilot CLI and VS Code, Claude Opus 5.5 used among the many fewest tokens and steps we measured. In VS Code, it solved extra terminal duties than Opus 5 in lower than half the steps. Greater than making particular person duties extra environment friendly, it’s making builders’ larger tasks extra achievable.”
Carl Bennett is CIO at Massive 4 accounting agency Deloitte Consulting LLP. He reviews, “Even at its lowest effort setting, Claude Opus 5.5 caught 72% of recognized bugs in our code opinions to Opus 5’s 56% at excessive effort, with fewer false alarms and a fraction of the output. On US consulting evaluation, low-thinking effort matched its higher-thinking settings on half the output and handed our high quality checks. When extra decrease considering efforts are deployed in manufacturing, that’s client-ready work delivered effectively.”
Additionally: This CIO doesn’t ‘hire engineers to write code’
So there you go. Extra energy, extra security, much less price, and fewer rambling. That’s loads of enchancment simply shy of two months after the final main launch.
What do you assume? Are you planning on stepping up from Opus 5 to Opus 5.5 as quickly because it’s out there? Tell us within the feedback beneath.
You possibly can comply with my day-to-day undertaking updates on social media. Remember to subscribe to my weekly update newsletter, and comply with me on Twitter/X at @DavidGewirtz, on Fb at Facebook.com/DavidGewirtz, on Instagram at Instagram.com/DavidGewirtz, on Bluesky at @DavidGewirtz.com, and on YouTube at YouTube.com/DavidGewirtzTV.





