Claude Opus 5.5 has been introduced as the initial model within the new Claude 5.5 family. This release marks the first model since the company's call for pacing the frontier of AI development. External evaluators, including Frontier Design and METR, tested the model prior to its public release.
Opus 5.5 shows a performance increase over Opus 5, becoming the new leading model. Early testers reported significant performance gains on complex tasks; one tester completed a 680,000-line code migration in under a day. The model also demonstrated effectiveness in identifying and correcting software inefficiencies, succeeding in 39 out of 40 attempts to reduce web app load times without altering behavior, compared to Opus 5's smaller improvements that sometimes affected app behavior. Another test showed Opus 5.5 scoring higher than other Claude models in graphics and polish when building a game from a single prompt.
Opus 5.5 achieved the highest scores to date on the automated behavioral audit, an alignment test suite that simulates thousands of scenarios. The model is less prone to irreversible actions and acting outside defined boundaries, and it exhibits increased resistance to prompt injection compared to Opus 5. Alignment testing has been expanded to include longer and impossible tasks, as well as scenarios based on real incidents. The full evaluation details are available in the Opus 5.5 System Card.
Opus 5.5 requires less computational power than Opus 5, which is reflected in its pricing. The model's performance is comparable to Claude Mythos 5.1 in biology and cybersecurity, leading to similar safeguards. Vetted organizations can apply to the Life Sciences Verification Program to use Opus 5.5 for biology research. Access to the Cyber Verification Program for verified cybersecurity practitioners will also be expanded in the coming weeks.
✨ This summary was generated by AI from the outlets' reporting listed below. It is not independently verified and may contain errors — check the original sources. How BrevFeed works →
One email each morning: the day's tech stories, clustered across outlets and summarized. No account needed.
One email a day. Unsubscribe in one click, any time.
Spend a few minutes, get the whole day. Every topic's top stories in one hands-free rundown — listen, watch, or read the transcript.
▶ Play today's briefNew every morning, and the back catalogue is archived by date.
Anthropic launched Claude Sonnet 5.5, the second model in its Claude 5.5 family, which runs over 30% faster and costs up to 30% less for most tasks compared to its predecessor. This model is designed for everyday tasks, bug fixing, and document creation, complementing the more complex work handled by Claude Opus 5.5.
A guide details prompting patterns for Claude Opus 5.5, which generates output tokens over 30% faster and uses fewer tokens than Claude Opus 5. The new model also demonstrates stronger agentic coding, code review, and sustained autonomous work capabilities.
Anthropic's Claude Opus 5.5 exhibits fewer AI writing patterns, shorter sentences, and simpler wording compared to Opus 5, according to analysis by Arena. This change indicates an evolution in how the AI model generates text, making its output less identifiable as AI-generated and potentially improving readability, though average answer length has increased.
Anthropic's Claude Opus 5.5 model is now accessible on Microsoft Foundry, allowing developers and enterprises to build AI applications and agents. This integration provides enhanced capabilities for agentic coding, knowledge work, and long-running tasks, with improvements in efficiency and cost.
Claude Opus 5.5, the first model in the Claude 5.5 family, has been released, performing at the level of Claude Fable 5.1 on most tasks while costing 40% less than Opus 5. This new model demonstrates significant performance improvements in complex tasks like code migration and software optimization, alongside enhanced safety features.