Home · Technology · Sep 26 archive

Anthropic Launches Opus 5.5 Amid AI Safety Debate

Confirmed

Technology Desk

In Short: Anthropic's disclosure comes as global leaders debate the pace of AI development, with some calling for independent safety evaluators to be embedded within AI labs.

Anthropic logo
Photo: Anthropic / Wikimedia Commons (Public domain)

Opus 5.5 is positioned as the most capable and expensive tier in the Claude lineup, with Sonnet in the middle and Haiku as the fastest and cheapest option.

According to Anthropic, Opus 5.5 was tested by external evaluators, including Frontier Design and METR, and performed the strongest on their automated behavioral audit.

YouTube — WeeklyHow YouTube

The new model is less likely to use jargon and provides important information upfront, making it more user-friendly.

Opus 5.5 also sets a new state-of-the-art in coding and knowledge work performance, outpacing the larger Fable model on many benchmarks.

The cost of Opus 5.5 is significantly reduced, with output tokens priced at $20 per million tokens, $5 less than the previous model.

Amid growing concerns about AI security, Anthropic announced that Claude now leads 26% of the company's R&D work, up from nearly none earlier this year.

More than 90% of Anthropic's AI R&D now involves Claude at least at a collaborative level, reflecting the rapid advancement of AI technology.

Anthropic's disclosure comes as global leaders debate the pace of AI development, with some calling for independent safety evaluators to be embedded within AI labs.

Anthropic CEO Dario Amodei said frontier AI labs should commit to embedding independent safety evaluators within their organizations to ensure safety practices are verified.

The company is unilaterally adopting this step immediately, granting ongoing, employee-level access to third-party evaluators to verify safety practices.

Anthropic's move to release Opus 5.5 amid calls for slowing AI development highlights the ongoing tension between innovation and safety in the AI industry.

What this adds

The release of Opus 5.5 by Anthropic and GPT-6 Sol by OpenAI within 90 minutes of each other underscores the rapid pace of AI development.

The cost reduction in Opus 5.5 pricing suggests a shift towards making advanced AI capabilities more accessible.

The involvement of Claude in a significant portion of Anthropic's R&D work indicates the increasing reliance on AI in the development process.

The push for independent safety evaluators reflects growing concerns about the potential risks associated with AI technology.

What's confirmed

What's still developing

Sources