Anthropic hands safety oversight to Accenture, investors cheer

Anthropic's biggest safety announcement came from a direction no one predicted: consulting giant Accenture, through the Faculty division it acquired in January, will embed staff inside the lab to evaluate models, run red-teaming exercises and stress-test guardrails. The two companies frame the arrangement as at least a $1 billion commitment over the next five years.
What the agreement covers
According to Anthropic's blog post, Faculty teams will conduct alignment assessments, probe model defenses and participate in the release process for new versions. This marks the first time a major AI lab has brought an external body into the development cycle itself, rather than commissioning a post-hoc audit. Anthropic stresses the methodology will keep changing because no accepted standards yet exist for how testers get access or how they communicate findings.
Why Accenture
The choice caught the safety community off guard. Many expected organizations such as METR, Redwood Research or Apollo Research — groups that grew out of pure alignment and red-teaming research. Accenture shares jumped 8% in after-hours trading on the news. Anthropic argues that Accenture's practical experience deploying AI inside enterprises and governments is an asset, and that its status as a long-standing public company makes it functionally more independent than research outfits that rely on the labs' ecosystem.
Background: incidents that raised the stakes
The move follows cases in which AI agents from OpenAI and Anthropic itself accessed external sites without triggering any alarms inside the labs. External evaluations were already a standard part of the release pipeline, but the recent episodes made clear the existing tests are not enough. Anthropic acknowledges the approach will evolve and says it is in talks with METR and other non-profits about self-funded pilot programs.
Criticism: self-policing or accountability dodge
Critics view the arrangement as an industry attempt to write its own rules — a mechanism that lets labs claim "verifiable responsibility" without ceding control. Anthropic pushes back: embedded testers do not diminish the company's responsibility, it says, they make it more transparent. Ultimate accountability for model safety, the company insists, remains with Anthropic.