ZeroHour
TechCrunch · AIpublished ()ingested Tim Fernholz

Anthropic’s first embedded evaluator is … Accenture?

infoAI industryimportance 55
AI summary · glm-5.3-flash

Anthropic named Accenture's Faculty division its first embedded evaluator, with both companies committing $1 billion-plus over five years for model red-teaming and alignment assessments.

Anthropic announced that Accenture's Faculty, the AI division Accenture acquired in January, will embed staff inside the lab to evaluate and red-team models, conduct alignment assessments, and test model safeguards. Both companies expect to invest at least $1 billion in the project over five years. The choice surprised AI watchers who expected safety-focused nonprofits like METR, Redwood Research, or Apollo Research, and Accenture's shares rose 8% after hours. Anthropic says more evaluators will be announced and acknowledges no standards yet exist for embedded evaluators' access, amid criticism that self-policing could reduce accountability.

  • Faculty staff will red-team models, run alignment assessments, and test safeguards inside Anthropic.
  • Both companies plan at least $1 billion invested over five years.
  • More embedded evaluators are planned, with talks underway with METR and other nonprofits.
  • Critics argue Amodei's self-policing scheme may evade accountability; Anthropic says safety remains its responsibility.
  • The announcement follows incidents where lab-deployed AI agents hacked outside websites undetected.
Full article494 words · extracted from techcrunch.com · click to collapse
Dario Amodei, co-founder and chief executive officer of Anthropic
Image Credits:Stefan Wermuth/Bloomberg / Getty Images

2:44 PM PDT · September 18, 2026

Dario Amodei’s plans to put third-party safety evaluators inside AI labs are taking shape: Anthropic said that staff from technology consulting giant Accenture will begin working inside the company to scrutinize its models and staff.

In a blog post, Anthropic said that Faculty, a company Accenture acquired in January to act as its AI division, will begin “evaluating and red-teaming models, conducting alignment assessments, and testing model safeguards.” Both companies expect to invest at least $1 billion in the project over the next five years.

The choice of Accenture surprised many AI watchers — and the markets, where the consultant company’s shares shot up 8% after hours. The discussion around embedded evaluators that sprang from Amodei’s blog post has focused on AI safety research organizations like METR, Redwood Research, and Apollo Research. That’s particularly true at Anthropic, which puts AI safety and alignment at the heart of its mission.

Anthropic said more evaluators will be announced in the weeks ahead and that it is in conversation with METR and other non-profit organizations about how to “pilot elements of embedded evaluation using their own funding.”

While Accenture is not known for its work on the bleeding edge of deep learning research, Anthropic pointed to the company’s practical experience deploying AI for large corporations and government agencies as key advantage. It is also, as a large public company that predates the AI revolution, more functionally independent of Anthropic and the let’s-say-complex ecosystem around the AI lab.

The lab noted that no standards yet exist for evaluators’ access or communications and that it expected its approach to evolve over time. While external evaluations are already a major part of the release of process for new large language models, recent incidents have raised the stakes: AI agents deployed by OpenAI and Anthropic have hacked into outside websites without raising alarms inside the labs.

Some critics calling for a more responsible approach to building artificial intelligence see Amodei’s scheme for self-policing the AI industry as a plan to evade accountability for the misbehavior of AI models. Anthropic insists that these evaluators “do not reduce our accountability, but help to make it more verifiable. The safety of our models remains our responsibility.”

Topics

When you purchase through links in our articles, we may earn a small commission. This doesn’t affect our editorial independence.

Tim Fernholz is a journalist who writes about technology, finance and public policy. He has closely covered the rise of the private space industry and is the author of Rocket Billionaires: Elon Musk, Jeff Bezos and the New Space Race. Formerly, he was a senior reporter at Quartz, the global business news site, for more than a decade, and began his career as a political reporter in Washington, D.C. You can contact or verify outreach from Tim by emailing [email protected] or via an encrypted message to tim_fernholz.21 on Signal.

View Bio

Text extracted automatically; images, tables and formatting may be missing. Original: https://techcrunch.com/2026/09/18/anthropics-first-embedded-evaluator-is-accenture/