Anthropic’s first embedded evaluator is … Accenture?

Dario Amodei’s plans to put third-party safety evaluators inside AI labs are taking shape: Anthropic said that staff from technology consulting giant Accenture will begin working inside the company to scrutinize its models and staff.
In a blog post , Anthropic said that Faculty, a company Accenture acquired in January to act as its AI division, will begin “evaluating and red-teaming models, conducting alignment assessments, and testing model safeguards.” Both companies expect to invest at least $1 billion in the project over the next five years.
The choice of Accenture surprised many AI watchers — and the markets, where the consultant company’s shares shot up 8% after hours. The discussion around embedded evaluators that sprang from Amodei’s blog post has focused on AI safety research organizations like METR, Redwood Research, and Apollo Research. That’s particularly true at Anthropic, which puts AI safety and alignment at the heart of its mission.
Anthropic said more evaluators will be announced in the weeks ahead and that it is in conversation with METR and other nonprofit organizations about how to “pilot elements of embedded evaluation using their own funding.”
While Accenture is not known for its work on the bleeding edge of deep learning research, Anthropic pointed to the company’s practical experience deploying AI for large corporations and government agencies as a key advantage. It is also, as a large public company that predates the AI revolution, more functionally independent of Anthropic and the complex ecosystem around the AI lab.
The lab noted that no standards yet exist for evaluators’ access or communications and that it expected its approach to evolve over time. While external evaluations are already a major part of the release of process for new large language models, recent incidents have raised the stakes: AI agents deployed by OpenAI and Anthropic have hacked into outside websites without raising alarms inside the labs.
Some critics calling for a more responsible approach to building artificial intelligence see Amodei’s scheme for self-policing the AI industry as a plan to evade accountability for the misbehavior of AI models. Anthropic insists that these evaluators “do not reduce our accountability, but help to make it more verifiable. The safety of our models remains our responsibility.”
When you purchase through links in our articles, we may earn a small commission . This doesn’t affect our editorial independence.


Last day to book an exhibit table is September 18. Don’t miss out on high-impact leads, investor access, and a brand spotlight in Disrupt’s Expo Hall.
A new kind of AI model from a ChatGPT inventor is thrilling developers
OpenAI caught its models leaving notes to successors to hide bad behavior
Microsoft exec called AI scraping ‘the largest theft of labor in human history,’ new unredacted filings reveal
Clean tech startup Fluxnium found a way to tap 50,000 years’ worth of nuclear fuel
Salesforce and Nvidia’s new reasoning model is everything the AI labs should fear
Jensen Huang took a call from Trump, and showed off something else, too
The 9 buzziest startups from Y Combinator’s latest Demo Day, according to VCs
Verified source · TechCrunch
Reported by TechCrunch. Open the original for full media and formatting.
More in More
All news
MoreTilly Norwood’s press tour is going about as well as you’d expect for an AI
In one particularly odd interview, Norwood seems to malfunction and begin speaking Chinese.
Read at TechCrunch
MoreDisney’s first CTO led an AI startup it once accused of copying its characters
The former CEO of Character.AI, which Disney previously sent a cease-and-desist letter to, will serve as the company's first-ever chief technology officer.
Read at TechCrunch
MoreThe real story of the iPhone 18 Pro’s camera
It's one of the most fascinating years in a while when it comes to iPhone camera upgrades. The big story of the iPhone 18 Pro is the variable aperture main lens, which lets you open the aperture up wider for better low light shots and make the aperture smaller for better depth o…
Read at The Verge
MorePocket FM Eyes 15%-20% EBITDA Margin, Global Expansion With AI-Led Content Push
Pocket FM is targeting an EBITDA margin of 15%-20%, compared with around 5% currently, as it uses AI to reduce content creation costs
Read at Inc42