Anthropic partners with Accenture to embed independent evaluators inside its AI development
Anthropic announced a partnership with Accenture on September 18 to embed independent evaluators directly inside its AI development process, with each company committing to invest at least $1 billion in building out the capability over five years.
What's new
The partnership will be led by Faculty, Accenture's specialist AI business, and covers evaluating and red-teaming Anthropic's models, running alignment assessments, and testing safeguards. As Anthropic put it: "Anthropic and Accenture each expect to invest at least $1 billion in building capacity in this area over the next five years."
What makes this different from existing third-party model testing, according to Anthropic, is access: "Unlike today's external evaluators, embedded evaluators will work inside AI companies, with access comparable to an employee's." That means Accenture's evaluators can watch models take shape during training, follow the internal decisions that govern how models are built and deployed, and talk directly to Anthropic staff — rather than only probing a model from the outside after release. Anthropic says it will fund Accenture's work directly, and is separately in talks with the nonprofit evaluator METR and other groups to pilot embedded evaluation under their own funding. The company describes the Accenture deal as non-exclusive: it plans to bring on other evaluators in the coming weeks, and Accenture will offer similar embedded-evaluation services to other AI developers.
Context
Anthropic frames this as a direct follow-through on CEO Dario Amodei's September 12 essay "We Must Pace the Frontier," which called for the industry to slow capability advances in favor of stronger safety practices and floated embedding evaluators inside labs as one concrete step. That essay has already reshaped this week's AI news cycle — Sam Altman, Elon Musk, and Demis Hassabis all publicly endorsed it the same day, and it's now the basis of a federal antitrust lawsuit accusing the four labs of illegally colluding to slow AI development. Anthropic also ties the move back to its June "Advanced AI Framework," in which it argued that independent evaluation ultimately needs pooled or government funding rather than lab-by-lab arrangements — a funding model it says doesn't exist yet, which is why it's paying Accenture directly for now.
Why it matters
External AI safety evaluation today mostly happens after a model ships, using whatever access a lab is willing to grant. Giving an outside evaluator employee-level visibility into training and internal deployment decisions is a meaningfully bigger commitment than publishing a model card or funding a red-team engagement, and the $1 billion-plus, five-year figure signals Anthropic wants this read as durable infrastructure rather than a one-off PR gesture. It also gives Anthropic a concrete answer to skeptics of the "pace the frontier" proposal — including Nvidia CEO Jensen Huang, who has called industry-coordinated slowdown proposals unnecessary, and lawmakers wary of handing frontier labs an antitrust exemption to coordinate on safety: rather than waiting on a government-brokered framework, Anthropic is funding embedded oversight itself while continuing to argue publicly that a permanent version of this needs to be paid for collectively.
Corroborating sources
- Anthropic
https://www.anthropic.com/news/accenture-embedded-evaluation
“Anthropic and Accenture each expect to invest at least $1 billion in building capacity in this area over the next five years.”