
Anthropic has hired Accenture to audit the safety of its AI models from inside the company. The arrangement makes Accenture an embedded evaluator, an outside expert who works within a firm and tests its systems.
Accenture’s Faculty unit will fill the role with access comparable to an employee’s. As a result, safety testing of advanced AI shifts from internal review toward outside scrutiny.
Anthropic Names Accenture as Its First Safety Evaluator
On the 18th of September, Anthropic announced the partnership. Faculty, Accenture’s specialist AI business, will lead the effort. Its team will red-team models, assess alignment with human values, and test safeguards.
For example, Red-teaming means attacking a model deliberately to expose weaknesses. Accenture’s Faculty brings experience because it has tested models for leading AI labs. Inside the company, evaluators will watch models during training and follow decisions about building and deploying them.
In addition, evaluators will talk to staff directly and work alongside internal teams and safety partners. Each company expects to invest at least $1 billion over five years.
Because no settled system for funding independent evaluation exists, the lab will pay Accenture directly. Accenture CEO Julie Sweet called embedded evaluation “an emerging area” and an important part of the safety landscape.
Amodei’s Slowdown Plan Drove the Deal
Furthermore, the deal grew from a proposal by Dario Amodei, Anthropic’s chief executive. He published a three-step plan in September to slow AI development. The plan seeks to do so without costing the United States its lead in AI. Independent evaluators with employee-like access form its first step.
Additionally, Anthropic committed to the step unilaterally, so the Accenture deal makes the pledge concrete. However, reactions split. Sam Altman and Elon Musk backed the framework, while Nvidia’s Jensen Huang opposed it.
Independent Oversight Raises the Stakes
Embedded access matters because evaluators can verify safety commitments and identify blind spots from inside the company. Faculty’s chief executive, Marc Warner, supports the approach, saying AI should be “safe by design, not safe by accident.”
However, embedded evaluation remains new, and no standards define evaluator access or reporting. Funding adds another tension, because Anthropic pays the firm assessing it.
Because Accenture also helps businesses and governments deploy AI, observers question whether it can stay fully independent. Anthropic responds by stressing it stays accountable for model safety despite the outside help.
Next Steps: Additional Partners and Funding Models
Anthropic also relies on variety. The agreement is non-exclusive, so more evaluators will follow in the coming weeks. On another end, Anthropic talks with METR and other nonprofits about piloting program elements using their own funding.
In the long run, it proposes pooled industry contributions or government funding to replace direct payment. On Accenture’s end, it will work with other AI developers in similar roles.
Consequently,early disclosure lets rival developers study the process, and if they follow, third-party audits could become standard. Credibility depends on whether evaluators can publish findings without Anthropic’s sign-off . Ultimately, the answer will show whether embedded evaluation delivers real accountability or only reassurance.
