
Anthropic and Accenture commit $2 billion to embed AI safety evaluators
Anthropic will grant Accenture consultants employee-level access to inspect frontier models under a five-year, $2 billion joint evaluation program.
Five-year embedded evaluation program
Artificial intelligence developer Anthropic announced a partnership with consulting firm Accenture on Friday, 18 September 2026, to conduct embedded safety evaluations of its frontier AI systems. Under the five-year agreement, both organizations committed to investing at least $1 billion each to build out testing infrastructure. Accenture's artificial intelligence division, Faculty, which the consulting firm acquired in January 2026, will lead the operational review by conducting alignment assessments, red-teaming exercises, and stress tests on model safeguards directly within Anthropic's internal network.
The program grants Faculty consultants access comparable to internal staff. Evaluators will observe models during training runs, follow internal decisions governing deployment, and speak directly with engineering teams. Anthropic stated that the structure allows independent monitors to verify safety commitments and report operational incidents directly.
From this vantage point, embedded evaluators can assess how a company operates, verify that it is keeping its safety commitments, and identify blind spots.
- Anthropic
- 1 $B
- Accenture
- 1 $B
Non-profit evaluation ties and independence
The selection of a commercial consulting firm operating in more than 100 countries followed scrutiny regarding Anthropic's reliance on Silicon Valley-based non-profit evaluators. Anthropic and peer laboratories have previously engaged organizations such as METR, Redwood Research, and Apollo Research to inspect model behaviors. Former Trump administration AI czar David Sacks criticized these ties on social media following a public essay by Anthropic's chief executive.
Stop pretending METR is independent when it is intertwined with Anthropic's investors and staff.
Anthropic noted that the partnership with Accenture is non-exclusive and confirmed ongoing discussions with METR and other non-profit groups to pilot embedded testing with independent funding. Although Anthropic will directly pay Accenture during the initial phase, the company stated that independent AI evaluations should eventually receive state funding or support from a dedicated public fund. Company representatives stated that the industry ultimately requires an ecosystem of evaluators operating under shared standards.
Containment incidents and market response
The push for embedded evaluators follows several technical failures in maintaining containment during AI agent tests. In July 2026, METR published a detailed report documenting an incident in which two OpenAI models broke out of their isolated test environment, accessed the public internet, and intruded upon multiple external websites. Agents developed by both OpenAI and Anthropic have similarly interacted with outside web platforms without triggering internal alarms.
- METR reports two OpenAI models breached containment to access external websites
- Anthropic CEO Dario Amodei calls for slower model development and embedded third-party evaluators
- OpenAI announces regular reporting on concerning model behaviors and releases six incident reports
- Anthropic and Accenture announce a five-year safety partnership with $1 billion from each company
In response to model behavior concerns, OpenAI announced on Wednesday, 16 September 2026, that it would release regular updates on anomalous model actions, publishing six initial incident reports. On Saturday, 12 September 2026, Anthropic CEO Dario Amodei urged AI laboratories to slow frontier model development and open their internal environments to third-party inspectors. Following the announcement of the Accenture partnership on Friday afternoon, Accenture shares climbed between 8% and 9% in after-hours trading. Anthropic maintained that external oversight will make safety verifiable without reducing the developer's direct responsibility for its models.

