Key Points
- Anthropic has selected Accenture and its AI business Faculty as its first embedded evaluator to assess AI models and test safety measures.
- The partnership follows CEO Dario Amodei’s proposal to slow the pace of AI development and strengthen safeguards as increasingly capable systems are developed.
- Anthropic and Accenture each plan to invest at least $1 billion in the initiative over five years, while Anthropic expects to add other independent evaluators.
Anthropic Begins Implementing AI Evaluation Plan
Anthropic has partnered with Accenture to establish an embedded evaluation program designed to provide independent testing and assessment of increasingly capable AI models.
The partnership follows a three-step proposal published by Anthropic CEO Dario Amodei on Sept. 12, calling for a more measured approach to AI development and stronger safeguards as the technology advances.
Anthropic said Accenture and its AI business Faculty will work on evaluating and red-teaming models, conducting alignment assessments and testing model safeguards.
The company said the exact structure of the embedded evaluation program is still being developed because the approach is relatively new.
Independent Evaluators Get Greater Access
A central element of Amodei’s proposal is the use of independent evaluators with access comparable to employees.
The objective is to allow outside teams to examine AI systems more deeply, identify potential weaknesses and assess whether safeguards function as intended before increasingly capable models are deployed.
Amodei has warned that AI development is accelerating partly because AI systems are increasingly capable of contributing to the development of subsequent generations of AI.
In his proposal, he described this process as recursive self-improvement and argued that unchecked development could potentially outpace society’s ability to understand and control increasingly capable systems.
The embedded-evaluation initiative represents the first step toward implementing that framework.
Accenture and Anthropic Commit $1 Billion Each
Anthropic and Accenture each expect to invest at least $1 billion in the project over the next five years, according to the companies.
Anthropic said there is currently no established system for consistently funding independent AI evaluation at the scale required.
The company has argued that long-term funding could eventually come from pooled industry resources or governments. Given the urgency it sees in developing evaluation capabilities, however, Anthropic plans to directly fund Accenture’s work under the initial partnership.
Accenture’s Faculty business has experience evaluating AI models and developing AI systems, including work related to safety and ethical deployment.
Partnership Will Not Be Exclusive
The agreement between Anthropic and Accenture is non-exclusive.
Anthropic said it expects to announce additional evaluators in the coming weeks, potentially creating a broader network of independent organizations capable of testing its models.
A larger evaluator network could provide multiple perspectives on model behavior, alignment and safety rather than relying on a single external organization.
The approach also allows Anthropic to begin developing the infrastructure and standards needed for embedded evaluation while the broader model continues to evolve.
AI Industry Debate Continues
Anthropic’s initiative comes amid differing views among technology leaders over the appropriate pace of AI development and regulation.
OpenAI CEO Sam Altman and SpaceX CEO Elon Musk responded positively to elements of Amodei’s proposal, while Nvidia CEO Jensen Huang has argued that additional regulation of AI development is not necessary.
The debate reflects a broader divide within the technology industry over how to balance rapid AI innovation with independent testing, safety mechanisms and oversight.
Anthropic’s partnership with Accenture does not itself establish a regulatory framework, but it creates an external evaluation mechanism within the company’s development process.
Outlook
Anthropic’s partnership with Accenture marks an initial implementation of its proposed embedded-evaluation model, with additional evaluators expected to join in the coming weeks. The effectiveness of the approach will depend on how much access independent teams receive, how evaluations are conducted and how findings influence model development and deployment. The planned $2 billion-plus combined investment over five years also signals a substantial commitment to building independent AI evaluation capacity as increasingly capable systems enter development.
Comparison, examination, and analysis between investment houses
Leave your details, and an expert from our team will get back to you as soon as possible