Facebook
Britain's News Portal
Around The Clock
BREAKING
Loading latest headlines…

Anthropic Taps Accenture Unit for Embedded AI Evaluations

Anthropic said staff from Accenture's AI division, Faculty, will work inside the company to evaluate and red-team its models. The two firms expect to invest at least $1 billion in the project over five years.

  • Anthropic said Accenture's Faculty unit will begin evaluating and red-teaming models, conducting alignment assessments and testing model safeguards.
  • Both companies expect to invest at least $1 billion in the project over the next five years.
  • Anthropic said more evaluators will be announced in the weeks ahead and that it is in conversation with METR and other non-profit organisations about piloting elements of embedded evaluation using their own funding.

Anthropic has said that staff from Accenture will begin working inside the company to scrutinise its models and staff, in the first embedded evaluator arrangement to emerge from chief executive Dario Amodei's plans to place third-party safety evaluators inside AI labs.

In a blog post, Anthropic said Faculty, a company Accenture acquired in January to act as its AI division, will begin "evaluating and red-teaming models, conducting alignment assessments, and testing model safeguards." Both companies expect to invest at least $1 billion in the project over the next five years.

The choice of Accenture surprised many AI watchers, and the markets, where the consulting company's shares rose 8% after hours. Discussion of embedded evaluators had previously focused on AI safety research organisations such as METR, Redwood Research and Apollo Research.

Anthropic pointed to Accenture's practical experience deploying AI for large corporations and government agencies as a key advantage, and noted that as a large public company predating the AI boom it is more functionally independent of Anthropic and the ecosystem around the lab. The lab said no standards yet exist for evaluators' access or communications and that it expects its approach to evolve over time.

Anthropic said more evaluators will be announced in the weeks ahead and that it is in conversation with METR and other non-profit organisations about how to "pilot elements of embedded evaluation using their own funding." Some critics see the self-policing scheme as a way to evade accountability for AI model misbehaviour. Anthropic insists the evaluators "do not reduce our accountability, but help to make it more verifiable," adding that "the safety of our models remains our responsibility."

Why this matters: The arrangement is an early test of whether third-party evaluators embedded inside AI labs can scrutinise frontier models, at a time when external evaluations are already part of the release process for new large language models.

Related Articles

Get the news that matters.

Join thousands of readers getting the best of British news straight to their inbox.