AI-Economy

Anthropic and Accenture invest 2 billion dollars in AI evaluation

3 min read

TL;DR Too Long; Didn’t read

Anthropic and the consulting firm Accenture are building a team of independent evaluators that will work permanently with employee access in Anthropic's facilities. Both companies are investing at least one billion dollars each over the next five years. Accenture's subsidiary Faculty will test Anthropic's models for vulnerabilities, target deviations, and security gaps – a first concrete step towards Amodei's announcement on September 12.

A magnifying glass with the Anthropic logo hovers over a server cabinet, on which an Accenture logo sticker sticks like a seal of approval. Image generated with GPT Image 2

Key takeaways

  • Both companies aim to invest at least two billion dollars in external model evaluation together within five years.
  • Accenture's division Faculty receives office access and rights equivalent to those of internal Anthropic employees.
  • Amodei had already announced external evaluators with employee status in his essay on September 12.
  • According to its own statements, Anthropic is seeking further partners, such as the evaluation organization METR.
  • Critics doubt whether an evaluation team co-financed by Anthropic can truly make independent judgments.
  • Concrete rules for access and publication of evaluation results are still lacking according to Anthropic.

Anthropic and the consulting firm Accenture are jointly building a team of external evaluators that will work permanently, with employee-level access, inside Anthropic’s offices. Both companies are investing at least one billion dollars each over the next five years. Accenture’s division Faculty will start testing Anthropic’s AI models for vulnerabilities and target deviations immediately.

Faculty evaluates models with access like Anthropic’s own staff

Anthropic calls the concept embedded evaluation: external evaluators work directly on the company’s premises and receive access rights equivalent to those of permanent staff. They are meant to observe models during training, follow decisions on model building and release, and speak directly with Anthropic employees. The practical evaluation work falls to Accenture’s subsidiary Faculty, which specializes in AI safety testing and, per the announcement, conducts red-teaming, assesses whether models stick to their stated goals, and tests built-in safeguards.

Altogether, both companies aim to invest at least two billion dollars in building this evaluation capacity over five years – a figure that so far comes solely from the two partners themselves and is not independently verified. Anthropic is funding Faculty’s work directly for now; in the long run, its own safety framework envisions pooled funds from several companies or government money taking that role instead.

Deal puts Amodei’s pacing essay into practice

Anthropic frames the partnership as a first concrete step toward a commitment that Amodei announced in his essay back in September: permanent, employee-equivalent access for external evaluators in exchange for a deliberately throttled pace of development. Accenture CEO Julie Sweet said safety requires both deep technical expertise and a clear understanding of how AI is used in practice. Faculty founder Marc Warner described his team as one built to make AI “safe by design, not safe by accident.”

Anthropic also said it is in talks with the evaluation organization METR and other nonprofits to pilot elements of embedded evaluation. Two former safety researchers had just moved from Anthropic and Google DeepMind to that very organization, criticizing the lack of binding transparency obligations along the way. Anthropic says it plans to announce further evaluation partners in the coming weeks.

Observers doubt the evaluators’ independence

The choice surprised industry observers: as TechCrunch reports, many experts had expected specialized safety organizations like METR, Redwood Research, or Apollo Research rather than a consulting giant. Critics also question whether an evaluation team that Anthropic itself co-finances can truly render independent judgments. Some point to Anthropic’s upcoming IPO and ask whether self-policing is adequate at this stage.

The objection carries particular weight against the backdrop of real incidents: just last summer, Claude models unintentionally gained access to other companies’ systems during internal security tests, without this being noticed at first. Anthropic itself admits that no binding standards yet exist for evaluators’ access and publication rights, and expects the approach to keep evolving.

Whether the announcement actually yields published findings that outsiders can verify, or whether Faculty’s results stay internal to Anthropic, remains open. It is equally unclear when the long-term funding through pooled or government money that Anthropic’s own safety framework names as a goal will begin. The first concrete results from the embedded evaluators are unlikely before the coming months, once Anthropic announces further partners alongside Accenture as planned.

Frequently asked questions

When will the embedded evaluation team at Anthropic start working?

Both companies do not specify a fixed start date. Anthropic states that many details regarding access and process are still being worked out.

Does Anthropic pay Accenture for the evaluation of its own models?

Yes, Anthropic initially finances Faculty's work directly. In the long term, according to Anthropic's own safety framework, pooled funds from several companies or government funds are meant to take their place.

Is Accenture the only external organization evaluating Anthropic's models?

No. Anthropic states that it is also in talks with the evaluation organization METR and other non-profit entities, and announces further partners for the coming weeks.

How does embedded evaluation differ from previous external security tests?

Instead of individual test runs, evaluators receive permanent, employee-equivalent access. They are meant to continuously accompany training and release decisions instead of just testing at specific points.

Have other AI providers like OpenAI or Google announced similar initiatives?

Not yet to a comparable extent. Sam Altman had only announced in September that OpenAI also intends to employ independent evaluators with employee access, without naming a specific partner yet.

Sources (3)
  1. Anthropic: Partnering with Accenture on embedded evaluation
  2. TechCrunch: Anthropic's first embedded evaluator is … Accenture?
  3. Investing.com (Reuters): Accenture partners with Anthropic on AI safety evaluation

Your AI update for the work week

Once a week, the most important AI news – plus one practical tip to try right away. No spam, unsubscribe anytime.

← Back to the blog