Anthropic Just Gave a Consulting Firm the Keys to Watch Its Own AI

AI customer support

Never lose a customer to a missed message

An AI agent trained on your own business, replying in seconds, in any language, on every channel your customers already use.

Try it free →replio.live

Anthropic announced on September 18 that it is bringing an outside firm inside its own walls to check its work. The Anthropic embedded evaluator arrangement gives Accenture’s specialist AI unit, Faculty, access comparable to an Anthropic employee’s, so it can red-team models, run alignment assessments and test safeguards from the inside rather than from a distance.

Both companies say they expect to invest at least $1 billion each over the next five years to build out this kind of independent evaluation capacity. It is one of the largest financial commitments any AI lab has made specifically to outside safety oversight.

The short version: Anthropic announced the deal September 18. Accenture’s Faculty unit leads the work. Each side is committing $1 billion. Evaluators get employee-level access. The arrangement is non-exclusive. Anthropic is talking to other evaluators too.

What an Anthropic Embedded Evaluator Actually Does

Rather than reviewing a model from outside after it ships, an embedded evaluator sits inside the company during training and deployment decisions. Faculty’s team will evaluate models, red-team them for weaknesses, run alignment assessments and stress-test the safeguards meant to keep systems behaving as intended.

Anthropic describes the access level as similar to what an employee would have. That is a significant departure from how most AI safety audits work today, where outside reviewers typically see a finished product rather than the process that built it.

Why Accenture’s Faculty Unit Got the Job

Faculty is Accenture’s specialist AI business, built around data science and applied AI work rather than general consulting. Anthropic said the partnership is non-exclusive. The company is also in talks with METR and other nonprofit evaluators to pilot similar embedded arrangements, meaning Faculty will not be the only outside group with this kind of access going forward.

The Billion-Dollar Commitment Behind the Deal

Anthropic frames the arrangement as a direct step toward a commitment CEO Dario Amodei made in a September 2026 essay on AI development pacing. Amodei has argued publicly that frontier labs need real, resourced outside scrutiny rather than voluntary self-reporting. Putting $1 billion behind that argument is meant to signal the commitment is more than words.

What This Means for AI Safety Oversight

Embedded evaluation does not replace regulation, and it is still a company paying for its own oversight, which raises an obvious independence question. Anthropic’s answer is transparency about the arrangement and a willingness to bring in multiple evaluators rather than just one. Whether that satisfies critics who want government-run auditing will depend on what Faculty’s team actually publishes about what it finds.

This is not oversight by a government regulator. It is not a court order. It is a private deal between two companies. That distinction matters to critics. It matters less to Anthropic, which argues speed beats waiting on legislation.

Critics of self-funded oversight point out that an evaluator paid by the company it evaluates has an incentive, even a subtle one, to avoid findings that damage the relationship. Anthropic’s counter-argument is that embedded access produces far more useful findings than an outside audit ever could, since Faculty’s team will see design decisions as they happen rather than reconstructing them after the fact.

How an Anthropic Embedded Evaluator Differs From a Traditional Audit

Traditional AI audits typically involve a third party testing a finished model against a checklist, producing a report weeks or months after the system has already shipped. An embedded evaluator instead sits alongside engineers during development, able to flag a concerning design choice before it becomes a shipped feature. Supporters say this catches problems earlier. Skeptics say it also means the evaluator becomes closer to the organization it is supposed to be scrutinizing, blurring the line between oversight and collaboration.

Simple to send.
Safe to verify.

OTPs over WhatsApp, one API call away

Try it free →replio.live

What Happens Next

Faculty’s evaluators are expected to begin embedded work in the coming months. Anthropic has not said whether findings will be published in full or summarized. Industry watchers will be looking for the first public report as the real test of whether this model produces meaningfully independent scrutiny or simply a more sophisticated form of self-review.

Other frontier labs are watching too. If Anthropic’s approach produces credible, publicly visible findings, competitors may face pressure to adopt something similar rather than rely on internal review alone.

The numbers at a glance: Announcement date: September 18. Anthropic’s commitment: at least $1 billion. Accenture’s commitment: at least $1 billion. Time frame: five years. Lead unit: Faculty. Access level: comparable to an employee’s.

Questions About the Anthropic-Accenture Deal

What is an embedded evaluator?
An outside reviewer given employee-level access inside a company to assess AI models during training and deployment, rather than reviewing only the finished product.

How much are Anthropic and Accenture committing?
Each company expects to invest at least $1 billion over five years in building this evaluation capacity.

Is Accenture the only embedded evaluator Anthropic will use?
No. The arrangement is non-exclusive, and Anthropic is in discussions with METR and other nonprofit evaluators too.

What prompted this partnership?
Anthropic ties it to a September 2026 essay by CEO Dario Amodei calling for embedded, resourced outside evaluation of frontier AI labs.

Will the findings be made public?
Anthropic has not detailed how much of Faculty’s evaluation work will be published.

Further Reading on This Story

Sources

  • Anthropic — Partnering With Accenture on Embedded Evaluation. anthropic.com
  • TechCrunch — Anthropic’s First Embedded Evaluator Is … Accenture? techcrunch.com
  • CNBC — Anthropic Selects Accenture as Its First Embedded AI Safety Evaluator. cnbc.com

WhatsApp OTP API

Verification your users actually receive.

Send one-time passcodes over WhatsApp with a single API call. Replio can generate, hash and verify the code for you.

Try it free →replio.live

Author: Francisca Samuel

Francisca Samuel is an editor at Tamara News, where she covers immigration, travel, business and technology news for readers across Africa and the Gulf.