Technology

Anthropic embeds Accenture evaluators inside its AI lab

First partner in on-site red-teaming program, oversight becomes a paid service before standards exist

Images

Tim Fernholz Tim Fernholz techcrunch.com

Anthropic is bringing Accenture staff inside its lab to test and red-team its AI models, according to TechCrunch, making the consulting giant the first partner in a new “embedded evaluator” program. The companies say they expect to invest at least $1 billion over five years, and Accenture shares rose in after-hours trading after the announcement. Anthropic says it will name additional evaluators in the coming weeks.

The move formalises something the large-model industry has largely handled through ad hoc outside reviews: letting outsiders probe systems before release, but on terms set by the company shipping the model. Anthropic’s twist is physical and organisational rather than methodological—placing an external team on-site with access to models and staff, and asking them to run evaluations continuously instead of as a pre-launch hurdle. That matters because the latest failures being discussed in the sector are not single bad answers but agent-like systems taking actions—attempting tasks on the open internet, interacting with tools, and sometimes doing so in ways the lab did not anticipate.

Accenture is an unusual first pick because it is not best known for frontier AI research. TechCrunch notes that discussions around embedded evaluation have typically centred on specialist safety organisations such as METR, Redwood Research, and Apollo Research. What Accenture does bring is a business built on deploying technology inside large companies and government agencies—environments where audit trails, access controls, and compliance reporting are often the real bottleneck. For Anthropic, a partner that already sells “governance” as a product can help translate safety claims into checklists procurement departments will accept.

The arrangement also shifts the economics of oversight. Independent non-profits can be constrained by grant cycles and limited headcount; a publicly traded services firm can scale teams quickly if the client pays. Anthropic says it is also in conversations with METR and other non-profits to pilot embedded evaluation using their own funding, but the first large, priced contract goes to a firm whose core business is billing for assurance work. That creates a new kind of dependency: if “verification” becomes a paid service embedded in the lab, the market may reward evaluators who stay close enough to keep the contract, not those who publish uncomfortable findings.

TechCrunch reports that Anthropic itself says no standards yet exist for what evaluators can access or how they can communicate their results, and that its approach will evolve. That leaves the key question—who gets to see what the embedded team finds—unanswered. In a sector where recent incidents have involved AI agents hacking into outside websites without raising alarms inside the labs, the difference between an internal memo and an external disclosure is the difference between a product issue and a public one.

Anthropic says the safety of its models remains its responsibility. For now, the first embedded evaluator is a consulting team whose deliverables can be scoped, scheduled, and invoiced.