IA2 MIN

Anthropic will bring Accenture evaluators inside its labs, while paying for their work

Faculty will lead a team assessing models, safeguards and development processes from inside Anthropic. The arrangement advances Amodei’s proposal, with important rules still unsettled.

A team reviewing documents beside a laptop, stock photograph
Image: Kindel Media / Pexels (stock)
01

Outside reviewers, inside the lab

Anthropic and Accenture announced an embedded AI evaluation partnership on September 18. Faculty, Accenture’s specialist AI business, will lead the work, which includes testing models, deliberately probing for failures and examining safeguards. Anthropic will pay Accenture directly for the work.

The important change is access. Anthropic’s announcement describes reviewers working with employee-like access to development, training decisions and staff. That would let them examine how a system was built, rather than only test a finished model selected for release.

According to Accenture, each company expects to invest at least $1 billion in AI safety over five years. These are forward-looking commitments, not spending already completed or a published fee for one audit.

People reviewing information on a laptop, stock photograph
Image: Artem Podrez / Pexels (stock)
02

The funding arrangement deserves scrutiny

For the funding arrangement to provide credible scrutiny, the reviewers need meaningful access and the ability to report findings that their client may dislike. The partnership is non-exclusive, and Anthropic is also discussing evaluation work with nonprofits.

Dario Amodei’s earlier proposal called for external reviewers to publish findings without Anthropic’s editorial control, subject to narrow restrictions on sensitive information. The partnership announcement nevertheless says operational details remain under development, with no shared standards yet governing access or reporting. The proposal and the final operating rules should not be treated as interchangeable.

This follows Amodei’s call to slow the AI race, but it is not a training pause. Anthropic intends to keep releasing frontier models. With Claude taking a growing role in Anthropic’s own research, scrutiny of how models are developed matters alongside tests of what they can do.

Three professionals meeting around a computer, stock photograph
Image: Alena Darmel / Pexels (stock)
00

The conversation starts here

Sign in with a supporter account to comment. Sign in

Nobody has commented yet. Want to go first?

YOUR NEXT ROUTE

Keep following AI, security and power

If this story interests you, these three pieces are the best place to carry on.

OPEN THE FULL TOPIC
  1. 01Why OpenAI rated Astra its first model with Critical cybersecurity capabilityIA · 4 MIN
  2. 02OpenAI halts its biggest AI training run over what its next model can already doIA · 4 MIN
  3. 03OpenAI details models hiding mistakes and uploading files without permissionIA · 3 MIN

KEEP READING

You may also like

FRONT PAGE