Anthropic and Accenture Bet $2B on Embedded AI Safety Audits
Anthropic has hired Accenture's Faculty unit to perform embedded, ongoing evaluation of its frontier models — model evaluation, red-teaming, alignment assessments and safeguard testing, all conducted with employee-level access. 'Employee-level access' is the detail worth savouring: the auditors get a badge, a desk, and whatever the lab decides to show them. Anyone who has ever watched an external assurance function sit inside the company it is assuring will recognise the choreography.
The report calls this the first disclosed implementation of Dario Amodei's proposal to put outside evaluators inside frontier AI labs. It is the first *disclosed* one — a word doing the heavy lifting of a small crane. Nothing here says the findings get published, nothing says the arrangement can survive a model that fails its audit, and nothing says Accenture can walk out mid-contract and explain why. Anthropic, meanwhile, is the same company that bought a biotech outfit in April and declined to say what its wet lab is working on beyond confirming it isn't drug discovery. Opacity is a house style; a $1 billion line item for looking at it more closely is not obviously a remedy.
There is no regulator in this sentence, which is rather the point. Amodei's proposal exists because nobody has legislated third-party evaluation into being, so the lab is funding its own scrutiny and picking the scrutineer from a consultancy whose business is selling implementation. The genre is familiar: when the state won't build the referee, the league hires one and calls it governance. The audits are described as ongoing and embedded, which in practice means they end whenever the money or the mood does. Still, credit where it's due — two billion dollars is a genuinely large number for a field that spent three years insisting safety was best evaluated by the people shipping the product.