India, Sept. 20 -- As frontier models become harder to assess from the outside, Anthropic is bringing independent evaluators closer to the systems being built

The AI industry is getting better at building increasingly capable models. It is still working out how to properly inspect them.

That gap is becoming harder to ignore. As frontier models move from generating text and code to handling longer, more complex tasks, companies need to know not just what a model can do, but how it behaves under pressure, where its safeguards fail, and whether the processes around its development are working as intended.

Anthropic and Accenture are now putting significant money behind one possible answer.

The two companies said on September 18 that they...