Anthropic is partnering with Accenture’s specialist AI business to conduct independent evaluation and red-teaming of its frontier models. This initiative supports Anthropic's commitment to embed evaluators within the company, allowing them to monitor model development and verify safety commitments.

  • The partnership involves evaluating models, conducting alignment assessments, and testing safeguards.
  • Both companies expect to invest at least $1 billion in this area over the next five years.
  • Embedded evaluators will have access comparable to employees to observe training processes and identify blind spots.
  • Anthropic will directly fund Accenture's work while exploring other funding arrangements for future evaluators.

This approach aims to make accountability more verifiable by allowing independent parties to assess how models are built and deployed, though standards for such embedded evaluation are still being developed.