Anthropic says it is partnering with Accenture to test the safety of its advanced AI models by embedding Accenture evaluators within Anthropic’s operations. The effort focuses on evaluating how the newest models behave and whether their safety measures function as intended.

Anthropic describes the Accenture team as conducting red-teaming and assessments of safeguards. NDTV reports that the evaluators will also examine alignment, including whether the technology acts in line with human objectives. Bloomberg likewise characterizes the arrangement as having evaluators placed directly inside Anthropic’s workflow to carry out safety testing, rather than relying on external review.

Across both accounts, the collaboration is framed as an internal testing and evaluation process covering AI safety performance, including safeguards and alignment-related checks. The outlets provide similar descriptions, with NDTV adding explicit reference to red-teaming and alignment assessment, while Bloomberg emphasizes the embedded nature of the evaluators and the safety testing goal.