Anthropic taps Accenture as first embedded evaluator for AI safety plan
Anthropic has selected Accenture as its first embedded evaluator, marking the initial implementation of CEO Dario Amodei's proposal to slow AI development. Both companies will invest at least $1 billion over five years, with Anthropic funding the work directly for now. The partnership embeds Accenture employees to test safeguards, red-team models, and assess alignment with human values. Anthropic committed unilaterally to granting third-party evaluators employee-level access, and is also in discussions with nonprofit METR and other parties. The move follows researcher warnings about catastrophic AI harm and Amodei's three-step plan, which drew mixed reactions from industry leaders. Anthropic emphasized it retains accountability for model safety and expects its approach to evolve.