Joe Benton calls for external accountability of frontier-AI companies

Former Anthropic safety-team member, writing in a personal Substack post.

Frontier AI companies are racing to build AI systems that can recursively self-improve.

Source and context

Original text

Joe Benton — Why I left Anthropic’s safety team to hold AI companies accountable (opens in a new tab)Opening paragraphs

About this source

Personal post by a former Anthropic safety-team member describing his assessment of frontier-AI incentives and his move to METR. It is attributed commentary from a former employee, not independent verification of company practices.

Before the quotation

Benton wrote that he had left Anthropic’s safety team two weeks earlier and would join the independent evaluator METR.

After the quotation

He argued that individual companies cannot be relied on to manage the associated societal risks without outside accountability.

Challenged the characterization
How this statement is classified

Benton explicitly challenges the characterization that frontier-AI companies are sufficiently pacing and resourcing safety work. His post is personal commentary from a former employee, not independent verification of internal practices.

Recorded on
Published here

More from this case

Anthropic

We must pursue RSI very carefully, if at all.
Read statement

OpenAI

OpenAI declined to comment when approached by CNBC
Read statement
Read the case