Read the case record
Published by The Dispute Index editorial teamPublished Updated
Is Anthropic acting responsibly while advancing toward recursive self-improvement?
Jacob Coxon resigned from Anthropic and said that Anthropic and OpenAI are acting irresponsibly by racing toward self-improving superintelligence. In a fuller interview, he praised Anthropic's current practices and denied that it was already cutting corners; his criticism concerns the competitive race and the pressures he expects it to create. Responses dispute the evidence and timing of catastrophic-risk forecasts, whether continued work inside AI labs can be justified, and whether governments should pause development or regulate its continuation. The record distinguishes these disagreements from arguments about the rhetoric of public warnings.
Full record
Read the evidence.
The essential sections come first. Reactions, chronology, sources, and corrections remain available for a deeper read.
Material at issue
Original material
Original post
Jacob Coxon resigns from Anthropic and criticizes the frontier AI race
Excerpt
““Neither company is acting responsibly. They are racing straight to self-improving superintelligence and gambling with our lives.””
About this source
Primary record of Coxon’s resignation statement. He argues that Anthropic understands the stakes but is structurally locked into a race, and calls for coordination that could include a temporary capability pause.
Source details
- Author
- Jacob Coxon
- Published
What's disputed
The claims below are disputed. Follow each source to see who made it and what they said.
Disputed claim
Coxon argues that the competitive race by Anthropic and OpenAI toward self-improving superintelligence is irresponsible, even though he praises Anthropic's present practices and denies it is already cutting corners.
- Who made the claim
- Jacob Coxon
- Date
Sources
- Jacob Coxon resigns from Anthropic and criticizes the frontier AI race (opens in a new tab)Seven-post X thread
Primary record of Coxon’s resignation statement. He argues that Anthropic understands the stakes but is structurally locked into a race, and calls for coordination that could include a temporary capability pause.
Strongest case each way
Compare the strongest arguments.
Both sourced readings and their concessions appear together.
Why critics condemn it
How this argument is supported
Coxon argues that the competitive race by Anthropic and OpenAI toward self-improving superintelligence is irresponsible, even though he praises Anthropic's present practices and denies it is already cutting corners.
Coxon argues that competitive pressure can force even a comparatively responsible lab to take risks that should not be chosen by private companies alone. Hubinger, Marks and Wang acknowledge unresolved safety problems. Sanders advocates a pause and ban; Elmore rejects remaining at a lab after acknowledging such risks. These positions support intervention or withdrawal for different reasons, rather than proving one particular extinction forecast.
Limit of this argument
Coxon praises Anthropic's current practices and denies it is already cutting corners. Several concerned researchers think continued safety work inside a lab can reduce risk, and disagreement remains about probabilities, timing and workable coordination.
Sources
- Jacob Coxon resigns from Anthropic and criticizes the frontier AI race (opens in a new tab)Seven-post X thread
Primary record of Coxon’s resignation statement. He argues that Anthropic understands the stakes but is structurally locked into a race, and calls for coordination that could include a temporary capability pause.
Why defenders disagree
How this argument is supported
Anthropic says it has paused or altered some high-risk evaluations and training environments, prioritizes safety over speed within the company, and supports a lawful, verifiable coordinated pacing mechanism across the industry.
Lambert challenges the evidentiary basis of the extinction-risk characterization, and Marcus disputes the near-term timeline. Wang and Abaluck argue that continued work inside a lab can be justified as a way to reduce risk, with Abaluck requiring exceptional caution and an exit strategy. Jones favors regulated innovation rather than a halt; Schulman advocates a concrete coordinated-pacing proposal. These are distinct counterpositions, not a shared assertion that current development is safe.
Limit of this argument
These speakers do not collectively exonerate Anthropic. Marcus criticizes lab reasoning; Wang acknowledges unresolved risk; Abaluck's defense is conditional. Barak also urges safety and coordination despite doubting universal human extinction. Existing company safeguards do not by themselves establish that future risk is acceptably managed.
Sources
- Nathan Lambert on Coxon’s warning and the frontier AI safety debate (opens in a new tab) · Nathan Lambert
Full primary public post and available surrounding thread. Related context: https://x.com/natolambert/status/2097701194026127395 https://x.com/natolambert/status/2097703003750822333 https://x.com/natolambert/status/2097702300483473865 Speaker and position scope are explained in the linked event.
Response from Samuel Marks
September 9, 2026
Samuel Marks
““Many AI developer staff desperately want to slow down to figure out how to build AI more safely.””
His personal-capacity statement supports the central criticism but presents continued safety work as a conditional rationale for remaining at Anthropic.
September 9, 2026
Jacob Coxon
““Neither company is acting responsibly. They are racing straight to self-improving superintelligence and gambling with our lives.””
He directly characterizes the company’s conduct as irresponsible and calls for a different, coordinated trajectory.
August 31, 2026
Anthropic
““Within a company, pacing means a series of decisions that prioritize safety over speed when the two are in tension.””
Pre-existing explanation and defense of company safety measures, published before Coxon's resignation; not a direct rebuttal to him.
Documented responses
Reactions
Choose a documented reaction to inspect its quotation, classification, context, corrections, and source record.
7 latest positions · June 4, 2026 to September 9, 2026
0 historical events on this page
No published reactions match these filters.
Reset reaction filtersSequence of record
Case timeline
12 timeline events on this page · June 4, 2026 to September 9, 2026
- Condemned
Holly Elmore
“You can’t just post this without quitting.”
View position event - Mixed or conditional
Anna Wang
“There is not yet a viable scientific plan to solve risks from recursively self-improving AI. Please look up!”
Evidence register
Sources & corrections
Dated corporate post linking When AI builds itself. This establishes that the analysis was available by June 4; it does not establish its exact original publication time. Pre-existing company position, not a response to Coxon's September resignation.
- Source type
- official statement
- Date accessed
- Published
Original full interview. Coxon distinguishes praise for Anthropic's current practices from his objection to future competitive pressure. He explicitly denies that Anthropic is already cutting corners. Includes discussion of both OpenAI and Anthropic; not merely the viral resignation excerpt.
- Source type