Mikita Balesni

Mikita Balesni is an AI safety researcher who studies language-model misalignment, evaluations and the monitoring of model reasoning. He was a founding member of Apollo Research and has worked at OpenAI. He previously was a MATS scholar working with Owain Evans and co-authored the 2024 ICLR paper The Reversal Curse.

Statements by Mikita Balesni

Mikita Balesni

Role at the time: Former OpenAI alignment researcher

Condemned

Response to this claim

Mikita Balesni says he did not share company IP and that his outside work was coordinated with his reporting line, research leadership and the board.

“If OpenAI has specific concerns, I invite them to write to us directly. I expect they will not, because our firing was pretextual.”
Read the original post (opens in a new tab)Mikita Balesni denies sharing company IP · Mikita BalesniOriginal thread post describing the exit call, information-sharing accusation and coordination with leadership.
Why we used this label

The conduct under scrutiny is his dismissal. His thread rejects the rationale and criticizes the firing as pretextual.

Topics in this coverage