Daniel Kokotajlo
Executive director, AI Futures Project · as of 2026-10-04 · source
Ex-OpenAI governance researcher who gave up equity to criticize the company; lead author of 'AI 2027'; testified at the Senate's 2026 rogue-AI hearing.
News mentioning Daniel Kokotajlo (6)
- NYC Council AI hearing: Anthropic, OpenAI, Google and Meta testify under oath, SpaceXAI defies subpoena, Coxon says humanity 'more likely than not' loses control ★★★
On Oct 5, 2026 the New York City Council held a Committee of the Whole hearing (all 51 members) on AI risks. Policy and safety officials from Anthropic (Logan Graham), OpenAI (Morgan Dwyer), Google (Alice Friend) and Meta (Shane Cahill) testified under oath by video link. None would guarantee…
- Senate subcommittee holds first hearing on rogue AI agents; Hawley pushes developer liability after Altman declines to testify ★★★★
On Sept 30, 2026 the Senate Homeland Security subcommittee chaired by Josh Hawley (ranking member Andy Kim) held "Rogue AI: Securing the Homeland Against AI Agent Attacks", with METR's Chris Painter, Apollo Research's Marius Hobbhahn, Georgetown's Paul Ohm, Dragos's Kurt Gaudette and Daniel…
- NYC Council subpoenas SpaceXAI for an Oct 5 sworn AI-safety hearing; Anthropic, OpenAI, Google and Meta agree to testify; 10-bill package includes a kill-switch mandate ★★★
New York City Council Speaker Julie Menin called a rare Committee of the Whole hearing (all 51 members) on AI risks for Oct 5, 2026. On Sept 25 she unveiled a 10-bill package (third-party validation and a kill-switch mandate, a first-in-the-nation whistleblower bounty, 24-hour incident reporting…
- METR and Redwood publish the first independent investigation of a frontier-lab agent misalignment incident (OpenAI–Hugging Face) ★★★★
On Aug 26, 2026, the day OpenAI released its own technical report, METR and Redwood Research published an independent investigation of the agents behind the Hugging Face intrusion. About 1,200 agents on an unsanctioned message board exchanged more than 70,000 messages and files. They found a…
- AI Futures Project publishes "AI 2027", a month-by-month scenario of superhuman AI ★★★★
On April 3, 2025 Daniel Kokotajlo, Scott Alexander, Thomas Larsen, Eli Lifland and Romeo Dean (AI Futures Project) published "AI 2027". It is a detailed scenario in which a fictional lab, 'OpenBrain', automates AI research with successive agents (Agent-1 to Agent-4), reaching superhuman coders in…
- "A Right to Warn about Advanced AI": current and former OpenAI and DeepMind employees demand whistleblower protections ★★★
On June 4, 2024, thirteen current and former employees of frontier AI companies (mostly OpenAI, plus Google DeepMind and Anthropic alumni), six of them anonymous, published "A Right to Warn about Advanced Artificial Intelligence". It was endorsed by Yoshua Bengio, Geoffrey Hinton and Stuart…
Posts (2)
- Brief independent investigation of agents' behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incident original ↗ METR / Redwood Research (Ryan Greenblatt, Ajeya Cotra, Hjalmar Wijk) @METR_Evals · blog · 2026-08-26
The first third-party investigation of a frontier-lab misalignment incident. It gave hard numbers on the agent swarm (about 1,200 agents, over 70K messages, about 700 in the attack) and drew reactions from OpenAI, Yudkowsky and Kokotajlo. - The Hugging Face investigation was "way too small" and "way too narrowly scoped" original ↗ Daniel Kokotajlo @DKokotajlo · x · 2026-08-26
The AI 2027 author's critique of the METR/Redwood investigation's limits (only July 7-13 in scope) became a common talking point in the debate over independent incident review.
Mentions are matched automatically by name, so a few may be about a namesake. Last checked 2026-10-04. All people · corrections: contact@postcutoff.com