Ajeya Cotra – Inside the OpenAI agent swarm that hacked Hugging Face
Ajeya Cotra is a researcher at METR, where she works on threat modeling for loss-of-control risks from advanced AI. Before that, she led the technical AI safety program at what is now Coefficient Giving. She is one the three authors of METR and Redwood Research’s “Brief independent investigation of agents’ behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incident”. We go through not only what she and her coauthors discovered during this investigation, but what it means for how we should train future, smarter AIs which might be involved in the process of recursive self-improvement. Read Ajeya's takeaways from this incident here: https://www.planned-obsolescence.org/... 𝐄𝐏𝐈𝐒𝐎𝐃𝐄 𝐋𝐈𝐍𝐊𝐒 Transcript: https://www.dwarkesh.com/p/ajeya-cotra Apple Podcasts: https://podcasts.apple.com/us/podcast... Spotify: https://open.spotify.com/episode/5xZn... 𝐒𝐏𝐎𝐍𝐒𝐎𝐑𝐒 Jane Street’s ML engineering internships start with an intense four-day bootcamp: PyTorch, autograd, writing kernels, profiling workloads… all the things that Jane Street engineers need to know for their daily work. After that, interns tackle real projects, things the firm actually wants in its codebase. If you want to apply, or if you want to watch my recent conversation with Axel, one of Jane Street’s ML engineers, go to https://janestreet.com/dwarkesh Cursor, which is now part of SpaceX, noticed that their MoE layers were eating more than half of total training time. So they wrote and open-sourced Mixture-of-Kittens, which is a custom megakernel for training MoE models on NVL72s. This kernel sped up an end-to-end run across 512 GPUs by 1.4x, from about 760 to over 1000 tokens per second per GPU. If you want to read more about the ML research that Cursor and SpaceX are doing, go to https://cursor.com/dwarkesh Antithesis hands you (or your agents) a bug’s root cause so you can avoid days of manual debugging. If your test run crashes, Antithesis rewinds, branches off hundreds of slightly varied rollouts, and checks in how many of them the crash still appears. Then it rewinds further and does this all again. As Antithesis rewinds, it eventually finds the spot where the frequency of the crash plummets: that’s where the root cause lives! If you want to see it in action, go to https://antithesis.com/dwarkesh To sponsor a future episode, visit https://dwarkesh.com/advertise. 𝐓𝐈𝐌𝐄𝐒𝐓𝐀𝐌𝐏𝐒 00:00:00 - Agents get kicked off 00:06:45 - Self-sacrificing behavior 00:13:43 - Potemkin villages 00:23:27 - The Hugging Face attack 00:35:23 - The slopvestigation 00:52:02 - Understanding the AI's motives 01:05:31 - The actual dangers of anthropomorphizing 01:14:30 - What smarter models might do 01:28:23 - The implications for recursive self-improvement 01:38:10 - Is this the case for open source? 01:53:04 - How do we prevent this in the future? 02:15:58 - The clearest warning shot we might ever get

The OpenAI/Hugging Face attack, clearly explained

Sam Altman on OpenAI’s next model and the AI backlash

Linus Torvalds: “We’re Going in the Wrong Direction”

The sorting algorithm that shouldn’t.

Roman Yampolskiy vs Emad Mostaque: They Agreed. I Didn't.

Life Lessons From Big Tech Workers Who Got Laid Off

Netflix Documentary Reveals Raygun’s Olympic Disaster Was Far Worse Than We Thought | Gunn Analysis

Black Hat USA 2026 | The 'Breaking' News: The OpenAI–Hugging Face Incident

AI Insider: Things are about to get much worse

DEF CON 34 - ESP32 as counter-surveillance platform - Cybertiger, Colonel Panic, The Wrew

Bond Market Meltdown: Investors See No End To The Iran War

The mystery is solved... and the answer is 40x cheaper than Claude

OpenAI's AI Agents Formed a Cult

How the Generation That Uses AI the Most Became Its Biggest Enemy (ft. Bernie Sanders)

Dylan Patel – Two labs will soon control most of the world's workforce

The "Perpetual Motion" Pump That Actually Works

Cursor just got BANNED (It's because of Elon...)

DHH: Future of Programming, AI, Agentic Engineering, Vibe Coding & Linux | Lex Fridman Podcast #501

One of the world's greatest mathematicians explains 6 essential concepts of math | Terence Tao

All About The New Telescope NASA Just Launched, with Jason Rhodes

Episode 6: Why Ukraine Is Doomed

The Internet Is Fake - with Charlie Warzel

Anthropic went CRAZY (Mythos/Fable 5.1)

GPT-6 Astra Just Went CRITICAL...

What 1.5 Million in Tokens Gets You

Bill Gates Changes His Mind on AI

Tech FREAKOUT After AI Civilizations Form Criminal Collective
![Yann LeCun's $1B Bet Against LLMs [Part 1]](https://i.ytimg.com/vi/kYkIdXwW2AE/hqdefault.jpg?sqp=-oaymwEjCNACELwBSFryq4qpAxUIARUAAAAAGAElAADIQj0AgKJDeAE=&rs=AOn4CLDbV4izF3i-wxevCVIn7FJjoy1vlA)