← Zpět

Ajeya Cotra – "This might be the clearest warning shot we ever get"

Dwarkesh Patel

Ajeya Cotra is a researcher at METR, where she works on threat modeling for loss-of-control risks from advanced AI. Before that, she led the technical AI safety program at what is now Coefficient Giving. She is one the three authors of METR and Redwood Research’s “Brief independent investigation of agents’ behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incident”. We go through not only what she and her coauthors discovered during this investigation, but what it means for how we should train future, smarter AIs which might be involved in the process of recursive self

Sarah Paine — Why wars are so difficult to end

This is the final lecture in our series with Sarah Paine, and it's a fitting place to end. Sarah tackles war termination: why wars are easy to start, hard to end, and even harder to end on the terms y