In just two months, a scrappy three-person team at OpenAI sprinted to fulfill what the entire AI field has been chasing for years—gold-level performance on the International Mathematical Olympiad problems. Alex Wei, Sheryl Hsu and Noam Brown discuss their unique approach using general-purpose reinforcement learning techniques on hard-to-verify tasks rather t ... Show More
Aug 18
Rich Sutton and Khurram Javed: Why AI Models Stop Learning, and How to Start It Again
Rich Sutton, who helped pioneer reinforcement learning and wrote the seminal AI essay The Bitter Lesson, has now cofounded Oak Lab with his former student Khurram Javed. Their goal: to build agents that continuously learn from their own experience rather than from us. Rich doesn' ... Show More
53m 43s
Nov 2024
AI and the Future of Math, with DeepMind’s AlphaProof Team
In this week’s episode of No Priors, Sarah and Elad sit down with the Google DeepMind team behind AlphaProof, Laurent Sartran, Rishi Mehta, and Thomas Hubert. AlphaProof is a new reinforcement learning-based system for formal math reasoning that recently reached a silver-medal st ... Show More
39m 21s
Sep 2024
OpenAI's Reasoning Machine + Instagram Teen Changes + Amazon RTO Drama
<p>Last week, OpenAI released a preview of its hotly anticipated new model, o1. We discuss what it has excelled at and how it could accelerate the timeline for building superintelligence. Then, we explain why Meta is making teenagers’ Instagram accounts private by default. And, f ... Show More
1h 6m
Aug 2025
OpenAI Goes Open Source!?
In this conversation, Jamie and Jaeden discuss the recent launch of OpenAI's new open-sourced AI reasoning models. They explore the significance of this development, its implications for businesses, and potential use cases for the models. The discussion also touches on the compet ... Show More
9m 44s