Today, we're joined by Maohao Shen, PhD student at MIT to discuss his paper, “Satori: Reinforcement Learning with Chain-of-Action-Thought Enhances LLM Reasoning via Autoregressive Search.” We dig into how Satori leverages reinforcement learning to improve language model reasoning—enabling model self-reflection, self-correction, and exploration of alternative ... Show More
Oct 6
Why Jev Is Changing How We Build With AI with Diogo Almeida - #779
In this episode, Diogo Almeida, co-founder and CEO of TypeSafe, joins us to discuss Jev, TypeSafe’s recently released model for bringing fast, reliable intelligence directly into software. We explore the idea of “machine-native intelligence” and why Diogo believes models optimize ... Show More
1h 30m
Sep 29
From Math Olympiads to Navier-Stokes: How Fast Is AI Progressing? with Greg Burnham - #778
AI systems have gone from struggling with grade-school math to helping solve research problems that have resisted mathematicians for decades, including Navier-Stokes. In this episode, Greg Burnham, who leads AI capabilities research at Epoch AI, joins us to examine what that prog ... Show More
1h 7m
Sep 17
From Voice Agents to AI Avatars with Alexander Smola - #777
Voice AI has gotten remarkably good, but natural conversation remains a high bar. Small delays, awkward interruptions, or the wrong tone can quickly break the illusion—and adding vision and visual presence only raises the stakes. In this episode, Alex Smola, co-founder and CEO of ... Show More
1h 4m
Jul 2023
#130 Mathew Lodge: The Future of Large Language Models in AI
<p data-pm-slice="1 1 []">Welcome to episode #130 of Eye on AI with <a href="https://www.linkedin.com/in/mathew/">Mathew Lodge</a>. In this episode, we explore the world of reinforcement learning and code generation. Mathew Lodge, the CEO of Diffblue, shares insights into how rei ... Show More
49m 44s
Dec 2024
Nature of Intelligence, Ep. 6: AI’s changing seasons
Guest: Melanie Mitchell, Resident Professor, Santa Fe InstituteHosts: Abha Eli PhobooProducer: Katherine MoncurePodcast theme music by: Mitch MignanoFollow us on:Twitter • YouTube • Facebook • Instagram • LinkedIn • BlueskyMore info:Tutorial: Fundamentals of Machine LearningLectu ... Show More
44m 1s
Sep 2023
The Defeat of the Winograd Schema Challenge
Our guest today is Vid Kocijan, a Machine Learning Engineer at Kumo AI. Vid has a Ph.D. in Computer Science at the University of Oxford. His research focused on common sense reasoning, pre-training in LLMs, pretraining in knowledge-based completion, and how these pre-trainings im ... Show More
31m 3s
<p>Hal Ashton, a PhD student from the University College of London, joins us today to discuss a recent work Causal Campbell-Goodhart's law and Reinforcement Learning.</p> <p>"Only buy honey from a local producer." - Hal Ashton</p> <p> </p> <p><strong>Works Mentioned:</strong></p> ... Show More