Is It Time to Rethink LLM Pre-Training? with Aditi Raghunathan - #747 - The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence) | Transcription & Insights

Description

Today, we're joined by Aditi Raghunathan, assistant professor at Carnegie Mellon University, to discuss the limitations of LLMs and how we can build more adaptable and creative models. We dig into her ICML 2025 Outstanding Paper Award winner, “Roll the dice & look before you leap: Going beyond the creative limits of next-token prediction,” which examines why LLMs struggle with generating truly novel ideas. We dig into the "Roll the dice" approach, which encourages structured exploration by injecting randomness at the start of generation, and the "Look before you leap" concept, which trains models to take "leaps of thought" using alternative objectives to create more diverse and structured outputs. We also discuss Aditi’s papers exploring the counterintuitive phenomenon of "catastrophic overtraining," where training models on more data improves benchmark performance but degrades their ability to be fine-tuned for new tasks, and dig into her lab's work on creating more controllable and reliable models, including the concept of "memorization sinks," an architectural approach to isolate and enable the targeted unlearning of specific information. The complete show notes for this episode can be found at https://twimlai.com/go/747.

Audio

Featured in this Episode

No persons identified in this episode.

Transcription

This episode hasn't been transcribed yet

Help us prioritize this episode for transcription by upvoting it.

0 upvotes

🗳️ Sign in to Upvote

Popular episodes get transcribed faster

Other recent transcribed episodes

Transcribed and ready to explore now

NPR News: 12-08-2025 2AM EST

08 Dec 2025

NPR News Now

NPR News: 12-07-2025 11PM EST

08 Dec 2025

NPR News Now

NPR News: 12-07-2025 10PM EST

08 Dec 2025

NPR News Now

Meidas Health: AAP President Strongly Pushes Back on Hepatitis B Vaccine Changes

08 Dec 2025

The MeidasTouch Podcast

Democrat Bobby Cole Discusses Race for Texas Governor

07 Dec 2025

The MeidasTouch Podcast

Fox News Crashes Out on Air Over Trump’s Rapid Fall

07 Dec 2025

The MeidasTouch Podcast

Comments

There are no comments yet.

Please log in to write the first comment.

The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence)

Is It Time to Rethink LLM Pre-Training? with Aditi Raghunathan - #747

This episode hasn't been transcribed yet

Other recent transcribed episodes

NPR News: 12-08-2025 2AM EST

NPR News: 12-07-2025 11PM EST

NPR News: 12-07-2025 10PM EST

Meidas Health: AAP President Strongly Pushes Back on Hepatitis B Vaccine Changes

Democrat Bobby Cole Discusses Race for Texas Governor

Fox News Crashes Out on Air Over Trump’s Rapid Fall

Login Required

Share this moment