GRPO

concept
9 mentions 5 recordings first heard Oct 2025 last heard 21 May ±0 this week

Reinforcement‑learning algorithm used in open‑source LLM training

Trend

mentions per day · public audio 7d30d6M12M
No mentions in the last 7d.Widen the range to see when GRPO was said.

Moments

newest first · ▶ plays the moment
Yann DuboisThe MAD Podcast with Matt Turck · OpenAI's Yann Dubois: Why AI Progress Suddenly Feels Real · 51:14 · 21 May
…They take GRPO, they apply it in like many different places and it just works.…
· Open transcript →
Yann DuboisThe MAD Podcast with Matt Turck · OpenAI's Yann Dubois: Why AI Progress Suddenly Feels Real · 42:21 · 21 May
…In the open source world, GRPO seems to be working very well.…
· Open transcript →
Unknown"The Cognitive Revolution" · The RL Fine-Tuning Playbook: CoreWeave's Kyle Corbitt on GRP... · 1:40:03 · 1 May
GRPO came through and pulled an unsatisfying play.…
· Open transcript →
Kyle Corbett"The Cognitive Revolution" · The RL Fine-Tuning Playbook: CoreWeave's Kyle Corbitt on GRP... · 16:51 · 1 May
…So first of all, like yeah, I think the reason……
Nathan Labenz"The Cognitive Revolution" · The RL Fine-Tuning Playbook: CoreWeave's Kyle Corbitt on GRP... · 1:16 · 1 May
…What distinguished the Deep Seek GRPO algorithm from its predecessors,……
Sebastian RaschkaThe MAD Podcast with Matt Turck · State of LLMs 2026: RLVR, GRPO, Inference Scaling — Sebast... · 1:02:46 · 29 Jan
…Uh I we had like a hopefully not too bad……
Sebastian RaschkaThe MAD Podcast with Matt Turck · State of LLMs 2026: RLVR, GRPO, Inference Scaling — Sebast... · 20:06 · 29 Jan
…And with that, they also introduced the GRPO algorithm you……
9 moments · sign in to read them all