DeepSeq R1

concept
1 mentions 1 recordings first heard Nov 2025 last heard 25 Nov ↓1 vs the 6 months before

Paper describing the DeepSeq R1 reinforcement learning approach

Trend

mentions per week · public audio 7d30d6M12M
No mentions in the last 6M.Widen the range to see when DeepSeq R1 was said.

Moments

newest first · ▶ plays the moment
DwarkeshDwarkesh Podcast · Ilya Sutskever — We're moving from the age of scaling to t... · 15:52 · 25 Nov
…this was in the DeepSeq R1 paper, is that the space of trajectories is so wide that maybe it's hard to learn a mapping from an intermediate trajectory and value.…
· Open transcript →
All 1 moments