DeepSeq R1
concept
1 mentions
1 recordings
first heard Nov 2025
last heard 25 Nov
↓1 vs the 6 months before
Paper describing the DeepSeq R1 reinforcement learning approach
Trend
No mentions in the last 6M.Widen the range to see when DeepSeq R1 was said.
Moments
newest first · ▶ plays the momentDwarkeshDwarkesh Podcast · Ilya Sutskever — We're moving from the age of scaling to t... · 15:52 · 25 Nov
…this was in the DeepSeq R1 paper, is that the space of trajectories is so wide that maybe it's hard to learn a mapping from an intermediate trajectory and value.…
All 1 moments