DeepSeq R1
concept
1 mentions
1 recordings
first heard Nov 2025
last heard 25 Nov
↑1 vs the 12 months before
Paper describing the DeepSeq R1 reinforcement learning approach
Trend
Peak in the week of 19 Nov — 1 mentions across 1 series.
Moments
newest first · ▶ plays the momentDwarkeshDwarkesh Podcast · Ilya Sutskever — We're moving from the age of scaling to t... · 15:52 · 25 Nov
…this was in the DeepSeq R1 paper, is that the space of trajectories is so wide that maybe it's hard to learn a mapping from an intermediate trajectory and value.…
All 1 moments