Rohin Shah

speaker
1,071 appearances 1 recordings 1 series first heard Jun 2026 last heard 2 Jun

Rohin Shah’s voice in public audio — every appearance, attributed to the second.

Trend

recordings per month · last 12 months
1 · Jun OctJan 26AprJulnow

Recordings per month over the last 12 months — 1 in all, peaking in Jun 2026 with 1.

Appearances

newest first · ▶ plays the moment
If you look at much of the work that was done at the time, things like debate,
Those are implicitly predicated on reinforcement learning being the method of choice for building powerful AI systems.
Debate just looks pretty pointless if you're not using reinforcement learning.
And so I was always expecting reinforcement learning to happen.
And from my perspective, everybody else was suddenly surprised by reinforcement learning happening.
Whereas I looked at it and I was like, well, it could have been the case that like once we do reinforcement learning,
It just generalizes beautifully to everything, the same way that instruction following really does generalize beautifully to everything.
And actually, that was not the case.
So I think mostly my timelines didn't change very much because I already thought it was reasonably likely RL wouldn't generalize to everything.
But I think if I had been tracking it sufficiently fine-grained, I would have probably updated slightly towards longer timelines on the release of 01 or 03.
I think the generalizability was the big deal for me.
Like, look, if you want to target some particular benchmark and apply machine learning to it, I think the lesson of machine learning is like, yes, you can do it.
People do choose the ones that are
that models are capable of doing.
So it's not like you can choose some arbitrary thing and just hit it with the machine learning hammer and succeed.
But I do think you have to be pretty careful about if somebody optimized for a specific thing,
How much should you update from that to like AGI, the full generality, the fully general intelligence?
It's like really quite a difficult kind of update to make and you should be looking quite a bit at the generalizability.
Yeah, I think that's right.
And you do see a little bit of this from reasoning models, to be clear.
Showing 881–900 of 1,071 · page 45 of 54 ← Previous Next →