Rob Wiblin
speaker
1,787 appearances
3 recordings
1 series
first heard Sep 2024
last heard May 2025
Rob Wiblin’s voice in public audio — every appearance, attributed to the second.
Trend
recordings per month · last 12 monthsNo recordings in the last 12 months.Older appearances are listed below; set an alert to hear about the next one.
Appearances
Trying to go go even f yeah, go even further on this thing that most people are going to regard as desirable and want to incorporate if it's if it's practical.
And then uh in terms of deciding whether this is actually a better path, uh, you know, is it actually going to have better consequences?
Um, I mean, I'm sure I'm sure people have made arguments that alignment might might backfire.
It could be, could be, could be, could be worse than worse than not aligning it.
You can imagine ways, but still unbalance of probabilities, it seems uh like a like a reasonable bet, uh, not something that people are.
super uncertain about, or at least that I feel really uncertain about.
So i I I mean, maybe this is the case that people focus on so much because it is among the better ones that people have ever have have have come up with for trying to pursue differential technology develop uh technological development.
And maybe there's lots of other ones that were left on the scrap heap because it wasn't clear that they were either viable or or desirable.
Uh what what do you think?
So suppose the the smart alloc response to uh reinforcement learning from human feedback being developed by alignment and safety focused people and then applied to make AI useful and uh you know economically um valuable across all kinds of different domains is to say, well, you've you've wasted your time.
Maybe you've even made things worse by by speeding them up, which you didn't want to do.
On the other hand, it seems like a reasonable reaction to say, well, like what what's your plan to not develop any of the technologies that actually makes AGI work?
That doesn't really seem like an alternative.
Uh surely that would
Well, that would just delay things at at best.
And we need to get to this and we need to get to the point that we're at now at some point sooner or later, right?
Um but I guess you're saying even if it's not actively detrimental, it could be kind of useless because well the market like, you know, uh GDM or some other group would have realized that we needed the equivalent of reinforcement learning from human feedback to make it work.
And so they just would have done it at some at some later time anyway.
So um the effect of your work has kind of just been undone.
Yeah, makes sense.
Showing 741–760 of 1,787 · page 38 of 90
← Previous
Next →