Rohin Shah
speaker
1,071 appearances
1 recordings
1 series
first heard Jun 2026
last heard 2 Jun
Rohin Shah’s voice in public audio — every appearance, attributed to the second.
Trend
recordings per month · last 12 monthsRecordings per month over the last 12 months — 1 in all, peaking in Jun 2026 with 1.
Appearances
But it does seem like probably the best thing you can do in order to actually increase the chance that your work gets used at a company, at least currently.
If they email me, they will probably not get very much of a reply because I get way too many emails.
And I think most of my reports, I think, would be fairly excited to talk to people about this.
I admit that I, at this point, get enough of these that it feels a little bit more like a burden than like an exciting thing.
But it definitely used to feel like an exciting thing before it became very common.
Okay, yeah.
Yeah, so other things...
I had a talk recently on how to theorize so empiricists will listen.
It could equally well have been called how to do safety research so that companies will listen.
The basic points from this, I think the first one, the most obvious one, but it's still worth asking yourself, is do they actually care?
Sometimes people do research on problems that we actually just don't care about and don't think matter.
I think the safety community is like fairly good about not doing this, but it does still happen.
And sometimes people might be a little bit surprised by what we do and don't care about.
So for example, take jailbreaks.
We care a lot about jailbreaks now because the models are looking to be strong enough that misuse could be a serious problem.
But if you look a year or two ago, people were doing a bunch of jailbreak research and saying, look, this means that the companies aren't very good at aligning their models because they're so susceptible to jailbreaks.
And I don't know about other companies, but at least at GDM,
basically our stance on jailbreaks was the models are not like, there's not any like actual misuse scenarios that we're particularly worried about given the model capabilities.
What safety is about is about protecting users from cases where the model does something unintended and harmful.
And so for a while, there was just a bunch of discourse about how models are always going to be jailbreakable, totally impossible to defend against it.
Showing 921–940 of 1,071 · page 47 of 54
← Previous
Next →