Rohin Shah
speaker
1,071 appearances
1 recordings
1 series
first heard Jun 2026
last heard 2 Jun
Rohin Shah’s voice in public audio — every appearance, attributed to the second.
Trend
recordings per month · last 12 monthsRecordings per month over the last 12 months — 1 in all, peaking in Jun 2026 with 1.
Appearances
Mostly, these do not involve changing the model weights.
because that's the thing that is most constrained.
That's the artifact that you have to produce one of it, and it's got a gazillion constraints imposing it.
But we have these out-of-the-model filters that we can change a bit more.
We can target them to specific prompts or specific problems.
And so those are a lot easier to update over time.
But not everything that you want to do is going to be solvable with an out-of-model filter.
So I guess going back to the original question you mentioned about, you know, our company's adversaries versus apathetic, I think basically my take is that, yeah, because of this like huge interaction of various constraints, it's just really the easiest way to model at least GDM, but probably most AI companies is
they can only do a few things at a time for safety.
And they will do them.
They do care.
But maybe you should just think of them as apathetic.
You should really try to really lay out, here's exactly what you need to do.
Here's why it's not going to hurt any of these other constraints that you care about.
Also, just because everyone is busy.
Yeah, I definitely think that is important to do.
I don't think it's very much in conflict with anything that I've said.
I think the way that this happens currently is we release models and then everybody sees how capable they are, which is by far the most important thing.
That's a kind of transparency that's needed.
In fact, we used to run this survey of researchers at GDM on what their views on X-risk and other kinds of safety were.
Showing 201–220 of 1,071 · page 11 of 54
← Previous
Next →