Rob Wiblin
speaker
1,787 appearances
3 recordings
1 series
first heard Sep 2024
last heard May 2025
Rob Wiblin’s voice in public audio — every appearance, attributed to the second.
Trend
recordings per month · last 12 monthsNo recordings in the last 12 months.Older appearances are listed below; set an alert to hear about the next one.
Appearances
And those include that: one, it's hard to trust that companies are gonna stick to their RSPs long term.
Two, it's difficult to truly measure what models can and can't do.
And an RSP is useless if it's mismeasuring model capabilities.
Three, it's pretty questionable whether profit-motivated companies will go out of their way to act in good faith and make their lives and their product release.
And four, that in the most important cases, we just don't have safeguards that are able to render new AI capabilities safe, even capabilities that could show up really soon.
At the end of the day, I come down thinking that responsible scaling policies are a really solid step forward from where we are now.
And they're a great way to test and learn what works and what feels practical from the experience of people who are working on the coal face of trying to make this technological revolution actually happen.
But I think in time, they're going to have to be put into legislation and operated by external groups or auditors rather than stay left just to companies themselves.
At least if if they're going to achieve their real potential or the or the full potential that I think is there.
Of course, uh Nick and I debate that take of mine as well.
If you want to let us let us know your reaction to this interview or any other interview that we do, then our inbox is always open at podcast at eighty thousand hours dot org.
But now, here's my interview with Nick Joseph, recorded on the thirtieth of may twenty twenty four.
Today I'm speaking with Nick Joseph.
Uh Nick is head of training at the major AI company Anthropic, where he manages a team of over 40 people focused on training Anthropic's large language models, including Claude, which I imagine uh many, many listeners have uh have heard of and and potentially used as well.
He was actually one of the relatively small group of people to uh to leave OpenAI alongside Dario and Danielle uh Amade, um who then went on to found uh Anthropic back in December of 2020.
Uh so
Thanks so much for coming on the podcast, Nick.
Thanks for having me.
I'm excited to be here.
Uh I'm really hoping to talk about how Anthropic is trying to prepare itself for training models capable enough that we're a little bit scared of what they what they might go and do.
Showing 1101–1120 of 1,787 · page 56 of 90
← Previous
Next →