Dwarkesh
speaker
1,722 appearances
5 recordings
1 series
first heard Apr 2025
last heard 25 Nov
Dwarkesh’s voice in public audio — every appearance, attributed to the second.
Trend
recordings per month · last 12 monthsRecordings per month over the last 12 months — 2 in all, peaking in Nov 2025 with 1.
Appearances
And everyone's just like, oh, yeah, AIs are like that.
What does your shirt say?
And I mean, I don't disagree with this.
I'm also in this position.
I see the AI is lying and it's obviously just like an artifact of the training process.
It's not anything sinister.
But I think this is just going to keep happening where no matter what evidence we get, people are going to think, oh yeah, that's not the AI turns evil thing that people have worried about.
That's not the Terminator scenario.
That's just one of these natural consequences of how we train it.
And I think that once a thousand of these natural consequences of training add up, the AI is evil in the same way that like once the AI can do chess and philosophy and all these other things, eventually you got to admit it's intelligent.
Yeah.
So I think that each individual failure, like maybe it will make the national news.
Maybe people say, oh, it's so strange that GPT-7 did this particular thing and then they'll train it away and then it won't do that thing.
And there will be some point at the process of becoming super intelligent at which it
I don't want to say makes the last mistake because you'll probably have like gradually decreasing number of mistakes to some asymptote, but the last mistake that anyone worries about.
And after that, it will be able to do its own thing.
Yeah, I think the alignment community did not really expect LLMs.
I mean, if you look in Bostrom Superintelligence, there's a discussion of Oracle AIs, which are sort of like LLMs.
I think that came as a surprise.
I think one of the reasons I'm more hopeful than I used to be is that LLMs are great for
Showing 1201–1220 of 1,722 · page 61 of 87
← Previous
Next →