Yoshua Bengio
speaker
2,216 appearances
5 recordings
5 series
first heard Oct 2018
last heard 7 May
Yoshua Bengio’s voice in public audio — every appearance, attributed to the second.
Trend
recordings per month · last 12 monthsRecordings per month over the last 12 months — 3 in all, peaking in May 2026 with 1.
Appearances
So you're always like pushed away from bad behavior.
And with some properties of how you train the system, like injecting noise into the training procedure in the stochastic gradient descent, you can get those guarantees.
That's right.
I mean, I think for the guardrail, you don't need as much compute because it's more specialized to predicting harm.
But when we get to the agentic scientist AI, for sure it has to be trained with similar compute and size of neural net, probably as the state of the art, which means my little nonprofit wouldn't be able to do that.
And there will be a need for either companies to take on this or governments or philanthropy to fund at a scale that we can do that.
But in order to convince all of these parties, we need to show on a small scale, for example, using fine tuning or using smaller models, that we do get these improvements in honesty and for the same size models that we don't lose in capability, for example.
It's mostly the mathematical work I've been doing in the last eight months, approximately, to go from the high level intuitions that I've had now for almost two years about how we could build a scientist AI
into something much more formal and much more precise about the conditions that are sufficient, maybe not even necessary, but sufficient at a mathematical level to get the kind of guarantees of vanishing small probability that something bad will happen.
And when I say something bad, I need to be a little bit more precise here.
This is not a guarantee that the AI won't be used for something bad by bad people.
it's a guarantee that the AI won't do something bad of its own accord, right?
Because of implicit goals or uncontrolled goals.
Besides loss of control, the other catastrophic possibility is humans using AI to construct a, you know, eventually worldwide dictatorship.
A small group of humans could concentrate all the power that AI will have, especially if we achieve AGI or super intelligence.
And it would be much harder to get rid of that kind of authoritarian power than what we've seen with fascism and what happened in the USSR.
because they didn't have this technology that is becoming more and more feasible of surveillance and even shaping public opinion.
So AI is becoming really good at persuasion and there are studies showing that progress, if I can call it this way in that direction,
The people who control these systems will be able to shape public opinion, to detect and kill off their opponents, to develop weapons that can destroy the countries that disagree with them.
And that is why I'm spending a large part of my time in explaining the issues more broadly of the risks that powerful AI brings, including the power concentration
Showing 201–220 of 2,216 · page 11 of 111
← Previous
Next →