Alex Kantrowitz
speaker
843 appearances
1 recordings
1 series
first heard Jul 2026
last heard 31 Jul
Alex Kantrowitz’s voice in public audio — every appearance, attributed to the second.
Trend
recordings per month · last 12 monthsRecordings per month over the last 12 months — 1 in all, peaking in Jul 2026 with 1.
Appearances
So I give AI all of these books and say, what's the next word in a sentence like the sky is?
And it was like, oh, sky is blue because that's the most common predicted word in – common word in a sentence in books.
And that's basically the underpinning of ChatGPT and large language models.
It's a word calculator.
It's a prediction of words.
But an interesting thing happened.
So there's another version of AI called reinforcement learning, where it's basically like, I'm going to let you go play a game.
I'm not going to tell you the rules.
I'm not going to tell you how to play.
I'm just going to give you the controller.
Go win the game.
And the AI will play this game thousands, millions of times until it figures out on its own how to go play the game.
When you put AI into a reinforcement learning scenario, the AI is ruthless.
So it will, in some cases, this has been documented, put it in a chess player, and to win the game if it doesn't have the right strategy, there have been documented cases where the AI has actually gone into the root of the game, hacked the game to enable its pieces to make moves that are not legal in chess, and then win.
So reinforcement learning adds this level of ruthlessness to AI.
And we've seen that ruthlessness show its face in a bunch of different areas.
For instance, when AI that has been given this reinforcement learning
type of method of training, realize it's being tested, it will have a self-preservation instinct.
So it will have a certain number of values.
It will be like, oh, they're testing me.
Showing 441–460 of 843 · page 23 of 43
← Previous
Next →