Noam Brown
speaker
1,199 appearances
1 recordings
1 series
first heard Dec 2022
last heard Dec 2022
Noam Brown’s voice in public audio — every appearance, attributed to the second.
Trend
recordings per month · last 12 monthsNo recordings in the last 12 months.Older appearances are listed below; set an alert to hear about the next one.
Appearances
We're not 100% there, but we're getting closer at least.
Yeah, absolutely.
We've already started looking into this direction a bit.
So we tried to use the techniques that we've developed for diplomacy to make chess and go AIs.
And what we found is that it led to much more human-like strong chess and go players.
The way that...
AIs like Stockfish today play is in a very inhuman style.
It's very strong, but it's very different from how humans play.
And so we can take the techniques that we've developed for diplomacy.
We do something similar in chess and go, and we end up with a bot that's both strong and human-like.
But to elaborate on this a bit, one way to approach making a human-like AI for chess is to collect a bunch of human games, like a bunch of human grandmaster games, and just to supervise learning on those games.
But the problem is that if you do that, what you end up with is an AI that's substantially weaker than the human grandmasters that you've trained on.
because the neural net is not able to approximate the nuance of the strategy.
This goes back to the planning thing that I mentioned, the search thing that I talked about before, that these human grandmasters, when they're playing, they're using search and they're using planning.
And the neural net alone, unless you have a massive neural net that's like a thousand times bigger than what we have right now, it's not able to approximate those details very effectively.
And on the other hand, you can leverage search and planning very heavily, but then what you end up with is an AI that plays in a very different style from how humans play the game.
Now, if you strike this intermediate balance by setting the regularization parameters correctly and say, you can do planning, but try to keep it close to the human policy, then you end up with an AI that plays in both a very human-like style and a very strong style.
And you can actually even tune it to have a certain ELO rating.
So you can say, play in the style of like a 2800 ELO human.
Yeah, I think so.
Showing 1081–1100 of 1,199 · page 55 of 60
← Previous
Next →