Noam Brown
speaker
1,199 appearances
1 recordings
1 series
first heard Dec 2022
last heard Dec 2022
Noam Brown’s voice in public audio — every appearance, attributed to the second.
Trend
recordings per month · last 12 monthsNo recordings in the last 12 months.Older appearances are listed below; set an alert to hear about the next one.
Appearances
Now, if you're playing a game like chess, the idea that you're going to search always to the end of the game is kind of unimaginable, right?
There's just so many situations where you just won't be able to use search in that case, or the cost would be prohibitive.
And
This technique allowed us to leverage search and without having to pay such a huge computational cost for it and be able to apply it more broadly.
So we actually did not use neural nets at all for Libratus or Pluribus.
And a lot of people found this surprising back in 2017.
I think they find it surprising today that we were able to do this without using any neural nets.
And I think the reason for that, I mean, I think neural nets are incredibly powerful and the techniques that are used today, even for poker AIs do rely quite heavily on neural nets, but it wasn't the main challenge for poker.
Like I think what neural nets are really good for is,
If you're in a situation where finding features for a value function is really difficult, then neural nets are really powerful.
And this was the problem in Go.
The problem in Go was that, or the final problem in Go at least, was that nobody had a good way of looking at a board and figuring out who was winning or losing, and describing through a simple algorithm who was winning or losing.
And so there neural nets were super helpful because you could just feed in a ton of different board positions into this neural net and it would be able to predict then who was winning or losing.
But in poker, the features weren't the challenge.
The challenge was how do you design a scalable algorithm that would allow you to find this balanced strategy that would understand that you have to bluff with the right probability.
Yeah, so the way the value functions work in poker, like the latest and greatest poker AIs, they do use neural nets for the value function.
The way it's done is very different from how it's done in a game like chess or go, because in poker, you have to reason about beliefs.
And so the value of a state depends on the beliefs that players have about what the different cards are.
Like if you have pocket aces,
then whether that's a really, really good hand or just an okay hand depends on whether you know I have pocket aces.
Showing 601–620 of 1,199 · page 31 of 60
← Previous
Next →