Mike Knoop
speaker
927 appearances
4 recordings
1 series
first heard Mar 2025
last heard 18 Nov
Mike Knoop’s voice in public audio — every appearance, attributed to the second.
Trend
recordings per month · last 12 monthsRecordings per month over the last 12 months — 1 in all, peaking in Nov 2025 with 1.
Appearances
so working on full details of how that's gonna look next year.
Um but uh yeah, we're we're sort of like in the throes of it.
We're definitely using some of these frontier systems to do red teaming against the benchmark, just to s you know, assert that like, yeah, these games are still hard for AI and we're still finding that to be the case even with things like Gemini three.
Um but uh but yeah, that's uh we're still in progress with development right now.
If anyone at Google is uh listening to this and could sort of give me access to CEMA two, I would love to test it on V three.
This is actually something that uh that's that we haven't done yet in a ball too
We haven't done yet in a volume
You you read the marketing material and it's like, okay, that seems like it should solve V3 before it exists.
So like if that's the case, we should know that.
Uh and so, but yeah, I haven't got haven't gone hands-on with it yet.
So I can't sort of make any the statement either way on the claims.
that this is one true you you've mentioned something really something true about V three, which is that it's still a relatively short time horizon tasks and they're self-contained.
It does add some new complexity where you have to deal with interactivity because you have to do goal acquisition,
have to do exploration.
We'll have a really nice action efficiency comparison between humans and AI, which we we haven't been able to get before on the V one, V two domain.
So we're gonna get a lot of new signal, I think, on V3.
Um, but yeah, I think as you sort of look even further out into the future, things that are more open ended are the things I think we're starting to get excited about trying to like understand.
Like what does it mean to put one of these AI systems in an open-end environment and then look back on the system?
you know, 10 minutes into the future, a hundred minutes in the future, a thousand minutes in the future.
And can you look at the environment that that AI system has been like how it's manipulated its environment and like, you know, say something interesting about how intelligent the system is based on that like observation and open ended sense.
Showing 121–140 of 927 · page 7 of 47
← Previous
Next →