Gideon Lewis-Kraus

speaker
557 appearances 2 recordings 2 series first heard Feb 2026 last heard 27 Feb

Gideon Lewis-Kraus’s voice in public audio — every appearance, attributed to the second.

Trend

recordings per month · last 12 months
2 · Feb OctJan 26AprJulnow

Recordings per month over the last 12 months — 2 in all, peaking in Feb 2026 with 2.

Appearances

newest first · ▶ plays the moment
anxiety and associations with performance.
You know, when you kind of looked inside, you could see that some part of it was making associations with a sort of playful, performative exchange, which is to say that it seems like Claude recognized that it was participating in a game.
Well, one doesn't have to go quite so far as to say that it's conscious of itself.
As to suggest, you know, one of the ways to look at this is that what these things are very good at are recognizing the genre that they are in and picking up on all of these small linguistic context clues that suggest like, oh, you know, this is not actually like a serious academic discussion of quantum mechanics.
That like what is happening here is,
a playful exchange between people where one person is like kind of hiding something but winking that they're not really hiding it and that like that's the genre in which it is operating.
So it doesn't have to be conscious in order to do that.
It just has to be a very good reader and replicator of genre conventions.
I mean, this is a great question.
And this is where one kind of runs up against the limits of what can be known and what can be said at this point.
I mean, he was basically saying, you know, look, I understand what's going on in here, that this is just a lot of matrix multiplication, that these are
tens of thousands of tiny numbers being multiplied together, that there's nothing really spooky happening here, that there's no ghost in the machine.
But what he was saying was, with models up to a certain point, he was able, using kind of a similar tool to the one Josh Batson used, instead of looking at what the model was, so to speak, thinking, he could incept an idea into the model.
He could say, right at this point where you are having an association with the Eiffel Tower, we're going to
put in an association with cheese and see what happens.
And so then the model would respond by saying something about cheese.
And he would say something similar to what Batson said, which was like, why did you add that thing about cheese that I didn't ask about?
And the model would basically just look back at the entire conversation that they had been having and then try to kind of retcon an explanation.
But what Jack has found more recently is that when he incepts these ideas into the model, instead of the model purely looking at its own external behavior to try to figure out why it had done something,
that actually these models could very dimly perceive that something strange had gone on internally, that someone was monkeying with, you know, the neurons inside the model to make it do something different.
Showing 421–440 of 557 · page 22 of 28 ← Previous Next →