Judd Rosenblatt
speaker
589 appearances
1 recordings
1 series
first heard Oct 2024
last heard Oct 2024
Judd Rosenblatt’s voice in public audio — every appearance, attributed to the second.
Trend
recordings per month · last 12 monthsNo recordings in the last 12 months.Older appearances are listed below; set an alert to hear about the next one.
Appearances
What is the real problem at the end of the day, even if you do stuff which we're excited about, like approvably safe AI, guaranteed safe AI, that's awesome and really great and we want to support that.
The the ultimate problem you need to solve is how do you make it such that when AI is smarter and more capable than us, how do you make it not kill us at that point?
And if
You create AI that has a negative alignment tax.
It's able to beat the AI without the negative alignment tax.
Interestingly, a lot of the work which labs have done so far does seem to have a negative alignment tax.
The theory of Mirias that, of course, it only has like that it has that and everything's great until all of a sudden it kills you.
I think that's a a valid objection potentially here too.
But even still, what we would like to do is find more negligible.
approaches which do in fact have this negative alignment tax.
So far we've looked out by finding a couple that that do seem to have them.
And I think that people tend to underestimate the possibility that you could
build something like this.
I think there are quite a lot of other neuro inspired approaches potentially that would also have negative alignment taxes and substantially move things forward as well.
Yeah, I do just want to emphasize that like Mike talked about, it is extremely important to do further consciousness research in the first place.
And that itself is a neglected thing to better understand what's up with consciousness.
It's worth doing consciousness research with AI, and it's much better to do it on a smaller scale rather than have it emerge at a larger scale where there's greater moral patienthood concern and uh greater X risk.
Yeah, our Let's Rung Post does a pretty good job going through our thinking about this and a lot of different possible neglected approaches that we do advise and are most interested in, specifically with biological inspired things.
We think there's a lot more work that can be done in terms of reverse engineering pro sociality, and that may be associated with negative alignment taxes too.
So that's something we're quite interested in and intend to do a lot of more work on.
Showing 541–560 of 589 · page 28 of 30
← Previous
Next →