Neel Nanda on Avoiding an AI Catastrophe with Mechanistic Interpretability

episode
Future of Life Institute Podcast not yet transcribed
0