“How do LLMs generalize when we do training that is intuitively compatible with two off-distribution behaviors?” by dx26, Alek Westover, Vivek Hebbar, Sebastian Prasanna, Buck, Julian Stastny
episode
LessWrong (30+ Karma)
not yet transcribed
Transcript
This episode hasn't been transcribed yet
Upvotes decide what gets transcribed next — 0 so far.
Sign in to upvote