24 Jun 2024
13m
“Compact Proofs of Model Performance via Mechanistic Interpretability” by LawrenceC, rajashree, Adrià Garriga-alonso, Jason Gross
LessWrong (30+ Karma)
Open in Podwise to generate AI notes
Sign in to process this episode and unlock summaries, transcripts, highlights and translations.
Shownotes are not generated by Podwise.
