06 May 2024
2m
[Linkpost] “Uncovering Deceptive Tendencies in Language Models: A Simulated Company AI Assistant” by Olli Järviniemi, evhub
LessWrong (30+ Karma)
Open in Podwise to generate AI notes
Sign in to process this episode and unlock summaries, transcripts, highlights and translations.
Shownotes are not generated by Podwise.
