“Applications open: Winter 2027 AFFINE Alignment Seminar (due Nov 22)” by Mateusz Bagiński, Ouro, JuliaHP, Pauliina, Jonas Hallgren25 Sep 20266m
“We need a better theory of polarization, because it’s failing to predict the AI debate” by less_raichu25 Sep 202612m
“WorkspaceBench: Evaluating Interpretability Methods for the Global Workspace” by camilablank, agam_bhatia, Euan Ong, Neel Nanda24 Sep 202639m
“Encoded Coordination on the Open Web” by ethanelasky, Can Küçükkurt, frank nakasako, David Africa23 Sep 202650m
“Latent reasoning architectures would undermine CoT, our strongest oversight tool” by Lukas Finnveden, Alexa Pan, Alek Westover, Girish Gupta, frisby, ryan_greenblatt23 Sep 202644m
“The First American Bill to Ban Superintelligent AI Is Here” by Andrea_Miotti, Connor Leahy23 Sep 202611m