Boston Computation Club podcast

06/26/26: Tracing Introspection Across Model Depth, Zach Maas

0:00
48:13
Manda indietro di 15 secondi
Manda avanti di 15 secondi

Zach Maas is an independent AI safety & mechanistic interpretability researcher in Boulder, Colorado, funded by Coefficient Giving. Today Zach joined us to talk about some of his recent work tracing introspection across model depth. This is, I think, the first mech interp talk we've hosted other than ChessGPT, and it was a good one! We hope you enjoy it as much as we did!

Altri episodi di "Boston Computation Club"