Boston Computation Club podcast

06/26/26: Tracing Introspection Across Model Depth, Zach Maas

0:00
48:13
Spola tillbaka 15 sekunder
Spola framåt 15 sekunder

Zach Maas is an independent AI safety & mechanistic interpretability researcher in Boulder, Colorado, funded by Coefficient Giving. Today Zach joined us to talk about some of his recent work tracing introspection across model depth. This is, I think, the first mech interp talk we've hosted other than ChessGPT, and it was a good one! We hope you enjoy it as much as we did!

Fler avsnitt från "Boston Computation Club"