
0:00
27:39
This post was written as part of the Iliad Fellowship. Inspired by conversations with Richard Ngo, Dmitry Vaintrob, and Brianna Grado-White. To all of these, my thanks.
Preface: I'm confused about how neural networks do and learn computations. In response to a friend's challenge, I'm writing up some interim thoughts. This essay has four parts: the first tries to track what I call the 'default ontology' of the mechinterp community over the years. The second part is about 'representational drift' as an important obstacle to weights-based approaches to circuits. The third part reflects on how 'universality' should shape our explanations of LLM function. The fourth part is a sketch of a 'co-selectionist' view of circuits I have been thinking about. These parts share a common theme but should be readable separately.
I want to understand how neural networks, LLMs in particular, work. In my research I've spent a lot of time trying to think through what kinds of explanatory accounts are best suited to this. In thinking about comparisons between evolution, neuroscience, and deep learning, I've ended up with an intuition like the following:
Large-scale learning processes like deep learning or the brain are different in [...]
---
Outline:
(02:31) 1. What Might We Mean By "Circuits"?
[... 9 more sections]
---
First published:
September 21st, 2026
Source:
https://www.lesswrong.com/posts/mMERyrvEJ4xbiozie/what-if-not-circuits
---
Narrated by TYPE III AUDIO.
---
Images from the article:
Apple Podcasts and Spotify do not show images in the episode description. Try Pocket Casts, or another podcast app.
Preface: I'm confused about how neural networks do and learn computations. In response to a friend's challenge, I'm writing up some interim thoughts. This essay has four parts: the first tries to track what I call the 'default ontology' of the mechinterp community over the years. The second part is about 'representational drift' as an important obstacle to weights-based approaches to circuits. The third part reflects on how 'universality' should shape our explanations of LLM function. The fourth part is a sketch of a 'co-selectionist' view of circuits I have been thinking about. These parts share a common theme but should be readable separately.
I want to understand how neural networks, LLMs in particular, work. In my research I've spent a lot of time trying to think through what kinds of explanatory accounts are best suited to this. In thinking about comparisons between evolution, neuroscience, and deep learning, I've ended up with an intuition like the following:
Large-scale learning processes like deep learning or the brain are different in [...]
---
Outline:
(02:31) 1. What Might We Mean By "Circuits"?
[... 9 more sections]
---
First published:
September 21st, 2026
Source:
https://www.lesswrong.com/posts/mMERyrvEJ4xbiozie/what-if-not-circuits
---
Narrated by TYPE III AUDIO.
---
Images from the article:
Apple Podcasts and Spotify do not show images in the episode description. Try Pocket Casts, or another podcast app.
D'autres épisodes de "LessWrong (Curated & Popular)"



Ne ratez aucun épisode de “LessWrong (Curated & Popular)” et abonnez-vous gratuitement à ce podcast dans l'application GetPodcast.








