“Finding Features in Neural Networks with the Empirical NTK” by jylin04

Update: 2025-10-17

Description

Audio note: this article contains 63 uses of latex notation, so the narration may be difficult to follow. There's a link to the original text in the episode description.

Summary

Kernel regression with the empirical neural tangent kernel (eNTK) gives a closed-form approximation to the function learned by a neural network in parts of the model space. We provide evidence that the eNTK can be used to find features in toy models for interpretability. We show that in Toy Models of Superposition and a MLP trained on modular arithmetic, the eNTK eigenspectrum exhibits sharp cliffs whose top eigenspaces align with the ground-truth features. Moreover, in the modular arithmetic experiment, the evolution of the eNTK spectrum can be used to track the grokking phase transition. These results suggest that eNTK analysis may provide a new practical handle for feature discovery and for detecting phase changes in small models.

[...]

---

Outline:

(00:23 ) Summary

(01:25 ) Background

(04:57 ) Results

(05:10 ) Toy models of Superposition

(06:34 ) Modular arithmetic

(10:00 ) Next steps

The original text contained 9 footnotes which were omitted from this narration.

---

First published:

October 16th, 2025

Source:

https://www.lesswrong.com/posts/cpFqDDjhvhbaoyHnd/finding-features-in-neural-networks-with-the-empirical-ntk-1

---

Narrated by TYPE III AUDIO.

---

Images from the article:

Apple Podcasts and Spotify do not show images in the episode description. Try Pocket Casts, or another podcast app.

Comments

In Channel

“Meditation is dangerous” by Algon

2025-10-1807:27

“I’m an EA who benefitted from rationality” by juliawise

2025-10-1704:36

“Finding Features in Neural Networks with the Empirical NTK” by jylin04

2025-10-1711:07

“Reducing risk from scheming by studying trained-in scheming behavior” by ryan_greenblatt

2025-10-1720:08

“Rogue internal deployments via external APIs” by Fabien Roger, Buck

2025-10-1611:39

“Cheap Labour Everywhere” by Morpheus

2025-10-1603:39

“[draft amnesty] A New Global Risk: Large Comet’s Impact on Sun Could Cause Fires on Earth” by avturchin

2025-10-1603:40

“How I Became a 5x Engineer with Claude Code” by Gordon Seidoh Worley

2025-10-1511:56

“That Mad Olympiad” by Tomás B.

2025-10-1526:42

“It will cost you nothing to ‘bribe’ a Utilitarian” by Gabriel Alfour

2025-10-1509:15

“The Biochemical Beauty of Retatrutide: How GLP-1s Actually Work” by Elizabeth

2025-10-1514:45

“Recontextualization Mitigates Specification Gaming Without Modifying the Specification” by vgillioz, TurnTrout, cloud, ariana_azarbal

2025-10-1421:17

“The ‘Length’ of ‘Horizons’” by Adam Scholl

2025-10-1414:16

“Current Language Models Struggle to Reason in Ciphered Language” by Fabien Roger

2025-10-1413:46

“The Mom Test for AI Extinction Scenarios” by Taylor G. Lunt

2025-10-1409:29

“How AI Manipulates—A Case Study” by Adele Lopez

2025-10-1432:49

“If Anyone Builds It Everyone Dies, a semi-outsider review” by dvd

2025-10-1426:02

“Making legible that many experts think we are not on track for a good future, barring some international cooperation” by Mateusz Bagiński, Ishual

2025-10-1324:15

“OpenAI #15: More on OpenAI’s Paranoid Lawfare Against Advocates of SB 53” by Zvi

2025-10-1345:53

“Sublinear Utility in Population and other Uncommon Utilitarianism” by Alice Blair

2025-10-1312:57

00:00

“Finding Features in Neural Networks with the Empirical NTK” by jylin04

#box-pro-ellipsis-176086482349260{-webkit-line-clamp:2;}“Finding Features in Neural Networks with the Empirical NTK” by jylin04

“Finding Features in Neural Networks with the Empirical NTK” by jylin04

“Finding Features in Neural Networks with the Empirical NTK” by jylin04