“How AI Is Learning to Think in Secret” by Nicholas Andresen

Update: 2026-01-06

Description

On Thinkish, Neuralese, and the End of Readable Reasoning

In September 2025, researchers published the internal monologue of OpenAI's GPT-o3 as it decided to lie about scientific data. This is what it thought:

Pardon? This looks like someone had a stroke during a meeting they didn’t want to be in, but their hand kept taking notes.

That transcript comes from a recent paper published by researchers at Apollo Research and OpenAI on catching AI systems scheming. To understand what's happening here - and why one of the most sophisticated AI systems in the world is babbling about “synergy customizing illusions” - it first helps to know how we ended up being able to read AI thinking in the first place.

That story starts, of all places, on 4chan.

In late 2020, anonymous posters on 4chan started describing a prompting trick that would change the course of AI development. It was almost embarrassingly simple: instead of just asking GPT-3 for an answer, ask it instead to show its work before giving its final answer.

Suddenly, it started solving math problems that had stumped it moments before.

To see why, try multiplying 8,734 × 6,892 in your head. If you’re like [...]

The original text contained 3 footnotes which were omitted from this narration.

---

First published:

January 6th, 2026

Source:

https://www.lesswrong.com/posts/gpyqWzWYADWmLYLeX/how-ai-is-learning-to-think-in-secret

---

Narrated by TYPE III AUDIO.

---

Images from the article:

Apple Podcasts and Spotify do not show images in the episode description. Try Pocket Casts, or another podcast app.

Comments

In Channel

“Public intellectuals need to say what they actually believe” by Aaron Bergman

2026-01-0832:19

“Broadening the training set should help with alignment” by Seth Herd

2026-01-0716:51

“Mainstream approach for alignment evals is a dead end” by Igor Ivanov

2026-01-0710:34

“How hard is it to inoculate against misalignment generalization?” by Jozdien

2026-01-0712:16

“The Evolution Argument Sucks” by peralice

2026-01-0614:33

“How AI Is Learning to Think in Secret” by Nicholas Andresen

2026-01-0637:45

“On Owning Galaxies” by Simon Lermen

2026-01-0605:38

“Exploring Reinforcement Learning Effects on Chain-of-Thought Legibility” by Julian H, RohanS, Baram Sosis, vedant-badoni, The-Turtle

2026-01-0623:02

“Oversight Assistants: Turning Compute into Understanding” by jsteinhardt

2026-01-0616:25

“Axiological Stopsigns” by JenniferRM

2026-01-0626:49

“Claude Wrote Me a 400-Commit RSS Reader App” by Brendan Long

2026-01-0605:11

[Linkpost] “The Thinking Machine” by PeterMcCluskey

2026-01-0605:04

“The economy is a graph, not a pipeline” by anithite

2026-01-0609:09

“Dos Capital” by Zvi

2026-01-0631:56

“The Technology of Liberalism” by L Rudolf L

2026-01-0555:14

“AI Risk timelines: 10% chance (by year X) should be the headline (and deadline), not 50%. And 10% is _this year_!” by Greg C

2026-01-0501:44

“The inaugural Redwood Research podcast” by Buck, ryan_greenblatt

2026-01-0503:28

“In My Misanthropy Era” by jenn

2026-01-0413:52

“AI #149: 3” by Zvi

2026-01-0445:28

“The bio-pirate’s guide to GLP-1 agonists” by quiet_NaN

2026-01-0309:42

00:00

“How AI Is Learning to Think in Secret” by Nicholas Andresen

#box-pro-ellipsis-176784532880772{-webkit-line-clamp:2;}“How AI Is Learning to Think in Secret” by Nicholas Andresen

“How AI Is Learning to Think in Secret” by Nicholas Andresen

“How AI Is Learning to Think in Secret” by Nicholas Andresen