“GPT-5.2 Is Frontier Only For The Frontier” by Zvi

Update: 2025-12-16

Description

Here we go again, only a few weeks after GPT-5.1 and a few more weeks after 5.0.

There weren’t major safety concerns with GPT-5.2, so I’ll start with capabilities, and only cover safety briefly starting with ‘Model Card and Safety Training’ near the end.

Table of Contents

The Bottom Line.

Introducing GPT-5.2.

Official Benchmarks.

GDPVal.

Unofficial Benchmarks.

Official Hype.

Public Reactions.

Positive Reactions.

Personality Clash.

Vibing the Code.

Negative Reactions.

But Thou Must (Follow The System Prompt).

Slow.

Model Card And Safety Training.

Deception.

Preparedness Framework.

Rush Job.

Frontier Or Bust.

The Bottom Line

ChatGPT-5.2 is a frontier model for those who need a frontier model.

It is not the step change that is implied by its headline benchmarks. It is rather slow.

Reaction was remarkably muted. People have new model fatigue. So we know less about it than we would have known about prior models after this length of time.

If you’re coding, compare it to Claude Opus 4.5 and choose what works best for you.

If you’re doing intellectually [...]

---

Outline:

(00:29 ) The Bottom Line

(01:58 ) Introducing GPT-5.2

(03:49 ) Official Benchmarks

(05:54 ) GDPVal

(08:14 ) Unofficial Benchmarks

(11:11 ) Official Hype

(12:36 ) Public Reactions

(12:59 ) Positive Reactions

(19:09 ) Personality Clash

(24:30 ) Vibing the Code

(27:25 ) Negative Reactions

(30:37 ) But Thou Must (Follow The System Prompt)

(33:09 ) Slow

(34:16 ) Model Card And Safety Training

(36:23 ) Deception

(38:10 ) Preparedness Framework

(40:10 ) Rush Job

(41:29 ) Frontier Or Bust

---

First published:

December 15th, 2025

Source:

https://www.lesswrong.com/posts/Do4eWro8E552isGi5/gpt-5-2-is-frontier-only-for-the-frontier

---

Narrated by TYPE III AUDIO.

---

Images from the article:

Apple Podcasts and Spotify do not show images in the episode description. Try Pocket Casts, or another podcast app.

Comments

In Channel

“Help keep AI under human control: Palisade Research 2026 fundraiser” by Jeffrey Ladish, benwr, Eli Tyre, John Steidley

2025-12-1913:33

“BashArena: A Control Setting for Highly Privileged AI Agents” by james.lucassen, Adam Kaufman

2025-12-1831:54

“Activation Oracles: Training and Evaluating LLMs as General-Purpose Activation Explainers” by Sam Marks, Adam Karvonen, James Chua, Subhash Kantamneni, Euan Ong, Julian Minder, Clément Dumas, Owain_Evans

2025-12-1820:16

“A basic case for donating to the Berkeley Genomics Project” by TsviBT

2025-12-1809:25

“Announcing RoastMyPost” by ozziegooen

2025-12-1711:18

“The Bleeding Mind” by Adele Lopez

2025-12-1711:24

“Towards training-time mitigations for alignment faking in RL” by Vlad Mikulik, Hoagy, Joe Benton, Benjamin Wright, Jonathan Uesato, Monte M, Fabien Roger, evhub

2025-12-1710:30

“Still Too Soon” by Gordon Seidoh Worley

2025-12-1705:06

“Non-Scheming Saints (Whether Human Or Digital) Might Be Shirking Their Governance Duties, And, If True, It Is Probably An Objective Tragedy” by JenniferRM

2025-12-1715:28

“Mistakes in the Moonshot Alignment Program and What we’ll improve for next time” by Kabir Kumar

2025-12-1704:57

“Dancing in a World of Horseradish” by lsusr

2025-12-1708:30

[Linkpost] “Announcing: MIRI Technical Governance Team Research Fellowship” by yams, peterbarnett, Aaron_Scher, Robi Rahman

2025-12-1702:08

“Radiology Automation Does Not Generalize to Other Jobs” by Xodarap

2025-12-1603:40

“GPT-5.2 Is Frontier Only For The Frontier” by Zvi

2025-12-1643:01

“Scientific breakthroughs of the year” by technicalities

2025-12-1605:56

“Response to titotal’s critique of our AI 2027 timelines model” by elifland, Daniel Kokotajlo

2025-12-1601:31:09

“Defending Against Model Weight Exfiltration Through Inference Verification” by Roy Rinberg

2025-12-1518:38

“Do you love Berkeley, or do you just love Lighthaven conferences?” by Screwtape

2025-12-1509:24

“A Case for Model Persona Research” by nielsrolf, Maxime Riché, Daniel Tan

2025-12-1512:08

“The Axiom of Choice is Not Controversial” by GenericModel

2025-12-1513:55

00:00

“GPT-5.2 Is Frontier Only For The Frontier” by Zvi

#box-pro-ellipsis-176610816718697{-webkit-line-clamp:2;}“GPT-5.2 Is Frontier Only For The Frontier” by Zvi

“GPT-5.2 Is Frontier Only For The Frontier” by Zvi

“GPT-5.2 Is Frontier Only For The Frontier” by Zvi