Less Wrong

@less-wrong 🧩 Philosophy
📰 671 articles 🔄 Updated Aug 17, 2026 lesswrong.com

Latest Articles

Untie Squared ReLU variant
1 IntroMy collaborator Michael Bukatin came up with an idea: in the hidden projection of MLP, instead of using ReLU, he
LessWrong · Aug 17, 2026 Philosophy
0 8
Study Update: Does post-training quantization change welfare-relevant indicators in open-weight language models?
This content will not make sense without reading the preregistration found in the original postWe ask whether welfare-re
LessWrong · Aug 16, 2026 Philosophy
0 11
Q2.5 2026 Timelines Update: Uplift and Revenue
Tl;dr: Our timelines haven’t changed much (they got slightly shorter) but our modeling and evidence base have noticeably
LessWrong · Aug 16, 2026 Philosophy
0 9
Case for Funding AI Safety in Japan
Intro and tl;drI'm Esa Koskinen, working as a volunteer director of AI Safety Tokyo.Epistemic status: estimates by an in
LessWrong · Aug 16, 2026 Philosophy
0 8
Will There Be an AI Hegemon? A Mental Model for AI Power Concentration
On the Baker-Anthropic Conversation, Power Acquisition and Escape VelocityWill there be an AI hegemon, i.e. an entity th
LessWrong · Aug 16, 2026 Philosophy
0 9
The Doomsday Argument is Reasonable and Mostly Points to Longevity
The Doomsday argument was proposed in a 1983 lecture by Brandon Carter and elaborated on in John Leslie's 1996 book The
LessWrong · Aug 16, 2026 Philosophy
0 10
Three thoughts on civilisational handoff
What happens when humans put AIs in charge of civilisationally important decisions? A frontier AI company might hand ove
LessWrong · Aug 16, 2026 Philosophy
0 9
Should Less Wrong add subtitles?
If Less Wrong wants people to be sharing more of their intellectual output on this website, we should probably be lookin
LessWrong · Aug 16, 2026 Philosophy
0 3
Are questions allowed on LessWrong?
Sometimes I want to post things like:[thing i'm wondering about][here’s my initial stab at it][but i have no idea if thi
LessWrong · Aug 16, 2026 Philosophy
0 4
Does DiffusionGemma do latent reasoning?
TL;DR Google DeepMind's recent model DiffusionGemma (DG) generates text via diffusion, meaning many diffusion steps happ
LessWrong · Aug 16, 2026 Philosophy
0 3
Mom's Advice For Hosting A Class Reunion
Pour more money and effort into them than you think is reasonable. Treasure them, because you can't actually host that m
LessWrong · Aug 15, 2026 Philosophy
0 7
I'm starting a interview series of people working in Lean / formal methods / math formalization
I think the topics of discussion would be of interest to a lot of people here, so I thought I'd share the first episode:
LessWrong · Aug 15, 2026 Philosophy
0 5
Nuclear physics of Alex Zhao's comment for "Pacing the Frontier"
Very recently, the "Pacing the Frontier" petition was published. I want to focus on the comment from Alex Zhao, research
LessWrong · Aug 15, 2026 Philosophy
0 3
Learning new facts can change LLM behaviour
TL:DR: I use synthetic document fine-tuning to train an LLM to believe that in 2027 ‘long-horizon’ frontier LLMs count a
LessWrong · Aug 15, 2026 Philosophy
0 3
All Utilitarians Should Be Classical Utilitarians
This is a crosspost from my blog post.I recently had the great joy of meeting a group of utilitarians, but, to my comple
LessWrong · Aug 15, 2026 Philosophy
0 4
On Dwarkesh Patel’s Podcast With Ryan Greenblatt
Some podcasts are self-recommending enough that I look to break them down if I have the chance. This, as a debate about
LessWrong · Aug 15, 2026 Philosophy
0 3
Rerunning AI safety papers on every frontier release would be pretty easy and valuable
tl;dr: Some important AI safety research is never rerun on the newest models. There are probably cases where this would
LessWrong · Aug 15, 2026 Philosophy
0 5
Metaphilosophy II: Empirical Flywheels
1.4 Two philosophical methods1.4.1 Philosophy consists of updating the highest-level concepts of the mind. As discussed
LessWrong · Aug 15, 2026 Philosophy
0 5
Red vs Blue, but for Evals
🔵 The blue team proposes an evaluation protocol for some capability/propensity of interest. This consists of a suite of
LessWrong · Aug 15, 2026 Philosophy
0 4
Toy Model of Activation Obfuscation
I completed this work as part of the BlueDot Impact Technical AI Safety Project. This linkpost is a somewhat condensed v
LessWrong · Aug 14, 2026 Philosophy
0 5