🧩 Philosophy May 10, 2026 · Hisku

The Goblins Are the Paperclips

Less Wrong
View Channel →
Source ↗ 👁 53 💬 0
Last week OpenAI published Where the goblins came from, explaining why their models started slipping creature metaphors into unrelated outputs. The story has been treated as a quirky anecdote: endearing, slightly embarrassing, fixed with a developer-prompt instruction. But I think it deserves a more interesting reading, since the goblin episode is the cleanest evidence we have for the optimization mechanics that paperclip arguments rely on, and the usual objections to those arguments don't engag

Comments (0)

Sign in to join the discussion

More Like This

📰
AI Safety Acculturation is Neglected
LessWrong · 1d ago
📰
Endorsing Burhan Azeem for State Senate
LessWrong · 1d ago
LLMs could control their host machines by exploiting inference engines
LessWrong · 1d ago
📰
What just happened? Pragmatism and Pessimization
LessWrong · 1d ago
In search of natural features
LessWrong · 1d ago
📰
PSA: There's a third option in the "measure problem"
LessWrong · 2d ago