🧩 Philosophy 12h ago · Abhimanyu Pallavi Sudhir

Reward is hyperstitional information

Less Wrong
View Channel →
Source ↗ 👁 0 💬 0
(This article broadly explains mirror ascent, continuous Bayesian inference and information geometry in full. Title refers to the result in section 3.)
The logarithmic scoring rule


mjx-container[jax="CHTML"] {
line-height: 0;
}

mjx-container [space="1"] {
margin-left: .111em;
}

mjx-container [space="2"] {
margin-left: .167em;
}

mjx-container [space="3"] {
margin-left: .222em;
}

mjx-container [space="4"] {
margin-left: .278em;
}

mjx-container [space="5"] {
margin-left: .333em;

Comments (0)

Sign in to join the discussion

More Like This

Model Organisms of Sandbagging in the Wild
LessWrong · 2h ago
Function vectors as a model diffing tool: 17 heads repair a bad fine-tune
LessWrong · 5h ago
How to define P(doom) and why it matters
LessWrong · 5h ago
AI #180: No Longer In Charge
LessWrong · 6h ago
The Open Problems of the AI Alignment Field and their Cruxes
LessWrong · 7h ago
Matryoshka NLAs: training activation verbalizers to frontload reconstruction-relevant information
LessWrong · 10h ago