Cover image

Your Reward Model Is Also a Voting Rule: What Social Choice Changes About RLHF

TL;DR for operators When thousands of annotators disagree about two acceptable model responses, the pipeline still has to decide how that disagreement becomes model behavior. That decision is often hidden inside familiar technical choices: how comparisons are sampled, whether votes are collapsed into majority labels, which reward-model class is fitted, and how aggressively a policy is optimized against the resulting score. ...

September 16, 2026 · 8 min · Zelina
Cover image

When Democracy Meets the Algorithm: Auditing Representation in the Age of LLMs

When Democracy Meets the Algorithm: Auditing Representation in the Age of LLMs Agenda-setting is where participation quietly becomes power. Anyone can invite hundreds of people to submit questions. That part is now cheap. The difficult part arrives ten minutes later, when an expert panel has time to answer only seven of them, and someone has to decide which seven. This is the small administrative hinge on which democratic legitimacy loves to swing. A moderator chooses. A platform ranks. An LLM summarises. Everyone else is told, usually with a straight face, that their concerns were “reflected”. ...

November 7, 2025 · 16 min · Zelina