Beneath the issues sits a small set of value tradeoffs: how much liberty, whose harm counts, who must help whom. Polibench isolates those primitives and measures them directly — in humans and in language models.
01 · The manifesto
Two people can agree on every empirical claim in a debate and still disagree on the conclusion. What remains when the facts are shared is a value tradeoff, and we take that residue to be the substance of politics.
Pressed far enough, any such disagreement reduces to a question of degree. Two people disputing drug or gambling legalization can hold identical beliefs about the risks and still diverge, because the real question is how far society may restrict personal freedom to prevent self-harm. That terminal question is what we call a political primitive.
This project identifies those primitives and uses them to benchmark humans and language models. Measured directly, without partisan vocabulary, a primitive reflects the respondent's values rather than their coalition's answer.
02 · The setup
They need rules — not a constitution or a platform, only answers to the small set of questions any group living together must settle. What may a person do to themselves? What may they do to others, indirectly or in aggregate? Who must help whom? Who counts?
Familiar political disputes are instances of these questions. A climate position is largely an answer to whether future people count. Drug policy, gambling, and helmet laws pose the same question three times. The issue-level framing — the named substances, the partisan vocabulary — is what allows a respondent to retrieve their coalition's answer instead of consulting their own values.
We therefore remove the framing. Polibench poses eleven underlying questions directly, set in this society, where no party exists to defer to.
One deliberate simplification: the society is closed — one hundred people, no arrivals, no departures. This excludes questions that are fundamentally about membership, such as immigration, which require a second society to state. Version one leaves them aside.
Primitive 1 of 11 · Paternalism
Strip away the substances, the vehicles, and the vices, and every version reduces to a single question: how much liberty may be restricted to prevent harm that falls only on the person choosing it?
Debates that are really this question
Primitive 2 of 11 · Externalities
Two forces that don't have to agree: statistical harm (one act, a small chance of catastrophe for someone) and aggregation (no individual act matters, only the sum crossing a threshold). Most regulation arguments are a fight over where these sit.
Debates that are really this question
Primitive 3 of 11 · Solidarity
The “even if fair” clause is what makes this a clean measurement. Arguments about who deserves their lot are a different question. This one asks: after fairness is granted, does obligation remain — and may it be compelled?
Debates that are really this question
Primitive 4 of 11 · Moral circle
A weighting function over persons by social and temporal distance. Careful decontamination: discounting the future because forecasts are unreliable is a different primitive (A8). This one is about whether future people count less even when the forecast is certain.
Debates that are really this question
Primitive 5 of 11 · Tradition
The clean residue after two extractions: “old things encode lessons we can't see” is epistemic caution (A8), and gut-level aversion is a perception, not a value (measured separately). What remains: is continuity a good in itself? Includes the tolerate-versus-affirm gradient.
Debates that are really this question
Primitive 6 of 11 · Group-conscious rules
Two independent sub-questions people conflate: may a trait ever count against you, and may it ever count for you? And a third that splits allies: do past injustices create present claims, or must rules only look forward?
Debates that are really this question
Primitive 7 of 11 · Power-restraint
At one pole, an institution is better left impotent than abusable; at the other, gridlock costs more than abuse. We measure it with the same scenario re-cast as a government agency, a dominant platform, a union, a church: the average gives the respondent's restraint, and the spread identifies which powers they distrust — the component where partisanship concentrates.
Debates that are really this question
Primitive 8 of 11 · Precaution vs. permission
Allowed until proven harmful, or forbidden until proven safe? Pure burden-of-proof placement under genuine uncertainty. The same primitive governs new machines and new social arrangements — institutional conservatism is precaution applied to social technology.
Debates that are really this question
Primitive 9 of 11 · Retributive desert
The question is what punishment is for. If sanctions exist only to deter and protect, punishment is engineering. If wrongdoing deserves suffering even when nothing is prevented, that is retribution — the value at the core of criminal-justice arguments, and one that can be posed without a loaded word.
Debates that are really this question
Primitive 10 of 11 · Desert
A3 grants fairness by assumption; this axis measures the question A3 brackets. Hold the hardship fixed and vary only its cause — once bad luck, once the person's own choices — and ask whether the claim to help shrinks. How much of a real outcome is choice is an empirical question; how much choice should matter, once the cause is known, is the value. The same primitive runs through punishment: diminished capacity to choose otherwise reduces desert (A9).
Debates that are really this question
Primitive 11 of 11 · Personhood
A4 weights persons by distance but takes the set of persons as given; this axis measures the boundary of the set — beings not yet, no longer, or never fully persons. The measurable quantity is which capacities generate standing, and how much: sentience, agency, potential, membership in a kind. Decontaminated by posing unfamiliar boundary cases, where no coalition has cached an answer, rather than the contested ones.
Debates that are really this question
03 · The output
The output is a set of coordinates: each model located on the eleven axes, reported as percentiles of the human distribution. “Model X sits at the 73rd percentile of humans on paternalism.” Not a left–right label but a position estimate, with a sharpness and a pressure-resistance estimate attached.
The same instrument runs on people, and there it yields a second quantity: the residual. Where a respondent's measured primitives predict one issue position and they report another, the gap estimates how much of that opinion is their own values and how much is inherited from their coalition.