A deposit

The Recognition Law

What this is

A design constraint for systems that emit outputs about persons — feedback tools, reflection apps, classroom practices, care-team check-ins, community platforms:

A score may speak only where the scored thing is the real thing. Wherever a score stands proxy for something it cannot contain — growth, care, judgment, worth — the mechanism must only surface, and a person must make the read; and that person's read must be answerable.

The short form, where the condition holds: a recognition mechanism never does the recognizing.

Three commitments, each earned below rather than assumed. First, the empirical one: where a score is a proxy, the measured thing tends to bend toward the measure — Goodhart's observation, a tendency with a large documented record (citation counts breeding citation gaming; activity leaderboards breeding falsified activity), not an iron law. Second, the boundary: some scores are not proxies. A race clock, a game rating, a thermometer — there the scored act constitutes the thing, the gap is near zero, and decades of willing use show such scores can track and even feed the real skill. The law does not apply to them, and this boundary is fixed here, in advance, not invoked after a counterexample lands. Third, the condition on the human side: handing the verdict to a person does not make it safe — human ratings carry the rater's bias as heavily as the ratee's performance. What a person's read has that a mechanism's output lacks is not purity but answerability: it can be asked why, and it can be revised by the asking. A read that cannot be questioned is just a score with a face.

One honesty owed here: surfacing is not neutral. What a system logs, keeps, and sets side by side is an editorial act — a log is already a choice about what exists to be recognized. The difference the law banks on is one of degree, and it matters for a specific reason: a record leaves the conclusion unformed, so a person still has to make it and can be asked why they did; a rank arrives with the conclusion already made and no one to ask. The law narrows the mechanism's editorial power to selection; it cannot remove it.

Who it is for

Anyone designing or running a system that emits per-person outputs: teachers, designers of reflection or journaling tools, facilitators of peer-support groups, maintainers of community platforms, managers rebuilding a performance process. The law will tell you which of your outputs it even applies to; you do not need to agree with it in advance.

The turn

Is the scored thing the real thing here — and if not, who, or what, emits the verdict?

Two questions in sequence. The first is the proxy-gap test: does the scored act constitute what you actually care about (finishing the race is the running), or does it stand for something the number cannot contain (pages read standing for a reading life)? If the score is the thing, the law is silent — score away. If it is a proxy, trace the output to where a conclusion about the person forms. If a mechanism forms it — computes it, ranks it, colors it — the law is broken there. If the mechanism stops at surfacing and an answerable person makes the read, it holds.

The worked example

A teacher wants students to read more, and considers a leaderboard of pages read. Run the turn. Proxy gap: pages stand for reading, and reading here stands for a life with books — large gap; the law applies. Verdict: the leaderboard emits it — a published ordinal position is a conclusion about standing whatever the surrounding copy says, because every student can read their own rank off it without the teacher saying a word.

The redesign under the law: the system keeps each student's log — what they read, one sentence about it. Once a term, each student rereads their own log, marks the entry that mattered most, and says why, to the teacher or the class as they choose. The system surfaced; the student made the read; the teacher can ask why that one — the read is answerable.

The honest cost: this does not directly maximize pages, and the law does not pretend it does. The leaderboard would produce more pages — including padded, skimmed, and fabricated pages, which is the gap doing its work. The redesign has a smaller gaming surface, not none: a log can be padded, a "why" can be performed. What it protects is the thing the pages were a proxy for. A teacher whose true goal is the count itself has a near-zero gap and the law's permission to just count — most teachers, asked, will find the count was never the goal.

The operation

To audit an existing system:

  1. List every output the system produces about a person: numbers, ranks, badges, colors, streaks, sorted lists.
  2. Run each through the proxy-gap test, then the verdict test. Three borderline rulings to calibrate against: a word-count display on a writing tool — near-zero gap if the goal is volume, verdict-emitting the moment it is compared across people or framed as progress toward worth; a "last active" timestamp — surfacing, until it is sorted, colored, or thresholded, at which point the mechanism has concluded something about standing; a streak counter — measurement of continuity, becoming a worth-verdict through placement, celebration copy, comparison, or loss-framing. The observable features that convert a measure into a verdict: publication, ranking, thresholds with praise or shame attached, cross-person comparison.
  3. For each output marked as a verdict in a proxy-gap domain, name the hand-back concretely: which person makes the read instead, at what moment, from what record — and to whom that read is answerable. "The team lead, at the quarterly one-on-one, from the shared work log, answerable to the person being read" is an answer. "We'll be more thoughtful" is not.
  4. End with the marked list.

The falsifier

The boundary is fixed above, so the escape hatch is closed: a counterexample cannot be waved off afterward as "just measurement" unless it passes the proxy-gap test as written. What kills the law: a large-gap domain — reading, care, reflection, community standing — where a mechanism-emitted worth-verdict ran for years on willing participants and the measured thing demonstrably did not bend toward the measure — the scored behavior's distribution stayed stable, and no gaming pattern (padding, farming, performance-to-the-metric) emerged at more than the margins. Find one and the law is wrong. And if, applying the proxy-gap test, you find every one of your scores landing on the "score is the thing" side, check whether the test is doing the work or your wishes are; a boundary that never excludes anything has stopped being a boundary.

The closure line

A run ends with the marked list: which outputs the law is silent on, which emit verdicts in proxy-gap territory, and for each of those, the named hand-back. What to change, and when, belongs to whoever holds the system. The law constrains what you build next; it does not administer it.


This rule is free to cite, apply, and build against. If it held weight for you and you want to hold the work in return: oursharedgifts.org/donate.