Armchair · Scholar
Specification Paper · 2026 · July

The Praise and the Praiseworthiness.

The Wage and the Worth ended owing a measurement: it named recognized standing as the good automation's settlements must re-route, and conceded that no instrument existed to detect its loss. This paper builds the specification — and names the assumption that explains why the column is missing. The measurement of non-material goods has a settled method: ask the person. But standing is conferred, not felt — a property of the social field, not the psyche — and Smith's oldest distinction, praise received against praiseworthiness deserved, marks the gap between what the survey reads and what the society does. Cross felt against conferred, categorical against positional, and every existing instrument fills one of three cells: mattering scales, status ladders, prestige rankings. The fourth — whether a person is counted a contributor, measured from the society's side — has no occupant, though it has close neighbours this paper names rather than elides: three decades of welfare-deservingness research, and a factorial-survey tradition recovering society-side judgment functions since 1977. What is new is not the machinery and not the judge-side stance; it is the object, the categorical gate, the population statistics, and the linkage of the conferred measure to the felt one on the same people. The paper specifies the instruments that would fill the cell — a factorial-survey spectator; a triangle of stated, private, and coordination elicitations that keeps social desirability apart from pluralistic ignorance; an informant design pointing sociometry at contribution; and a desert arm benchmarked against a deliberative mini-public — with three numeric tripwires stating in advance how the programme dies, and one rule held constitutive rather than prudential: measure the society's counting, never administer the person's worth.

StatusUnder review
ThreadMotivation
Pages · Reading48 pp · ~57 min
SubjectsmeasurementrecognitionmatteringAdam Smithsocial indicatorspsychometricsfactorial surveydeservingnesssocial normsbasic incomeoccupational prestigeGoodhart's lawautomation
Contents

The Praise and the Praiseworthiness

Imprint manuscript · July 2026 · accepted at Round 4 of house review (Paper No. 009, Thread I — Motivation, cross-listed Thread IV — Moral Judgment); four adversarial rounds, response letters and dossier amendments filed with the archive. The instrument dossier ships as a supplementary artifact; the corpus index carries the verification flags still open at imprint.

· · ·

“Man naturally desires, not only to be loved, but to be lovely; or to be that thing which is the natural and proper object of love. He naturally dreads, not only to be hated, but to be hateful; or to be that thing which is the natural and proper object of hatred.” — Adam Smith, The Theory of Moral Sentiments, III.ii.1

· · ·

Abstract

The Wage and the Worth ended owing a measurement. It argued that automation’s settlements will succeed or fail on whether they re-route recognized standing along with income, and it conceded that no instrument exists to detect the loss it predicted. This paper builds the specification, and begins by naming the assumption that explains why the column is missing. The measurement of non-material goods has a settled method: ask the person. The wellbeing tradition and its critics quarrel over whether the answers deserve a place in the national accounts, and the quarrel conceals what both sides assume — that the respondent is a sufficient informant on their own social goods. For happiness, plausibly so. For standing, the assumption fails by construction: standing is conferred, a property of the social field, and the person’s belief about the field’s verdict is a second variable — related, fallible, and measurably distinct wherever the two have been paired. Cross felt against conferred and categorical against positional, and the existing instruments fill three cells: mattering scales, status ladders, prestige rankings. The fourth — whether a person is counted a contributor, measured from the society’s side — has no occupant, though it has close neighbors this revision names rather than elides: three decades of welfare-deservingness research have measured, from judge samples, whom the public counts as deserving of support, and the factorial-survey tradition has recovered society-side judgment functions since the 1970s. What is new here is not the machinery and not the judge-side stance; it is the object — contributor status itself, over ways of living waged and non-waged — the categorical gate, the population statistics built on it, and the linkage of the conferred measure to the felt one on the same people. The paper establishes that the adjacent literatures license divergence between felt and conferred standing (r ≈ .2–.5 wherever paired, with a documented mechanism: calibration is manufactured by feedback and thins with it), and it specifies the instruments: a factorial-survey spectator design recovering the culture’s conferral function, with a triangle of elicitations that keeps three quantities apart — stated judgment, private judgment under incentive-compatible cover, and the coordination-elicited perceived shared standard, which is the operative social fact and is measured knowing it may be a misperception, since the gap between what a culture privately judges and what it believes itself to judge is itself a first-class estimand; an informant design pointing sociometry at contribution; a linked design whose estimand is the joint distribution of felt and conferred standing — the calibration gap; and a desert arm, judges under idealized-information and reflective instructions, benchmarked against a small deliberative mini-public because instructed idealization approximates deliberation in name and not mechanism — the approximation, stated as approximation, that lets a paper titled with Smith’s two words keep its title honestly. A validation program is stated to pre-registration grade with named tripwires, including the one this revision hardened: shared-standard verdicts now require agreement across strata of judges, not merely reliable pooled averages, because a national conferral function could be a reliable mixture corresponding to no one’s operative field. One rule is constitutive: the column measures a society’s counting and never administers a person’s worth — no individual scores, no gates, no administrative reuse — because esteem dispensed by formula is esteem destroyed. The contribution is methodological in the strict sense: not a finding but the instrument the finding requires — offered with the concession that the cheapest way for this paper to die, felt sufficing after all, is also the one it most owes the field to test.

· · ·

§ IThe Question and the Empty Cell

Every paper in this archive tries to name the load-bearing assumption a dominant framework will not. This one begins closer to home, with a debt recorded in the archive’s own ledger. The Wage and the Worth (No. 008) reframed the automation debate from the quantity of displacement to the architecture of re-routing, and its framework turned on a variable it could define but not measure: recognized standing, the categorical facet of the worth — whether one is counted a contributor at all. It committed to measurement before menu; it specified an index from adjacent instruments — the worth index of its §6, the phrase this paper’s subtitle carries and this paper’s §6 builds out as four statistics and a function — and conceded, under refereeing, that the specification measured the wrong thing exactly where it mattered — a self-report scale captures whether a person feels counted, where the theory is built on whether a person is counted, as a fact about the surrounding society. It closed by saying we will not know whether we have built a place on the stairs “until we agree to measure the thing the stairs were carrying.” This paper is that agreement, drafted.

Begin with the two words, and with a discipline about them that this revision states at the outset rather than discovering at the end. Smith split the desire for approval in two: the desire for praise, satisfied by approval received, deserved or not; and the desire for praiseworthiness, satisfied only by deserving approval, received or not (TMS III.ii). The words differ by a suffix and the difference carries his moral psychology: unmerited applause gives no solid joy, merited obscurity keeps its consolation, and the two verdicts are issued by different tribunals. Now the discipline. A survey of actual spectators — however many, however well sampled — recovers praise: the field’s verdict as it stands, deserved or not. It does not recover praiseworthiness, which is adjudicated not by counting judges but by qualifying them — the well-informed, impartial, reflective standpoint Smith builds his spectator from. This paper’s primary instruments therefore measure the gap between two praise-side quantities: regard as felt by the person and regard as conferred by the field. That gap — the calibration gap — is the paper’s measured object, and every sentence that reports it says “counted,” not “worthy.” The second word of the title names the paper’s direction of travel, not its primary estimand: a specified extension arm (§6) moves from what the society does count toward what it would count under idealized information and reflection — an approximation toward praiseworthiness along Smith’s own dimensions, labeled approximation wherever it appears. The title is a debt the paper works off in stages, and it says so.

Here is the assumption stated so that it can be doubted. The respondent is a sufficient informant on their own social goods: the fact a survey needs is always inside the person surveyed. Both parties to the wellbeing-measurement quarrel hold it. The boosters hold it when they argue that subjective wellbeing is reliable enough to steer policy: the feeling is the fact, so measure the feeling. The skeptics hold it when they argue that feelings are noise dressed as data: the feeling would be the fact if it were worth measuring, which it is not. Neither asks whether, for some goods, the fact is not a feeling at all. For happiness the assumption is true by definition. For standing it is false by construction: standing consists in how the surrounding society counts a person, and the person’s belief about that counting is a second, separate variable. A census of the judged cannot see what only the judges know.

The claim, stated so it can be wrong — and resized, in this revision, to what the adjacent traditions leave genuinely unclaimed. Recognized contribution — the categorical, conferred standing on which No. 008's framework turns — is a measurable social fact. Judge-side instruments in the factorial-survey lineage can recover, with stated reliability and stated cross-stratum agreement, whether a society counts an activity as contribution and a person as contributing. The conferred measure will diverge measurably from felt mattering, with structure rather than noise. At current baseline the conferral function will weight waged employment heavily; the loading is a validity check and a finding, since the culture presently routes recognition through the wage, and the instrument’s job is to be able to see that loading change — with its signaling component separated from its constitutive component by design, because a judge who reads “employed” infers industry and competence, and the wage’s weight in the counting is not the same thing as its weight as evidence. The validated column converts No. 008's first falsifier from a specification into a runnable study for the activity-mediated channel, and says plainly which channel it cannot reach. Three observations would disconfirm the claim, derived in §6 with numeric tripwires: convergence (felt was a sufficient statistic all along), incoherence (the judges share no standard — now tested across strata, not just in the pooled average), and positional collapse (conferral cannot be separated from rank).

A word on method. Every non-trivial sentence belongs to one of three registers — what recognition is (conceptual, passage-level, constructions labeled), what the instruments show (magnitude, venue, contest), what can be expected (trend, assumption, falsifier). Three standing rules: the no-prophecy rule — the column is motivated by a conditional future but specified and justified entirely in the present tense; the no-administration rule — the instrument measures a society’s counting and never administers a person’s worth, a rule §8 shows is constitutive rather than prudential; and the no-reification rule — until the validation program returns, the column is a specification, and no sentence treats the unfielded index as data.

§ IITwo Caricatures to Clear

Two pictures dominate any conversation about measuring a soft construct, and each must be cleared before the argument can run, in the house manner: by concession first.

The first is the operationalist: a construct is whatever its scale measures, so if you want mattering, add the mattering items — the wellbeing toolbox is mature, and the missing column is filled by renaming a full one. The concession is real. The felt instruments are good: reliabilities from .74 to .98 across forty-five years, a negative pole, workplace and community variants, meta-analytic links to mental health (§4). If the paper’s claim were that self-report is sloppy, the operationalist would win. The claim is different: the felt scales measure a different object — regard as perceived — and no reliability in measuring one variable makes it a measurement of another. Whether the two objects coincide in practice is an empirical bet this paper converts into a named falsifier and predicts will lose, on the record of every adjacent pairing. What the operationalist cannot do is win by definition.

The second is the mystic: standing, dignity, recognition are unmeasurable in principle — quantification profanes them, and the column should not exist. The concessions, again, are real: a number is not the good; bad measurement of sacred things has a body count (§4 carries the self-esteem debacle as a first-class exhibit); and any instrument whose output could become a score attached to a person deserves suspicion — a suspicion this paper adopts as a design rule rather than talks anyone out of (§8). But the in principle proves too much, and this revision can make the rebuttal stronger than the first draft knew. Sociology has measured conferred regard society-side for seventy years (occupational prestige); developmental psychology has measured person-level conferred standing in every peer-nomination study ever fielded; the justice tradition has recovered society-side judgment functions over person-profiles since Jasso and Rossi in 1977; and welfare-deservingness research has spent three decades asking judge samples whom the public counts as deserving. The conferred side of social life is not merely measurable in principle — it is one of the most measured things in social science. What has never been measured is one specific object, and the mystic has mistaken an empty cell for a sacred one.

What the two caricatures share is the assumption that measurement validity is settled by fiat — granted by definition or denied by definition — when validity has been, since Cronbach and Meehl (1955), something earned by validation, construct by construct, test by named test. The paper’s two wagers — the cell is empty, in a sense §5 now states precisely; the cell is fillable — are both empirical, and clearing the caricatures is the first statement of the thesis.

§ IIIThe Construct Named, I: What Recognition Is

This section names the construct in the conceptual register, from witnesses read at passage level, and builds the paper’s central construction in the open. The would/wouldn’t discipline governs throughout.

Smith supplies the distinction and the design constraint. The desire to be loved and the desire to be lovely are different desires with different satisfaction conditions: praise can be received undeserved, and “the love of praise-worthiness is by no means derived altogether from the love of praise” — the man applauded for what he did not do finds it “more mortifying than any censure” (TMS III.ii). Two features matter to measurement. First, praiseworthiness is not a feeling of the agent, and it is not a tally of the agent’s actual audience either: it is the verdict of a qualified standpoint — informed, impartial, reflective — which is why Smith’s two tribunals (the man without, the man within) can conflict, and why no poll of the men without, however large, is a measurement of the man within. This revision holds that line strictly: the paper’s judge samples recover conferral — aggregated actual praise — and the approach to praiseworthiness is a separate, labeled arm that qualifies the judges instead of multiplying them (§6). Second, the negative pole is symmetrical: the dread of being hateful — a proper object of resentment — is as load-bearing as the desire to be lovely, which is why the instrument carries a burden item and not only a contribution item.

Weber, added to the witness list in this revision, supplies the sociological name the first draft should have carried: status honor — standing as social estimation, attached to ways of living, conferred and withdrawn by communities, convertible with but never identical to market position (“Class, Status, Party”). The construct this paper measures is Weberian status honor at its categorical floor: not the honor gradations of status groups but the threshold estimation — counted a contributor at all — beneath them.

Now the construction, labeled as such. For a person in a social field there are three verdicts, not two: felt standing — the person’s perception of being counted, which is what every self-report scale reads; conferred standing — the field’s actual counting, a fact about the judges rather than the judged; and deserved standing — praiseworthiness proper. The paper’s primary instruments target the second. The second is evidence about the third — aggregating spectators removes the self’s flattery and the lone judge’s caprice — but it is not the third, a culture can confer wrongly, and the instrument recovers the culture’s counting with its biases intact. The paper polices its verbs: counted is what the primary instruments measure; would be counted on reflection is what the desert arm approximates; worthy is what no instrument in this paper measures.

Durkheim supplies the warrant that the second verdict is measurable at all. A social fact is a way of acting and judging “external to the individual, and endowed with a power of coercion,” to be treated as a thing (Rules, ch. 1); Suicide is the tradition’s proof that an inner state can be read from social-side data. And Durkheim’s coercion clause points at something this revision takes more seriously than the first draft did: the counting operates — on reputations, marriages, benefit take-up, obituaries — which means stated judgments are not its only trace, and §6 now carries behavioral validation channels alongside the stated ones.

Rawls locates the good outside the respondent — the tradition’s own admission. The social bases of self-respect are “perhaps the most important primary good,” and they are bases: institutional facts, the appreciation of one’s person and deeds by associates (§67; wording verified at copy-edit). Self-respect, in the tradition’s own philosophy, normally depends on the respect of others; the felt state rides on a conferred fact. A measurement program that stops at the felt state has measured the shadow on the wall of the thing it cares about.

Honneth and Taylor make the conferred good categorical: recognition in the esteem sphere is accorded for contributions to shared goals (The Struggle for Recognition, ch. 5); misrecognition “can inflict harm, can be a form of oppression” (“The Politics of Recognition,” p. 25); and being counted a contributor is not a rank — everyone in a field can hold it at once. That non-scarcity distinguishes the categorical cell from the positional one, inherited from No. 008 §6 and cited, not re-argued. One refinement this revision adds under referee pressure (§6): if the counting turns out psychometrically continuous, non-scarcity survives as threshold semantics — a claim about scoring, not about latent structure — and the paper owns that reading rather than smuggling it.

Akerlof and Kranton, added this round, give the construct its microfoundation inside economics: identity economics puts category membership directly into the utility function — who one is counted as being carries payoffs — which is the discipline’s own shelf for what this paper measures. And Brennan and Pettit give the conferred good its economics — esteem as a commodity that cannot be bought or commanded, only elicited under observation, the intangible hand — with two implications the paper builds on: a distribution exists, so the statistical object is coherent; and esteem pursued or dispensed too directly is forfeited — the teleological paradox that seeds §8's constitutive rule. Bénabou and Tirole sharpen the economics into a warning the design must answer: esteem is inference — observers read conduct under an information set — so any judgment weight on a signal like employment bundles what the signal constitutes with what it merely evidences, and §6 separates the two by design rather than assumption.

The section’s yield. Received and deserved regard are distinct, and the deserved is reached by qualifying judges, not counting them (Smith); the field’s actual counting is status honor at its floor (Weber), a measurable social fact that also operates behaviorally (Durkheim); the tradition’s own philosophy locates the good in that counting (Rawls); the counting is categorical and non-scarce, at least as semantics (Honneth, Taylor); it sits in the utility function (Akerlof–Kranton) as inference under observation (Bénabou–Tirole) in an economy that forbids its administration (Brennan–Pettit). Cross felt/conferred with categorical/positional and the measurement space has four cells. What occupies them is §4's question.

§ IVThe Construct Named, II: What the Instruments Show

This is the sobriety chapter. It audits the three filled cells at full strength, assembles the divergence record the paper’s wager extrapolates, then runs the deflation, which writes the design constraints for everything that follows.

The felt-categorical cell first, because it is the best-instrumented and the most misread. The mattering tradition runs from Rosenberg and McCullough (1981) through the General Mattering Scale, Elliott, Kao and Grant’s three-facet Mattering Index (2004), Flett’s synthesis (2018) and its negative pole, Prilleltensky’s multidimensional MIDLS — whose architecture, feeling valued crossed with adding value across relationships, work, and community, is the closest existing structure to this paper’s construct — and workplace and university variants with an explicit societal-mattering factor (Jung & Heppner 2017). The internal-consistency record is uniformly strong: alphas from .74 to .94, a general-factor omega of .98, across adolescents, students, working adults, homeless men, and community samples in multiple languages. And the validation record has a single, remarkable shape, which this revision states at the strength the search supports rather than as omniscience: we found no mattering instrument ever validated against the reports of the people one supposedly matters to — a systematic search across informant, partner, dyadic, round-robin, and peer-nomination pairings, protocol in the corpus index, one dyadic near-miss examined and found to survey only the perceiver (Mak & Marshall 2004). The nomological net — self-esteem, belonging, support, loneliness, depression — is woven from the same respondent. This is not a scandal: the tradition set out to measure a perception and measured it well. But it means the correspondence of felt mattering to conferred regard is not weak — it is untested by the field that owns the construct, and multi-informant machinery routine in personality psychology since the 1990s (Vazire 2010; Connelly & Ones 2010) has never been pointed at it.

The felt-positional cell is consequential and visibly diverges from its objective anchors — instructively. MacArthur-ladder subjective status correlates with objective position at r ≈ .25–.34 by indicator, pooled r = .32 across 357 studies and 2.35 million respondents (Tan et al. 2020; Cundiff & Matthews 2017); in Whitehall II, entered simultaneously with objective grade, only the subjective measure remains associated with health and decline (Singh-Manoux et al. 2005). Felt standing is real, predictive, and its own variable — findings this paper endorses because they cut in its favor: the ladder’s power net of objective position is the live demonstration that felt and anchored standing come apart.

The conferred-positional cell proves feasibility and the wrong tool in one apparatus. Occupational prestige has been measured society-side for seventy years with remarkable cross-national stability (Treiman 1977; the ISEI lineage), and the 2012 GSS restudy fielded the full machinery — 1,001 raters, 860 occupation titles, a nine-rung ladder. Society-side judgment at national scale is a solved survey problem. Its object, almost always, is occupations — the tradition did once rate the homemaker (Nilson 1978; Bose 1980 — a trim this revision owes to its referee: the universal was overstated) — and its output is rank: a scale that orders cannot register a good that is not an ordering. The apparatus exists; it has been pointed at the wrong grammatical object.

And the near-neighbors this revision restores to the map, because §5's precision depends on them. The deservingness tradition — van Oorschot (2000) and three decades of successors — has measured, from judge samples and often with vignettes, whom the public counts as deserving of support, structured by criteria (control, attitude, reciprocity, identity, need) of which reciprocity — has contributed, will contribute — sits nearest this paper’s gate. The factorial-survey justice tradition — Jasso and Rossi (1977) forward — has recovered society-side judgment functions over person-profiles for half a century; Jasso’s just-reward function is, structurally, this paper’s conferral function pointed at pay. Ridgeway’s status-construction program has measured conferral experimentally in interaction for decades. These traditions are judge-side, vignette-fluent, and categorical in operation, and the first draft’s claim that “no one has crossed the existing methods with the missing construct” was false as to lineage. What none of them measures — §5 states it precisely — is the object itself.

The fourth method exists too. Sociometry — peer nomination aggregated across a bounded community — has supplied person-level conferred measures for decades, and its comparison with self-perception opens the divergence record: sociometric and perceived popularity correlate at about r = .40 (Parkhurst & Hopmeyer 1998), with the correlation declining through adolescence to non-significance for ninth-grade girls (Cillessen & Mayeux 2004). The record generalizes. Meta-analytic self–other agreement on personality runs r ≈ .29–.51 by trait, .21 in low-acquaintance designs (Connelly & Ones 2010; Kenny & West 2010); the asymmetry has structure — the self is the better informant on internal states, others on evaluative, externally visible traits (Vazire 2010) — and standing is evaluative and external. Metaperception completes the pattern: beliefs about how others see us track how we see ourselves (r = .87; Kenny & DePaulo 1993), knowledge of being liked runs as low as ≈ .17 dyadically at short acquaintance, and the systematic error is underestimation — the liking gap (Boothby et al. 2018). One exception organizes all of it: in face-to-face groups, self-perceptions of status are accurate (.46–.63) because accuracy there is enforced — over-claimers are punished with lost acceptance and conflict — while unenforced perceptions in the same groups run at .16–.20 (Anderson et al. 2006). Calibration, where it exists, is a social product, manufactured by feedback. Where feedback thins, there is no manufacturer.

Now the deflation, at full strength, because it writes the constraints. First: self-report performs well in its own regime — no global self-enhancement against informants (mean-level δ = −.038 across 152 samples; Kim, Di Domenico & Connelly 2019, flagged at verification grade in the endnote), and the felt measures carry predictive weight the conferred ones may not (Whitehall again). Any version of this paper that reads as “self-report bad” has failed its own evidence. Second: a column is not influence — the beyond-GDP record is a record of columns built, adopted, and largely idle; the paper’s claim is decidability, not power. Third: the cautionary tale is real — the self-esteem movement elevated a felt construct to policy before validation, and the definitive review found the causal benefits largely absent (Baumeister et al. 2003); validation before deployment, deployment before enthusiasm. Fourth: the survey methods nearest this paper’s designs have known fragile joints, and this revision names both inheritances correctly. From the factorial-survey tradition — which is what the spectator design is — the fragilities are attribute imputation (judges infer what profiles omit), deck effects, and description-language framing (“stay-at-home mother” and “provides forty hours a week of childcare” will not score alike, so phrasing choices are substantive and the design treats them as manipulations, not defaults). From the anchoring-vignette tradition — inherited only for the calibration and scale-use component — the fragile assumption is vignette equivalence, which failed cross-culturally in the largest health applications while response consistency survived; the design consequence is within-culture first, equivalence tested by design. Fifth, and newly weighted in this revision: stated judgment is not operative judgment. A zero-stakes rating under full social monitoring invites the judge to perform virtue, and desirability pressure will inflate the stated counting of caregivers and volunteers exactly where the paper’s baseline projection expects discounting — a bias running against the paper’s own hypothesis, which is the dangerous direction, since it mimics disconfirmation. The design answer (§6) is a second conferral method borrowed from experimental economics — coordination-incentivized norm elicitation (Krupka & Weber 2013), which rewards accuracy about the shared standard rather than sincerity about one’s own virtue — and behavioral validation channels where the counting operates on conduct: the welfare-stigma and take-up literature, in which eligible households leave money unclaimed at rates that reveal a conferred-standing cost entering utility (Moffitt 1983; the modern take-up experiments are its measurement arm).

The chapter’s honest summary. Three cells are filled and well-behaved; the near-neighbors measure deservingness-of-support, just rewards, and interactional status, but not contributor status itself; the adjacent record licenses an expectation of divergence — r ≈ .2–.5, structured by feedback density — and no measured domain has shown felt/conferred identity. Whether standing is the exception is what the instrument exists to decide.

§ VThe Documented Present

Everything in this section is measured or dated. This revision runs it in three halves: the apparatus, the adjacent traditions, and the nearest tests.

The apparatus, swept — with its scope stated. The universal negative at the paper’s base is scoped to what was actually swept: no official statistical agency or major survey program carries a conferred-categorical measure, asserted after an attempt to break it across fourteen instrument families as of mid-2026 — the ONS4 and the reorganized UK wellbeing dashboard (fifty-nine measures, February 2026); the OECD’s How’s Life? 2024 (eleven dimensions; the 2024 additions were loneliness, energy poverty, temperature exposure, and pain); Eurostat and EU-SILC; the UNECE 2025 guidelines; Gallup’s World Poll and Q12; the World Happiness Report 2026; the ESS’s dormant wellbeing module; the GSS and its worklife module; the ISSP; New Zealand’s Living Standards Framework; Bhutan’s GNH; MIDUS; the Global Flourishing Study (2024–25); the satellite accounts and volunteering statistics. The sweep’s yield is a near-miss field, named: the closest item in existence is Eurofound’s exclusion item — “I feel the value of what I do is not recognised by others” (2011–2016; the successor survey fields in autumn 2026, retention unknown — a watch item, flagged); behind it, the ESS’s “appreciated by the people you are close to,” the ONS4's “worthwhile,” Gallup’s seven-day workplace recognition item, the ISSP’s “my job is useful to society,” MIDUS’s contribution subscale. Every near-miss is felt-form; the one society-side apparatus in official statistics ranks job titles.

The adjacent traditions — the second half the first draft owed. Academic research, outside official statistics, has been nearer the cell than the wellbeing dashboards ever came, and the honest map shows three neighbors and their exact distances. The deservingness tradition (van Oorschot 2000; the CARIN criteria; three decades of cross-national vignette studies) measures society-side, categorical-in-operation judgments — but its object is deservingness of public support: a conditional, distributive attitude, entangled with need and with views of the welfare state, fielded at attitude-survey grade with no person-linkage, no felt-side pairing, and no population column. Reciprocity, its criterion nearest this paper’s gate, asks whether support is earned — not whether the person is counted a contributor, full stop, benefits aside. The factorial-survey justice tradition (Jasso & Rossi 1977 forward) recovers judgment functions over person-profiles — the machinery this paper’s spectator design belongs to, said plainly — but its recovered functions concern just rewards and fair earnings, conferral’s positional cousin, not the categorical count. Ridgeway’s status-construction and expectation-states program measures conferral experimentally — how status beliefs form and attach to categories in interaction — the lab-scale dynamics of exactly the phenomenon whose population-scale statics this paper wants; it has never been asked to produce a national statistic. And a fourth neighbor arrived inside this paper’s own sweep window: the pandemic’s key-worker moment, in which categorical esteem visibly moved without pay moving — a natural experiment in conferral now carrying an empirical literature (canon settled at copy-edit, flagged). The resized claim, stated once: the machinery is inherited; the judge-side stance is inherited; what has no occupant is the object — contributor status itself, over ways of living waged and non-waged, with a categorical gate, population statistics built on it, and linkage to the felt side on the same people.

The corrected frontier. The parent reported that the cash literature had never measured the worth channel; that was true when written and requires narrowing now. Since 2023, Penn’s guaranteed-income research center has fielded an adult mattering scale across its American studies: St. Paul (2023), Cambridge (2024), Gainesville (February 2025), Oakland (April 2025) — with results mixed in an instructive way: elevated mattering in Oakland and Gainesville, and in Cambridge a statistically significant negative movement on the Importance subscale at twelve months (verification-flagged), the first live hint that cash and felt mattering can move independently. Germany’s three-year pilot reported purpose in life up 0.25 SD (April 2025) alongside no labor-market withdrawal. The flagship American trial — employment paper accepted at the Quarterly Journal of Economics, health paper conditionally accepted at the American Economic Review — still carries no standing construct, with wellbeing gains reverting after year one. And a result this revision adds where it always belonged: the closest existing experimental separation of work’s non-pecuniary channel — Hussam, Kelley, Lane and Zahra’s randomized employment-versus-cash design in Rohingya refugee camps (AER, 2022), where employment outperformed equal-value cash on psychosocial outcomes and a meaningful share of participants would pay to work. Its outcomes are felt-side, so it does not fill the cell; it is the strongest evidence in the record that the channel the cell would measure is real. The frontier, precisely: the felt column has entered the trials; the conferred column exists nowhere — not in the trials, not in the accounts, not in the panels.

The nearest test, read at its size. The Austrian job guarantee — the Marienthal experiment, evaluated with a pre-registered randomized waitlist inside a pre-registered synthetic control, forthcoming at AEJ: Economic Policy (2026) — instrumented Jahoda’s latent functions at last (the twelve-item LAMB battery). Two results organize this paper’s reading, both verification-flagged pending the published version. Starting work moved felt social recognition (+0.138 on a zero-to-one scale, p = .029; composite +0.108, p = .001, placebos passing). And the social-status ladder did not respond to starting work (−0.013, p = .614) while rising for those merely covered by the guarantee (+0.115 anticipation; +0.132 total a year on): the certificate moved before the activity did. This revision demotes the reading from keystone to suggestive, for a reason the first draft did not carry: the design was saturated — the whole town knew the guarantee existed — so the coverage effect may be a community-level shift in the local norm field rather than individual certification, and staggered entry embeds anticipation and selection in the split. What survives demotion is real: the worth bundle is trial-instrumentable, and “recognized standing” has structure — certificate versus activity — that felt instruments only partly resolve, in a design whose every measure was nonetheless a self-report. Sixty-two experimental participants, one town, a pandemic-era labor market.

The section’s structural observation, one notch more precise than the parent’s. The felt half of the ledger is filling in — mattering in the cash trials, latent functions in the job guarantee, psychosocial outcomes in the refugee-camp experiment — and the adjacent traditions have surrounded the conferred half for decades: deservingness of support, just rewards, interactional status. The object itself — counted a contributor, population-scale, linked to the felt side — has no instrument anywhere. The empty cell is smaller than the first draft’s rhetoric and exactly as empty. Everything that follows targets it.

§ VIThe Hinge: The Fourth Cell

This section is the framework and the instrument, labeled as such: not a finding but the structure that organizes the findings and generates the falsifiers. Its operational form — item pools, grammars, thresholds, power — ships as the instrument dossier (v1.1, amended after Round 1); this section states what a reader needs to evaluate the argument.

The quantities. For a person: FC, felt standing — the mirrored self-battery plus the mattering anchors; CC, conferred standing — the field’s counting; FP and CP, the felt and conferred positional measures the existing instruments supply. Over a population, the column is four statistics and one function. C(p) — the conferral function: the culture’s recovered weighting over activity profiles, waged and non-waged; not a score for anyone, but the standard itself, made visible — and made visible twice, because this revision distinguishes the national function from the stratum-specific ones, for a reason given below. κ — the conferral share: the proportion of the adult population whose actual ways of living are counted (threshold rules pre-registered, expected ranges pre-stated — this revision carries a ceiling prior on the gate item and treats divergence across threshold rules as a finding). λ_w — the wage-loading: the marginal weight of waged status in the counting, reported as a reduced-form judgment weight with a designed decomposition, because a judge who reads “employed” infers competence, health, and reliability, and the wage’s constitutive weight in the counting is not its signaling weight as evidence about unstated attributes (Bénabou–Tirole); the grammar therefore carries reason-for-status at full resolution — laid off, chose caregiving, unable, independently wealthy — and an information-set manipulation whose difference identifies the signaling share. Dispersion of CC, reported on its own bounded scale with its functional stated in the dossier — the first draft’s “reported the way income inequality is reported” invited a false equivalence and is withdrawn. And the calibration gap — felt against conferred on the same people. The first draft called this Δ and treated it as an estimand; this revision keeps the name for exposition and kills the estimand, on a sixty-year-old critique that applies in full (Cronbach & Furby 1970): a difference of within-sample standardizations is mean-zero by construction and compounds the unreliability of both parts. The estimand is the joint distribution of (FC, CC) — bivariate latent model, stratum-moderated, response-surface analysis as robustness — from which the gap’s structure is read: who runs felt-above-conferred, who conferred-above-felt, and where the two decouple entirely. A refinement rides along: the felt battery’s metaperception item (whether one believes one is counted) separates not knowing the field’s verdict from disagreeing with it.

Which function the payoff needs — the question this revision answers instead of leaving implicit. No. 008's worth channel is standing in the person’s operative field: the reference structures that P3 below says govern felt standing are local, and Lamont’s ethnography predicts the conferral criteria themselves are class- and race-differentiated. A national C(p) could therefore be a reliable mixture corresponding to no one’s operative standard — reliable in the average, incoherent in the field. The design meets this head-on: stratum-specific conferral functions are primary estimands, not robustness checks; the national function serves the accounts (κ, drift, the public column); and the parent’s framework payoff — does re-routed contribution get counted where the person lives — is carried by the operative-field instruments, not by the national average. Both objects are worth having; they are not the same object; the paper says which is which.

The ruling is level-indexed — the propagation the constitutive claim requires. Declaring the coordination arm CC-constitutive raises an immediate question the framework must answer rather than leave for a hostile reader to find: if the operative social fact is the perceived standard, why does the linked design pair felt standing with the informant measure, whose informants report first-order verdicts rather than perceived norms? Because constitutive is indexed to a level. At the norm level — the culture’s standard over ways of living, the object of C(p) and κ — coercion is anticipatory: it runs on what people believe others count, so the perceived-shared standard the coordination arm recovers is the operative fact, misperception included (the Durkheimian ground of the ruling). At the person level — the realized counting of a particular person in a bounded community, the object of the informant design — the informants are the field: their first-order verdicts both constitute the standing the person actually holds and administer its sanctions, the invitations extended and withheld, the regard shown and denied, so the informant aggregate is itself operative there and needs no second-order detour. The linked design therefore pairs FC with the informant CC by design, not by oversight: the person-level operative measure is the first-order aggregate of the people who are that person’s field. The operative-field payoff thus has two carriers because it is two objects at two levels — stratum-target coordination for the operative norm, the informant aggregate for the operative person-level counting — and the paper indexes each claim to its level wherever it is made.

The instruments. The spectator design, named by its true lineage: a factorial survey (Jasso & Rossi; Auspurg & Hinz), inheriting that tradition’s known fragilities — attribute imputation, deck effects, description-language framing, each answered in the dossier by design (reason-for-status attributes; information-set arms; pre-registered dual phrasings treated as manipulations) — with an anchoring-vignette inheritance confined to what it actually covers, the calibration and scale-use component (two fixed calibration profiles first, judge-level recentring, equivalence diagnostics across judge strata, within-culture scope). Eighteen hundred judges is not the binding constraint and the paper stops headlining it; the binding constraints are vignettes-per-judge, factorial coverage, and D-efficiency on the wage-by-activity interactions, stated in the dossier and here. And one honest sentence the Round 1 additions accumulated a debt for: the arms cannot share judges — stated, private, coordinated, and desert elicitations contaminate each other, so allocation is between subjects — and the full program’s judge demand is roughly three and a half times the de-headlined figure, about sixty-four hundred across arms, with the allocation table in the dossier; the additions were not free, and the paper prices them. Judges rate anonymized profiles on a conferral index whose items name contribution without naming payment, a categorical gate (with pre-registered harder variants, because the soft variant plausibly runs at ceiling), and a separate ladder item scored to the positional construct. The judges are the culture’s actual spectators — a quota sample is neither informed nor impartial, and this revision drops the first draft’s flourish to the contrary; what the anonymized design buys is narrower and real: no partiality toward the judged, who cannot be known, flattered, or lobbied. The norm-elicitation arm, added last round as the second conferral method and completed this round into a triangle, because its Round 1 form conflated two things the paper’s commitments require it to separate. Coordination-incentivized judgment (Krupka & Weber) — judges rewarded for matching a reference population’s modal rating — recovers the perceived shared standard: second-order beliefs, not first-order judgment. Where private judgments and perceived norms diverge, the coordination arm faithfully recovers the misperception — and the canonical modern demonstration that they can diverge massively on exactly this class of object is Bursztyn, González and Yanagizawa-Drott (2020): Saudi men privately supported their wives working, systematically underestimated other men’s support, and behavior tracked the misperceived norm until information corrected it. The design therefore runs three elicitations, between subjects: stated public rating (the A1 instrument), private judgment under incentive-compatible cover (a list-experiment variant of the gate, randomized response as robustness), and coordinated perceived norms — with the pairwise differences identified in the dossier: stated-minus-private is desirability distortion; private-minus-coordinated is the norm wedge, pluralistic ignorance about the culture’s own counting. The framework then rules, rather than leaving implicit, which arm is CC-constitutive: the coordination arm, on Durkheimian ground — the counting’s coercion runs on what people believe others count, so the operative social fact is the perceived standard even where it is a misperception — with private judgment measuring the culture’s latent standard and the wedge reported beside every operative estimate, because the wedge’s R2 relevance is direct: a wage-heavy operative norm sitting on a care-egalitarian private standard is fixable by norm-information intervention and needs no re-routing architecture, while wage-heavy private judgment is the vacuum the parent diagnosed. And the coordination target is specified in text, not only in the dossier: national-target coordination measures the perceived national norm and serves the accounts; stratum-target coordination measures the perceived local norm and carries the operative-field payoff at the norm level (the layering paragraph above). Three disciplines the dossier attaches to the private leg belong in text, because they set its grade. The list experiment identifies prevalence in aggregate, per profile (Blair & Imai 2012, with the standard design and placement diagnostics named there): the wedge therefore exists at profile-by-stratum resolution over a covered subset of the grammar — a low-resolution overlay on the high-resolution stated function, not a third full-resolution arm, and the paper stops implying otherwise. The sensitive direction flips across profiles — for a sympathetic non-waged profile the concealed answer is “not counted,” elsewhere the reverse — which constrains list construction profile by profile. And a focality diagnostic is pre-registered against a failure the wedge’s whole interpretation rides on: coordination can settle on focal answers — scale midpoints, stereotype-consistent responses — rather than perceived typicality, so placebo profiles and attribute variations where no plausible norm exists are fielded, and strong coordination there flags salience contamination rather than pluralistic ignorance. The informant design: sociometry pointed at contribution — bounded communities, consented targets across waged and non-waged lives, five-plus informants each across nominated and roster channels, reliability-qualified — the person-level conferred measure, and the operative-field carrier at the person level, where (per the layering paragraph) the informants are the field and their first-order aggregate is the realized counting itself. The linked design: both halves on the same people, mirrored item-for-item, joint distribution as estimand. The desert arm, added last round and disciplined this round by its true lineage. The question it asks — what would the public count under information and reflection? — has a standing method: deliberative polling and the mini-publics tradition (Fishkin), which produce considered judgment through actual information and actual deliberation. Instructed idealization — assume you know everything relevant; set aside how popular this way of living is; ought it count? — approximates that method in name, not mechanism, and its known failure mode is that the instruction telegraphs which answer the researcher deems reflective: the desirability inflation the coordination arm purges from the descriptive side re-enters the normative side wearing a philosopher’s coat, and it inflates exactly the quantity this arm exists for — whether the culture’s considered standard outruns its practice. Three disciplines follow, all in the dossier: the arm’s expected inflation direction is pre-registered (instructed idealization is expected to over-count non-waged contribution relative to its benchmark), a small deliberative mini-public rating a profile subset after structured information and facilitated discussion serves as the arm’s validation criterion — so normative conferral has a benchmark the way stated conferral now has the triangle — and the philosophical honesty sentence is carried in text: whether any idealized-attitude construction tracks desert has been disputed since Firth (1952), so the arm approximates Smith’s named dimensions — information, impartiality, reflection — and claims no further. Its output remains normative conferral, labeled, never desert measured. And vignette-imputation, the element that scales, now sold at its true grade: any respondent’s activity profile can be scored through C(p) by judges who never met them — but imputed CC is a deterministic function of the profile, so it detects only activity-mediated standing changes and is structurally blind to changes running through community perception without activity change. The blindness is stated here in tripwire-grade type because the first draft made the reader find it by arithmetic.

The validation program. The multitrait–multimethod logic has to be handled with care here, because the triangle broke the naive version of it: MTMM reads same-trait cross-method convergence as validity, but the paper has just spent a page establishing that stated, private, and coordinated elicitations measure distinct quantities whose divergences are estimands — so a matrix that lumped them together would pre-register the triangle’s own findings as validity failures. The matrix is therefore partitioned. The validity matrix proper holds the convergence-expected pairs, where cross-method agreement is evidence the trait is one thing: informant CC against imputed CC (the multimethod test of the person-level measure — with its criterion state-dependent, because the imputation-blindness paragraph below binds here too: the pair must clear a pre-registered predictive floor at baseline cross-section, where profiles should predict realized standing or the person-level measure has no multimethod anchor, while residual and dynamic divergence between them — local reputation moving without activity moving — is logged as the reputation channel, the very structure the blindness paragraph predicts, and not as validity failure); stated against private where the profile carries low desirability stakes; the felt battery against its established anchors. The estimand set holds the divergence-meaningful pairs, where a gap is a finding and not a failure: stated versus private (desirability), private versus coordinated (the norm wedge). The coordination column is entered as its own trait — the perceived norm, constitutive at the norm level — not as a method for measuring first-order conferral, which is the distinction the level-indexed ruling drew. Around the partitioned matrix: an invariance ladder across waged and non-waged strata, metric invariance the floor; known-groups sensitivity — the instrument must distinguish a thirty-hour caregiver, a sustained volunteer, a commons maintainer, a civic retiree from a long-term-idle comparison, or return for redesign as an employment detector, with variance decomposition separating instrument insensitivity (a defect) from the culture’s verdict (a finding), and with clearance run per conferral arm: the stated and coordination arms clear independently, the coordination arm’s clearance is the one P2's scoring waits on, and the private arm is cleared at the aggregate list-experiment level within its power limits (a coarser test than the graded arms get, and named as such); graded-response IRT calibration of the full index, with the categorical claim carried as threshold semantics if the latent structure runs continuous; test–retest; holdout replication; and the behavioral channels, conditioned this round by the literature that decomposes them — take-up gaps bundle stigma with hassle and information frictions, separated experimentally in the modern record (Finkelstein & Notowidigdo 2019), so the conferral prediction registers against the stigma component where decomposed and carries its own falsifier (local conferral functions fail to predict experimentally isolated stigma), which keeps a hassle-dominated null from reading as conferral-irrelevance; the obituary channel remains exploratory and labeled. Every threshold is numeric and pre-stated in the dossier.

The falsifiers, derived — one rebuilt this round. The framework makes three load-bearing claims; each falsifier is a negation with its tripwire fixed in advance. The cells are distinct: negated if the disattenuated felt–conferred association reaches .70 in every stratum including the non-waged, with the joint distribution showing no stratum structure (convergence — the cell was filled all along; the mattering battery suffices; this paper collapses into a short memo). The conferred cell is measurable — rebuilt last round, completed this round: the shared-standard verdict requires both the pooled thresholds and cross-stratum agreement, with the agreement statistic now computed over profile-level predicted conferral across the covered profile space (bootstrap intervals), not the rank order of a handful of attribute effects; and the verdict map no longer leaves the plausible modal outcome unnamed — at or above .70, a shared standard; between .40 and .70, partial sharing, the pre-registered expectation: stratum functions become the primary estimands, the national function is reported as a weighted mixture with a heterogeneity index, and κ is reported stratum-wise; below .40, or below the coordination-rate floor, incoherence (incoherence — the cell is empty and unfillable as a shared standard; what remains measurable, and the paper would report it as the finding it is, are the strata’s separate countings: Lamont vindicated at national scale). And conferral is separable from rank: negated if the categorical and positional conferred measures reach latent .85 or one-factor equivalence (positional collapse — Smith’s non-positional praiseworthiness is philosophy without an instrument, and R2 loses its target).

What the parent inherits back, at its true grade. When the validation program passes, No. 008's first falsifier converts from specification to runnable study for the activity-mediated channel: the trial module — roughly seven participant-minutes plus central scoring, powered arithmetic stated in the dossier — detects whether re-routed income changes what recipients do in ways the culture counts. The channel it cannot reach in dispersed trials is the one the Austrian result suggests exists: standing moving through perception without activity change. That channel needs informant designs, which do not ride in dispersed samples; the module’s grammar does carry “recipient of unconditional income” as a profile attribute, which measures the norm-level standing of grant receipt — a different, valuable object, labeled as such. And the Round-4 proxy concession the parent carried resolves into an estimand: the within-sample association of conferred standing with earnings is the first designed measurement of the quantity the relative-income literature has been assuming — reported with its signaling decomposition, per the λ_w discipline above.

§ VIIWhat We Can Expect, I: The Shape of the Gap

Everything in this section and the next is projection. Each carries its trend, its assumption, and its falsifier; none borrows the authority of the measured record; and none requires a displacement forecast, because every projection is decidable in the present.

Projection 1 — the calibration gap exists and is structured. Trend: every adjacent felt/conferred pairing ever measured has come apart — popularity at .40 and declining with age, personality at .29–.51, status against its anchors at .32, dyadic knowledge of being liked at .17 among new acquaintances — and the one accurate case, face-to-face status, is accurate because accuracy there is enforced, while unenforced perceptions in the same groups drift (§4). Assumption: standing behaves like its adjacent evaluative domains — evaluative, external, informant-advantaged. Expectation: felt and conferred standing associate positively and imperfectly, with the joint distribution structured — tightest where contribution is publicly legible and feedback dense, loosest where feedback has thinned. Falsifier: the convergence tripwire (§6). If it trips, the paper concedes in a paragraph what it argued in fifty pages, and the concession is cheap for the field: the existing scales suffice.

Projection 2 — the baseline loading: the culture counts the wage. Trend: the conferred tradition has overwhelmingly rated occupations (§4); the near-miss items are employment-anchored or generic (§5); the deservingness literature’s reciprocity criterion polices earning; the one experimental datum available shows an institutional certificate moving standing (§5, at its demoted grade). Assumption: judges’ counting tracks institutionally legible certificates, of which the wage is the paradigm. Expectation: the conferral function weights waged profiles heavily at baseline — with the constitutive/signaling decomposition run before the number is read; with the desirability caveat (stated judgment plausibly inflates the counting of caregivers and volunteers, so the coordination arm, not the stated arm, carries the scoring); and, this round, with the triangle’s disambiguation attached, because a wage-heavy coordination result is consistent with two states of the world that have opposite implications for the parent: wage-heavy private judgment — the vacuum as diagnosed — or care-egalitarian private judgment trapped under a misperceived wage norm, which is fixable by norm-information intervention and needs no re-routing architecture at all. P2 is therefore scored on the coordination arm and read against the wedge: operative loading is the projection’s subject, and the wedge decides what kind of problem the loading is. Its disconfirmation — a culture already counting care and commons work at par in the operative norm — would be good news for the parent’s R2 and bad news for this paper’s account of the vacuum’s invisibility, and would be reported as exactly that; with one symmetric caveat the wedge forces, since the disambiguation cuts both ways: operative parity sitting on a wage-heavy private standard is unstable good news, an equilibrium the same norm-information dynamics could erode from beneath, so the good-news branch is read against the wedge exactly as the bad-news branch is. Falsifier: the factorial survey itself, scored on the coordination arm, after that arm’s own known-groups clearance.

Projection 3 — the felt proxy degrades exactly where the parent needs it not to. Trend: metaperception is self-anchored (r = .87 with self-view); calibration is manufactured by feedback, and the waged life is a feedback factory — evaluations, pay, title, the easy answer to “what do you do” — while the non-waged life thins every channel at once; joblessness hurts less where the reference group shares it (Clark 2003; the fixed-effects misery baseline in Winkelmann & Winkelmann 1998), evidence that standing perception tracks local reference structure; and the first live trial hint points the same way — cash improving material security while a felt-Importance subscale moved negative in one site (§5, Cambridge, flagged). Assumption: felt standing tracks conferred standing materially worse in non-waged strata. Expectation: the mattering scales prove adequate proxies among the employed and degrade among the displaced and non-waged — meaning the populations automation’s settlements most need to see are the ones self-report sees worst, and the trials now adding felt-mattering batteries are instrumenting the half of the construct that fails where it matters. Falsifier: uniform convergence across strata — invariance holds and the felt–conferred association is equal everywhere — in which case self-report suffices even for the displaced and the trial module simplifies to a questionnaire.

Projection 4 — the certificate confers before the activity does, rebuilt at its true grade. Trend: the Austrian split (§5): the status ladder rose for everyone covered by the guarantee before they worked a day and did not respond to employment onset; felt recognition rose with the work itself. Assumption, now carrying the confound the first draft omitted: the split reflects individual certification rather than the saturated design’s alternative reading — a community-level shift in the town’s norm field, with anticipation and selection embedded in staggered entry. The projection is that certification has an individual component: institutionally legible coverage — a recognized place held for one — confers standing partly independent of realized activity and partly independent of what the whole community knows. Expectation: recognition architectures work partly through certification, which is cheap and buildable, not only through activity, which is expensive. Falsifier, redesigned because the first draft’s could fail for design reasons while the hypothesis stood: individually randomized certification under community-saturation control — designs where coverage varies within a shared information environment (randomized waitlists whose status is privately known, or certification arms in dispersed trials via the module’s grant-receipt attribute at the norm level) — finding no coverage effect. Saturated-design nulls no longer count against P4, and saturated-design positives no longer count for it.

§ VIIIWhat We Can Expect, II: Can the Column Survive Being Official?

The constructive question and the darkest one are the same question run in two directions: what happens to this measure when institutions touch it.

The officialization path, read from the record. Statistical categories are built, not found — the unemployment rate was an act of construction, and the history of official numbers (Desrosières) is a history of contested categories hardening into furniture. The wellbeing precedent is sobering in a specific way: columns built, adopted, and mostly idle. The path is therefore stated modestly and in order: trial instrumentation first (the module, the cheapest decidability gain in the program); panel adoption second; official statistics last, if ever, and only behind validation. The claim is decidability, not influence: with the column, a wage-only settlement that empties the worth channel becomes visible. Whether anyone acts on the sight is a different variable, and the record leans pessimistic.

The Goodhart turn, at full strength — now in both directions. Suppose the column succeeded — watched like inflation. Every documented pathology of high-stakes measurement applies. Measures converted to targets degrade (Goodhart; Campbell); the felt-construct precedent ended in the self-esteem debacle; a standing metric feeding any allocation would summon contribution-theater, credentialed mattering, the gaming of legibility by exactly the populations the column was built to see. The answer is design, and the design is constitutive: distributions, never league tables; strata, never scores; no individual’s conferred standing computed for consequence, returned, or released; administrative reuse barred contractually; any future officialization bound to a written Goodhart audit. §3 supplied the deeper reason the rule is constitutive rather than prudential: esteem dispensed by formula is forfeited — a gated column would not merely corrupt the measurement, it would destroy the good being measured. Only an un-gated measure measures anything. And this revision adds the direction the first draft missed: publication is itself an intervention. A headline wage-loading becomes a political fact — quoted, campaigned on, taught — and later waves of judges are downstream of it, which threatens the drift time series §10 sells as a feature. The design response is stated at its honest size: the estimation protocol is fixed across waves, drift analysis is pre-registered against a holdout attribute set never published, and exposure sensitivity is reported where measurable — a firewall, not an immunity, and the caveat travels with every drift estimate.

The dark genealogy, faced — and one reclassification owed. “The state measures who counts as contributing” is a sentence with a history. The social-credit architectures document what administered standing becomes at state scale; the welfare-conditionality record documents deservingness testing at case-worker scale — sanction, stigma, the audit of the poor; credit scoring documents the commercial version, reputation hardened into a gate. These are this paper’s inversion — person-level scores, administratively reused, gating benefits — and the line between them and it is the no-administration rule, held by custody design: statistical agencies under disclosure rules, not benefit agencies under budget rules; research custody under review boards until then; the dual-use concession stated (publishing the specification lowers the cost of building the inverted version). The reclassification: the first draft filed the deservingness literature here, among the administrative practices, and its referee was right to object — van Oorschot’s tradition is a measurement literature, this paper’s nearest methodological kin, and it now stands in §5 where it belongs. What remains in the genealogy is deservingness administered — the conditionality apparatus — which is precisely the distinction the no-administration rule exists to hold.

Projection 5 — the gated column corrupts; the audit-grade column can survive. Trend: the Goodhart/Campbell record; the self-esteem precedent; the conditionality genealogy; the teleological paradox. Assumption: standing metrics behave like other high-stakes measures under gating, and audit-grade custody — not the construct’s virtue — is what preserves validity. Expectation: any administered variant degrades measurably against informant ground truth; the un-gated variant, custody held, retains validity within the publication-endogeneity limits stated above. Falsifier: a gated recognition system that stays calibrated under independent audit — which would weaken the necessity of the no-administration rule while leaving its ethics standing on the genealogy.

§ IXThe Ledger: What We Claim and What We Do Not

In the house manner, the register audit — regenerated at Round 3, not patched, since a ledger that has gone stale against the sections it audits is worse than none.

What recognition is (Register C, §3): the praise/praiseworthiness distinction with its tribunal discipline — aggregated actual spectators recover praise, and desert is approached by qualifying judges, not counting them; the three-way construction — felt, conferred, deserved — labeled the paper’s own, verbs policed, with the desert arm carrying its approximation status; Weber’s status honor as the construct’s sociological name; the Durkheimian warrant including its behavioral clause; the Rawlsian location of the good outside the respondent; the Honneth–Taylor categorical reading, with non-scarcity held as threshold semantics if the latent structure runs continuous; the Akerlof–Kranton microfoundation; the Brennan–Pettit economics and the Bénabou–Tirole inference warning that disciplines every judgment weight.

What the instruments show (Register E, §§4–5): the felt tradition’s reliability record and its validation circularity, stated at search-grade — we found no informant validation, protocol in the corpus index — never as omniscience; the felt-positional cell’s real predictive power; the conferred-positional apparatus as feasibility proof and wrong tool, its universal trimmed (the homemaker was rated); the divergence record at r ≈ .2–.5 with its enforcement mechanism; the adjacent traditions at their true distances — deservingness-of-support, just rewards, interactional status construction, key-worker esteem — machinery and stance inherited, object unoccupied; the deflation in full, including the two inheritances correctly labeled (factorial-survey fragilities; anchoring-vignette calibration only) and the stated-versus-operative gap with its against-the-thesis bias direction; the fourteen-family sweep, scoped, with watch items; the corrected trial frontier including the refugee-camp experiment; and the Austrian result at its demoted, saturation-confounded, verification-flagged grade.

What can be expected (Register P, §§7–8): five projections, each trend/assumption/falsifier, none requiring a displacement forecast; P2 scored on the coordination arm and read against the norm wedge on both branches, its good news as conditional as its bad; P4 rebuilt so saturated designs neither confirm nor disconfirm it.

What is framework (§6): the 2×2; the quantities, with the joint distribution as estimand and the calibration gap as exposition; the national/operative-field distinction, and the level-indexed constitutive ruling that resolves it — coordination arm constitutive at the norm level, informant aggregate at the person level; six instruments (the factorial-survey spectator, the private-elicitation and coordination arms completing the triangle, the informant design, the linked design, the desert arm) plus the vignette-imputation mechanism that scales; the partitioned validation program — validity matrix for the convergence-expected pairs, estimand set for the divergence-meaningful ones — with its behavioral channels; the three falsifiers with the rebuilt, profile-level incoherence test and its named middle verdict — labeled framework throughout, shipped with dossier v1.3.

What the paper does not claim. No claim to measure desert — the desert arm approximates it under stated idealizations, and worthy appears in no result. No claim that the index should gate any benefit, ever. No claim that self-report is worthless — three cells are its property. No claim that the culture’s counting is correct — the instrument records what is counted; diagnosis, not endorsement. No claim that the column changes policy — decidability only. No claim of cross-cultural validity — domestic first, invariance tested. No claim that the national conferral function is anyone’s operative field — the stratum-target coordination function carries the operative norm and the informant aggregate the operative person-level counting, the constitutive ruling being level-indexed and not a single carrier. No claim that the trial module reaches the perception channel — activity-mediated only, stated in tripwire type. No claim of method novelty — the factorial survey is inherited, the deservingness tradition is kin, and the paper’s originality is its object, its gate, its statistics, and its linkage. No claim that coordination results reveal private judgment — the coordination arm measures the perceived norm, constitutive by ruling, and the wedge between private and perceived is an estimand, not an embarrassment. No claim that instructed idealization is deliberation — the desert arm carries its benchmark and its Firth sentence. No prophecy. The paper claims an empty cell at its resized dimensions; a fillable cell, staked on stated tests including cross-stratum agreement; a structured gap, extrapolated from every adjacent pairing; and a rule — measure the counting, never the worth — derived from the construct itself.

§ XImplications: The Column, Specified

The house tradition ends its arguments with what to measure. This paper is that ending run to specification, so its implications can be a build sheet — with each item now carrying its grade.

For the statistical agencies, eventually, and the panels, sooner: the column is four statistics and a function — the conferral function C(p), national for the accounts and stratum-specific for the fields people live in, re-estimated at intervals with the publication-endogeneity caveat attached to every drift estimate; the conferral share κ, with its threshold rules and ceiling priors pre-registered; the wage-loading λ_w, reported only with its constitutive/signaling decomposition; the dispersion of conferred standing on its own bounded scale; and the joint felt/conferred distribution by stratum. These would make a wage-only settlement’s failure mode visible while it forms — in the channel the instruments can see.

For the trials, now, at the module’s true grade: roughly seven participant-minutes plus central judge-panel scoring buys the felt battery, the activity profile, and imputed conferral — which detects the activity-mediated component of the worth channel and is structurally blind to standing that moves through community perception without activity change. The blindness is not a footnote; it is the reason the informant design exists, and dispersed trials cannot carry informant designs. What a flagship-scale trial that adopts the module converts is therefore No. 008's first falsifier for the channel the module sees — a real and decidable question, sold at exactly that size. The grant-receipt attribute in the grammar adds the norm-level standing of receiving unconditional income — a different object, worth having, labeled. The job-guarantee evaluations get the sharper offer: bounded sites can carry the informant design, and the Austrian team’s successor studies — individually randomized certification under saturation control — are the cleanest test of P4 on the table.

For the mattering tradition: the validation this paper specifies is the one the construct has deserved for forty-five years — felt mattering benchmarked against conferred regard. Whatever the result, the tradition wins: convergence crowns the existing scales; divergence gives the field its next twenty years.

And the refusal, in the house tradition: no policy menu. The column evaluates the menus others write. The paper does not say what a society should count as contribution — that is R2's question, and it belongs to the parent thread’s successors and to politics. It says that the counting is real, that it is measurable from the side where it lives, that the traditions nearest it have measured everything around it but not it, and that every future argument about work, worth, and the settlements of automation will be conducted either with this column or in the dark next to it. One binding travels with the build sheet: officialization carries a written Goodhart audit first, and the no-administration rule travels with the instrument into every custody it enters.

§ XIA Research Program

This paper is one move in a longer argument, conducted in the open, and the move is deliberately unglamorous: between the parent’s framework and its falsifiers stood a missing instrument, and the archive’s response is to specify the instrument rather than lower the bar. The specification is complete in the sense that matters — a competent lab could field the factorial survey, the norm-elicitation arm, the informant study, the linked design, and the desert arm from the dossier alone, thresholds and all — and incomplete in the sense that also matters: it is a specification, and the no-reification rule holds until the data exist. The revision has made the paper smaller and harder to kill: the vacuum resized to its true dimensions, the estimand rebuilt on the joint distribution, the incoherence test moved to where incoherence would actually live, the module sold at its grade.

The invitations, named — and grown by a round of refereeing. To the psychometricians: the MTMM matrix, the invariance ladder, and the IRT calibration are yours; the paper borrowed your machinery and has already taken corrections with thanks. To the deservingness tradition: you have been this paper’s nearest kin for thirty years, and the gate item is close to your reciprocity criterion pointed at membership instead of support — the crossing is offered as collaboration, not competition. To the factorial-survey and justice-function researchers: C(p) is your method aimed at a new object; Jasso’s function for what people should earn has waited fifty years for its sibling, what people are counted as giving. To the norm-elicitation economists: the coordination arm is your instrument carrying this paper’s construct — and the misperception economists behind the wedge (the Bursztyn program) have, in this design, a new object on which private judgment and perceived norm may famously disagree. To the deliberative-polling tradition: the desert arm’s benchmark is a mini-public, and the question it would deliberate — what ought to count as contribution — may be the most consequential one the method has not yet been handed. To the mattering tradition: the other-report benchmark is offered as completion, not critique. To the sociometrists: your method, pointed at contribution in adult populations, carries the operative-field payoff at the person level — where, in a bounded community, your informants are the field itself; nothing here works without it. To the experimentalists — the Penn center whose mattering battery is already in the field, the flagship teams whose next waves could carry seven more minutes, the refugee-camp team that separated work from cash experimentally, and the Austrian evaluators whose successor designs could settle P4: the module is written for you. To the statistical agencies: the European survey that carried the nearest item in existence fields again in autumn 2026 — keep the item; the column is drafted for the day the validation record warrants it. And to the critics of measurement: §8 was written for you, the rule it derives is constitutive, and the paper is safer with you watching.

The thread placement, last. This paper belongs to Motivation because standing is the motive Smith placed at the center — the thread’s wager has been that motivation is a substrate to be engineered, and an engineering discipline that cannot measure its target variable is not yet a discipline. It is cross-listed to Moral Judgment because the instrument is the spectator, externalized: Nos. 006 and 007 argued that algorithmic markets need spectatorship built into them, and this paper builds the spectator as a survey — judges sampled where Smith’s are imagined, qualified where his are idealized, with the distance between the two stated as the distance between what a culture counts and what it would count on reflection. Smith needed two words for what one feeling and one fact kept separate. This paper maps the relation between the feeling and the fact, and specifies the arm that reaches from the fact toward the second word — the praise a person carries, the counting a society performs, and the reflective verdict neither of them settles: written down, all three, precisely enough to be wrong about.


Notes

On the naming of the falsifiers. This paper's three falsifiers — convergence, incoherence, positional collapse — are claims about the instrument. The parent paper's substitution falsifier, a claim about cash, is written throughout as "No. 008-F1," and the trial bolt-on that would make it runnable is "the No. 008-F1 module." The disambiguation is fixed for the corpus. [§§1, 6, 10]

On the level-index. The ruling that the coordination arm is constitutive holds at the norm level, where coercion is anticipatory and runs on beliefs about the standard. At the person level — one person's realized counting in a bounded community — the informant first-order aggregate is constitutive, because the informants are the field: their verdicts both constitute the standing and administer its sanctions. The linked design therefore pairs felt standing with the informant measure by design. No unconditional "the coordination arm is constitutive" sentence stands anywhere in this paper or its artifacts without its level index. [§6]

On the search behind a universal negative. The claim that no mattering instrument has been validated against an other-report criterion is stated at search grade — we found none — not as omniscience. The search covered informant, partner, dyadic, round-robin, and peer-nomination pairings across the Rosenberg–McCullough, Elliott, Flett, Prilleltensky, France–Finney, and Jung–Heppner lineages; the protocol is filed in the corpus index. One dyadic near-miss (Mak and Marshall, 2004) was examined and found to survey only the perceiver; that design detail is inferred from the abstract and secondary descriptions, and carries a verification flag pending full-text confirmation. [§4]

On the sweep behind the empty cell. The claim that no official statistical agency or major survey programme carries a conferred-categorical measure is scoped to what was swept: fourteen instrument families, as of mid-2026, enumerated in §5. The near-miss field is named there rather than waved at, because a universal negative asserted without its scope is not a finding but a posture. Watch items: the Eurofound successor survey fields in autumn 2026, retention of its recognition item unknown; the Wales quantitative report and two American guaranteed-income full reports remained unlocated at the time of writing. [§5]

On figures carrying verification flags. The following are flagged at imprint and are to be discharged against primary sources before external submission: the Glasgow-edition paragraph references beyond TMS III.ii.1; the wording of Rawls §67; the MAGMA decomposition figures, pending the published AEJ: Policy version; the Cambridge Importance-subscale result; the informant-agreement figure δ = −.038; the Nilson and Bose homemaker-prestige citations; and the canon of the key-worker esteem literature. [§§3, 4, 5]

On the dossier. The operational form of §6 — item pools, the profile grammar, sampling designs, the multitrait–multimethod partition, invariance and known-groups thresholds, the list-experiment and coordination arms, the desert arm's mini-public benchmark, power arithmetic for the trial module, and the custody protocol implementing the no-administration rule — ships as a supplementary artifact at pre-registration grade, CC-BY, item pools included. Nothing in this paper is fielded by the imprint; the no-reification rule holds until the data exist. [§§6, 10, 11]


References

Adler, N. E., Epel, E. S., Castellazzo, G., & Ickovics, J. R. (2000). Relationship of subjective and objective social status. Health Psychology, 19(6).

Akerlof, G. A., & Kranton, R. E. (2000). Economics and identity. Quarterly Journal of Economics, 115(3).

Anderson, C., Ames, D. R., & Gosling, S. D. (2008). Punishing hubris. Organizational Behavior and Human Decision Processes, 107.

Anderson, C., Hildreth, J. A. D., & Howland, L. (2015). Is the desire for status a fundamental human motive? Psychological Bulletin, 141(3).

Anderson, C., Srivastava, S., Beer, J. S., Spataro, S. E., & Chatman, J. A. (2006). Knowing your place: Self-perceptions of status in face-to-face groups. Journal of Personality and Social Psychology, 91(6).

Auspurg, K., & Hinz, T. (2015). Factorial Survey Experiments. Sage.

Bago d'Uva, T., Lindeboom, M., O'Donnell, O., & van Doorslaer, E. (2011). Education-related inequity in healthcare with heterogeneous reporting of health. Journal of Human Resources, 46(4).

Baumeister, R. F., Campbell, J. D., Krueger, J. I., & Vohs, K. D. (2003). Does high self-esteem cause better performance? Psychological Science in the Public Interest, 4(1).

Bénabou, R., & Tirole, J. (2006). Incentives and prosocial behavior. American Economic Review, 96(5).

Blair, G., & Imai, K. (2012). Statistical analysis of list experiments. Political Analysis, 20(1).

Bohmann, S., Fiedler, S., Kasy, M., Schupp, J., & Schwerter, F. (2026). Experimental evaluation of a basic income pilot in Germany. Working papers. [DIW Wochenbericht 15/2025]

Boothby, E. J., Cooney, G., Sandstrom, G. M., & Clark, M. S. (2018). The liking gap in conversations. Psychological Science, 29.

Bose, C. E. (1980). Social status of the homemaker. [flagged for verification]

Brennan, G., & Pettit, P. (2004). The Economy of Esteem. Oxford University Press.

Bursztyn, L., González, A. L., & Yanagizawa-Drott, D. (2020). Misperceived social norms: Women working outside the home in Saudi Arabia. American Economic Review, 110(10).

Campbell, D. T., & Fiske, D. W. (1959). Convergent and discriminant validation by the multitrait-multimethod matrix. Psychological Bulletin, 56(2).

Card, D., Mas, A., Moretti, E., & Saez, E. (2012). Inequality at work: The effect of peer salaries on job satisfaction. American Economic Review, 102(6).

Castro, A., & West, S. (2023). Journal of Urban Health, 100(2).

Chen, F. F. (2007). Sensitivity of goodness of fit indexes to lack of measurement invariance. Structural Equation Modeling, 14(3).

Cheung, G. W., & Rensvold, R. B. (2002). Evaluating goodness-of-fit indexes for testing measurement invariance. Structural Equation Modeling, 9(2).

Cillessen, A. H. N., & Marks, P. E. L. (2011). Conceptualizing and measuring popularity. In Popularity in the Peer System. Guilford.

Cillessen, A. H. N., & Mayeux, L. (2004). From censure to reinforcement: Developmental changes in the association between aggression and social status. Child Development, 75(1).

Cillessen, A. H. N., & Rose, A. J. (2005). Understanding popularity in the peer system. Current Directions in Psychological Science, 14(2).

Citron, D. K., & Pasquale, F. (2014). The scored society. Washington Law Review, 89.

Clark, A. E. (2003). Unemployment as a social norm. Journal of Labor Economics, 21(2).

Connelly, B. S., & Ones, D. S. (2010). An other perspective on personality. Psychological Bulletin, 136(6). [per-trait values flagged]

Creemers, R. (2018). China's social credit system: An evolving practice of control. SSRN.

Cronbach, L. J., & Furby, L. (1970). How we should measure "change" — or should we? Psychological Bulletin, 74(1).

Cronbach, L. J., & Meehl, P. E. (1955). Construct validity in psychological tests. Psychological Bulletin, 52.

Cundiff, J. M., & Matthews, K. A. (2017). Is subjective social status a unique correlate of physical health? Health Psychology, 36(12).

Desrosières, A. (1998). The Politics of Large Numbers. Harvard University Press.

Durkheim, É. (1895). The Rules of Sociological Method; (1897). Suicide.

Eid, M., Lischetzke, T., Nussbeck, F. W., & Trierweiler, L. I. (2003). Separating trait effects from trait-specific method effects in multitrait-multimethod models. Psychological Methods, 8(1).

Eisenkraft, N., Elfenbein, H. A., & Kopelman, S. (2017). We know who likes us, but not who competes against us. Psychological Science, 28.

Elliott, G., Kao, S., & Grant, A.-M. (2004). Mattering: Empirical validation of a social-psychological concept. Self and Identity, 3(4). [subscale alphas flagged]

Eurofound. (2012; 2016). European Quality of Life Survey — item Q29g. [successor wave fields autumn 2026]

Finkelstein, A., & Notowidigdo, M. J. (2019). Take-up and targeting: Experimental evidence from SNAP. Quarterly Journal of Economics, 134(3).

Firth, R. (1952). Ethical absolutism and the ideal observer. Philosophy and Phenomenological Research, 12(3).

Fishkin, J. S. (2009). When the People Speak. Oxford University Press.

Flett, G. L. (2018). The Psychology of Mattering. Academic Press.

Flett, G. L., Nepon, T., Goldberg, J. O., Rose, A. L., Atkey, S. K., & Zaki-Azat, J. (2022). The Anti-Mattering Scale. Journal of Psychoeducational Assessment, 40(1).

Fourcade, M., & Healy, K. (2017). Seeing like a market. Socio-Economic Review, 15(1).

France, M. K., & Finney, S. J. (2009). What matters in the measurement of mattering? Measurement and Evaluation in Counseling and Development, 42(2).

Ganzeboom, H. B. G., et al. The International Socio-Economic Index of occupational status (ISEI).

Grol-Prokopczyk, H., Verdes-Tennant, E., McEniry, M., & Ispány, M. (2015). Promises and pitfalls of anchoring vignettes in health survey research. Demography, 52(5).

Hacking, I. (1999). The Social Construction of What? Harvard University Press.

Hackman, J. R., & Oldham, G. R. (1980). Work Redesign. Addison-Wesley.

Hainmueller, J., Hopkins, D. J., & Yamamoto, T. (2014). Causal inference in conjoint analysis. Political Analysis, 22(1).

Harper, et al. (2026). How well do we know how others see us? A systematic review of meta-accuracy. Journal of Personality.

Honneth, A. (1995). The Struggle for Recognition. Polity/MIT Press.

Hopkins, D. J., & King, G. (2010). Improving anchoring vignettes. Public Opinion Quarterly, 74(2).

Hussam, R., Kelley, E. M., Lane, G., & Zahra, F. (2022). The psychosocial value of employment: Evidence from a refugee camp. American Economic Review, 112(11).

Jasso, G., & Rossi, P. H. (1977). Distributive justice and earned income. American Sociological Review, 42.

Jung, A.-K., & Heppner, M. J. (2017). Development and validation of a Work Mattering Scale. Journal of Career Assessment, 25(3).

Kangas, O., et al. (2020). Evaluation of the Finnish Basic Income Experiment. Ministry of Social Affairs and Health, 2020:15.

Kapteyn, A., Smith, J. P., & van Soest, A. (2007). Vignettes and self-reports of work disability in the United States and the Netherlands. American Economic Review, 97(1).

Kasy, M., & Lehner, L. (2026). Employing the unemployed of Marienthal: Evaluation of a guaranteed job program. Forthcoming, American Economic Journal: Economic Policy. AEARCTR-0006706. [figures flagged pending publication]

Kenny, D. A., & DePaulo, B. M. (1993). Do people know how others view them? Psychological Bulletin, 114(1).

Kenny, D. A., & West, T. V. (2010). Similarity and agreement in self- and other perception. Personality and Social Psychology Review, 14.

Keyes, C. L. M. Social contribution subscale, social well-being. [MIDUS]

Kim, H., Di Domenico, S. I., & Connelly, B. S. (2019). Self–other agreement in personality reports. Psychological Science, 30(1). [δ flagged]

King, G., Murray, C. J. L., Salomon, J. A., & Tandon, A. (2004). Enhancing the validity and cross-cultural comparability of measurement in survey research. American Political Science Review, 98(1).

King, G., & Wand, J. (2007). Comparing incomparable survey responses. Political Analysis, 15(1).

Krupka, E. L., & Weber, R. A. (2013). Identifying social norms using coordination games. Journal of the European Economic Association, 11(3).

Lamont, M. (2000). The Dignity of Working Men. Harvard University Press / Russell Sage.

Lauer, J. (2017). Creditworthy. Columbia University Press.

Luttmer, E. F. P. (2005). Neighbors as negatives: Relative earnings and well-being. Quarterly Journal of Economics, 120(3).

Mak, L., & Marshall, S. K. (2004). Perceived mattering in young adults' romantic relationships. Journal of Social and Personal Relationships, 21(4). [design detail flagged]

Marshall, S. K. (2001). Do I matter? Journal of Adolescence, 24(4).

Meredith, W. (1993). Measurement invariance, factor analysis and factorial invariance. Psychometrika, 58.

Moffitt, R. (1983). An economic model of welfare stigma. American Economic Review, 73(5).

Nilson, L. B. (1978). The social standing of a housewife. Journal of Marriage and the Family. [flagged for verification]

Noble, K. G., et al. (2025). Baby's First Years, four-year follow-up. NBER 33844.

OECD. (2024). How's Life? 2024.

Office for National Statistics. (2011– ). ONS4 personal well-being measures; national well-being dashboard (updated February 2026).

van Oorschot, W. (2000). Who should get what, and why? Policy & Politics, 28(1).

Parkhurst, J. T., & Hopmeyer, A. (1998). Sociometric popularity and peer-perceived popularity. Journal of Early Adolescence, 18(2).

Paul, M., Darity, W., & Hamilton, D. On a federal job guarantee.

Prilleltensky, I. See Scarpa, Zopluoglu, & Prilleltensky.

Rawls, J. (1971; rev. 1999). A Theory of Justice, §67. Belknap/Harvard. [wording flagged]

Ridgeway, C. L. Status construction theory and the expectation-states program.

Rosenberg, M., & McCullough, B. C. (1981). Mattering: Inferred significance and mental health among adolescents. Research in Community and Mental Health, 2.

Rossi, P. H., & Anderson, A. B. (1982). The factorial survey approach. In Measuring Social Judgments.

Ryff, C. D. (1989). Happiness is everything, or is it? Journal of Personality and Social Psychology, 57(6).

Scarpa, M. P., Zopluoglu, C., & Prilleltensky, I. (2022). Assessing multidimensional mattering (MIDLS). Journal of Community Psychology, 50(3).

Schieman, S., & Taylor, J. (2001). Statuses, roles, and the sense of mattering. Sociological Perspectives, 44(4).

Singh-Manoux, A., Marmot, M. G., & Adler, N. E. (2005). Does subjective social status predict health and change in health status better than objective status? Psychosomatic Medicine, 67.

Smith, A. (1759). The Theory of Moral Sentiments, Glasgow Edition (Raphael & Macfie, 1976). [paragraph references beyond III.ii.1 flagged]

Smith, T. W., & Son, J. (2014). Measuring occupational prestige on the 2012 General Social Survey. GSS Methodological Report 122.

van Soest, A., Delaney, L., Harmon, C., Kapteyn, A., & Smith, J. P. (2011). Validating the use of anchoring vignettes for the correction of response scale differences. Journal of the Royal Statistical Society: Series A, 174(3).

Stiglitz, J. E., Sen, A., & Fitoussi, J.-P. (2009). Report by the Commission on the Measurement of Economic Performance and Social Progress.

Tan, J. J. X., Kraus, M. W., Carpenter, N. C., & Adler, N. E. (2020). The association between objective and subjective socioeconomic standing and subjective well-being. Psychological Bulletin, 146(11).

Taylor, C. (1994). The politics of recognition. In Multiculturalism. Princeton University Press.

Taylor, J., & Turner, R. J. (2001). A longitudinal study of the role and significance of mattering. Journal of Health and Social Behavior, 42(3).

Tcherneva, P. R. (2020). The Case for a Job Guarantee. Polity.

Treiman, D. J. (1977). Occupational Prestige in Comparative Perspective. Academic Press.

Vazire, S. (2010). Who knows what about a person? The self–other knowledge asymmetry (SOKA) model. Journal of Personality and Social Psychology, 98(2).

Vivalt, E., Bartik, A., Miller, S., Broockman, D., & Rhodes, E. (2024–26). The OpenResearch Unconditional Income Study. NBER 32719, 32711, 32784; employment paper accepted, Quarterly Journal of Economics.

Watts, B., & Fitzpatrick, S. (2018). Welfare Conditionality. Routledge.

Weber, M. Class, status, party. In Economy and Society.

West, S., & Castro, A. (2021). Stockton Economic Empowerment Demonstration, first-year report.

Winkelmann, L., & Winkelmann, R. (1998). Why are the unemployed so unhappy? Economica, 65(257).

· · ·

Companion papers in the archive: No. 002, Beyond the Binary (the reward architecture this paper's target variable serves); No. 006, Smith's Operating System and No. 007, The Spectator and the Agent (the spectator as formal role, here built as survey method); No. 008, The Wage and the Worth (the parent, whose measurement debt this paper pays).

The 17-slide companion deck — The Praise and the Praiseworthiness — sits beside this paper; the instrument dossier (v1.3), corpus index, claim outline, framework set, four referee reports and their response letters, and the five prior manuscript versions are filed under cowork. The dossier is offered for adoption rather than submission: the factorial-survey spectator, the elicitation triangle, the informant design, and the desert arm can be fielded from it as written. Comments — and corrections to the flagged citations — welcome at contactme​@​marshallcahill.com.

APA
Cahill, M. (2026). The Praise and the Praiseworthiness: Standing as a social fact, the worth index, and the column the accounts cannot yet read. Armchair Scholar Working Papers, No. 009.
BibTeX
@techreport{armchair-scholar-009,
  author      = {Cahill, Marshall},
  title       = {The Praise and the Praiseworthiness: Standing as a social fact, the worth index, and the column the accounts cannot yet read},
  institution = {Armchair Scholar},
  number      = {009},
  year        = {2026},
  month       = {July},
  type        = {Working Paper},
  pages       = {48}
}

Slides · accompanying lecture

2026 · 07

References