EconBase
← Back to paper

The econometrics of happiness: Are we underestimating the returns to education and income?

Extracted main text — title through conclusion, appendix excluded. This is what our citation measures are computed over, published so the extraction can be checked by eye.

143,510 characters · 25 sections · 52 citation commands

Rendered from LaTeX for readability, not typeset faithfully. Citation keys are highlighted; maths is left as source; figures, tables and equation environments are summarised rather than reproduced; unrecognised commands are greyed out so nothing is silently dropped. Email addresses are removed.

The econometrics of happiness: Are we underestimating the returns to education and income?

\doparttoc \dopartlof \dopartlot

abstractThis paper describes a fundamental and empirically conspicuous problem inherent to surveys of human feelings and opinions in which subjective responses are elicited on numerical scales. The paper also proposes a solution. The problem is a tendency by some individuals --- particularly those with low levels of education --- to simplify the response scale by considering only a subset of possible responses such as the lowest, middle, and highest. In principle, this “focal value rounding” (FVR) behavior renders invalid even the weak ordinality assumption often used in analysis of such data. With “happiness” or life satisfaction data as an example, descriptive methods and a multinomial logit model both show that the effect is large and that education and, to a lesser extent, income level are predictors of FVR behavior. A model simultaneously accounting for the underlying wellbeing and for the degree of FVR is able to estimate the latent subjective wellbeing, i.e. the counterfactual full-scale responses for all respondents, the biases associated with traditional estimates, and the fraction of respondents who exhibit FVR. Addressing this problem helps to resolve a longstanding puzzle in the life satisfaction literature, namely that the returns to education, after adjusting for income, appear to be small or negative. Due to the same econometric problem, the marginal utility of income in a subjective wellbeing sense has been consistently underestimated.\\ { {\sc keywords:} happiness; subjective wellbeing; life satisfaction; modeling; income; education; welfare}

\listoffigures \listoftables \pagestyle{fancy} \fancyhead[LO,RE]{Econometrics of happiness ({\em J. Public Econ}, 2024)} \fancyhead[CO,CE]

Introduction

Now firmly entrenched in the economics literature, in national statistical agency data collection, and in the dialogue about progress and wellbeing, survey-based subjective evaluations of life\footnote{The life satisfaction question and close cousins such as the Cantril Ladder question are posed in numerous national and international social surveys, both cross-sectional and panel. The U.K. Treasury's Green Book includes instructions for how to use compensating differentials, estimated from life satisfaction, to carry out cost/benefit calculations for central government UK-Treasury-2021-GreenBook-supplement-wellbeing,MacLennan-Stead-Rowlat-UK-Treasury-2021-life-satisfaction-approach. } are the basis for estimating welfare benefits and costs of everything from inflation and unemployment, to air pollution and being married Blanchflower-et-al-JMCB2014-unemployment-inflation,Levinson-JPubE2012-happiness-pollution,Stutzer-Frey-JSE2006-marriage-causality-happiness. Estimates of the psychological benefit of increased income, using this approach, are five decades old, and those evaluating the net individual return of additional education have been carried out for at least three decades. In terms of optimally allocating human resources, not much could be more central than knowing the marginal utility of income and of education.

Responding to life evaluation questions

However, the coherence and value of subjective evaluations of life rely on a series of considerable cognitive tasks to be performed in short order by the respondent. When asked,\footnote{Typically, cognitive evaluations of life consist of a single, subjective, quantitative question like this one, and responses are used directly as a cardinal or ordinal proxy for welfare, i.e., “experienced” utility Easterlin-1974.} “Overall, how satisfied are you with life as a whole these days, measured on a scale of 0 to 10?” a respondent must in some sense

inparaenum[(i)] • conceive of the domains, expectations, aspirations or other criteria salient to her sense of experienced life quality or satisfaction; • assemble evidence pertaining to each ideal, such as recent affective (emotional) states, significant events, and objective outcomes; • appropriately weight and aggregate this evidence according to its importance to overall life quality, and • project the result onto the discrete numerical scale specified in the question.

This is without doubt a tall order, and any embrace of subjective wellbeing (\hypertarget{defSWB}{SWB}) data, and especially the headline measure of life satisfaction (\hypertarget{defSWL}{\hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace}), rests on their remarkable reproducibility and apparent cardinal comparability, possibly along with the principle that any objective indicator of experienced wellbeing must ultimately be accountable to a subjective one. While various studies have sought, with limited success, to find differences in interpretation of the \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace question or norms of expression across cultures and languages Helliwell-Barrington-Leigh-Harris-Huang-2010,Exton-Smith-Vandendriessche-OECD2015,Lau-Cummins-Mcpherson-SIR2005-cultural-bias-SWB,Clark-Etile-Postel-Vinay-Senik-VanderStraeten-EJ2005-latent-class-coefficients-SWB, an important fact is that, uniformly across cultures, responding to the question is cognitively demanding. This paper focuses specifically on the consequences of an apparent heterogeneity across respondents in their ease with the final, quantitative step in the process outlined above.

The crux is that people with less facility with numbers may simplify the numerical response scale for themselves. In particular, the evidence below shows that some respondents restrict the set of numerical options under consideration to a three-point scale consisting of the bottom, middle, and top options, rather than the full set offered. This can be expected to introduce complex biases in mean life satisfaction and in estimated marginal effects on life satisfaction, in particular with respect to education and other correlates of numerical literacy itself.

I present evidence of the prevalence and quantitative significance of this problem, with implications for the interpretation and analysis of all numerical, subjective response scales. The language and empirical examples all focus on the case of single-item \pdftooltip{\hyperlink{defSWB}{SWB}}{Subjective Well-Being (See page (ref))}\xspace questions, mostly \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace Cheung-Lucas-QoLR2014-single-item-SWL-vs-SWLS-validity, which underlie the field of the “economics of happiness”. While most empirical studies make use of a cardinal interpretation of the response scale in the life satisfaction question, and at least an ordinality assumption is universal,\footnote{A common finding is that models assuming cardinality give similar results to those assuming only ordinality Ferrer-i-Carbonell-Frijters-EJ2004.} the “focal value rounding” (\hypertarget{defFVR}{FVR}) behavior, described above, introduces a conspicuous violation of the ordinality of response options. Because a number of governments are gearing up to carry out cost/benefit analyses using regressions of \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace data for budgeting and program evaluation Frijters-Clark-Krekel-Layard-BPP2020-Happy-Choice-SWB-as-goal-for-government,Frijters-Krekel-2021-SWB-policy-handbook,happiness-research-institute-2020-WALYs,Grimes-2021-chapter-budgeting-for-wellbeing,FinanceCanada-QoL-framework-2021-April,UK-Treasury-2021-GreenBook-supplement-wellbeing, proper econometric accounting for \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace may have practical importance.

In order to estimate the size of systematic biases associated with widely used methods of inference based on \pdftooltip{\hyperlink{defSWB}{SWB}}{Subjective Well-Being (See page (ref))}\xspace reports, I present a model which accounts for the \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace phenomenon and which shows why biases on estimates can be large or small and positive or negative. The model also quantifies the fraction of respondents in a sample who have chosen an alternate, simplified response scale, a value I call the Focal Value Rounding Index, or \hypertarget{defFVRI}{FVRI}.

The rest of this paper proceeds as follows. The remainder of the Introduction reviews some stylized facts related to education and wellbeing in the happiness literature, and mentions some points of history in the development of \pdftooltip{\hyperlink{defSWB}{SWB}}{Subjective Well-Being (See page (ref))}\xspace survey questions like \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace. Next, \secref{Motivational-evidence} will convince the reader that there is a measurement problem with quantitative, subjective scales like \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace that is conspicuous, ubiquitous, and strongly correlated with educational attainment and that it has a natural explanation supported by the behavioral evidence. Then \secref{Cognitive-model} presents the formal model in which a mixture of high- and low-numeracy respondents treat the response scale differently. \secref{simulation} validates the estimation and identification approach using synthetic data and explores the complexity of biases that can result from \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace. \secref{empirical-estimates} presents the main empirical estimates of the relationship between education, income, and wellbeing, using a large social survey from Canada. \secref{Applications} reexamines three previously published studies, along with a ranking of U.S. states, as applications to investigate the extent of bias in existing published literature as well as in popular happiness rankings. In these empirical applications, previously anomalous but reproducible findings include evidence that a disadvantaged population reports high life satisfaction, and that the return to extra years of education after primary school are negative, especially when conditioned on income. These surprising findings are overturned when taking into account focal response behavior. A summary and a perspective on future directions are in \secref{Conclusion}.

Effects of education and income on subjective wellbeing

Education may be expected to confer welfare benefits not just through higher income but also through better health behaviors and enhanced social capital of various forms with intrinsic benefit Helliwell-Putnam-EEJ2007,Powdthavee-Lekfuangfu-Wooden-JBEE2015-education-SWL-Australia-HILDA, as well as through some kind of psychological capital which captures intrinsic benefits of learning or knowledge, or which complements other consumption (for instance, possibly literature, fine art, or the night sky). However, among the more surprising stylized facts in the economics of happiness is that formal education does not help much to explain \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace once income\footnote{Studies routinely control for current income rather than wealth when modeling \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace. Current income may be an especially poor proxy for lifetime income in this context because choosing to pursue extra education entails a trade-off between short-term income and future income. This leaves educational attainment as a positive proxy for unmeasured future income expectations, making the low coefficients measured on education even more surprising.} is accounted for Layard-2011-happiness-lessons-new-science,Frijters-Krekel-2021-SWB-policy-handbook.\footnote{This generalization hides considerable variation in the literature. Since the various channels and directions of influence are not easily identified, and the relationship may not even be monotonic Stutzer-JEBO2004, estimates vary from slightly positive to substantially negative overall effects of having extra education.}

Similarly, although the literature on the importance of income and income growth on \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace is enormous and involves a large potential role of consumption externalities Barrington-Leigh-EQLWBR2014-consumption-externalities, one may summarize the findings by saying that income has been found to be a weak predictor of \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace in comparison to other, less market-mediated parts of life Blanchflower-Oswald-JPubE2004,Hamilton-Helliwell-Woolcock-NBER2016-social-capital-wealth,Layard-2011-happiness-lessons-new-science,Frijters-Krekel-2021-SWB-policy-handbook.\footnote{That is, the large compensating differentials found for having positive social relationships Powdthavee-JSE2008-valuing-social-relationships,Helliwell-Barrington-Leigh-2010-social-capital-worth,Helliwell-Putnam-PTRSL2004 reflect a small denominator, i.e. the value of income for increasing life satisfaction. }

This paper does not aim to explore all the reasons for this well-established evidence about quality of life from subjective response data. Instead, it characterizes a measurement error in which those with lower education and, as a proxy, those with lower income, may be more likely to under-utilize the \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace response options in such a way that tends to bias their reported life satisfaction. In general, resulting biases on marginal effects could exist in either direction, but as described below they are more likely to be downward, meaning that they may go some way to explaining the education anomaly and to revise upward, if modestly, the estimated importance of income for supporting \pdftooltip{\hyperlink{defSWB}{SWB}}{Subjective Well-Being (See page (ref))}\xspace.

Evolution of precision in subjective, quantitative reports

The history of survey questions on subjective assessments mirrors in part technological norms. Early innovators in monitoring \pdftooltip{\hyperlink{defSWB}{SWB}}{Subjective Well-Being (See page (ref))}\xspace in social and household surveys tended to use a four point or five point scale, typically with Likert-style verbal response options. In such questions, the numbers were not meant as cues for the respondent. In some populations, most respondents chose one of the top two options, limiting the variation, or precision. As limitations of paper survey media have been erased by the adoption of computer aided interviews, the resolution of these subjective scales has expanded. However, with more than five or seven response options, verbal cues are typically not provided except for the highest and lowest response options. Responses instead become numerical. For instance, after many years of asking \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace questions with a variety of scales, Statistics Canada settled over a decade ago on a particular wording with an 11 point scale.\footnote{However, Conti-Pudney-RES2011-SWB-survey-design-BHPS-distribution describe an evolution in the opposite direction, away from unlabeled response options, in 1992 in the British Household Panel Survey.}

The OECD-2013-guidelines-measuring-SWB has also developed recommendations for standardizing the way such questions are asked by all national statistical agencies. The de facto standard for \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace now is an 11-point scaling, from 0 to 10, with the lower extreme meaning, for example, “not at all satisfied”, the upper signifying “completely satisfied”, and the interpretation of the remaining values left up to the respondent.

An older literature sought to determine the optimal number of response options in survey questions with verbal cues for each option. For instance, it may be that in an oral interview, i.e., with no visual cues, four or five responses are the maximum that can be handled without confusion or overload Bradburn-Norman-Sudman-Wansink-2004-questionnaire-design.

When the scale is explicitly numeric, as with modern \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace measures, there also arises a trade-off between the cognitive load imposed by a scale and the precision it allows. From the respondent's point of view, this trade-off is between the opportunity for self-expression and the cost of cognitive processing. The survey designer wishes to allow for precise responses in order to capture variability among respondents and over time, while not demanding too much. Overburdening would result, at best, in the respondent not fully optimizing her answer or not properly interpreting or using the given range of responses OECD-2013-guidelines-measuring-SWB. Various studies on this balance have tended to favor 11-point quantitative scales over coarser option sets (e.g. 7-point scales) as well as over nearly continuous options Alwin-SMR1997-survey-question-scales,Kroh-DRAFT2006,Saris-vanWijk-Scherpenzeel-SIR1998,OECD-2013-guidelines-measuring-SWB,Weng-EPM2004-Likert-labels-and-number-of-options.\footnote{Interestingly, government surveys in the USA have tended to stick with 3 or 4-point scales for \pdftooltip{\hyperlink{defSWB}{SWB}}{Subjective Well-Being (See page (ref))}\xspace questions.}

Descriptive evidence

A small number of studies have remarked in some way on the use of focal values, but without a full account or explanation.\footnote{There is also a psychometrics literature which refers to the tendency towards top and bottom value responses as “extreme response style” and tendency towards the central value as “moderate response style” Hamamura-Heine-Paulhus-PID2007-cultural-differences-response-styles,Khorramdel-vonDavier-Pokropek-BJMSP2019-extreme-response-styles-model. That literature, known in psychology as “item response theory”, is motivated by an interest in personality type, as classified by responses to a battery of Likert questions with all-verbal response options. These studies have not considered cognitive ability as an explanatory factor. Giustinelli-Manski-Molinari-JEconometrics2020-rounding also study rounding of reported quantitative beliefs which relies on observing multiple questions for each respondent.} Dolan-Layard-Metcalfe-ONS2011-SWB-public-policy mention that \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace ratings in one study are positively associated with life circumstances as one would expect, except at the top of the scale, where “those rating their life satisfaction as `ten out of ten' are older, have less income and less education than those whose life satisfaction is nine out of ten”. They speculate a reason unrelated to cognitive limitations for this observation but declare that “This issue warrants further research”. Conti-Pudney-RES2011-SWB-survey-design-BHPS-distribution describe focal value behavior as a response to the existence of verbal cues, present on only three out of seven response options. Landua-SIR1992-panel-focal-values analyses response transition probabilities in the German Socio-Economic Panel, and Frick-Goebel-Schechtman-Wagner-Yitzhaki-SMR2006-SOEP-attrition-endpoints confirm his report that respondents have a tendency to move away from the endpoints over time. In fact, this could be driven largely by the \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace behavior diminishing as panel participants, especially those with low numeracy, gain familiarity and comfort with the scale.

Educational attainment

Simply inspecting their \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace distributions, stratified by education level, might have led these authors to the hypothesis developed in this paper. For illustrative purposes I appeal to one cycle from the Canadian Community Health Survey (CCHS), a large annual cross-section which includes an 11-point life satisfaction question as well as educational attainment.\footnote{In a repeated cross-section, most respondents are facing the \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace question for the first time. In the CCHS, education is recorded in four categories: less than high school graduate, high school graduate, some post-secondary training, and completed post-secondary training, but relatively few respondents report the third category, so I combine the top two. In the 2017-2018 wave, out of a total of 113289 respondents, 97604 reported their age as at least 25 years, and 93043 of those answered the \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace, educational attainment, and income questions.

} Conditioning \pdftooltip{\hyperlink{defSWB}{SWB}}{Subjective Well-Being (See page (ref))}\xspace responses on educational attainment reveals a striking feature (\figref{CCHS-17-18:SWL-by-education}). The relative frequencies of each focal value (0, 5, and 10) decrease with increasing education level. While the lowest education category shows four peaks, the distribution of responses in the highest education category features what would be a unimodal distribution around SWL=8, except for a slight enhancement at SWL=0. In addition to \figref{CCHS-17-18:SWL-by-education}, several other lines of evidence support the interpretation that scale simplification is a specific response to cognitive challenge, a model to be formalized in \secref{Cognitive-model}.

Numeracy

figure*[figure* omitted — 878 chars of source]

Educational attainment is a widely-available characteristic in social survey data, but may capture attributes which are relevant to latent wellbeing or to reporting functions, but which are different from cognitive ability with numbers. To support the numeracy interpretation of \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace, Appendix Figure F.1 makes use of the 2018 Programme for International Student Assessment (PISA) survey of 15 year old enrolled school students in 72 countries. This survey includes a measure of mathematical ability, along with 11-point life satisfaction. Separating respondents according to an overall math score,\footnote{This math score is a “plausible value” of the individual's underlying latent ability, appropriate for modeling relationships such as this one; see Wu-SEE2005-plausible-values-PISA.} the same feature as in \figref{CCHS-17-18:SWL-by-education} is observed: the relative frequency of each focal value response decreases with math proficiency, as does the mean \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace response. This is true of the global distribution, as well as for individual countries such as the USA (see Appendix Figure F.1).

Difficulty responding

Another indication that the apparent tendency to simplify the response scale has to do with the difficulty of answering the question, as it is posed, comes from noticing that respondents with less education are more likely to refuse to answer the \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace question at all. Appendix Table F.1 shows that response rates to the \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace question, although close to 100%, are strictly increasing with educational attainment.

Unordered choice model

The existence of \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace behavior implies that \pdftooltip{\hyperlink{defSWB}{SWB}}{Subjective Well-Being (See page (ref))}\xspace response scales cannot safely be assumed to be ordinal. For example, those with lower education may, all else equal, experience lower life satisfaction but be systematically inclined to report a higher value due to rounding up from a 3 or 4 to 5, or from 8 or 9 to 10. It is possible, therefore, that on average those reporting 9 could be happier than those reporting 10.

Traditional methods used in econometric inference from \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace --- such as OLS, ordered logit, ordered probit, and related time series and instrumented analogues --- leverage strong assumptions about the symmetry of effects of explanatory variables on each step of the response scale, as well as assuming cardinality or at least ordinality among response values. Those models are therefore not flexible enough to account for the heterogeneous influence of predictors like education on focal and non-focal response values.\footnote{Note that, unlike OLS, ordered logit and ordered probit naturally account for multi-peaked distributions such as is shown in \figref{CCHS-17-18:SWL-by-education}. That is due to the flexibility given by the cut points in those models, which can squeeze together or stretch apart in order to decrease or increase (respectively) an option's predicted response probability.}

An alternative approach is to relax the ordinality assumption for response options, and model the probability of each response independently, subject only to the constraint that the probabilities add up to one. The multinomial (polytomous) logit model\footnote{The multinomial logit model, in its latent variable formulation, consists of a system of equations generating scores $Y_{i,j}^{*}$ for each individual $i$ and response option $j\in\{0\dots10\}$ as $Y_{i,j}^{\ast}=\boldsymbol{\beta}_{j}\cdot\mathbf{x}_{i}+\varepsilon_{j}\,$ where $\varepsilon_{j}\sim\text{EV}_{1}(0,1),$ i.e., the error terms have standard type-1 extreme value distributions. Then observation probabilities are given by $P(Y_{i}=s_{j})=e^{\boldsymbol{\beta}_{j}\cdot\mathbf{x}_{i}}/\sum_{k\ne j}e^{\boldsymbol{\beta}_{k}\cdot\mathbf{x}_{i}}$ with one necessary normalization like $\beta_{0}=0.$ This model is clearly also misspecified for \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace data, since the formal independence of irrelevant alternatives (IIA) assumption is violated by the numbered options of an \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace question. The IIA requirement is frequently overlooked in applications of the multinomial logit.} does this.

figure*[figure* omitted — 824 chars of source]

\global\long\def\Pswl#1{P_{SWL=#1}}

\figref{mlogit-marginal-effects-multivariate} shows marginal effects of education and income on response probabilities of each of the 11 points in the \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace scale, from a multinomial logit model using education, logarithmic income, age, and age$^{2}$ as predictors for the sample shown in \figref{CCHS-17-18:SWL-by-education}. Under an ordinality assumption, one might expect marginal effects to rise monotonically with response value, since a better circumstance like education or income should lead to an increase in the relative probability of response $s+1$ as compared with response $s$. Indeed, other than the focal response values 0, 5, and 10, the marginal effect of one step higher educational attainment (for instance, graduating from high school) is weakly increasing in reported \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace. By contrast, the effects on the focal value responses are, with high statistical significance, negative\footnote{For each explanatory variable, the sum of all marginal effects on probabilities is zero by construction.} outliers far below what would be expected based on the pattern of adjacent values. They show that more education significantly reduces the probabilities of each focal value response. Multinomial logit estimations provide a diagnostic tool for detecting predictors of focal value behaviour. However, the effect sizes are hard to interpret because they are averages over the entire sample. For instance, the education coefficient for \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace=10 is an average effect over high types, for whom higher education increases the chance of reporting 10, and low types, for whom higher education decreases that chance. In order to separate those effects, a more structured mixture model approach, described below, is required.

Precision and self-expression

As a final piece of empirical motivation for the modeling approach developed below, I note that when excess precision is offered in an \pdftooltip{\hyperlink{defSWB}{SWB}}{Subjective Well-Being (See page (ref))}\xspace scale, respondents appear to make a costly effort to choose round numbers. Specifically, Appendix Figure F.2 shows the distribution of responses from a computer-based \pdftooltip{\hyperlink{defSWB}{SWB}}{Subjective Well-Being (See page (ref))}\xspace survey question framed on a 0--10 scale but with an available resolution of 0.1. There are clearly favored responses at every integer and half-integer value. The response interface was a graphical slider which gave no preference for any particular values. Thus, the prevalence of rounded values indicates that extra effort in the form of fine manual control was exerted in order to leave the slider precisely on a half- or whole-integer value. This can be interpreted as evidence of effort to faithfully communicate a mental result, motivated by the drive for self-expression Alwin-SMR1997-survey-question-scales,OECD-2013-guidelines-measuring-SWB.\footnote{It also suggests that respondents have introspective knowledge of their degree of precision in answering a numerical \pdftooltip{\hyperlink{defSWB}{SWB}}{Subjective Well-Being (See page (ref))}\xspace question --- knowledge which is not typically elicited in surveys.}

Cognitive mixture model

\global\long\global\long\global\long\def\Prob#1{P\left(#1\right)} \global\long\global\long\def\Prg#1#2{\Prob{#1\mid\boldsymbol{x}#2}} \global\long\def\Prgz#1#2{\Prob{#1\mid\boldsymbol{z}#2}}

Motivated by the evidence above, the enhanced use of focal values can be interpreted as an indication that respondents have simplified their cognitive task by coarsening the numerical scale. Because \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace behavior is inversely associated with education and math skills, I focus on “numeracy” as one major influence on scale choice. The two-type mixture model below is based on the assumption that the cognitive processes of respondents differ only in the execution of step (iv) described in the second paragraph of \secref{introduction}. That is, an internal representation of overall wellbeing exists in a similar way across the two groups, who subsequently project that assessment onto either the full scale (high numeracy respondents) or a subset consisting of the bottom, central, and top values (low numeracy respondents).

For each of the two types, latent wellbeing is mapped onto a discrete response scale as in a standard, i.e. canonical, ordered logit model. That is, given a continuous, latent subjective assessment $S^{\star}$ modeled in terms of individual characteristics $\boldsymbol{x}$ as $S^{\star}=\boldsymbol{x'\boldsymbol{\beta}_{s}}+\varepsilon$, the cumulative probability of discrete responses $k$ is given by:

align[align omitted — 250 chars of source]

where $\alpha_{k}$ are a sequence of threshold values $\alpha_{k}^{H}$ separating the full set of observed responses $\left\{ 0,1,\dots,10\right\} $, or $\alpha_{k}^{L}$ for the focal subset $\left\{ 0,5,10\right\} $, and $\Phi(\cdot)$ is the cumulative distribution function of the unexplained portion $\varepsilon$ of $S^{\star}$. Use of the logistic distribution for $\Phi(\cdot)$ makes this an ordered logit model.

The high and low alternative ordered logit outcomes are combined using a simple dichotomous logit model. If $\boldsymbol{z}$ is a vector of individual characteristics, possibly overlapping with $\boldsymbol{x}$, which serve as a measure of numeracy, then

equation[equation omitted — 112 chars of source]

There is no explicit consideration of costs and benefits to the respondent.\footnote{Conceptually, the individual benefits of using the full scale are self-expression and performing one's best at fulfilling the purpose of the survey; the costs are those of the cognitive calculation. For more on the tradeoff between self-expressive capacity and cognitive capacity in the design of response scales, see Alwin-SMR1997-survey-question-scales,Kroh-DRAFT2006,Saris-vanWijk-Scherpenzeel-SIR1998,OECD-2013-guidelines-measuring-SWB. There is no evidence that respondents' value of time, which might for instance vary with income, is also a major factor in scale choice in the face of the inclination for self-expression. The value of time likely affects the choice to participate in a survey, but once asked the \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace question, respondents answer it quickly, i.e., in a few seconds. The indication in evidence shown above and below is that higher income predicts higher resolution in responses, which seems contradictory to a hypothesis of behavior driven by the market value of respondents' time..}

\newif\iftwocolumncpbl\twocolumncpbltrue Together, (ref) form a mixture model. The probability of observing response $k$ is \iftwocolumncpbl

align[align omitted — 211 chars of source]

\else

equation[equation omitted — 166 chars of source]

\fi The model is similar to the ordinal-outcome “finite mixture model” of Boes-Winkelmann-ASA2006-ordered-response except that here the mixing probability is dependent on individual characteristics Everitt-Merette-JAS1990-mixture-model-ordered-logit,Everitt-SPL1988-mixture-model-ordered-logit,Uebersax-APM1999-mixture-model-ordinal. A more detailed account of the model is presented in Appendix A..

Identification

Are the parameters in this model point-identified in principle?\footnote{Point-identification, typically referred to simply as “identification”, is also called frequentist identification and is a frequentist concept. In Bayesian estimation, as is used in the empirics to follow, parameters are assumed to have distributions, not point values. Using a Bayesian estimation method with a broad enough prior, alternate sets of values which account for the data are simply reflected in multimodal (or suitably broad) estimates of the parameters Lewbel-JEL2019-identification-meaning.} Identification is a challenge because the same predictors may be used to predict the latent numeracy variable and to predict the latent wellbeing variable. As a result, one might fear that more than one set of parameters could equally well explain observations for a given sample. Excluding the columns of $z$ from $x$ in (ref) would overcome this problem. However, for an all-encompassing subjective outcome such as latent wellbeing, it is safer to assume that everything could be a determinant. More specifically, a particular interest motivating this study is to assess the bias on estimates of the wellbeing effect of education, and education is also the primary available predictor of numeracy.

With stronger assumptions, an alternative strategy to the mixture model may be able to identify parameters for latent \pdftooltip{\hyperlink{defSWB}{SWB}}{Subjective Well-Being (See page (ref))}\xspace by avoiding \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace altogether. One approach would be through thin set identification, if respondents with some level of education were known never to exhibit \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace. One standard problem with this kind of identification is that it relies on an assumption of a uniform effect of a covariate across its support, as well as the absence of interaction effects with other covariates. By contrast, the mixture model approach of (ref), which leverages the entire sample, has the advantage of generalizability to explicitly estimate interaction terms or other functional forms to allow for non-uniform effects.

Another approach would be through selection on the dependent variable; that is, by restricting the sample to the subset of high types who did not respond with a focal value. Because no “5”s are observed in this group, it would consist of two subsamples: those with observed $s\in\{1,2,3,4\}$ and those with $s\in\{6,7,8,9\}$. In fact, if the symmetries required for this approach to be unbiased were believed, then one could likely estimate coefficients for latent wellbeing using binary models like logit and sample subsets of respondents who answered one of only two consecutive, non-focal response options.

Returning to (ref), within each of the two ordered logit formulations nested in the model, identification of the set of parameters (with no constant term) is standard. This still leaves us with incomplete identification, in general, of the parameters on variables common to $\boldsymbol{x}$ and $\boldsymbol{z}$. One can imagine extreme distributions of \pdftooltip{\hyperlink{defSWB}{SWB}}{Subjective Well-Being (See page (ref))}\xspace, for instance all near 10, in which \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace only acts to convert 9s to 10s. In this case, discriminating between the effect of common variables on latent wellbeing or \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace would not be possible, especially if the sign of coefficients is not constrained. However, more typically, with a broader \pdftooltip{\hyperlink{defSWB}{SWB}}{Subjective Well-Being (See page (ref))}\xspace distribution, \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace will be distinguishable from effects on latent wellbeing. That is, successful identification rests on having sufficient independent, explainable variance in latent \pdftooltip{\hyperlink{defSWB}{SWB}}{Subjective Well-Being (See page (ref))}\xspace across low types in order that there is also variation in their observed response. For instance, if the latent wellbeing of low types is sufficiently spread out that they sometimes round down and sometimes round up, then the influence of education on \pdftooltip{\hyperlink{defSWB}{SWB}}{Subjective Well-Being (See page (ref))}\xspace, controlling for other influences, is separately identified from the influence of education on the reporting function, i.e., on the likelihood of being a low type. Put differently, identification comes from the response of the observed distribution to changes in numeracy, driven by some variable, being different from the response of the observed distribution to changes in latent wellbeing, driven by the same variable. This is assured when there is nontrivial variation in the latent wellbeing of low types. This conceptual argument is best corroborated quantitatively through simulation, which demonstrates, in \secref{simulation}, that $\boldsymbol{\beta}_{S}$ and $\boldsymbol{\beta}_{N}$ are simultaneously recovered when estimating (ref).

Focal Value Rounding Index

As shown below, net biases on some estimated moments and model coefficients may be zero due to offsetting effects, even when \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace behavior is prominent. Therefore, to express straightforwardly the magnitude of the numeracy problem in a sample of respondents, another estimated value is helpful. This is the Focal Value Rounding Index (\pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding Index (See page (ref))}\xspace), which is an estimate of the fraction of the population who restrict their answer to a set of focal values --- i.e., the estimated fraction of low types. This value is well identified whenever $\boldsymbol{\beta}_{N}$ is.

Counterfactual SWL distribution

The mixture model provides a posterior estimate of the latent \pdftooltip{\hyperlink{defSWB}{SWB}}{Subjective Well-Being (See page (ref))}\xspace distribution, i.e., that which would have been reported had respondents all used the full scale. This represents a “correction” to the reported distribution of \pdftooltip{\hyperlink{defSWB}{SWB}}{Subjective Well-Being (See page (ref))}\xspace. This is a distribution of predicted, counterfactual, discrete responses on the 0--10 scale, not an estimate of the latent variable $S^{\star}.$ The next section demonstrates through simulation that the model successfully recovers (identifies) this counterfactual distribution, along with the \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding Index (See page (ref))}\xspace, means, and coefficients.

Model validation

This section, supplemented by several appendices, reports on the use of simulated data to validate the computational approach\footnote{Estimation of the mixture model was carried out using the no-U-turn sampler (NUTS) variant of a Hamiltonian Monte Carlo (HMC) algorithm, which is in turn a Markov Chain Monte Carlo (MCMC) method Stan2018,pystan2.19.1.1,Carpenter-Gelman-Hoffman-Lee-Goodrich-Betancourt-Brubaker-Guo-Li-Riddell-JournalStatSoftware2017-Stan and handles the non-concave objective well. Appendix B provides more detail on estimation, including analytic derivations of the gradient and Hessian for a log likelihood approach.} and the model's ability to identify simultaneous influences of a predictor, like education, on \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace and on latent \pdftooltip{\hyperlink{defSWB}{SWB}}{Subjective Well-Being (See page (ref))}\xspace. A large battery of simulations demonstrates the complexity and scope of possible biases, due to \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace, in conventional estimates of \pdftooltip{\hyperlink{defSWB}{SWB}}{Subjective Well-Being (See page (ref))}\xspace means and of marginal effects.\footnote{To reiterate, this bias is, conceptually, the difference between naively estimated values and those which would be obtained in the counterfactual case that all respondents had used the full numeric response scale, i.e. were of “high numeracy” type. This counterfactual can be perfectly calculated using synthetic data, but is of course unobservable in traditional empirical data.}

Synthetic data validation results

Simulated \pdftooltip{\hyperlink{defSWB}{SWB}}{Subjective Well-Being (See page (ref))}\xspace data are generated by a model in which a scalar $z$ partly determines numeracy through (ref) while $z$ and a second scalar, $y$, partly determine the latent wellbeing $S^{\star}$ (thus $\boldsymbol{x}\equiv

bmatrix[bmatrix omitted — 19 chars of source]

$ in \eqref{cumulativeOLogit}). A non-zero correlation, parameterized by $\chi$, may exist between $z$ and $y$. Conceptually, and for comparison with the empirical estimates to follow, $z$ is meant to represent education and $y$ represents other variables, such as income, which are not direct measures of numeracy (do not cause \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page \pageref{def:FVR})}\xspace). In order to reveal the possible scope of biases for plausible distributions of \pdftooltip{\hyperlink{defSWB}{SWB}}{Subjective Well-Being (See page \pageref{def:SWB})}\xspace, a number of parameters of the synthetic data generation process were varied systematically. These include $\chi,$ $\boldsymbol{\boldsymbol{\beta}}_{N},$ and two parameters determining the scale and offset of the cut points.\footnote{See Appendix C for details of the parameters used in the synthetic data generating function, Appendix D for some detail from estimates using the simulated data, showing an example of the complicated dependence of biases on attributes of the distribution, and Appendix E for propositions on the maximum possible scope of these biases.}

figure*[figure* omitted — 1,259 chars of source]

\figref{v3figa-bc} shows one example of a simulated distribution of \pdftooltip{\hyperlink{defSWB}{SWB}}{Subjective Well-Being (See page (ref))}\xspace. In (a), shaded bars represent simulated responses on a 0 to 10 scale. Unlike in real data, we are able to identify which respondents (among those giving a 0, 5, or 10) used \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace. This portion of responses, labeled “low type”, are shaded pink. Also because the data are synthetic, we are able to construct the latent (“true”) wellbeing levels and thus the counterfactual 0--10 responses which would have been given if everyone reported without \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace. This counterfactual distribution, including both low and high types, is shown split into two groups based on education level. Although the true wellbeing distribution of this sample is centered around 7.5, equidistant from the focal values of 5 and 10, there is a net negative bias of $-0.08$ in mean reported \pdftooltip{\hyperlink{defSWB}{SWB}}{Subjective Well-Being (See page (ref))}\xspace. This is because the distribution of the lower educated component is generally closer to “5” than to “10”. Thus, the amount of rounding up is less than the amount of rounding down.

Simulated biases in regression coefficients are obtained by estimating a traditional ordered logit model on the synthetic data, and comparing those estimates to the true values used in constructing the data, $\beta_{S}^{z}=\beta_{S}^{y}=1$. In the case shown in \figref{v3figa-bc}, these biases are also both negative, namely $-14\%$ and $-23\%$ respectively.\footnote{The origin of these negative biases is slightly more subtle; see Appendix D and Appendix E for details.}

Simulations were carried out for a wide range of parameters, generating cases with both positive and negative biases on mean \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace and on $\beta_{S}^{z}$ much larger than in this example. Simulated biases on $\beta_{S}^{y}$, by contrast, tend to be negative.\footnote{ See Appendix D for an explanation. In the simulations, $y$ has no extra effect on (information about) numeracy, after taking $z$ into account. In real data, income is likely to contain variance that is informative for \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace but orthogonal to available measures of education. In this case, income in empirical applications will exhibit a blend of the bias features attributed to $z$ and $y$ in these simulations.} More generally, the bias on mean \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace can be as large as $\pm2$ points (see Appendix Proposition E.1 in Appendix E) and the bias on $\beta_{S}^{z}$ may be even larger (Appendix Proposition E.1). In all cases, the true distribution, fraction of low-types, and effects of $z$ and $y$ on latent wellbeing are identified and correctly estimated by the \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace mixture model. As an example, \figref{v3figa-bc}(b) shows estimated coefficients and cut points for the same case shown in (a).

Variance of SWL (“happiness inequality”)

Although not a focus of this paper, it is also worth noting that variance of \pdftooltip{\hyperlink{defSWB}{SWB}}{Subjective Well-Being (See page (ref))}\xspace, which has attracted interest as a measure of inequality Goff-Helliwell-Mayraz-EI2018-SWB-inequality,Hasegawa-Ueda-JSE2011-SWB-inequality,Stevenson-Wolfers-NBER2008-happiness-inequality-United-States, also suffers from bias due to \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace, as of course do other moments and other measures of dispersion. For a relatively narrow distribution of \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace centred around 5, \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace behavior decreases the variance. For a wider distribution, focal values of 0 and 10 would become prominent, and the variance could be biased upwards instead.

Empirical estimates of \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace bias and \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding Index (See page (ref))}\xspace

With the above evidence of parameter identification from simulated data, the rest of this paper turns to empirical estimates. The distributions of \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace for different levels of education, shown in \figref{CCHS-17-18:SWL-by-education}, indicate the significance of focal value rounding behavior in the CCHS sample. Using the mixture model, the role of education in supporting \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace can be estimated, despite the strong relationship between education and the focal value bias. Columns (1) and (2) of \tabref{CCHSmixtureEstimates} show the results of conventional, or “naive” estimates of the following simple individual-level cross-sectional OLS model for \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace, \iftwocolumncpbl

align[align omitted — 162 chars of source]

\else \[ \text{SWL}_{i}=c+\sum_{j}\beta_{S}^{h_{j}}\text{education}_{j,i}+\beta_{S}^{I}\log\left(\text{HH income}\right)_{i}+\varepsilon_{i} \] \fi as well as its ordered logit counterpart. Educational attainment is captured by a set of cumulative dummies, so that $\beta^{h_{j}}$ is the impact of having completed education level $j$ or higher.

The naive estimated coefficient on completing secondary education is near-zero or distinctly negative in the two estimates. The ordered logit coefficients predict that completion of high school reduces the odds of a higher \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace by more than 7%, and that even a university education reduces those odds by nearly 4% as compared with someone who has less than a high school education. These values are economically large; using the simultaneously-estimated coefficient on log income, the former effect is estimated to be equivalent to a 13% reduction in income.\footnote{The values in this paragraph are calculated as $e^{-.075}-1=-0.072\approx-7\%$; $e^{-.075+.038}-1=-0.036\approx-4\%$; and $e^{-.075/0.53}-1=-0.13\approx-13\%.$}

table*[table* omitted — 6,120 chars of source]

When constrained to disallow focal value behavior, the mixture model's estimate, shown in column (3) of \tabref{CCHSmixtureEstimates}, reproduces the ordered logit values, as it should. However, when the full model is estimated, a significantly positive value ($\sim$0.06) is found for the \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace benefit of completion of secondary school, and an additional 0.17 for those completing post-secondary.

The bias in a conventional estimate of the income coefficient is also large: the mixture model strongly rejects the naive estimated value of $\sim$0.53, in favor of a value of $\sim$0.62. This represents a 17% difference in the most studied value in happiness economics. Combining these coefficients implies that, after controlling for income, the true benefit of college completion, as compared with an otherwise-similar respondent without high school completion, is equivalent to an additional 45% of income. High school completion by itself confers a benefit equivalent to more than 10% of income, after controlling for differences in actual income.

The specification in \tabref{CCHSmixtureEstimates} includes both education and income as predictors of \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace. Appendix Table F.2 shows that alternate models with only education in the \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace equation, or with additional controls, give highly consistent results.

Next to \tabref{CCHSmixtureEstimates} are visualizations of several sets of distributions, showing the model's ability to predict observed response patterns while estimating the distribution of “underlying” or “true” \pdftooltip{\hyperlink{defSWB}{SWB}}{Subjective Well-Being (See page (ref))}\xspace.

Applications

Hundreds of empirical papers estimating models of life satisfaction and other extended-Likert-like scales could be revisited in light of the significant possibility of biases identified above. Those focusing on effects of socioeconomic status, gender, and age, and those which particularly address populations with low levels of numeracy, especially invite reanalysis. Here I reproduce estimates from three papers to exemplify the important changes that may result from such analysis, and to show that the often-reproduced “paradox” of negative benefits to education may be largely resolved by the cognitive mixture model.

U.K.: Clark and Oswald (1996)

The first of these papers, with over 1500 citations, is a relatively early contribution in the modern study of relative income concerns but also prominently points out the anomalously low estimated returns to wellbeing from education Clark-Oswald-JPubE1996. It was also recently cited as one of 11 studies in the major accumulated evidence on the life satisfaction benefits from additional education Clark-Fleche-Layard-Powdthavee-Ward-2019-origins-of-happiness. In fact, the paper uses data from the British Household Panel Survey (BHPS) prior to its inclusion of \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace, so it uses instead responses to 7-point satisfaction with pay and satisfaction with job questions.

Clark-Oswald-JPubE1996 did not examine the distributions of these subjective response variables according to formal educational attainment.\footnote{The description from Clark-Oswald-JPubE1996 reads: “Table 5 contains two ordered probits, in each of which three dummies for educational attainment are included as well as a control for income. The dummies are for a college degree, advanced high school (A-level approximately), and intermediate high school (O-level approximately). The omitted category is for no or low qualifications. These four categories are for achieved paper certificates and not merely for years of schooling”.} Doing so reveals dramatic \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace behavior which roughly diminishes with education (\figref{BHPS-SWB-by-education}). The distribution of satisfaction with pay is wider and more central (i.e., near “4”) than that of job satisfaction, and features unmistakable evidence of all three focal values (1, 4, and 7) for groups with lower academic certifications. For job satisfaction, the upper focal value is most obvious but all three are evident on inspection. If those with A-levels but no College are excluded, then the group means and the prevalence of each focal value all decrease monotonically with education.

figure*[figure* omitted — 506 chars of source]

\tabref{jobSatis-BHPS-main-estimates} shows raw coefficients for model estimates of overall satisfaction with job. The first three models are conventional estimation approaches, including an ordered probit model, which nearly reproduces the published values\footnote{The coefficient shown on log income (.016) strongly disagrees with the published value (.50) in Clark-Oswald-JPubE1996. Upon contacting the authors, it was determined that a typo in production of the original work resulted in a reporting of 0.50 rather than the estimated 0.05 for this coefficient (personal communication, Andrew Clark, 2021). This error has not been previously reported. Because of the error, the authors did not address the surprisingly low coefficient on log of household income. The set of regional, health, race, industry, and occupation dummies are excluded in \tabref{paySatis-BHPS-main-estimates} because the exact definitions from the 1996 work are not available.} and retained sample size (4730 in all my estimates) of the main estimate in Clark-Oswald-JPubE1996.\footnote{For easier comparison with their table, the education categories are mutually exclusive, rather than cumulative, as in the CCHS data and the data to follow.} In ordered probit, OLS, and ordered logit models, academic attainment is strongly predictive of lower satisfaction after adjusting for log of household income. The implied effect is enormous. As compared with someone with primary education only, an advanced high school graduate (A-levels) is less satisfied with their job by as much as they would be with a 3-fold decrease in wage.\footnote{The coefficients on log job hours and on A-levels education are nearly identical, implying that, having already adjusted for income, a unit log, or factor $\sim$2.7, increase in hours worked predicts a similar change to job satisfaction as does the educational attainment.} As already shown in \figref{BHPS-SWB-by-education}, even the raw mean job satisfaction is decreasing across the first three education groups. Clark-Oswald-JPubE1996 speculate that their findings of low satisfaction of the higher educated may be related to a recent recession that particularly hit the middle class in the UK, but also cite several earlier studies which corroborate the negative or negligible benefits from education on job satisfaction.

table*[table* omitted — 4,777 chars of source]

Equally surprising in these results is the nil effect of income on job satisfaction. The 95% confidence interval for the coefficient of log income in column (3) is $-$0.10 to $+$0.13, with the upper limit implying that a doubling of income would increase the odds of a higher satisfaction response by less than 10%.

table*[table* omitted — 4,699 chars of source]

Column (4) simply shows that the cognitive mixture model reproduces an ordered logit estimate when focal value behavior is turned off, while the key result lies in Column (5). When focal value behavior is accounted for, the income coefficient increases to a confidently positive value, and the strongly negative coefficients on O-level and College completion are eliminated. Respondents who finished A-levels but stopped there for some reason, i.e., did not complete college, are still predicted to be less satisfied with their jobs, but the effect is half as large as in the naive model. Estimates of other coefficients remain statistically unchanged. Both formal education and reported income prove significant in predicting focal value behavior. The estimated fraction of respondents, overall, who restricted their answers to focal values is 28%. The model also estimates a significant bias in the mean reported job satisfaction, from a latent value of 5.3 which would have obtained had all respondents used the full scale, to the observed value of 5.5. The model estimates that the low-numeracy (\pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace) respondents reported an average job satisfaction of 5.9, and that the high-numeracy respondents reported an average of 5.3.

\tabref{paySatis-BHPS-main-estimates} parallels \tabref{jobSatis-BHPS-main-estimates} but relates to the other column in Clark-Oswald-JPubE1996's Table 5 --- an estimate for satisfaction with pay rather than with the job overall. In this case, increased income is a strong predictor of satisfaction even in naive estimates. On the other hand, higher education again strongly predicts lower satisfaction, after adjusting for household income, in conventional models. This may make sense if the primary effect of education in this context is to set expectations about pay. In any case, for satisfaction with pay, the \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace mixture model corroborates the estimates of the naive ordered logit model.

How can the coefficient estimates remain relatively unchanged in the presence of such a high degree of \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace? While column (5) of \tabref{paySatis-BHPS-main-estimates} shows that income and higher education levels predict lower propensity for \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace, and that 31% of respondents used a simplified response scale for answering this question, the net effect of \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace on the estimated coefficients is small. This can be understood by considering the distribution of latent wellbeing values, with reference to the discussion in \secref{Validation} and the Remark for Appendix Proposition E.2. For this sample, the number of respondents rounding up from 6 to 7 or from 3 to 4 is balanced by the number rounding down from 2 to 1 or from 5 to 4.\footnote{As discussed earlier and as this example shows, there is no simple relationship between the extent of \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace and the size of net biases, due to the possibility of offsetting contributions to bias and the importance of detailed distributional features of the sample. It is also worth noting that the model is capable of accounting for a high fraction of extreme values (1s and 7s, here) as scale boundary effects rather than \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace. Indeed, it is also capable of accounting for a central peak (here, at 4) without appealing to the existence of any \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace. Instead, the model estimate suggests that respondents were simplifying the scale, and that the independently-estimated fractions of respondents who did so were the same (28% and 31%) for the two questions.} Appendix Figure F.3 shows the estimated distributions of responses which would have been given in the absence of any \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace (second row), for both job and pay satisfaction. All education levels exhibit broad distributions of latent pay satisfaction, and all carried out some degree of \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace.

Australia: Powdthavee-Lekfuangfu-Wooden-JBEE2015-education-SWL-Australia-HILDA

More recently, Powdthavee-Lekfuangfu-Wooden-JBEE2015-education-SWL-Australia-HILDA have shed some further light on the apparent negative or insignificant returns to education in life satisfaction regressions. They articulate a more considered causal model for the impact of educational attainment on overall life evaluations, taking into account several of the multiple non-monetary channels through which education is expected or known to affect life. In particular, they allow for mediating effects of education through health, marriage, child-rearing, and employment, in addition to income. They conclude that “education is likely to be positively related to overall life satisfaction through many different channels, even when ceteris paribus education itself has a negative and statistically significant relationship with overall life satisfaction”. Thus, while identifying some positive indirect effects of education, their analysis does not account for the overall negative effect of education on life satisfaction.

Here I do not integrate their panel data mediation pathways into the \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace model, which would go beyond the scope of this paper. Instead, I use one cycle (2010) of the HILDA survey Powdthavee-Lekfuangfu-Wooden-JBEE2015-education-SWL-Australia-HILDA to test the same questions as above: how much of the negative overall association between education and life satisfaction is accounted for by \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace behaviour? and how biased is the income coefficient when \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace is ignored?

\figref{histograms-HILDA-by-education-25yrold} shows a familiar pattern in weighted life satisfaction response distributions when separated by education level. Here the focal value enhancements are more subtle, but anomalously high response fractions for 5 and 10 are noticeable at least in the lowest education group, and the proportions of each focal value decrease across education groups.

figure*[figure* omitted — 359 chars of source]

\tabref{HILDA-main-estimates} shows the comparison in now-familiar form of the naive estimates of income and education effects on life satisfaction in Australia (columns 1, 2, and 3) with an estimate of the \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace model in column (4). In the \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace-aware model, the coefficient on income approximately doubles, jumping by 4 standard errors. The additional effect of college degree attainment after finishing high school becomes weakly positive, and the effect of high school graduation climbs by 5 standard errors.

table*[table* omitted — 4,626 chars of source]

First Nations and Métis in Canada

Next I pick on my own prior work by re-examining a paper which reported an anomalously low benefit of income for a sample of Indigenous (First Nations and Métis) peoples in Canada Barrington-Leigh-Sloman-IIPJ2016-aboriginal. In addition to estimating a negative effect of income on life satisfaction, we found an average life satisfaction among Indigenous respondents that was equivalent to that of the general population, despite the stark objective challenges faced by the former groups, including disproportionate levels of discrimination and socioeconomic disadvantage with respect to the rest of the Canadian population. Barrington-Leigh-Sloman-IIPJ2016-aboriginal suggested as a possible interpretation that total income is not well measured by the standard income question for this group, but remain “cautious and skeptical” about the data overall.

figure*[figure* omitted — 587 chars of source]

This case study relates to the importance of being able to use life satisfaction data across diverse cultural and economic circumstances. It also demonstrates the use of the mixture model on a small sample. The data come from two Canadian surveys: the national Equality, Security and Community survey (General ESC, $N=3725$) and its follow-up small sample of on- (70%) and off-reserve (30%) First Nations and Métis peoples\footnote{Both of these groups are considered Aboriginal (and, along with the Inuit, Indigenous).} in the Canadian Prairies (Aboriginal ESC, $N=446$). As can be seen in the first panel of \figref{indigenous-distributions}, an enhancement at \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace=10 in the Aboriginal ESC sample makes it the modal response value and may go some way to explaining the high mean reported \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace. Indeed, this is likely the first report of a \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace distribution with such a dominant response at its top value. On the other hand, respondents also gave plenty of 7s, 8s, and 9s, each with higher frequency than \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace=5. Below I use the cognitive mixture model to assess how much this distribution might be biased by \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace, and whether the anomalous estimates in Barrington-Leigh-Sloman-IIPJ2016-aboriginal are reversed.

The first column of \tabref{ESCaboriginal} shows a conventional ordered logit estimate of 10-point life satisfaction of the Aboriginal sample. For consistency with Barrington-Leigh-Sloman-IIPJ2016-aboriginal, the education variable is a more continuous variable than in the previous two applications, being measured on ten steps ranging from no primary school to a PhD or professional degree. Once again, and despite a sample size of only 446, a significantly negative coefficient on education shows that, after adjusting for income, those with higher education report lower life satisfaction. In addition, the coefficient on log household income is estimated to be most likely negative, with a 95% confidence interval between $-$0.40 and $+0.08$.

The second column reports the estimate of a cognitive mixture model. Education strongly predicts numeracy, i.e., use of the full response scale. Most importantly, the education anomaly in the earlier analysis is resolved when \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace is taken into account: the confidently negative education coefficient is replaced by a weakly positive point estimate with a 95% confidence interval between $-$.08 and +.17. The weaker anomaly of a negative income coefficient is also partly resolved; in its place is one centered closely on zero with similar precision.

In order to address the surprisingly high average life satisfaction reported by Indigenous respondents, I next use a pooled model to compare groups after controlling for income and education. Pooled estimates of the Canada-wide respondents and the First Nations/Métis sample are shown in Columns (3) and (4) of \tabref{ESCaboriginal}. Adjusting for income and education, the Aboriginal respondents report 0.30 higher life satisfaction than non-Aboriginal. Although the explanatory variables here are few and the model is simple, this positive boost is counterintuitive for the reasons described above. However, when \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace is accounted for (Column 4), this situation is reversed, with the Aboriginal respondents reporting a weakly lower life satisfaction than others with similar income and education. In this model, education, income, and Aboriginal status are all allowed to predict \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace behavior. Education and income positively predict lower propensity for \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace behavior, as expected, while Aboriginal status has the equivalent effect on \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace as a two-point reduction in educational attainment level, for instance from completing a technical or community college certification to completing only high school.

For the pooled sample, the mixture model estimates a 70% higher income coefficient and corrects the strongly negative education effect of the naive model estimate with a weakly positive one.

table*[table* omitted — 4,614 chars of source]

Ranking of U.S. states by happiness

The United States is somewhat exceptional in that there are no prominent domestic surveys assessing subjective wellbeing with more than a 4-point response,\footnote{However, two international datasets, the World Values Survey and the Gallup World Poll, do so on 10 and 11 point scales, respectively.} with the exception of the Gallup Daily Poll, which poses the Cantril Ladder question on an 11-point scale.

In this section, I investigate the extent to which a ranking of states by average reported life evaluations is biased by focal-value response behavior. I find that state-level differences in educational attainment relate to state-level differences in \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace. Applying the cognitive mixture model to these data provides a counterfactual “latent” or “corrected” mean life evaluation for each state, allowing for a comparison between a naive ranking and a corrected ranking of states.

Ranking of happiness around the world garners considerable attention, with over one million visits and downloads of the World Happiness Report each year. Below, the USA case demonstrates that a bias in rankings occurs when mean responses are taken at face value.

\figref{dailyPoll-isTen}(a)'s horizontal axis shows the distribution of state mean responses to the Cantril Ladder framing of life evaluation\footnote{See Appendix G.7 for the precise wording of the question.} in the 2019 (final) wave of the Gallup Daily Poll. Counter-intuitively, these means are uncorrelated with the fraction of respondents in each state who provided the answer “10” on the 0--10 scale (vertical axis). \figref{dailyPoll-isTen}(b) gives some suggestion as to why. States with higher high school completion rates have lower tendency to answer “10”. \figref{dailyPoll-isTen}(c) and (d) show an example of how much states can differ in terms of \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace. The weighted response distribution for Misssissippi, which has a high incidence of answer “10”, is remarkably different from that of Washington DC,\footnote{The survey also covered Washington DC. It is aggregated here alongside the states, even though it is entirely urban, unlike any state.} with the lowest incidence, even though their mean responses are similar.

figure*[figure* omitted — 1,195 chars of source]

With this motivation, Appendix Table F.7 presents estimates of a version of the mixture model Appendix Equation A.1 explaining individual responses with $\boldsymbol{x}=\boldsymbol{z}$ comprised of the logarithm of household income, along with a set of indicators for a five-level educational attainment question. As before, several parameters and posteriors of interest are: the fraction (\pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding Index (See page (ref))}\xspace) of respondents estimated to be using a simplified focal value scale; a mean of the latent life evaluation which would have been observed had all respondents chosen to use the full scale; and coefficients for the effect of income and education levels on the underlying (latent) life evaluations. These values are estimated separately for each state and can be compared in Appendix Table F.7 to the naive model, equivalent to an ordered logit, in which focal value behavior is not acknowledged.

figure*[figure* omitted — 636 chars of source]

\figref{histograms-three-biases-dailypoll} presents the distributions of biases in mean life evaluation and effects of high school completion and family income on life evaluations, obtained by comparing the ordered logit and mixture models. It shows that the Cantril Ladder question is in most states estimated to elicit highly positively-biased responses. In other words, the effect of “rounding up” to 10 (or to 5) outweighs any rounding down to 5 (or to 0), and is large. In many cases, the raw mean report is 0.1--0.2 higher than that inferred with the focal value correction, which is large given that the standard deviation of Cantril ladder means is 0.13 among states, and the standard deviation of individual responses nationally is only 1.89. This bias is larger for states with lower educational attainment.

\figref{histograms-three-biases-dailypoll} also shows that the distributions of biases in education effects and in income effects are both uniformly downwards at the state level. Reassuringly, the mixture-model estimated effects of educational attainment on wellbeing are overwhelmingly positive after the correction (Appendix Table F.7).

Lastly, \figref{daily-poll-ranks} presents state rankings for both the raw reported life evaluation and the estimated latent life evaluation. The overlapping estimate ranges reflect the typically imprecise nature of this kind of ranking, especially given the small sample size in some states (see Appendix Table F.6). There is also significant consistency (correlation 0.70) between the corrected and uncorrected rankings. Nevertheless, the shifts are considerable: more than a quarter of states shift by more than a quartile in the distribution (despite the overall correlation), 65% of states shift positions by 5 or more, and 37% shift by 10 or more.

figure*[figure* omitted — 467 chars of source]

Discussion and conclusion

The contributions of this paper are to

inparaenum[(i)] • explain a prominent feature of many subjective scale response distributions as the result of respondents simplifying the scale; • identify education and other proxies of numeracy as predictors of this “focal value rounding” behavior; • formulate a model and estimation strategy for predicting life satisfaction responses from individual and contextual circumstances which properly takes into account a mixture of reporting behavior used by respondents; • explore theoretically the biases possible due to the effect; • provide a way to estimate the degree (\pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding Index (See page (ref))}\xspace) of focal value rounding behavior; and • demonstrate the application of the estimation method and its significant impact for four published studies and surveys.

Clark-Oswald-JPubE1996 write “Counter to what neoclassical economic theory might lead one to expect, highly educated people appear less content. The effect is monotonic and well-defined”. This contradiction with neoclassical economic theory has generally held up to subsequent analysis over two decades but is partly resolved with the model described here, which takes into account a conspicuous empirical feature of the subjective wellbeing response function.

Income effects have been a focus in the study of wellbeing in economics since the field's inception, and an enormous literature exists around the magnitude of the income coefficient Easterlin-1974,Deaton-JEP2008-GWP,Clark-Frijters-Shields-JEL2008,Dolan-Peasgood-White-JEPsych2008,Easterlin-JEBO1995,Easterlin-EI2013-SWB-growth-public-policy,Ferrer-i-Carbonell-JPubE2005-veblen-GSOEP,Kapteyn-vanPraag-vanHerwaarden-EL1978,Luttmer-QJE2005,Senik-JES2005,VanPraag-EER-1973fei. Almost every economic study of \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace includes an estimate of the income effect, and typically other influences on life satisfaction are quantified in terms of their income “compensating differential”, i.e., the ratio between a coefficient of interest and the coefficient on income. Thus, the large corrections estimated here for the income coefficient indicate that material supports are slightly more effective for raising human wellbeing, as compared with the other --- especially social --- dimensions of life, than the literature has shown so far. According to the simulations, some downward bias can also occur for those other coefficients, especially when those dimensions of life are correlated with education, but there is little empirical evidence for this in the estimates carried out in this paper.

One next step for research is to examine international and cultural patterns in response functions. Effects will differ across countries according to where the average \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace level lies on the scale, and according to the income and education distribution. There may be additional international differences in the tendency to use focal values. Therefore, using the mixture model approach, both differences in education systems and more cultural drivers of \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace can be incorporated into international comparisons of \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace. Flexibly modeling each possible \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace response so as to allow for non-ordinal relationships between them, carried out here using multinomial logit, is a good starting point for detecting such response biases driven by cultural norms as well as numeracy. Despite the general evidence of good comparability of \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace patterns across cultures Helliwell-Barrington-Leigh-Harris-Huang-2010, it may still be possible to identify response biases towards central values or away from “extreme” values. One natural extension of the model described in this paper is to allow for the inclination to round (\pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace) to vary separately for each focal value, effectively creating a mixture of eight “types” in the case of three focal values.

A deeper analysis of panel data will also be important, through an extension to incorporate fixed effects into the model developed in this paper. Preliminary analysis of panel data with a 5-point scale for \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace, treating values 1, 3, and 5 as focal values, shows that the probability of \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace changing from the middle value is decreasing in education. Traditional 1st-differences approaches for panel fixed effects are invalid because, for instance, the dependence of the $3\rightarrow4$ transition is not the mirror of the $4\rightarrow3$ transition.

Another extension of the model used in this paper will be to incorporate instrumental variables. Fortunately, this is relatively straightforward in Bayesian estimation frameworks, in which a single-step estimation procedure for instrumental variables is natural, subject to the normal exclusion restrictions Dreze-Econometrica1976-Bayesian-Instrumental-Variables,Kleibergen-Zivot-JEconometrics2003-Bayesian-Instrumental-Variables.

As a proof of principle and in light of the descriptive evidence, this paper focuses on the idea of numeracy and on education as a primary predictor of \pdftooltip{\hyperlink{defFVR}{FVR}}{Focal Value Rounding (See page (ref))}\xspace. Understanding the role of secondary influences, such as other demographic variables, fatigue, the cost of time, or motivation with respect to the survey, may help to identify other biases or to design better surveys.

Survey and questionnaire interface design is a further topic of future work. While the present study carries out an ex post determination of how respondents have used a subjective numerical scale, it may make sense to give respondents this choice up front. An interactive survey interface could dynamically offer different degrees of precision or resolution in responses, thus accommodating variation in cognitive capacity and other differences in the confidence of respondents' answers. Open-ended graphical scales may be one means to accomplish this, but further research into ways to elicit a statement of precision from respondents would be valuable. The potential for creativity and innovation is high, given the increasing availability of technology during an interview.

Depending on one's perspective, the present findings on response behavior, happiness income coefficients, and mean response biases may be taken as a warning of how difficult it would be to realize the most ambitious implementations of \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace as a guide to policy Frijters-Clark-Krekel-Layard-BPP2020-Happy-Choice-SWB-as-goal-for-government,Frijters-Krekel-2021-SWB-policy-handbook,Barrington-Leigh-Escande-SIR2017-review-indicators,Barrington-Leigh-2016-CICbook-SWB-community-indicators,happiness-research-institute-2020-WALYs,Barrington-Leigh-SNSS2021-budgeting-for-happiness,UK-Treasury-2021-GreenBook-supplement-wellbeing,MacLennan-Stead-Rowlat-UK-Treasury-2021-life-satisfaction-approach or, conversely, as another reassuring example of the robustness of \hyperlink{defSWL}{\pdftooltip{LS}{Life Satisfaction (See page (ref))}}\xspace inference to potential flaws inherent in its cognitive complexity, and possibly even a defense of the rough magnitudes of estimated effects that have become so reproducible in study after study. I take away both of these messages.

Acknowledgements

I am grateful for discussion and comments from Fabian Lange, Kevin Lang, Andrew Oswald, Idrissa Ouili, Nadia DeLeon, two helpful referees, co-editor Keith Marzilli Ericson, and audiences in Vancouver, Chicago, Oxford, Cambridge, Montreal, and Winnipeg, among others. This work was supported by Canada's Social Sciences and Humanities Research Council (SSHRC) grant 435-2016-0531.

thebibliography{70} \expandafter\ifx\csname natexlab\endcsname\relax\def\natexlab#1{#1}\fi \ifx\xfnm\relax \def\xfnm[#1]{\unskip,\space#1}\fi \bibitem[{Alwin(1997)}]{Alwin-SMR1997-survey-question-scales} Alwin, D.F., 1997. \newblock {Feeling thermometers versus 7-point scales: Which are better?} \newblock Sociological Methods & Research 25, 318--340. \bibitem[{Barrington-Leigh(2014)}]{Barrington-Leigh-EQLWBR2014-consumption-externalities} Barrington-Leigh, C., 2014. \newblock Consumption Externalities, in: Michalos, A.C. (Ed.), Encyclopedia of Quality of Life and Well-Being Research. Springer, pp. 1248--1252. \newblock {\sc url}: http://wellbeing.ihsp.mcgill.ca/?p=pubs\#ConsumptionExternalities. \bibitem[{Barrington-Leigh(2016)}]{Barrington-Leigh-2016-CICbook-SWB-community-indicators} Barrington-Leigh, C., 2016. \newblock The role of subjective well-being as an organizing concept for community indicators, in: Meg Holden, R.P., Stevens, C. (Eds.), {Community Quality of Life and Wellbeing: Best Cases VII}. Springer. Community Quality-of-Life Indicators, pp. 19--34. \newblock {\sc doi}: 10.1007/978-3-319-54618-6. \bibitem[{Barrington-Leigh(2021)}]{Barrington-Leigh-SNSS2021-budgeting-for-happiness} Barrington-Leigh, C., 2021. \newblock Life satisfaction and sustainability: a policy framework. \newblock SN Social Sciences {\sc doi}: 10.1007/s43545-021-00185-8. \bibitem[{Barrington-Leigh and Escande(2018)}]{Barrington-Leigh-Escande-SIR2017-review-indicators} Barrington-Leigh, C., Escande, A., 2018. \newblock Measuring progress and well-being: A comparative review of indicators. \newblock Social Indicators Research 135, 893--925. \newblock {\sc doi}: 10.1007/s11205-016-1505-0. \bibitem[{Barrington-Leigh and Sloman(2016)}]{Barrington-Leigh-Sloman-IIPJ2016-aboriginal} Barrington-Leigh, C., Sloman, S., 2016. \newblock Life satisfaction among Aboriginals in the Canadian Prairies: Evidence from the Equality, Security and Community survey. \newblock The International Indigenous Policy Journal 7. \newblock {\sc doi}: 10.18584/iipj.2016.7.2.2. \bibitem[{Blanchflower and Oswald(2004)}]{Blanchflower-Oswald-JPubE2004} Blanchflower, D., Oswald, A., 2004. \newblock {Well-Being Over Time in Britain and the USA}. \newblock Journal of Public Economics 88, 1359--1386. \newblock {\sc doi}: 10.1016/S0047-2727(02)00168-8. \bibitem[{Blanchflower et al.(2014)Blanchflower, Bell, Montagnoli and Moro}]{Blanchflower-et-al-JMCB2014-unemployment-inflation} Blanchflower, D.G., Bell, D.N., Montagnoli, A., Moro, M., 2014. \newblock The happiness trade-off between unemployment and inflation. \newblock Journal of Money, Credit and Banking 46, 117--141. \newblock {\sc doi}: 10.1111/jmcb.12154. \bibitem[{Boes and Winkelmann(2006)}]{Boes-Winkelmann-ASA2006-ordered-response} Boes, S., Winkelmann, R., 2006. \newblock Ordered response models. \newblock Allgemeines Statistisches Archiv 1, 167--181. \newblock {\sc doi}: 10.1007/s10182-006-0228-y. \bibitem[{Bradburn et al.(2004)Bradburn, Sudman and Wansink}]{Bradburn-Norman-Sudman-Wansink-2004-questionnaire-design} Bradburn, N.M., Sudman, S., Wansink, B., 2004. \newblock Asking questions: the definitive guide to questionnaire design --- for market research, political polls, and social and health questionnaires. \newblock John Wiley & Sons. \bibitem[{Ferrer-i Carbonell(2005)}]{Ferrer-i-Carbonell-JPubE2005-veblen-GSOEP} Ferrer-i Carbonell, A., 2005. \newblock {Income and well-being: an empirical analysis of the comparison income effect}. \newblock Journal of Public Economics 89, 997--1019. \newblock {\sc doi}: 10.1016/j.jpubeco.2004.06.003. \bibitem[{Carpenter et al.(2017)Carpenter, Gelman, Hoffman, Lee, Goodrich, Betancourt, Brubaker, Guo, Li and Riddell}]{Carpenter-Gelman-Hoffman-Lee-Goodrich-Betancourt-Brubaker-Guo-Li-Riddell-JournalStatSoftware2017-Stan} Carpenter, B., Gelman, A., Hoffman, M.D., Lee, D., Goodrich, B., Betancourt, M., Brubaker, M., Guo, J., Li, P., Riddell, A., 2017. \newblock Stan: A probabilistic programming language. \newblock Journal of statistical software 76. \newblock {\sc doi}: 10.18637/jss.v076.i01. \bibitem[{Cheung and Lucas(2014)}]{Cheung-Lucas-QoLR2014-single-item-SWL-vs-SWLS-validity} Cheung, F., Lucas, R.E., 2014. \newblock Assessing the validity of single-item life satisfaction measures: results from three large samples. \newblock Quality of Life Research 23, 2809--2818. \newblock {\sc doi}: 10.1007/s11136-014-0726-4. \bibitem[{Clark et al.(2005)Clark, Etil\'e, Postel-Vinay, Senik and Van der Straeten}]{Clark-Etile-Postel-Vinay-Senik-VanderStraeten-EJ2005-latent-class-coefficients-SWB} Clark, A., Etil\'e, F., Postel-Vinay, F., Senik, C., Van der Straeten, K., 2005. \newblock Heterogeneity in reported well-being: Evidence from twelve european countries. \newblock The Economic Journal 115, C118--C132. \newblock {\sc doi}: 10.1111/j.0013-0133.2005.00983.x. \bibitem[{Clark and Oswald(1996)}]{Clark-Oswald-JPubE1996} Clark, A., Oswald, A., 1996. \newblock {Satisfaction and comparison income}. \newblock Journal of Public Economics 61, 359--381. \newblock {\sc doi}: 10.1016/0047-2727(95)01564-7. \bibitem[{Clark et al.(2019)Clark, Fl{\`e}che, Layard, Powdthavee and Ward}]{Clark-Fleche-Layard-Powdthavee-Ward-2019-origins-of-happiness} Clark, A.E., Fl{\`e}che, S., Layard, R., Powdthavee, N., Ward, G., 2019. \newblock The origins of happiness: the science of well-being over the life course. \newblock Princeton University Press. \bibitem[{Clark et al.(2008)Clark, Frijters and Shields}]{Clark-Frijters-Shields-JEL2008} Clark, A.E., Frijters, P., Shields, M.A., 2008. \newblock {Relative Income, Happiness, and Utility: An Explanation for the Easterlin Paradox and Other Puzzles}. \newblock The Journal of Economic Literature 46, 95--144. \newblock {\sc doi}: 10.1257/jel.46.1.95. \bibitem[{Conti and Pudney(2011)}]{Conti-Pudney-RES2011-SWB-survey-design-BHPS-distribution} Conti, G., Pudney, S., 2011. \newblock Survey design and the analysis of satisfaction. \newblock Review of Economics and Statistics 93, 1087--1093. \newblock {\sc doi}: 10.1162/REST_a_00202. \bibitem[{Deaton(2008)}]{Deaton-JEP2008-GWP} Deaton, A., 2008. \newblock {Income, health and wellbeing around the world: Evidence from the Gallup World Poll}. \newblock Journal of Economic Perspectives 22, 53. \newblock {\sc doi}: 10.1257/jep.22.2.53. \bibitem[{{Department of Finance}(2021)}]{FinanceCanada-QoL-framework-2021-April} {Department of Finance}, 2021. \newblock Measuring What Matters: Toward a Quality of Life Strategy for Canada. \newblock {\sc url}: https://www.canada.ca/en/department-finance/services/publications/measuring-what-matters-toward-quality-life-strategy-canada.html. \bibitem[{Dolan et al.(2011)Dolan, Layard and Metcalfe}]{Dolan-Layard-Metcalfe-ONS2011-SWB-public-policy} Dolan, P., Layard, R., Metcalfe, R., 2011. \newblock Measuring subjective well-being for public policy. \newblock Office for National Statistics paper . \bibitem[{Dolan et al.(2008)Dolan, Peasgood and White}]{Dolan-Peasgood-White-JEPsych2008} Dolan, P., Peasgood, T., White, M., 2008. \newblock {Do we really know what makes us happy? A review of the economic literature on the factors associated with subjective well-being}. \newblock Journal of Economic Psychology 29, 94--122. \newblock {\sc doi}: 10.1016/j.joep.2007.09.001. \bibitem[{Dr\`eze(1976)}]{Dreze-Econometrica1976-Bayesian-Instrumental-Variables} Dr\`eze, J.H., 1976. \newblock Bayesian limited information analysis of the simultaneous equations model. \newblock Econometrica 44, 1045--1075. \newblock {\sc url}: http://www.jstor.org/stable/1911544. \bibitem[{Easterlin(1974)}]{Easterlin-1974} Easterlin, R., 1974. \newblock {Does Economic Growth Improve the Human Lot? Some Empirical Evidence}, in: David, P., Reder, M. (Eds.), Nations and Households in Economic Growth: Essays in Honour of Moses Abramovitz. Academic Press, pp. 98--125. \bibitem[{Easterlin(1995)}]{Easterlin-JEBO1995} Easterlin, R.A., 1995. \newblock Will raising the incomes of all increase the happiness of all? \newblock Journal of Economic Behavior & Organization 27, 35--47. \bibitem[{Easterlin(2013)}]{Easterlin-EI2013-SWB-growth-public-policy} Easterlin, R.A., 2013. \newblock {Happiness, growth, and public policy}. \newblock Economic Inquiry 51, 1--15. \newblock {\sc doi}: 10.1111/j.1465-7295.2012.00505.x. \bibitem[{Everitt(1988)}]{Everitt-SPL1988-mixture-model-ordered-logit} Everitt, B.S., 1988. \newblock A finite mixture model for the clustering of mixed-mode data. \newblock Statistics & probability letters 6, 305--309. \bibitem[{Everitt and Merette(1990)}]{Everitt-Merette-JAS1990-mixture-model-ordered-logit} Everitt, B.S., Merette, C., 1990. \newblock The clustering of mixed-mode data: a comparison of possible approaches. \newblock Journal of Applied Statistics 17, 283--297. \bibitem[{Exton et al.(2015)Exton, Smith and Vandendriessche}]{Exton-Smith-Vandendriessche-OECD2015} Exton, C., Smith, C., Vandendriessche, D., 2015. \newblock Comparing happiness across the world: Does culture matter? \newblock OECD Statistics Working Papers {\sc doi}: 10.1787/5jrqppzd9bs2-en. \bibitem[{{Ferrer-i-Carbonell} and Frijters(2004)}]{Ferrer-i-Carbonell-Frijters-EJ2004} {Ferrer-i-Carbonell}, A., Frijters, P., 2004. \newblock {How Important is Methodology for the estimates of the determinants of Happiness?} \newblock The Economic Journal 114, 641--659. \newblock {\sc doi}: 10.1111/j.1468-0297.2004.00235.x. \bibitem[{Frick et al.(2006)Frick, Goebel, Schechtman, Wagner and Yitzhaki}]{Frick-Goebel-Schechtman-Wagner-Yitzhaki-SMR2006-SOEP-attrition-endpoints} Frick, J.R., Goebel, J., Schechtman, E., Wagner, G.G., Yitzhaki, S., 2006. \newblock {Using analysis of Gini (ANOGI) for detecting whether two subsamples represent the same universe: The German Socio-Economic Panel Study (SOEP) experience}. \newblock Sociological Methods & Research 34, 427--468. \newblock {\sc doi}: 10.1177/0049124105283109. \bibitem[{Frijters et al.(2020)Frijters, Clark, Krekel and Layard}]{Frijters-Clark-Krekel-Layard-BPP2020-Happy-Choice-SWB-as-goal-for-government} Frijters, P., Clark, A.E., Krekel, C., Layard, R., 2020. \newblock A happy choice: wellbeing as the goal of government. \newblock Behavioural Public Policy 4, 126--165. \newblock {\sc doi}: 10.1017/bpp.2019.39. \bibitem[{Frijters and Krekel(2021)}]{Frijters-Krekel-2021-SWB-policy-handbook} Frijters, P., Krekel, C., 2021. \newblock {A Handbook for Wellbeing Policy-Making: History, Theory, Measurement, Implementation, and Examples}. \newblock Oxford University Press. \newblock {\sc doi}: 10.1093/oso/9780192896803.001.0001. \bibitem[{Giustinelli et al.(2020)Giustinelli, Manski and Molinari}]{Giustinelli-Manski-Molinari-JEconometrics2020-rounding} Giustinelli, P., Manski, C.F., Molinari, F., 2020. \newblock Tail and center rounding of probabilistic expectations in the health and retirement study. \newblock Journal of Econometrics {\sc doi}: 10.1016/j.jeconom.2020.03.020. \bibitem[{Goff et al.(2018)Goff, Helliwell and Mayraz}]{Goff-Helliwell-Mayraz-EI2018-SWB-inequality} Goff, L., Helliwell, J.F., Mayraz, G., 2018. \newblock Inequality of subjective well-being as a comprehensive measure of inequality. \newblock Economic Inquiry 56, 2177--2194. \newblock {\sc doi}: 10.1111/ecin.12582. \bibitem[{Grimes(2021)}]{Grimes-2021-chapter-budgeting-for-wellbeing} Grimes, A., 2021. \newblock Budgeting for wellbeing, in: A Modern Guide to Wellbeing Research. Edward Elgar Publishing, pp. 266--281. \bibitem[{Hamamura et al.(2008)Hamamura, Heine and Paulhus}]{Hamamura-Heine-Paulhus-PID2007-cultural-differences-response-styles} Hamamura, T., Heine, S.J., Paulhus, D.L., 2008. \newblock Cultural differences in response styles: The role of dialectical thinking. \newblock Personality and Individual differences 44, 932--942. \newblock {\sc doi}: 10.1016/j.paid.2007.10.034. \bibitem[{Hamilton et al.(2016)Hamilton, Helliwell and Woolcock}]{Hamilton-Helliwell-Woolcock-NBER2016-social-capital-wealth} Hamilton, K., Helliwell, J.F., Woolcock, M., 2016. \newblock {Social Capital, Trust and Well-being in the Evaluation of Wealth} {\sc doi}: 10.3386/w22556. \bibitem[{{Happiness Research Institute}(2020)}]{happiness-research-institute-2020-WALYs} {Happiness Research Institute}, 2020. \newblock Wellbeing Adjusted Life Years: A universal metric to quantify the happiness return on investment. \newblock {\sc url}: https://www.happinessresearchinstitute.com/waly-report. \bibitem[{Hasegawa and Ueda(2011)}]{Hasegawa-Ueda-JSE2011-SWB-inequality} Hasegawa, H., Ueda, K., 2011. \newblock Measuring inequality of subjective well-being: A bayesian approach. \newblock The Journal of Socio-Economics 40, 700--708. \newblock {\sc doi}: 10.1016/j.socec.2011.05.009. \bibitem[{Helliwell and Barrington-Leigh(2011)}]{Helliwell-Barrington-Leigh-2010-social-capital-worth} Helliwell, J., Barrington-Leigh, C., 2011. \newblock How much is social capital worth?, in: Jetten, J., Haslam, C., Haslam, S.A. (Eds.), The Social Cure: Identity, Health, and Well-being. Taylor and Francis, pp. 55--71. \newblock {\sc url}: http://wellbeing.research.mcgill.ca/publications/w16025-and-appendix.pdf. \bibitem[{Helliwell et al.(2010)Helliwell, Barrington-Leigh, Harris and Huang}]{Helliwell-Barrington-Leigh-Harris-Huang-2010} Helliwell, J., Barrington-Leigh, C., Harris, A., Huang, H., 2010. \newblock International evidence on the social context of well-being, in: Diener, E., Helliwell, J., Kahneman, D. (Eds.), International Differences in Well-Being. Oxford University Press, pp. 213--229. \bibitem[{Helliwell and Putnam(2004)}]{Helliwell-Putnam-PTRSL2004} Helliwell, J., Putnam, R., 2004. \newblock {The social context of well-being.} \newblock Philos Trans R Soc Lond B Biol Sci 359, 1435--46. \newblock {\sc doi}: 10.1098/rstb.2004.1522. \bibitem[{Helliwell and Putnam(2007)}]{Helliwell-Putnam-EEJ2007} Helliwell, J., Putnam, R., 2007. \newblock {Education and Social Capital}. \newblock Eastern Economic Journal 33, 1--19. \newblock {\sc doi}: 10.1057/eej.2007.1. \bibitem[{Kapteyn et al.(1978)Kapteyn, van Praag and van Herwaarden}]{Kapteyn-vanPraag-vanHerwaarden-EL1978} Kapteyn, A., van Praag, B., van Herwaarden, F., 1978. \newblock {Individual welfare functions and social reference spaces}. \newblock Economics Letters 1, 173--177. \bibitem[{Khorramdel et al.(2019)Khorramdel, von Davier and Pokropek}]{Khorramdel-vonDavier-Pokropek-BJMSP2019-extreme-response-styles-model} Khorramdel, L., von Davier, M., Pokropek, A., 2019. \newblock Combining mixture distribution and multidimensional irtree models for the measurement of extreme response styles. \newblock British Journal of Mathematical and Statistical Psychology 72, 538--559. \newblock {\sc doi}: 10.1111/bmsp.12179. \bibitem[{Kleibergen and Zivot(2003)}]{Kleibergen-Zivot-JEconometrics2003-Bayesian-Instrumental-Variables} Kleibergen, F., Zivot, E., 2003. \newblock Bayesian and classical approaches to instrumental variable regression. \newblock Journal of Econometrics 114, 29--72. \newblock {\sc doi}: 10.1016/S0304-4076(02)00219-1. \bibitem[{Kroh et al.(2006)}]{Kroh-DRAFT2006} Kroh, M., et al., 2006. \newblock An experimental evaluation of popular well-being measures . \bibitem[{Landua(1992)}]{Landua-SIR1992-panel-focal-values} Landua, D., 1992. \newblock {An attempt to classify satisfaction changes: Methodological and content aspects of a longitudinal problem}. \newblock Social Indicators Research 26, 221--241. \newblock {\sc doi}: 10.1007/BF00286560. \bibitem[{Lau et al.(2005)Lau, Cummins and McPherson}]{Lau-Cummins-Mcpherson-SIR2005-cultural-bias-SWB} Lau, A., Cummins, R., McPherson, W., 2005. \newblock {An Investigation into the Cross-Cultural Equivalence of the Personal Wellbeing Index}. \newblock Social Indicators Research 72, 403 -- 430. \newblock {\sc doi}: 10.1007/s11205-004-0561-z. \bibitem[{Layard(2011)}]{Layard-2011-happiness-lessons-new-science} Layard, P.R.G., 2011. \newblock Happiness: Lessons from a new science. \newblock Penguin UK. \bibitem[{Levinson(2012)}]{Levinson-JPubE2012-happiness-pollution} Levinson, A., 2012. \newblock Valuing public goods using happiness data: The case of air quality. \newblock Journal of Public Economics 96, 869--880. \newblock {\sc doi}: 10.1016/j.jpubeco.2012.06.007. \bibitem[{Lewbel(2019)}]{Lewbel-JEL2019-identification-meaning} Lewbel, A., 2019. \newblock The identification zoo: Meanings of identification in econometrics. \newblock Journal of Economic Literature 57, 835--903. \newblock {\sc doi}: 10.1257/jel.20181361. \bibitem[{Luttmer(2005)}]{Luttmer-QJE2005} Luttmer, E.F.P., 2005. \newblock Neighbors as negatives: Relative earnings and well-being. \newblock Quarterly Journal of Economics 120, 963--1002. \newblock {\sc doi}: 10.1093/qje/120.3.963. \bibitem[{MacLennan et al.(2021)MacLennan, Stead and Rowlatt}]{MacLennan-Stead-Rowlat-UK-Treasury-2021-life-satisfaction-approach} MacLennan, S., Stead, I., Rowlatt, A., 2021. \newblock Wellbeing discussion paper: monetisation of life satisfaction effect sizes: A review of approaches and proposed approach. \newblock Technical Report. UK Government. \newblock {\sc url}: https://www.gov.uk/government/publications/green-book-supplementary-guidance-wellbeing. \bibitem[{OECD(2013)}]{OECD-2013-guidelines-measuring-SWB} OECD, 2013. \newblock {OECD Guidelines on Measuring Subjective Well-being}. \newblock OECD Publishing. \newblock {\sc doi}: 10.1787/9789264191655-en. \bibitem[{Powdthavee(2008)}]{Powdthavee-JSE2008-valuing-social-relationships} Powdthavee, N., 2008. \newblock Putting a price tag on friends, relatives, and neighbours: Using surveys of life satisfaction to value social relationships. \newblock Journal of Socio-Economics 37, 1459--1480. \newblock {\sc doi}: 10.1016/j.socec.2007.04.004. \bibitem[{Powdthavee et al.(2015)Powdthavee, Lekfuangfu and Wooden}]{Powdthavee-Lekfuangfu-Wooden-JBEE2015-education-SWL-Australia-HILDA} Powdthavee, N., Lekfuangfu, W.N., Wooden, M., 2015. \newblock What's the good of education on our overall quality of life? a simultaneous equation model of education and life satisfaction for australia. \newblock Journal of Behavioral and Experimental Economics 54, 10--21. \newblock {\sc doi}: 10.1016/j.socec.2014.11.002. \bibitem[{Riddell et al.(2020)Riddell, Hartikainen and Carter}]{pystan2.19.1.1} Riddell, A., Hartikainen, A., Carter, M., 2020. \newblock {PyStan (2.19.1.1)}. \newblock PyPI. \bibitem[{Saris et al.(1998)Saris, Van Wijk and Scherpenzeel}]{Saris-vanWijk-Scherpenzeel-SIR1998} Saris, W., Van Wijk, T., Scherpenzeel, A., 1998. \newblock {Validity and reliability of subjective social indicators}. \newblock Social indicators research 45, 173--199. \bibitem[{Senik(2005)}]{Senik-JES2005} Senik, C., 2005. \newblock {Income distribution and well-being: what can we learn from subjective data?} \newblock Journal of Economic Surveys 19, 43--63. \newblock {\sc doi}: 10.1111/j.0950-0804.2005.00238.x. \bibitem[{{Stan Development Team}(2018)}]{Stan2018} {Stan Development Team}, 2018. \newblock Stan modeling language users guide and reference manual, version 2.18.0. \newblock {\sc url}: \texttt{http://mc-stan.org/}. \bibitem[{Stevenson and Wolfers(2008)}]{Stevenson-Wolfers-NBER2008-happiness-inequality-United-States} Stevenson, B., Wolfers, J., 2008. \newblock Happiness inequality in the united states {\sc url}: \texttt{http://www.nber.org/papers/w14220}. \bibitem[{Stutzer(2004)}]{Stutzer-JEBO2004} Stutzer, A., 2004. \newblock {The Role of Income Aspirations in Individual Happiness}. \newblock Journal of Economic Behavior and Organization 54, 89--109. \newblock {\sc doi}: 10.1016/j.jebo.2003.04.003. \bibitem[{Stutzer and Frey(2006)}]{Stutzer-Frey-JSE2006-marriage-causality-happiness} Stutzer, A., Frey, B.S., 2006. \newblock Does marriage make people happy, or do happy people get married? \newblock The Journal of Socio-Economics 35, 326--347. \newblock {\sc doi}: 10.1016/j.socec.2005.11.043. the Socio-Economics of Happiness. \bibitem[{Uebersax(1999)}]{Uebersax-APM1999-mixture-model-ordinal} Uebersax, J.S., 1999. \newblock Probit latent class analysis with dichotomous or ordered category measures: conditional independence/dependence models. \newblock Applied Psychological Measurement 23, 283--297. \newblock {\sc doi}: 10.1177/01466219922031400. \bibitem[{{UK Treasury}(2021)}]{UK-Treasury-2021-GreenBook-supplement-wellbeing} {UK Treasury}, 2021. \newblock {Wellbeing Guidance for Appraisal: Supplementary Green Book Guidance}. \newblock Technical Report. UK Government. \newblock {\sc url}: \texttt{https://www.gov.uk/government/publications/green-book-supplementary-guidance-wellbeing}. \bibitem[{Van Praag and Kapteyn(1973)}]{VanPraag-EER-1973fei} Van Praag, B., Kapteyn, A., 1973. \newblock {Further evidence on the individual welfare function of income: An empirical investigation in The Netherlands}. \newblock European Economic Review 4, 33--62. \bibitem[{Weng(2004)}]{Weng-EPM2004-Likert-labels-and-number-of-options} Weng, L.J., 2004. \newblock Impact of the number of response categories and anchor labels on coefficient alpha and test-retest reliability. \newblock Educational and Psychological Measurement 64, 956--972. \newblock {\sc doi}: 10.1177/0013164404268674. \bibitem[{Wu(2005)}]{Wu-SEE2005-plausible-values-PISA} Wu, M., 2005. \newblock The role of plausible values in large-scale surveys. \newblock Studies in Educational Evaluation 31, 114--128. \newblock {\sc doi}: 10.1016/j.stueduc.2005.05.005. measurement, Evaluation, and Statistical Analysis.