EconBase
← Back to paper

Structural models for policy-making: Coping with parametric uncertainty

Extracted main text — title through conclusion, appendix excluded. This is what our citation measures are computed over, published so the extraction can be checked by eye.

75,123 characters · 12 sections · 75 citation commands

Rendered from LaTeX for readability, not typeset faithfully. Citation keys are highlighted; maths is left as source; figures, tables and equation environments are summarised rather than reproduced; unrecognised commands are greyed out so nothing is silently dropped. Email addresses are removed.

Structural models for policy-making

\setcounter{page}{1} \thispagestyle{empty}

abstractThe ex-ante evaluation of policies using structural econometric models is based on estimated parameters as a stand-in for the true parameters. This practice ignores uncertainty in the counterfactual policy predictions of the model. We develop a generic approach that deals with parametric uncertainty using uncertainty sets and frames model-informed policy-making as a decision problem under uncertainty. The seminal human capital investment model by Keane.1997 provides a well-known, influential, and empirically-grounded test case. We document considerable uncertainty in the models's policy predictions and highlight the resulting policy recommendations obtained from using different formal rules of decision-making under uncertainty.\\
tabular[tabular omitted — 157 chars of source]

\setcounter{page}{1} \FloatBarrier

Introduction

Structural microeconometricians use highly parameterized computational models to investigate economic mechanisms, predict the impact of proposed policies, and inform optimal policy-making Wolpin.2013. These models represent deep structural relationships of theoretical economic models invariant to policy changes Hood.1953. The sources of uncertainty in such an analysis are ubiquitous Saltelli.2020. For example, models are often misspecified, there are numerical approximation errors in their implementation, and model parameters are uncertain. Therefore, most disciplines require a proper account of uncertainty before using computational models to inform decision-making Council.2012,SAPEA.2019.\\

The following study focuses on parametric uncertainty in structural microeconometric models that are estimated on observed data. Researchers often do not account for parametric uncertainty and conduct an as-if analysis in which the point estimates serve as a stand-in for the true model parameters. They then continue to study the implications of their models at the point estimates Adda.2017,Blundell.2016,Eckstein.2019,Eisenhauer.2015b and rank competing policy proposals based on the point predictions alone Blundell.2012,Cunha.2010,Gayle.2019,Todd.2006. In fact, Keane.2011d states in their handbook article that they are unaware of any applied work that reports the distribution of policy predictions under parametric uncertainty. To the best of our knowledge, this statement remains true more than a decade later. Consequently, economists risk accepting fragile findings as facts, ignoring the trade-off between model complexity and prediction uncertainty, and neglecting to frame policy advice as a decision problem under uncertainty.\\

To mitigate these shortcomings, we develop an approach that copes with parametric uncertainty in structural microeconometric models and embeds model-informed policy-making in a decision-theoretic framework. Ideally, policy-makers fix the parameter space ex-ante and then evaluate the policy options according to decision rules. However, this approach is often computationally intractable. We, therefore, follow Manski.2021's suggestion and, instead of using the parameter estimates as-if they were true, incorporate uncertainty in the analysis by treating the estimated confidence set as-if it is correct. We use the confidence set to construct an uncertainty set that is anchored in empirical estimates, statistically meaningful, and computationally tractable Ben-Tal.2013. Instead of just focusing on the point estimates, we evaluate counterfactual policies based on all parametrizations within the uncertainty set.\\

We draw on statistical decision theory Manski.2013 to deal with the uncertainty in counterfactual predictions. This approach promotes a well-reasoned and transparent policy process. Before a decision, it clarifies trade-offs between choices Gilboa.2018. Afterward, decision-theoretic principles allow constituents to scrutinize the coherence of choices Gilboa.2020, ease the ex-post justification Berger.2021, and facilitate the communication of uncertainty Manski.2019.\\

We tailor our approach to the class of Eckstein-Keane-Wolpin (EKW) models Aguirregabiria.2010. Labor economists often use EKW models to learn about human capital investment and consumption-saving decisions and predict the impact of proposed reforms to education policy and welfare programs Keane.2011d,Low.2017,Blundell.2017. The analysis of these models poses serious computational challenges. During estimation, EKW models are solved thousands of times and even a single solution often takes several minutes. Thus, a decision-theoretic ex-ante analysis of alternative decision rules across the whole parameter space, as intended by Wald.1950, is infeasible. Instead we construct an uncertainty set, a subset of the whole parameter space, and deal with the ex-post uncertainty after estimating the model. This compromise allows us to garner the benefits of using statistical decision theory to shape policy-making under uncertainty while ensuring the computational tractability of our analysis.\\

As an example of our approach, we analyze the seminal human capital investment model by Keane.1997 as a well-known, empirically grounded, and computationally demanding test case. We follow the authors and estimate the model on the National Longitudinal Survey of Youth 1979 (NLSY79) NLSY.2019 using the original dataset and reproduce all core results. We revisit their predictions for the impact of a tuition subsidy on completed years of schooling. The economics of the model implies that the nonlinear mapping between the model parameters and predictions is truncated at zero, and we thus use the Confidence Set (CS) bootstrap Woutersen.2019 to estimate the confidence set for the counterfactuals. We document considerable uncertainty in the policy predictions and highlight the resulting policy recommendations from different formal rules on decision-making under uncertainty.\\

Our work extends existing research exploring the sensitivity of implications and predictions to parametric uncertainty in macroeconomics and climate economics. For example, Harenberg.2019 study uncertainty propagation and sensitivity analysis for a standard real business cycle model. Cai.2019 examine how uncertainties and risks in economic and climate systems affect the social cost of carbon. However, neither of them estimates their model on data. Instead, they rely on expert judgments to inform the degree of parametric uncertainty. They do not investigate the consequences of uncertainty for policy decisions in a decision-theoretic framework.\\

We complement a burgeoning literature on the sensitivity analysis of policy predictions in light of model or moment misspecification. For example, Andrews.2017 and Andrews.2020 treat the model specification as given and then analyze the sensitivity of the parameter estimates to the misspecification of the moments used for estimation. Christensen.2019 study global sensitivity of the model predictions to misspecification of the distribution of unobservables. Jorgensen.2021 provides a local measure for the sensitivity of counterfactuals to model parameters that are fixed before the estimation of the model.\footnote{For other examples, see Armstrong.2021, Bonhomme.2020, Bugni.2019, and Mukhin.2018.} This literature does not embed the counterfactual predictions in a decision-theoretic setting. Recent work by Kalouptsidi.2020, Kalouptsidi.2021, and Norets.2014 studies (partial) identification and inference on counterfactuals. However, they all adopt the setup outlined in Rust.1987 and exploit the additive separability of the immediate utility function between observed and unobserved state variables, which does not apply to EKW models. In related work, Blesch.2021 conduct a decision-theoretic ex-ante analysis to determine optimal decision rules in Rust.1987's stochastic dynamic investment model where the decision-maker directly accounts for uncertainty in the model's transition dynamics. They only consider uncertainty in a subset of the model's parameters which are estimated outside the model and remain fixed to their point estimates during the analysis.\\

In Section (ref), we describe the decision-theoretic framework for making model-informed decisions under parametric uncertainty using an illustrative example. After summarizing the empirical setting of Keane.1997 in Section (ref), we present our results in Section (ref). We complete our analysis in Section (ref) with a brief conclusion and outlook.

\FloatBarrier

Structural models for policy-making

In the following section, we discuss uncertainty propagation and the common practice of using estimated parameters as a plug-in replacement for the true model parameters. We then explore the limitations of this strategy and introduce our alternative approach, in which we implement estimated confidence sets to construct uncertainty sets. In so doing, we are able to cope with uncertain policy predictions in a proper decision-theoretic framework. \\

At a high level, a structural microeconometric model provides a mapping $\mathcal{M}(\bm{\theta})$ between the $l$ model parameters $\bm{\theta} \in \bm{\Theta}$ and a quantity $y$ that is of interest to policy-makers.

align*[align* omitted — 102 chars of source]

A policy $g \in \mathcal{G}$ changes the mapping to $\mathcal{M}_g(\bm{\theta})$ and produces a counterfactual $y_g$.\\

Estimation of a baseline model $\mathcal{M}(\bm{\theta})$ describing the status-quo on observed data allows researchers to learn about the true parameters. Frequentist estimation procedures such as maximum likelihood estimation and the method of simulated moments produce a point estimate $\hat{\bm{\theta}}$. However, uncertainty about the true parameters remains.\\

Previewing our empirical analysis of Keane.1997, our $\mathcal{M}$ is provided by a dynamic model of human capital accumulation, which we estimate on observed schooling and labor market decisions using simulated maximum likelihood estimation. The policy $g$ is the implementation of a college tuition subsidy, and the counterfactual is the level of completed schooling in the population. Example parameters that drive the economics of the model are time preferences of individuals, the return to schooling, and the transferability of work experience across occupations.\\

The following illustrative example highlights our key points. We consider two policies $g \in\{1, 2\}$ that result in two different mappings $ (\mathcal{M}_1, \mathcal{M}_2)$ of the same scalar $\theta$ to a counterfactual $y_g$. Higher values of $y_g$ are more desirable for a policy-maker. The point estimate $\hat{\theta}$ is determined by estimating a baseline model on an observed dataset. We denote the probability density function of its sampling distribution by $f_{\hat{\theta}}$.\\

Under the first policy, the counterfactual is an increasing nonlinear function of $\theta$. In the case of the second policy, the relationship is decreasing and linear.

figure[figure omitted — 365 chars of source]

\FloatBarrier Figure (ref) traces the counterfactual from both models over a range of the parameter. At the point estimate, both models yield the same value for the counterfactual. Once we account for uncertainty in our estimates of the true parameter, deciding which policy to adopt becomes less straightforward: for higher values of $\theta$, the first policy is preferred, while the opposite is true for lower values.

Uncertainty sets

Manski.2021 suggests acknowledging parametric uncertainty by working with estimated confidence sets instead of point estimates. A confidence set $\bm{\Theta}(\alpha) \subset \bm{\Theta}$ covers the true parameters, from an ex-ante point of view, with a predetermined coverage probability of $(1 - \alpha)$. Proceeding with our analysis, we refine the status quo procedure, in which estimated parameter values serve as a stand-in for the model's true parametrization. Instead, we assume the estimated confidence set for the parameters $\hat{\bm{\Theta}}(\alpha)$ and the counterfactual $\hat{\bm{\Theta}}_{y_g}(\alpha)$ are correct and analyze policy decisions accordingly. \\

Based on the estimated confidence sets, we construct so-called uncertainty sets for the parameters $\mathcal{U}(\alpha)$ and the prediction $\mathcal{U}_{y_g}(\alpha)$ by only considering parameterizations that we cannot reject based on a hypothesis test with confidence level $1 - \alpha$. This approach ensures the tractability of our decision-theoretic analysis, as the uncertainty set of the parameters is much smaller than the whole parameter space of the model. We adopt this procedure from the literature on data-driven robust optimization in operations research Ben-Tal.2013,Bertsimas.2018.

Statistical decision theory

In our setting, a policy-maker relies on a structural model with an uncertain parametrization to map alternative policies to counterfactual predictions. In most cases, the preferred policy depends on the model's uncertain true parameters. We, therefore, draw on statistical decision theory to organize the decision-making process Gilboa.2009,Marinacci.2015.\\

Returning to our example, we rank the two policies according to alternative statistical decision rules using an uncertainty set derived from a confidence set with a $90\%$ coverage probability. In what follows, we postulate a simple linear utility function $U(y_g)$ to describe the policy-maker's preferences.\footnote{We assume that the sampling distribution of the point estimate is normal with a mean of three and a standard deviation of three-fourths. We can derive the uncertainty sets directly and simply consider realizations of $\theta\in[1.76, 4.23]$.}\\

Figure (ref) shows the implied sampling distribution of the predictions for the two alternative policies and the corresponding uncertainty sets $\mathcal{U}_{y_g}(0.1)$. The mapping $\mathcal{M}_1$ is highly nonlinear, while the mapping $\mathcal{M}_{2}$ is linear. When evaluated at the point estimate, the counterfactual is the same under both policies, so a policy-maker is indifferent. However, the spread of the uncertainty set differs considerably.\\

figure[figure omitted — 168 chars of source]

\FloatBarrier

Decision theory proposes a variety of different rules for reasonable decisions in this setting. We explore the following four: (1) as-if optimization, (2) maximin criterion, (3) minimax regret rule, and (4) subjective Bayes.\\

As-if optimization describes the predominant practice. The estimation of the model produces point estimates that serve as a plug-in for the true parameters. The decision maximizes the utility at the point estimate. More formally,

align*[align* omitted — 92 chars of source]

Given our example, an as-if policy-maker is indifferent between the two policies, since both policies result in the same counterfactual at the point estimates as indicated by the dashed line in Figure (ref).\\

The maximin criterion and minimax regret rule are two common alternatives that favor actions that work uniformly well over all possible parameters in the uncertainty set. This approach departs from as-if optimization, which only considers a policy's performance at a single point in the uncertainty set. The maximin decision Gilboa.1989, Wald.1950 is determined by computing the minimum utility for each policy within the uncertainty set and choosing the one with the highest worst-case outcome. Stated concisely,

align*[align* omitted — 127 chars of source]

Returning to Figure (ref), a maximin policy-maker prefers $g_2$ as the worst-case outcome. Within the uncertainty set, $\underline{y}_2$ is better than under the alternative policy, $g_1$.\\

The minimax regret rule Manski.2004, Niehans.1948 computes the maximum regret for each policy over the whole uncertainty set and chooses the policy that minimizes the maximum regret. The regret of choosing a policy $g$ for a given parameterization of the model is the difference between the maximum possible utility achieved from adopting $\tilde{g} \in \mathcal{G}$ and the actual utility obtained. The decision maximizes:

align*[align* omitted — 238 chars of source]

Figure (ref) compares our two policy examples over the uncertainty sets. A policy-maker adopting policy $g_1$ regrets his choice for small values of the model parameter, while the opposite is true for larger values. The regret of each policy is maximized at the boundaries of the uncertainty set. Maximum regret is minimized when a policy-maker chooses $g_1$. It corresponds to the difference in the counterfactual at the lower boundary of the uncertainty set instead of the larger difference at its upper bound. This outcome contradicts the maximin decision in which policy $g_2$ is preferred.

figure[figure omitted — 183 chars of source]

\FloatBarrier

Each decision rule presented so far focuses on a single point in the uncertainty set as the policy's relevant performance measure. Bayesian approaches aggregate a policy's performance over the complete uncertainty set.\\

Maximization of the subjective expected utility Savage.1954 requires the policy-maker to place a subjective probability distribution $f_{\bm{\theta}}$ over the parameters in the uncertainty set. A policy-maker then selects the alternative with the highest expected subjective utility. Formally,

align*[align* omitted — 153 chars of source]

Applying a uniform distribution to our example, a policy-maker chooses $g_1$, which performs well for high values of $\bm{\theta}$ and still reasonably well for low values.

\FloatBarrier

Eckstein-Keane-Wolpin models

We now present the general structure of Eckstein-Keane-Wolpin (EKW) models Aguirregabiria.2010 and their solution approach. We then turn to the customized version used by Keane.1997 to study the career decisions of young men and investigate the consequences of parametric uncertainty in this empirically-grounded and computationally demanding setting. We outline their model's basic setup, provide some descriptive statistics of the empirical data used in our estimation, and then discuss the core findings.

General structure

EKW models describe sequential decision-making under uncertainty Gilboa.2009, Machina.2014. At time $t = 1, \hdots, T$ each individual observes the state of their choice environment $s_t\in S$ and chooses an action $a_t$ from the set of admissible actions $\mathcal{A}$. The decision has two consequences: an individual receives an immediate utility $u_t(s_t, a_t)$ and their environment evolves to a new state $s_{t + 1}$. The transition from $s_t$ to $s_{t + 1}$ is affected by the action but remains uncertain. Since individuals are forward-looking, they do not simply choose the alternative with the highest immediate utility. Instead, they take the future consequences of their actions into account.\\

A policy $\pi =(d^\pi_1, \hdots, d^\pi_T)$ provides the individual with instructions for choosing an action in any possible future state. It is a sequence of decision rules $d^\pi_t$ that specify the action $d^\pi_t(s_t) \in \mathcal{A}$ at a particular time $t$ for any possible state $s_t$ under $\pi$. The implementation of a policy generates a sequence of utilities that depends on the objective transition probability distribution $p_t(s_t, a_t)$ for the evolution from state $s_t$ to $s_{t + 1}$ induced by the model.\\

Figure (ref) depicts the timing of events for two generic periods. At the beginning of period $t$, an individual fully learns about each action's immediate utility, selects one of the alternatives, and receives its immediate utility. Then, the state evolves from $s_t$ to $s_{t + 1}$, and the process repeats itself in $t + 1$.\\

figure[figure omitted — 2,757 chars of source]

Individuals make their decisions facing uncertainty about the future and seek to maximize their expected total discounted utilities over all decision periods given all available information. They have rational expectations Muth.1961, so their subjective beliefs about the future agree with the objective probabilities for all possible future events provided by the model. Immediate utilities are separable between periods Kahneman.1997, and a discount factor $\delta$ parameterizes a preference for immediate over future utilities Samuelson.1937.\\

Equation ((ref)) formally describes the individual's objective. Given an initial state $s_1$, they implement a policy $\pi$ that maximizes the expected total discounted utilities over all decision periods given the information available at the time.

align[align omitted — 146 chars of source]

EKW models are set up as a standard Markov decision process (MDP) Puterman.1994,Rust.1994,White.1993 that can be solved by a simple backward induction procedure. In the final period $T$, there is no future to consider, and the optimal action is choosing the alternative with the highest immediate utility in each state. With the decision rule for the final period, we can determine all other optimal decisions recursively. We use our group's open-source research code \verb+respy+ Gabler.2020b, which allows for the flexible specification, simulation, and estimation of EKW models. Detailed documentation of the software and its numerical components is available at \url{http://respy.readthedocs.io}.

The career decisions of young men

Keane.1997 specialize the model above to explore the career decisions of young men regarding their schooling, work, and occupational choices using the National Longitudinal Survey of Youth 1979 (NLSY79) NLSY.2019 for the estimation of the model. We restrict ourselves to a basic summary of their setup. Further documentation of the model specification and the observed dataset is available in the Appendix.\\

Keane.1997 follows individuals over their working life from young adulthood at age 16 to retirement at age 65. Each decision period $t = 16, \dots, 65$ represents a school year. Figure (ref) illustrates the initial decision problem as individuals select one of five alternatives from the set of admissible actions $a\in\mathcal{A}$. They can decide to either work in a blue-collar or a white-collar occupation ($a = 1, 2$), serve in the military $(a = 3)$, attend school $(a = 4)$, or stay at home $(a = 5)$.\\

figure[figure omitted — 7,722 chars of source]

\FloatBarrier

Individuals are already heterogeneous when entering the model. They differ with respect to their level of initial schooling $h_{16}$, and have one of four different $\mathcal{J} = \{1, \hdots, 4\}$ alternative-specific skill endowment types $\bm{e} = \left(e_{j,a}\right)_{\mathcal{J} \times \mathcal{A}}$.\\

The immediate utility $u_a(\cdot)$ of each alternative consists of a non-pecuniary utility $\zeta_a(\cdot)$ and, at least for the working alternatives, an additional wage component $w_a(\cdot)$. Both depend on the level of human capital as measured by their alternative-specific skill endowment $\bm{e}$, their years of completed schooling $h_t$, and their occupation-specific work experience $\bm{k_t} = \left(k_{a,t}\right)_{a\in\{1, 2, 3\}}$. The immediate utilities are influenced by last-period choices $a_{t -1}$ and alternative-specific productivity shocks $\bm{\epsilon_t} = \left(\epsilon_{a,t}\right)_{a\in\mathcal{A}}$ as well. Their general form is given by:

align*[align* omitted — 348 chars of source]

Work experience $\bm{k_t}$ and years of completed schooling $h_t$ evolve deterministically. There is no uncertainty about grade completion Altonji.1993 and no part-time enrollment. Schooling is defined by time spent in school, not by formal credentials acquired. Once individuals reach a certain amount of schooling, they acquire a degree.

align*[align* omitted — 176 chars of source]

The productivity shocks $\bm{\epsilon_t}$ are uncorrelated across time and follow a multivariate normal distribution with mean $\bm{0}$ and covariance matrix $\bm{\Sigma}$. Given the structure of the utility functions and the distribution of the shocks, the state at time $t$ is $s_t = \{\bm{k_t}, h_t, t, a_{t -1}, \bm{e},\bm{\epsilon_t}\}$.\\

Skill endowments $\bm{e}$ and initial schooling $h_{16}$ are the only sources of persistent heterogeneity in the model. All remaining differences in life-cycle decisions result from different transitory shocks $\bm{\epsilon_t}$ that occur over time.\\

Theoretical and empirical research from specialized disciplines within economics informs the specification of each $u_a(\cdot)$. As an example, we provide the exact functional form of the non-pecuniary utility from schooling in Equation ((ref)). Further details on the specification of the utility functions are available in the Appendix.

align[align omitted — 556 chars of source]

There is a direct cost in the form of tuition for continuing education after high school $\beta_{tc_1}$ and college $\beta_{tc_2}$. The decision to leave school is reversible, but entails re-enrollment costs that differ by schooling category ($\beta_{rc_1}, \beta_{rc_2}$).\\

We analyze the original dataset used by Keane.1997. We only provide a brief description and relegate further details to the Appendix. The authors construct their sample based on the NLSY79, a nationally representative sample of young men and women living in the United States in 1979 and born between 1957 and 1964. Individuals were followed from 1979 onwards and repeatedly interviewed about their schooling decisions and labor market experiences. Based on this information, individuals are assigned to either working in one of the three occupations, attending school, or simply staying at home.\\

Keane.1997 restrict attention to white men, who turned 16 between 1977 and 1981, and exploit information collected between 1979 and 1987. Thus, individuals in the sample range in age between 16 and 26 years old. While the sample initially consists of 1,373 individuals at age 16, this number drops to 256 at the age of 26 due to sample attrition and missing data. Overall, the final sample consists of 12,359 person-period observations.\\

Figure (ref) summarizes the evolution of choices and wages over the sample period. Roughly 86% of individuals initially enroll in school, but this share steadily declines with age. Nevertheless, about 39% pursue some form of higher education and obtain more than a high school degree. As individuals leave school, most of them initially pursue a blue-collar occupation. However, the relative share of white-collar workers increases as individuals entering the labor market later gain access to higher levels of schooling. At age 26, about 48% work in a blue-collar occupation and 34% in a white-collar occupation. The share of individuals in the military peaks around age 20 at 8%. At its maximum around age 18, approximately 20% of individuals stay at home.\\

figure[figure omitted — 546 chars of source]

\FloatBarrier

For an individual, the average wage starts at about \$10,000 at age 16 and increases considerably up to about \$25,000 by the age of 26. While starting wages for blue-collar workers are about \$10,286, wages in white-collar occupations and the military start around \$9,000. However, wages for white-collar occupations increase sharply over time, overtaking blue-collar wages around age 21. By the end of the observation period, wages for white-collar occupations are about 50% higher than blue-collar wages at \$32,756 compared to only \$20,739. Military wages remain lowest throughout.\\

We consider observations for $i = 1, \hdots, N$ individuals in each time period $t = 1, \dots, T_i$. For every observation $(i, t)$ in the data, we observe the action $a_{it}$, some components $\bar{u}_{it}$ of the utility, and a subset $\bar{s}_{it}$ of the state $s_{it}$. Therefore, from an economist's point of view, we must distinguish between two types of state variables $s_{it} = \{\bar{s}_{it}, \bm{e},\bm{\epsilon_t}\}$. At time $t$, the economist and individual both observe $\bar{s}_{it}$, while $\{ \bm{e},\bm{\epsilon_t}\}$ is only observed by the individual.\\

We use simulated maximum likelihood Fisher.1922,Manski.1977 estimation and determine the $88$ model parameters $\hat{\bm{\theta}}$ that maximize the likelihood function $\mathcal{L}(\bm{\theta}\mid\mathcal{D})$. As we only observe a subset $\bar{s}_t = \{\bm{k_t}, h_t, t, a_{t -1}\}$ of the state, we can determine the probability $p_{it}(a_{it}, \bar{u}_{it} \mid \bar{s}_{it}, \bm{\theta})$ of individual $i$ at time $t$ in $\bar{s}_{it}$ choosing $a_{it}$ and receiving $\bar{u}_{it}$ given parametric assumptions about the distribution of $\bm{\epsilon_t}$. The objective function takes the following form:

align*[align* omitted — 248 chars of source]

Overall, our parameter estimates are in broad agreement with the results reported in the original paper and the related literature. For example, individuals discount future utilities by $6\%$ per year. The returns to schooling vary according to occupation. While wages for white-collar occupations increase by about $6\%$ with each additional year of schooling, they only increase by $2\%$ for those working blue collar jobs. Skills are transferable across occupations as work experience increases wages in both blue and white-collar occupations.\\

Figure (ref) shows the overall agreement between the empirical data and a dataset simulated using the estimated model parameters. We show average wages and the share of individuals choosing a blue-collar occupation over time. The results are based on a simulated sample of $10,000$ individuals. Additional model fit statistics are available in the Appendix.

figure[figure omitted — 238 chars of source]

\FloatBarrier

We adhere to the procedure outlined by the authors of the original paper and use the estimated model to conduct the ex-ante evaluation of a $\$2,000$ tuition subsidy on educational attainment. We simulate a sample of $10,000$ individuals using the point estimates and compare completed schooling to a sample of the same size, but with a reduction of $\hat{\beta}_{tc_1}$ by $\$2,000$. The subsidy increases average final schooling by 0.65 years. College graduation increases by 13 percentage points and high school graduation rates improve by 4 percentage points.

Confidence set bootstrap

The construction of confidence sets for counterfactuals in many structural models poses two distinct challenges. First, the computational burden of even a single estimation of the model is considerable. This makes the application of a standard bootstrap approach Efron.1979 infeasible. Second, the nonlinear mapping from the parameters of the model to the counterfactual predictions often has kinks or is truncated. For example, in our case, the predicted impact of a tuition subsidy is bounded from below by zero. This violates the smoothness requirements of the delta method.\\

We use the Confidence Set (CS) bootstrap to construct the confidence set of the counterfactual. Although the CS bootstrap was originally proposed in Rao.1973, it has only recently been formalized by Woutersen.2019. Its application does not require repeated estimations of the model, as it uses the asymptotic normal distribution of the estimator for $\hat{\bm{\theta}}$. Furthermore, its validity does not depend on the differentiability of the prediction function.\footnote{See Reich.2020 for a critical assessment of confidence sets based on asymptotic arguments. They advocate the use of likelihood-ratio confidence intervals instead and set up their computation as a constraint optimization problem.}\\

Algorithm (ref) provides a concise description of the steps involved, where $\chi_l^2(1 - \alpha)$ is the quantile function for probability $1 - \alpha$ of the chi-square distribution with $l$ degrees of freedom.\\

\floatname{algorithm}{ Algorithm}

algorithm[algorithm omitted — 717 chars of source]

\FloatBarrier

To summarize, we draw a large sample of $M$ parameters from the estimated asymptotic normal distribution of our estimator with mean $\hat{\bm{\theta}}$ and covariance matrix $\hat{\boldsymbol{\Sigma}}$, accepting only those draws that are elements of the confidence set of the model parameters. We then compute the counterfactual for all remaining draws and calculate the confidence set for the counterfactual based on its lowest and highest value.\\

The CS bootstrap poses a considerable computational challenge. In many applications, including our own, a single prediction of a counterfactual takes several minutes. At the same time, the number of parameter samples must be large to ensure that the minimum and maximum values for the counterfactual prediction are reliable. However, the algorithm is amenable to parallelization using modern high-performance computational resources by processing each of the $M$ parameter draws independently.\\

Our uncertainty sets then take the following form:

align*[align* omitted — 464 chars of source]

\FloatBarrier

Results

Turning to the presentation of our results, we focus on the impact of a $\$2,000$ tuition subsidy on completed schooling and use the 90% uncertainty set to measure the degree of uncertainty. All our results potentially depend on the size of the uncertainty set. In practice, policy-makers choose the uncertainty set's size in line with their underlying preferences - the more desirable protection against unfavorable outcomes is, the larger the uncertainty set will be.\footnote{In a different setting, Blesch.2021 conduct an ex-ante performance evaluation of the statistical decision functions over the whole parameter space Wald.1950,Manski.2021.}\\

All results are based on $30,000$ draws from the asymptotic normal distribution of our parameter estimates. We follow Keane.1997 and start by analyzing the prediction for a general subsidy. Then we turn to the situation where we use endowment types for policy targeting. Throughout our analysis, we postulate a linear utility function for the policy-maker.

General subsidy

Figure (ref) explores the impact prediction for a general tuition subsidy. We show the point prediction, its sampling distribution, and the uncertainty set. At the point estimate, average schooling increases by $0.65$ years. However, there is considerable uncertainty about the prediction, as the uncertainty set ranges from $0.15$ to $1.10$ years.\\

figure[figure omitted — 140 chars of source]

\FloatBarrier In Figure (ref), we trace the effect of the discount rate $\delta$ on the subsidy's impact over the uncertainty set, while keeping all other parameters at their point estimate. Initially, as $\delta$ increases, so does the policy's impact as individuals value the long-term benefits from increasing their level of schooling more and more. However, for high levels of the discount factor, the policy's impact starts to decrease as most individuals already complete a high school or college degree even without the subsidy.

figure[figure omitted — 131 chars of source]

\FloatBarrier

Targeted subsidy

So far, we restricted the analysis to a general subsidy available to the whole population and the average predicted impact. We now examine the setting in which a policy-maker can target individuals based on the type of their initial endowment. The importance of early endowment heterogeneity in shaping economic outcomes over the life-cycle is the most important finding from Keane.1997. It served as motivation for a host of subsequent research on the determinants of skill heterogeneity among adolescents Caucutt.2020,Erosa.2010,Todd.2007.\\

To ease the exposition, we initially focus our discussion of results on Type 1 and Type 3 individuals. We later rank policies targeting either of the four types based on the different decision-theoretic criteria. Additional results are available in our Appendix.\\

Figure (ref) confirms that life-cycle choices differ considerably by initial endowment type. On the left, we show the number of periods the two types spend on average in each of the five alternatives. Those characterized as Type 1 individuals spend more than six years on their education even after entering the model. Type 3 individuals, on the other hand, extend their academic pursuits for only an additional two years. This difference translates into very different labor market experiences. While Type 1 individuals work for about 35 years in a white-collar occupation, Type 3 workers switch more frequently between white and blue-collar occupations and spend a comparable amount of time working in either occupation -- approximately 44 years split equally among white and blue-collar occupations. Both types only spend a short time at home.

figure[figure omitted — 322 chars of source]

\FloatBarrier On the right, we show the distribution of final schooling for both types. Years of schooling are considerably higher for Type 1 individuals with an average of more than 16 years compared to only 12 years for those identified as Type 3 individuals. Nearly all Type 1 individuals enroll in college and most graduate with a degree.\\

Figure (ref) provides a visualization of our core results for a targeted subsidy. At the point estimates, the predicted impact is considerably lower for Type 1 than Type 3. However, the prediction uncertainty is much larger for Type 3 compared to Type 1. The uncertainty set for Type 3 ranges all the way from $0$ to $1.2$ years, while the prediction for Type 1 is between $0.18$ and $0.75$.\\

figure[figure omitted — 286 chars of source]

\FloatBarrier This heterogeneity in impact and prediction uncertainty follows directly from the underlying economics of the model. Type 1 individuals are already more likely to have a college degree before the subsidy, and thus, the predicted impact is smaller. Alternatively, Type 1 individuals affected by the subsidy are in the middle of pursuing a college education and thus directly benefit from it. Since Type 3 individuals are at the lower end of the schooling distribution, a tuition subsidy can considerably increase their level of schooling. Whether the subsidy succeeds in doing so, however, remains uncertain.\\

We now consider the policy option to target Type 2 and Type 4 as well. Their point predictions are actually highest with an additional $0.81$ years on average for Type 2 and $0.75$ years for Type 4. However, both predictions are fraught with uncertainty. For Type 2 the uncertainty set ranges from $0.17$ to $1.3$, while for Type 4 it starts at zero and spans all the way to $1.18$.\\

Figure (ref) shows the policy alternative's ranking by the decision-theoretic criteria we discussed in Section (ref). Ranking alternatives using as-if optimization is straightforward. A policy targeting Type 2 is the most preferred alternative, while a focus on Type 1 is the least attractive. However, once we account for the presence of uncertainty in the predictions, a more nuanced picture emerges. Moving from as-if optimization to a subjective Bayes criterion using a uniform distribution over the uncertainty set does not change the ordering. However, once a decision-maker is concerned with performance across the whole range of values in the uncertainty set -- we move to the minimax regret or maximin criterion -- a policy targeting Type 1 becomes more and more attractive despite its low point prediction because its worst-case utility is highest.\\

figure[figure omitted — 132 chars of source]

\FloatBarrier In general, framing policy advice as a decision problem under uncertainty shows that there are many different ways of making reasonable decisions. The ranking of policies varies depending on the decision criteria. Not only that, but due to the necessary ex-post nature of our implementation, the ranking for a given criteria also depends on the choice of $\alpha$. The selection of $\alpha$ is part of the decision problem: the more a policy-maker is concerned about worst-case scenarios, the smaller the appropriate value for $\alpha$ will be. After deciding on a preferred decision rule, we suggest performing a sensitivity analysis around the selected $\alpha$ value by checking how much the policy ranking varies within a neighborhood.

\FloatBarrier

Conclusion

We develop a generic approach that addresses parametric uncertainty when using models to inform policy-making. We propose a decision-theoretic analysis of computationally demanding structural models based on uncertainty sets. We construct the uncertainty sets from empirical estimates and ensure their computational tractability by using the confidence set bootstrap. We revisit the seminal work by Keane.1997 to document the empirical relevance of prediction uncertainty and showcase our analysis. Focusing on their ex-ante evaluation of a tuition subsidy, we report considerable uncertainty in the policy's impact on completed schooling. We show how a policy-maker's preferred policy depends on the choice of alternative formal rules for decision-making under uncertainty.\\

In our ongoing research, we pursue three avenues for further improvements. First, we link our work with the literature on inference under (local) model misspecification to refine the construction of our uncertainty sets. For example, Armstrong.2021 and Bonhomme.2020 propose different methods for taking misspecification into account when constructing confidence sets. Second, we incorporate ideas from the literature on global sensitivity analysis Razavi.2021 to identify the parameters most responsible for uncertainty in predictions. The attribution of importance based on Shapely values, familiar to economists from game theory, appears promising Owen.2014, Shapley.1953 as well. Third, we address our analysis's computational burden using surrogate modeling Forrester.2008, which emulates the full model's behavior at a negligible cost per run and allows us to determine prediction uncertainty using a nonparametric bootstrap procedure.

thebibliography\bibitem[Adda et al., 2017]{Adda.2017} Adda, J., Dustmann, C., and Stevens, K. (2017). \newblock The career costs of children. \newblock {\em Journal of Political Economy}, 125(2):293--337. \bibitem[Aguirregabiria and Mira, 2010]{Aguirregabiria.2010} Aguirregabiria, V. and Mira, P. (2010). \newblock Dynamic discrete choice structural models: A survey. \newblock {\em Journal of Econometrics}, 156(1):38--67. \bibitem[Altonji, 1993]{Altonji.1993} Altonji, J. (1993). \newblock The demand for and return to education when education outcomes are uncertain. \newblock {\em Journal of Labor Economics}, 11(1):48--83. \bibitem[Andrews et al., 2017]{Andrews.2017} Andrews, I., Gentzkow, M., and Shapiro, J. M. (2017). \newblock Measuring the sensitivity of parameter estimates to estimation moments. \newblock {\em The Quarterly Journal of Economics}, 132(4):1553--1592. \bibitem[Andrews et al., 2020]{Andrews.2020} Andrews, I., Gentzkow, M., and Shapiro, J. M. (2020). \newblock On the informativeness of descriptive statistics for structural estimates. \newblock {\em Econometrica}, 88(6):2231--2258. \bibitem[Armstrong and Koles{\'a}r, 2021]{Armstrong.2021} Armstrong, T. B. and Koles{\'a}r, M. (2021). \newblock Sensitivity analysis using approximate moment condition models. \newblock {\em Quantitative Economics}, 12(1):77--108. \bibitem[Ben-Tal et al., 2013]{Ben-Tal.2013} Ben-Tal, A., den Hertog, D., De Waegenaere, A., Melenberg, B., and Rennen, G. (2013). \newblock Robust solutions of optimization problems affected by uncertain probabilities. \newblock {\em Management Science}, 59(2):341--357. \bibitem[Berger et al., 2021]{Berger.2021} Berger, L., Berger, N., Bosetti, V., Gilboa, I., Hansen, L. P., Jarvis, C., Marinacci, M., and Smith, R. D. (2021). \newblock Rational policymaking during a pandemic. \newblock {\em Proceedings of the National Academy of Sciences}, 118(4). \bibitem[Bertsimas et al., 2018]{Bertsimas.2018} Bertsimas, D., Gupta, V., and Kallus, N. (2018). \newblock Data-driven robust optimization. \newblock {\em Mathematical Programming}, 167(2):235--292. \bibitem[Blesch and Eisenhauer, 2021]{Blesch.2021} Blesch, M. and Eisenhauer, P. (2021). \newblock Robust decision-making under risk and ambiguity. \newblock {\em arXiv Working Paper}. \bibitem[Blundell, 2017]{Blundell.2017} Blundell, R. (2017). \newblock What have we learned from structural models? \newblock {\em American Economic Review}, 107(5):287--92. \bibitem[Blundell et al., 2016]{Blundell.2016} Blundell, R., Costa Dias, M., Meghir, C., and Shaw, J. (2016). \newblock Female labor supply, human capital, and welfare reform. \newblock {\em Econometrica}, 84(5):1705--1753. \bibitem[Blundell and Shephard, 2012]{Blundell.2012} Blundell, R. and Shephard, A. (2012). \newblock {Employment, hours of work and the optimal taxation of low-income families}. \newblock {\em The Review of Economic Studies}, 79(2):481--510. \bibitem[Bonhomme and Weidner, 2020]{Bonhomme.2020} Bonhomme, S. and Weidner, M. (2020). \newblock Minimizing sensitivity to model misspecification. \newblock {\em arXiv Working Paper}. \bibitem[Bugni and Ura, 2019]{Bugni.2019} Bugni, F. and Ura, T. (2019). \newblock Inference in dynamic discrete choice problems under local misspecification. \newblock {\em Quantitative Economics}, 10(1):67--103. \bibitem[{Bureau of Labor Statistics}, 2019]{NLSY.2019} {Bureau of Labor Statistics} (2019). \newblock {\em National Longitudinal Survey of Youth 1979 cohort, 1979-2016 (rounds 1-27)}. \newblock Center for Human Resource Research, Ohio State University, Columbus, OH. \bibitem[Cai and Lontzek, 2019]{Cai.2019} Cai, Y. and Lontzek, T. S. (2019). \newblock The social cost of carbon with economic and climate risks. \newblock {\em Journal of Political Economy}, 127(6):2684--2734. \bibitem[Caucutt and Lochner, 2020]{Caucutt.2020} Caucutt, E. M. and Lochner, L. (2020). \newblock Early and late human capital investments, borrowing constraints, and the family. \newblock {\em Journal of Political Economy}, 128(3):1065--1147. \bibitem[Christensen and Connault, 2019]{Christensen.2019} Christensen, T. and Connault, B. (2019). \newblock Counterfactual sensitivity and robustness. \newblock {\em arXiv Working Paper}. \bibitem[Cunha et al., 2010]{Cunha.2010} Cunha, F., Heckman, J. J., and Schennach, S. (2010). \newblock Estimating the technology of cognitive and noncognitive skill formation. \newblock {\em Econometrica}, 78(3):883--931. \bibitem[Eckstein et al., 2019]{Eckstein.2019} Eckstein, Z., Keane, M., and Lifshitz, O. (2019). \newblock Career and family decisions: Cohorts born 1935-1975. \newblock {\em Econometrica}, 87(1):217--253. \bibitem[Efron, 1979]{Efron.1979} Efron, B. (1979). \newblock Bootstrap methods: Another look at the jackknife. \newblock {\em The Annals of Statistics}, 7(1):1--26. \bibitem[Eisenhauer et al., 2015]{Eisenhauer.2015b} Eisenhauer, P., Heckman, J. J., and Mosso, S. (2015). \newblock Estimation of dynamic discrete choice models by maximum likelihood and the simulated method of moments. \newblock {\em International Economic Review}, 56(2):331--357. \bibitem[Erosa et al., 2010]{Erosa.2010} Erosa, A., Koreshkova, T., and Restuccia, D. (2010). \newblock How important is human capital? {A} quantitative theory assessment of world income inequality. \newblock {\em The Review of Economic Studies}, 77(4):1421--1449. \bibitem[Fisher, 1922]{Fisher.1922} Fisher, R. A. (1922). \newblock On the mathematical foundations of theoretical statistics. \newblock {\em Philosophical Transactions of the Royal Society of London. Series A, Containing Papers of a Mathematical or Physical Character}, 222(594-604):309--368. \bibitem[Forrester et al., 2008]{Forrester.2008} Forrester, D. A. I. J., Sobester, D. A., and Keane, P. A. J. (2008). \newblock {\em Engineering design via surrogate modelling: A practical guide}. \newblock John Wiley & Sons, Chichester, England. \bibitem[Gabler and Raabe, 2020]{Gabler.2020b} Gabler, J. and Raabe, T. (2020). \newblock respy - {A} framework for the simulation and estimation of {E}ckstein-{K}eane-{W}olpin models. \bibitem[Gayle and Shephard, 2019]{Gayle.2019} Gayle, G.-L. and Shephard, A. (2019). \newblock Optimal taxation, marriage, home production, and family labor supply. \newblock {\em Econometrica}, 87(1):291--326. \bibitem[Gilboa, 2009]{Gilboa.2009} Gilboa, I. (2009). \newblock {\em Theory of decision under uncertainty}. \newblock Cambridge University Press, New York City, NY. \bibitem[Gilboa, 2020]{Gilboa.2020} Gilboa, I. (2020). \newblock What were you thinking? {R}evealed preference theory as coherence test. \newblock {\em Working Paper}. \bibitem[Gilboa et al., 2018]{Gilboa.2018} Gilboa, I., Rouziou, M., and Sibony, O. (2018). \newblock Decision theory made relevant: Between the software and the shrink. \newblock {\em Research in Economics}, 72(2):240--250. \bibitem[Gilboa and Schmeidler, 1989]{Gilboa.1989} Gilboa, I. and Schmeidler, D. (1989). \newblock Maxmin expected utility with non-unique prior. \newblock {\em Journal of Mathematical Economics}, 18(2):141--153. \bibitem[Harenberg et al., 2019]{Harenberg.2019} Harenberg, D., Marelli, S., Sudret, B., and Winschel, V. (2019). \newblock Uncertainty quantification and global sensitivity analysis for economic models. \newblock {\em Quantitative Economics}, 10(1):1--41. \bibitem[Hood and Koopmans, 1953]{Hood.1953} Hood, W. and Koopmans, T. (1953). \newblock {\em Studies in econometric method}. \newblock John Wiley & Sons, New York City, NY. \bibitem[J{\o}rgensen, 2021]{Jorgensen.2021} J{\o}rgensen, T. H. (2021). \newblock Sensitivity to calibrated paramters. \newblock {\em Review of Economics and Statistics}, forthcoming. \bibitem[Kahneman et al., 1997]{Kahneman.1997} Kahneman, D., Wakker, P. P., and Sarin, R. (1997). \newblock Back to {B}entham? {E}xplorations of experienced utility. \newblock {\em The Quarterly Journal of Economics}, 112(2):375--406. \bibitem[Kalouptsidi et al., 2020]{Kalouptsidi.2020} Kalouptsidi, M., Kitamura, Y., Lima, L., and {Souza-Rodrigues}, E. A. (2020). \newblock Partial identification and inference for dynamic models and counterfactuals. \newblock {\em NBER Working Paper Series}, No. 26761. \bibitem[Kalouptsidi et al., 2021]{Kalouptsidi.2021} Kalouptsidi, M., Scott, P. T., and {Souza-Rodrigues}, E. (2021). \newblock Identification of counterfactuals in dynamic discrete choice models. \newblock {\em Quantitative Economics}, 12(2):351--403. \bibitem[Keane et al., 2011]{Keane.2011d} Keane, M. P., Todd, P. E., and Wolpin, K. I. (2011). \newblock The structural estimation of behavioral models: Discrete choice dynamic programming methods and applications. \newblock In Ashenfelter, O. and Card, D., editors, {\em Handbook of Labor Economics}, pages 331--461. Elsevier Science, Amsterdam, Netherlands. \bibitem[Keane and Wolpin, 1997]{Keane.1997} Keane, M. P. and Wolpin, K. I. (1997). \newblock The career decisions of young men. \newblock {\em Journal of Political Economy}, 105(3):473--522. \bibitem[Low and Meghir, 2017]{Low.2017} Low, H. and Meghir, C. (2017). \newblock The use of structural models in econometrics. \newblock {\em Journal of Economic Perspectives}, 31(2):33--58. \bibitem[Machina and Viscusi, 2014]{Machina.2014} Machina, M. J. and Viscusi, K. (2014). \newblock {\em Handbook of the economics of risk and uncertainty}. \newblock North-Holland Publishing Company, Amsterdam, Netherlands. \bibitem[Manski, 2004]{Manski.2004} Manski, C. F. (2004). \newblock Statistical treatment rules for heterogeneous populations. \newblock {\em Econometrica}, 72(4):1221--1246. \bibitem[Manski, 2013]{Manski.2013} Manski, C. F. (2013). \newblock {\em Public policy in an uncertain world: Analysis and decisions}. \newblock Harvard University Press, Cambridge, MA. \bibitem[Manski, 2019]{Manski.2019} Manski, C. F. (2019). \newblock Communicating uncertainty in policy analysis. \newblock {\em Proceedings of the National Academy of Sciences}, 116(16):7634--7641. \bibitem[Manski, 2021]{Manski.2021} Manski, C. F. (2021). \newblock Econometrics for decision making: Building foundations sketched by {H}aavelmo and {W}ald. \newblock {\em Econometrica}, forthcoming. \bibitem[Manski and Lerman, 1977]{Manski.1977} Manski, C. F. and Lerman, S. R. (1977). \newblock The estimation of choice probabilities from choice based samples. \newblock {\em Econometrica}, 45(8):1977--1988. \bibitem[Marinacci, 2015]{Marinacci.2015} Marinacci, M. (2015). \newblock Model uncertainty. \newblock {\em Journal of the European Economic Association}, 13(6):1022--1100. \bibitem[Mukhin, 2018]{Mukhin.2018} Mukhin, Y. (2018). \newblock Sensitivity of regular estimators. \newblock {\em arXiv Working Paper}. \bibitem[Muth, 1961]{Muth.1961} Muth, J. F. (1961). \newblock Rational expectations and the theory of price movements. \newblock {\em Econometrica}, 29(3):315--335. \bibitem[{National Research Council}, 2012]{Council.2012} {National Research Council} (2012). \newblock {\em Assessing the reliability of complex models: Mathematical and statistical foundations of verification, validation, and uncertainty quantification}. \newblock The National Academies Press, Washington, DC. \bibitem[Niehans, 1948]{Niehans.1948} Niehans (1948). \newblock Zur {P}reisbildung bei ungewissen {E}rwartungen. \newblock {\em Swiss Journal of Economics and Statistics}, 84(5):433--456. \bibitem[Norets and Tang, 2014]{Norets.2014} Norets, A. and Tang, X. (2014). \newblock Semiparametric inference in dynamic binary choice models. \newblock {\em The Review of Economic Studies}, 81(3):1229--1262. \bibitem[Owen, 2014]{Owen.2014} Owen, A. B. (2014). \newblock {S}obol' indices and {S}hapley value. \newblock {\em Journal on Uncertainty Quantification}, 2(1):245--251. \bibitem[Puterman, 1994]{Puterman.1994} Puterman, M. L. (1994). \newblock {\em {M}arkov decision processes: Discrete stochastic dynamic programming}. \newblock John Wiley & Sons, New York City, NY. \bibitem[Rao, 1973]{Rao.1973} Rao, C. R. (1973). \newblock {\em Linear statistical inference and its applications}. \newblock John Wiley & Sons, New York City, NY. \bibitem[Razavi et al., 2021]{Razavi.2021} Razavi, S., Jakeman, A., Saltelli, A., Prieur, C., Iooss, B., Borgonovo, E., Plischke, E., Piano, S. L., Iwanaga, T., Becker, W., et al. (2021). \newblock The future of sensitivity analysis: An essential discipline for systems modeling and policy support. \newblock {\em Environmental Modeling & Software}, 137:104954. \bibitem[Reich and Judd, 2020]{Reich.2020} Reich, G. and Judd, K. L. (2020). \newblock Efficient likelihood ratio confidence intervals using constrained optimization. \newblock {\em SSRN Working Paper}. \bibitem[Rust, 1987]{Rust.1987} Rust, J. (1987). \newblock Optimal replacement of {GMC} bus engines: An empirical model of {Harold Zurcher}. \newblock {\em Econometrica}, 55(5):999--1033. \bibitem[Rust, 1994]{Rust.1994} Rust, J. (1994). \newblock Structural estimation of {M}arkov decision processes. \newblock In Engle, R. and McFadden, D., editors, {\em Handbook of {E}conometrics}, pages 3081--3143. North-Holland Publishing Company, Amsterdam, Netherlands. \bibitem[Saltelli et al., 2020]{Saltelli.2020} Saltelli, A., Bammer, G., Bruno, I., Charters, E., Di Fiore, M., Didier, E., Espeland, W. N., Kay, J., Piano, S. L., Mayo, D., et al. (2020). \newblock Five ways to ensure that models serve society: {A} manifesto. \newblock {\em Nature}, 582:482--484. \bibitem[Samuelson, 1937]{Samuelson.1937} Samuelson, P. A. (1937). \newblock A note on measurement of utility. \newblock {\em Review of Economic Studies}, 4(2):155--161. \bibitem[SAPEA, 2019]{SAPEA.2019} SAPEA (2019). \newblock Making sense of science for policy under conditions of complexity and uncertainty. \bibitem[Savage, 1954]{Savage.1954} Savage, L. J. (1954). \newblock {\em The foundations of statistics}. \newblock John Wiley & Sons, New York City, NY. \bibitem[Shapley, 1953]{Shapley.1953} Shapley, L. S. (1953). \newblock {\em A value for n-person games}. \newblock Princeton University Press, Princeton, NJ. \bibitem[Todd and Wolpin, 2006]{Todd.2006} Todd, P. E. and Wolpin, K. I. (2006). \newblock Assessing the impact of a school subsidy program in {M}exico: Using a social experiment to validate a dynamic behavioral model of child schooling and fertility. \newblock {\em American Economic Review}, 96(5):1384--1417. \bibitem[Todd and Wolpin, 2007]{Todd.2007} Todd, P. E. and Wolpin, K. I. (2007). \newblock The production of cognitive achievement in children: Home, school and racial test score gaps. \newblock {\em Journal of Human Capital}, 1(1):91--136. \bibitem[Wald, 1950]{Wald.1950} Wald, A. (1950). \newblock {\em Statistical decision functions}. \newblock John Wiley & Sons, New York City, NY. \bibitem[White, 1993]{White.1993} White, D. J. (1993). \newblock {\em {M}arkov decision processes}. \newblock John Wiley & Sons, New York City, NY. \bibitem[Wolpin, 2013]{Wolpin.2013} Wolpin, K. I. (2013). \newblock {\em The limits to inference without theory}. \newblock MIT University Press, Cambridge, MA. \bibitem[Woutersen and Ham, 2019]{Woutersen.2019} Woutersen, T. and Ham, J. (2019). \newblock Confidence sets for continuous and discontinuous functions of parameters. \newblock {\em SSRN Working Paper}.