EconBase
← All papers

Dynamic Decision-Making under Model Misspecification

Xinyu Dai

arXiv 20 May 2025 · Econometrics

arXiv:2505.14913 · PDF · DOI · OpenAlex · Extracted main text

Abstract

In this study, I investigate the dynamic decision problem with a finite parameter space when the functional form of conditional expected rewards is misspecified. Traditional algorithms, such as Thompson Sampling, guarantee neither an $O(e^{-T})$ rate of posterior parameter concentration nor an $O(T^{-1})$ rate of average regret. However, under mild conditions, we can still achieve an exponential convergence rate of the parameter to a pseudo truth set, an extension of the pseudo truth parameter concept introduced by White (1982). I further characterize the necessary conditions for the convergence of the expected posterior within this pseudo-truth set. Simulations demonstrate that while the maximum a posteriori (MAP) estimate of the parameters fails to converge under misspecification, the algorithm's average regret remains relatively robust compared to the correctly specified case. These findings suggest opportunities to design simple yet robust algorithms that achieve desirable outcomes even in the presence of model misspecifications.

Citation extraction

23
references
28
in-text mentions
23
distinct cited
0
self-citations
6,566
main-text words

appendix boundary found by appendix_titled_section at “Appendix” · 82% of the source is main text. Read the extracted text to check this.

Most heavily cited references

The works this paper leans on most, across its whole bibliography — not restricted to papers in our corpus. Ranked by composite intensity, which combines how often a work is mentioned, how many sections mention it, and how much of that falls in the main text rather than the appendix.

ReferenceIntensityMentionsSectionsMain text
1White, Halbert (1982) Maximum likelihood estimation of misspecified models0.84333100%
2Kim, Michael Jong (2017) Thompson sampling for stochastic control: The finite parameter case0.5113233%
3Fan, Lin, Glynn, Peter W (2021) Diffusion Approximations for Thompson Sampling0.51121100%
4Esponda, Ignacio, Pouzo, Demian, Yamamoto, Yuichi (2021) Asymptotic behavior of Bayesian learners with misspecified models0.40511100%
5Foster, Dylan J, Gentile, Claudio, Mohri, Mehryar, Zimmert, Julian,… (2020) Adapting to Misspecification in Contextual Bandits0.40511100%
6Bogunovic, Ilija, Krause, Andreas, Ranzato, M., Beygelzimer, A., Dau… (2021) Misspecified Gaussian Process Bandit Optimization0.40511100%
7Adusumilli, Karun (2021) Risk and optimal policies in bandit experiments0.40511100%
8Andrews, Isaiah, Barahona, Nano, Gentzkow, Matthew, Rambachan, Ashes… (2023) Structural estimation under misspecification: theory and implications for practice0.40511100%
9Armstrong, Timothy, Kline, Patrick M, Sun, Liyang (2024) Adapting to Misspecification0.40511100%
10Ba, Cuimin (2023) Robust Misspecified Models and Paradigm Shifts0.40511100%

Showing the top 10 of 23 scored citations.