EconBase
← All papers

Reinforcement Learning in High-frequency Market Making

Yuheng Zheng, Zihan Ding

arXiv 14 Jul 2024 · Finance — Trading · 3 citations (OpenAlex)

arXiv:2407.21025 · PDF · DOI · OpenAlex · Extracted main text

Abstract

This paper establishes a new and comprehensive theoretical analysis for the application of reinforcement learning (RL) in high-frequency market making. We bridge the modern RL theory and the continuous-time statistical models in high-frequency financial economics. Different with most existing literature on methodological research about developing various RL methods for market making problem, our work is a pilot to provide the theoretical analysis. We target the effects of sampling frequency, and find an interesting tradeoff between error and complexity of RL algorithm when tweaking the values of the time increment $\Delta$ $-$ as $\Delta$ becomes smaller, the error will be smaller but the complexity will be larger. We also study the two-player case under the general-sum game framework and establish the convergence of Nash equilibrium to the continuous-time game equilibrium as $\Delta\rightarrow0$. The Nash Q-learning algorithm, which is an online multi-agent RL method, is applied to solve the equilibrium. Our theories are not only useful for practitioners to choose the sampling frequency, but also very general and applicable to other high-frequency financial decision making problems, e.g., optimal executions, as long as the time-discretization of a continuous-time markov decision process is adopted. Monte Carlo simulation evidence support all of our theories.

Citation extraction

41
references
80
in-text mentions
41
distinct cited
0
self-citations
12,511
main-text words

appendix boundary found by appendix_command · 65% of the source is main text. Read the extracted text to check this.

Most heavily cited references

The works this paper leans on most, across its whole bibliography — not restricted to papers in our corpus. Ranked by composite intensity, which combines how often a work is mentioned, how many sections mention it, and how much of that falls in the main text rather than the appendix.

ReferenceIntensityMentionsSectionsMain text
1Avellaneda, M. and Stoikov, S (2008) High-frequency trading in a limit order book0.87462100%
2Cont, R. and Xiong, W (2022) Dynamics of market making algorithms in dealer markets: Learning and tacit collusion0.87452100%
3Even-Dar, E., Mansour, Y., and Bartlett, P (2003) Learning rates for q-learning0.8115280%
4Filar, J. and Vrieze, K (2012) Competitive Markov decision processes0.7817271%
5Guéant, O., Lehalle, C.-A., and Fernandez-Tapia, J (2013) Dealing with the inventory risk: a solution to the market making problem0.73732100%
6Hu, J. and Wellman, M. P (2003) Nash q-learning for general-sum stochastic games0.73732100%
7Luo, J. and Zheng, H (2021) Dynamic equilibrium of market making with price competition0.73732100%
8Guo, X. and Hernández-Lerma, O (2005) Nonzero-sum games for continuous-time markov chains with unbounded discounted payoffs0.6445240%
9Gihman, I. I. and Skorohod, A. V (2012) Controlled stochastic processes0.6443267%
10Sutton, R. S. and Barto, A. G (2018) Reinforcement learning: An introduction0.64422100%

Showing the top 10 of 41 scored citations.