← All authors Zhuoran Yang Yale University (per OpenAlex) · ORCID · OpenAlex
39 papers in scope · 38 published · 1 on the econ.EM arXiv · 721 citations · h-index 14 (over the papers listed here)
Related authors The 20 authors closest to this one in our weighted citation graph, most related first.
Rui Miao Cong Shi Lin Lin Zhengling Qi Zeqi Wu Nishanth Dikkala Lester Mackey Demián Pouzo Greg Lewis Chunrong Ai Xiaohong Chen Zheng Zhang Siyu Chen Vasilis Syrgkanis Andrew Bennett Nathan Kallus Masatoshi Uehara Whitney K. Newey Victor Chernozhukov Wei-Chen Wang Proximity is measured over citations between two papers we both hold, weighted by how heavily one leans on the other, and is symmetric — it does not distinguish citing from being cited. Authors without a profile here are skipped, and a genuinely close colleague can be missing simply because their work is not in our arXiv corpus. Method: docs/06-citations-pipeline.md .
Papers Show only papers in our arXiv econ.EM corpus (1 of 39)
Risk-Sensitive Deep RL: Variance-Constrained Actor-Critic Provably Finds Globally Optimal Policy
published 2025 · Journal of the American Statistical Association · 5 citations · first circulated 2020
Sample-Efficient Reinforcement Learning From Human Feedback via Information-Directed Sampling
published 2025 · IEEE Transactions on Information Theory
with Han Qi, Haochen Yang, Qiaosheng Zhang, Qi Han
Offline Reinforcement Learning for Human-Guided Human-Machine Interaction with Private Information
published 2025 · Management Science · 2 citations · first circulated 2022
working paper 2025 · arXiv
Efficient and assured reinforcement learning-based building HVAC control with heterogeneous expert-guided training
published 2025 · Scientific Reports · 17 citations
with Shichao Xu, Yangyang Fu, Yixuan Wang, Chao Huang, Zheng O’Neill, Zhaoran Wang, Qi Zhu
Contextual Dynamic Pricing with Strategic Buyers
published 2024 · Journal of the American Statistical Association · 5 citations · first circulated 2023
Provably Efficient Reinforcement Learning with Linear Function Approximation
published 2023 · Mathematics of Operations Research · 219 citations · first circulated 2019
A Two-Timescale Stochastic Algorithm Framework for Bilevel Optimization: Complexity Analysis and Application to Actor-Critic
published 2023 · 54 citations · first circulated 2020
Sequential Information Design: Markov Persuasion Process and Its Efficient Reinforcement Learning
published 2022 · Proceedings of the 23rd ACM Conference on Economics and Computation · 6 citations
Online Bootstrap Inference For Policy Evaluation In Reinforcement Learning
published 2022 · Journal of the American Statistical Association · 21 citations · first circulated 2021
Learning Zero-Sum Simultaneous-Move Markov Games Using Function Approximation and Correlated Equilibrium
published 2022 · Mathematics of Operations Research · 26 citations · first circulated 2020
with Qiaomin Xie, Yudong Chen, Zhaoran Wang
Understanding Implicit Regularization in Over-Parameterized Single Index Model
published 2022 · Journal of the American Statistical Association · 14 citations · first circulated 2020
Pessimism Meets Invariance: Provably Efficient Offline Mean-Field Multi-Agent RL
published 2021 · Neural Information Processing Systems · 25 citations · first circulated 2020
with Minshuo Chen, Yan Li, Ethan Wang, Zhaoran Wang, Tuo Zhao, Ying Jin
Provably Efficient Causal Reinforcement Learning with Confounded Observational Data
published 2021 · Neural Information Processing Systems · 13 citations · first circulated 2020
with Lingxiao Wang, Zhaoran Wang
Offline Constrained Multi-Objective Reinforcement Learning via Pessimistic Dual Value Iteration
published 2021 · Neural Information Processing Systems
with Runzhe Wu, Yufeng Zhang, Zhaoran Wang
no link
BooVI: Provably Efficient Bootstrapped Value Iteration
published 2021 · Neural Information Processing Systems
with Boyi Liu, Qi Cai, Zhaoran Wang
no link
A robust and efficient variable selection method for linear regression
published 2021 · Journal of Applied Statistics · 5 citations
Learning While Playing in Mean-Field Games: Convergence and Optimality
published 2021 · International Conference on Machine Learning · 9 citations
with Qiaomin Xie, Zhaoran Wang, Andreea Minca
no link
Doubly Robust Off-Policy Actor-Critic: Convergence and Optimality
published 2021 · International Conference on Machine Learning · 8 citations
with Tengyu Xu, Zhaoran Wang, Yingbin Liang
Risk-Sensitive Reinforcement Learning with Function Approximation: A Debiasing Approach
published 2021 · International Conference on Machine Learning · 6 citations
with Yingjie Fei, Zhaoran Wang
no link
Reinforcement Learning for Cost-Aware Markov Decision Processes
published 2021 · International Conference on Machine Learning · 2 citations
with Wesley A. Suttle, Kaiqing Zhang, Ji Liu, David N. Kraemer
no link
An efficient Gehan-type estimation for the accelerated failure time model with clustered and censored data
published 2021 · Lifetime Data Analysis · 4 citations · first circulated 2020
Provably Efficient Actor-Critic for Risk-Sensitive and Robust Adversarial RL: A Linear-Quadratic Case
published 2021 · International Conference on Artificial Intelligence and Statistics · 4 citations
with Yufeng Y. Zhang, Zhaoran Wang
no link
A Near-Optimal Algorithm for Stochastic Bilevel Optimization via Double-Momentum
published 2021 · Neural Information Processing Systems · 26 citations
Efficient and doubly-robust methods for variable selection and parameter estimation in longitudinal data analysis
published 2020 · Computational Statistics · 3 citations
Provably Efficient Neural Estimation of Structural Equation Models: An Adversarial Approach
published 2020 · Neural Information Processing Systems · 12 citations
Dynamic regret of policy optimization in non-stationary environments
published 2020 · neural information processing systems · 11 citations
with Yingjie Fei, Zhaoran Wang, Qiaomin Xie
Provably Efficient Reinforcement Learning with Kernel and Neural Function Approximations
published 2020 · Neural Information Processing Systems · 18 citations
no link
High-dimensional Varying Index Coefficient Models via Stein's Identity
published 2019 · Journal of Machine Learning Research · 9 citations · first circulated 2018
with Sen Na, Zhaoran Wang, Mladen Kolar
Provably Global Convergence of Actor-Critic: A Case for Linear Quadratic Regulator with Ergodic Cost
published 2019 · neural information processing systems · 58 citations
Generalized estimating equations for analyzing multivariate survival data
published 2019 · Communications in Statistics - Simulation and Computation · 1 citations
with Liya Fu, Jun Zhang, Anle Long, Yan Zhou
On the statistical rate of nonlinear recovery in generative models with heavy-tailed data
published 2019 · International Conference on Machine Learning · 17 citations
with Xiaohan Wei, Zhaoran Wang
no link
Policy Optimization Provably Converges to Nash Equilibria in Zero-Sum Linear Quadratic Games
published 2019 · Neural Information Processing Systems · 49 citations
with Kaiqing Zhang, Tamer Başar
Statistical-Computational Tradeoff in Single Index Models
published 2019 · Neural Information Processing Systems · 1 citations
with Lingxiao Wang, Zhaoran Wang
no link
Nonlinear Structured Signal Estimation in High Dimensions via Iterative Hard Thresholding
published 2018 · International Conference on Artificial Intelligence and Statistics · 6 citations
with Kaiqing Zhang, Zhaoran Wang
no link
Provable Gaussian Embedding with One Observation
published 2018 · Neural Information Processing Systems · 3 citations
with Ming Yu, Tuo Zhao, Mladen Kolar, Zhaoran Wang
Estimating High-dimensional Non-Gaussian Multiple Index Models via Stein’s Lemma
published 2017 · Neural Information Processing Systems · 14 citations
no link
High-dimensional Non-Gaussian Single Index Models via Thresholded Score Function Estimation
published 2017 · International Conference on Machine Learning · 37 citations
no link
Learning non-Gaussian multi-index model via second-order Stein's method
published 2017 · Neural Information Processing Systems · 11 citations
with Krishna Balasubramanian, Zhaoran Wang, Han Liu
no link
Assembled from arXiv and OpenAlex. Duplicate records for the same paper are merged, and the published version is shown where we could identify one. Corrections welcome.
Built from arXiv and OpenAlex. Supported by UKRI grant APP47921 (Martin Weidner, UCL · Francis J. DiTraglia, Oxford).