EconBase
← All authors

Emma Brunskill

Stanford University (per OpenAlex) · ORCID · OpenAlex

35 papers in scope · 34 published · 1 on the econ.EM arXiv · 814 citations · h-index 17 (over the papers listed here)

Related authors

The 20 authors closest to this one in our weighted citation graph, most related first.

  1. Yiwei Sun
  2. Kentaro Kawato
  3. Chen Qiu
  4. Raphaël Langevin
  5. Brenda Prallon
  6. Jörg Stoye
  7. Shosei Sakaguchi
  8. José Luis Montiel Olea
  9. Andres Fernandez
  10. Henry Zhu
  11. Lezhi Tan
  12. Jiaqi Huang
  13. José Blanchet
  14. Toru Kitagawa
  15. Patrik Guggenberger
  16. Charles F. Manski
  17. Emily Breza
  18. Arun G. Chandrasekhar
  19. Yuichi Kitamura
  20. Aleksey Tetenov

Proximity is measured over citations between two papers we both hold, weighted by how heavily one leans on the other, and is symmetric — it does not distinguish citing from being cited. Authors without a profile here are skipped, and a genuinely close colleague can be missing simply because their work is not in our arXiv corpus. Method: docs/06-citations-pipeline.md.

Papers

(1 of 35)

A statistical test for the benefits of personalizing interventions
published2026 · Science
with Zhaoqi Li
Cost-Aware Near-Optimal Policy Learning
published2025 · Proceedings of the AAAI Conference on Artificial Intelligence
with Joy He-Yueya, Jonathan Lee, Matthew Jörke
Evaluating Treatment Prioritization Rules via Rank-Weighted Average Treatment Effects
published2024 · Journal of the American Statistical Association · 51 citations · first circulated 2021
with Steve Yadlowsky, Scott L. Fleming, Nigam H. Shah, Stefan Wager, Scott Fleming
MedAlign: A Clinician-Generated Dataset for Instruction Following with Electronic Medical Records
published2024 · Proceedings of the AAAI Conference on Artificial Intelligence · 46 citations · first circulated 2023
with Scott L. Fleming, Alejandro Lozano, William J. Haberkorn, Jenelle Jindal, Eduardo Pontes Reis, Rahul Thapa, Louis Blankemeier, Julian Z. Genkins, Ethan Steinberg, Ashwin Nayak, Birju Patel, Chia-Chun Chiang, …
working paper2024 · arXiv · 2 citations
Reinforcement learning tutor better supported lower performers in a math task
published2024 · Machine Learning · 26 citations · first circulated 2023
with Sherry Ruan, Allen Nie, William Steenbergen, Jiayu He, J. Q. Zhang, Meng Guo, Yao Liu, Kyle Dang Nguyen, Catherine Y. Wang, Rui Ying, James A. Landay, JQ Zhang, …
Model-Based Offline Reinforcement Learning with Local Misspecification
published2023 · Proceedings of the AAAI Conference on Artificial Intelligence
with Kefan Dong, Yannis Flet-Berliac, Allen Nie
Constraint Sampling Reinforcement Learning: Incorporating Expertise for Faster Learning
published2022 · Proceedings of the AAAI Conference on Artificial Intelligence · 6 citations · first circulated 2021
with Tong Mu, Georgios Theocharous, David Arbour
Reinforcement Learning with State Observation Costs in Action-Contingent Noiselessly Observable Markov Decision Processes
published2021 · Neural Information Processing Systems · 7 citations
with Hyunji Nam, Scott L. Fleming
Learning When-to-Treat Policies
published2020 · Journal of the American Statistical Association · 48 citations · first circulated 2019
Provably Good Batch Reinforcement Learning Without Great Exploration
published2020 · Neural Information Processing Systems · 36 citations
with Yao Liu, Adith Swaminathan, Alekh Agarwal
Learning Near Optimal Policies with Low Inherent Bellman Error
published2020 · International Conference on Machine Learning · 38 citations
with Andrea Zanette, Alessandro Lazaric, Mykel J. Kochenderfer
Understanding the Curse of Horizon in Off-Policy Evaluation via Conditional Importance Sampling
published2020 · International Conference on Machine Learning · 8 citations · first circulated 2019
with Yao Liu, Pierre-Luc Bacon
Sublinear Optimal Policy Value Estimation in Contextual Bandits
published2020 · International Conference on Artificial Intelligence and Statistics · 3 citations · first circulated 2019
with Weihao Kong, Gregory Valiant
Being Optimistic to Be Conservative: Quickly Learning a CVaR Policy
published2020 · Proceedings of the AAAI Conference on Artificial Intelligence · 36 citations
with Ramtin Keramati, Christoph Dann, Alex Tamkin
Off-policy Policy Evaluation For Sequential Decisions Under Unobserved Confounding
published2020 · Neural Information Processing Systems · 18 citations
with Hongseok Namkoong, Ramtin Keramati, Steve Yadlowsky
Provably Good Batch Off-Policy Reinforcement Learning Without Great Exploration
published2020 · Neural Information Processing Systems · 18 citations
with Yao Liu, Adith Swaminathan, Alekh Agarwal
Provably Efficient Reward-Agnostic Navigation with Linear Value Iteration
published2020 · Neural Information Processing Systems · 13 citations
with Andrea Zanette, Alessandro Lazaric, Mykel J. Kochenderfer
Off-Policy Policy Gradient with State Distribution Correction
published2019 · Uncertainty in Artificial Intelligence · 46 citations
with Yao Liu, Adith Swaminathan, Alekh Agarwal
Limiting Extrapolation in Linear Approximate Value Iteration
published2019 · Neural Information Processing Systems · 17 citations
with Andrea Zanette, Alessandro Lazaric, Mykel J. Kochenderfer
Offline Contextual Bandits with High Probability Fairness Guarantees
published2019 · Neural Information Processing Systems · 27 citations
with Blossom Metevier, Stephen Giguere, Sarah Brockman, Ari Kobren, Yuriy Brun, Philip S. Thomas
Almost Horizon-Free Structure-Aware Best Policy Identification with a Generative Model
published2019 · Neural Information Processing Systems · 13 citations
with Andrea Zanette, Mykel J. Kochenderfer
Fake It Till You Make It: Learning-Compatible Performance Support.
published2019 · Uncertainty in Artificial Intelligence · 3 citations
with Jonathan Bragg
Decoupling Gradient-Like Learning Rules from Representations.
published2018 · International Conference on Machine Learning · 2 citations · first circulated 2017
with Philip S. Thomas, Christoph Dann
The Misidentified Identifiability Problem of Bayesian Knowledge Tracing.
published2017 · Grantee Submission · 4 citations
with Shayan Doroudi
Importance Sampling with Unequal Support
published2017 · Proceedings of the AAAI Conference on Artificial Intelligence · 8 citations
with Philip S. Thomas
Where to Add Actions in Human-in-the-Loop Reinforcement Learning
published2017 · Proceedings of the AAAI Conference on Artificial Intelligence · 42 citations
with Travis Mandel, Yun-En Liu, Zoran Popović
Predictive Off-Policy Policy Evaluation for Nonstationary Decision Problems, with Applications to Digital Marketing
published2017 · Proceedings of the AAAI Conference on Artificial Intelligence · 31 citations
with Philip S. Thomas, Georgios Theocharous, Mohammad Ghavamzadeh, Ishan Durugkar
Unifying PAC and Regret: Uniform PAC Bounds for Episodic Reinforcement Learning
published2017 · Neural Information Processing Systems · 103 citations
with Christoph Dann, Tor Lattimore
Latent contextual bandits and their application to personalized recommendations for new users
published2016 · International Joint Conference on Artificial Intelligence · 22 citations
with Li Zhou
Efficient Bayesian clustering for reinforcement learning
published2016 · International Joint Conference on Artificial Intelligence · 14 citations
with Travis Mandel, Yun-En Liu, Zoran Popović
Energetic natural gradient descent
published2016 · International Conference on Machine Learning · 8 citations
with Philip S. Thomas, Bruno Castro da Silva, Christoph Dann
Offline Evaluation of Online Reinforcement Learning Algorithms
published2016 · Proceedings of the AAAI Conference on Artificial Intelligence · 12 citations
with Travis Mandel, Yun-En Liu, Zoran Popović
The Queue Method: Handling Delay, Heuristics, Prior Data, and Evaluation in Bandits
published2015 · Proceedings of the AAAI Conference on Artificial Intelligence · 36 citations
with Travis Mandel, Yun-En Liu, Zoran Popović
PUMA: Planning Under Uncertainty with Macro-Actions
published2010 · Proceedings of the AAAI Conference on Artificial Intelligence · 70 citations
with Ruijie He, Nicholas Roy

Assembled from arXiv and OpenAlex. Duplicate records for the same paper are merged, and the published version is shown where we could identify one. Corrections welcome.