Convergent Policy Optimization for Safe Reinforcement Learning
Jan 1, 2019·
,
,·
0 min read
Ming Yu
Zhuoran Yang
Mladen Kolar
Zhaoran Wang

Authors
PhD (2016-2020)
Ming received his PhD in Econometrics and Statistics at University of Chicago, Booth School of Business in March 2020. His research interests include high dimensional statistical inference, non-convex optimization, and reinforcement learning, with a focus on developing novel methodologies with both practical applications and theoretical guarantees.

Authors
Professor of Data Sciences and Operations
Mladen Kolar is a Professor of Data Sciences and Operations at the University of Southern California Marshall School of Business and a Visiting Professor of Statistics and Data Science at Mohamed bin Zayed University of Artificial Intelligence. Before joining USC, he was on the faculty of the University of Chicago Booth School of Business. His research is focused on high-dimensional statistical methods, graphical models, varying-coefficient models and data mining, driven by the need to uncover interesting and scientifically meaningful structures from observational data. He is a Fellow of the Institute of Mathematical Statistics.