Date & Time:
December 6, 2019 10:30 am – 11:30 am
Location:
TTIC 526, 6045 S. Kenwood Ave., Chicago, IL,
12/06/2019 10:30 AM 12/06/2019 11:30 AM America/Chicago Zhaoran Wang (Northwestern) – Computational and Statistical Efficiency of Policy Optimization UChicago CS/Toyota Technological Institute of Chicago Machine Learning Seminar Series TTIC 526, 6045 S. Kenwood Ave., Chicago, IL,

On the Computational and Statistical Efficiency of Policy Optimization in (Deep) Reinforcements Learning

Coupled with powerful function approximators such as neural networks, policy optimization plays a key role in the tremendous empirical successes of deep reinforcement learning. In sharp contrast, the theoretical understandings of policy optimization remain rather limited from both the computational and statistical perspectives. From the perspective of computational efficiency, it remains unclear whether policy optimization converges to the globally optimal policy in a finite number of iterations, even given infinite data. From the perspective of statistical efficiency, it remains unclear how to attain the globally optimal policy with a finite regret or sample complexity.

To address the computational question, I will show that, under suitable conditions, natural policy gradient/proximal policy optimization/trust-region policy optimization (NPG/PPO/TRPO) converges to the globally optimal policy at a sublinear rate, even when it is coupled with neural networks. To address the statistical question, I will present an optimistic variant of NPG/PPO/TRPO, namely OPPO, which incorporates exploration in a principled manner and attains a sqrt{T}-regret. 

(Joint work with Qi Cai, Chi Jin, Jason Lee, Boyi Liu, Zhuoran Yang) 

Host: Mladen Kolar

Zhaoran Wang

Assistant Professor of Industrial Engineering and Management Sciences, Northwestern University

Zhaoran Wang is an assistant professor at Northwestern University, working at the interface of machine learning, statistics, and optimization. He is the recipient of the AISTATS (Artificial Intelligence and Statistics Conference) notable paper award, ASA (American Statistical Association) best student paper in statistical learning and data mining, INFORMS (Institute for Operations Research and the Management Sciences) best student paper finalist in data mining, and the Microsoft fellowship.

Related News & Events

headshots
In the News

Quantum Materials, Built By AI Robot

Apr 22, 2025
UChicago CS News

New Research Explores Augmented Breathing Through Thermal Feedback

Apr 21, 2025
headshot
UChicago CS News

University of Chicago’s Fred Chong Awarded $2 Million for Innovative Quantum Computing Cancer Research Project

Apr 04, 2025
simulated Roblox chat
UChicago CS News

Helping Elementary School Children Learn About Digital Privacy and Security With Micro-Lessons

Mar 25, 2025
grant ho writing on white board
UChicago CS News

New Study Reveals Gaps in Common Types of Cybersecurity Training

Mar 24, 2025
headshot
UChicago CS News

Jasmine Lu on Sustainable Computing: Rethinking E-Waste and Innovation

Mar 18, 2025
Pedro giving speech
UChicago CS News

Pedro Lopes Honored with 2025 IEEE VGTC Virtual Reality Significant New Researcher Award

Mar 13, 2025
ai generated network traffic
UChicago CS News

University of Chicago Researchers Revolutionize Network Traffic Generation with AI Breakthrough

Mar 12, 2025
UChicago CS News

Federal budget cuts threaten to decimate America’s AI superiority—and other countries are watching

Feb 25, 2025
Netflix logo on phone screen
UChicago CS News

The Hidden Cost of Netflix’s Autoplay: A Study on Viewing Patterns and User Control

Feb 25, 2025
Raul Castro Fernandez
UChicago CS News

Raul Castro Fernandez among six UChicago scientists awarded prestigious Sloan Fellowships in 2025

Feb 18, 2025
UChicago CS News

Quantum Leap: New Research Reveals Secrets of Random Quantum Circuits

Feb 04, 2025
arrow-down-largearrow-left-largearrow-right-large-greyarrow-right-large-yellowarrow-right-largearrow-right-smallbutton-arrowclosedocumentfacebookfacet-arrow-down-whitefacet-arrow-downPage 1CheckedCheckedicon-apple-t5backgroundLayer 1icon-google-t5icon-office365-t5icon-outlook-t5backgroundLayer 1icon-outlookcom-t5backgroundLayer 1icon-yahoo-t5backgroundLayer 1internal-yellowinternalintranetlinkedinlinkoutpauseplaypresentationsearch-bluesearchshareslider-arrow-nextslider-arrow-prevtwittervideoyoutube