Department of Mathematics - Seminar on Statistics and Data Science - Towards Optimal Sample Complexities in Offline Reinforcement Learning and Markov Games

12月7日

10:30am - 11:30am

研討會, 演講, 講座

Emerging reinforcement learning (RL) applications necessitate the design of sample-efficient solutions in order to accommodate the explosive growth of problem dimensionality. Despite the empirical success, however, our understanding about the statistical limits of RL remains highly incomplete. In this talk, I will present some recent progress towards settling the sample complexity in two RL scenarios. The first one is concerned with offline or batch RL, which performs learning using only pre-collected data without further exploration. We prove that model-based offline RL --- a plug-in approach that leverages the pessimism principle with Bernstein-style penalty --- achieves minimal-optimal sample complexity without any burn-in cost. The second scenario is concerned with multi-agent RL in zero-sum Markov games, assuming access to a generative model (a.k.a. simulator). We develop a new algorithm --- built upon the integration of adaptive sampling, online learning, and the optimism principle --- that overcomes the curse of multi-agents and the barrier of long horizon simultaneously. Our results emphasize the prolific interplay between high-dimensional statistics, online learning, and game theory. (See https://arxiv.org/abs/2204.05275 and https://arxiv.org/abs/2208.10458 for more details).

This is based on joint work with Gen Li, Laixi Shi, Yuling Yan, Yuejie Chi, Jianqing Fan, and Yuting Wei.

12月7日

10:30am - 11:30am

立即登記

地點

https://hkust.zoom.us/j/95008648547 (Passcode: hkust)

講者/表演者

Prof. Yuxin CHEN
University of Pennsylvania

主辦單位

Department of Mathematics

聯絡方法

付款詳情

對象

Alumni, Faculty and staff, PG students, UG students

語言

英語

其他活動

7月14日

研討會, 演講, 講座

IAS / School of Science Joint Lecture - Boron Clusters

Abstract The study of carbon clusters led to the discoveries of fullerenes, carbon nanotubes, and graphene. Are there other elements that can form similar nanostructures? To answer this questio...

5月15日

研討會, 演講, 講座

IAS / School of Science Joint Lecture - Laser Spectroscopy of Computable Atoms and Molecules with Unprecedented Accuracy

Abstract Precision spectroscopy of the hydrogen atom, a fundamental two-body system, has been instrumental in shaping quantum mechanics. Today, advances in theory and experiment allow us to ext...