{"661867":{"#nid":"661867","#data":{"type":"event","title":"AI4OPT Seminar Series: Chi Jin","body":[{"value":"\u003Cp\u003E\u003Cstrong\u003EAI4OPT Seminar Series\u003C\/strong\u003E\u003C\/p\u003E\r\n\r\n\u003Cp\u003EDate: Thursday, October 6, 2022\u003C\/p\u003E\r\n\r\n\u003Cp\u003ELocation: Virtual Meeting\u003C\/p\u003E\r\n\r\n\u003Cp\u003ETime: Noon \u0026ndash; 1:00 pm\u003C\/p\u003E\r\n\r\n\u003Cp\u003EMeeting Link:\u0026nbsp;\u003Ca href=\u0022https:\/\/gatech.zoom.us\/j\/99381428980\u0022 tabindex=\u0022-1\u0022 title=\u0022https:\/\/gatech.zoom.us\/j\/99381428980\u0022\u003Ehttps:\/\/gatech.zoom.us\/j\/99381428980\u003C\/a\u003E\u003C\/p\u003E\r\n\r\n\u003Cp\u003E\u003Cstrong\u003ESpeaker: Chi Jin (\u91d1\u9a70)\u003C\/strong\u003E\u003C\/p\u003E\r\n\r\n\u003Cp\u003E\u003Cstrong\u003EWhen Is Partially Observable Reinforcement Learning Not Scary?\u003C\/strong\u003E\u003C\/p\u003E\r\n\r\n\u003Cp\u003E\u003Cstrong\u003EAbstract:\u003C\/strong\u003E\u0026nbsp;Partially observability is ubiquitous in applications of Reinforcement Learning (RL), in which agents learn to make a sequence of decisions despite lacking complete information about the latent states of the controlled system. Partially observable RL is notoriously difficult in theory---well-known information-theoretic results show that learning partially observable Markov decision processes (POMDPs) requires an exponential number of samples in the worst case. Yet, this does not rule out the possible existence of interesting subclasses of POMDPs, which include a large set of partial observable applications in practice while being tractable.\u0026nbsp;In this talk we identify a rich family of tractable POMDPs, which we call weakly revealing POMDPs. This family rules out the pathological instances of POMDPs where observations are uninformative to a degree that makes learning hard. We prove that for weakly revealing POMDPs, a simple algorithm combining optimism and Maximum Likelihood Estimation (MLE) is sufficient to guarantee a polynomial sample complexity. To the best of our knowledge, this gives the first line of provably sample-efficient results for learning from interactions in POMDPs. This is based on joint works with Qinghua Liu, Alan Chung, Akshay Krishnamurthy, Sham Kakade, and Csaba Szepesvari.\u003C\/p\u003E\r\n\r\n\u003Cp\u003E\u003Cstrong\u003EBio:\u0026nbsp;\u003C\/strong\u003EChi Jin is an assistant professor at the Electrical and Computer Engineering department of Princeton University. He obtained his Ph.D. in Computer Science at University of California, Berkeley, advised by Michael I. Jordan. His research mainly focuses on theoretical machine learning, with special emphasis on nonconvex optimization and reinforcement learning. His representative work includes proving noisy gradient descent escape saddle points efficiently and proving the efficiency of Q-learning and least-squares value iteration when combined with optimism in reinforcement learning.\u003C\/p\u003E\r\n\r\n\u003Cp\u003E\u003Cstrong\u003EAbout Seminar Series:\u003C\/strong\u003E\u003C\/p\u003E\r\n\r\n\u003Cp\u003EArtificial Intelligence Institute for Advances in Optimization (\u003Ca href=\u0022https:\/\/www.ai4opt.org\/\u0022\u003EAI4OPT\u003C\/a\u003E) is an NSF funded AI institute jointly between Georgia Tech and several other institutions. Starting this Fall, the institute is kicking off a new seminar series, broadly on AI and Optimization. The weekly seminar announcements will be sent in the new ai4opt-seminars mailing list. To receive these announcements, please subscribe here: \u003Ca href=\u0022https:\/\/lists.isye.gatech.edu\/mailman\/listinfo\/ai4opt-seminars\u0022\u003Ehttps:\/\/lists.isye.gatech.edu\/mailman\/listinfo\/ai4opt-seminars\u003C\/a\u003E\u003C\/p\u003E\r\n\r\n\u003Cp\u003EWe have the following lineup of speakers for the Fall semester (with a few more that will be added).\u003C\/p\u003E\r\n\r\n\u003Cul\u003E\r\n\t\u003Cli\u003EHamsa Bastani, Wharton School of Business, U. Penn\u003C\/li\u003E\r\n\t\u003Cli\u003ESatyen Kale,\u0026nbsp;Google Research\u003C\/li\u003E\r\n\t\u003Cli\u003EChi Jin, Princeton\u003C\/li\u003E\r\n\t\u003Cli\u003ESpyros Chatzivasileiadis, Technical University of Denmark\u003C\/li\u003E\r\n\t\u003Cli\u003EKarthyek Murthy, Singapore University of Technology and Design\u003C\/li\u003E\r\n\t\u003Cli\u003EDylan Foster, Microsoft Research\u003C\/li\u003E\r\n\t\u003Cli\u003ESoroosh Shafieezadeh Abadeh, CMU\u003C\/li\u003E\r\n\t\u003Cli\u003ESubhabrata Sen, Harvard\u003C\/li\u003E\r\n\u003C\/ul\u003E\r\n","summary":null,"format":"limited_html"}],"field_subtitle":"","field_summary":[{"value":"\u003Cp\u003EArtificial Intelligence Institute for Advances in Optimization (\u003Ca href=\u0022https:\/\/www.ai4opt.org\/\u0022\u003EAI4OPT\u003C\/a\u003E) is a NSF funded AI institute jointly between Georgia Tech and several other institutions. The institutes seminar series continues presentations on\u0026nbsp;topics\u0026nbsp;broadly surrounding AI and Optimization.\u003C\/p\u003E\r\n","format":"limited_html"}],"field_summary_sentence":[{"value":"AI4OPT\u0027s Seminar Series continues this Thursday with a virtual presentation by Chi Jin"}],"uid":"36348","created_gmt":"2022-10-05 17:35:47","changed_gmt":"2022-10-05 17:37:08","author":"Breon Martin","boilerplate_text":"","field_publication":"","field_article_url":"","field_event_time":{"event_time_start":"2022-10-06T13:00:00-04:00","event_time_end":"2022-10-06T14:00:00-04:00","event_time_end_last":"2022-10-06T14:00:00-04:00","gmt_time_start":"2022-10-06 17:00:00","gmt_time_end":"2022-10-06 18:00:00","gmt_time_end_last":"2022-10-06 18:00:00","rrule":null,"timezone":"America\/New_York"},"extras":[],"groups":[],"categories":[],"keywords":[],"core_research_areas":[],"news_room_topics":[],"event_categories":[{"id":"1795","name":"Seminar\/Lecture\/Colloquium"}],"invited_audience":[{"id":"78761","name":"Faculty\/Staff"},{"id":"174045","name":"Graduate students"},{"id":"78751","name":"Undergraduate students"}],"affiliations":[],"classification":[],"areas_of_expertise":[],"news_and_recent_appearances":[],"phone":[],"contact":[{"value":"\u003Cp\u003E\u003Cstrong\u003EGeneral Information\u003C\/strong\u003E\u003C\/p\u003E\r\n\r\n\u003Cp\u003ECami Douglass\u003C\/p\u003E\r\n\r\n\u003Cp\u003EExecutive Assistant\u003C\/p\u003E\r\n\r\n\u003Cp\u003E404-894-5953\u003C\/p\u003E\r\n\r\n\u003Cp\u003E\u0026nbsp;\u003C\/p\u003E\r\n\r\n\u003Cp\u003E\u003Cstrong\u003EPartnering with the Institute\u003C\/strong\u003E\u003C\/p\u003E\r\n\r\n\u003Cp\u003EStephanie Sigler\u003C\/p\u003E\r\n\r\n\u003Cp\u003ECorporate Relations Manager\u003C\/p\u003E\r\n\r\n\u003Cp\u003E404-894-4307\u003C\/p\u003E\r\n\r\n\u003Cp\u003E\u0026nbsp;\u003C\/p\u003E\r\n\r\n\u003Cp\u003E\u003Cstrong\u003EMedia Relations\u003C\/strong\u003E\u003C\/p\u003E\r\n\r\n\u003Cp\u003EBreon Martin\u003C\/p\u003E\r\n\r\n\u003Cp\u003EDirector of Communications\u003C\/p\u003E\r\n\r\n\u003Cp\u003E404-590-2220\u003C\/p\u003E\r\n","format":"limited_html"}],"email":[],"slides":[],"orientation":[],"userdata":""}}}