from    
to    
search  

 


Phase transition of random plaquette models
Tiling with Electrons: fractionalization and emergent symmetry
Pyroptosis & Innate Immunity: Mechanisms & Therapeutics Potentials
【数学之美-杰出学者讲坛】2024年第6期 || Some recent results on conformally in...
报告题目:
Building Generalizable Agents by Learning to Plan
 报告人:
Yi Wu
Ph.D. candidate
报告时间:
2018-10-22 10:00
报告地点:
FIT 1-222
主办单位:
交叉信息研究院
  简介:

Abstract: Despite the tremendous successes by deep reinforcement learning (DRL), one critical issue for existing DRL works is generalization. A DRL agent is typically evaluated in the same environment as where it was trained. Therefore, the learned policy can be extremely specialized to the training scenarios and easily fail when the agent is tested in a new environment. In contrast, humans have the ability to adapt to new environments easily without further training. This generalization issue indicates a fundamental challenge towards bringing learning agents from lab to the real world.

 

This talk presents several recent progresses on this challenge by enabling the DRL agents to have two crucial capabilities: (1) the ability of performing long-term planning instead of merely memoizing the training experiences; (2) the ability of utilizing prior knowledge of the real world to derive better plans. We will show the agents with these planning capabilities generalize significantly better than classical DRL agents on a variety of challenging tasks.



Bio: Yi Wu is now a 5-th year Ph.D. candidate at UC Berkeley advised by Prof. Stuart Russell. He received his B.E. from the special pilot class (Yao class) from Institute for Interdisciplinary Information Sciences, Tsinghua University. Yi's research focuses on how to effectively incorporate human knowledge into AI models to produce both interpretable and generalizable solutions. He is now working on a variety of projects, including deep reinforcement learning, natural language processing and probabilistic programming. 



今日相关信息
Defects on carbon for electrocatalysis
Large Scale/Massive Antenna Arrays fo...
清华信息大讲堂179讲:Radio Access Netw...
清华信息大讲堂180讲:New scenarios fo...
 
同类别相关信息
量子计算+化学小型研讨会
清华大学材料科学与工程研究院《材料科学...
清华大学材料科学与工程研究院《材料科学...
Pseudo-criticality and its implicat...
Implications of superadditive algeb...
学术活动