from    
to    
search  

 


AIR学术沙龙第36期|人工智能医疗保健和虚拟世界的可穿戴传感器和触觉技术
先立后破?实现双碳目标
迎接生物药制造的第四次浪潮-
支撑未来海量资源接入,电力系统通用信息模型(CIM)发展探讨
报告题目:
Policy Iteration Algorithms
 报告人:
Uri Zwick
Prof. Tel Aviv University
报告时间:
2009-09-21 09:45
报告地点:
FIT楼多功能厅
主办单位:
清华大学理论计算机科学研究中心
  简介:

Abstract

 

The policy iteration algorithms is a simple family of algorithms that can be applied in many different settings, ranging from the relatively simple problem of finding a minimum mean weight cycle in a graph, the more challenging solution of Markov Decision Processes (MDPs), to the solution of 2-player full information stochastic games, also known as Simple Stochastic Games (SSGs).
It was recently shown by Fridmann that the worst case running time of a natural deterministic version of the policy iteration algorithm, when applied to Parity Games (PGs), is exponential. It is still open, however, whether deterministic policy iteration algorithm can solve Markov Decision Processes in polynomial time, and whether randomized policy iteration algorithms can solve Simple Stochastic Games in polynomial time.
The talk will survey what is known regarding policy iteration algorithms and mention many intriguing open problems.

 

Bio of the Speaker

 

Uri Zwick received his B.Sc. degree in Computer Science from the Technion, Israel Institute of Technology, and his M.Sc. and Ph.D. degrees in Computer Science from Tel Aviv University. He is currently a Professor of Computer Science in Tel Aviv University. His main research interests are: algorithms and complexity, combinatorial optimization, mathematical games, and recreational mathematics.
今日相关信息
清华大学科学社会学与政策学沙龙第55期:...
Mobility in Ad Hoc Wireless Networks:...
English in Switzerland in the 21st Ce...
张福运基金会年度学术讲座:法律推动社会变...
排泄地球大气层二氧化碳——在对流层和同温...
 
同类别相关信息
Testing Gravity on Cosmological Scales
Coordination in Multi-Agent Systems...
Revealing the nature of dark matter...
AI for Imperfect-Information Games:...
Reverberation mapping of AGNs: role...
学术活动