from    
to    
search  

 


第478期“工物学术论坛”:X射线探测器领域的行业发展情况和机遇
天文系 Colloquium: Interstellar X-ray Dust Scattering: Current Research andFu...
【图书馆系列讲座】开题与立项前的文献调研概述(理工类)
【图书馆系列讲座】开题与立项前的文献调研概述(社科类)
报告题目:
Policy Iteration Algorithms
 报告人:
Uri Zwick
Prof. Tel Aviv University
报告时间:
2009-09-21 09:45
报告地点:
FIT楼多功能厅
主办单位:
清华大学理论计算机科学研究中心
  简介:

Abstract

 

The policy iteration algorithms is a simple family of algorithms that can be applied in many different settings, ranging from the relatively simple problem of finding a minimum mean weight cycle in a graph, the more challenging solution of Markov Decision Processes (MDPs), to the solution of 2-player full information stochastic games, also known as Simple Stochastic Games (SSGs).
It was recently shown by Fridmann that the worst case running time of a natural deterministic version of the policy iteration algorithm, when applied to Parity Games (PGs), is exponential. It is still open, however, whether deterministic policy iteration algorithm can solve Markov Decision Processes in polynomial time, and whether randomized policy iteration algorithms can solve Simple Stochastic Games in polynomial time.
The talk will survey what is known regarding policy iteration algorithms and mention many intriguing open problems.

 

Bio of the Speaker

 

Uri Zwick received his B.Sc. degree in Computer Science from the Technion, Israel Institute of Technology, and his M.Sc. and Ph.D. degrees in Computer Science from Tel Aviv University. He is currently a Professor of Computer Science in Tel Aviv University. His main research interests are: algorithms and complexity, combinatorial optimization, mathematical games, and recreational mathematics.
今日相关信息
清华大学科学社会学与政策学沙龙第55期:...
Mobility in Ad Hoc Wireless Networks:...
English in Switzerland in the 21st Ce...
张福运基金会年度学术讲座:法律推动社会变...
排泄地球大气层二氧化碳——在对流层和同温...
 
同类别相关信息
Quantum computing and quantum machi...
AIR学术工作坊第1期 | AI赋能基因分析...
Real bordism, Real orientations, an...
AIR学术沙龙第2期|人工智能赋能个体化...
清华信息前沿交叉论坛
学术活动