from    
to    
search  

 


Symmetry restoration and quantum Mpemba effects in chaotic andlocalization sy...
Quantum Gases 2024
Stories of Fermions in an Optical Box
Contractive Unitary and Classical Shadow Tomography
报告题目:
Policy Iteration Algorithms
 报告人:
Uri Zwick
Prof. Tel Aviv University
报告时间:
2009-09-21 09:45
报告地点:
FIT楼多功能厅
主办单位:
清华大学理论计算机科学研究中心
  简介:

Abstract

 

The policy iteration algorithms is a simple family of algorithms that can be applied in many different settings, ranging from the relatively simple problem of finding a minimum mean weight cycle in a graph, the more challenging solution of Markov Decision Processes (MDPs), to the solution of 2-player full information stochastic games, also known as Simple Stochastic Games (SSGs).
It was recently shown by Fridmann that the worst case running time of a natural deterministic version of the policy iteration algorithm, when applied to Parity Games (PGs), is exponential. It is still open, however, whether deterministic policy iteration algorithm can solve Markov Decision Processes in polynomial time, and whether randomized policy iteration algorithms can solve Simple Stochastic Games in polynomial time.
The talk will survey what is known regarding policy iteration algorithms and mention many intriguing open problems.

 

Bio of the Speaker

 

Uri Zwick received his B.Sc. degree in Computer Science from the Technion, Israel Institute of Technology, and his M.Sc. and Ph.D. degrees in Computer Science from Tel Aviv University. He is currently a Professor of Computer Science in Tel Aviv University. His main research interests are: algorithms and complexity, combinatorial optimization, mathematical games, and recreational mathematics.
今日相关信息
“中日民事诉讼法的制度与理论比较”国际研...
Succinct Data Structures
Chemically Enabled Fabrication of Mol...
Nanopatterns and nanomaterials: Synth...
The Science and Engineering of Protoc...
 
同类别相关信息
【数学之美-杰出学者讲坛】2024年第6期...
【数学之美-杰出学者讲坛】2024年第3期...
人工智能拓展火灾安全研究的进展
第四届清华信息前沿交叉论坛
【数学之美-杰出学者讲坛】2024年第2期...
学术活动