科研速览 · Science Skim继续刷下去 · Keep skimming →
◆ Cambridge University Press eBooks2026-07-31· Markov decision process

Discounted Markov Decision Processes

Shie Mannor, Yishay Mansour, Aviv Tamar

原始摘要(英文原文)· Original abstract
This chapter provides a comprehensive treatment of stationary infinite-horizon MDPs with discounted return criterion. The chapter establishes that optimal stationary deterministic policies exist and develops two fundamental algorithms: value iteration and policy iteration. The theoretical foundation rests on contraction mapping theory, with the Banach fixed point theorem guaranteeing unique fixed points and geometric convergence. Error bounds and stopping criteria are derived.
读原文 · Read the paper ↗

AI 追问PRO

登录后使用 AI 追问

讨论区

登录后参与讨论

相关论文 · Related

Discounted Markov Decision Processes — 科研速览 Science Skim