科研速览 · Science Skim继续刷下去 · Keep skimming →
◆ Cambridge University Press eBooks2026-07-31· Reinforcement learning

Introduction and Overview

Shie Mannor, Yishay Mansour, Aviv Tamar

原始摘要(英文原文)· Original abstract
This chapter introduces reinforcement learning (RL) as the discipline of learning and acting in environments where sequential decisions are made. It traces the origins of RL across multiple disciplines including optimal control, operations research, neuroscience and psychology. The chapter motivates the field through landmark successes in game playing (checkers, backgammon, go, chess), Atari video games, robotics and language model fine-tuning. The primary mathematical model – the Markov Decision Process (MDP) – is introduced as the framework for handling uncertainty in dynamics, actions and knowledge. The chapter outlines the book’s organization into two main themes: planning (optimal decision making with known models) and learning (decision making with unknown models), and presents the key tradeoff between exploitation and exploration inherent to RL problems.
读原文 · Read the paper ↗

AI 追问PRO

登录后使用 AI 追问

讨论区

登录后参与讨论

相关论文 · Related

Introduction and Overview — 科研速览 Science Skim