RESEARCH PAPERYear 1998

Reinforcement Learning: An Introduction

Richard S. Sutton; Andrew G. Barto

Classification

View four quadrants
Major category
Foundational work
Architecture
Not applicable
Prediction paradigm
Not applicable
Source review status
Explicit in survey

Category review. The 1998 RL textbook establishes MDPs, dynamic programming and Dyna integration of planning, acting and model learning. These are general mathematical/algorithmic foundations rather than one WAM architecture. Reading evidence

AT A GLANCE

Contribution

A contribution summary has not been added yet.

Abstract

An abstract has not been added yet.

Affiliations

Not listed in the collection.

BibTeX