RESEARCH PAPERYear 1998
Reinforcement Learning: An Introduction
Classification
View four quadrants- Major category
- Foundational work
- Quadrant
- Not applicable
- Architecture
- Not applicable
- Prediction paradigm
- Not applicable
- Source review status
- Explicit in survey
Category review. The 1998 RL textbook establishes MDPs, dynamic programming and Dyna integration of planning, acting and model learning. These are general mathematical/algorithmic foundations rather than one WAM architecture. Reading evidence
Contribution
A contribution summary has not been added yet.
Abstract
An abstract has not been added yet.
Affiliations
Not listed in the collection.