RESEARCH PAPERYear 2019
When to Trust Your Model: Model-Based Policy Optimization
Classification
View four quadrants- Major category
- Foundational work
- Quadrant
- Not applicable
- Architecture
- Not applicable
- Prediction paradigm
- Not applicable
- Source review status
- Explicit in survey
Category review. MBPO is pre-2026 general theory and methodology for controlling model bias through short model rollouts from real states. It teaches how learned dynamics augment a separate policy learner, a foundational MBRL contribution. Reading evidence
Contribution
A contribution summary has not been added yet.
Abstract
An abstract has not been added yet.
Affiliations
Not listed in the collection.