RESEARCH PAPERYear 2019

When to Trust Your Model: Model-Based Policy Optimization

Michael Janner; Justin Fu; Marvin Zhang; Sergey Levine

Classification

View four quadrants
Major category
Foundational work
Architecture
Not applicable
Prediction paradigm
Not applicable
Source review status
Explicit in survey

Category review. MBPO is pre-2026 general theory and methodology for controlling model bias through short model rollouts from real states. It teaches how learned dynamics augment a separate policy learner, a foundational MBRL contribution. Reading evidence

AT A GLANCE

Contribution

A contribution summary has not been added yet.

Abstract

An abstract has not been added yet.

Affiliations

Not listed in the collection.

BibTeX