RESEARCH PAPERYear 2022
Temporal Difference Learning for Model Predictive Control
Classification
View four quadrants- Major category
- Foundational work
- Quadrant
- Not applicable
- Architecture
- Not applicable
- Prediction paradigm
- Not applicable
- Source review status
- Verified from primary sources
Category review. TD-MPC establishes a pre-2026 general integration of task-oriented latent dynamics, terminal value learning and model-predictive action optimization. It is a canonical foundation explicitly used in the manuscript WM/MBRL thread. Reading evidence
Contribution
A contribution summary has not been added yet.
Abstract
An abstract has not been added yet.
Affiliations
Not listed in the collection.