RESEARCH PAPERYear 2022

Temporal Difference Learning for Model Predictive Control

Nicklas A Hansen; Hao Su; Xiaolong Wang

Classification

View four quadrants
Major category
Foundational work
Architecture
Not applicable
Prediction paradigm
Not applicable
Source review status
Verified from primary sources

Category review. TD-MPC establishes a pre-2026 general integration of task-oriented latent dynamics, terminal value learning and model-predictive action optimization. It is a canonical foundation explicitly used in the manuscript WM/MBRL thread. Reading evidence

AT A GLANCE

Contribution

A contribution summary has not been added yet.

Abstract

An abstract has not been added yet.

Affiliations

Not listed in the collection.

BibTeX