RESEARCH PAPER
Motubrain: An Advanced World Action Model for Robot Control
Classification
View four quadrants- Major category
- WAMs
- Quadrant
- Not assigned
- Architecture
- Not assigned
- Prediction paradigm
- Not assigned
- Source review status
- Not assigned
Category review. Language, video and action streams exchange information in a unified generative policy, with joint denoising followed by accelerated action completion and robot execution. Reading evidence
Contribution
Motubrain combines language, video and action streams in a unified generative robot policy. It transfers video priors into relative end-effector control, then uses asymmetric attention and asynchronous action chunks for deployment. Simulation, world-prediction and real-robot evaluations support different capabilities; the autoregressive attention specification remains internally inconsistent.
Abstract
An abstract has not been added yet.
Affiliations
Not listed in the collection.