RESEARCH PAPER
MIND-V: Hierarchical World Model for Long-Horizon Robotic Manipulation with RL-based Physical Alignment
Classification
View four quadrants- Major category
- Benchmarks & simulators
- Quadrant
- Not applicable
- Architecture
- Not applicable
- Prediction paradigm
- Not applicable
- Source review status
- Verified from primary sources
Category review. MIND-V decomposes instructions into object/arm trajectories and generates candidate manipulation videos with a specialized CogVideoX-based renderer. Its iterative feedback judges generated futures; the separate robot-policy experiment uses synthetic visual goals. The main deliverable is task-conditioned neural world simulation rather than a generic pretrained backbone. Reading evidence
Contribution
A contribution summary has not been added yet.
Abstract
An abstract has not been added yet.
Affiliations
Not listed in the collection.