RESEARCH PAPER
World Simulation with Video Foundation Models for Physical AI
Classification
View four quadrants- Major category
- Components of WAMs
- Quadrant
- Not applicable
- Architecture
- Not applicable
- Prediction paradigm
- Not applicable
- Source review status
- Verified from primary sources
Category review. Cosmos-Predict2.5 is a general text/visual-conditioned video foundation backbone with VAE latents, text embedding and flow-transformer generation; controlled and robot-action variants are adaptations. This technical report supplies the transferable Cosmos backbone explicitly intended by the component category. Reading evidence
Contribution
A contribution summary has not been added yet.
Abstract
An abstract has not been added yet.
Affiliations
Not listed in the collection.