RESEARCH PAPERYear 2025
CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer
Classification
View four quadrants- Major category
- Components of WAMs
- Quadrant
- Not applicable
- Architecture
- Not applicable
- Prediction paradigm
- Not applicable
- Source review status
- Verified from primary sources
Category review. CogVideoX provides a general causal spatiotemporal VAE and text/video diffusion transformer. Its outputs are generated videos and its learned representation/backbone can be reused by WAM systems; it is not a task-specific robotic controller. Reading evidence
Contribution
A contribution summary has not been added yet.
Abstract
An abstract has not been added yet.
Affiliations
Not listed in the collection.