Nvidia introduced Cosmos, an open-source, open-weight Video World Model
It's trained on 20M hours of videos and weighs from 4B to 14B. Cosmos offers two flavors: diffusion (continuous tokens) and autoregressive (discrete tokens); and two generation modes: text->video and text+video->video.
Physical AI has a big data problem. Synthetic data to the rescue.
Nvidia apply Cosmos to large-scale synthetic data generation for robotics and autonomous driving, and now you can too.
Post #2893
865