
A structured survey of how diffusion models extend from image to video generation, unpacking the temporal-consistency and data-scarcity problems that define the current research frontier — useful for engineers building or evaluating generative video systems.
“The task itself is a superset of the image case, since an image is a video of 1 frame”
Checking sign-in…
Loading comments…