Video synthesis is the generation of novel video sequences from text, image, or other conditioning signals using generative models, most commonly diffusion models extended across a temporal dimension to maintain coherence between frames. It is used in creative tools for content generation, simulation, and film pre-visualisation. Key challenges include maintaining temporal consistency, object permanence, and physically plausible motion across generated frames.

Provenance