Generating Videos with Scene Dynamics

Vondrick, Carl; Pirsiavash, Hamed; Torralba, Antonio

doi:10.48550/arxiv.1609.02612

preprintarXiv (Cornell University)Sep 8, 2016GREEN OA

Generating Videos with Scene Dynamics

CVCarl Vondrick HPHamed Pirsiavash ATAntonio Torralba

Massachusetts Institute of Technology · University of Maryland, Baltimore County

Indexed inarxivdatacite

Abstract

We capitalize on large amounts of unlabeled video in order to learn a model of scene dynamics for both video recognition tasks (e.g. action classification) and video generation tasks (e.g. future prediction). We propose a generative adversarial network for video with a spatio-temporal convolutional architecture that untangles the scene's foreground from the background. Experiments suggest this model can generate tiny videos up to a second at full frame rate better than simple baselines, and we show its utility at predicting plausible futures of static images. Moreover, experiments and visualizations show the model internally learns useful features for recognizing actions with minimal supervision, suggesting…

Citation impact

848

total citations

FWCI: —
Percentile: —
References: 34

Citations per year

Authors

3

Topics & keywords

Topics

Keywords

Computer science
Artificial intelligence
Representation (politics)
Dynamics (music)
Generative grammar
Convolutional neural network
Generative model
Frame rate

No related works found for this paper.