Generating Videos with Scene Dynamics

Vondrick, Carl; Pirsiavash, Hamed; Torralba, Antonio

Computer Science > Computer Vision and Pattern Recognition

arXiv:1609.02612 (cs)

[Submitted on 8 Sep 2016 (v1), last revised 26 Oct 2016 (this version, v3)]

Title:Generating Videos with Scene Dynamics

Authors:Carl Vondrick, Hamed Pirsiavash, Antonio Torralba

View PDF

Abstract:We capitalize on large amounts of unlabeled video in order to learn a model of scene dynamics for both video recognition tasks (e.g. action classification) and video generation tasks (e.g. future prediction). We propose a generative adversarial network for video with a spatio-temporal convolutional architecture that untangles the scene's foreground from the background. Experiments suggest this model can generate tiny videos up to a second at full frame rate better than simple baselines, and we show its utility at predicting plausible futures of static images. Moreover, experiments and visualizations show the model internally learns useful features for recognizing actions with minimal supervision, suggesting scene dynamics are a promising signal for representation learning. We believe generative video models can impact many applications in video understanding and simulation.

Comments:	NIPS 2016. See more at this http URL
Subjects:	Computer Vision and Pattern Recognition (cs.CV); Graphics (cs.GR); Machine Learning (cs.LG)
Cite as:	arXiv:1609.02612 [cs.CV]
	(or arXiv:1609.02612v3 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.1609.02612

Submission history

From: Carl Vondrick [view email]
[v1] Thu, 8 Sep 2016 22:29:52 UTC (1,926 KB)
[v2] Mon, 17 Oct 2016 03:13:10 UTC (1,927 KB)
[v3] Wed, 26 Oct 2016 13:58:10 UTC (1,927 KB)

Full-text links:

Access Paper:

view license

Current browse context:

cs.CV

< prev | next >

new | recent | 2016-09

Change to browse by:

cs
cs.GR
cs.LG

References & Citations

DBLP - CS Bibliography

listing | bibtex

Carl Vondrick
Hamed Pirsiavash
Antonio Torralba

export BibTeX citation

Computer Science > Computer Vision and Pattern Recognition

Title:Generating Videos with Scene Dynamics

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Generating Videos with Scene Dynamics

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators