HyperAIHyperAI

Command Palette

Search for a command to run...

5 months ago

Generating Videos with Scene Dynamics

Carl Vondrick; Hamed Pirsiavash; Antonio Torralba

Generating Videos with Scene Dynamics

Abstract

We capitalize on large amounts of unlabeled video in order to learn a model of scene dynamics for both video recognition tasks (e.g. action classification) and video generation tasks (e.g. future prediction). We propose a generative adversarial network for video with a spatio-temporal convolutional architecture that untangles the scene's foreground from the background. Experiments suggest this model can generate tiny videos up to a second at full frame rate better than simple baselines, and we show its utility at predicting plausible futures of static images. Moreover, experiments and visualizations show the model internally learns useful features for recognizing actions with minimal supervision, suggesting scene dynamics are a promising signal for representation learning. We believe generative video models can impact many applications in video understanding and simulation.

Benchmarks

BenchmarkMethodologyMetrics
self-supervised-action-recognition-on-ucf101VideoGan (C3D)
3-fold Accuracy: 52.1
Frozen: false
Pre-Training Dataset: UCF101
video-generation-on-ucf-101-16-framesVGAN
Inception Score: 8.18
video-generation-on-ucf-101-16-frames-64x64VGAN
Inception Score: 8.18

Build AI with AI

From idea to launch — accelerate your AI development with free AI co-coding, out-of-the-box environment and best price of GPUs.

AI Co-coding
Ready-to-use GPUs
Best Pricing
Get Started

Hyper Newsletters

Subscribe to our latest updates
We will deliver the latest updates of the week to your inbox at nine o'clock every Monday morning
Powered by MailChimp
Generating Videos with Scene Dynamics | Papers | HyperAI