未来を予測して動画を生成 – Generating Videos with Scene Dynamics

LINK

arXiv

一枚の画像から、前景背景を推定し、その画像の次の1〜2秒の動画を生成する研究です。

Arxiv(2016 11月公開)

We capitalize on large amounts of unlabeled video in order to learn a model of scene dynamics for both video recognition tasks (e.g. action classification) and video generation tasks (e.g. future prediction). We propose a generative adversarial network for video with a spatio-temporal convolutional architecture that untangles the scene’s foreground from the background. Experiments suggest this model can generate tiny videos up to a second at full frame rate better than simple baselines, and we show its utility at predicting plausible futures of static images. Moreover, experiments and visualizations show the model internally learns useful features for recognizing actions with minimal supervision, suggesting scene dynamics are a promising signal for representation learning. We believe generative video models can impact many applications in video understanding and simulation.

人工知能と表現の今

未来を予測して動画を生成 – Generating Videos with Scene Dynamics –

LINK

関連

TAG

SHARE US

SEARCH

キーワード検索

タグ検索