Location Dependency in Video Prediction

Hafez Farazi; Niloofar Azizi; Sven Behnke

arxiv: 1810.04937 · v2 · pith:GXXQHNPEnew · submitted 2018-10-11 · 💻 cs.CV

Location Dependency in Video Prediction

Niloofar Azizi , Hafez Farazi , Sven Behnke This is my paper

classification 💻 cs.CV

keywords videoconvolutionalpredictionspatiallyinvariantlocationlocation-dependentnetworks

0 comments

read the original abstract

Deep convolutional neural networks are used to address many computer vision problems, including video prediction. The task of video prediction requires analyzing the video frames, temporally and spatially, and constructing a model of how the environment evolves. Convolutional neural networks are spatially invariant, though, which prevents them from modeling location-dependent patterns. In this work, the authors propose location-biased convolutional layers to overcome this limitation. The effectiveness of location bias is evaluated on two architectures: Video Ladder Network (VLN) and Convolutional redictive Gating Pyramid (Conv-PGP). The results indicate that encoding location-dependent features is crucial for the task of video prediction. Our proposed methods significantly outperform spatially invariant models.

This paper has not been read by Pith yet.

Location Dependency in Video Prediction

discussion (0)