Deep Convolutional Inverse Graphics Network

Joshua B. Tenenbaum; Pushmeet Kohli; Tejas D. Kulkarni; Will Whitney

arxiv: 1503.03167 · v4 · pith:PTI2NEJWnew · submitted 2015-03-11 · 💻 cs.CV · cs.GR· cs.LG· cs.NE

Deep Convolutional Inverse Graphics Network

Tejas D. Kulkarni , Will Whitney , Pushmeet Kohli , Joshua B. Tenenbaum This is my paper

classification 💻 cs.CV cs.GRcs.LGcs.NE

keywords modelgraphicsconvolutiondc-igndeepimagesinverselighting

0 comments

read the original abstract

This paper presents the Deep Convolution Inverse Graphics Network (DC-IGN), a model that learns an interpretable representation of images. This representation is disentangled with respect to transformations such as out-of-plane rotations and lighting variations. The DC-IGN model is composed of multiple layers of convolution and de-convolution operators and is trained using the Stochastic Gradient Variational Bayes (SGVB) algorithm. We propose a training procedure to encourage neurons in the graphics code layer to represent a specific transformation (e.g. pose or light). Given a single input image, our model can generate new images of the same object with variations in pose and lighting. We present qualitative and quantitative results of the model's efficacy at learning a 3D rendering engine.

This paper has not been read by Pith yet.

discussion (0)

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

Learning to Theorize the World from Observation
cs.LG 2026-05 unverdicted novelty 6.0

NEO induces compositional latent programs as world theories from observations and executes them to enable explanation-driven generalization.