Pith. sign in

REVIEW 1 cited by

Spatiotemporally Consistent Indoor Lighting Estimation with Diffusion Priors

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2508.08384 v1 pith:VK6IACKH submitted 2025-08-11 cs.GR cs.AIcs.CV

Spatiotemporally Consistent Indoor Lighting Estimation with Diffusion Priors

classification cs.GR cs.AIcs.CV
keywords lightingestimationdiffusionimageindoorlightspatiotemporallyvideo
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

Indoor lighting estimation from a single image or video remains a challenge due to its highly ill-posed nature, especially when the lighting condition of the scene varies spatially and temporally. We propose a method that estimates from an input video a continuous light field describing the spatiotemporally varying lighting of the scene. We leverage 2D diffusion priors for optimizing such light field represented as a MLP. To enable zero-shot generalization to in-the-wild scenes, we fine-tune a pre-trained image diffusion model to predict lighting at multiple locations by jointly inpainting multiple chrome balls as light probes. We evaluate our method on indoor lighting estimation from a single image or video and show superior performance over compared baselines. Most importantly, we highlight results on spatiotemporally consistent lighting estimation from in-the-wild videos, which is rarely demonstrated in previous works.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Video Generation Models Are Inherent Lighting Estimators

    cs.CV 2026-07 conditional novelty 6.0

    Video diffusion models recover temporally coherent HDR environment maps from single in-the-wild videos by chrome-ball inpainting plus an HDR-aware VAE and LoRA adaptation.