REVIEW 4 cited by
Deep Learning for Omnidirectional Vision: A Survey and New Perspectives
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Omnidirectional image (ODI) data is captured with a 360x180 field-of-view, which is much wider than the pinhole cameras and contains richer spatial information than the conventional planar images. Accordingly, omnidirectional vision has attracted booming attention due to its more advantageous performance in numerous applications, such as autonomous driving and virtual reality. In recent years, the availability of customer-level 360 cameras has made omnidirectional vision more popular, and the advance of deep learning (DL) has significantly sparked its research and applications. This paper presents a systematic and comprehensive review and analysis of the recent progress in DL methods for omnidirectional vision. Our work covers four main contents: (i) An introduction to the principle of omnidirectional imaging, the convolution methods on the ODI, and datasets to highlight the differences and difficulties compared with the 2D planar image data; (ii) A structural and hierarchical taxonomy of the DL methods for omnidirectional vision; (iii) A summarization of the latest novel learning strategies and applications; (iv) An insightful discussion of the challenges and open problems by highlighting the potential research directions to trigger more research in the community.
Forward citations
Cited by 4 Pith papers
-
InSpace: Structure-Aware 3D Indoor Scene Generation from a Single 360{\deg} Image
InSpace generates complete structure-aware 3D indoor scenes (layout plus textured assets) from a single equirectangular 360° image via three-stage flow matching with view- and asset-selective attention.
-
Seam360GS: Seamless 360{\deg} Gaussian Splatting from Real-World Omnidirectional Images
Training 3D Gaussian splatting with a learnable dual-fisheye distortion model, then turning it off at inference, renders seamless 360-degree novel views from imperfect panoramas.
-
Panoramic Scene Understanding: A Survey from Distortion-Aware Engineering to Sphere-Native Modeling
Survey organizing panoramic scene analysis literature by architectural design and training paradigm, identifying the absence of methods achieving both strict spherical equivariance and full reuse of perspective-pretra...
-
SO3UFormer: Learning Intrinsic Spherical Features for Rotation-Robust Panoramic Dense Prediction
SO3UFormer removes global latitude cues, adds quadrature-weighted spherical attention and gauge-pooled relative bias, and uses an SO(3)-consistency regularizer, retaining ~70.7 mIoU under arbitrary rotations where Sph...
Discussion (0). Sign in to comment.