Spatial Prediction pretext task learns spatial structure in self-supervised learning by regressing relative position and scale between image views, yielding more structured representations and better generalization.
Learning to see through a baby’s eyes: Early visual diets enable robust visual intelligence in humans and machines
2 Pith papers cite this work. Polarity classification is still indexing.
2
Pith papers citing it
citation-role summary
background 1
citation-polarity summary
fields
cs.CV 2years
2026 2verdicts
UNVERDICTED 2roles
background 1polarities
background 1representative citing papers
PRISM is a pyramid vision architecture using iterative slot memory for progressive reasoning that reports competitive performance on classification, detection, and segmentation with improved robustness to occlusions.
citing papers explorer
-
Learning to Perceive "Where": Spatial Pretext Tasks for Robust Self-Supervised Learning
Spatial Prediction pretext task learns spatial structure in self-supervised learning by regressing relative position and scale between image views, yielding more structured representations and better generalization.
-
PRISM: Progressive Reasoning through Iterative Slot Memory for Vision
PRISM is a pyramid vision architecture using iterative slot memory for progressive reasoning that reports competitive performance on classification, detection, and segmentation with improved robustness to occlusions.