Discrete-JEPA learns discrete semantic image tokens through latent predictive coding without pixel reconstruction, and achieves stable long-horizon prediction on synthetic symbolic tasks.
Data2vec: A general framework for self-supervised learning in speech, vision and language
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CV 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Discrete JEPA: Learning Discrete Token Representations without Reconstruction
Discrete-JEPA learns discrete semantic image tokens through latent predictive coding without pixel reconstruction, and achieves stable long-horizon prediction on synthetic symbolic tasks.