Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T13:59:36.446502Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 68 of 68 outbound references and 0 inbound Pith citation observations for arXiv:2411.15787.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T13:59:36.446502Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
68 of 68 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c0c18bca-5231-4214-b82b-0e51ceb643c7 · outbound
Multi-Token Enhancing for Vision Representation Learning Asano, Christian Rupprecht, and Andrea Vedaldi
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation fb9c483a-0a2e-4aa0-95d1-237f9ae77cf2 · outbound
Multi-Token Enhancing for Vision Representation Learning BEit: BERT pre-training of image transformers
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 81363a80-e5b7-4c4c-93aa-0416cae652cb · outbound
Multi-Token Enhancing for Vision Representation Learning Cascade r-cnn: Delving into high quality object detection
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 17b47991-8b28-4e0b-9141-b7f273e49974 · outbound
Multi-Token Enhancing for Vision Representation Learning Unsupervised learn- ing of visual features by contrasting cluster assignments
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 0eb08d6c-8a62-4bb8-87af-39508d7990a7 · outbound
Multi-Token Enhancing for Vision Representation Learning Emerg- ing properties in self-supervised vision transformers
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation e5b8be40-8107-4031-972d-4d6a660ce06f · outbound
Multi-Token Enhancing for Vision Representation Learning Mixed autoencoder for self-supervised visual representation learning
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation bae4412f-d772-4e0f-80d8-54654691c16e · outbound
Multi-Token Enhancing for Vision Representation Learning A simple framework for contrastive learning of visual representations
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 4987cec4-5022-4483-86b2-723da00bcfe6 · outbound
Multi-Token Enhancing for Vision Representation Learning Exploring simple siamese rep- resentation learning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 7b450849-3c64-41c3-b76d-bf7cd4ad6571 · outbound
Multi-Token Enhancing for Vision Representation Learning An empiri- cal study of training self-supervised vision transformers
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 1c7a8955-0b96-42f2-b548-25581cf233fe · outbound
Multi-Token Enhancing for Vision Representation Learning Convit: Improving vision transformers with soft convolutional inductive biases
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d5345f7c-f55e-466c-8da6-f9b683ef5111 · outbound
Multi-Token Enhancing for Vision Representation Learning Ensemble methods in machine learn- ing
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 7c14abbd-a1ad-4f6c-9577-b8b7051979d7 · outbound
Multi-Token Enhancing for Vision Representation Learning An image is worth 16x16 words: Transformers for image recognition at scale
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 59e7332b-e982-491d-9d09-8a3b37c71d86 · outbound
Multi-Token Enhancing for Vision Representation Learning Whitening for self-supervised representation learning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f324fea3-b88c-4cbd-95d9-b320c9086d2e · outbound
Multi-Token Enhancing for Vision Representation Learning Seed: Self-supervised dis- tillation for visual representation
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f8a93134-6acc-4218-9dec-85847ba44a80 · outbound
Multi-Token Enhancing for Vision Representation Learning Evolved part masking for self-supervised learning
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 1df1351e-b688-472e-8008-58ad63218e85 · outbound
Multi-Token Enhancing for Vision Representation Learning Ganaie, Minghui Hu, A.K
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation cc5b256c-db5d-4777-a2d4-2980d658d30e · outbound
Multi-Token Enhancing for Vision Representation Learning Large-scale un- supervised semantic segmentation
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 87d73ed9-4a22-4f77-a574-d3764fe50431 · outbound
Multi-Token Enhancing for Vision Representation Learning Richemond, Elena Buchatskaya, Carl Doersch, Bernardo ´Avila Pires, Zhaohan Guo, Moham- mad Gheshlaghi Azar, Bilal Piot, Koray Kavukcuoglu, R´emi Munos, and Michal Valko
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 0007e657-5c59-4277-9ac2-129af3be0f57 · outbound
Multi-Token Enhancing for Vision Representation Learning Visual Attention Network
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 202ee373-4dee-40f8-b622-5c150b8e821b · outbound
Multi-Token Enhancing for Vision Representation Learning Hansen and P
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation ec2b9bb3-731d-42c2-8a6a-304c77232e7a · outbound
Multi-Token Enhancing for Vision Representation Learning Training independent subnetworks for robust prediction
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 6f0a9f0a-c6a9-420a-b565-579b2f1db22f · outbound
Multi-Token Enhancing for Vision Representation Learning Deep residual learning for image recognition
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7307fe2b-58c4-4e12-be64-f897cce54176 · outbound
Multi-Token Enhancing for Vision Representation Learning Momentum contrast for unsupervised visual rep- resentation learning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation fec79ee1-7fa5-4f2b-9c1b-320cee2d22b1 · outbound
Multi-Token Enhancing for Vision Representation Learning Masked autoencoders are scalable vision learners
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation ae5b7a43-a9ee-4420-8f01-4c65f282ca8a · outbound
Multi-Token Enhancing for Vision Representation Learning Distilling the Knowledge in a Neural Network
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f91d5ffd-e3db-4dee-ac1b-1705c7950631 · outbound
Multi-Token Enhancing for Vision Representation Learning Conv2Former: A Simple Transformer-Style ConvNet for Visual Recognition
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c8310df-939a-47f3-ba8c-54e875f15300 · outbound
Multi-Token Enhancing for Vision Representation Learning Averaging Weights Leads to Wider Optima and Better Generalization
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 008a7f77-e186-44e7-8e9b-595500a28f5c · outbound
Multi-Token Enhancing for Vision Representation Learning Similarity of neural network representa- tions revisited
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 6b864517-9da5-4755-9925-88bcbaf7e41c · outbound
Multi-Token Enhancing for Vision Representation Learning Fractalnet: Ultra-deep neural networks without residuals
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation cf0c168f-5f8b-47c9-b486-193d522a2456 · outbound
Multi-Token Enhancing for Vision Representation Learning Why M Heads are Better than One: Training a Diverse Ensemble of Deep Networks
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49c894ea-0fc9-4cbd-be3b-b64fc9434d9c · outbound
Multi-Token Enhancing for Vision Representation Learning Sere: Exploring feature self-relation for self-supervised trans- former
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 06f74ad0-5d71-4f4f-8297-eb3141445aee · outbound
Multi-Token Enhancing for Vision Representation Learning Enhancing representa- tions through heterogeneous self-supervised learning
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 7de05d14-dcb8-4fab-be4f-7cea56036150 · outbound
Multi-Token Enhancing for Vision Representation Learning Microsoft coco: Common objects in context
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 1e928978-8500-4391-873d-2a7a81ef28c1 · outbound
Multi-Token Enhancing for Vision Representation Learning Swin trans- former: Hierarchical vision transformer using shifted win- dows
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation cd04de15-3d21-4fa2-bd93-99c50f2f8caa · outbound
Multi-Token Enhancing for Vision Representation Learning A convnet for the 2020s
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation fe97687c-bdf2-4f79-a15a-b052e55841bd · outbound
Multi-Token Enhancing for Vision Representation Learning Decoupled weight decay regularization
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 81876898-bc47-4c23-9244-b503f4ac24ce · outbound
Multi-Token Enhancing for Vision Representation Learning Representation uncertainty in self-supervised learning as variational inference
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 474571af-d00b-4c85-9f9e-7093000d1393 · outbound
Multi-Token Enhancing for Vision Representation Learning Simreg: Regression as a sim- ple yet effective tool for self-supervised knowledge distilla- tion
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 58b883a7-a95f-485c-965d-8b9157c8bdb5 · outbound
Multi-Token Enhancing for Vision Representation Learning Representation Learning with Contrastive Predictive Coding
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e92f2de2-92a5-4f0b-90d8-4b6528726863 · outbound
Multi-Token Enhancing for Vision Representation Learning Imagenet large scale visual recognition challenge
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30511b61-22e1-4182-bfb2-190f132ed2ca · outbound
Multi-Token Enhancing for Vision Representation Learning Learning common rationale to improve self-supervised rep- resentation for fine-grained visual recognition problems
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 6c11bba4-cc9b-4d59-b780-d51ed92533c1 · outbound
Multi-Token Enhancing for Vision Representation Learning Unresolved cited work
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 1397c6fb-8ffa-482b-9707-6664ac7e44ae · outbound
Multi-Token Enhancing for Vision Representation Learning Multi- mode online knowledge distillation for self-supervised visual representation learning
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 39d4be1e-013c-4f9a-ac70-79b84717ee06 · outbound
Multi-Token Enhancing for Vision Representation Learning Semantics-consistent feature search for self-supervised visual representation learning
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 01ba14a2-f08d-40cd-be01-5faafd0dab63 · outbound
Multi-Token Enhancing for Vision Representation Learning Dropout: A simple way to prevent neural networks from overfitting
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 0ffb2b9b-743a-455d-9305-745c44fb2bd1 · outbound
Multi-Token Enhancing for Vision Representation Learning Siamese image modeling for self-supervised vision represen- tation learning
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 5160aa41-7918-4047-808c-b87342537b04 · outbound
Multi-Token Enhancing for Vision Representation Learning Un- derstanding self-supervised learning dynamics without con- trastive pairs
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d4306c5c-5350-40be-a1a1-e9d2f647d87e · outbound
Multi-Token Enhancing for Vision Representation Learning The inaturalist species classification and de- tection dataset
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation fd3cc78a-db4f-4275-9230-f92fded43f8e · outbound
Multi-Token Enhancing for Vision Representation Learning Pyramid vision transformer: A versatile backbone for dense prediction without convolutions
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c21843a7-1cc4-4fb6-b6d0-a03cbfb301fd · outbound
Multi-Token Enhancing for Vision Representation Learning Dense contrastive learning for self-supervised visual pre-training
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 0b6c0f23-d613-4c04-a131-7ab4fed5a8bf · outbound
Multi-Token Enhancing for Vision Representation Learning Masked Feature Prediction for Self-Supervised Visual Pre-Training
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f63cfc6-8ea1-4a29-8c9c-fbd761f90e77 · outbound
Multi-Token Enhancing for Vision Representation Learning Batchensemble: an alternative approach to efficient ensemble and lifelong learning
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation fd8cabfb-f044-4e51-ae64-d8cf6949b0e4 · outbound
Multi-Token Enhancing for Vision Representation Learning Con- vnext v2: Co-designing and scaling convnets with masked autoencoders
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 3f8ca473-1b5c-482f-ad48-83bae95b1767 · outbound
Multi-Token Enhancing for Vision Representation Learning Cvt: Introducing con- volutions to vision transformers
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d1933b00-fa25-46b6-9444-27b306a7c594 · outbound
Multi-Token Enhancing for Vision Representation Learning P2T: Pyramid pooling transformer for scene understanding
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 85ef4907-c63c-486a-81d0-55e8f7d0a51e · outbound
Multi-Token Enhancing for Vision Representation Learning Unified perceptual parsing for scene understand- ing
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b3a14ab8-cde0-4d9a-b440-18f5bb7d8396 · outbound
Multi-Token Enhancing for Vision Representation Learning Detco: Unsu- pervised contrastive learning for object detection
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation e9c44f02-2f1e-4160-ba2c-222945df8fcb · outbound
Multi-Token Enhancing for Vision Representation Learning Self-Supervised Learning with Swin Transformers
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 586afe26-a476-4a99-b125-432adefc9de2 · outbound
Multi-Token Enhancing for Vision Representation Learning Propagate yourself: Exploring pixel-level consistency for unsupervised visual representation learning
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 2262c200-fa96-41d5-9fda-4b9c8c84af89 · outbound
Multi-Token Enhancing for Vision Representation Learning Simmim: A simple framework for masked image modeling
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6aab7386-a01d-4f1e-a02d-c4deaeeaa838 · outbound
Multi-Token Enhancing for Vision Representation Learning Bag of instances aggregation boosts self-supervised distillation
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d73d5289-4ec1-4238-b808-479c9960f467 · outbound
Multi-Token Enhancing for Vision Representation Learning Joint unsuper- vised learning of deep representations and image clusters
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b611e978-ca13-4602-a3a5-b61a99a31d0e · outbound
Multi-Token Enhancing for Vision Representation Learning Decoupled contrastive learning
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation da304176-fd51-44a5-8855-0610b6db12f6 · outbound
Multi-Token Enhancing for Vision Representation Learning Online deep clustering for unsupervised representation learning
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d104ba81-a887-4d37-9d54-85f9fd059ddd · outbound
Multi-Token Enhancing for Vision Representation Learning Scene parsing through ade20k dataset
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 59d1a300-238d-4504-bd43-6c4df54e38fb · outbound
Multi-Token Enhancing for Vision Representation Learning ibot: Image bert pre-training with online tokenizer
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 3209ecb5-4892-4949-a2d5-7eb2ad445a7b · outbound
Multi-Token Enhancing for Vision Representation Learning Mugs: A Multi-Granular Self-Supervised Learning Framework
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1d707bd-7ebc-4ab5-baa6-8d2b9186b074 · outbound
Multi-Token Enhancing for Vision Representation Learning Multi-label self- supervised learning with scene images
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
No inbound Pith citation observations are available.