Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T23:19:02.251395Z
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 0 inbound Pith citation observations for arXiv:2412.20682.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T23:19:02.251395Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
60 of 60 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c6080f1f-968c-4258-abc4-92ad33c29ca7 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Learning transferable visual models from natural language supervision,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 24a9c0f8-f2fe-4376-a37e-59c0a3822ab3 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Scaling up visual and vision-language representation learning with noisy text supervision,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 1496bf16-a508-4a7d-9e0c-b9a0d8b127c3 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Sigmoid loss for language image pre-training,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 1d7dc89c-8759-4ea7-bd9f-7f6e10e24a54 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Sgva-clip: Semantic- guided visual adapting of vision-language models for few-shot image classification,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f125e2cf-d345-4ff8-8bf6-1d65bf841a07 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Clip-vg: Self-paced curriculum adapting of clip for visual grounding,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 4f0ff8ac-de93-4d10-bae8-4f677414e50c · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Effective end-to-end vision language pre- training with semantic visual loss,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 39e2d1f9-77ae-45d4-8827-e40e13faa139 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Neural logic vision language explainer,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 45f4ac16-7638-4079-96fe-a911d77bd1f4 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Lovm: Language- only vision model selection,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ed98bd14-cf8f-4e55-a8c6-7b2f33da4b48 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Bridge the Modality and Capability Gaps in Vision-Language Model Selection
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9663b70-939f-42e3-b691-b47a749e5d47 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Imagenet large scale visual recognition challenge,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59c80df0-e92e-4a65-bf70-fb67a9d8e0d8 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks GPT-4 Technical Report
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1291799-f858-46f5-b114-950b4b26af82 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Leveraging unlabeled data to predict out-of-distribution performance,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation a13c2413-be53-44db-b299-acae43f47066 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Are labels always necessary for classifier accuracy evaluation?
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 327ab1b6-d597-49f8-b18e-7b59a2ddeee1 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Predicting out-of- distribution error with the projection norm,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c3384c91-811e-458a-8cee-ce8ae80128a6 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Data determines distributional robustness in contrastive language image pre-training (clip),
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 18584f6a-cff5-4044-b263-60498db4077c · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Does clip’s generalization performance mainly stem from high train- test similarity?
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 3bde1082-2188-4a4f-a509-69890d4088ec · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks A Survey on Evaluation of Out-of-Distribution Generalization
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23247191-a828-4b72-9d61-913dc3761483 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Which Model to Transfer? A Survey on Transferability Estimation
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67e21d47-42b9-47c7-841a-68ede3983428 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Rankme: Assessing the downstream performance of pretrained self-supervised representations by their rank,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 1a5b9b43-afd2-4f53-83e0-2523ef975313 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Identifying useful learnwares for heterogeneous label spaces,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation bac8c173-40be-4e03-930c-158b51cd1e34 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Etran: Energy-based transferability estimation,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 27d233fe-7d78-4b5d-a545-499a6c8ed392 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Predicting out-of-distribution error with confidence optimal transport,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e3096a9c-fbdd-48cf-9bb4-7fa431d6cd33 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Data analysis and regression. a second course in statistics,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 07e1f4b7-5bf3-4985-9975-d8648163fea9 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Tune it the right way: Unsupervised validation of domain adaptation via soft neighborhood density,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 50eff583-92e6-4574-bb0a-40ebbff8ca86 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Covariate shift adap- tation by importance weighted cross validation
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 591eb3d8-46f2-42c8-9f10-df2da74637e7 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Towards accurate model selection in deep unsupervised domain adaptation,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 7b35aab9-69dc-4412-a0f5-3156d5e7e5c2 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Stochastic gradient methods for dis- tributionally robust optimization with f-divergences,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation add8cb75-5990-468f-ae31-0288f3206bb9 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Invariant Risk Minimization
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c547453-dc72-408a-92b8-62b718f6e057 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Stable learning via sample reweighting,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation a61de187-c254-4af9-8795-bcda9ebcb386 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks A baseline for detecting misclassified and out-of-distribution examples in neural networks,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation b578e508-0b05-421a-8fdb-39e58819caf8 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks What does rotation prediction tell us about classifier accuracy under varying testing environments?
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 42141123-3547-4e66-940d-9e732916013f · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Agreement-on-the- line: Predicting the performance of neural networks under distribution shift,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 69d6d889-6245-47ff-84ed-d09684bf3972 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d0258e10-b13b-47bd-aa46-36d5a80fa574 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks I. mathematical contributions to the theory of evolu- tion.—vii. on the correlation of characters not quantitatively measur- able,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 3a956975-6cb9-4cde-a6d6-06c8787a739b · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Learning multiple layers of features from tiny images,
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7101e1b9-9972-4dd6-9fd8-8c9d9db1b391 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Cats and dogs,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 2c700529-10b4-4c1e-8269-d91c1adf3467 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Automated flower classification over a large number of classes,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ba33f10b-0d1f-43da-aa42-52ceab64c3e8 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Reading digits in natural images with unsupervised feature learning,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation fc7e523e-c7f9-45ea-ac45-d9a27606758f · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Detection of traffic signs in real-world images: The german traffic sign detection benchmark,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 7ca0654a-f47a-4828-81f5-6b5bfc88c71a · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Describing textures in the wild,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e08b47f2-ef10-484a-b07f-6fc1400b7c45 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Yfcc100m: The new data in multimedia research,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 0c4880aa-d471-4885-a240-66142a8a7251 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Sun database: Large-scale scene recognition from abbey to zoo,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 2ad95db0-05d0-42b8-b66c-0ea03d373dca · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Gradient-based learning applied to document recognition,
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation a387ba64-2dae-4eb9-a1ef-9d21e78f5e04 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Challenges in representation learning: Facial expression recognition challenge,
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 45258e42-4956-4e3a-aa1a-87627e74055d · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks On the importance of feature separability in predicting out-of-distribution error,
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 315a85bc-873d-463f-97c0-b37bf5325664 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Unsupervised representation learning by predicting image rotations,
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation a88f3d46-786e-4e23-acdb-27219d5b6f06 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks The use of multiple measurements in taxonomic prob- lems,
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 218e96b0-d58b-4d22-9b7a-d5f41067cfa0 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Silhouettes: a graphical aid to the interpretation and validation of cluster analysis,
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c39edd1e-2beb-494c-b54b-d9f797f5ae56 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Deep residual learning for image recognition,
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7995dc37-9bf7-441f-b9fc-c302ac82e98e · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks An image is worth 16x16 words: Transformers for image recognition at scale,
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 119a2684-2c0c-439d-9f05-922bad6cc5cc · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks A convnet for the 2020s,
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 675996f6-57fd-4a99-ae23-b26f0b48b90a · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Laion-400m: Open dataset of clip-filtered 400 million image-text pairs,
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 82cc4b91-9e7e-4ad3-8f3c-c26e6c5d8102 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks AltCLIP: Altering the language encoder in CLIP for extended language capabilities,
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d7237b80-b802-4761-859c-448fcacc02af · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Groupvit: Semantic segmentation emerges from text supervision,
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation a3b48329-45f1-4ab2-856a-bec9f04e4f0e · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Learning Generalized Zero-Shot Learners for Open-Domain Image Geolocalization
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b63e8fcc-ea48-4330-8d0f-23b2ae3a5fef · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Demystifying CLIP Data
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8829e10d-c0a3-4c87-a026-2ede33fedede · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60cb4d52-1c1d-4b93-8581-ff0c846425f1 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Quilt-1M: One Million Image-Text Pairs for Histopathology
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e2650a0-8233-468f-8cc3-09fc29526389 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks BioCLIP: A vision foundation model for the tree of life,
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation cb36a3e1-a99f-4dcc-8527-2511724e5102 · outbound
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Gpt-4: Generative pre-trained transformer,
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
No inbound Pith citation observations are available.