Pith. sign in

Paper Citation Record · LEDGER

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks

As of 12 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 1 inbound Pith citation observation for arXiv:2412.20682.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.20682 v1

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T23:19:02.251395Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T00:32:20.408599Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-12T00:32:21.103085Z

Reference resolution

60 of 60 outbound references displayed

  • verified exact0
  • verified fuzzy47
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c6080f1f-968c-4258-abc4-92ad33c29ca7 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Learning transferable visual models from natural language supervision,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.853203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.062335Z digest=sha256:d089de60b2c374902eb5892e7b36e0b2a4794c955cadd14948d3ade90b493cbe

Observation 24a9c0f8-f2fe-4376-a37e-59c0a3822ab3 · outbound

This paper cites Scaling up visual and vision-language representation learning with noisy text supervision,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Scaling up visual and vision-language representation learning with noisy text supervision,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.844441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.067085Z digest=sha256:8d588ee9cb5b3df46c3c6ea94aca3d62f3e8a92f408108e0dd1620e6841b3569

Observation 1496bf16-a508-4a7d-9e0c-b9a0d8b127c3 · outbound

This paper cites Sigmoid loss for language image pre-training,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Sigmoid loss for language image pre-training,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.836193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.070476Z digest=sha256:c859afd343b1a02bd7ad199daa8ead6e8fc02ff0e9ef848548fc74a8094c3305

Observation 1d7dc89c-8759-4ea7-bd9f-7f6e10e24a54 · outbound

This paper cites Sgva-clip: Semantic- guided visual adapting of vision-language models for few-shot image classification,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Sgva-clip: Semantic- guided visual adapting of vision-language models for few-shot image classification,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.827397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.074512Z digest=sha256:25c2e94b79074282c63e68ca92e54bf9f0c774ead36fd47cc419c8e49cfce9e0

Observation f125e2cf-d345-4ff8-8bf6-1d65bf841a07 · outbound

This paper cites Clip-vg: Self-paced curriculum adapting of clip for visual grounding,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Clip-vg: Self-paced curriculum adapting of clip for visual grounding,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.818810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.078024Z digest=sha256:97a06656d2a7c72607d26de2a4ef900f1698643463ff2589be3a6676a0c4abb1

Observation 4f0ff8ac-de93-4d10-bae8-4f677414e50c · outbound

This paper cites Effective end-to-end vision language pre- training with semantic visual loss,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Effective end-to-end vision language pre- training with semantic visual loss,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.809768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.081342Z digest=sha256:b4f5267bb4b935bb0cc562b3b69ba8a6435b97fc8dc32aa5cbc6e57dc4a9c926

Observation 39e2d1f9-77ae-45d4-8827-e40e13faa139 · outbound

This paper cites Neural logic vision language explainer,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Neural logic vision language explainer,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.799422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.085792Z digest=sha256:327f30728e27d85d0b4f8e277c85f17ac1caf82476d983ad70a16054e0324de6

Observation 45f4ac16-7638-4079-96fe-a911d77bd1f4 · outbound

This paper cites Lovm: Language- only vision model selection,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Lovm: Language- only vision model selection,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.788305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.088657Z digest=sha256:09665b71c5465958925d6566b53d3d2a7dadf05767da3c2cd5031a244c984bd5

Observation ed98bd14-cf8f-4e55-a8c6-7b2f33da4b48 · outbound

This paper cites Bridge the Modality and Capability Gaps in Vision-Language Model Selection.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Bridge the Modality and Capability Gaps in Vision-Language Model Selection

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.092268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.092268Z digest=sha256:b9494dd0bd687050382c01a8c0a3a4b06fbf7a63517997ec176d99edc1831c6a

Observation a9663b70-939f-42e3-b691-b47a749e5d47 · outbound

This paper cites Imagenet large scale visual recognition challenge,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Imagenet large scale visual recognition challenge,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.096576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.096576Z digest=sha256:7854b0530c8966489deb48836bb4812bc2ba9a962876d081655dee1a207d8469

Observation 59c80df0-e92e-4a65-bf70-fb67a9d8e0d8 · outbound

This paper cites GPT-4 Technical Report.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks GPT-4 Technical Report

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.099108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.099108Z digest=sha256:293def4f101ae1e4b93be40e99bf345f0d4e128c76a076e380e65ef37335e9c8

Observation a1291799-f858-46f5-b114-950b4b26af82 · outbound

This paper cites Leveraging unlabeled data to predict out-of-distribution performance,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Leveraging unlabeled data to predict out-of-distribution performance,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.776644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.102861Z digest=sha256:6bcc3a076d631b9a3e0ed68af5d1fe42b986937e265803a5626bc8ce15f8aa4f

Observation a13c2413-be53-44db-b299-acae43f47066 · outbound

This paper cites Are labels always necessary for classifier accuracy evaluation?.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Are labels always necessary for classifier accuracy evaluation?

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.769117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.106942Z digest=sha256:2250247df243815e63ca5b2e5b0cd73a7086e5e07a0136efd6d858c50f66817d

Observation 327ab1b6-d597-49f8-b18e-7b59a2ddeee1 · outbound

This paper cites Predicting out-of- distribution error with the projection norm,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Predicting out-of- distribution error with the projection norm,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.760868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.109667Z digest=sha256:c75146ccf4805498f51da14a82e0ada9abf06d8132f4d1f7b444b65198eb6163

Observation c3384c91-811e-458a-8cee-ce8ae80128a6 · outbound

This paper cites Data determines distributional robustness in contrastive language image pre-training (clip),.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Data determines distributional robustness in contrastive language image pre-training (clip),

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.752385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.112361Z digest=sha256:be63443bf6cfa6b98db9503451e242d2b433344ada119fc218aba34850a0e560

Observation 18584f6a-cff5-4044-b263-60498db4077c · outbound

This paper cites Does clip’s generalization performance mainly stem from high train- test similarity?.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Does clip’s generalization performance mainly stem from high train- test similarity?

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.744442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.115218Z digest=sha256:7dcf1f66e91127bbcf9b447ce2c5a43019d36c075b4a251e40a9e7ee80602bce

Observation 3bde1082-2188-4a4f-a509-69890d4088ec · outbound

This paper cites A Survey on Evaluation of Out-of-Distribution Generalization.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks A Survey on Evaluation of Out-of-Distribution Generalization

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.117741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.117741Z digest=sha256:d113c6f2652ceff2050abb289b8b207effe2c6a2c9b7495fa028df4db9b80124

Observation 23247191-a828-4b72-9d61-913dc3761483 · outbound

This paper cites Which Model to Transfer? A Survey on Transferability Estimation.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Which Model to Transfer? A Survey on Transferability Estimation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.120616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.120616Z digest=sha256:faf1f88a4289617c6711025ffd544e9f2a61eaf634a7e1c11fcf29248f3cbb7a

Observation 67e21d47-42b9-47c7-841a-68ede3983428 · outbound

This paper cites Rankme: Assessing the downstream performance of pretrained self-supervised representations by their rank,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Rankme: Assessing the downstream performance of pretrained self-supervised representations by their rank,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.736740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.123697Z digest=sha256:511a31ec15485bbdd864451ab90a554336baa7653e8cb974a83fc2b662d7416f

Observation 1a5b9b43-afd2-4f53-83e0-2523ef975313 · outbound

This paper cites Identifying useful learnwares for heterogeneous label spaces,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Identifying useful learnwares for heterogeneous label spaces,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.727669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.126635Z digest=sha256:0b57e272b64df089ed9699e1248a98ee8406d2f6cd64b3c93abce16bef3a31d3

Observation bac8c173-40be-4e03-930c-158b51cd1e34 · outbound

This paper cites Etran: Energy-based transferability estimation,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Etran: Energy-based transferability estimation,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.720403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.129072Z digest=sha256:e0daee8725f89be598103a3295a87a824294a5ac48de9ac6eba41d80fac251d2

Observation 27d233fe-7d78-4b5d-a545-499a6c8ed392 · outbound

This paper cites Predicting out-of-distribution error with confidence optimal transport,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Predicting out-of-distribution error with confidence optimal transport,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.713537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.132874Z digest=sha256:28534c2a73b2c357e885be2e0f8d3ced27a7c9108e03deb9240eed9e15bd980a

Observation e3096a9c-fbdd-48cf-9bb4-7fa431d6cd33 · outbound

This paper cites Data analysis and regression. a second course in statistics,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Data analysis and regression. a second course in statistics,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.706480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.136115Z digest=sha256:7a0211a201beadef5214cc7a8fc9a69519e413fdb6bd3e642b4f60cf204df3f4

Observation 07e1f4b7-5bf3-4985-9975-d8648163fea9 · outbound

This paper cites Tune it the right way: Unsupervised validation of domain adaptation via soft neighborhood density,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Tune it the right way: Unsupervised validation of domain adaptation via soft neighborhood density,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.699122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.138546Z digest=sha256:b18865f89c696f2ab6519e622b1f5a3f8e67c2adce78e780291d7ecf06d17249

Observation 50eff583-92e6-4574-bb0a-40ebbff8ca86 · outbound

This paper cites Covariate shift adap- tation by importance weighted cross validation.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Covariate shift adap- tation by importance weighted cross validation

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.689740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.142516Z digest=sha256:47ca7a611657f7b5c73b7bee812503113cf4036d1d884f6f5e90bfc606bda8f8

Observation 591eb3d8-46f2-42c8-9f10-df2da74637e7 · outbound

This paper cites Towards accurate model selection in deep unsupervised domain adaptation,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Towards accurate model selection in deep unsupervised domain adaptation,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.680320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.146091Z digest=sha256:6709532c8f3e10cfa97875d8a9e9f687ceb00b2ec214e4a316e151a966454344

Observation 7b35aab9-69dc-4412-a0f5-3156d5e7e5c2 · outbound

This paper cites Stochastic gradient methods for dis- tributionally robust optimization with f-divergences,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Stochastic gradient methods for dis- tributionally robust optimization with f-divergences,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.672181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.148593Z digest=sha256:2087af9ed707f1c381089123f65da875e82c82b45e49db2c8885c874fd33a89b

Observation add8cb75-5990-468f-ae31-0288f3206bb9 · outbound

This paper cites Invariant Risk Minimization.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Invariant Risk Minimization

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.151663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.151663Z digest=sha256:c4e5e0b08128e0d33e31d341161c756351429f43135c0aaabf741483444653be

Observation 4c547453-dc72-408a-92b8-62b718f6e057 · outbound

This paper cites Stable learning via sample reweighting,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Stable learning via sample reweighting,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.664007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.155701Z digest=sha256:ea0c2e00d03a03d3a3374135d05ee102487568cc68a702685787687622092150

Observation a61de187-c254-4af9-8795-bcda9ebcb386 · outbound

This paper cites A baseline for detecting misclassified and out-of-distribution examples in neural networks,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks A baseline for detecting misclassified and out-of-distribution examples in neural networks,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.655597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.159008Z digest=sha256:61aef528a5cfed5be000c5b8c26c16ed5f86c6b188d63078d848632df9065287

Observation b578e508-0b05-421a-8fdb-39e58819caf8 · outbound

This paper cites What does rotation prediction tell us about classifier accuracy under varying testing environments?.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks What does rotation prediction tell us about classifier accuracy under varying testing environments?

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.647830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.161587Z digest=sha256:742305911de3850e8dd6e6c76e04e5ed041cfdabefbcac0d9813419fc8a32d55

Observation 42141123-3547-4e66-940d-9e732916013f · outbound

This paper cites Agreement-on-the- line: Predicting the performance of neural networks under distribution shift,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Agreement-on-the- line: Predicting the performance of neural networks under distribution shift,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.637822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.164953Z digest=sha256:5dec0475c239b8fbb97aadae773b0eef741ecf78b5c1c71de00934480fae038f

Observation 69d6d889-6245-47ff-84ed-d09684bf3972 · outbound

This paper cites Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.630112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.168153Z digest=sha256:4c01043bf0fc7d631690b6061395c5ab99e7a36126b73b228202ca8f7d9cdece

Observation d0258e10-b13b-47bd-aa46-36d5a80fa574 · outbound

This paper cites I. mathematical contributions to the theory of evolu- tion.—vii. on the correlation of characters not quantitatively measur- able,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks I. mathematical contributions to the theory of evolu- tion.—vii. on the correlation of characters not quantitatively measur- able,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.621254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.171155Z digest=sha256:88d6ad51a3fd047a949223d38513e7846c3f004f6bcafc0d4755a167b610a153

Observation 3a956975-6cb9-4cde-a6d6-06c8787a739b · outbound

This paper cites Learning multiple layers of features from tiny images,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Learning multiple layers of features from tiny images,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.173527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.173527Z digest=sha256:f7782ab5d1575426b2b77f080365150e6c527cc10707cb16dc97512b4819e6e5

Observation 7101e1b9-9972-4dd6-9fd8-8c9d9db1b391 · outbound

This paper cites Cats and dogs,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Cats and dogs,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.610138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.176231Z digest=sha256:a0afd15dec82a5a0a67aec7259f934dc25386bb25acc7fb5c1f13d3c98e016dc

Observation 2c700529-10b4-4c1e-8269-d91c1adf3467 · outbound

This paper cites Automated flower classification over a large number of classes,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Automated flower classification over a large number of classes,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.603200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.178912Z digest=sha256:696edcc8ec09734a5499eb362c79e24cf8cb649431698d5c5e8ec6863e1c9a95

Observation ba33f10b-0d1f-43da-aa42-52ceab64c3e8 · outbound

This paper cites Reading digits in natural images with unsupervised feature learning,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Reading digits in natural images with unsupervised feature learning,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.595212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.181346Z digest=sha256:65708ddc765a3516be6a1019550f9f85811720cbd012234daebad24d9398695a

Observation fc7e523e-c7f9-45ea-ac45-d9a27606758f · outbound

This paper cites Detection of traffic signs in real-world images: The german traffic sign detection benchmark,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Detection of traffic signs in real-world images: The german traffic sign detection benchmark,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.585413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.185130Z digest=sha256:0749142af127c127441611b2d8b58ebe85322beef2df21b824bbd4551d68b276

Observation 7ca0654a-f47a-4828-81f5-6b5bfc88c71a · outbound

This paper cites Describing textures in the wild,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Describing textures in the wild,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.576201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.189062Z digest=sha256:5c35f7bb509167108c86c921f6a8178d7455943592c5fa1dd494dcffd6398dfd

Observation e08b47f2-ef10-484a-b07f-6fc1400b7c45 · outbound

This paper cites Yfcc100m: The new data in multimedia research,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Yfcc100m: The new data in multimedia research,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.566206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.192584Z digest=sha256:2686ff9a25e6d3461edcb236c108105270ebd716fdf59ad07e3f9296472a1a9c

Observation 0c4880aa-d471-4885-a240-66142a8a7251 · outbound

This paper cites Sun database: Large-scale scene recognition from abbey to zoo,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Sun database: Large-scale scene recognition from abbey to zoo,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.554864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.198741Z digest=sha256:14e5eec9937c6a3e511eeb7f5bdd45be32dbf5b98889f0cc51f657a6570a9c78

Observation 2ad95db0-05d0-42b8-b66c-0ea03d373dca · outbound

This paper cites Gradient-based learning applied to document recognition,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Gradient-based learning applied to document recognition,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.544234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.201847Z digest=sha256:f5df78e54afc4e14d6588c1b67645960dfec3caa17559eca0a071bf01d440438

Observation a387ba64-2dae-4eb9-a1ef-9d21e78f5e04 · outbound

This paper cites Challenges in representation learning: Facial expression recognition challenge,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Challenges in representation learning: Facial expression recognition challenge,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.535101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.205503Z digest=sha256:c8e794c9e358491dcc82c0e3c0314158c85d33e97fb18ba11112ae2dd5e78cf3

Observation 45258e42-4956-4e3a-aa1a-87627e74055d · outbound

This paper cites On the importance of feature separability in predicting out-of-distribution error,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks On the importance of feature separability in predicting out-of-distribution error,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.527167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.209061Z digest=sha256:89f9eeb56f837e1d929152286671c63645dba4db7c9f89446c6a012ac6a8d059

Observation 315a85bc-873d-463f-97c0-b37bf5325664 · outbound

This paper cites Unsupervised representation learning by predicting image rotations,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Unsupervised representation learning by predicting image rotations,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.517108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.212042Z digest=sha256:aa13f522f1428c152ea6d795e0938dfd6a1eca35440d3977cbfe1a754503ea74

Observation a88f3d46-786e-4e23-acdb-27219d5b6f06 · outbound

This paper cites The use of multiple measurements in taxonomic prob- lems,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks The use of multiple measurements in taxonomic prob- lems,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.508560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.214634Z digest=sha256:2211cc9c92fc6de014e498392d1300b2310ef28375d3c22a0afd2f5d2eae702e

Observation 218e96b0-d58b-4d22-9b7a-d5f41067cfa0 · outbound

This paper cites Silhouettes: a graphical aid to the interpretation and validation of cluster analysis,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Silhouettes: a graphical aid to the interpretation and validation of cluster analysis,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.217025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.217025Z digest=sha256:3ea670d2840debdc249019482286af305e07236d58d711fbc329ed194c008ff9

Observation c39edd1e-2beb-494c-b54b-d9f797f5ae56 · outbound

This paper cites Deep residual learning for image recognition,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Deep residual learning for image recognition,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.219537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.219537Z digest=sha256:4bd20930178de1f0e3e1ced8a14af3da7ba7a9c7b7942035ad55723424b08298

Observation 7995dc37-9bf7-441f-b9fc-c302ac82e98e · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks An image is worth 16x16 words: Transformers for image recognition at scale,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.422639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.222327Z digest=sha256:d4f9bd731fb10ee26953bde028a962cff9aafc7ef9088e65940b163d80701e37

Observation 119a2684-2c0c-439d-9f05-922bad6cc5cc · outbound

This paper cites A convnet for the 2020s,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks A convnet for the 2020s,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.382172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.225312Z digest=sha256:4343d2417b0338eea570aba2ea3806f74a10193ed58a65964cc69e5fd59ef2d5

Observation 675996f6-57fd-4a99-ae23-b26f0b48b90a · outbound

This paper cites Laion-400m: Open dataset of clip-filtered 400 million image-text pairs,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Laion-400m: Open dataset of clip-filtered 400 million image-text pairs,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.372675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.227883Z digest=sha256:03033009487002520067bfc0e15a791f3168ddd2574b719f35cdafabbb891d51

Observation 82cc4b91-9e7e-4ad3-8f3c-c26e6c5d8102 · outbound

This paper cites AltCLIP: Altering the language encoder in CLIP for extended language capabilities,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks AltCLIP: Altering the language encoder in CLIP for extended language capabilities,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.362356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.230309Z digest=sha256:d24bc29de380c8f7849f4d9e73648e16e69b941a66f76f3fc56d52656201e6ba

Observation d7237b80-b802-4761-859c-448fcacc02af · outbound

This paper cites Groupvit: Semantic segmentation emerges from text supervision,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Groupvit: Semantic segmentation emerges from text supervision,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.350740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.232910Z digest=sha256:dfa2e55af81d6541943184596772dd051226d8e14fd648dda66bb139c8bb8108

Observation a3b48329-45f1-4ab2-856a-bec9f04e4f0e · outbound

This paper cites Learning Generalized Zero-Shot Learners for Open-Domain Image Geolocalization.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Learning Generalized Zero-Shot Learners for Open-Domain Image Geolocalization

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.235754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.235754Z digest=sha256:0b4cb9abbd834866db4b06766d28eb50548f4d9dd32f09705b1d4ecf01f39f4e

Observation b63e8fcc-ea48-4330-8d0f-23b2ae3a5fef · outbound

This paper cites Demystifying CLIP Data.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Demystifying CLIP Data

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.239160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.239160Z digest=sha256:e6e1fba744db4f6c936d8c0091587e71183ee8e2ca912be728f40600f61f4d26

Observation 8829e10d-c0a3-4c87-a026-2ede33fedede · outbound

This paper cites BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.242350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.242350Z digest=sha256:a13a0f80d8b911509a663e384cff333174a90e87e20f824f2fb0590a38e3fc13

Observation 60cb4d52-1c1d-4b93-8581-ff0c846425f1 · outbound

This paper cites Quilt-1M: One Million Image-Text Pairs for Histopathology.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Quilt-1M: One Million Image-Text Pairs for Histopathology

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.245892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.245892Z digest=sha256:fb5c9bd191ea39933b22f77e68b361398e003088e78d90104b672a3a161fa442

Observation 8e2650a0-8233-468f-8cc3-09fc29526389 · outbound

This paper cites BioCLIP: A vision foundation model for the tree of life,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks BioCLIP: A vision foundation model for the tree of life,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.341422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.248744Z digest=sha256:4a9e848e4795804199989258cd46871e3de3245a16d532abfa09682de767306e

Observation cb36a3e1-a99f-4dcc-8527-2511724e5102 · outbound

This paper cites Gpt-4: Generative pre-trained transformer,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Gpt-4: Generative pre-trained transformer,

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.331857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T23:19:02.251395Z digest=sha256:ea46a4d402a695a6c4f904271cc7c5b882bfff0cb9c683ebdbb924384459741c

Pith citing papers

Observation 12f4859a-834e-4817-a2cf-bfb3385c5b4e · inbound

HugSelect: An Explainable Multi-Criteria Decision-Support Framework for foundation-model selection cites this paper.

HugSelect: An Explainable Multi-Criteria Decision-Support Framework for foundation-model selection Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks

Reference 8

Resolution
metadata mismatch
local_arxiv, observed 2026-08-12T00:32:21.110763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T00:32:20.408599Z digest=sha256:c1ba9cee422618fa45121763b7cec0b08c75f30d59e896e162cfcc2b3dfc800b