Pith. sign in

Paper Citation Record · LEDGER

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks

As of 11 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 0 inbound Pith citation observations for arXiv:2412.20682.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.20682 v1

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T23:19:02.251395Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

60 of 60 outbound references displayed

  • verified exact0
  • verified fuzzy47
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c6080f1f-968c-4258-abc4-92ad33c29ca7 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Learning transferable visual models from natural language supervision,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.853203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.062335Z digest=sha256:c2c76c05392779ed57086c2860c0341e5532a577a8a0ff0b66566c3fa6087334

Observation 24a9c0f8-f2fe-4376-a37e-59c0a3822ab3 · outbound

This paper cites Scaling up visual and vision-language representation learning with noisy text supervision,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Scaling up visual and vision-language representation learning with noisy text supervision,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.844441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.067085Z digest=sha256:da6044d178806488fb071320056cf698fd75205a388b009e9d8b66da6809b621

Observation 1496bf16-a508-4a7d-9e0c-b9a0d8b127c3 · outbound

This paper cites Sigmoid loss for language image pre-training,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Sigmoid loss for language image pre-training,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.836193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.070476Z digest=sha256:5f9458d62c974e510516c137f8801fd5469c9afd1d57d14a41d5c75cce60c11e

Observation 1d7dc89c-8759-4ea7-bd9f-7f6e10e24a54 · outbound

This paper cites Sgva-clip: Semantic- guided visual adapting of vision-language models for few-shot image classification,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Sgva-clip: Semantic- guided visual adapting of vision-language models for few-shot image classification,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.827397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.074512Z digest=sha256:fc54580609e84bec12dd7aa79bda13bee4a2924e491cd26f208ede843cacd3c9

Observation f125e2cf-d345-4ff8-8bf6-1d65bf841a07 · outbound

This paper cites Clip-vg: Self-paced curriculum adapting of clip for visual grounding,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Clip-vg: Self-paced curriculum adapting of clip for visual grounding,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.818810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.078024Z digest=sha256:800c60242cf0798fa157829977f1d041f988acf21b7819055cd188da71ef9913

Observation 4f0ff8ac-de93-4d10-bae8-4f677414e50c · outbound

This paper cites Effective end-to-end vision language pre- training with semantic visual loss,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Effective end-to-end vision language pre- training with semantic visual loss,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.809768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.081342Z digest=sha256:0dad61e68fd0b9e2fde8bfc21af786a914299e801fefa38ecaf7855bd6233979

Observation 39e2d1f9-77ae-45d4-8827-e40e13faa139 · outbound

This paper cites Neural logic vision language explainer,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Neural logic vision language explainer,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.799422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.085792Z digest=sha256:6e55ba9d98a0a65875de7b87611896d9ff86def6cf3e8f3ef378a2d7bb5c970b

Observation 45f4ac16-7638-4079-96fe-a911d77bd1f4 · outbound

This paper cites Lovm: Language- only vision model selection,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Lovm: Language- only vision model selection,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.788305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.088657Z digest=sha256:e34e713543c4631c684a52dd5bdc119b43a2af29a9fa9ae398667b23e3caac88

Observation ed98bd14-cf8f-4e55-a8c6-7b2f33da4b48 · outbound

This paper cites Bridge the Modality and Capability Gaps in Vision-Language Model Selection.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Bridge the Modality and Capability Gaps in Vision-Language Model Selection

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.092268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.092268Z digest=sha256:993041d79047659200ff6366318b85c6a9b01efccb9a13b2e3e61c8658e53a2a

Observation a9663b70-939f-42e3-b691-b47a749e5d47 · outbound

This paper cites Imagenet large scale visual recognition challenge,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Imagenet large scale visual recognition challenge,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.096576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.096576Z digest=sha256:984e8ba576f583d6a89d29262a9fb84f61cf9b35412236e74bf382b76f2dd551

Observation 59c80df0-e92e-4a65-bf70-fb67a9d8e0d8 · outbound

This paper cites GPT-4 Technical Report.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks GPT-4 Technical Report

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.099108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.099108Z digest=sha256:4f9cac2389952b25bded41ca27f5c667893ead65df8afc34ab01aaae507ffbbd

Observation a1291799-f858-46f5-b114-950b4b26af82 · outbound

This paper cites Leveraging unlabeled data to predict out-of-distribution performance,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Leveraging unlabeled data to predict out-of-distribution performance,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.776644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.102861Z digest=sha256:104ab7b5c936175a9934854513ff99297a67d5b151bd77f6c32fb8205be52e78

Observation a13c2413-be53-44db-b299-acae43f47066 · outbound

This paper cites Are labels always necessary for classifier accuracy evaluation?.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Are labels always necessary for classifier accuracy evaluation?

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.769117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.106942Z digest=sha256:111c51d3d334b8df37e650389b94602153a95104f62a9239810b35f9f2b4b15c

Observation 327ab1b6-d597-49f8-b18e-7b59a2ddeee1 · outbound

This paper cites Predicting out-of- distribution error with the projection norm,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Predicting out-of- distribution error with the projection norm,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.760868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.109667Z digest=sha256:e36ba18fceccbb6b278d263a3aa83b49e31f89a3d4c7e93f8370a71d86404d12

Observation c3384c91-811e-458a-8cee-ce8ae80128a6 · outbound

This paper cites Data determines distributional robustness in contrastive language image pre-training (clip),.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Data determines distributional robustness in contrastive language image pre-training (clip),

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.752385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.112361Z digest=sha256:1e7a15753c1b06a7175d0c3c294c14cd115f5b5072746b1cdbd070c52a5228da

Observation 18584f6a-cff5-4044-b263-60498db4077c · outbound

This paper cites Does clip’s generalization performance mainly stem from high train- test similarity?.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Does clip’s generalization performance mainly stem from high train- test similarity?

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.744442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.115218Z digest=sha256:9ebd889b4c8f40c7698339d630a6ea3921ef2cc102304d53740f290f6200a757

Observation 3bde1082-2188-4a4f-a509-69890d4088ec · outbound

This paper cites A Survey on Evaluation of Out-of-Distribution Generalization.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks A Survey on Evaluation of Out-of-Distribution Generalization

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.117741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.117741Z digest=sha256:2a32b4a715902f99726e0fdc34b7376a460f22f00f7e77e2df9761be8a6b9212

Observation 23247191-a828-4b72-9d61-913dc3761483 · outbound

This paper cites Which Model to Transfer? A Survey on Transferability Estimation.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Which Model to Transfer? A Survey on Transferability Estimation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.120616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.120616Z digest=sha256:77e9829a07f1013e8c9dd43aa74206da77a6f0dc3b3e9dd40a726d3cf8fa44a4

Observation 67e21d47-42b9-47c7-841a-68ede3983428 · outbound

This paper cites Rankme: Assessing the downstream performance of pretrained self-supervised representations by their rank,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Rankme: Assessing the downstream performance of pretrained self-supervised representations by their rank,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.736740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.123697Z digest=sha256:66f391e46046bbab36fb5604d76e11c765914b899bd69ecaf42f72e82369745a

Observation 1a5b9b43-afd2-4f53-83e0-2523ef975313 · outbound

This paper cites Identifying useful learnwares for heterogeneous label spaces,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Identifying useful learnwares for heterogeneous label spaces,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.727669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.126635Z digest=sha256:a21e0d0f17ce49b1c286c975b5bcf276e7e8155e3f62e231716ab99a842dbc7b

Observation bac8c173-40be-4e03-930c-158b51cd1e34 · outbound

This paper cites Etran: Energy-based transferability estimation,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Etran: Energy-based transferability estimation,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.720403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.129072Z digest=sha256:5aaded529e2543bf58a90b5795b8833791fef3c4af22701fa6480118841464a1

Observation 27d233fe-7d78-4b5d-a545-499a6c8ed392 · outbound

This paper cites Predicting out-of-distribution error with confidence optimal transport,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Predicting out-of-distribution error with confidence optimal transport,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.713537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.132874Z digest=sha256:5c2ea7843d6201129ea7fd7636edac2fb1c4cdf28007742942d77b661fe1830f

Observation e3096a9c-fbdd-48cf-9bb4-7fa431d6cd33 · outbound

This paper cites Data analysis and regression. a second course in statistics,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Data analysis and regression. a second course in statistics,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.706480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.136115Z digest=sha256:79ee6ea42db7b4f94935682dbe2f30c44cf505f6d1342f713ed49ad6584e2924

Observation 07e1f4b7-5bf3-4985-9975-d8648163fea9 · outbound

This paper cites Tune it the right way: Unsupervised validation of domain adaptation via soft neighborhood density,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Tune it the right way: Unsupervised validation of domain adaptation via soft neighborhood density,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.699122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.138546Z digest=sha256:f5689cb4d5b81295296d271808ad24a1266f8267f733a355e2f5d37ace217858

Observation 50eff583-92e6-4574-bb0a-40ebbff8ca86 · outbound

This paper cites Covariate shift adap- tation by importance weighted cross validation.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Covariate shift adap- tation by importance weighted cross validation

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.689740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.142516Z digest=sha256:d46a46915ce80ede29b130521cc7cacd1af8ed1e60f1c4be9ea2979b36874861

Observation 591eb3d8-46f2-42c8-9f10-df2da74637e7 · outbound

This paper cites Towards accurate model selection in deep unsupervised domain adaptation,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Towards accurate model selection in deep unsupervised domain adaptation,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.680320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.146091Z digest=sha256:fc6531ad737f43a5298b854434c97eac230da33220b6ca1ec779c0681f09784d

Observation 7b35aab9-69dc-4412-a0f5-3156d5e7e5c2 · outbound

This paper cites Stochastic gradient methods for dis- tributionally robust optimization with f-divergences,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Stochastic gradient methods for dis- tributionally robust optimization with f-divergences,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.672181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.148593Z digest=sha256:0dac14887c3aa18d3c213d38d45d430ed72d010ac9ca6290ad508f82989cba7b

Observation add8cb75-5990-468f-ae31-0288f3206bb9 · outbound

This paper cites Invariant Risk Minimization.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Invariant Risk Minimization

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.151663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.151663Z digest=sha256:4276ecebd9782b232dd06ef1450555dc6c90e59a0169dbd6afa9c42cbe2c0785

Observation 4c547453-dc72-408a-92b8-62b718f6e057 · outbound

This paper cites Stable learning via sample reweighting,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Stable learning via sample reweighting,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.664007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.155701Z digest=sha256:8f7a76179684b1557c4116db79e9a65d958aa1704525a5cd60a451e03a3e5540

Observation a61de187-c254-4af9-8795-bcda9ebcb386 · outbound

This paper cites A baseline for detecting misclassified and out-of-distribution examples in neural networks,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks A baseline for detecting misclassified and out-of-distribution examples in neural networks,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.655597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.159008Z digest=sha256:6ee92d170b838b6fadddd4b7ef90e80744f7ae9e78be1fb87a519e50d3ef89f5

Observation b578e508-0b05-421a-8fdb-39e58819caf8 · outbound

This paper cites What does rotation prediction tell us about classifier accuracy under varying testing environments?.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks What does rotation prediction tell us about classifier accuracy under varying testing environments?

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.647830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.161587Z digest=sha256:9357b8054e69d3b47a8275cbe7d3177d495a2a691a3371adc4b6b3c18d041326

Observation 42141123-3547-4e66-940d-9e732916013f · outbound

This paper cites Agreement-on-the- line: Predicting the performance of neural networks under distribution shift,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Agreement-on-the- line: Predicting the performance of neural networks under distribution shift,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.637822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.164953Z digest=sha256:6abda5784b444b5761bb2ff0a7fd8c415ecc9b76f7c40c0e6fd6ce699fbdf406

Observation 69d6d889-6245-47ff-84ed-d09684bf3972 · outbound

This paper cites Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.630112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.168153Z digest=sha256:cffb2ce14586a6ae75caf81c3c86806155b2cd530731595443e94c9ab6136d7c

Observation d0258e10-b13b-47bd-aa46-36d5a80fa574 · outbound

This paper cites I. mathematical contributions to the theory of evolu- tion.—vii. on the correlation of characters not quantitatively measur- able,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks I. mathematical contributions to the theory of evolu- tion.—vii. on the correlation of characters not quantitatively measur- able,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.621254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.171155Z digest=sha256:80e179df63872c0810d93d4cac3c0bd0603ad52917658290c15c2da51d5d52c1

Observation 3a956975-6cb9-4cde-a6d6-06c8787a739b · outbound

This paper cites Learning multiple layers of features from tiny images,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Learning multiple layers of features from tiny images,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.173527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.173527Z digest=sha256:9b58110070f56c31e1c21c7ae68f7af4ed3a2928c3a5fe8174fa33e0e030e79c

Observation 7101e1b9-9972-4dd6-9fd8-8c9d9db1b391 · outbound

This paper cites Cats and dogs,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Cats and dogs,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.610138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.176231Z digest=sha256:5ff263ceb0ae08e1b2cdfcdb6e12bf7ac8b69c6df81f412e54a83da824174cc7

Observation 2c700529-10b4-4c1e-8269-d91c1adf3467 · outbound

This paper cites Automated flower classification over a large number of classes,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Automated flower classification over a large number of classes,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.603200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.178912Z digest=sha256:b79ec57478cd0d397e943c0c11f966312475a089aca75f898aa7220b572c9eb0

Observation ba33f10b-0d1f-43da-aa42-52ceab64c3e8 · outbound

This paper cites Reading digits in natural images with unsupervised feature learning,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Reading digits in natural images with unsupervised feature learning,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.595212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.181346Z digest=sha256:a6acd51392ccbe5e0ef4515e205fdbb84d3af969baa246de51b274b9bff82b4e

Observation fc7e523e-c7f9-45ea-ac45-d9a27606758f · outbound

This paper cites Detection of traffic signs in real-world images: The german traffic sign detection benchmark,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Detection of traffic signs in real-world images: The german traffic sign detection benchmark,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.585413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.185130Z digest=sha256:dda90f916f6ef7e35df0458e05faac3b502f8ca2976c0ccc0b3a358dd290a44d

Observation 7ca0654a-f47a-4828-81f5-6b5bfc88c71a · outbound

This paper cites Describing textures in the wild,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Describing textures in the wild,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.576201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.189062Z digest=sha256:d70119ffaffeb60527f425f1a6f75d6b9cfbec2b048d5b82b10d2a38022b5965

Observation e08b47f2-ef10-484a-b07f-6fc1400b7c45 · outbound

This paper cites Yfcc100m: The new data in multimedia research,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Yfcc100m: The new data in multimedia research,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.566206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.192584Z digest=sha256:19e52b3af6582e43f5d8b7f39e7301caef9548c9d91e39d36c47c71b5cbb5cb8

Observation 0c4880aa-d471-4885-a240-66142a8a7251 · outbound

This paper cites Sun database: Large-scale scene recognition from abbey to zoo,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Sun database: Large-scale scene recognition from abbey to zoo,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.554864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.198741Z digest=sha256:a825bc206a6e2718c5dcd773dbb1bbbfaea5cd13907d33911e9f2fd12ec54443

Observation 2ad95db0-05d0-42b8-b66c-0ea03d373dca · outbound

This paper cites Gradient-based learning applied to document recognition,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Gradient-based learning applied to document recognition,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.544234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.201847Z digest=sha256:2ddcd87cdc7b46db2f2c7553125d6cabd9cf26a8298acbb6c4b140c7d153c53c

Observation a387ba64-2dae-4eb9-a1ef-9d21e78f5e04 · outbound

This paper cites Challenges in representation learning: Facial expression recognition challenge,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Challenges in representation learning: Facial expression recognition challenge,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.535101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.205503Z digest=sha256:e48d6e2ad23e3b794749bf6b5a772de008345962bef0cd35bdb0cec08f9913b9

Observation 45258e42-4956-4e3a-aa1a-87627e74055d · outbound

This paper cites On the importance of feature separability in predicting out-of-distribution error,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks On the importance of feature separability in predicting out-of-distribution error,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.527167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.209061Z digest=sha256:d408e621e7269916629bdcb0ae6b6429cffa8127fbe5f3e9f266727f8327b93b

Observation 315a85bc-873d-463f-97c0-b37bf5325664 · outbound

This paper cites Unsupervised representation learning by predicting image rotations,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Unsupervised representation learning by predicting image rotations,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.517108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.212042Z digest=sha256:a52c7266be6932dc22142b1a462758633c088152ed635145824e95e2d6df5120

Observation a88f3d46-786e-4e23-acdb-27219d5b6f06 · outbound

This paper cites The use of multiple measurements in taxonomic prob- lems,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks The use of multiple measurements in taxonomic prob- lems,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.508560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.214634Z digest=sha256:d82797cac22bf2d32cc22bdf9bcb44896aa068632a33f2aca2df947b449951aa

Observation 218e96b0-d58b-4d22-9b7a-d5f41067cfa0 · outbound

This paper cites Silhouettes: a graphical aid to the interpretation and validation of cluster analysis,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Silhouettes: a graphical aid to the interpretation and validation of cluster analysis,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.217025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.217025Z digest=sha256:776de6d7ebcaa279a4e52f991847226fb1a25c182732b5022511ff30d8914ff5

Observation c39edd1e-2beb-494c-b54b-d9f797f5ae56 · outbound

This paper cites Deep residual learning for image recognition,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Deep residual learning for image recognition,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.219537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.219537Z digest=sha256:e126be2a2b1ffa9d4b7efbd3df2a35fa81e52017269ccbeaf606ef75c5a0d8f0

Observation 7995dc37-9bf7-441f-b9fc-c302ac82e98e · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks An image is worth 16x16 words: Transformers for image recognition at scale,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.422639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.222327Z digest=sha256:230d11f1f7c3643234b7288976bfe9ca1e357018569f8d63e4bb461662ea886a

Observation 119a2684-2c0c-439d-9f05-922bad6cc5cc · outbound

This paper cites A convnet for the 2020s,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks A convnet for the 2020s,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.382172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.225312Z digest=sha256:553ce5e7381e153ce9f93478aa6649bb75fbdaa4d7999cea675f718c433ff3b5

Observation 675996f6-57fd-4a99-ae23-b26f0b48b90a · outbound

This paper cites Laion-400m: Open dataset of clip-filtered 400 million image-text pairs,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Laion-400m: Open dataset of clip-filtered 400 million image-text pairs,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.372675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.227883Z digest=sha256:926013ed9f45aac41cd177918b91577caaa98443ca8ad5654cb94cc42b7c515d

Observation 82cc4b91-9e7e-4ad3-8f3c-c26e6c5d8102 · outbound

This paper cites AltCLIP: Altering the language encoder in CLIP for extended language capabilities,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks AltCLIP: Altering the language encoder in CLIP for extended language capabilities,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.362356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.230309Z digest=sha256:7b2c08092424c30367fc9cac0e0f3fefc1334919a1caeb86155365cf06126dc6

Observation d7237b80-b802-4761-859c-448fcacc02af · outbound

This paper cites Groupvit: Semantic segmentation emerges from text supervision,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Groupvit: Semantic segmentation emerges from text supervision,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.350740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.232910Z digest=sha256:1e828e7dfe49daa0f1a72fbf788d5bc76650ddc6e969f6396a3daff8eaae2f01

Observation a3b48329-45f1-4ab2-856a-bec9f04e4f0e · outbound

This paper cites Learning Generalized Zero-Shot Learners for Open-Domain Image Geolocalization.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Learning Generalized Zero-Shot Learners for Open-Domain Image Geolocalization

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.235754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.235754Z digest=sha256:ac3d1ce0a71f2fd59e0ccac746acc02e67a644e025f6eb328b0c8512479b73b2

Observation b63e8fcc-ea48-4330-8d0f-23b2ae3a5fef · outbound

This paper cites Demystifying CLIP Data.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Demystifying CLIP Data

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.239160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.239160Z digest=sha256:d3e09fd3e7baa0498d6e6b53f6f3c3a67cb9f117051dac27f3b2262f483f4afc

Observation 8829e10d-c0a3-4c87-a026-2ede33fedede · outbound

This paper cites BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.242350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.242350Z digest=sha256:e12d1ec58eb42dcf10749afb43f7d6d8cc6f9958515295c69ff31ed085fce2ba

Observation 60cb4d52-1c1d-4b93-8581-ff0c846425f1 · outbound

This paper cites Quilt-1M: One Million Image-Text Pairs for Histopathology.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Quilt-1M: One Million Image-Text Pairs for Histopathology

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.245892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.245892Z digest=sha256:ab7d78895476983355076a4828ae3e3e86b8b6ce345f9b1ecacc5c74d8096005

Observation 8e2650a0-8233-468f-8cc3-09fc29526389 · outbound

This paper cites BioCLIP: A vision foundation model for the tree of life,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks BioCLIP: A vision foundation model for the tree of life,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.341422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.248744Z digest=sha256:40b163f57f98b879f001f27e78f2f4296f2706bc0c732305da44a811d4230e43

Observation cb36a3e1-a99f-4dcc-8527-2511724e5102 · outbound

This paper cites Gpt-4: Generative pre-trained transformer,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Gpt-4: Generative pre-trained transformer,

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.331857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T23:19:02.251395Z digest=sha256:cfb293c85a44fb1c7a673d394d5c6b305cc10775f5222314f17afc6db4d9631b

Pith citing papers

No inbound Pith citation observations are available.