Pith. sign in

Paper Citation Record · LEDGER

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves

As of 14 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 1 inbound Pith citation observation for arXiv:2412.11509.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.11509 v2

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T14:55:41.073787Z

measured 45 of 45 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:26:34.671641Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T19:26:35.642817Z

Reference resolution

44 of 44 outbound references displayed

  • verified exact0
  • verified fuzzy19
  • unresolved25
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 50196bb7-1941-4002-848c-039290e6c9a1 · outbound

This paper cites GPT-4 Technical Report.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T14:55:40.866543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:55:40.866543Z digest=sha256:21e41137c82fcd9a9099701760a84b432bc0cbeac2cb2052fc30dcce73286658

Observation a5a51f63-df43-42ec-aba2-7c1ea5438fa4 · outbound

This paper cites Food-101–mining discriminative components with random forests.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Food-101–mining discriminative components with random forests

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T14:55:40.872498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:55:40.872498Z digest=sha256:12d38db615b6ba5025bf6b37c6bb7c4e74681c9b0282bc284afc15e31127ec09

Observation f1d5b6ac-5348-4101-a2ea-f9d5e6efdd64 · outbound

This paper cites Describing textures in the wild.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Describing textures in the wild

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T14:55:40.877302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:55:40.877302Z digest=sha256:0ac23b453f3c2f4e1443c69d3ebf0b75b71073f4215fd2b6d230aed57d01ca67

Observation 12676e42-8686-4727-899c-c51b06e29ad6 · outbound

This paper cites Imagenet: A large-scale hierarchical image database.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Imagenet: A large-scale hierarchical image database

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T14:55:40.882391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:55:40.882391Z digest=sha256:104f8fa270e385b07d45a4e764c77020a3124fcf2f0758ff1ae6b00ef07356a5

Observation fc7cf552-8be5-4f6f-ac23-fd1b29b84956 · outbound

This paper cites Learning gener- ative visual models from few training examples: An incre- mental bayesian approach tested on 101 object categories.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Learning gener- ative visual models from few training examples: An incre- mental bayesian approach tested on 101 object categories

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:55:41.658028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T14:55:40.887409Z digest=sha256:7bd9c2cf641d4c20eb65d704c3b91194d40041b6a0bdbf1bb46be03fe3eb63d5

Observation 0f96c8ec-4c78-4462-81ce-30069592af70 · outbound

This paper cites Prompt- det: Towards open-vocabulary detection using uncurated im- ages.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Prompt- det: Towards open-vocabulary detection using uncurated im- ages

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:55:41.641803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T14:55:40.892303Z digest=sha256:8e635c3fc443bddd3eed22875c9b8ed5725e268d05c7dba3009e5c7de0a72112

Observation c4eb3178-c79a-496d-a013-837de9c261ec · outbound

This paper cites Generalized meta-fdmixup: Cross-domain few-shot learning guided by labeled target data.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Generalized meta-fdmixup: Cross-domain few-shot learning guided by labeled target data

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:55:41.626322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T14:55:40.897467Z digest=sha256:b759b086b3d3336ddd64e57b2fff7d49ae47ec33bdee278ae49a39e482979834

Observation 36f32ad2-6175-4f97-90b5-2f35af77f568 · outbound

This paper cites Styleadv: Meta style adversarial training for cross-domain few-shot learning.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Styleadv: Meta style adversarial training for cross-domain few-shot learning

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:55:41.610661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T14:55:40.902416Z digest=sha256:04dac241242fc17069c487f985603e948290bb7e0dcec9b11c0bb78a8f98c977

Observation 69625c8a-da00-4531-82f3-b721d9cf621f · outbound

This paper cites Cross-domain few-shot object detection via enhanced open-set object detector.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Cross-domain few-shot object detection via enhanced open-set object detector

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:55:41.594743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T14:55:40.907812Z digest=sha256:32779ba03631f296b0e1774f062d808348e3e1ec7e75763ee788da2ba998bfcc

Observation 39a99d8c-286a-44e7-a172-e881a96deafd · outbound

This paper cites Clip-adapter: Better vision-language models with feature adapters.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Clip-adapter: Better vision-language models with feature adapters

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:55:41.579042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T14:55:40.912369Z digest=sha256:5d9e291ba552c3a7b43e63123c78f378b3e000b53b3b6bbbca5c0f8c9bac2f91

Observation 038038e9-71d8-4733-8b9f-3a472b58f70f · outbound

This paper cites Open-vocabulary Object Detection via Vision and Language Knowledge Distillation.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Open-vocabulary Object Detection via Vision and Language Knowledge Distillation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T14:55:40.917060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:55:40.917060Z digest=sha256:5d4e8a3a49be971fa68398033ddba903d4164de36e1e90ebb4a85a27b4f13bb7

Observation 31dd6fd7-3d72-4954-83e2-f2c83fff7ddc · outbound

This paper cites Eurosat: A novel dataset and deep learning benchmark for land use and land cover classification.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Eurosat: A novel dataset and deep learning benchmark for land use and land cover classification

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T14:55:40.922596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:55:40.922596Z digest=sha256:dd51a2db26a6196d153a635f35c915e239418def78535bebed5860721a81095f

Observation 8e503ee4-dba8-447f-8507-3be414a1399c · outbound

This paper cites The many faces of robust- ness: A critical analysis of out-of-distribution generalization.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves The many faces of robust- ness: A critical analysis of out-of-distribution generalization

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T14:55:40.927495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:55:40.927495Z digest=sha256:1672f4bf3b455bc863668b2ac75c11ceeb83153ced87df1d7880e807ddfd7949

Observation 08e6694f-6f22-4b42-abc0-b67bd7c2c8eb · outbound

This paper cites Natural adversarial examples.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Natural adversarial examples

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T14:55:40.932178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:55:40.932178Z digest=sha256:cf25fdd3d242d0f3921221f1043904eab83de46c4ca94196534445177bb3e68d

Observation d7d6ad56-eb48-4648-be15-67e927630919 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves LoRA: Low-Rank Adaptation of Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T14:55:40.937188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:55:40.937188Z digest=sha256:479aba4e2dbfafa00ec3bc7802a749057c128e994e5de27fb0e4b277f9c97566

Observation c7bd8b7d-e88d-42fa-b6d5-1e4c6f88fcd4 · outbound

This paper cites Scaling up visual and vision-language representa- tion learning with noisy text supervision.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Scaling up visual and vision-language representa- tion learning with noisy text supervision

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T14:55:40.941812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:55:40.941812Z digest=sha256:da6cfe9707d3670212bbec3f7601c3d1ec1da306b087da8e6a9dfb7b3954709a

Observation 218717cc-5862-437d-8a66-23f6df721752 · outbound

This paper cites Maple: Multi-modal prompt learning.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Maple: Multi-modal prompt learning

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:55:41.525993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T14:55:40.946668Z digest=sha256:e903803436bcd678d15ff32abfb180eb35c204de9e99e17f597669455fcea333

Observation b510d689-7869-45c0-ad7c-0fc0019fa416 · outbound

This paper cites Self-regulating prompts: Foundational model adaptation without forgetting.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Self-regulating prompts: Foundational model adaptation without forgetting

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:55:41.510474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T14:55:40.951194Z digest=sha256:e7f4e49f5d09db0134d6ffa8b26806d3b9701986cca0873d229709611855c49f

Observation cae21571-0a4d-46a5-9316-9ff8c2c0746c · outbound

This paper cites 3d object representations for fine-grained categorization.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves 3d object representations for fine-grained categorization

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T14:55:40.955491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:55:40.955491Z digest=sha256:bf20c530d9bc5abe2b712e12958edfe8e26efdafd64d11739bcd039b596bd4e7

Observation c4dedc22-9636-4ae1-8f2c-ddd769184ec5 · outbound

This paper cites Image segmenta- tion using text and image prompts.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Image segmenta- tion using text and image prompts

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:55:41.484465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T14:55:40.960226Z digest=sha256:9416bd824b357d3626ab29db73225f5611d77fb9e559c5ff96d7c25444bb7dce

Observation fffef978-e826-4180-9f1e-bddedd282271 · outbound

This paper cites Fine-Grained Visual Classification of Aircraft.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Fine-Grained Visual Classification of Aircraft

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T14:55:40.964760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:55:40.964760Z digest=sha256:2086f684930710728789aae27b19c41eada911ce43e5ac372509ca70c8c51be7

Observation fb8a4d52-e292-41c2-818e-2e7e6c8b81a7 · outbound

This paper cites Automated flower classification over a large number of classes.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Automated flower classification over a large number of classes

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T14:55:40.969740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:55:40.969740Z digest=sha256:89fa468e3f91ae8bff6a6cdfc259647368bb5df8596ec09895b28c0989255b5b

Observation 105a9afd-9f48-4627-ab4d-9f0d5c3b3872 · outbound

This paper cites Cats and dogs.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Cats and dogs

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T14:55:40.973859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:55:40.973859Z digest=sha256:7aee65d36e851d4dc5358a18d9b078b2c35451d53d31a90c8ac67b46c910539a

Observation f9433952-39a0-483a-939d-ae6e48940013 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Learning transferable visual models from natural language supervi- sion

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:55:41.451261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T14:55:40.978466Z digest=sha256:9b285815aaedd314aeb090b3f10b8e859489fedc15798a1177a8b7db6bede9bb

Observation 6c0aa5d0-9c55-4141-a48f-25ba5ede9a99 · outbound

This paper cites Denseclip: Language-guided dense prediction with context- aware prompting.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Denseclip: Language-guided dense prediction with context- aware prompting

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:55:41.436276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T14:55:40.982944Z digest=sha256:446ebfc7ace85a5590fb978e771a6013947b45da4b0093f863e60b15595e68b1

Observation 28f4a049-da1a-453c-be0c-ac38c72a895e · outbound

This paper cites Do imagenet classifiers generalize to im- agenet? In International conference on machine learning , pages 5389–5400.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Do imagenet classifiers generalize to im- agenet? In International conference on machine learning , pages 5389–5400

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T14:55:40.987519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:55:40.987519Z digest=sha256:d8c4e7b8d22d652f218b94dc6cba6f73de89d2caee239b9f781869a9eff0759b

Observation 9f46c994-a1c0-4747-b71a-05dc7bd05873 · outbound

This paper cites Consistency-guided Prompt Learning for Vision-Language Models.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Consistency-guided Prompt Learning for Vision-Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T14:55:40.992129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:55:40.992129Z digest=sha256:c95b75248060aab1ef617e1757f3193ed1e6eff593576e970266cf6e086a4d36

Observation 17a0e2ac-bd65-4368-9a73-64d9dbf85c5b · outbound

This paper cites UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T14:55:40.997354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:55:40.997354Z digest=sha256:c0489c09f7d24b42a03f6068baa4feecb4b75b26bf22a13a5c6489218b8259f3

Observation ee3d80c4-0f9e-4a51-a8e6-01f409d394a0 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves LLaMA: Open and Efficient Foundation Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T14:55:41.002297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:55:41.002297Z digest=sha256:177d384d7833ac36a292876f8c9247734403df4695f1b367c0ead658cf8ff636

Observation feb61d8c-569c-4f08-9f2d-99e32086f4db · outbound

This paper cites Learning robust global representations by penalizing local predictive power.Advances in Neural Information Pro- cessing Systems, 32, 2019.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Learning robust global representations by penalizing local predictive power.Advances in Neural Information Pro- cessing Systems, 32, 2019

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T14:55:41.006861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:55:41.006861Z digest=sha256:9772c1dc09419616e0fe7696acd0d2ef0465457fd03ba30644be5e1873c6cf95

Observation edecc4d4-a53d-48d8-9f9a-c9565a98389d · outbound

This paper cites Sun database: Large-scale scene recognition from abbey to zoo.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Sun database: Large-scale scene recognition from abbey to zoo

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T14:55:41.011614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:55:41.011614Z digest=sha256:d1536cb5113c08586649b57a0cfd31627cd2cc3dacb613caf5fb1fe3998e6d91

Observation 12187ccf-abcf-4ef3-80e2-21c2c2f10cac · outbound

This paper cites Visual- language prompt tuning with knowledge-guided context op- timization.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Visual- language prompt tuning with knowledge-guided context op- timization

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T14:55:41.016126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:55:41.016126Z digest=sha256:af55ca57d72b7aa91f92f5d4daa0b041825e44078870867376ef1b506b976ac7

Observation 4650617c-a3aa-4a58-ab49-1c8bd2c5f95b · outbound

This paper cites Tcp: Textual- based class-aware prompt tuning for visual-language model.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Tcp: Textual- based class-aware prompt tuning for visual-language model

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:55:41.383096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T14:55:41.021112Z digest=sha256:ed4575643d40fc72b16d3f790cd6b193847b3ab735b492eb1fb13876f599b005

Observation 2c36153d-63d1-48ad-a815-2d6ffac7da70 · outbound

This paper cites FILIP: Fine-grained Interactive Language-Image Pre-Training.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves FILIP: Fine-grained Interactive Language-Image Pre-Training

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T14:55:41.025736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:55:41.025736Z digest=sha256:b5de31adaca92b32a599aefb25762492c340973c46c8ee87c70fba876ce086ae

Observation 4ac206f5-5e74-47ed-b819-89b882264cb8 · outbound

This paper cites Open-vocabulary detr with conditional matching.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Open-vocabulary detr with conditional matching

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:55:41.366888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T14:55:41.030573Z digest=sha256:6d51898bcbe0e42e60115db023aeb0264ecf99fc50e67905fbe2a5c2b549e13f

Observation e7e505bc-1423-44f1-8ed7-7dde50754993 · outbound

This paper cites Lit: Zero-shot transfer with locked-image text tuning.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Lit: Zero-shot transfer with locked-image text tuning

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:55:41.351706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T14:55:41.035892Z digest=sha256:f9c937668a07cdd595ae21f8c5779f7d2c548ee9c70d74a8f5bb14c0ca5a8c72

Observation 346028ab-2921-4926-8abb-299c6892f5be · outbound

This paper cites Free-lunch for cross-domain few-shot learning: Style-aware episodic training with robust contrastive learning.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Free-lunch for cross-domain few-shot learning: Style-aware episodic training with robust contrastive learning

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:55:41.336546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T14:55:41.040584Z digest=sha256:75c5a7b9d59733a0f4cfbc8bac191b9ed083d59c7bd7045cc3f3fc3abfa7304c

Observation 6f45b139-4843-4ac7-82ec-1c8c5d73f536 · outbound

This paper cites Deta: Denoised task adaptation for few-shot learning.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Deta: Denoised task adaptation for few-shot learning

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:55:41.320563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T14:55:41.045210Z digest=sha256:235ff5c8bf475852ba3921e207d14097ac3e3a1541b63f7df4959f5f0d830e57

Observation d4ed0a32-f3a7-4109-9317-d6e362a5639b · outbound

This paper cites Dept: Decoupled prompt tuning.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Dept: Decoupled prompt tuning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:55:41.304739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T14:55:41.049780Z digest=sha256:3c5d50d6cb0e9af0e5bd2bc3857c319365ea503b57a23cd0b5f01e75ac93eaf2

Observation 4ca619a2-5012-4f47-8b5e-409ad5fb1973 · outbound

This paper cites Tip-Adapter: Training-free CLIP-Adapter for Better Vision-Language Modeling.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Tip-Adapter: Training-free CLIP-Adapter for Better Vision-Language Modeling

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T14:55:41.054173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:55:41.054173Z digest=sha256:1a4d656e24f04b3812734c805d2eb1ab416d189a949f9ccc800fe59774b260ef

Observation 9e95b12f-2456-4299-8e55-fea5e2a463ae · outbound

This paper cites Conditional prompt learning for vision-language mod- els.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Conditional prompt learning for vision-language mod- els

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-11T14:55:41.059049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:55:41.059049Z digest=sha256:f7bfe4215945eeeccd8fe611acab605bfb20b744d3095d120bcd0b33b72d64f1

Observation 479f12e9-6d32-4f0b-8675-7875d882cb47 · outbound

This paper cites Learning to prompt for vision-language models.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Learning to prompt for vision-language models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-11T14:55:41.064119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:55:41.064119Z digest=sha256:4612b207e9893964300b4817d19884f18e294d5b6cd17061d768240703c38628

Observation c14771ea-0ddb-4fb0-9f07-08ce44532234 · outbound

This paper cites Detecting twenty-thousand classes using image-level supervision.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Detecting twenty-thousand classes using image-level supervision

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:55:41.269575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T14:55:41.069003Z digest=sha256:a5673843d8952a7652de38614851bf021cb8ccaf4a3925f49c025fe6ec74f780

Observation 03474990-154b-415d-90fd-87ffdc0dd1b6 · outbound

This paper cites Prompt-aligned gradient for prompt tuning.

Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves Prompt-aligned gradient for prompt tuning

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:55:41.253467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T14:55:41.073787Z digest=sha256:95b621c9ef1017a97f2c5a988324a7d7bccb3dcffba5011bc817d64f4e856862

Pith citing papers

Observation a41e6341-a21e-4246-aefd-e225f5a2fa8e · inbound

Dynamic Rank Adaptation for Vision-Language Models cites this paper.

Dynamic Rank Adaptation for Vision-Language Models Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-08-06T19:26:35.649789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T19:26:34.671641Z digest=sha256:bd9f2489a7f6bcbaea403a22467b1c5866b291bac27de88d0a9690b24ffad636