Pith. sign in

Paper Citation Record · LEDGER

Variational Adapter for Cross-modal Similarity Representation

As of 19 August 2026, this Paper Citation Record lists 16 of 16 outbound references and 0 inbound Pith citation observations for arXiv:2605.30968.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.30968 v1

Coverage vector

measured 16 of 16 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-28T23:04:14.351521Z

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

16 of 16 outbound references displayed

  • verified exact8
  • verified fuzzy0
  • unresolved7
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2b563439-ea02-4ae3-898b-1aec1e37d9fe · outbound

This paper cites Microsoft COCO Captions: Data Collection and Evaluation Server.

Variational Adapter for Cross-modal Similarity Representation Microsoft COCO Captions: Data Collection and Evaluation Server

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:16:00.174405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T23:04:14.351521Z digest=sha256:3ac8fec9f9f4b70d6e92a4c921cb272050f359024526d872e6e0af9676cb62ab

Observation fbc31ecb-526e-434e-8303-1d2d37adcf6a · outbound

This paper cites VSE++: Improving Visual-Semantic Embeddings with Hard Negatives.

Variational Adapter for Cross-modal Similarity Representation VSE++: Improving Visual-Semantic Embeddings with Hard Negatives

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:16:00.157696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T23:04:14.351521Z digest=sha256:78580f54596a3879be687f946d9840437e19b566c8f1e502d38d50c57882a0b1

Observation 0e81e320-1d72-430d-9fba-7448eabe7cdf · outbound

This paper cites Learning generative visual models from few training examples: An incremen- tal bayesian approach tested on 101 object categories.

Variational Adapter for Cross-modal Similarity Representation Learning generative visual models from few training examples: An incremen- tal bayesian approach tested on 101 object categories

Reference 3

Resolution
unresolved
no resolver link, observed 2026-06-28T23:04:14.351521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T23:04:14.351521Z digest=sha256:428a9799ba02185987f7094cc49e9ba2f6d586a435304d4d287452c6ca839ede

Observation bdb8f07f-9cad-4f81-bdd2-46f9bd55d584 · outbound

This paper cites Auto-Encoding Variational Bayes.

Variational Adapter for Cross-modal Similarity Representation Auto-Encoding Variational Bayes

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:16:00.170432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T23:04:14.351521Z digest=sha256:c045f7a5ded81467775cb673f158afea7f83169a15aea87c17474aa2ad5e7553

Observation 10bebb61-460d-4435-82e2-c612871c838d · outbound

This paper cites Fine-Grained Visual Classification of Aircraft.

Variational Adapter for Cross-modal Similarity Representation Fine-Grained Visual Classification of Aircraft

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:16:00.171923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T23:04:14.351521Z digest=sha256:4e834cb7416cb81b7371e49187091c98c9e72d6f8283ff89168eee526144a9cf

Observation c973190d-b69f-430a-9a31-e1af1bca4a65 · outbound

This paper cites Crisscrossed Captions: Extended Intramodal and Intermodal Semantic Similarity Judgments for MS-COCO.

Variational Adapter for Cross-modal Similarity Representation Crisscrossed Captions: Extended Intramodal and Intermodal Semantic Similarity Judgments for MS-COCO

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:16:00.168002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T23:04:14.351521Z digest=sha256:121a4d7988ed6b50cd9e0c84802b0094b606f45ca7c1e1e82f06ae64f60eb8bc

Observation 38b61483-6ae9-48e5-b3c7-05dc67544cc1 · outbound

This paper cites Consistency-guided Prompt Learning for Vision-Language Models.

Variational Adapter for Cross-modal Similarity Representation Consistency-guided Prompt Learning for Vision-Language Models

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:16:00.162367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T23:04:14.351521Z digest=sha256:9c6867351da291be38c1779f6e7056af83fc4c0ed4b3264d1050c56de71ebca1

Observation 2e8fe5c2-f876-46cf-a546-4b863c1a7096 · outbound

This paper cites UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild.

Variational Adapter for Cross-modal Similarity Representation UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:16:00.159049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T23:04:14.351521Z digest=sha256:3db188458671163feb34e39eee63591dc515f341e2cc45559c9c5a8fc649529b

Observation b8a1df95-1f29-4a19-9706-7db4a60409b8 · outbound

This paper cites Probvlm: Probabilistic adapter for frozen vison-language models.

Variational Adapter for Cross-modal Similarity Representation Probvlm: Probabilistic adapter for frozen vison-language models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-28T23:04:14.351521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T23:04:14.351521Z digest=sha256:92f6a16a3e12db81ddf0a1844bab819f5bfd2b0daa0c64a4454fab2730b363ac

Observation 85604e52-8341-4db0-863f-673c1eb56d2c · outbound

This paper cites Towards Instance-wise Personalized Federated Learning via Semi-Implicit Bayesian Prompt Tuning.

Variational Adapter for Cross-modal Similarity Representation Towards Instance-wise Personalized Federated Learning via Semi-Implicit Bayesian Prompt Tuning

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:16:00.175678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T23:04:14.351521Z digest=sha256:1caea798adfdac67d89bd5c8f455b1ab1e8b5babaecadb6d269cabf7667d3204

Observation 2e869e0b-8e4b-4480-923e-6131f33a2db7 · outbound

This paper cites C., and Liu, Z.

Variational Adapter for Cross-modal Similarity Representation C., and Liu, Z

Reference 11

Resolution
unresolved
no resolver link, observed 2026-06-28T23:04:14.351521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T23:04:14.351521Z digest=sha256:3417db4f18599a1c59a9f7a03607b1dfea1ffe50cd9f6dfa5f478fd380b53459

Observation 8cc7b580-0128-408f-bf6a-9a7b7cd755d7 · outbound

This paper cites Upadhyay et al.(Upadhyay et al.,.

Variational Adapter for Cross-modal Similarity Representation Upadhyay et al.(Upadhyay et al.,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-06-28T23:04:14.351521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T23:04:14.351521Z digest=sha256:79af7a24972a46df89adbeb53a738324881bbd5a8abbac108ceacc2402ead7eb

Observation c8176680-ae13-416e-aa11-3f3aba5675ad · outbound

This paper cites Collectively, these approaches expand the set of potential results, constructing a richer semantic retrieval space.

Variational Adapter for Cross-modal Similarity Representation Collectively, these approaches expand the set of potential results, constructing a richer semantic retrieval space

Reference 13

Resolution
unresolved
no resolver link, observed 2026-06-28T23:04:14.351521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T23:04:14.351521Z digest=sha256:2ec04bfd7c4f0bc32cd20620bb882995ff5298ded6ba5b39ddbec505496409cc

Observation ed0e5583-2841-40a2-9fc2-87fbf0dc992a · outbound

This paper cites We employ a 16-shot setting and use the template ”a photo of a <category>” for the word embeddings.

Variational Adapter for Cross-modal Similarity Representation We employ a 16-shot setting and use the template ”a photo of a <category>” for the word embeddings

Reference 14

Resolution
unresolved
no resolver link, observed 2026-06-28T23:04:14.351521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T23:04:14.351521Z digest=sha256:029a62765d9b7f2104babbfba1c212948c555a83091bc4b7ec6c794f6a7d8c5d

Observation 4c5c5e61-3753-4fce-a9cb-435a77b34b69 · outbound

This paper cites an unresolved cited work.

Variational Adapter for Cross-modal Similarity Representation Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-06-28T23:04:14.351521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T23:04:14.351521Z digest=sha256:9793082260762ae29dfc1701e761b70113e9543bef3758b5f6a8b11301a6cda5

Observation 11ceda50-bd12-426f-8f46-7c4d4f6081d3 · outbound

This paper cites the sample may not actually be a negative sample.

Variational Adapter for Cross-modal Similarity Representation the sample may not actually be a negative sample

Reference 16

Resolution
malformed identifier
arxiv_id, observed 2026-07-01T19:16:00.173027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T23:04:14.351521Z digest=sha256:611765d26a2d8d034c188ecb9198d297c73e8966a24773695f540f1a376543b2

Pith citing papers

No inbound Pith citation observations are available.