Pith. sign in

Paper Citation Record · LEDGER

Mitigate the Gap: Investigating Approaches for Improving Cross-Modal Alignment in CLIP

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2406.17639.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.17639 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:09:14.869726Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T15:24:50.126060Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 49c0afc5-d534-494e-bdaf-b4c25c4ef3eb · inbound

Mind the Gap: Preserving and Compensating for the Modality Gap in CLIP-Based Continual Learning cites this paper.

Mind the Gap: Preserving and Compensating for the Modality Gap in CLIP-Based Continual Learning Mitigate the Gap: Investigating Approaches for Improving Cross-Modal Alignment in CLIP

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T18:09:14.869726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:09:14.869726Z digest=sha256:199d9671e4f874f6fc87f2db364bb55a2e7135b1fb7742c4ce3be77f8bfb4aed

Observation 92d90656-2777-48e8-84c1-4c84c71e51a7 · inbound

CLIP-RD: Relative Distillation for Efficient CLIP Knowledge Distillation cites this paper.

CLIP-RD: Relative Distillation for Efficient CLIP Knowledge Distillation Mitigate the Gap: Investigating Approaches for Improving Cross-Modal Alignment in CLIP

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T00:23:22.861383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T00:20:36.975106Z digest=sha256:e79b17fdc04cdc78c1d1d92e4f2b47b41052e6d26af1a22d010af4fadb4ed730

Observation 7b68ec84-d598-4413-8618-f5190302e054 · inbound

Reviving In-domain Fine-tuning Methods for Source-Free Cross-domain Few-shot Learning cites this paper.

Reviving In-domain Fine-tuning Methods for Source-Free Cross-domain Few-shot Learning Mitigate the Gap: Investigating Approaches for Improving Cross-Modal Alignment in CLIP

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:27:02.550470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T01:24:22.361181Z digest=sha256:67fea7ad2ee83afeee96d5b331d1d65fc62e295a3f7deafcfa4a83d344b818b7

Observation f63168dc-4e44-449f-8bb0-1daf752c84d4 · inbound

STAMBRIDGE: Spectral-Temporal Amplitude-aware Mid-Feature Bridge for EEG Visual Decoding cites this paper.

STAMBRIDGE: Spectral-Temporal Amplitude-aware Mid-Feature Bridge for EEG Visual Decoding Mitigate the Gap: Investigating Approaches for Improving Cross-Modal Alignment in CLIP

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-25T03:35:18.157957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T03:30:38.528341Z digest=sha256:c68bcd0b2b8f53c34a8da995cb8ae943d06a68e5494ea42ff84eef457522da29

Observation ed936728-3c94-45f7-a784-a2542342eb02 · inbound

STAMBRIDGE: Spectral-Temporal Amplitude-aware Mid-Feature Bridge for EEG Visual Decoding cites this paper.

STAMBRIDGE: Spectral-Temporal Amplitude-aware Mid-Feature Bridge for EEG Visual Decoding Mitigate the Gap: Investigating Approaches for Improving Cross-Modal Alignment in CLIP

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-06-30T15:24:50.127560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T15:18:41.350333Z digest=sha256:015a84290761fe477a5f4d78714c06af068a1d2f70c0fe26227d50c2f95a9185

Observation c343c166-52ce-4d12-824e-6580531435fc · inbound

On the modality gap and the contrastive loss in multi-modal representation learning cites this paper.

On the modality gap and the contrastive loss in multi-modal representation learning Mitigate the Gap: Investigating Approaches for Improving Cross-Modal Alignment in CLIP

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-14T09:55:43.412471Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T09:55:43.412471Z digest=sha256:8631a63504a703b14b3c82090d2301a1f0f03744c4aab5a6d347805a536feadd

Observation b21e126c-c892-4099-aaee-3e4bbc1b0100 · inbound

AspectCLIP: Optimizing CLIP Representation Space via Aspect-Guided Consistency Regularization cites this paper.

AspectCLIP: Optimizing CLIP Representation Space via Aspect-Guided Consistency Regularization Mitigate the Gap: Investigating Approaches for Improving Cross-Modal Alignment in CLIP

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T03:43:47.345797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:43:47.345797Z digest=sha256:e77b817dcf79232b9faa028376e02310f168d688f9b3ca65dcacfe6253174a2b

Observation d7cc4406-92e3-47ed-834d-16aa5d5e93b9 · inbound

ScalablePromptus: Scalable and High-Fidelity Prompt-Based Video Streaming cites this paper.

ScalablePromptus: Scalable and High-Fidelity Prompt-Based Video Streaming Mitigate the Gap: Investigating Approaches for Improving Cross-Modal Alignment in CLIP

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-01T02:23:45.550343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:23:45.550343Z digest=sha256:89fb18e6ee7adcb82161ab21cdfa870b2f9314a7721cc86bfe26a711327978f0

Observation 20e46a5d-9c0b-4505-9b39-08dc69a291c2 · inbound

TokenSwap: Benchmarking and Reducing the Modality Gap in Multimodal LLMs cites this paper.

TokenSwap: Benchmarking and Reducing the Modality Gap in Multimodal LLMs Mitigate the Gap: Investigating Approaches for Improving Cross-Modal Alignment in CLIP

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-03T00:55:23.673276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T00:55:23.673276Z digest=sha256:fef260a404846b8efc2b60e4b8bbc63c2f9019fde4dc35d9ce1c2bb4cde021b4