Pith. sign in

Paper Citation Record · LEDGER

SLIP: Self-supervision meets Language-Image Pre-training

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2112.12750.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2112.12750 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T01:06:20.852001Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-24T11:09:22.375297Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a08922a2-2835-489c-83c5-e3a7ccc6d006 · inbound

Hierarchical Text-Conditional Image Generation with CLIP Latents cites this paper.

Hierarchical Text-Conditional Image Generation with CLIP Latents SLIP: Self-supervision meets Language-Image Pre-training

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T16:55:57.837312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T16:55:57.612364Z digest=sha256:22c407013eb3192702db0e551926fe0d247084088a706952aa15ecf998348f84

Observation 2af4e0bb-9e70-4c8c-ba09-1940b83fb103 · inbound

DetailCLIP: Injecting Image Details into CLIP's Feature Space cites this paper.

DetailCLIP: Injecting Image Details into CLIP's Feature Space SLIP: Self-supervision meets Language-Image Pre-training

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-24T11:09:22.379272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-24T11:08:20.298043Z digest=sha256:ba402909c5c7ecfbd497f66a056f4227527bec4c05e04ad6da38c43c6c274d63

Observation aea50e06-111b-4c59-92a7-a0ed59beed2e · inbound

Demystifying CLIP Data cites this paper.

Demystifying CLIP Data SLIP: Self-supervision meets Language-Image Pre-training

Reference 79

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T09:20:20.320666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-16T09:20:20.143143Z digest=sha256:a6600b1fe0a4e44fed899ceca76d7bcdbece5c4ad9b0aa4feed8a17b68f8d36d

Observation c9fd9868-860c-4876-ba31-ae48f8af8efd · inbound

Visual Pre-Training on Unlabeled Images using Reinforcement Learning cites this paper.

Visual Pre-Training on Unlabeled Images using Reinforcement Learning SLIP: Self-supervision meets Language-Image Pre-training

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T01:06:20.852001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T01:06:20.852001Z digest=sha256:4066ec9650ec8d78373ad208243d5031bc3a9045a8c8fb29f104f90eb2fd96cf

Observation 3d51cdbc-cacf-482e-9d25-a439e1f36727 · inbound

CXR-CML: Improved zero-shot classification of long-tailed multi-label diseases in Chest X-Rays cites this paper.

CXR-CML: Improved zero-shot classification of long-tailed multi-label diseases in Chest X-Rays SLIP: Self-supervision meets Language-Image Pre-training

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T14:24:51.259176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:24:51.259176Z digest=sha256:9a808a98d2ad50e282b8ff01e938e19427d60ae25bfdd09c77792ae43eb4235c

Observation 026b57e2-414c-42b3-ac8e-6a91ed6ce1f7 · inbound

Meta CLIP 2: A Worldwide Scaling Recipe cites this paper.

Meta CLIP 2: A Worldwide Scaling Recipe SLIP: Self-supervision meets Language-Image Pre-training

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T12:08:23.020444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:08:23.020444Z digest=sha256:fec2fdb26beea9879d2f0c087ff6407df35bb8b405235b02b9275c2d142d83f3

Observation 7b19e53f-3bc4-47e8-b13b-5492d1d5e755 · inbound

Bottleneck Tokens for Unified Multimodal Retrieval cites this paper.

Bottleneck Tokens for Unified Multimodal Retrieval SLIP: Self-supervision meets Language-Image Pre-training

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T11:01:03.336386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T15:14:00.615638Z digest=sha256:e2d8f5a98b3c714f8115a542a19fceded7fc5bc17ea4bbbc5533ffbd53fc6f99