Pith. sign in

Paper Citation Record · LEDGER

SLIP: Self-supervision meets Language-Image Pre-training

As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2112.12750.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2112.12750 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T01:06:20.852001Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-24T11:09:22.375297Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a08922a2-2835-489c-83c5-e3a7ccc6d006 · inbound

Hierarchical Text-Conditional Image Generation with CLIP Latents cites this paper.

Hierarchical Text-Conditional Image Generation with CLIP Latents SLIP: Self-supervision meets Language-Image Pre-training

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T16:55:57.837312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T16:55:57.612364Z digest=sha256:c7a8e609e5903f2d5f3c5993a9a9ada8bce23c8e8ef09d07ec0e2c1a239d217c

Observation 2af4e0bb-9e70-4c8c-ba09-1940b83fb103 · inbound

DetailCLIP: Injecting Image Details into CLIP's Feature Space cites this paper.

DetailCLIP: Injecting Image Details into CLIP's Feature Space SLIP: Self-supervision meets Language-Image Pre-training

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-24T11:09:22.379272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-24T11:08:20.298043Z digest=sha256:81e2467117c41584022eb3f34a10d3a05fb4ff91377ee6b32aef6a9bbe9aa28b

Observation aea50e06-111b-4c59-92a7-a0ed59beed2e · inbound

Demystifying CLIP Data cites this paper.

Demystifying CLIP Data SLIP: Self-supervision meets Language-Image Pre-training

Reference 79

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T09:20:20.320666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-16T09:20:20.143143Z digest=sha256:ac80afa06e434822f8c858828533733d1670a2cb4c6245e347ad1e723dab7be5

Observation c9fd9868-860c-4876-ba31-ae48f8af8efd · inbound

Visual Pre-Training on Unlabeled Images using Reinforcement Learning cites this paper.

Visual Pre-Training on Unlabeled Images using Reinforcement Learning SLIP: Self-supervision meets Language-Image Pre-training

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T01:06:20.852001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T01:06:20.852001Z digest=sha256:1efcc4ba65dd8ece8827a18336153aafc10c5bdd697157a660a5222e937d168c

Observation 3d51cdbc-cacf-482e-9d25-a439e1f36727 · inbound

CXR-CML: Improved zero-shot classification of long-tailed multi-label diseases in Chest X-Rays cites this paper.

CXR-CML: Improved zero-shot classification of long-tailed multi-label diseases in Chest X-Rays SLIP: Self-supervision meets Language-Image Pre-training

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T14:24:51.259176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:24:51.259176Z digest=sha256:26c714ac1cbbaef790ba8589b30f610687521c836f82452244165ba5f51792ef

Observation 026b57e2-414c-42b3-ac8e-6a91ed6ce1f7 · inbound

Meta CLIP 2: A Worldwide Scaling Recipe cites this paper.

Meta CLIP 2: A Worldwide Scaling Recipe SLIP: Self-supervision meets Language-Image Pre-training

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T12:08:23.020444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:08:23.020444Z digest=sha256:48a7e0584bb8ef2b6c244541de1de5ca1e531c6c0e1af7b1e83c7e4c89afa502

Observation 7b19e53f-3bc4-47e8-b13b-5492d1d5e755 · inbound

Bottleneck Tokens for Unified Multimodal Retrieval cites this paper.

Bottleneck Tokens for Unified Multimodal Retrieval SLIP: Self-supervision meets Language-Image Pre-training

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T11:01:03.336386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T15:14:00.615638Z digest=sha256:5705a7238b763f28df4647b3a999b49467391b341052db05aa51f592b9d1f312