Pith. sign in

Paper Citation Record · LEDGER

Evaluating CLIP: Towards Characterization of Broader Capabilities and Downstream Implications

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2108.02818.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2108.02818 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T21:57:09.104016Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T02:49:24.782109Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c41121ab-e8dc-4a04-8546-051b9a7c9493 · inbound

CLIPScore: A Reference-free Evaluation Metric for Image Captioning cites this paper.

CLIPScore: A Reference-free Evaluation Metric for Image Captioning Evaluating CLIP: Towards Characterization of Broader Capabilities and Downstream Implications

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-12T22:24:10.077645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-12T22:24:10.018933Z digest=sha256:0abc081fd363c70b4877eaa714cd629b64ad98e43a60fe9f9da615fbb3c7c435

Observation 78a346d0-ddee-4761-b2b5-033aa1c51182 · inbound

Detecting Content Rating Violations in Android Applications: A Vision-Language Approach cites this paper.

Detecting Content Rating Violations in Android Applications: A Vision-Language Approach Evaluating CLIP: Towards Characterization of Broader Capabilities and Downstream Implications

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-08T21:57:09.104016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:57:09.104016Z digest=sha256:db83d3d94cc33f6364e182c10db41f13c5fe076831e491591a3b9e3a7c0219ba

Observation ab35c4c4-5ab4-49f7-8158-0a0153740591 · inbound

PRISM: Reducing Spurious Implicit Biases in Vision-Language Models with LLM-Guided Embedding Projection cites this paper.

PRISM: Reducing Spurious Implicit Biases in Vision-Language Models with LLM-Guided Embedding Projection Evaluating CLIP: Towards Characterization of Broader Capabilities and Downstream Implications

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T18:17:06.459886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:17:06.459886Z digest=sha256:3208bacfbccd2e0b89bd2af0eced12a513b194e91133a354e7eb4ec0c829f4cd

Observation 833e9928-9cfc-4302-bae0-3e319dab862f · inbound

Advancing Reliable Test-Time Adaptation of Vision-Language Models under Visual Variations cites this paper.

Advancing Reliable Test-Time Adaptation of Vision-Language Models under Visual Variations Evaluating CLIP: Towards Characterization of Broader Capabilities and Downstream Implications

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T18:00:23.348598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:00:23.348598Z digest=sha256:22a03e2964695f686cc83ed0cc315b669eecdc5df6bb365c681443295381e456

Observation 4a56f9e5-172e-4ad3-a4da-dcecce4dfbad · inbound

Learning from Limited and Imperfect Data cites this paper.

Learning from Limited and Imperfect Data Evaluating CLIP: Towards Characterization of Broader Capabilities and Downstream Implications

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T13:09:48.854564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:09:48.854564Z digest=sha256:bf122a593b3e465fa37a7abf68c359cf2643bbc09238820570f413edc89d6e19

Observation e1bda523-21bc-4b8a-b8b2-d5126c285824 · inbound

Constrained Prompt Enhancement for Improving Zero-Shot Generalization of Vision-Language Models cites this paper.

Constrained Prompt Enhancement for Improving Zero-Shot Generalization of Vision-Language Models Evaluating CLIP: Towards Characterization of Broader Capabilities and Downstream Implications

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T16:56:22.655634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:56:22.655634Z digest=sha256:1f9916fa41e51a728fb903a0caff7017f298f9e3999343bfe877bf26f50060ea

Observation 7decb8c3-6288-451f-9b3d-8597d0a09106 · inbound

Bias at the End of the Score cites this paper.

Bias at the End of the Score Evaluating CLIP: Towards Characterization of Broader Capabilities and Downstream Implications

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:11:03.198747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T15:36:38.278484Z digest=sha256:5ae72d7a3228f6e0fec1977e767518bab13d9257e7c923b34554e27929fab9d9

Observation c4def6bc-9e66-45d1-bcec-060ef37659c7 · inbound

An Attribute-Based Measure of Video Complexity cites this paper.

An Attribute-Based Measure of Video Complexity Evaluating CLIP: Towards Characterization of Broader Capabilities and Downstream Implications

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-06-28T19:02:34.094095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T19:00:54.718177Z digest=sha256:cb05ecb4a4c5ac02fef194ac4f88259d7a3e7aedf853c9a8d0724cd326d2bc76

Observation ed544556-5e65-41bf-afe6-a850e4016607 · inbound

The Market in the Model: Latent Diffusion as Neural Economy cites this paper.

The Market in the Model: Latent Diffusion as Neural Economy Evaluating CLIP: Towards Characterization of Broader Capabilities and Downstream Implications

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-04T02:49:24.783813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T19:05:25.386543Z digest=sha256:c8b07f82216fc7ffbe2a67ead0daaa11f34ef34d54271c645dcf07a20f367924

Observation 40e1dafc-9eb6-48fd-8cea-724cda48e559 · inbound

ComMem: Complementary Memory Systems for Test-Time Adaptation of Vision-Language Models cites this paper.

ComMem: Complementary Memory Systems for Test-Time Adaptation of Vision-Language Models Evaluating CLIP: Towards Characterization of Broader Capabilities and Downstream Implications

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-06-30T12:54:40.355067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-30T09:57:02.377866Z digest=sha256:ffebb13424f0b6f3152dece5e2c78ee92e38b99e2cc744f676b29a1e518d80a9

Observation c8674f4a-97a1-4e4a-a841-1463dfaca796 · inbound

Between Gradient and Natural Gradient: A Continuum of LoRA Initializations cites this paper.

Between Gradient and Natural Gradient: A Continuum of LoRA Initializations Evaluating CLIP: Towards Characterization of Broader Capabilities and Downstream Implications

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-01T00:26:41.634015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T00:26:41.634015Z digest=sha256:91dadf157fcaa9dd2e03f226c1fc64404c65fc4bde7cc42c101debebd0092ed0

Observation c26edcce-a659-4162-9b48-bfefc04f5c4d · inbound

Scaling Vision-Language Models Is Not Enough to Mitigate Bias cites this paper.

Scaling Vision-Language Models Is Not Enough to Mitigate Bias Evaluating CLIP: Towards Characterization of Broader Capabilities and Downstream Implications

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-31T14:32:40.279767Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T14:32:40.279767Z digest=sha256:fee024c4e9f6dd37208c6569220dd9d0e9ddfe6d4ecc2aa0299354a85189377b