Pith. sign in

Paper Citation Record · LEDGER

ClearCLIP: Decomposing CLIP Representations for Dense Vision-Language Inference

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2407.12442.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.12442 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:58:56.316279Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T21:43:28.684801Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1d73ca2e-2d06-433e-bf66-3189fc8f657f · inbound

ReCLIP++: Learn to Rectify the Bias of CLIP for Unsupervised Semantic Segmentation cites this paper.

ReCLIP++: Learn to Rectify the Bias of CLIP for Unsupervised Semantic Segmentation ClearCLIP: Decomposing CLIP Representations for Dense Vision-Language Inference

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-23T21:43:28.691444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-23T21:42:34.305164Z digest=sha256:36e8626898e373b5a840737c4d04fa1aee1edb8cfa6d8719e936751fb12dbd60

Observation dd7ce8bf-859c-472a-8597-aa47c66b4327 · inbound

Rethinking the Global Knowledge of CLIP in Training-Free Open-Vocabulary Semantic Segmentation cites this paper.

Rethinking the Global Knowledge of CLIP in Training-Free Open-Vocabulary Semantic Segmentation ClearCLIP: Decomposing CLIP Representations for Dense Vision-Language Inference

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:35:23.278056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-23T04:33:51.883348Z digest=sha256:badc7801d97b18103c670f7ec8d241e60b2e1190178d6811485e3147ad5dea6d

Observation 41ac7d0c-10f7-4c94-82b0-dd1517596ae3 · inbound

Single Domain Generalization for Few-Shot Counting via Universal Representation Matching cites this paper.

Single Domain Generalization for Few-Shot Counting via Universal Representation Matching ClearCLIP: Decomposing CLIP Representations for Dense Vision-Language Inference

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:58:56.316279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:58:56.316279Z digest=sha256:418204d1d5c4527d02b1517f64cf9d5ed6121d3bf6aa9253af2f52de3e2cad11

Observation 8e55e445-7107-4088-b075-769e426d8bab · inbound

MARBLE: Material Recomposition and Blending in CLIP-Space cites this paper.

MARBLE: Material Recomposition and Blending in CLIP-Space ClearCLIP: Decomposing CLIP Representations for Dense Vision-Language Inference

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T10:26:00.474524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:26:00.474524Z digest=sha256:a0fbfa3d22e21e07045b0251721fe34d22c5e36b1ad72572d5c3d221639324cc

Observation 1a54c288-6e1e-4028-a99c-ad479f2bbcc4 · inbound

Bridge Feature Matching and Cross-Modal Alignment with Mutual-filtering for Zero-shot Anomaly Detection cites this paper.

Bridge Feature Matching and Cross-Modal Alignment with Mutual-filtering for Zero-shot Anomaly Detection ClearCLIP: Decomposing CLIP Representations for Dense Vision-Language Inference

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T17:27:15.609442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:27:15.609442Z digest=sha256:3baa64fbf1def1c03814ebfd9597e135dcfed35fffecad2b29ebad0e46fa6013

Observation c1621a8f-eb6d-481f-8c7a-ff0aab5ac66a · inbound

Annotation-Free Open-Vocabulary Segmentation for Remote-Sensing Images cites this paper.

Annotation-Free Open-Vocabulary Segmentation for Remote-Sensing Images ClearCLIP: Decomposing CLIP Representations for Dense Vision-Language Inference

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-05T16:41:45.551955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:41:45.551955Z digest=sha256:67eb93cc653bba31c975de3002a07814e48c36122683ade523da656f79c095ef

Observation 603f27bc-08f2-4cdf-ad3a-dea061a70b7d · inbound

Plug-in Feedback Self-adaptive Attention in CLIP for Training-free Open-Vocabulary Segmentation cites this paper.

Plug-in Feedback Self-adaptive Attention in CLIP for Training-free Open-Vocabulary Segmentation ClearCLIP: Decomposing CLIP Representations for Dense Vision-Language Inference

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:44.355578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:18:44.355578Z digest=sha256:18a2b612c770cb23c0bfbc5f8421fdeb9a77f394de23f44f2011081ebaa60d6f

Observation 8b2aff63-856e-425b-8163-be6df4b3c8f2 · inbound

SegEarth-OV3: Exploring SAM 3 for Open-Vocabulary Semantic Segmentation in Remote Sensing Images cites this paper.

SegEarth-OV3: Exploring SAM 3 for Open-Vocabulary Semantic Segmentation in Remote Sensing Images ClearCLIP: Decomposing CLIP Representations for Dense Vision-Language Inference

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-16T23:53:42.296301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T23:53:27.350335Z digest=sha256:ae0d900e5ac375c0523b9171f591884511e1a686e24e3ac82264fc520c716f39

Observation 33b53525-e1cc-405f-8c12-c178e005f3dd · inbound

The Hyperspherical Geometry of CLIP Latent Space: A Semantic Mixture Model cites this paper.

The Hyperspherical Geometry of CLIP Latent Space: A Semantic Mixture Model ClearCLIP: Decomposing CLIP Representations for Dense Vision-Language Inference

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T04:33:53.021547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:33:53.021547Z digest=sha256:f224960e9cf2e4910bf54901e1c3f883191b08b146b5c3f6c5c1304034ab4444