Pith. sign in

Paper Citation Record · LEDGER

Vision-by-Language for Training-Free Compositional Image Retrieval

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2310.09291.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2310.09291 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T05:54:15.956672Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T16:27:09.023654Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation d2cafc08-83e6-44b9-a357-09c7a8aa6376 · inbound

E5-V: Universal Embeddings with Multimodal Large Language Models cites this paper.

E5-V: Universal Embeddings with Multimodal Large Language Models Vision-by-Language for Training-Free Compositional Image Retrieval

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-16T22:52:20.974516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T22:52:20.935555Z digest=sha256:8459bda186ef0622a6c4837baaa674dd48f8fd4a53899c531f6a4863e8636ff1

Observation a2540e4b-15b0-4ffc-b70e-47acf8f5fe44 · inbound

UniCoRN: Unified Commented Retrieval Network with LMMs cites this paper.

UniCoRN: Unified Commented Retrieval Network with LMMs Vision-by-Language for Training-Free Compositional Image Retrieval

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-08T05:54:15.956672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T05:54:15.956672Z digest=sha256:83ea409a7ff1004cbeec112624dcc477afd3504f32e8fbd091add902d70960ed

Observation 703b7aba-149f-44df-9c13-06255edc21b9 · inbound

DetailFusion: A Dual-branch Framework with Detail Enhancement for Composed Image Retrieval cites this paper.

DetailFusion: A Dual-branch Framework with Detail Enhancement for Composed Image Retrieval Vision-by-Language for Training-Free Compositional Image Retrieval

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T14:43:59.484503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:43:59.484503Z digest=sha256:b31683f1ea25bb4900be79660149022d1035db8695856c89daaa4ee26c7e6bcc

Observation 085e0c1d-7a55-444c-b67d-40552bca0c92 · inbound

FACap: A Large-scale Fashion Dataset for Fine-grained Composed Image Retrieval cites this paper.

FACap: A Large-scale Fashion Dataset for Fine-grained Composed Image Retrieval Vision-by-Language for Training-Free Compositional Image Retrieval

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T19:10:00.073729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:10:00.073729Z digest=sha256:adc8aa7d44816cf2e1b124b5b2a05452821974a74e398f5d9cadffb926d0f8d1

Observation b8ea230b-1b02-4063-89f3-427c1cb390ea · inbound

MCoT-RE: Multi-Faceted Chain-of-Thought and Re-Ranking for Training-Free Zero-Shot Composed Image Retrieval cites this paper.

MCoT-RE: Multi-Faceted Chain-of-Thought and Re-Ranking for Training-Free Zero-Shot Composed Image Retrieval Vision-by-Language for Training-Free Compositional Image Retrieval

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T16:42:40.156102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:42:40.156102Z digest=sha256:371065dad1c581be16289ce0b5a0074d83c898cf975ee76af28de1bc8a74d259

Observation cda78990-b7ed-4ae7-9ea9-05acbf4101ff · inbound

Beyond Simple Edits: Composed Video Retrieval with Dense Modifications cites this paper.

Beyond Simple Edits: Composed Video Retrieval with Dense Modifications Vision-by-Language for Training-Free Compositional Image Retrieval

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T18:50:58.144723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:50:58.144723Z digest=sha256:7bc7ae74b36986f0e3c99e455b88135ef61d80aab092c19ccb499b01bcf0afda

Observation d93c6bb1-380b-46be-a1ee-b864874ec20d · inbound

A Sanity Check on Composed Image Retrieval cites this paper.

A Sanity Check on Composed Image Retrieval Vision-by-Language for Training-Free Compositional Image Retrieval

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:36:05.594025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T15:24:27.497284Z digest=sha256:916d1c7ee995e884260bc70945dbc815a1d8f9dbb97ecf05393fa0277b493ad7

Observation 943d49a4-0d22-4e0a-a289-188b5de38372 · inbound

G-MIXER: Geodesic Mixup-based Implicit Semantic Expansion and Explicit Semantic Re-ranking for Zero-Shot Composed Image Retrieval cites this paper.

G-MIXER: Geodesic Mixup-based Implicit Semantic Expansion and Explicit Semantic Re-ranking for Zero-Shot Composed Image Retrieval Vision-by-Language for Training-Free Compositional Image Retrieval

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:50:19.952322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T11:50:00.850857Z digest=sha256:2866d73ac2c9454049116a40397087f4d6b17447a1649742fce79a4d0ca2ef58

Observation d8923319-ac51-4b3d-b01c-9f7b0179ea20 · inbound

STiTch: Semantic Transition and Transportation in Collaboration for Training-Free Zero-Shot Composed Image Retrieval cites this paper.

STiTch: Semantic Transition and Transportation in Collaboration for Training-Free Zero-Shot Composed Image Retrieval Vision-by-Language for Training-Free Compositional Image Retrieval

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:33:58.857727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T05:30:08.145613Z digest=sha256:bc37fcca9589070ec80205fe5152f55b55f06e2a6c05b6d88789849201ad8872

Observation cac05558-6818-457b-b095-d2c2dfe52d60 · inbound

Never Seen Before: Benchmarking Genuine Zero-Shot Composed Image Retrieval with Consistent Video-Sourced Datasets cites this paper.

Never Seen Before: Benchmarking Genuine Zero-Shot Composed Image Retrieval with Consistent Video-Sourced Datasets Vision-by-Language for Training-Free Compositional Image Retrieval

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-02T16:27:09.025375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T22:44:36.059582Z digest=sha256:b2c6e39bed1b751fd14483868f7ac9b3751d5392a29e755e73f3edb7af93602c

Observation bbc903ec-637b-47e2-b7f3-8f0cebfa60bc · inbound

CR-Refiner: An Object-Centric Optimal Transport Reranker for Edit-Conditioned 3D Scene Retrieval cites this paper.

CR-Refiner: An Object-Centric Optimal Transport Reranker for Edit-Conditioned 3D Scene Retrieval Vision-by-Language for Training-Free Compositional Image Retrieval

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T13:28:40.030166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:28:40.030166Z digest=sha256:8f3d0e7f4e97b58694a7a859334d46b2b7b0cd703482a27b24195e11d013607e