Pith. sign in

Paper Citation Record · LEDGER

Rethinking Text-Based Image Retrieval in Specific Domain

As of 16 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 0 inbound Pith citation observations for arXiv:2608.10524.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.10524 v1

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T14:23:02.416523Z

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

17 of 17 outbound references displayed

  • verified exact0
  • verified fuzzy3
  • unresolved14
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3e2f9ccd-7aa1-41ba-bb9c-0b1b6bbf2840 · outbound

This paper cites jina-embeddings-v4: Universal Embeddings for Multimodal Multilingual Retrieval.

Rethinking Text-Based Image Retrieval in Specific Domain jina-embeddings-v4: Universal Embeddings for Multimodal Multilingual Retrieval

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.359256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.359256Z digest=sha256:066ee43669d04364b159a24cbd8d3c50a002d4ddcc047109fe1c0863a9b77d38

Observation 3d1a5540-a3b5-4556-a3ee-d2cd1a8f0b59 · outbound

This paper cites arXiv:2510.12798.

Rethinking Text-Based Image Retrieval in Specific Domain arXiv:2510.12798

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.378328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.378328Z digest=sha256:206f35d81082c1c33f605f51d898c460b95a5cc83a7a9c42fe3f427d5dc3837b

Observation 2a8bfb3a-4832-45a1-bde0-5d95e0a3e28f · outbound

This paper cites Qwen3-VL-Embedding and Qwen3-VL-Reranker: A Unified Framework for State-of-the-Art Multimodal Retrieval and Ranking.

Rethinking Text-Based Image Retrieval in Specific Domain Qwen3-VL-Embedding and Qwen3-VL-Reranker: A Unified Framework for State-of-the-Art Multimodal Retrieval and Ranking

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.382460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.382460Z digest=sha256:28b5ad68b08b5c609ba47d7651f34387f6fb516f214cf27ab1fac83afd4ab0a7

Observation 4fdafe87-c3a2-4063-aece-58d4500ae319 · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

Rethinking Text-Based Image Retrieval in Specific Domain DINOv2: Learning Robust Visual Features without Supervision

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.394978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.394978Z digest=sha256:88988ac6d4a8f1a0a7d225b2051af1873267b3367b0ef556644e63c4b1d18729

Observation 2b4c60c8-a704-4b3c-a2a7-f02aa086b66f · outbound

This paper cites DINOv3.

Rethinking Text-Based Image Retrieval in Specific Domain DINOv3

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.399110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.399110Z digest=sha256:9abb5254808c467aef4a41144e674e87d6b8ee3a00d2588f4c7d864cc9741a03

Observation 0b3638de-b2f7-44ec-88c5-5deb2570c30f · outbound

This paper cites SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features.

Rethinking Text-Based Image Retrieval in Specific Domain SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.403717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.403717Z digest=sha256:1fa2bb2c7d0c3a7d4170df05ac2627c51fcf085bf3e4dd169a5fd69f948148c5

Observation f7e661eb-b70a-42f9-85e1-b675b0a4ffa6 · outbound

This paper cites When and why vision-language models behave like bags-of-words, and what to do about it?.

Rethinking Text-Based Image Retrieval in Specific Domain When and why vision-language models behave like bags-of-words, and what to do about it?

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.412149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.412149Z digest=sha256:4db082ac8cf9a11e12d704f24e3d2be0e15113d462ed0acbb3789e6798055c98

Observation ec61eab2-8c51-4266-afc2-ad636b4317c9 · outbound

This paper cites PLIP: Language-Image Pre-training for Person Representation Learning.

Rethinking Text-Based Image Retrieval in Specific Domain PLIP: Language-Image Pre-training for Person Representation Learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.416523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.416523Z digest=sha256:46c78e7af8d12ee80d12ae23c8dd978821d026f59821e5771d2bfa161a138d28

Observation 3887df14-79ad-4484-9829-6ef6c9ccd390 · outbound

This paper cites Loshchilov,I.;andHutter,F.2019.

Rethinking Text-Based Image Retrieval in Specific Domain Loshchilov,I.;andHutter,F.2019

Reference 755

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:23:02.843138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T14:23:02.391050Z digest=sha256:464394c9bde93cb5695f1a4c9039d44046db08857bf8de22b8815f5b890ca310

Observation 912a4f61-9f0f-450d-b93c-320f2191e70d · outbound

This paper cites InComputer Vision– ECCV 2014: 13th European Conference, Zurich, Switzer- land, September 6-12, 2014, Proceedings, Part V 13, 740–.

Rethinking Text-Based Image Retrieval in Specific Domain InComputer Vision– ECCV 2014: 13th European Conference, Zurich, Switzer- land, September 6-12, 2014, Proceedings, Part V 13, 740–

Reference 2014

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:23:02.856676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T14:23:02.386965Z digest=sha256:b148f5d70e3d62c5e7699597d1567be44d4458a60e68ca6972915fef98a44e36

Observation 0661cb2a-c744-424b-9b69-11ac60e9a177 · outbound

This paper cites Automatic Spatially-aware Fashion Concept Discovery.

Rethinking Text-Based Image Retrieval in Specific Domain Automatic Spatially-aware Fashion Concept Discovery

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.363970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.363970Z digest=sha256:2edb7e775f0bf5964daa00e62ae6dc8aaf05908f7fb7285d2ed12811f0cf1794

Observation 8fe333e9-05dd-4a34-9b72-628ffeaa3e19 · outbound

This paper cites InProceedings of the International Conference on Machine Learning (ICML), 4904–4916.

Rethinking Text-Based Image Retrieval in Specific Domain InProceedings of the International Conference on Machine Learning (ICML), 4904–4916

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.374137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.374137Z digest=sha256:084870a927f82e9f77e733a7234ca26ffd0ca82b5ecf2376c11cfd866fd91d57

Observation 5d455686-0178-496b-a77c-152bfa02d758 · outbound

This paper cites InProceedings of the 2022 Conference on Empirical Methods in Natural Language Processing, 3876–3887.

Rethinking Text-Based Image Retrieval in Specific Domain InProceedings of the 2022 Conference on Empirical Methods in Natural Language Processing, 3876–3887

Reference 2022

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:23:02.829637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T14:23:02.408018Z digest=sha256:b1cf16f7e79cfd7edb1a328a144705a986e97d5f28b14e83226e35c07e486b2a

Observation 9f2b1b85-6a13-40a9-9293-72d1a9a4f6d1 · outbound

This paper cites Semantically Self-Aligned Network for Text-to-Image Part-aware Person Re-identification.

Rethinking Text-Based Image Retrieval in Specific Domain Semantically Self-Aligned Network for Text-to-Image Part-aware Person Re-identification

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.350551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.350551Z digest=sha256:b2b1d1eef0cef96931793a237077d20e6f6168733931e4ae1b6fc94ae2a3891e

Observation b0064054-8f47-4ae7-8996-b2b2f8ce3188 · outbound

This paper cites InProceedings of the AAAI ConferenceonArtificialIntelligence,volume38,1860–1868.

Rethinking Text-Based Image Retrieval in Specific Domain InProceedings of the AAAI ConferenceonArtificialIntelligence,volume38,1860–1868

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.355262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.355262Z digest=sha256:4d9090984402472bd556c81963f2dc96e2ef6a79bd1c9ef6b6eb56c05295cfe5

Observation d3e4e7d7-997b-4ada-b890-2a30680b3923 · outbound

This paper cites Qwen3-VL Technical Report.

Rethinking Text-Based Image Retrieval in Specific Domain Qwen3-VL Technical Report

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.345821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.345821Z digest=sha256:487975b6ab4e6450ce565f37a8d33bba283ed184747ce999ec7a360046480fe7

Observation 67d2c312-7f7a-459f-bd96-0aae8c262a64 · outbound

This paper cites Jia, C.;Yang, Y.; Xia,Y.; Chen, Y.-T.;Parekh, Z.; Pham,H.; Le, Q.; Sung, Y.-H.; Li, Z.; and Duerig, T.

Rethinking Text-Based Image Retrieval in Specific Domain Jia, C.;Yang, Y.; Xia,Y.; Chen, Y.-T.;Parekh, Z.; Pham,H.; Le, Q.; Sung, Y.-H.; Li, Z.; and Duerig, T

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.369622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.369622Z digest=sha256:7e1512435496a04bdf4dc59181a6e26f4e1b38da42a13f5ae9fb5472566c6b5e

Pith citing papers

No inbound Pith citation observations are available.