Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-14T16:20:19.874172Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 0 inbound Pith citation observations for arXiv:2607.04605.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-14T16:20:19.874172Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
61 of 61 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation d69e9b98-ba0a-4fb7-8bb8-3afc01558228 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Proceedings of the 43rd International ACM SIGIR Conference on Research and Development in Information Retrieval , pages=
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3db9897-eddf-421d-913c-50aeaa3018d4 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Proceedings of the 2022 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies , pages=
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9abbfd0-264e-45b9-b99f-5ba875813938 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval ColPali: Efficient Document Retrieval with Vision Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 614baf50-3197-4e6c-a184-e83cc3385e28 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6459299-2443-4e5d-bf89-3fddb94388f4 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval International Conference on Learning Representations , year=
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72881047-42d5-48b4-ae9d-5f032e5f9f77 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Towards Storage-Efficient Visual Document Retrieval: An Empirical Study on Reducing Patch-Level Embeddings
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b44ce88-7803-4bd0-bb2d-2987fc49004f · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval International conference on machine learning , pages=
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5d226f0-ce47-4fe1-8b2f-30c48a48e1e0 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Advances in neural information processing systems , volume=
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7703ee07-54d3-43de-b8f2-4f0b2f26142f · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval International conference on machine learning , pages=
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc1234fc-5b86-48a6-be92-0163d25f7af0 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval International conference on machine learning , pages=
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e80213c8-c79a-44db-a381-bb22539273e3 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Proceedings of the 2021 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies , pages=
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ac8931e-7d9d-4a42-8e35-459216666c32 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Proceedings of the 44th International ACM SIGIR Conference on Research and Development in Information Retrieval , pages=
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09587334-a962-4f58-b348-dfc3f3705406 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval SPLADE v2: Sparse Lexical and Expansion Model for Information Retrieval
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48358023-8a27-4865-8203-861a2b2f349a · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Proceedings of the 31st ACM International Conference on Information & Knowledge Management , pages=
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1bdc37e-1621-412a-8f23-79f1bf2b187f · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Proceedings of the 48th international ACM SIGIR conference on research and development in information retrieval , pages=
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9266dbf4-b5f7-4987-9eda-3006827c5273 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9449c600-a6b1-46f4-970d-158ea7559a7c · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Findings of the Association for Computational Linguistics: NAACL 2025 , pages=
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39c3bc25-6d97-4a42-8a7d-5398cd0f68d6 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 670d5b3d-9b21-4e27-9358-59f543957974 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval ZipVL: Efficient Large Vision-Language Models with Dynamic Token Sparsification
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba733c62-eea7-47ec-ae6c-25dc6e4a0a15 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval CLIP Tricks You: Training-free Token Pruning for Efficient Pixel Grounding in Large VIsion-Language Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ac10463-127a-45d7-b11c-e78b436dfdff · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval The Eleventh International Conference on Learning Representations , year=
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 809f288d-ee8b-41b1-851b-d3d07b433fbd · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Qwen3-VL Technical Report
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12098874-ec8f-4ca9-988a-52382304447c · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Transactions of the Association for Computational Linguistics , volume=
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70a783c4-b83b-49c6-b3a1-bf9c97b06cb4 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Flickr30k Entities: Collecting Region-to-Phrase Correspondences for Richer Image-to-Sentence Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60b43be1-021a-4c63-8c88-cc95ab52d0b8 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval European Conference on Computer Vision , year =
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ad16f05-d915-4fa0-b319-46bfaae20e41 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision , year =
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0eb666d-a6bc-4003-b48a-ee38e51cb817 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Image Retrieval from Contextual Descriptions
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1cac47bf-ff03-43c1-868c-ccc2e7c41ce8 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , year =
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation add979f7-4818-4453-af91-c5d782d59624 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval European Conference on Computer Vision , year =
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8513aa62-b565-4949-a652-732b1da7900e · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval International conference on machine learning , pages=
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1735af3-c2dd-4dae-816b-f95839cfb5de · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b15c0d7a-21d6-4694-ab37-c72993e0f7f5 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Demystifying
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fce11417-3c5a-4d9b-8770-bb7346f16fab · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval EVA-CLIP: Improved Training Techniques for CLIP at Scale
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3284f6fb-af5d-41e2-abd0-011f00031f52 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Data Filtering Networks
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be27854f-dab2-4f0f-9551-9a0f353d78fc · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e11db6d0-a67e-4419-82ce-a54400f900a8 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval 2026 , url=
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a4f119f-b87f-4274-a54c-eb44ec712845 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval GME: Improving Universal Multimodal Retrieval by Multimodal LLMs
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2093b4af-f076-44b3-be8c-cac7b1b3fc25 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Unresolved cited work
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 89aff243-e945-455a-badc-a74de29dede0 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Proceedings of the AAAI conference on artificial intelligence , volume=
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6f69e71-bdc3-4e4c-be96-b8a33365b224 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Q-GroundCAM: Quantifying Grounding in Vision Language Models via GradCAM
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63c43d9d-f918-4abb-bc8d-6d9fe6106649 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval International Journal of Computer Vision , volume=
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e07c2b9-3d19-496d-9eb2-d3df3499d24a · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval European Conference on Computer Vision , pages=
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca750301-05a1-4f8f-b567-07d07ef89b58 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Proceedings of the IEEE conference on computer vision and pattern recognition , pages=
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33d6b6a4-3f95-482f-8501-340b202a0e67 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval International journal of computer vision , volume=
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1643704-35b0-4bae-a2fa-af4f52a4ac79 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics , pages=
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ada24e3-bb8a-437b-9214-f508d1bb776a · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Advances in Neural Information Processing Systems , year=
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bec6475e-e5dd-41ea-a8d4-a8d073976247 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Advances in Neural Information Processing Systems , volume=
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5f3c0f1-6f25-4cf9-b2af-10f80fbbf73b · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval International Conference on Learning Representations , year=
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 855dbae7-fd2f-492f-8af8-23e202e544a5 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval European Conference on Computer Vision , year=
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5fccd8b-8d63-44a8-ba83-20b216e24dae · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision , year=
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 507d476f-c002-4c6e-9e12-cd8785215232 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b592cfa8-028d-4429-84a4-754274ce5bb4 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval European Conference on Computer Vision , pages=
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffea91b0-0d09-4c19-9b2e-94e75984bced · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f1faae1-c479-4901-b8b4-7231f7803286 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Proceedings of the AAAI Conference on Artificial Intelligence , volume=
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd18aa89-36ae-415c-bdd9-d87f34618c89 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Towards Storage-Efficient Visual Document Retrieval: An Empirical Study on Reducing Patch-Level Embeddings
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8426ecb0-5c54-48a3-ac15-08cc4b119d94 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Hierarchical Patch Compression for ColPali: Efficient Multi-Vector Document Retrieval with Dynamic Pruning and Quantization
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ce0c88e-ec0b-4238-ad56-e06169975b3b · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval 2026 , eprint=
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8c7517a-ecdb-4aa0-8fe8-e89abb0f064e · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval arXiv preprint arXiv:2602.21202 , year=
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33f6898f-e8c1-46b6-a794-b7fb5ad3a4ab · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Representation Learning with Contrastive Predictive Coding
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6cf9eae-3f4a-466a-b2c9-dd54554f0904 · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Advances in Neural Information Processing Systems , volume=
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d298ddc2-cea4-40df-ae3d-2ae99672d71e · outbound
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval 2025 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) , pages=
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.