Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-03T01:02:35.739213Z
Paper Citation Record · LEDGER
As of 5 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 0 inbound Pith citation observations for arXiv:2602.10809.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-03T01:02:35.739213Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
39 of 39 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation e703c280-ad41-4414-b796-56356ad3d700 · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Introducing claude opus 4.5, 2025 a
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c3cd124-f102-4b22-b40c-74306fe7aab0 · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Introducing claude sonnet 4.5, 2025 b
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b349e52a-5994-4ac3-97de-6beed00b7849 · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Qwen3-VL Technical Report
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e93fd0d-42b5-4259-bd2b-04f63bc3beaf · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Seed1.6-embedding
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12d175b6-6b6f-42b9-9584-86eab0bb3b58 · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories MoCa: Modality-aware Continual Pre-training Makes Better Bidirectional Multimodal Embeddings
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf2732b4-67c7-4f6a-95c7-0b56b7d21034 · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories mme5: Improving multimodal multilingual embeddings via high-quality synthetic data
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3526ed00-8740-452d-a0e9-f03ea2bb870b · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Generative thinking, corrective action: User-friendly composed image retrieval via automatic multi-agent collaboration
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fdce8250-ad50-4153-8923-823f3152cbf8 · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories EmbodiedEval: Evaluate Multimodal LLMs as Embodied Agents
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 445417bb-7f86-409f-ba75-5d3366f6297d · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories N., Awasthi, A., Pan, X., Ahuja, C., Mishra, S
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83553ad4-7500-4cc2-b70d-d83a80b48ce5 · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Mind2web: Towards a generalist agent for the web
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a166859-1437-4990-b6e9-a382c3b0fc2b · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Colpali: Efficient document retrieval with vision language models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1fe1adda-0ff8-4ee8-810d-001bc1a4585f · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories A new era of intelligence with Gemini 3
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20dd5299-2df9-4aa2-b4a3-347238f418c1 · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Mind2Web 2: Evaluating Agentic Search with Agent-as-a-Judge
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02532e51-fb6a-4e92-8739-f284b0ef4f0b · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories GPT-4o System Card
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f24c0db-43fc-4f5d-bd7b-a6602f470807 · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories V., Sung, Y., Li, Z., and Duerig, T
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31d3e9fd-d003-4240-b176-73b5528a99b8 · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Vlm2vec: Training vision-language models for massive multimodal embedding tasks
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc69d2e7-97ed-4dad-90db-293002dc2100 · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Crafting papers on machine learning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15aa4e4e-bea2-4fb9-bdb9-a3d9b3927088 · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Unresolved cited work
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ded2419-e48c-4485-958d-a9b5d6c385d7 · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Qwen3-VL-Embedding and Qwen3-VL-Reranker: A Unified Framework for State-of-the-Art Multimodal Retrieval and Ranking
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32b42022-5d2f-4b8e-9ab1-c7de86b259c1 · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories MM-BrowseComp: A Comprehensive Benchmark for Multimodal Browsing Agents
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38ffbf4b-2e8d-4ba8-ae1a-df91b7c7e20e · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Mm-embed: Universal multimodal retrieval with multimodal LLMS
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16166113-7826-420e-87a1-7f1c84adfb6b · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories VLM2Vec-V2: Advancing Multimodal Embedding for Videos, Images, and Visual Documents
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ac97c6c-3015-4164-ae98-6d82a0434c9c · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Introducing GPT-5.2 : The most advanced frontier model for professional work and longrunning agents
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1f48c7d-b670-41b2-9661-828190fdb8db · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., Krueger, G., and Sutskever, I
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b387237-72f7-43ce-87b4-bb5e9b8e8374 · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Glm-4.5v and glm-4.1v-thinking: Towards versatile multimodal reasoning with scalable reinforcement learning, 2025
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36d67bfd-7b05-455f-8669-7ddc259737ba · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories A., Friedland, G., Elizalde, B., Ni, K., Poland, D., Borth, D., and Li, L
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b1c5e39-0d5b-436d-98fc-ec0b8f51f774 · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Multimodal Reasoning Agent for Zero-Shot Composed Image Retrieval
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef29b82e-b6f1-423d-ab3d-17a186663cef · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Uniir: Training and benchmarking universal multimodal information retrievers
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9fe77158-cac5-4238-a4e8-22e0d2f817b1 · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories J., Cheng, Z., Shin, D., Lei, F., Liu, Y., Xu, Y., Zhou, S., Savarese, S., Xiong, C., Zhong, V., and Yu, T
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc4dfd74-5688-4ac3-9319-be2578710aad · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories A survey on agentic multimodal large language models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 980d3209-a6fa-4191-80b3-f613ddd9f0b7 · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories A Survey on Multimodal Large Language Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49d1665b-b035-4f00-a2a9-87125fc03ff8 · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Sigmoid loss for language image pre-training
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b10a885-a4ba-4dfb-95e3-06b3ed909129 · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Magiclens: Self-supervised image retrieval with open-ended instructions
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e00c3eb1-e0f5-4c09-8fb1-7a4805c5d85b · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories GME: Improving Universal Multimodal Retrieval by Multimodal LLMs
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b411d52-5745-4bfb-862b-b6cb155c6133 · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories V-MAGE: A Game Evaluation Framework for Assessing Vision-Centric Capabilities in Multimodal Large Language Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ef789e4-2270-4b6f-a5bc-0b33acecb204 · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories J., and Lian, D
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ac36528-f433-4fc8-b53f-0c0e46b615e9 · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories F., Zhu, H., Zhou, X., Lo, R., Sridhar, A., Cheng, X., Ou, T., Bisk, Y., Fried, D., Alon, U., and Neubig, G
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b43014d-2eb1-4f58-95f2-60bce1e64fa0 · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Large language models for information retrieval: A survey
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 933a7374-a2f8-4049-b55b-4cb929bf10af · outbound
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories write newline
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.