Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T06:02:30.468386Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 100 of 114 outbound references and 2 inbound Pith citation observations for arXiv:2506.06279.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T06:02:30.468386Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-11T01:49:15.136031Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-11T16:01:23.018232Z
100 of 114 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 79d773ee-f286-40d3-bd8a-4f084176938d · outbound
CoMemo: LVLMs Need Image Context with Image Memory write newline
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b37105c-5bfe-48d5-a435-308641a2b8b8 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Nocaps: Novel object captioning at scale
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7646e2aa-2a92-4878-a2de-6d3de73342b7 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Flamingo: a visual language model for few-shot learning
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f750bb2f-0dcb-4ec5-bf3c-732c7d48c15b · outbound
CoMemo: LVLMs Need Image Context with Image Memory MathQA: Towards Interpretable Math Word Problem Solving with Operation-Based Formalisms
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation def89704-aae5-45d8-a5ac-d1f50160c601 · outbound
CoMemo: LVLMs Need Image Context with Image Memory A., Datla, V
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e30b908-a5a0-41d9-9b70-8ab804a9e781 · outbound
CoMemo: LVLMs Need Image Context with Image Memory F., Tito, R., Mafla, A., Gomez, L., Rusinol, M., Valveny, E., Jawahar, C., and Karatzas, D
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8d49dac-16a7-4aa6-825c-c39d28fffe1e · outbound
CoMemo: LVLMs Need Image Context with Image Memory Coyo-700m: Image-text pair dataset
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a9ee46c-ef58-4c1f-ae7b-45ef3d6f2469 · outbound
CoMemo: LVLMs Need Image Context with Image Memory and Xiao, J
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05435e20-be77-4581-9970-3db59625d8a3 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Textocr-gpt4v
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea0cc5cf-0707-45eb-84f8-9d044095792f · outbound
CoMemo: LVLMs Need Image Context with Image Memory MapQA: A Dataset for Question Answering on Choropleth Maps
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e07abd60-3b54-47fa-9573-77bf4aa61664 · outbound
CoMemo: LVLMs Need Image Context with Image Memory ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59f67add-d782-4305-b878-8925ec32c11e · outbound
CoMemo: LVLMs Need Image Context with Image Memory UniGeo: Unifying Geometry Logical Reasoning via Reformulating Mathematical Expression
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae2a3243-366e-4ca9-bc2f-638aae7bf3b1 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa14a1d1-29be-4aeb-89cf-76468b19416b · outbound
CoMemo: LVLMs Need Image Context with Image Memory EVLM: An Efficient Vision-Language Model for Visual Understanding
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bbc65fc0-80bd-44e0-a9d6-e376db1066f4 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de2765e1-0281-4a57-8c39-37fe9e9ebe9e · outbound
CoMemo: LVLMs Need Image Context with Image Memory Complicated Table Structure Recognition
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b42769e-724d-4fc0-9d0c-1b09306a9c0f · outbound
CoMemo: LVLMs Need Image Context with Image Memory K., Liu, Y., Sun, Y., Ng, C
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd94c8a4-0aba-436e-a7a4-f13cea61dff6 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Simple and Effective Multi-Paragraph Reading Comprehension
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33df9bd5-09bf-4ca6-8a80-9c1818466996 · outbound
CoMemo: LVLMs Need Image Context with Image Memory NVLM: Open Frontier-Class Multimodal LLMs
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff4a96ac-f9d2-4b61-bde0-f07a2c0002e3 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Deep visual template-free form parsing
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1af4d204-7229-42ac-9f8c-b5f8dca7c0f4 · outbound
CoMemo: LVLMs Need Image Context with Image Memory The Llama 3 Herd of Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ae6ab93-5751-4fb2-98a7-9ddc968e5796 · outbound
CoMemo: LVLMs Need Image Context with Image Memory MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 763727de-789a-4026-a5f0-6c930777ce2c · outbound
CoMemo: LVLMs Need Image Context with Image Memory A., Ma, W.-C., and Krishna, R
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4544efaf-0813-4a67-9ba4-cbd5ad03d612 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Making the v in vqa matter: Elevating the role of image understanding in visual question answering
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12fc4599-f778-4b46-a339-87cc2d6b3369 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Wukong: A 100 million large-scale chinese cross-modal pre-training benchmark
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52e019b4-ae1e-4d76-a27d-ed70e62f19f6 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Eaten: Entity-aware attention for single shot visual text extraction
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2430febf-ee02-49cd-b57f-6ac84cc8b47f · outbound
CoMemo: LVLMs Need Image Context with Image Memory Icpr2018 contest on robust reading for multi-type web images
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 209f2471-82e4-450e-8fe7-45bb600d254e · outbound
CoMemo: LVLMs Need Image Context with Image Memory PathVQA: 30000+ Questions for Medical Visual Question Answering
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 669db1a4-b54e-49e4-8bfe-4d8b997731fb · outbound
CoMemo: LVLMs Need Image Context with Image Memory Unresolved cited work
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ca98ae8-e705-436f-9692-368ec67430e2 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Koniq-10k: An ecologically valid database for deep learning of blind image quality assessment
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13ce348c-6f4e-4ff0-a75c-36eed2ec4bee · outbound
CoMemo: LVLMs Need Image Context with Image Memory mPLUG-DocOwl 1.5: Unified Structure Learning for OCR-free Document Understanding
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c75405c-2f20-4413-a079-1700b6a00b72 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Medical-diff-vqa: a large-scale medical dataset for difference visual question answering on chest x-ray images, 2023
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation beed86e7-8ef0-4c38-b874-721f76f683c1 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Movienet: A holistic dataset for movie understanding
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3283491-6cde-4272-a8e9-6a8b47581690 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Icdar2019 competition on scanned receipt ocr and information extraction
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4dfa8682-7aa0-4224-8ea3-1847b10f9a60 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Unresolved cited work
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 174f77cd-8052-47e7-9770-9eba5b8372c6 · outbound
CoMemo: LVLMs Need Image Context with Image Memory MANTIS: Interleaved Multi-Image Instruction Tuning
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80d25d45-dc03-4fc4-a455-9dc0d7bf3c4c · outbound
CoMemo: LVLMs Need Image Context with Image Memory Clevr: A diagnostic dataset for compositional language and elementary visual reasoning
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5cdb447-51fa-47f8-a580-c592747f2c7e · outbound
CoMemo: LVLMs Need Image Context with Image Memory Dvqa: Understanding data visualizations via question answering
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8be9ef5-2188-4f56-a41a-7957667c4179 · outbound
CoMemo: LVLMs Need Image Context with Image Memory FigureQA: An Annotated Figure Dataset for Visual Reasoning
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1474023-82b6-4b32-be90-265dbf797445 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Chart-to-Text: A Large-Scale Benchmark for Chart Summarization
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8277c4b0-01fe-4f01-b02a-6ea0cf332505 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Referitgame: Referring to objects in photographs of natural scenes
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f63e000-09da-4d75-827b-ff1de325a7e9 · outbound
CoMemo: LVLMs Need Image Context with Image Memory A diagram is worth a dozen images
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1cc699c-fcae-4550-87ec-cf8c16069995 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Are you smarter than a sixth grader? textbook question answering for multimodal machine comprehension
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57ad395e-beeb-4c37-a87b-ed1205ca2522 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Visual information extraction in the wild: practical dataset and end-to-end solution
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 810c6589-e670-4112-a806-0b314790c112 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Laion-gpt4v dataset
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dfc59059-e7aa-4d95-a722-6d38b63e3d24 · outbound
CoMemo: LVLMs Need Image Context with Image Memory J., Gayen, S., Ben Abacha, A., and Demner-Fushman, D
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a24351c-1ad4-46be-9297-e4b072587cfb · outbound
CoMemo: LVLMs Need Image Context with Image Memory What matters when building vision-language models?
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc7246f4-6e8f-4389-a348-e5da0dbb0458 · outbound
CoMemo: LVLMs Need Image Context with Image Memory G., and Lov \'o n Melgarejo, J
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36f591ec-7866-4929-8473-e3c6ff0d02ff · outbound
CoMemo: LVLMs Need Image Context with Image Memory Chemvlm: Exploring the power of multimodal large language models in chemistry area
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9322a4e8-564e-40c4-a543-064dc566e5e2 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Unresolved cited work
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd1e125a-4572-4e4e-b84a-48a5cf860a21 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Vila: On pre-training for visual language models
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0efdde03-aabb-4c5e-9981-273822210940 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Unresolved cited work
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01704a5f-c125-49bc-881a-eb17e9f8b86c · outbound
CoMemo: LVLMs Need Image Context with Image Memory Unresolved cited work
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99dea663-00ac-46ea-babb-7b35d882486d · outbound
CoMemo: LVLMs Need Image Context with Image Memory Slake: A semantically-labeled knowledge-enhanced dataset for medical visual question answering
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b189077f-4b5e-4840-ad8e-bea1bbe31d7d · outbound
CoMemo: LVLMs Need Image Context with Image Memory Casia online and offline chinese handwriting databases
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 946e01d2-f093-45ca-a824-c0bf5ccc3cc5 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Visual spatial reasoning
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b07f0ca7-744d-4623-b1e8-e0f98c805451 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Mitigating hallucination in large multi-modal models via robust instruction tuning
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 17c71ad7-9f4f-4159-9ce9-1a6764bd694f · outbound
CoMemo: LVLMs Need Image Context with Image Memory MMC: Advancing Multimodal Chart Understanding with Large-scale Instruction Tuning
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f00e6b2e-9063-4c1a-8cd1-2a7c6a92d1a5 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Unresolved cited work
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 604930f5-0b88-4c59-95d0-1ee410d9de32 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Unresolved cited work
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8546b904-8a57-41ab-a652-42fed20649b7 · outbound
CoMemo: LVLMs Need Image Context with Image Memory F., Lin, K., Hewitt, J., Paranjape, A., Bevilacqua, M., Petroni, F., and Liang, P
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90241c51-b471-4207-a6b0-416a39b7b9e8 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Paying more attention to image: A training-free method for alleviating hallucination in lvlms
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 733505ed-4a4d-490f-9430-7783d7a73374 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Mmbench: Is your multi-modal model an all-around player? In European conference on computer vision, pp.\ 216--233
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5695cc77-06fc-4aca-a66e-e9a988521821 · outbound
CoMemo: LVLMs Need Image Context with Image Memory MMDU: A Multi-Turn Multi-Image Dialog Understanding Benchmark and Instruction-Tuning Dataset for LVLMs
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3b3ff6b-8025-4c8a-98d7-14cdf5e688b2 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Inter-GPS: Interpretable Geometry Problem Solving with Formal Language and Symbolic Reasoning
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e92e147-407f-4e5b-8b8f-1e49ef1e8d61 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Learn to explain: Multimodal reasoning via thought chains for science question answering
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4f30c604-bac2-41b0-9774-f0e2832b62cd · outbound
CoMemo: LVLMs Need Image Context with Image Memory Dynamic Prompt Learning via Policy Gradient for Semi-structured Mathematical Reasoning
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 156f6993-0a26-418b-bb6f-867c05b90e02 · outbound
CoMemo: LVLMs Need Image Context with Image Memory MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f661c154-e39b-42b0-980a-4f0457f3041a · outbound
CoMemo: LVLMs Need Image Context with Image Memory Deepart: Learning joint representations of visual arts
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c77c4483-ffa4-4447-b75b-c3042ff0f6a1 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Ok-vqa: A visual question answering benchmark requiring external knowledge
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dfa4350b-4f3c-46ca-9875-93bcd26166eb · outbound
CoMemo: LVLMs Need Image Context with Image Memory and Bunke, H
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 18c6213e-1545-47cb-916a-ab4fe68274be · outbound
CoMemo: LVLMs Need Image Context with Image Memory L., Tan, J
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 137e982f-9a0d-49a5-8a9e-134d0c8a0659 · outbound
CoMemo: LVLMs Need Image Context with Image Memory ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be2f4648-53c9-4e91-8ab0-fc059295805c · outbound
CoMemo: LVLMs Need Image Context with Image Memory UniChart: A Universal Vision-language Pretrained Model for Chart Comprehension and Reasoning
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6bcc7c4-897b-429b-ad41-109aa8beff90 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Infographicvqa
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9ff00e35-d786-456c-8814-5f0078660c84 · outbound
CoMemo: LVLMs Need Image Context with Image Memory M., and Kumar, P
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3ecfc09d-1ebb-4353-a622-1da18e7a13fe · outbound
CoMemo: LVLMs Need Image Context with Image Memory Opengvlab/internvl-chat-v1-2-sft-data, Jan 2024
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 63479621-6e1a-4838-8694-93568d5fa6e7 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Training language models to follow instructions with human feedback
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c1e3acc-6270-4b13-aa75-e1b031b50306 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Kosmos-2: Grounding Multimodal Large Language Models to the World
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01a410e3-8cf4-4fae-b973-5d36686e14b5 · outbound
CoMemo: LVLMs Need Image Context with Image Memory A., Wang, L., Cervantes, C
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1cc2ab6b-5754-4d27-8241-182a2418c82a · outbound
CoMemo: LVLMs Need Image Context with Image Memory Laion-5b: An open large-scale dataset for training next generation image-text models
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bb4b8522-8214-418c-a2ff-1c533289bb58 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Laion coco: 600m synthetic captions from laion2b-en
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1b84a50c-ef90-4ba9-bdef-b3783f87e3ad · outbound
CoMemo: LVLMs Need Image Context with Image Memory Solving geometry problems: Combining text and diagram interpretation
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fbe9a7b8-59c9-41b3-99d9-a4dfb01ad904 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Unresolved cited work
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1a3ad93e-ed8e-4a09-9d05-0bf9b7515aec · outbound
CoMemo: LVLMs Need Image Context with Image Memory Objects365: A large-scale, high-quality dataset for object detection
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0a2e401d-22de-4408-b203-8ce576a0d7e5 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Towards vqa models that can read
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation be159f74-34f0-4efd-9ca0-9bfb09e11a71 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Towards vqa models that can read
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3179f98c-0cf1-44ad-ab44-ee10b2d0dbce · outbound
CoMemo: LVLMs Need Image Context with Image Memory Textocr: Towards large-scale end-to-end reasoning for arbitrary-shaped scene text
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d554ce3f-0191-4126-aca7-caf3d993cddc · outbound
CoMemo: LVLMs Need Image Context with Image Memory MileBench: Benchmarking MLLMs in Long Context
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15374b95-0c4e-410f-af9a-b0853b967627 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Transformer roadmap: 2
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9d98cf6c-e5e3-4c5c-980b-b5d1a4852d4f · outbound
CoMemo: LVLMs Need Image Context with Image Memory C., Han, J., Ding, E., Liu, J., Karatzas, D., et al
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 18ec7216-01ed-42d5-85ed-341a4c6cd833 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Internvl2: Better than the best—expanding performance boundaries of open-source multimodal models with the progressive scaling strategy, 2024
Reference 93
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 295ed525-12db-4af6-bcb7-6dfc706f61ef · outbound
CoMemo: LVLMs Need Image Context with Image Memory Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 94
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1f5b464-6f14-461f-aedd-9d6cbb42dc60 · outbound
CoMemo: LVLMs Need Image Context with Image Memory COCO-Text: Dataset and Benchmark for Text Detection and Recognition in Natural Images
Reference 95
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f93a81d-03ac-4199-8d57-73741ee9121c · outbound
CoMemo: LVLMs Need Image Context with Image Memory V3det: Vast vocabulary visual detection dataset
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ae545fb5-9c31-4421-9b57-625e6aeba795 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Measuring Multimodal Mathematical Reasoning with MATH-Vision Dataset
Reference 97
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c69544c-f78e-4907-bb82-9e20ca304583 · outbound
CoMemo: LVLMs Need Image Context with Image Memory Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 98
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74ab545b-117c-4047-824a-57ee643cddd5 · outbound
CoMemo: LVLMs Need Image Context with Image Memory The All-Seeing Project: Towards Panoptic Visual Recognition and Understanding of the Open World
Reference 99
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f7dfd5e-5b7a-4f5b-90cb-ac18b352031d · outbound
CoMemo: LVLMs Need Image Context with Image Memory Needle In A Multimodal Haystack
Reference 100
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60177022-5d53-47f2-baf6-075177bef644 · outbound
CoMemo: LVLMs Need Image Context with Image Memory C., Luo, C., Jin, L., Chan, C
Reference 101
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 141c552d-38db-4546-9c1b-cf0bfe83f149 · inbound
Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs CoMemo: LVLMs Need Image Context with Image Memory
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a6c78e64-9adc-4225-94c0-35ecef142c6c · inbound
Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs CoMemo: LVLMs Need Image Context with Image Memory
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.