Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:08:08.869281Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 59 of 59 outbound references and 0 inbound Pith citation observations for arXiv:2505.22613.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:08:08.869281Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
59 of 59 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 90463c61-bf5e-443d-8851-0b8611e30f46 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction Flamingo: a Visual Language Model for Few-Shot Learning
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc5ff3fa-1c18-47dd-b333-692732d69e67 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction SPICE: Semantic Propositional Image Caption Evaluation
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 714fbf1d-f07e-4025-aea3-127cae2dfbc4 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction Qwen Technical Report
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc6bcb00-ea6c-45ef-a8e1-9e2c5364cc79 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e803da03-b8c5-45fa-9602-7b9c93be180c · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction Hallucination of Multimodal Large Language Models: A Survey
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e97fa7e0-4965-4cb8-8b22-5fc9a85af19b · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction Unresolved cited work
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e059b8d5-4d35-4768-99d0-84a0df065497 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction ShareGPT4V: Improving Large Multi-Modal Models with Better Captions
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5cb60e30-a5c0-4049-9cb6-7d47de177c7f · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction Gonzalez, Ion Stoica, and Eric P
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec6b558f-1628-4982-9668-2ca3c7f3a36b · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 285fe182-3d78-4df4-860c-79512ddbfdd4 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction Benchmarking and Improving Detail Image Caption
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93f5842e-428b-43d3-8fb6-f97863357a38 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction A Survey on In-context Learning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0697e11-d0e0-4a02-872b-e8111bb91916 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction Improving CLIP Training with Language Rewrites
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1c9293f-1173-4e62-b229-6156df03b9d9 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction Unresolved cited work
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9b1095e-54f8-46ee-9570-65993b2f0c2d · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction GenEval: An Object-Focused Framework for Evaluating Text-to-Image Alignment
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f17a7886-2c4c-4fc3-b881-22885e2569bc · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction CLIPScore: A Reference-free Evaluation Metric for Image Captioning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a683d7db-d212-4e2f-92ee-460c5bb7e6bc · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction CogVLM2: Visual Language Models for Image and Video Understanding
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e721d09-db49-4d03-ad35-f21644d1adc5 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction LoRA: Low-Rank Adaptation of Large Language Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67873c65-c2b7-47b4-93ef-2441ad0781a2 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction ELLA: Equip Diffusion Models with LLM for Enhanced Semantic Alignment
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab514049-398b-44d1-9f7b-4630afa36023 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-image Generation
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65f33dcf-9ea4-497b-a508-52d258668a74 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction Unresolved cited work
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e7fa958-1490-4002-81da-fac08bbf48bf · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction Unresolved cited work
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae4bd084-619c-4959-a82d-421bd37155db · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction VeCLIP: Improving CLIP Training via Visual-enriched Captions
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4af59865-4ae8-42ba-a9c6-40fe3663865b · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a14b41b1-eede-4c43-b21c-bf8d2717c0e5 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction BLIP: Bootstrapping Language-Image Pre-training for Unified Vision-Language Understanding and Generation
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50aa06cf-5d1b-49a6-8515-95435e9fd6d2 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction FACTUAL: A Benchmark for Faithful and Consistent Textual Scene Graph Parsing
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f82f29cd-725b-4492-89dd-b8a0c02d954b · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction Unresolved cited work
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f79b9c52-9e3f-46e8-9e9a-96587df0c713 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction Evaluating Text-to-Visual Generation with Image-to-Text Generation
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25002f22-4f2d-4fbf-9222-f49370bb2c88 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27ef73e7-64f7-4ca8-92d9-a7a406674719 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction Improved Baselines with Visual Instruction Tuning
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3c8e907-8dcc-4ff2-93ed-c5653997fe54 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction Visual Instruction Tuning
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35c61650-6da7-4968-a50d-1fe8be2d38ea · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction Decoupled Weight Decay Regularization
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 050d1fec-5a17-4cbb-a3dd-6661b90d05d4 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction Benchmarking Large Vision-Language Models via Directed Scene Graph for Comprehensive Image Captioning
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7fbd89d8-0707-44bc-b076-3e5942c8d3fb · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction GPT-4o System Card
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e12b594c-8055-4723-88d4-332c0e529590 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction Unresolved cited work
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1f97d57-0233-458c-8d8c-b6f5ad7bcb87 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction Training language models to follow instructions with human feedback
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e3b42fd-d060-407b-961f-0a91638baa29 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction Unresolved cited work
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8243aab-30e4-4d78-8449-60e07b5b33d0 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction Patch Matters: Training-free Fine-grained Image Caption Enhancement via Local Perception
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b96e1fe6-ade3-4e55-99f6-14941409d313 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction Unresolved cited work
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db01e3cc-8589-40c3-811e-42278e299bee · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0bbf7527-2e0a-43dd-8862-846ec2812208 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction TimeChat: A Time-sensitive Multimodal Large Language Model for Long Video Understanding
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b91d8922-fafe-4c54-8071-bc1d40044a68 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction Unresolved cited work
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a85e9afc-a2df-4d8a-9b1a-719b8539f95b · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction Gemini: A Family of Highly Capable Multimodal Models
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24b18474-4d6c-4d2b-a31f-3b23fe409bb8 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1fe5ad3a-fbac-417a-9cd5-dfb1c964319b · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction LLaMA: Open and Efficient Foundation Language Models
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f58cc58-b22b-45a3-ad0c-4314b2fb9870 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction CIDEr: Consensus-based Image Description Evaluation
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30f5fc9a-3a27-4395-af93-e472c428b700 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction AMBER: An LLM-free Multi-dimensional Benchmark for MLLMs Hallucination Evaluation
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da252b43-a893-4478-abe5-2eb5a55d4d6e · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10e2b154-614f-429e-bd7e-aceb10d84581 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction CogVLM: Visual Expert for Pretrained Language Models
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0946a2de-47c2-4566-92ca-7f82c54c7637 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction LaDiC: Are Diffusion Models Really Inferior to Autoregressive Counterparts for Image-to-Text Generation?
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c8d4c1e2-bb6e-428d-984e-240281321dc1 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction Chain-of-Thought Prompting Elicits Reasoning in Large Language Models
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b48f0b6e-904e-436b-ace1-e253840de87c · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction Altogether: Image Captioning via Re-aligning Alt-text
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a1361b2-4b68-4d2e-adf4-9a1239c4f5a1 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction Unresolved cited work
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e07b8e6d-61fa-4e78-9970-982a6a4e66c1 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction CapEnrich: Enriching Caption Semantics for Web Images via Cross-modal Pre-trained Knowledge
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 264c1c5c-0e8a-4328-a5a9-f06044edce12 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction Painting with Words: Elevating Detailed Image Captioning with Benchmark and Alignment Learning
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 282cb1db-1d9d-4e8f-86a7-92e62621b0dd · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction CapsFusion: Rethinking Image-Text Data at Scale
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b41ccc6b-4fa4-493f-9796-3b6bb1bd2b2f · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction LlamaFactory: Unified Efficient Fine-Tuning of 100+ Language Models
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5967b2d-3470-487a-9840-053c41c5eff8 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38de8f72-b661-4f69-92fa-c19e8eda71ef · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction online" 'onlinestring :=
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b10c9615-855e-404a-ac3e-97799e4bf9d0 · outbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction write newline
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.