Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2410.09733.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-12T14:23:44.140834Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation e61be1d3-bed9-4b26-a8b1-bdf0a84e9d41 · inbound
FINECAPTION: Compositional Image Captioning Focusing on Wherever You Want at Any Granularity MMCOMPOSITION: Revisiting the Compositionality of Pre-trained Vision-Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72696188-7aeb-4108-aa46-b636a46f7385 · inbound
Visual Compositional Tuning MMCOMPOSITION: Revisiting the Compositionality of Pre-trained Vision-Language Models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation a27e9937-2230-4ba7-bf44-1d7987b2a080 · inbound
MMIG-Bench: Towards Comprehensive and Explainable Evaluation of Multi-Modal Image Generation Models MMCOMPOSITION: Revisiting the Compositionality of Pre-trained Vision-Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e34ec5b-9fbe-4501-974f-fa54928b9f1e · inbound
Point-RFT: Improving Multimodal Reasoning with Visually Grounded Reinforcement Finetuning MMCOMPOSITION: Revisiting the Compositionality of Pre-trained Vision-Language Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08c3f604-a780-46fe-b6b7-cb5b067c7cd4 · inbound
Impact of Pretraining Word Co-occurrence on Compositional Generalization in Multimodal Models MMCOMPOSITION: Revisiting the Compositionality of Pre-trained Vision-Language Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1cc8abd9-bf34-4c5a-b4da-50f6edddfa9c · inbound
Evaluating Compositional Generalisation in VLMs and Diffusion Models MMCOMPOSITION: Revisiting the Compositionality of Pre-trained Vision-Language Models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eca0d95a-d76d-416e-bf40-aa99013712fe · inbound
ChartNet: A Million-Scale, High-Quality Multimodal Dataset for Robust Chart Understanding MMCOMPOSITION: Revisiting the Compositionality of Pre-trained Vision-Language Models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 27372a24-07a9-4c29-9881-909e16265e62 · inbound
MementoGUI: Learning Agentic Multimodal Memory Control for Long-Horizon GUI Agents MMCOMPOSITION: Revisiting the Compositionality of Pre-trained Vision-Language Models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 94c3e4ce-b0ce-468f-a1b9-62f70045d21c · inbound
Agent Skills Should Go Beyond Text: The Case for Visual Skills MMCOMPOSITION: Revisiting the Compositionality of Pre-trained Vision-Language Models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 026b55e2-a04a-44d3-be46-718b85d65819 · inbound
MemoBench: Benchmarking World Modeling in Dynamically Changing Environments MMCOMPOSITION: Revisiting the Compositionality of Pre-trained Vision-Language Models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 30318801-6c61-45bb-a607-f13110a5bcf6 · inbound
MemoBench: Benchmarking World Modeling in Dynamically Changing Environments MMCOMPOSITION: Revisiting the Compositionality of Pre-trained Vision-Language Models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 7104efbb-ce55-4f2d-ab61-b9bca81830da · inbound
MemoBench: Benchmarking World Modeling in Dynamically Changing Environments MMCOMPOSITION: Revisiting the Compositionality of Pre-trained Vision-Language Models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 4aafa573-09b4-42ca-982f-898331afc419 · inbound
MemoBench: Benchmarking World Modeling in Dynamically Changing Environments MMCOMPOSITION: Revisiting the Compositionality of Pre-trained Vision-Language Models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 00aa0689-4103-4c6e-891b-8480f87b4465 · inbound
MemoBench: Benchmarking World Modeling in Dynamically Changing Environments MMCOMPOSITION: Revisiting the Compositionality of Pre-trained Vision-Language Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84daaa27-8626-4168-b9ee-cc3c1c2e9e5e · inbound
MemoBench: Benchmarking World Modeling in Dynamically Changing Environments MMCOMPOSITION: Revisiting the Compositionality of Pre-trained Vision-Language Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42c656a9-cb30-441c-9236-974d2d20ed31 · inbound
Learning to Compose: Revisiting Proxy Task Design for Zero-Shot Composed Image Retrieval MMCOMPOSITION: Revisiting the Compositionality of Pre-trained Vision-Language Models
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 7b4dbd12-f7ba-49db-8bed-d4a396b798c9 · inbound
VideoArgus: Agentic Rubric-Grounded Unified Evaluation for Video Generation and Editing MMCOMPOSITION: Revisiting the Compositionality of Pre-trained Vision-Language Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.