Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T20:01:50.856287Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 0 inbound Pith citation observations for arXiv:2507.04151.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T20:01:50.856287Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
32 of 32 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 36ea33ba-720d-4b0d-a522-6a82941ee6aa · outbound
Unlocking Compositional Control: Self-Supervision for LVLM-Based Image Generation Draw ALL Your Imagine: A Holistic Benchmark and Agent Framework for Complex Instruction-based Image Generation
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0237661-a292-42e9-b8b0-f7c521d28f7a · outbound
Unlocking Compositional Control: Self-Supervision for LVLM-Based Image Generation Score: Story coherence and retrieval enhancement for ai narratives,
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 712b153c-1902-48dc-9661-2acf37a83993 · outbound
Unlocking Compositional Control: Self-Supervision for LVLM-Based Image Generation Enhancing Code LLMs with Reinforcement Learning in Code Generation: A Survey
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 471daa47-9b5e-439c-9265-31438a124170 · outbound
Unlocking Compositional Control: Self-Supervision for LVLM-Based Image Generation Diffusion models beat gans on image synthesis,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c1544d88-8c49-4e86-89c5-bce1d492e2cb · outbound
Unlocking Compositional Control: Self-Supervision for LVLM-Based Image Generation Triple sequence generative adversarial nets for unsupervised image captioning,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0380e3cc-e4dd-4c90-8ed4-29e8044e2766 · outbound
Unlocking Compositional Control: Self-Supervision for LVLM-Based Image Generation Visual in-context learning for large vision-language models,
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f51352c-1325-4fa7-8205-16168498195e · outbound
Unlocking Compositional Control: Self-Supervision for LVLM-Based Image Generation Weak to strong generalization for large language models with multi-capabilities,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c2e7e97-51ae-48d3-89ca-cb45602dea6d · outbound
Unlocking Compositional Control: Self-Supervision for LVLM-Based Image Generation Improving Medical Large Vision-Language Models with Abnormal-Aware Feedback
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e769d58f-c0b2-436b-b2a5-4884d2b33751 · outbound
Unlocking Compositional Control: Self-Supervision for LVLM-Based Image Generation LAION-400M: Open Dataset of CLIP-Filtered 400 Million Image-Text Pairs
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a64bd966-0714-446e-baf7-df3e2de89e81 · outbound
Unlocking Compositional Control: Self-Supervision for LVLM-Based Image Generation Conceptual captions: A cleaned, hypernymed, image alt-text dataset for automatic image cap- tioning,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e983ef3-f2c4-4479-84e6-ed3e42306a8d · outbound
Unlocking Compositional Control: Self-Supervision for LVLM-Based Image Generation Microsoft COCO: common objects in context,
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b51647f4-05da-472d-8aa9-a3e55bbafaea · outbound
Unlocking Compositional Control: Self-Supervision for LVLM-Based Image Generation Beyond empathy: Integrating diagnostic and therapeutic reasoning with large language models for mental health counseling,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4bb82b5-5d46-4de5-a61e-31fa8521f3f0 · outbound
Unlocking Compositional Control: Self-Supervision for LVLM-Based Image Generation Multimodal event transformer for image-guided story ending generation,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 52b8e76b-5685-4a07-9ed3-d7699c55c7d8 · outbound
Unlocking Compositional Control: Self-Supervision for LVLM-Based Image Generation An image is worth 16x16 words: Transformers for image recognition at scale,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 303dce78-eb93-4e1d-af3f-46553e5732e9 · outbound
Unlocking Compositional Control: Self-Supervision for LVLM-Based Image Generation Language models with image descriptors are strong few- shot video-language learners,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2b9ee77-6052-4fa3-a5ca-aa9cdc7d36a8 · outbound
Unlocking Compositional Control: Self-Supervision for LVLM-Based Image Generation Learning transferable visual models from natural language supervision,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6b2fe0d-0c95-403f-b6f8-29463105be68 · outbound
Unlocking Compositional Control: Self-Supervision for LVLM-Based Image Generation BLIP: bootstrapping language-image pre-training for unified vision-language understanding and generation,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11c69b06-b6ac-46c8-b741-6f2949ee603e · outbound
Unlocking Compositional Control: Self-Supervision for LVLM-Based Image Generation Flamingo: a visual language model for few-shot learning,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c8d1758d-90c1-4e67-b954-f884343bde56 · outbound
Unlocking Compositional Control: Self-Supervision for LVLM-Based Image Generation Cross-lingual transfer of large language model by visually- derived supervision toward low-resource languages,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bbb1fff5-f82a-4315-a287-ce1fe51623f6 · outbound
Unlocking Compositional Control: Self-Supervision for LVLM-Based Image Generation Coca: Contrastive captioners are image-text foundation models,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a8f4c49-5753-4a90-8d2b-349c24d96839 · outbound
Unlocking Compositional Control: Self-Supervision for LVLM-Based Image Generation Tx-llava: Large language and vision assistant for temporal changes in chest x-rays,
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0f02167-b11d-450b-84ee-efd249df8742 · outbound
Unlocking Compositional Control: Self-Supervision for LVLM-Based Image Generation MAM: Modular Multi-Agent Framework for Multi-Modal Medical Diagnosis via Role-Specialized Collaboration
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 862c957f-faa2-484a-8d51-e71cc023b4db · outbound
Unlocking Compositional Control: Self-Supervision for LVLM-Based Image Generation GLIDE: towards photorealistic image generation and editing with text-guided diffusion models,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 297a636c-188f-4afa-b695-eec47deb6d44 · outbound
Unlocking Compositional Control: Self-Supervision for LVLM-Based Image Generation Taming transformers for high-resolution image synthesis,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 58fbf2b1-988a-45fb-aa34-34272b0e1b4e · outbound
Unlocking Compositional Control: Self-Supervision for LVLM-Based Image Generation Hierarchical Text-Conditional Image Generation with CLIP Latents
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9bfb984-33e0-4f54-a912-407aa9600c49 · outbound
Unlocking Compositional Control: Self-Supervision for LVLM-Based Image Generation BLIP-2: bootstrapping language-image pre-training with frozen image encoders and large language models,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 23f68944-0b40-4f1c-8bc8-6956db206c4c · outbound
Unlocking Compositional Control: Self-Supervision for LVLM-Based Image Generation Self-rewarding large vision- language models for optimizing prompts in text-to-image generation,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a03b9272-b4da-457a-93c2-b024dea37138 · outbound
Unlocking Compositional Control: Self-Supervision for LVLM-Based Image Generation Evolvedirector: Approaching advanced text-to-image generation with large vision- language models,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d773c6ae-7ce0-4acc-90e3-fc520c889a19 · outbound
Unlocking Compositional Control: Self-Supervision for LVLM-Based Image Generation Improving Compositional Text-to-image Generation with Large Vision-Language Models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce90422c-0cd8-4b94-aee9-40bbe469c126 · outbound
Unlocking Compositional Control: Self-Supervision for LVLM-Based Image Generation 12 888–12 900
Reference 162
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36bc1f5d-8b4d-4268-81a7-3680fabc8de5 · outbound
Unlocking Compositional Control: Self-Supervision for LVLM-Based Image Generation 12 873–12 883
Reference 2021
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e42b8339-0647-416f-ab6f-96e8422d85ea · outbound
Unlocking Compositional Control: Self-Supervision for LVLM-Based Image Generation Available: http://papers.nips.cc/paper\ files/paper/2022/ hash/960a172bc7fbf0177ccccbb411a7d800-Abstract-Conference.html 11
Reference 2022
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
No inbound Pith citation observations are available.