Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-25T08:15:12.947854Z
Paper Citation Record · LEDGER
As of 12 August 2026, this Paper Citation Record lists 70 of 70 outbound references and 0 inbound Pith citation observations for arXiv:2506.03530.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-25T08:15:12.947854Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
70 of 70 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 95d9c0ea-7a99-48a9-9c9a-7f27333dbbcf · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Incomplete multimodality-diffused emotion recognition
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 81e53b55-8c3e-4c3c-b416-2906b9dfbca7 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Smil: Multimodal learning with severely missing modality
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f3fcc976-2e50-4d47-9ccd-826cc65c5303 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Are multi- modal transformers robust to missing modality?
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 591521bb-543d-4aaf-93cd-b1ef47aaa961 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? M3care: Learning with missing modalities in multimodal healthcare data
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 3591e894-f0cd-4280-8dc1-cc4a2518090c · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Emu3: Next-Token Prediction is All You Need
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 8aa2407d-110a-40d5-bafe-800413dd2d33 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 92f9d9f1-1d00-4d5d-97b5-8a65d31f697a · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? GPT-4o System Card
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation e35e6e17-188b-4144-bb42-4ec1b804f7d4 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Qwen2.5-Omni Technical Report
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f1f84e6f-9410-43e2-b209-f4aca9e639ee · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Show-o: One Single Transformer to Unify Multimodal Understanding and Generation
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation e7634bad-bdd2-4147-a13f-7b1be3b9624e · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Knowledge bridger: Towards training-free missing multi-modality completion
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 562bb69e-53cd-4995-9c54-67d5ce617449 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? High- resolution image synthesis with latent diffusion models
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 17e81f9e-19e0-4059-a506-e41c56fb8682 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Gen- erative adversarial text to image synthesis
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 81165763-dabb-4b4d-83fd-5a059ea595bf · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Hierarchical Text-Conditional Image Generation with CLIP Latents
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation fbe66dd3-b9f1-4061-9032-6aedf59e49c6 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Scaling Autoregressive Models for Content-Rich Text-to-Image Generation
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 833355d6-6bc7-45e7-b768-abac05b45178 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 23d4d13f-90ac-4812-8a5d-2f401024cf76 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Imagebind: One embedding space to bind them all
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f932e840-318a-45e1-9f6f-875b17f5d0a0 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Can We Generate Images with CoT? Let's Verify and Reinforce Image Generation Step by Step
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a8431434-d89a-4f89-a59b-cb58546d02bf · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? ComfyGen: Prompt-Adaptive Workflows for Text-to-Image Generation
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 27f16c1e-7fa0-443d-966b-2e4399a6cbc5 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Can Test-Time Scaling Improve World Foundation Model?
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f7291bbb-a83d-484b-94ec-91c2c5f66080 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Training strategies to handle missing modalities for audio-visual expression recognition
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 98f8b18d-b2f3-4f45-9aef-4ecb947d9439 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Deep partial multi-view learning
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 6aac7bab-27d2-41cc-be8f-d7806573f6e7 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Multi-modal learning with missing modality via shared-specific feature modelling
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 1306f14e-3cf1-4765-9f3f-1fb0693173d1 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Gcnet: Graph completion network for incomplete multimodal learning in conversation
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 2ef93831-50f4-415e-85ec-907762700387 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Found in translation: Learning robust joint representations by cyclic translations between modalities
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation e8c84306-a58c-4762-9674-9b57a8f996fd · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Multimodal prompting with missing modalities for visual recognition
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation add58491-431b-4211-8890-da59ada4d926 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Multimodal prompt learning with missing modalities for sentiment analysis and emotion recognition
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation e2f4b396-b82b-48c6-a569-c625b56b0796 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Multi-modal modality- masked diffusion network for brain mri synthesis with random modality missing
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b2101404-ec82-4372-8f1a-70f9e1900ce0 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Fgc2f-udiff: Frequency-guided and coarse-to-fine unified diffusion model for multi-modality missing mri synthesis
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation be771c8c-aed5-4680-a4db-49327325439f · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Qwen2.5 Technical Report
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 1db78713-cf77-4beb-bd20-28d49d14a03a · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? LLaMA: Open and Efficient Foundation Language Models
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 75032a97-bd9b-4ea7-bcdd-07c71ed1b506 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Unresolved cited work
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a5eeaf16-9350-4e60-8ff0-3ee6c0751137 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Stable audio open
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 8697f9f9-c670-42f7-a4ab-c026daef522a · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Audioldm 2: Learning holistic audio generation with self-supervised pretraining
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation dcc7b4c1-f8da-40f1-b550-492ab4deecfc · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Next-gpt: Any-to-any multimodal llm
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a32b42bb-0b00-453e-8f43-b79ace42a4fc · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Generative adversarial networks
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 1fb165c4-dc5e-4ef8-814c-6a963cfa0674 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Conditional Generative Adversarial Nets
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 305dd6a6-685e-4c9b-be12-83fff1068f11 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Large Scale GAN Training for High Fidelity Natural Image Synthesis
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c7b0bb86-372f-48cf-9581-89151f26efe8 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Classifier-Free Diffusion Guidance
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c4843018-e672-42d7-9409-4c395484750c · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? AMM-Diff: Adaptive Multi-Modality Diffusion Network for Missing Modality Imputation
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b5d2ff7b-b71d-42b3-91a3-db5ebee5e88e · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? MissDiff: Training Diffusion Models on Tabular Data with Missing Values
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b908170d-de34-46b1-ab09-f6274f8d5e66 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Generating with fairness: A modality-diffused counterfactual framework for incomplete multimodal recommendations
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 11215dfd-8e4d-494a-8f0b-3799655c7a20 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Agent AI: Surveying the Horizons of Multimodal Interaction
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation ade6246f-7ee3-4cc8-be8f-f7db2bcb9ff6 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Agent S: An Open Agentic Framework that Uses Computers Like a Human
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 43924d15-bd65-4640-85e1-d15346248ae5 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? A Comprehensive Survey of Agents for Computer Use: Foundations, Challenges, and Future Directions
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation ff716df3-c2ac-4cf8-b95c-34098fcf9d58 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Solving Math Word Problems via Cooperative Reasoning induced Language Models
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation ec7f7ddf-16f5-487e-94fa-ebd976594f1a · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Encouraging Divergent Thinking in Large Language Models through Multi-Agent Debate
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b8aa8cf8-10d0-488f-929d-34fe3433ce68 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 02427c67-1d3e-47fe-9033-d3a9a90df43a · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Agen- tic ai software engineer: Programming with trust
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 6a9032a2-fc2d-42b3-80f4-9d213b3ca745 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Building Living Software Systems with Generative & Agentic AI
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 067affa2-1d29-4a23-964a-b00040c00145 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Toolformer: Language models can teach themselves to use tools
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 4afbdc9a-44cf-48ed-9cd9-0edbd520751f · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Hugginggpt: Solving ai tasks with chatgpt and its friends in hugging face
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation dfdc6810-ce1a-4bac-8552-d320aa651c67 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Navgpt: Explicit reasoning in vision- and-language navigation with large language models
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 0ab2517b-68f6-48ab-8557-ccb0ef8e018c · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 77770db0-3625-4086-b62b-52bdcb2715db · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Mind2web: Towards a generalist agent for the web
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 75e366fe-e82b-49f6-8c59-fead979f0204 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Mllm-as-a-judge: Assessing multimodal llm-as- a-judge with vision-language benchmark
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f7d87118-9734-4b47-8731-438c378ff97d · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Qwen2.5-VL Technical Report
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 1029e453-ebe7-4fbd-8816-ed5d7afdc1e8 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Vggsound: A large- scale audio-visual dataset
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f1565391-2e04-4d59-bfa7-93f57bfb3891 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Msr-vtt: A large video description dataset for bridging video and language
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 35213dd1-9ad4-4139-83f4-d7a39c585e8d · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Audiocaps: Generating captions for audios in the wild
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 8910431c-8771-455b-826f-e19084f1d670 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Microsoft coco: Common objects in context
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 9ab16854-af37-4199-bba3-7fc35cfd0d89 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Gans trained by a two time-scale update rule converge to a local nash equilibrium
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation aa6201c1-5006-4f59-a19a-3b9d15bd1357 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Learning transferable visual models from natural language supervision
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d19118ca-92d3-48ff-bcc1-dade5239dcd9 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? From wer and ril to mer and wil: improved evaluation measures for connected speech recognition
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation e5dba6e3-7d84-4d71-a510-1bf3c92c8645 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Tasnet: time-domain audio separation network for real-time, single-channel speech separation
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b4a385fd-9de8-4333-824e-2fc364433eb1 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Perceptual evaluation of speech quality (pesq): An objective method for end-to-end speech quality assessment of narrow- band telephone networks and speech codecs
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 42014936-3b41-4833-9fdf-6179158f65fc · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Best Practices and Lessons Learned on Synthetic Data
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a784138f-fbcc-4a56-80a8-9d14f20a69fe · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Does Reinforcement Learning Really Incentivize Reasoning Capacity in LLMs Beyond the Base Model?
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f7dccf6c-b24b-4b1b-b01d-b5acb5812876 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 8b3a0c79-3445-43f2-a813-5c2e0971c014 · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? Lora: Low-rank adaptation of large language models
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c1fd2402-f37d-4ebd-a6aa-821fb6e344ea · outbound
How Far Are We from Generating Missing Modalities with Foundation Models? The Power of Scale for Parameter-Efficient Prompt Tuning
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
No inbound Pith citation observations are available.