Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 45 inbound Pith citation observations for arXiv:2412.15188.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-08T17:08:55.803846Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 4281b462-436e-40f7-899c-6eae8f1ce502 · inbound
UniMoD: Efficient Unified Multimodal Transformers with Mixture-of-Depths LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82064ae3-1537-41f5-954b-ed80812f159e · inbound
I Think, Therefore I Diffuse: Enabling Multimodal In-Context Reasoning in Diffusion Models LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df39cbdd-3381-441e-a123-c2e256285cdc · inbound
WISE: A World Knowledge-Informed Semantic Evaluation for Text-to-Image Generation LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b00e0bf0-cad8-4fb7-9243-39b51f257aa3 · inbound
Transfer between Modalities with MetaQueries LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f1969a78-60b6-4416-b4b6-16c8b833dcbf · inbound
BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 51b2df35-9b9c-4a5f-92ae-eaf04c263e21 · inbound
Emerging Properties in Unified Multimodal Pretraining LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 05dac12b-2ec1-4114-b49e-baa9bdda25cf · inbound
UniRL: Self-Improving Unified Multimodal Models via Supervised and Reinforcement Learning LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e7650c6-7a75-4f80-a7ee-22d476024b72 · inbound
OpenUni: A Simple Baseline for Unified Multimodal Understanding and Generation LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31f5f6e2-7a61-446a-af72-fef9c92bf846 · inbound
Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a98f8182-ac5b-424c-b2ac-401da3970f5a · inbound
LaTtE-Flow: Layerwise Timestep-Expert Flow-based Transformer LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58e5e446-5684-49ed-a9fb-a57566674917 · inbound
MADFormer: Mixed Autoregressive and Diffusion Transformers for Continuous Image Generation LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a9b2bc1-764f-4d42-96fb-7ae08c69dbe3 · inbound
Dreamland: Controllable World Creation with Simulator and Generative Models LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73a3cffc-4bfb-4dea-88c1-15afa9fa889f · inbound
Show-o2: Improved Native Unified Multimodal Models LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 95
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 301bff5f-135a-4629-bd06-220f1bcd1fea · inbound
OmniGen2: Towards Instruction-Aligned Multimodal Generation LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation cd09d3a2-d46b-49ef-be2a-5baa8f857fa3 · inbound
WordCon: Word-level Typography Control in Scene Text Rendering LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72d32ee3-09c9-486d-a573-65be95ab85cf · inbound
Rethinking Discrete Tokens: Treating Them as Conditions for Continuous Autoregressive Image Synthesis LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2c473d0-10b9-4e03-938d-35c50e31c5ef · inbound
FreeLoRA: Enabling Training-Free LoRA Fusion for Autoregressive Multi-Subject Personalization LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 594a0dbc-5e29-48b0-89c5-62155ea44e51 · inbound
X-Omni: Reinforcement Learning Makes Discrete Autoregressive Image Generative Models Great Again LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b17a1063-74ee-4200-8211-e01266c29ff6 · inbound
Galaxea Open-World Dataset and G0 Dual-System VLA Model LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93875204-2588-4310-b563-7bab02756f4c · inbound
Interleaving Reasoning for Better Text-to-Image Generation LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b562f445-e176-474e-9c2d-9b13fbd5371d · inbound
Reconstruction Alignment Improves Unified Multimodal Models LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef364fbd-f8fd-4b0c-bf37-103f5d74a52b · inbound
UniVideo: Unified Understanding, Generation, and Editing for Videos LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 527a0908-9afd-40e7-91e4-9cdb4c2ca112 · inbound
SRUM: Fine-Grained Self-Rewarding for Unified Multimodal Models LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70e91f46-7163-4878-a42c-6e83fdf87912 · inbound
FasterVAR: Plug-and-Play Acceleration for Visual Autoregressive Models LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71837f0d-f365-4e9a-814e-146d94d08b85 · inbound
CG-MLLM: Captioning and Generating 3D content via Multi-modal Large Language Models LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8854c87b-93c1-45fc-926c-f848d8caded8 · inbound
ChatUMM: Robust Context Tracking for Conversational Interleaved Generation LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c89b769c-3ee7-42e6-8247-0d57e1dcebd4 · inbound
LLaMo: Scaling Pretrained Language Models for Unified Motion Understanding and Generation with Continuous Autoregressive Tokens LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation aa3453fd-dc74-4795-95b3-f947d0d1f384 · inbound
Demystifying Video Reasoning LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f54f030-7751-43f7-9d1a-645e1db41827 · inbound
Demystifying Video Reasoning LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10617339-23d3-4317-89c2-12d670fc9617 · inbound
LMGenDrive: Bridging Multimodal Understanding and Generative World Modeling for End-to-End Driving LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b0800302-88df-4745-bf7b-dd720fc3f116 · inbound
Free Lunch for Unified Multimodal Models: Enhancing Generation via Reflective Rectification with Inherent Understanding LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 06ef9691-e5ef-45bd-8e01-14279bf16a8f · inbound
Co-generation of Layout and Shape from Text via Autoregressive 3D Diffusion LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c235ae80-05e8-4e10-8822-a65b8dec7d74 · inbound
MMCORE: MultiModal COnnection with Representation Aligned Latent Embeddings LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 66ee8cee-fdc4-41e5-ab3a-66281e56cb1d · inbound
Meta-CoT: Enhancing Granularity and Generalization in Image Editing LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 51bd0ddd-055f-465b-a524-64859693240b · inbound
SpatialFusion: Endowing Unified Image Generation with Intrinsic 3D Geometric Awareness LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 73068788-af9b-4a9b-89b8-001858cd0320 · inbound
STARFlow2: Bridging Language Models and Normalizing Flows for Unified Multimodal Generation LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c63185a6-3916-499f-8a62-05a744587d3b · inbound
Reversing the Flow: Generation-to-Understanding Synergy in Large Multimodal Models LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8fe1fa5e-1efa-4d0a-a3b9-8971c93eae2c · inbound
Semantic Generative Tuning for Unified Multimodal Models LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e554ff05-2944-404f-b193-242c8dc4740e · inbound
Semantic Generative Tuning for Unified Multimodal Models LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 86284e46-aff6-4509-b9a9-68049f183b84 · inbound
Where to Refine, When to Stop: Rethinking Redundancy via Latent Discrepancy for Efficient Visual Autoregressive Generation LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a6e04f3e-7839-49c8-b13a-3a891a308702 · inbound
Polaris: Scaling Up Instruction-Guided Image Generation Towards Millions of Personalized Style Needs LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 74f42b3d-7ebf-4053-b30b-be90316c1160 · inbound
UniTac: A Unified Multimodal Model for Cross-Sensor Tactile Understanding and Generation LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 583ff840-3bba-428a-894e-007a3b67114d · inbound
Streaming Multi-Agent Autoregressive Diffusion Model with World State Registers LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 126
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c8c05c8-dcfb-46f2-961d-2d9147766e61 · inbound
Towards Physics of Multimodal Pretraining: Knowledge Flow, Modality Synergy, Early Unification, and Recipes LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 100
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b47279d-1775-4e41-80f8-d3ed0ac8850b · inbound
Towards Physics of Multimodal Pretraining: Knowledge Flow, Modality Synergy, Early Unification, and Recipes LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 100
Source-reported events for the cited work
Unavailable: canonical work link unavailable.