Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T22:01:21.610269Z
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 26 of 26 outbound references and 34 inbound Pith citation observations for arXiv:2506.23044.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T22:01:21.610269Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T00:31:05.911595Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-05T16:51:14.255860Z
26 of 26 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation b1f26940-a111-4509-9044-f93b68e15997 · outbound
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1dee3b2c-2436-46d1-b713-42165f67ae27 · outbound
Ovis-U1 Technical Report Ming-Lite-Uni: Advancements in Unified Architecture for Natural Multimodal Interaction
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bfebae62-49c6-4f63-9b2a-40380c2b9677 · outbound
Ovis-U1 Technical Report HunyuanVideo: A Systematic Framework For Large Video Generative Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7f30d90-fd70-4685-adca-9ceaab2d380d · outbound
Ovis-U1 Technical Report Generating multi-image synthetic data for text-to-image customization
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff295d1e-38c2-4eb2-8bb2-3530cd5c1afd · outbound
Ovis-U1 Technical Report FLUX.1 Kontext: Flow Matching for In-Context Image Generation and Editing in Latent Space
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1c27ebc-14f2-4f97-ba9e-1b17c909838e · outbound
Ovis-U1 Technical Report UniWorld-V1: High-Resolution Semantic Encoders for Unified Visual Understanding and Generation
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de2cc211-29d7-4ada-92a8-c56fefde7204 · outbound
Ovis-U1 Technical Report Step1X-Edit: A Practical Framework for General Image Editing
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6966aa3f-f1fa-4f2b-9869-9325d6035023 · outbound
Ovis-U1 Technical Report MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2aad56e1-aa05-4038-b813-07d979938afa · outbound
Ovis-U1 Technical Report Ovis: Structural Embedding Alignment for Multimodal Large Language Model
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4c83747-2718-486e-b673-aa6780cd42c5 · outbound
Ovis-U1 Technical Report Exploring the Role of Large Language Models in Prompt Encoding for Diffusion Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 832f4d8a-2664-4068-8465-4905474e6b9b · outbound
Ovis-U1 Technical Report OminiControl: Minimal and Universal Control for Diffusion Transformer
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19406f29-e030-43b2-a19e-d4750f808bb6 · outbound
Ovis-U1 Technical Report Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df0d8873-5fd2-4ce5-9671-a01b95f3a8e0 · outbound
Ovis-U1 Technical Report OmniGen2: Towards Instruction-Aligned Multimodal Generation
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8cf80178-8f4f-4ce6-b773-27d2b2e069ee · outbound
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6b919d7-06ec-46af-943d-2ae1d285fe96 · outbound
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34755a93-0a9f-4799-a54b-94c234da339c · outbound
Ovis-U1 Technical Report ImgEdit: A Unified Image Editing Dataset and Benchmark
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37e62b70-b871-4fba-8415-04f74c3b61b2 · outbound
Ovis-U1 Technical Report Unified multimodal understanding and generation models: Advances, challenges, and opportunities
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27412aec-328a-47fb-9fbc-581d57699664 · outbound
Ovis-U1 Technical Report InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f4da6bf-77e9-4a35-829a-8cbaf1ad1744 · outbound
Ovis-U1 Technical Report SEED-Data-Edit Technical Report: A Hybrid Dataset for Instructional Image Editing
Reference 2016
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b14cbdff-083d-445e-8ddb-7ce25a85e739 · outbound
Ovis-U1 Technical Report UniControl: A Unified Diffusion Model for Controllable Visual Generation In the Wild
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0cc0b9b-74db-4cd5-b9c9-2db2a0241973 · outbound
Ovis-U1 Technical Report StyleBooth: Image Style Editing with Multimodal Instruction
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7fffe3c9-409d-4de7-ae26-296c0c759451 · outbound
Ovis-U1 Technical Report ELLA: Equip Diffusion Models with LLM for Enhanced Semantic Alignment
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c54a4482-8029-4f7e-8993-d4e00c5153ac · outbound
Ovis-U1 Technical Report BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99b5d5ba-653a-4696-b463-c2707af7e4f3 · outbound
Ovis-U1 Technical Report Emerging Properties in Unified Multimodal Pretraining
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d02e96e-c3c7-4bd1-86f5-1a7e2fc38f73 · outbound
Ovis-U1 Technical Report A diagram is worth a dozen images
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation cda295d1-e7e2-4768-962b-fe867de61575 · outbound
Ovis-U1 Technical Report Scalable Vision Language Model Training via High Quality Data Curation
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c05e8fc-c730-486f-bdc0-f7ccc969fa2d · inbound
X-Omni: Reinforcement Learning Makes Discrete Autoregressive Image Generative Models Great Again Ovis-U1 Technical Report
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68f62def-fac4-4a54-b902-6a6326680544 · inbound
Reconstruction Alignment Improves Unified Multimodal Models Ovis-U1 Technical Report
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4e1981a-d8f5-4e2e-a892-9b6253aac7b7 · inbound
Few-Shot Synthetic Image Attribution: Identifying Unseen Generators with Limited Samples Ovis-U1 Technical Report
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 388417b9-e942-4903-b2e5-7420ab00a381 · inbound
InfoTok: Information-Theoretic Regularization for Capacity-Constrained Shared Visual Tokenization in Unified MLLMs Ovis-U1 Technical Report
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation c63b1aef-16f2-4797-ac9a-c00c190a6fbf · inbound
PlanViz: Evaluating Planning-Oriented Image Generation and Editing for Computer-Use Tasks Ovis-U1 Technical Report
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation c0b3b0ae-925d-4c11-a52b-13e69920a4fd · inbound
OmniFysics: Towards Physical Intelligence Evolution via Omni-Modal Signal Processing and Network Optimization Ovis-U1 Technical Report
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 72272e35-bc7a-4e93-98b7-74e476b2700d · inbound
UniICL: Systematizing Unified Multimodal In-context Learning through a Capability-Oriented Taxonomy Ovis-U1 Technical Report
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f735d185-2c74-49ba-98d9-a19ea1605fe2 · inbound
TRACE: High-Fidelity 3D Scene Editing via Tangible Reconstruction and Geometry-Aligned Contextual Video Masking Ovis-U1 Technical Report
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c9f300d-95f8-4824-a1c9-a6074134a83c · inbound
Training-Free Image Editing with Visual Context Integration and Concept Alignment Ovis-U1 Technical Report
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 111eae68-f7db-4cd1-a53d-15684b85ba06 · inbound
Learning Preference-Based Objectives from Clinical Narratives for Dynamic Sepsis Treatment Ovis-U1 Technical Report
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 769cea6d-9b80-4643-bc85-90fff50cd8ce · inbound
Pseudo-Unification: Entropy Probing Reveals Divergent Information Patterns in Unified Multimodal Models Ovis-U1 Technical Report
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 58d72faf-f4a2-43e7-b1fa-44eb7d8d53e1 · inbound
UniGeo: Unifying Geometric Guidance for Camera-Controllable Image Editing via Video Models Ovis-U1 Technical Report
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 6141e825-fa89-48f2-9b39-1e23e0800b4f · inbound
UniGeo: Unifying Geometric Guidance for Camera-Controllable Image Editing via Video Models Ovis-U1 Technical Report
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 1337908f-8e02-472f-b4a9-bad8915ae98a · inbound
UniGeo: Unifying Geometric Guidance for Camera-Controllable Image Editing via Video Models Ovis-U1 Technical Report
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 975c441b-446c-404d-b0be-d206e36050e7 · inbound
UniCSG: Unified High-Fidelity Content-Constrained Style-Driven Generation via Staged Semantic and Frequency Disentanglement Ovis-U1 Technical Report
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation d6f94271-ac7c-42dd-993c-e634b5a8b48d · inbound
Tuna-2: Pixel Embeddings Beat Vision Encoders for Multimodal Understanding and Generation Ovis-U1 Technical Report
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation ed7f43f2-3360-479d-bc81-2260ec90bc9f · inbound
Tuna-2: Pixel Embeddings Beat Vision Encoders for Multimodal Understanding and Generation Ovis-U1 Technical Report
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 3efdaf0f-fbd8-4c7d-9902-b255e18ebbb9 · inbound
UniPath: Adaptive Coordination of Understanding and Generation for Unified Multimodal Reasoning Ovis-U1 Technical Report
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation aadd1ce5-a9dd-4642-93a2-69c3c47df8ba · inbound
SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-unify Architecture Ovis-U1 Technical Report
Reference 130
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation bfd24f7b-55d3-4bca-9c8e-6d77ce01bf63 · inbound
Lance: Unified Multimodal Modeling by Multi-Task Synergy Ovis-U1 Technical Report
Reference 112
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 18da38ba-8842-46f5-a654-b8adc6741fc5 · inbound
Lance: Unified Multimodal Modeling by Multi-Task Synergy Ovis-U1 Technical Report
Reference 113
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation fc274802-491c-4c43-a684-3257e5aeeccf · inbound
ProductWebGen: Benchmarking Multimodal Product Webpage Generation Ovis-U1 Technical Report
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 02cbe20d-9d72-462f-b43a-dd01b43dd8d6 · inbound
Is This Edit Correct? A Multi-Dimensional Benchmark for Reasoning-Aware Image Editing Ovis-U1 Technical Report
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1aa08193-def3-42b2-85cf-1652be0eb18c · inbound
InterleaveThinker: Reinforcing Agentic Interleaved Generation Ovis-U1 Technical Report
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 86578b31-29d0-469b-b3f0-e0c6be404263 · inbound
ImageWAM: Do World Action Models Really Need Video Generation, or Just Image Editing? Ovis-U1 Technical Report
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation ee6c667c-5e32-4925-b143-4c580b0bc9b1 · inbound
Boogu-Image-0.1: Boosting Open Agentic Multimodal Generation via Understanding under a Minimal Budget Ovis-U1 Technical Report
Reference 127
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85c6cb5c-2878-4b36-bba8-fc6a7f6c33e4 · inbound
SciForma: Structure-Faithful Generation of Scientific Diagrams Ovis-U1 Technical Report
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7feddf75-2d9c-4cbd-9b90-0bcd4e1bda71 · inbound
Mage-Flow: An Efficient Native-Resolution Foundation Model for Image Generation and Editing Ovis-U1 Technical Report
Reference 94
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67dbe261-a3d7-43e7-9ac1-134dc8cd0f8f · inbound
Mage-VL: An Efficient Codec-Native Streaming Multimodal Foundation Model Ovis-U1 Technical Report
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 808d42c2-22d5-4d08-a6a9-fbfcffebc0da · inbound
Instruction-based Image Editing: A Survey on Data, Models, Evaluation, and Applications Ovis-U1 Technical Report
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cbf59592-eca8-4bfc-9597-a5d824d2a55b · inbound
Towards Physics of Multimodal Pretraining: Knowledge Flow, Modality Synergy, Early Unification, and Recipes Ovis-U1 Technical Report
Reference 123
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1086410-3241-4a04-985b-2c38ff909417 · inbound
Towards Physics of Multimodal Pretraining: Knowledge Flow, Modality Synergy, Early Unification, and Recipes Ovis-U1 Technical Report
Reference 123
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23ad3e4c-6cc5-435c-ad12-8a2e9c55ece3 · inbound
UniSpace: Unified Visual Representation and Scalable Multimodal Modeling Ovis-U1 Technical Report
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f9d7d50-5a78-435e-817c-0e09c284d389 · inbound
A Model-Internal Protocol for Assessing Multimodal Models as Integrated Systems Ovis-U1 Technical Report
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.