Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T14:34:02.117406Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 18 inbound Pith citation observations for arXiv:2502.06734.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T14:34:02.117406Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:21:41.308248Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-01T07:55:30.838800Z
35 of 35 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 3968fcc7-b0b5-41bb-b816-ec748c4a6f34 · outbound
Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual Inversion
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e8f7153-f15e-4aa8-a82d-36d60faf3cd5 · outbound
Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists TokenFlow: Consistent Diffusion Features for Consistent Video Editing
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9ed8e8d-e873-464c-8a7c-0fb7e7e326fe · outbound
Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ddf45ffd-7775-4269-a74d-f6ceaeacb56e · outbound
Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists commas and used as input prompts for Grounded-SAM2 (Liu et al., 2023a; Ravi et al., 2024)
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 56ceb27b-0a23-4396-b94a-19cb1c46bd9f · outbound
Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists VIVID-10M: A Dataset and Baseline for Versatile and Interactive Video Local Editing
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc93fe88-f379-40c1-834c-e3698e829b37 · outbound
Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists HQ-Edit: A High-Quality Dataset for Instruction-based Image Editing
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e5ceaf7-b428-4401-8fea-8974b1755fe7 · outbound
Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists Video Diffusion Models are Strong Video Inpainter
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a110379d-fa15-4e05-bd0f-be30a42b816b · outbound
Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists Stablev2v: Stablizing shape consistency in video-to-video editing
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65508319-28ab-45a2-9db6-7b693d1e3e26 · outbound
Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists The video on the left depicts the original video, while the video on the right displays the edited videos
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2943a5e1-0cbd-4409-a54b-0ca3edd5d155 · outbound
Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists SDEdit: Guided Image Synthesis and Editing with Stochastic Differential Equations
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 872c7789-ff4e-46c1-b721-5248140c7282 · outbound
Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists ReVideo: Remake a Video with Motion and Content Control
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be7ebc54-aa6f-48b9-b111-c69517bea17c · outbound
Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists Zero-shot image-to-image translation
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c20652ba-8f08-4c35-964a-287171ab34c9 · outbound
Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists The 2017 DAVIS Challenge on Video Object Segmentation
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c8114e1-09fb-4ad0-8faf-59d53d2c1a9c · outbound
Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists DreamFusion: Text-to-3D using 2D Diffusion
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb7bec9a-b80e-4707-93b8-610c8f73af7e · outbound
Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists FateZero: Fusing Attentions for Zero-shot Text-based Video Editing
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3fc7e47a-e520-4b09-be19-aa1676ca7d34 · outbound
Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists SAM 2: Segment Anything in Images and Videos
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 157abfa0-e1ee-43ce-992a-c0f3a5034dc2 · outbound
Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists Plug- and-play diffusion features for text-driven image-to-image translation
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 77b33201-cf32-4aad-a065-9ea7f40b1021 · outbound
Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists OmniEdit: Building Image Editing Generalist Models Through Specialist Supervision
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d671e7a-02ea-40a6-bbd1-da6aacf34d06 · outbound
Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists Zhang, K., Mo, L., Chen, W., Sun, H., and Su, Y
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc54726e-b55e-448e-a843-fcd2e0f400be · outbound
Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists CoCoCo: Improving Text-Guided Video Inpainting for Better Consistency, Controllability and Compatibility
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c28dd6a3-e0f2-4b05-ba95-b98c0f9cde40 · outbound
Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists This limitation prevents us from applying techniques such as ControlNet to repaint a video effectively
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 19183171-3c17-41da-8f31-ae79636fa012 · outbound
Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists Table 5 shows that our expert model outperforms all baselines, achieving the lowest Ewarp (9.02), highest CLIPScore (0.3145), and best Temporal Consistency (0.9781)
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5d8ed50c-0bc1-46d7-be75-99d457685518 · outbound
Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists The best results are boldfaced
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 310b5d8b-27a2-45d2-8c50-e7828f4592c1 · outbound
Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists Bottom: The data construction pipeline for Señorita-2M using our inpainter
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 19ae2d9c-29ad-4530-a975-96bab3585997 · outbound
Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists Unresolved cited work
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4a10c795-5672-4f74-8382-50f9dcd94ef0 · outbound
Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists For object recognition, we utilize CogVLM-video-llama3-chat (Hong et al.,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 29e4c9da-64d4-4c24-b5ea-61148795e91a · outbound
Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists We set the maximum token length to 120 and use six frames per video
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f552e8c8-f3b5-4e00-a369-3fd06334b793 · outbound
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e6efe788-38df-4990-8af4-7998ff0f43bf · outbound
Reference 592
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e9b5e5a9-be66-4812-ba75-4cafddfc9199 · outbound
Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists DPM-Solver++: Fast Solver for Guided Sampling of Diffusion Probabilistic Models
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a8d4ad2-12ee-4734-8eb7-2b0ed1cf73de · outbound
Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists Hierarchical Text-Conditional Image Generation with CLIP Latents
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ccf3b396-b07e-4a72-b983-5689db8bcda2 · outbound
Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists CogVLM2: Visual Language Models for Image and Video Understanding
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cda3b44a-09aa-412d-9c2f-51313623c117 · outbound
Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists VideoCrafter2: Overcoming Data Limitations for High-Quality Video Diffusion Models
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4dbe9b51-1bda-4641-b1c8-78a8ddde0fed · outbound
Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists The Llama 3 Herd of Models
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 430243a2-6521-4863-93bf-61ca84c74603 · outbound
Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists HunyuanVideo: A Systematic Framework For Large Video Generative Models
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6467c033-e05e-4c69-ac81-3574c016c6ed · inbound
Step1X-Edit: A Practical Framework for General Image Editing Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 71e8ad13-ce70-4bc8-8281-4caf719b07be · inbound
MiniMax-Remover: Taming Bad Noise Helps Video Object Removal Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6148350f-6329-420b-b9f1-cefb4e840c79 · inbound
O-DisCo-Edit: Object Distortion Control for Unified Realistic Video Editing Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5dd8ca7-7353-4efb-966f-0828de6cccde · inbound
EditVerse: Unifying Image and Video Editing and Generation with In-Context Learning Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7253c983-e39d-4c33-a4fa-3823e6bc0878 · inbound
VideoCoF: Unified Video Editing with Temporal Reasoner Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3fa7ad1d-f7b6-47ba-9276-1a48340ccc55 · inbound
Under One Sun: Multi-Object Generative Perception of Materials and Illumination Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f6801e6-67af-49a5-a181-e1107e417339 · inbound
InsEdit: Towards Instruction-based Visual Editing via Data-Efficient Video Diffusion Models Adaptation Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e47282b7-5dac-4410-8ec7-7ffa7add2e6a · inbound
Omni-Fake: Benchmarking Unified Multimodal Social Media Deepfake Detection Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists
Reference 104
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 424be639-87c3-4e15-9fcd-e89a3e71a1be · inbound
LIVEditor-14B: Lightning Unified Video Editing via In-Context Sparse Attention Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation afcd0946-6219-4eef-ad83-30a35095024f · inbound
Sparkle: Realizing Lively Instruction-Guided Video Background Replacement via Decoupled Guidance Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4e5b4a75-01ad-40b5-aa12-3a8d9528dd7e · inbound
InstructAV2AV: Instruction-Guided Audio-Video Joint Editing Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e92d4077-2ccc-4a26-bdcf-63d68177fe92 · inbound
Aurora: Unified Video Editing with a Tool-Using Agent Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b513661f-8be0-4224-aa5e-04e27f7f2e72 · inbound
StreamEdit: Training-Free Video Editing via Few-Step Streaming Video Generation Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists
Reference 99
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c6ab60b4-a4a9-4d6d-afcc-873b1cfe21a1 · inbound
StreamEdit: Training-Free Video Editing via Few-Step Streaming Video Generation Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists
Reference 99
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d23f9d27-f0fb-40c4-b004-66f6d2470acb · inbound
Reasoning to Align: Implicit Reasoning in Diffusion Transformers for Video Editing Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d61e88f4-3fe8-450f-bc2e-7d6c2322071a · inbound
SpongeBob: Sync-Aware Harmonious Audio-Visual Generative Editing Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 47fa062d-9b6a-43d9-8126-df62952a8cfc · inbound
ElasticTTT: Prior-Preserving Test-Time Tuning for Video Editing Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 361c599b-b00a-4c0b-8e2f-6eb2e1900059 · inbound
ElasticTTT: Prior-Preserving Test-Time Tuning for Video Editing Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.