Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T15:14:36.146697Z
Paper Citation Record · LEDGER
As of 16 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 0 inbound Pith citation observations for arXiv:2608.01113.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T15:14:36.146697Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
53 of 53 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 657b10a4-a0bb-42b0-8f78-5823b143d706 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Scaling instruction-based video editing with a high-quality synthetic dataset.arXiv preprint arXiv:2510.15742, 2025
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1856cfa5-fccf-4d0a-825a-d496f3c5a391 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing In- structpix2pix: Learning to follow image editing instructions
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 408e0bc6-58b0-4827-b1be-96613547c5e7 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing HunyuanVideo-Avatar: High-Fidelity Audio-Driven Human Animation for Multiple Characters
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18f9807a-5d72-4114-a3c3-70770a3eedaf · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Consistent Video-to-Video Transfer Using Synthetic Dataset
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b75ff8c6-70e0-4f0d-9f67-4e751ab27f24 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Rass: Improving denoising diffusion sam- plers with reinforced active sampling scheduler
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation db7eded1-6605-4037-9244-54308783348a · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Why compress what you can generate? when gpt-4o generation ushers in image compression fields
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation c7e8d454-32fa-46ec-9c31-56a81201e8ce · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing SEED-Data-Edit Technical Report: A Hybrid Dataset for Instructional Image Editing
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb512c56-8fb9-4959-b4af-05742885a328 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Clipscore: A reference-free evaluation met- ric for image captioning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 48d98d9a-6c21-44d9-86c0-995bf19d8587 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Imagen Video: High Definition Video Generation with Diffusion Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8cba0fc-3fa7-4c3d-93bf-6345952e11fa · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Video dif- fusion models.Advances in neural information processing systems, 35:8633–8646, 2022
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25ef2ba6-7d42-469c-831d-7876d2fdfc32 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1b5049f-49ca-4103-864e-8cf6af35b0b1 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Animate anyone: Consistent and controllable image- to-video synthesis for character animation
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 49884281-ad6b-42ba-a8a2-3f56dd32d611 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7cde7cc7-1d24-4b9c-aac6-7a87e4f26e5f · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Vbench: Comprehensive bench- mark suite for video generative models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4dd9d16c-bfc1-46db-894b-b7216878607e · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Anyedit: Edit any knowledge encoded in language models.arXiv preprint arXiv:2502.05628, 2025
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4b2e429-8f5b-4a6f-b212-61d7d3bcaf88 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Vace: All-in-one video creation and editing
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2c88d85-f4ca-499b-becf-0f87bb95e566 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Text2video-zero: Text- to-image diffusion models are zero-shot video generators
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0242e969-b0d2-454e-8224-5ea3437cee65 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing HunyuanVideo: A Systematic Framework For Large Video Generative Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05eb5eef-d43d-4443-912a-ad663df13610 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing AnyV2V: A Tuning-Free Framework For Any Video-to-Video Editing Tasks
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3f22532-7f1b-41d9-85ae-d90033f338a7 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing OmniV2V: Versatile Video Generation and Editing via Dynamic Content Manipulation
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c7772c1-8450-4621-a2fe-7362429c370a · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Grounding 3D Scene Affordance From Egocentric Interactions
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1fa8719-768e-4bd2-bd97-7155ad412c6f · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Stablev2v: Stabilizing shape consistency in video-to- video editing.IEEE Transactions on Circuits and Systems for Video Technology, 2025
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 409ef355-6b9c-41e2-945a-6a7dd9d41699 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing The health- wealth gradient in labor markets: Integrating health, in- surance, and social metrics to predict employment density
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 4f40e9e3-6641-4c18-a16c-9f04c1aa84a9 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Video-p2p: Video editing with cross-attention control
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34839ddd-5a94-46ac-b8ce-f8a987f2cca6 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Follow your pose: Pose- guided text-to-video generation using pose-free videos
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5be0a624-c720-4b35-ab97-43354a0f784d · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Magic- stick: Controllable video editing via control handle transfor- mations
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 6d48c170-3a29-4bf3-98e0-67d67d71e691 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing In- structx: Towards unified visual editing with mllm guidance
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2585ad4-dec7-4090-9764-6c2391d8a21d · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Occluded video instance segmentation: A bench- mark.International Journal of Computer Vision, 130(8): 2022–2039, 2022
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 98cde1d6-e308-4324-ab13-0f0a5d78d10a · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Instructvid2vid: Controllable video editing with natural language instructions
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 64bf5aa5-8938-4b18-8474-807bb8e824c3 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Urvos: Unified referring video object segmentation network with a large-scale benchmark
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 2141106f-c39d-43ae-bfa7-3cc3fee87fb2 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Make-A-Video: Text-to-Video Generation without Text-Video Data
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3fcd91b3-7f33-49e9-8c2f-baff1f40219b · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Stylegan-v: A continuous video generator with the price, image quality and perks of stylegan2
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 83fff1b4-c3a0-4f43-bf40-4a6491434b31 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Omni-video: Democratizing uni- fied video understanding and generation.arXiv preprint arXiv:2507.06119, 2025
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af04fc58-2e23-4c2a-bbe7-cc285382811b · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Lucy edit: Open-weight text-guided video editing, 2025
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation d0dc72c7-8972-4b1e-bc7e-45e9454f15ff · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Gemini: A Family of Highly Capable Multimodal Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57023dca-7e67-4dd7-8704-606aa4574292 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Fvd: A new metric for video generation
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 9ab29126-2f77-4800-ae0e-3ff4fbd845f6 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Phenaki: Variable Length Video Generation From Open Domain Textual Description
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3cb4d851-d6ff-425f-9239-27bf6f9efad8 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing ModelScope Text-to-Video Technical Report
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e663e69f-bf22-4cbb-9f2b-755e9db47ae8 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Koala-36m: A large-scale video dataset improving consistency between fine-grained conditions and video content
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 8fb1b249-ffac-4e93-a955-c98c5a7a2aa0 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Tiv-diffusion: Towards object-centric movement for text-driven image to video gen- eration
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation f013a9ff-17b8-40ca-9547-f6c71f7c7fc6 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Re-attentional con- trollable video diffusion editing
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 43379ae1-bdc5-4091-8111-ffe361a8a947 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Training-free controllable text-guided video editing.IEEE Transactions on Circuits and Systems for Video Technology, 2026
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation d7a7aec6-83e1-4638-adb3-63faecad4f77 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Phrasecut: Language-based image segmen- tation in the wild
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation c685c6db-e0dd-4fca-b8ee-7b27984456af · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Tune-a-video: One-shot tuning of image diffusion models for text-to-video generation
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 2f70e7ea-a7e9-4628-b819-5d9dc8fb09d1 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Insvie-1m: Effective instruction-based video editing with elaborate dataset construction
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 67b4273c-b3c0-4edc-8f4b-aa0bce3e75e6 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Veg- gie: Instructional editing and reasoning video concepts with grounded generation
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 222f1b9e-0d29-495a-98a4-845715a0059e · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Editworld: Simulating world dynamics for instruction- following image editing
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 8ba8b5ea-f2e3-400f-bd20-5da9728d9eab · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing I2VGen-XL: High-Quality Image-to-Video Synthesis via Cascaded Diffusion Models
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 89c0ee69-c1c6-411c-8e1c-6e5af7d15788 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing EffiVED:Efficient Video Editing via Text-instruction Diffusion Models
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 106458a7-e319-4faf-9c0d-7a77c611b3f4 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Motionpro: A precise mo- tion controller for image-to-video generation
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 892ca003-1808-4665-9ca2-76f4c6b06c01 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Ultraedit: Instruction-based fine-grained im- age editing at scale.Advances in Neural Information Pro- cessing Systems, 37:3058–3093, 2024
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 6937549a-7058-4ae0-a562-31ce580c7c73 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Semantic under- standing of scenes through the ade20k dataset.International journal of computer vision, 127(3):302–321, 2019
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation f2bbfc9a-9c3a-44b6-b89a-7f1cfbfb6813 · outbound
CoT-Edit: Let CoT Guide Instruction Video Editing Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.