Pith. sign in

Paper Citation Record · LEDGER

AnyI2V: Animating Any Conditional Image with Motion Control

As of 8 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 1 inbound Pith citation observation for arXiv:2507.02857.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.02857 v1

Coverage vector

measured 57 of 57 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:25:29.212390Z

measured 58 of 58 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-03T16:19:51.348169Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T16:28:38.518422Z

Reference resolution

57 of 57 outbound references displayed

  • verified exact3
  • verified fuzzy25
  • unresolved29
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1fed75c3-b20c-4f6b-96ac-78b1a717cd86 · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

AnyI2V: Animating Any Conditional Image with Motion Control Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:24.234995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:24.234995Z digest=sha256:859f5d549cebf2499357432e19e13b80c8919bc6232f653836662203ed12a1c1

Observation 88234fc6-27ef-4f33-83cd-dc35b5f64f2a · outbound

This paper cites Align your latents: High-resolution video synthesis with latent diffusion models.

AnyI2V: Animating Any Conditional Image with Motion Control Align your latents: High-resolution video synthesis with latent diffusion models

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:34.460141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:25:24.323877Z digest=sha256:eb9d5728f1d6f0ba24169114ef95c0a508f1baa1ef3ff36b6719bafce231e60a

Observation efdf9c10-8737-47a9-a29a-4edda0d43b45 · outbound

This paper cites A unified 3d human motion synthesis model via conditional variational auto-encoder.

AnyI2V: Animating Any Conditional Image with Motion Control A unified 3d human motion synthesis model via conditional variational auto-encoder

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:34.328860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:25:24.413431Z digest=sha256:c8eddb1259d44098a19d2e9e95c4c2af3865f14aff66add6de791c0fcc7496be

Observation 47fb6f8e-8c13-427a-8c2c-c82e2e41c664 · outbound

This paper cites VideoCrafter1: Open Diffusion Models for High-Quality Video Generation.

AnyI2V: Animating Any Conditional Image with Motion Control VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:24.503120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:24.503120Z digest=sha256:42744b50929ef41207d0b78516903cef0cdec95c2bdcaa6a4a79abe695c3685f

Observation 537051c3-c895-478e-8dab-e699d1e73474 · outbound

This paper cites Videocrafter2: Overcoming data limitations for high-quality video diffusion models.

AnyI2V: Animating Any Conditional Image with Motion Control Videocrafter2: Overcoming data limitations for high-quality video diffusion models

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:34.189521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:25:24.594083Z digest=sha256:a26098bcb9ee4d0f2bd59526a5fc62c810defc8828c7e40766722e4ff0e221b9

Observation 78f7342e-fe75-436f-8721-6dfd29aad377 · outbound

This paper cites MeViS: A large-scale benchmark for video segmentation with motion expressions.

AnyI2V: Animating Any Conditional Image with Motion Control MeViS: A large-scale benchmark for video segmentation with motion expressions

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:33.997665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:25:24.679423Z digest=sha256:6f019558e253d54309106d8255fc0ab8afb3ec7359e8c1c02ab3b84f64da73b7

Observation 8991ac9c-0e5e-49cd-86b0-9fe1b6ca30f8 · outbound

This paper cites AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning.

AnyI2V: Animating Any Conditional Image with Motion Control AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:24.769350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:24.769350Z digest=sha256:21bc326bef5ce0545f2bffe70e9726bb38e303c21fb698c4cc3233a056bbbf65

Observation 35b70089-05a0-4fa1-8c2b-22c36faedc0e · outbound

This paper cites Sparsectrl: Adding sparse controls to text-to-video diffusion models.

AnyI2V: Animating Any Conditional Image with Motion Control Sparsectrl: Adding sparse controls to text-to-video diffusion models

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:33.854495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:25:24.884766Z digest=sha256:a35d462ec13a4e8ce5387adf0fdcdc78a067354dc13a451c3827e07ccd402e8c

Observation 84520573-1143-4c07-b930-d1d95e56baa1 · outbound

This paper cites CameraCtrl: Enabling Camera Control for Text-to-Video Generation.

AnyI2V: Animating Any Conditional Image with Motion Control CameraCtrl: Enabling Camera Control for Text-to-Video Generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:24.979478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:24.979478Z digest=sha256:623d6f76683d5bf11064525aadc42107d568bfbd0e6ceec608d6307e6cfa48b2

Observation 00467e71-d160-4147-9adc-40188cf446b9 · outbound

This paper cites Latent Video Diffusion Models for High-Fidelity Long Video Generation.

AnyI2V: Animating Any Conditional Image with Motion Control Latent Video Diffusion Models for High-Fidelity Long Video Generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:25.057827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:25.057827Z digest=sha256:4b41cc340775d6322c4b892211e09ecf7bd39055c61b6859148536d081e34ba3

Observation c424eb1b-8aa8-43b4-960f-411cbbdaba4e · outbound

This paper cites Prompt-to-Prompt Image Editing with Cross Attention Control.

AnyI2V: Animating Any Conditional Image with Motion Control Prompt-to-Prompt Image Editing with Cross Attention Control

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:25.140993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:25.140993Z digest=sha256:42a16b745b3b511222bb3150b12917a13964dcf752fc7c79ce3b4ae63224c992

Observation 4a3287ac-ad55-4777-a748-4f898526342c · outbound

This paper cites Video diffusion models.

AnyI2V: Animating Any Conditional Image with Motion Control Video diffusion models

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:33.684728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:25:25.240688Z digest=sha256:dc8e2b1918c85430f00d64a75edfe39a5a0a168885e95d44297dab73020eae21

Observation c3b78684-d66f-476e-bafb-695ab7fa9ace · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

AnyI2V: Animating Any Conditional Image with Motion Control LoRA: Low-Rank Adaptation of Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:25.318105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:25.318105Z digest=sha256:755144acc316f88a11593319ee2c1aa62af1764bbae87e2ac0487d058641b7ee

Observation 9fddd0ee-b922-46d1-9b73-51842255c65b · outbound

This paper cites Cocktail: Mixing multi-modality control for text-conditional image generation.

AnyI2V: Animating Any Conditional Image with Motion Control Cocktail: Mixing multi-modality control for text-conditional image generation

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:33.505365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:25:25.380684Z digest=sha256:bf75230a15b84811f154df7a3daf677c0422d62e830662e6dd5c2184745b6e21

Observation 475fe6c3-411c-4d7c-be8a-a8000066e717 · outbound

This paper cites VideoControlNet: A Motion-Guided Video-to-Video Translation Framework by Using Diffusion Model with ControlNet.

AnyI2V: Animating Any Conditional Image with Motion Control VideoControlNet: A Motion-Guided Video-to-Video Translation Framework by Using Diffusion Model with ControlNet

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:25.466570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:25.466570Z digest=sha256:77d03fc7ead8fef5a91bdfa4b8530e9b0ce83b961075565637e455058734bd78

Observation bc66d003-8b0c-4769-bb91-666dc3eb6548 · outbound

This paper cites Arbitrary style transfer in real-time with adaptive instance normalization.

AnyI2V: Animating Any Conditional Image with Motion Control Arbitrary style transfer in real-time with adaptive instance normalization

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:33.367870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:25:25.558706Z digest=sha256:d0ede10476657fee81603ef2e510d0ad85c351945987b44bdcded1040bfbc26e

Observation d19d4810-6fc5-4c9f-a904-61b8c942574f · outbound

This paper cites Cotracker: It is better to track together.

AnyI2V: Animating Any Conditional Image with Motion Control Cotracker: It is better to track together

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:33.236726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:25:25.679527Z digest=sha256:3d4a566595f5e688525ce2c14b73878ff9f43524c4b9fb25fdcae62aadc8c22f

Observation 79574eae-be01-4584-ad52-ea507b4c5876 · outbound

This paper cites Text2video-zero: Text- to-image diffusion models are zero-shot video generators.

AnyI2V: Animating Any Conditional Image with Motion Control Text2video-zero: Text- to-image diffusion models are zero-shot video generators

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:25.815397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:25.815397Z digest=sha256:5c9916ad6d90f737adfaf3308b42c3aa09e63c667af71511809d1f8e4c02cad5

Observation b8490783-78ba-40ff-bd46-73591113df79 · outbound

This paper cites DiffBlender: Composable and Versatile Multimodal Text-to-Image Diffusion Models.

AnyI2V: Animating Any Conditional Image with Motion Control DiffBlender: Composable and Versatile Multimodal Text-to-Image Diffusion Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:25.914025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:25.914025Z digest=sha256:81bbafadb198c96bf277854d9db8f6b1ed205d2479ff614b351eeabb172dd002

Observation f0a79e52-683e-471f-834b-a8b00b4bf615 · outbound

This paper cites Compose and Conquer: Diffusion-Based 3D Depth Aware Composable Image Synthesis.

AnyI2V: Animating Any Conditional Image with Motion Control Compose and Conquer: Diffusion-Based 3D Depth Aware Composable Image Synthesis

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:25:29.851895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:25:26.042201Z digest=sha256:237048e06b0848a5c7c657a1fd8777b56b77b8c35fe25380979d3c4b1c5b4fbc

Observation d3b52798-77ba-4689-a296-1a3266b80d7b · outbound

This paper cites Image Conductor: Precision Control for Interactive Video Synthesis.

AnyI2V: Animating Any Conditional Image with Motion Control Image Conductor: Precision Control for Interactive Video Synthesis

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:26.170396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:26.170396Z digest=sha256:9284c2c06b41dfbacbd7c39a1e908c1198fea2d6bc0da0b36e3bbd044587d52f

Observation 5feeba51-a878-426d-afe4-e1017b07d665 · outbound

This paper cites LOVECon: Text-driven Training-Free Long Video Editing with ControlNet.

AnyI2V: Animating Any Conditional Image with Motion Control LOVECon: Text-driven Training-Free Long Video Editing with ControlNet

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:25:29.716487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:25:26.239158Z digest=sha256:3cc133421a795d54acaf4f2271003b93a5757be9c62317124d586215ee682584

Observation 00524d34-72b6-4ca3-a573-6b5d1407a5dc · outbound

This paper cites Trailblazer: Trajectory control for diffusion-based video generation.

AnyI2V: Animating Any Conditional Image with Motion Control Trailblazer: Trajectory control for diffusion-based video generation

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:33.012633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:25:26.320914Z digest=sha256:419f924c29cf535b070520f5238573bead26be45d7992e7e73ea107587d7561f

Observation e796bd06-2fb9-498a-8b85-e3ccb432bbb7 · outbound

This paper cites Some methods for classification and analysis of multivariate observations.

AnyI2V: Animating Any Conditional Image with Motion Control Some methods for classification and analysis of multivariate observations

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:32.881148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:25:26.403459Z digest=sha256:e4c64a4416610cd24317da0a52124e638f9b06b21720f3cac76d7a0569dbeb19

Observation 661d2f45-91d9-4fc7-9bda-a2cced54d8c5 · outbound

This paper cites Large-scale video panoptic segmentation in the wild: A benchmark.

AnyI2V: Animating Any Conditional Image with Motion Control Large-scale video panoptic segmentation in the wild: A benchmark

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:32.735514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:25:26.474151Z digest=sha256:272ae5d2173a2278a810c843a41ece08d6a1187800900f97913ff47f3d7c636b

Observation d73ee21d-b3cd-43af-97c1-0bea7ea1e6ca · outbound

This paper cites Freecontrol: Training-free spatial control of any text-to-image diffusion model with any condition.

AnyI2V: Animating Any Conditional Image with Motion Control Freecontrol: Training-free spatial control of any text-to-image diffusion model with any condition

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:32.562331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:25:26.553258Z digest=sha256:e97d0c75d88183225dd3651ffe68b70826e9290eea03d44950c67edbb6f76083

Observation 8526d6aa-bf06-4f6b-85a0-d2c9887d5ca2 · outbound

This paper cites SG-I2V: Self-Guided Trajectory Control in Image-to-Video Generation.

AnyI2V: Animating Any Conditional Image with Motion Control SG-I2V: Self-Guided Trajectory Control in Image-to-Video Generation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:26.611739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:26.611739Z digest=sha256:05029f0ba497a47849998e91b684f081ca698e49abce2c2fd637161a47333c77

Observation 0afba912-7f85-4603-bfd0-d1ec67db126d · outbound

This paper cites Mofa-video: Controllable image animation via generative motion field adaptions in frozen image-to-video diffusion model.

AnyI2V: Animating Any Conditional Image with Motion Control Mofa-video: Controllable image animation via generative motion field adaptions in frozen image-to-video diffusion model

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:32.433431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:25:26.685839Z digest=sha256:e6d3d5cf72bf1f367071afe7e46eb8479d8ca5f83b712b0b8a9346887b4247ff

Observation 92e86803-561b-43a2-b383-9cdfb26ad7d8 · outbound

This paper cites Drag your gan: Interactive point-based manipulation on the generative image manifold.

AnyI2V: Animating Any Conditional Image with Motion Control Drag your gan: Interactive point-based manipulation on the generative image manifold

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:32.218979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:25:26.760190Z digest=sha256:a18dce2b2ae507ef9f6c3b4018be81c76073b623d2454aac6a2a4d63601ba6fb

Observation 38b44cb9-bcc3-40a5-8436-4c7176e5d212 · outbound

This paper cites UniControl: A Unified Diffusion Model for Controllable Visual Generation In the Wild.

AnyI2V: Animating Any Conditional Image with Motion Control UniControl: A Unified Diffusion Model for Controllable Visual Generation In the Wild

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:26.818684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:26.818684Z digest=sha256:3935c4bba580530b54cc8e384493a5b99a462b33b3c37fc32e9abb956ae099dc

Observation 42402254-5ab4-4997-be70-b199df2bf15c · outbound

This paper cites FreeTraj: Tuning-Free Trajectory Control in Video Diffusion Models.

AnyI2V: Animating Any Conditional Image with Motion Control FreeTraj: Tuning-Free Trajectory Control in Video Diffusion Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:26.881733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:26.881733Z digest=sha256:69b7dcfa80fd040fc94a9f8ebddf5b25a57b27ab5fae583042dde4e93158864b

Observation baed88d4-efb6-4686-a27d-15a66c5355a1 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

AnyI2V: Animating Any Conditional Image with Motion Control High-resolution image synthesis with latent diffusion models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:26.900124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:26.900124Z digest=sha256:a2c6716a268711a4363c2d83d55e51219eb3044de5e259ea6ba1c9bc0adb4d59

Observation fd175d70-34e2-448a-b6f4-b2577bf1b0ce · outbound

This paper cites Motion-i2v: Consistent and controllable image-to-video generation with explicit motion modeling.

AnyI2V: Animating Any Conditional Image with Motion Control Motion-i2v: Consistent and controllable image-to-video generation with explicit motion modeling

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:32.068727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:25:26.970264Z digest=sha256:2a06d4ef6bc34110df6879ad3ba4cb8d33faac4864bea8fd7dbb17bf7edc5d18

Observation 15eb4f40-dd58-403c-b2c4-ed011779c61b · outbound

This paper cites Dragdiffusion: Harnessing diffusion models for interactive point-based image editing.

AnyI2V: Animating Any Conditional Image with Motion Control Dragdiffusion: Harnessing diffusion models for interactive point-based image editing

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:31.963843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:25:27.078636Z digest=sha256:837f9c20fdf39090817720e01d653737a8d796a1ecd0253159fe36302063df40

Observation 40b743f7-bc70-444f-8b57-053c8fc6fc83 · outbound

This paper cites A Survey of Multimodal-Guided Image Editing with Text-to-Image Diffusion Models.

AnyI2V: Animating Any Conditional Image with Motion Control A Survey of Multimodal-Guided Image Editing with Text-to-Image Diffusion Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:27.267209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:27.267209Z digest=sha256:738f16e287d3c81c863b3d0ca890cd09289a971e5a515ad50f88147597d74890

Observation 32df7593-0501-4d21-bc58-60be5e71da60 · outbound

This paper cites Free-form motion control: A synthetic video generation dataset with controllable camera and object motions.

AnyI2V: Animating Any Conditional Image with Motion Control Free-form motion control: A synthetic video generation dataset with controllable camera and object motions

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:27.408167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:27.408167Z digest=sha256:1db7a747fb9460b0499ab4722e87caef9efd5c040ada3b806b212b2670a28d94

Observation d3201c98-aac7-4366-8a95-f8f259a0d87d · outbound

This paper cites Make-A-Video: Text-to-Video Generation without Text-Video Data.

AnyI2V: Animating Any Conditional Image with Motion Control Make-A-Video: Text-to-Video Generation without Text-Video Data

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:27.446167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:27.446167Z digest=sha256:ddb69becfc43aa3c0b0317cb566de26c35ae7a693943632ca70c2f5ad48b3c86

Observation b5292019-0b83-4721-be2b-ae7ff2d5fb28 · outbound

This paper cites Denoising Diffusion Implicit Models.

AnyI2V: Animating Any Conditional Image with Motion Control Denoising Diffusion Implicit Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:27.533479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:27.533479Z digest=sha256:e6b619ca19de6a8898928a5396fd05ccb065734c73b3dff013e5fb30b8bfdd48

Observation f6a8d720-8439-483d-a44e-83d33bed5481 · outbound

This paper cites Anycontrol: create your artwork with versatile control on text-to-image generation.

AnyI2V: Animating Any Conditional Image with Motion Control Anycontrol: create your artwork with versatile control on text-to-image generation

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:31.798937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:25:27.623943Z digest=sha256:343a0e75f382a0a2bdd18c9e67be9d5aa59c69e9caf0dc1946995e5a2895d9fd

Observation 27c0b93b-146f-400e-91c5-0aa28818b3fd · outbound

This paper cites Plug-and-play diffusion features for text-driven image-to-image translation.

AnyI2V: Animating Any Conditional Image with Motion Control Plug-and-play diffusion features for text-driven image-to-image translation

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:31.654118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:25:27.774197Z digest=sha256:e93cf52eb13cc62716a1e7166324041d3694ba7ac666354e384b7f96ea5883db

Observation d3771b2c-ec8f-4c5b-af71-83e19304621e · outbound

This paper cites ModelScope Text-to-Video Technical Report.

AnyI2V: Animating Any Conditional Image with Motion Control ModelScope Text-to-Video Technical Report

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:27.903864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:27.903864Z digest=sha256:741a30f868e12bdf5b32ed21ad709cde0b8d3faf7df8211644ca09e606dc2893

Observation ef2b81e3-e487-4b92-a6da-fb15b3fc5765 · outbound

This paper cites Boximator: Generating Rich and Controllable Motions for Video Synthesis.

AnyI2V: Animating Any Conditional Image with Motion Control Boximator: Generating Rich and Controllable Motions for Video Synthesis

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:28.032439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:28.032439Z digest=sha256:61ce229dc91a2adbcd87dc7650f6e4bf350806681a88b2e861619c4f77220b03

Observation d856d6b6-8afc-4eee-a71d-a74b60ec6519 · outbound

This paper cites Videocomposer: Compositional video synthesis with motion controllability.

AnyI2V: Animating Any Conditional Image with Motion Control Videocomposer: Compositional video synthesis with motion controllability

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:31.526763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:25:28.197918Z digest=sha256:e055d0cde582e7a142210da2ff65a30f987639dee77eaa05ee951a5900a80c6d

Observation 37898861-1e15-4fec-b544-e391ee84143d · outbound

This paper cites Lavie: High-quality video generation with cascaded latent diffusion models.

AnyI2V: Animating Any Conditional Image with Motion Control Lavie: High-quality video generation with cascaded latent diffusion models

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:31.177524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:25:28.282240Z digest=sha256:1a19d84bd780fd227d0023c40f4556d6933239f480348ea8562c9b38424bd8fe

Observation c9840de4-56e8-4fd6-9acd-8ff050e23d5c · outbound

This paper cites ObjCtrl-2.5D: Training-free Object Control with Camera Poses.

AnyI2V: Animating Any Conditional Image with Motion Control ObjCtrl-2.5D: Training-free Object Control with Camera Poses

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:28.341243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:28.341243Z digest=sha256:19e372443721a8afa5be21fcb4118b9ec335feec149e936f682a29d374e87aa5

Observation 91b33beb-3f62-4ddb-969e-f0e49dd75f68 · outbound

This paper cites Motionctrl: A unified and flexible motion controller for video generation.

AnyI2V: Animating Any Conditional Image with Motion Control Motionctrl: A unified and flexible motion controller for video generation

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:30.904545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:25:28.415133Z digest=sha256:75d5aa0fc4762c3bd62581f4e1b32cfc264c23f8bb29eff4ed7b6f5998f8e1ba

Observation 396d7879-b606-43e1-a74c-6cba50ac1ae1 · outbound

This paper cites Principal component analysis.

AnyI2V: Animating Any Conditional Image with Motion Control Principal component analysis

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:30.631228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:25:28.525885Z digest=sha256:797383d937a3d325187ce21bdc92eb7dd922cc6f008c14f87e95d146f1f94912

Observation 39f036ab-5588-464d-8392-59af76ab3bf2 · outbound

This paper cites MotionBooth: Motion-Aware Customized Text-to-Video Generation.

AnyI2V: Animating Any Conditional Image with Motion Control MotionBooth: Motion-Aware Customized Text-to-Video Generation

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:28.608683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:28.608683Z digest=sha256:a92202058af280f35432e53af875b96a0788e9a96def6008a34bbc0ce509c3d2

Observation 15b79da1-3c97-477f-b11c-507de6af68ec · outbound

This paper cites Draganything: Motion control for anything using entity representation.

AnyI2V: Animating Any Conditional Image with Motion Control Draganything: Motion control for anything using entity representation

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:30.383936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:25:28.647601Z digest=sha256:509d248d045efa859da06a64f29325dc08cca762daee073d759766ef99a28332

Observation d8787f87-8def-4463-8b24-f76278b73f0c · outbound

This paper cites Video Diffusion Models are Training-free Motion Interpreter and Controller.

AnyI2V: Animating Any Conditional Image with Motion Control Video Diffusion Models are Training-free Motion Interpreter and Controller

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:28.728755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:28.728755Z digest=sha256:063c73a3183032bb5362cd703d82987f56ac7989716a2410732249fd1d68cd62

Observation 4ab67a9c-61ea-49c4-a657-7d1bc1f651b0 · outbound

This paper cites Dynamicrafter: Animating open-domain images with video diffusion priors.

AnyI2V: Animating Any Conditional Image with Motion Control Dynamicrafter: Animating open-domain images with video diffusion priors

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:28.790311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:28.790311Z digest=sha256:a108444b8621b9d46bc648a687354fe5658bde7f302f7b5dc34018d76c68d1a9

Observation ec32a8f0-e334-4054-882a-570f449f538f · outbound

This paper cites Direct-a-video: Customized video generation with user- directed camera movement and object motion.

AnyI2V: Animating Any Conditional Image with Motion Control Direct-a-video: Customized video generation with user- directed camera movement and object motion

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:30.116321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:25:28.847157Z digest=sha256:3be4c19541f6ab2df6e72ae6832ecee5511b863009d6a11666b60120e07c98e7

Observation 695822a8-cc50-409c-8cf6-d9736a99a273 · outbound

This paper cites DragNUWA: Fine-grained Control in Video Generation by Integrating Text, Image, and Trajectory.

AnyI2V: Animating Any Conditional Image with Motion Control DragNUWA: Fine-grained Control in Video Generation by Integrating Text, Image, and Trajectory

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:28.946126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:28.946126Z digest=sha256:87ce633a695cbf5cca8fba78536ce736b386af13dd79d506716de517d294824b

Observation 5b748225-0f62-4468-8f00-57e762518ba8 · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

AnyI2V: Animating Any Conditional Image with Motion Control Adding conditional control to text-to-image diffusion models

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:28.992068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:28.992068Z digest=sha256:c6f3e36f080aaebe32efec825d58860197d6fa94e5f6e0d81a89627a4e497d9b

Observation 1dc34cde-99a8-4f00-8daa-3face3f52316 · outbound

This paper cites ControlVideo: Training-free Controllable Text-to-Video Generation.

AnyI2V: Animating Any Conditional Image with Motion Control ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:29.057301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:29.057301Z digest=sha256:63b13c85d860566bb0f5f6549274f60d30f677772892e3c0fdc0bbcfdf8f5dcd

Observation 1351d6ee-1961-4baa-95d3-85db563283d7 · outbound

This paper cites Tora: Trajectory-oriented Diffusion Transformer for Video Generation.

AnyI2V: Animating Any Conditional Image with Motion Control Tora: Trajectory-oriented Diffusion Transformer for Video Generation

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:29.127199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:29.127199Z digest=sha256:2fd842161c54faf2bf3a132beb55276e207d1d9dcd96c3c06b620e90ce5d4c06

Observation cdcdc99d-47e6-4ea1-8476-9548690d5220 · outbound

This paper cites TrackGo: A Flexible and Efficient Method for Controllable Video Generation.

AnyI2V: Animating Any Conditional Image with Motion Control TrackGo: A Flexible and Efficient Method for Controllable Video Generation

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:25:29.416173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:25:29.212390Z digest=sha256:29c6fe50a1d0d63267b1b0b8c65c8a58a16ceaaeaa30ec9010a9a88b45969791

Pith citing papers

Observation bd2c33f3-ad0b-4654-8b84-00132bccf2f5 · inbound

QWERTY: Training-Free Motion Control via Query-Warped Video Diffusion Transformers cites this paper.

QWERTY: Training-Free Motion Control via Query-Warped Video Diffusion Transformers AnyI2V: Animating Any Conditional Image with Motion Control

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-03T16:28:38.520077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-03T16:19:51.348169Z digest=sha256:b5fa366ab6d3d88974ec32d3bf6b7f42dfc211a9a29dce3740b7cc9c8eb23a65