Pith. sign in

Paper Citation Record · LEDGER

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation

As of 12 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 2 inbound Pith citation observations for arXiv:2412.05848.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.05848 v1

Coverage vector

measured 57 of 57 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T20:20:36.912049Z

measured 59 of 59 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:01:12.898019Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T21:09:16.226896Z

Reference resolution

57 of 57 outbound references displayed

  • verified exact0
  • verified fuzzy25
  • unresolved32
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 980a0a1a-aca7-4233-bb19-12526938c468 · outbound

This paper cites Frozen in time: A joint video and image encoder for end-to-end retrieval.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Frozen in time: A joint video and image encoder for end-to-end retrieval

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:20:37.749812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T20:20:36.430982Z digest=sha256:590bfd462b5e6e44e0141f8de1a0092fe5aee1a4f78349c1416cd651cb014d6a

Observation a334d119-9f6d-4022-9229-0a8b19a8ca4f · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T20:20:36.436348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:20:36.436348Z digest=sha256:7862e1c65106cabf8fdd7e1e75ef843a52585703c2349af7a44dfad920061ec7

Observation 78985cf2-7617-477e-a05b-bec2f208bef5 · outbound

This paper cites Align your latents: High-resolution video synthesis with latent diffusion models.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Align your latents: High-resolution video synthesis with latent diffusion models

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:20:37.735375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T20:20:36.441491Z digest=sha256:016ecd7d565da4f29201d8b6301acd0dc4398ed145dea526271eb7e2d578ac5f

Observation 0765079d-8002-41c3-a146-d1fb46381960 · outbound

This paper cites Emerg- ing properties in self-supervised vision transformers.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Emerg- ing properties in self-supervised vision transformers

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:20:37.720706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T20:20:36.446338Z digest=sha256:fedd7b6bb1702458a7d642b918737e8fb6e73580bd270256d87c32237d37f8c6

Observation 6d9176e7-3eea-48fc-9cef-76c1d79f06a5 · outbound

This paper cites Stablevideo: Text-driven consistency-aware diffusion video editing.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Stablevideo: Text-driven consistency-aware diffusion video editing

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:20:37.704685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T20:20:36.451031Z digest=sha256:5c5b137a4490e79d38333d51d38c8ddb38727d2da815eda2061f1655d00bda40

Observation d2dd61c4-2b06-4222-9bd4-f1e1530e8722 · outbound

This paper cites Magicpose: Realistic human poses and facial expressions retargeting with identity-aware dif- fusion.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Magicpose: Realistic human poses and facial expressions retargeting with identity-aware dif- fusion

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:20:37.687085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T20:20:36.455597Z digest=sha256:e31da20ab71795476f484a3e26e3694345cb78b6d409d1e6722999acc8bcd5c4

Observation b6045b52-7b44-4e9c-89b3-035e57e3dba5 · outbound

This paper cites PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T20:20:36.460860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:20:36.460860Z digest=sha256:39c50c47fb0892edde94b8e293a39747e25735ddc3ae937efa1fb6a270032dd4

Observation af723145-b00c-4415-ad34-b41a9aa85e6e · outbound

This paper cites Livephoto: Real image animation with text-guided motion control.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Livephoto: Real image animation with text-guided motion control

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:20:37.670885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T20:20:36.465603Z digest=sha256:bfd3c36601d706eb832f801c566b0691df68f76139d8b68497a7c44c48cd544e

Observation 9684d8ae-4941-4d56-a70c-d779c7a2df96 · outbound

This paper cites Time flies: Animating a still image with time-lapse video as reference.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Time flies: Animating a still image with time-lapse video as reference

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:20:37.655383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T20:20:36.470062Z digest=sha256:feb0ecd27dc77223bd1911994ceea4177986439aab53b015a3f9c6cc110f927a

Observation 75bb9476-a180-4527-acf0-b47e8ff417fa · outbound

This paper cites Animating pictures with stochastic motion textures.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Animating pictures with stochastic motion textures

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:20:37.640062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T20:20:36.474523Z digest=sha256:a3461e389f155d1fe0a23a361067c5cb8685831d606d12795f7d5cb147568b48

Observation 7dd910cc-6445-4dfc-b6ee-d588a634bb33 · outbound

This paper cites Animateanything: Fine-grained open domain image animation with motion guidance.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Animateanything: Fine-grained open domain image animation with motion guidance

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:20:37.624782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T20:20:36.479933Z digest=sha256:485c671dc2f72e0d547abee207a748ac58e0a69a50841516e3fe757e84c9ddc8

Observation 9428f561-5863-43aa-92a1-eb7d4041f472 · outbound

This paper cites Structure and content-guided video synthesis with diffusion models.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Structure and content-guided video synthesis with diffusion models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T20:20:36.484409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:20:36.484409Z digest=sha256:22b0e2553a2911b3b4c055a0476c2c361027294cc3fb8db49bffa8a4a0dd67f9

Observation 00872f25-3555-4bf4-9a17-c3350f7733fa · outbound

This paper cites Perceptual quality assessment of smartphone photography.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Perceptual quality assessment of smartphone photography

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:20:37.600126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T20:20:36.488559Z digest=sha256:98d0b52dbb2599df0fda6ccb7a9fd47ce217245757e3641afe300a039ec1fc33

Observation 8b34c3ec-711c-49a9-bd46-a26009a8df44 · outbound

This paper cites Lumina-T2X: Transforming Text into Any Modality, Resolution, and Duration via Flow-based Large Diffusion Transformers.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Lumina-T2X: Transforming Text into Any Modality, Resolution, and Duration via Flow-based Large Diffusion Transformers

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T20:20:36.493002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:20:36.493002Z digest=sha256:3d8e7834117bfee1f8e107b35e26e2654657f194aedcf53a24209bf7ffab15c1

Observation 37db5820-2d8f-4f00-b4f2-b9d1a74986e6 · outbound

This paper cites Check locate rectify: A training- free layout calibration system for text-to-image generation.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Check locate rectify: A training- free layout calibration system for text-to-image generation

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:20:37.585180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T20:20:36.497541Z digest=sha256:4e88dd32ae05e4bd9fd547785d088c433715ed241f9236c71abc13b00911747a

Observation e69f99eb-ed2e-4f36-bbb0-280a8498defb · outbound

This paper cites AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T20:20:36.501749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:20:36.501749Z digest=sha256:091df4dd98fa66d896af1259f25163bbc398180b71bf21ecd9bc69041cba5f60

Observation e5f5683c-4583-4461-875b-78e25997b877 · outbound

This paper cites Denoising dif- fusion probabilistic models.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Denoising dif- fusion probabilistic models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T20:20:36.506501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:20:36.506501Z digest=sha256:bed04d7d66652088a0b4b91207a987e3c3c399578f6d9c1f09f1483e80762a64

Observation f085d6af-4283-45cd-8ff9-03d92870e397 · outbound

This paper cites Animate anyone: Consistent and controllable image- to-video synthesis for character animation.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Animate anyone: Consistent and controllable image- to-video synthesis for character animation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T20:20:36.511450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:20:36.511450Z digest=sha256:0737bfdf26e6e116be26ac9fe5775d1379d936acbaf56aef950bb85f81538962

Observation 6e7ca8b5-ea40-41d9-9cfa-1d8e9d0921ab · outbound

This paper cites TAda! Temporally-Adaptive Convolutions for Video Understanding.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation TAda! Temporally-Adaptive Convolutions for Video Understanding

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T20:20:36.516428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:20:36.516428Z digest=sha256:ae44877514e27d8d09af0bc4a6437fafdf3a0d93dfb90dc03b49f321f126d2cb

Observation d1743d39-13de-4f3c-8012-fc3c4094b78f · outbound

This paper cites Vbench: Comprehensive bench- mark suite for video generative models.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Vbench: Comprehensive bench- mark suite for video generative models

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:20:37.551225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T20:20:36.524905Z digest=sha256:d42ed616ef03e00229971f9f1f211a3c613836b39adfce7028ed966ac908f2a1

Observation 8f4c6f9b-e355-4ccc-b585-470787608b67 · outbound

This paper cites Musiq: Multi-scale image quality transformer.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Musiq: Multi-scale image quality transformer

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:20:37.536914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T20:20:36.530086Z digest=sha256:32a363614700b13d156825dedf776c7b3a86c130e9ab878ba0f795e485928be5

Observation 0f742f86-0445-4137-b786-be186bdbd0fb · outbound

This paper cites WebVision Database: Visual Learning and Understanding from Web Data.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation WebVision Database: Visual Learning and Understanding from Web Data

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T20:20:36.535446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:20:36.535446Z digest=sha256:8ea7c6933b127534655f9f19026aaafc34a434d7797ce5e4a64faaa5091f22cc

Observation b06a2c8c-1a6b-40d0-96f6-767eca062b0f · outbound

This paper cites Amt: All-pairs multi-field transforms for efficient frame interpolation.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Amt: All-pairs multi-field transforms for efficient frame interpolation

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:20:37.522872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T20:20:36.541042Z digest=sha256:a5d822f390ee68062e59f5cf3b9b26a7fe93305e882cc18f9419785496f493ee

Observation 5235b3aa-d87f-4c22-925e-a05366230818 · outbound

This paper cites MagicEdit: High-Fidelity and Temporally Coherent Video Editing.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation MagicEdit: High-Fidelity and Temporally Coherent Video Editing

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T20:20:36.545889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:20:36.545889Z digest=sha256:0b455515ee830f097ecc0978f5a2f50538f36f6192dba7632aabf4ce3edde1e9

Observation cf8f166c-68ef-46fd-b72e-92db163939a8 · outbound

This paper cites Rankiqa: Learning from rankings for no-reference image quality assessment.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Rankiqa: Learning from rankings for no-reference image quality assessment

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:20:37.508141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T20:20:36.551671Z digest=sha256:df01935ef8ff736304ee38b89d9c230646c972a01b430442ecfcb85932ed64ac

Observation 83ddb5ed-a9ce-4102-9939-417706acd53d · outbound

This paper cites Cinemo: Consistent and Controllable Image Animation with Motion Diffusion Models.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Cinemo: Consistent and Controllable Image Animation with Motion Diffusion Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T20:20:36.556813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:20:36.556813Z digest=sha256:2347908568abb9e67176d424f7a60045586e8aa580b24fcf6db758ae5a001310

Observation 9614bddf-ab90-4f36-8783-b944b3519fa4 · outbound

This paper cites Latte: Latent Diffusion Transformer for Video Generation.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Latte: Latent Diffusion Transformer for Video Generation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T20:20:36.624161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:20:36.624161Z digest=sha256:1efe487097b5a9e4d7171a267b9117374367f6cb2d9fcae2d525b9c5d51ae3d7

Observation 0de5f171-5bcc-4013-8f7c-8de387ff6420 · outbound

This paper cites Controllable animation of fluid elements in still images.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Controllable animation of fluid elements in still images

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T20:20:36.692308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:20:36.692308Z digest=sha256:7a042c687a9afc913dbdb3bbddf150295826066c4421c09a27112ab725ca4a2b

Observation 822c1c3b-94c5-4808-89c8-836307a96fc0 · outbound

This paper cites ReVideo: Remake a Video with Motion and Content Control.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation ReVideo: Remake a Video with Motion and Content Control

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T20:20:36.765950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:20:36.765950Z digest=sha256:3a3d97375bcbef19afc5a6a8e18b344ddcc4b5fb39511425fe050f965822750e

Observation e8cd67a9-0f6b-4e8d-9e83-9bcc01870fde · outbound

This paper cites MOFA-Video: Controllable Image Animation via Generative Motion Field Adaptions in Frozen Image-to-Video Diffusion Model.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation MOFA-Video: Controllable Image Animation via Generative Motion Field Adaptions in Frozen Image-to-Video Diffusion Model

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T20:20:36.782193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:20:36.782193Z digest=sha256:f54f3af73c05371b36321e209e07b9268dd5482094fae78946f81018c0d497bc

Observation 6ed4f9d4-694f-4f90-9b9f-2f7bf487142d · outbound

This paper cites Animating pictures of fluid using video examples.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Animating pictures of fluid using video examples

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:20:37.484401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T20:20:36.788237Z digest=sha256:82b2398df7bdc468deff3873ad2eb4fc82ec81d1117012910695c7e806663b16

Observation 895a38e6-8e9a-4299-8934-6041e7e39948 · outbound

This paper cites Scalable diffusion models with transformers.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Scalable diffusion models with transformers

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T20:20:36.793530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:20:36.793530Z digest=sha256:3a26b4aba6cafb4d2896818096d7d7261a1f1bacf9779994273e66a4d1c41406

Observation 8b38ab25-68f8-495c-8082-a6258d822db5 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Learning transferable visual models from natural language supervi- sion

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:20:37.459181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T20:20:36.798873Z digest=sha256:0a6f1e1f13bfebaa24e1225c375636df08a336d23d3c13a8c55bd99a62d72ae3

Observation 02b0e638-f71c-48c4-94d1-711af8f7b426 · outbound

This paper cites ConsistI2V: Enhancing Visual Consistency for Image-to-Video Generation.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation ConsistI2V: Enhancing Visual Consistency for Image-to-Video Generation

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T20:20:36.804152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:20:36.804152Z digest=sha256:b88832fecd240fbbcd15e2f69f58a4dfbd5eda2fad39bc37a2b32661f024695e

Observation 0b10e7c8-5613-4fa9-9a51-acebb207c7c1 · outbound

This paper cites Image animation with perturbed masks.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Image animation with perturbed masks

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:20:37.444203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T20:20:36.809845Z digest=sha256:16ace7bd9a91187686a54370fa16a1917abca98edc52ae973c2c30ea7a0945e4

Observation 407a73a3-0285-4eda-9e68-c4a406000b9e · outbound

This paper cites ResMaster: Mastering High-Resolution Image Generation via Structural and Fine-Grained Guidance.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation ResMaster: Mastering High-Resolution Image Generation via Structural and Fine-Grained Guidance

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T20:20:36.814676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:20:36.814676Z digest=sha256:0a478f1796bdba8283f01ac815fa0b60b2ec6a5ff88fa54f23ac76b4427d5827

Observation e73fc26a-6487-49f0-867c-f5aeef149789 · outbound

This paper cites First order motion model for image animation.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation First order motion model for image animation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T20:20:36.819805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:20:36.819805Z digest=sha256:f8f6bb6f8ad83c9cc970c4f116a966eefba539c9f1ecf276c177f8054e60d257

Observation 88a0139a-db88-423d-ab11-aca855bd38bd · outbound

This paper cites Motion representations for articulated animation.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Motion representations for articulated animation

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:20:37.420030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T20:20:36.824564Z digest=sha256:e70b99e1496ab731a56d26769fa236a77ed244d19c9983c9dd435f9c938eebf0

Observation 1dc508fb-e8e9-4f7b-af0f-2454fb71e3d8 · outbound

This paper cites Denoising Diffusion Implicit Models.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Denoising Diffusion Implicit Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T20:20:36.829424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:20:36.829424Z digest=sha256:5d196cc5ea1ac9a9ded337ddf034d18c2726d07ead8a4b2c329d204f9ac62f3c

Observation 2762b3b4-87ca-4c02-b75f-2d07753087c3 · outbound

This paper cites Animate-X: Universal Character Image Animation with Enhanced Motion Representation.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Animate-X: Universal Character Image Animation with Enhanced Motion Representation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T20:20:36.834297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:20:36.834297Z digest=sha256:95e01f61b29dbe01d8f237927ef94f682bffccdcc0a45adae77cf9d71066958f

Observation c1026c99-ecad-4048-8882-3f664bf24a15 · outbound

This paper cites Raft: Recurrent all-pairs field transforms for optical flow.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Raft: Recurrent all-pairs field transforms for optical flow

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-11T20:20:36.839940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:20:36.839940Z digest=sha256:eb15f8940403d9a8974ccdd6e3ec9498becf9fccd5311f30a7dfd7c4c72fb15c

Observation c6b63351-bea6-4c7f-90a7-d59670ae9862 · outbound

This paper cites Videocomposer: Compositional video synthesis with motion controllability.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Videocomposer: Compositional video synthesis with motion controllability

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:20:37.395386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T20:20:36.844645Z digest=sha256:5969829e8eb1b61cea6906daa0b6b76fa6e61e45eb0ff49a1dce6949ed443dad

Observation 2293d014-ea73-4923-8c4c-6a7355b880fc · outbound

This paper cites UniAnimate: Taming Unified Video Diffusion Models for Consistent Human Image Animation.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation UniAnimate: Taming Unified Video Diffusion Models for Consistent Human Image Animation

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-11T20:20:36.849444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:20:36.849444Z digest=sha256:d3fa7ad6de3a892965d298ae25c80418003a2be2616ef402337979e892b9570c

Observation 85f9c955-9f79-4e06-9018-750d17553aab · outbound

This paper cites Latent Image Animator: Learning to Animate Images via Latent Space Navigation.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Latent Image Animator: Learning to Animate Images via Latent Space Navigation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-11T20:20:36.854413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:20:36.854413Z digest=sha256:94d56eb408f7424c61e039b97853401ab2d00db7cec79d317934c9822bc4c41d

Observation 12dcf13d-12f4-40ec-841e-8a3c35ecb70a · outbound

This paper cites LAVIE: High-Quality Video Generation with Cascaded Latent Diffusion Models.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation LAVIE: High-Quality Video Generation with Cascaded Latent Diffusion Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T20:20:36.859542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:20:36.859542Z digest=sha256:457ba00059ac11089b9bb26a695c6a4dbe63129aa6329995673982b666a78ce9

Observation 89d36ec6-5321-4ae6-bb04-feb5dba7d748 · outbound

This paper cites Image quality assessment: from error visibility to structural similarity.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Image quality assessment: from error visibility to structural similarity

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-11T20:20:36.864422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:20:36.864422Z digest=sha256:9c62e4b0a1c33f807bb06aa503003b6576e4f993cddf679ed4fad155cdac165c

Observation 12d676eb-19d5-418c-a6c3-9514f9feab9a · outbound

This paper cites Tune-a-video: One-shot tuning of image diffusion models for text-to-video generation.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Tune-a-video: One-shot tuning of image diffusion models for text-to-video generation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-11T20:20:36.869084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:20:36.869084Z digest=sha256:10d9765f43e21844a4a9f027dd674d790648d2f252a160bec7578196952323f5

Observation e9235b92-ae44-418c-aad8-f300c26b5402 · outbound

This paper cites Automatic animation of hair blowing in still portrait photos.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Automatic animation of hair blowing in still portrait photos

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:20:37.359847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T20:20:36.873823Z digest=sha256:a4068ac98c654fd77e184aef0d474e68dffd693f6cf17cf597dad02dfdc2f3f9

Observation 72089502-01b3-4c53-b271-636a43d486c9 · outbound

This paper cites Spatialtracker: Tracking any 2d pixels in 3d space.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Spatialtracker: Tracking any 2d pixels in 3d space

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:20:37.344564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T20:20:36.878490Z digest=sha256:675bbf2c2f49de7540880bfbf62ca31bb6c9c4eb02033ac6643bf18767918686

Observation 38a3b6cf-4da9-4bb3-8203-6ae15b32910d · outbound

This paper cites Moving Object Segmentation: All You Need Is SAM (and Flow).

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Moving Object Segmentation: All You Need Is SAM (and Flow)

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-11T20:20:36.882576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:20:36.882576Z digest=sha256:0bae040f3308fa6387994e7e71c7dfec2b4ac406d6e911cf6b712351573819ee

Observation 88c34f45-633e-46d5-872d-2a4a9b290d2d · outbound

This paper cites Dynamicrafter: Animating open-domain images with video diffusion priors.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Dynamicrafter: Animating open-domain images with video diffusion priors

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-11T20:20:36.886952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:20:36.886952Z digest=sha256:039eb171eaaca01bbfbd02bc0573d8f0c4fe72f503f4e315c1750950f8a3f2b7

Observation f2ad4be9-9cfa-4072-a30b-932c7f84f57e · outbound

This paper cites CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-11T20:20:36.890791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:20:36.890791Z digest=sha256:15e2b8627b886bc0547bb9f98cdcd44d08099dc14ec1837160656ab8ed0b5a11

Observation 77c497ee-8225-4536-8cb4-4466ed269763 · outbound

This paper cites I2VGen-XL: High-Quality Image-to-Video Synthesis via Cascaded Diffusion Models.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation I2VGen-XL: High-Quality Image-to-Video Synthesis via Cascaded Diffusion Models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-11T20:20:36.895068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:20:36.895068Z digest=sha256:861915a3a1bfe4f79b79e3cfaa659abfeb46b048a2b41cf07aa8c467cb40499b

Observation d13db79d-32f8-4983-9142-91d146723653 · outbound

This paper cites Pia: Your personalized image animator via plug-and-play modules in text-to-image models.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Pia: Your personalized image animator via plug-and-play modules in text-to-image models

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:20:37.319190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T20:20:36.899353Z digest=sha256:ebe84aeea22fdd8e09a0e2a5163bbc3fa7f8b0540c2f437d8e17ec367bcec6c0

Observation b885a085-1c37-4a5b-9b23-98a8c9e28fe5 · outbound

This paper cites Tora: Trajectory-oriented Diffusion Transformer for Video Generation.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Tora: Trajectory-oriented Diffusion Transformer for Video Generation

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-11T20:20:36.903443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:20:36.903443Z digest=sha256:305a1fced85d7cc28d2d5f378edc906388f50862d46cfbe2bc4f0558f411031e

Observation d4fdc5bd-f6a1-429b-883b-b3a2fd92f559 · outbound

This paper cites Thin-plate spline motion model for image animation.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Thin-plate spline motion model for image animation

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:20:37.303978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T20:20:36.907766Z digest=sha256:f018fb37d20ebac4f4d6f74c7900112c71505f75317f608e6f48fcf9363dbf04

Observation 76f30291-e566-49fe-b167-24b475a90d78 · outbound

This paper cites Camera zooms out. A penguin is dancing.

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation Camera zooms out. A penguin is dancing

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:20:37.289046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T20:20:36.912049Z digest=sha256:78fc1c86240e75bd2941d6dc6531f1788643509d9e95e3fcbeb0e336045e29a9

Pith citing papers

Observation e1c8c399-d6b0-42f8-ab63-1f5f4fd7713f · inbound

AniCrafter: Customizing Realistic Human-Centric Animation via Avatar-Background Conditioning in Video Diffusion Models cites this paper.

AniCrafter: Customizing Realistic Human-Centric Animation via Avatar-Background Conditioning in Video Diffusion Models MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T14:01:12.898019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:01:12.898019Z digest=sha256:69246983ebe7f22f2e05ab2b56c661d7d7e6c3645416250f2dd16bf7e153d3a3

Observation 22632463-f4e5-4499-93d4-3dcac8ddc498 · inbound

Animate-X++: Universal Character Image Animation with Dynamic Backgrounds cites this paper.

Animate-X++: Universal Character Image Animation with Dynamic Backgrounds MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-08-05T21:09:16.319417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T21:09:08.118424Z digest=sha256:ce83c3118e0d6010bc96f1cde107f83ee438d58f84e36a71ddb22f7e58bafa46