Pith. sign in

Paper Citation Record · LEDGER

Edit as You See: Image-guided Video Editing via Masked Motion Modeling

As of 11 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 0 inbound Pith citation observations for arXiv:2501.04325.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.04325 v1

Coverage vector

measured 54 of 54 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T21:41:23.282342Z

measured 54 of 54 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

54 of 54 outbound references displayed

  • verified exact1
  • verified fuzzy40
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 42b8a104-7d5a-4122-9737-66354e2c9150 · outbound

This paper cites Blended diffusion for text-driven editing of natural images.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Blended diffusion for text-driven editing of natural images

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.989202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.048302Z digest=sha256:6465a5af5d66f8bd3d5889ba319e71ebb635eba0fc556de32d80d575c5798874

Observation 9cb97f2d-8cf6-42a1-a1c2-d3eb9da35884 · outbound

This paper cites Text2live: Text-driven layered image and video editing.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Text2live: Text-driven layered image and video editing

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.978351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.053362Z digest=sha256:6b44a68551a2b0aaf21a26b7ff2845d2fad1eb121f56edfb0192157dbf065a82

Observation 0fcf37a5-6e79-4c3b-99c2-d52311add416 · outbound

This paper cites In- structpix2pix: Learning to follow image editing instructions.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling In- structpix2pix: Learning to follow image editing instructions

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.967675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.057864Z digest=sha256:7e843caa0c3402330c8ae8f7bd42561aa365924d3607403c794214d88adab0f6

Observation 3e7be476-1869-486d-9b28-9808f15909bb · outbound

This paper cites Pix2video: Video editing using image diffusion.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Pix2video: Video editing using image diffusion

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.956733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.063303Z digest=sha256:8616ad792af51fe0238222aa14862fc9cb89c7eae017327a8cc057e0d58a517e

Observation 3fa87c73-9e85-4344-8e4e-48911593c4de · outbound

This paper cites Stable- video: Text-driven consistency-aware diffusion video edit- ing.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Stable- video: Text-driven consistency-aware diffusion video edit- ing

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.944429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.069225Z digest=sha256:7a59226b709dc039d5968c22aa8fc9372ab1b700f0e73d7e68230925752d2137

Observation a0f5fde4-86e9-4423-b57e-e194863c68c0 · outbound

This paper cites Zero-shot Image Editing with Reference Imitation.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Zero-shot Image Editing with Reference Imitation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.074303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.074303Z digest=sha256:71dc86c74e2ff97cbf03c6f500a5cd5b92d8e534296d538a58aa557a1491a925

Observation 0e6f196c-371a-4f18-81ea-501dc8838046 · outbound

This paper cites Xmem: Long- term video object segmentation with an atkinson-shiffrin memory model.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Xmem: Long- term video object segmentation with an atkinson-shiffrin memory model

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.931085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.079713Z digest=sha256:77ec317a1e2e98d5058cf79cda60b44636839f91b81c39e741847bb5c8206cd6

Observation cc5a1e3d-7f05-402e-9b34-ebf73f9ca1d0 · outbound

This paper cites Compvis/stable-diffusion: A latent text-to- image diffusion model.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Compvis/stable-diffusion: A latent text-to- image diffusion model

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.918275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.084951Z digest=sha256:dadb82bf41ae5fa6935962d5caf906572b47a06aea30f3ebf775f9e89c351694

Observation 392b5eee-b0e5-4756-9ac6-ae9d16393a0d · outbound

This paper cites DiffEdit: Diffusion-based semantic image editing with mask guidance.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling DiffEdit: Diffusion-based semantic image editing with mask guidance

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.090284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.090284Z digest=sha256:c0c1412e37c735fca60dfba5ba407ef10e8aa0979d08c8d5b0a1dc1a6d167285

Observation 8e3650e1-78ba-4ee5-bcdf-3b5ca474cc2a · outbound

This paper cites Videdit: Zero-shot and spatially aware text-driven video editing.IEEE Trans.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Videdit: Zero-shot and spatially aware text-driven video editing.IEEE Trans

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.906371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.095632Z digest=sha256:4dc9104994de47cd502aa7dd8b7779a55fa736814678fb0113874829cfd55854

Observation 12e6d0ba-8e2b-4c86-b713-2e5465bd330a · outbound

This paper cites Diffusion models beat gans on image synthesis.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Diffusion models beat gans on image synthesis

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.894691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.100475Z digest=sha256:462d3af72f9e3ee1fed04939dbad400c30647a16d7d46e701a9af2ec477f4959

Observation b9ff7a3d-bf18-4496-8175-a42f42c738b5 · outbound

This paper cites Editanything: Empower- ing unparalleled flexibility in image editing and generation.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Editanything: Empower- ing unparalleled flexibility in image editing and generation

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.882900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.104235Z digest=sha256:342b991b154903f0504ae0c8617951f00be71c413123be654ebca25ffe9aaf3f

Observation 499e42d0-3b1c-4c68-be94-31de4bc362c1 · outbound

This paper cites TokenFlow: Consistent Diffusion Features for Consistent Video Editing.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling TokenFlow: Consistent Diffusion Features for Consistent Video Editing

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.107577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.107577Z digest=sha256:36d92c8ebf8ba15ca2b4ae0c20f4b955a629e69282122afef171c158f9c0f3c9

Observation c61d8e86-da4b-49e6-a4c5-6cf38dbfbb06 · outbound

This paper cites AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.111674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.111674Z digest=sha256:e453adaff27179c0504d8e965f2856d24e3717aea46b4bb17d474e1be2f26b81

Observation f1c5f390-96dd-4313-b408-6294732e2a52 · outbound

This paper cites Masked autoencoders are scalable vision learners.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Masked autoencoders are scalable vision learners

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.871666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.117261Z digest=sha256:9a08b702fd7f8fb6498c019658aa1a3e21dcfd77549d40f090d98e72ccc87bba

Observation 0bad9104-fd8f-40cb-b153-c7c5ddda50a0 · outbound

This paper cites Denoising dif- fusion probabilistic models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Denoising dif- fusion probabilistic models

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.861625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.121491Z digest=sha256:71e3ebbe42c435c8a74baca518737097bc857492ebc820a8eb828cf7512cf3eb

Observation 70106e01-8634-4d2a-a464-c6636c5f02f8 · outbound

This paper cites Gritsenko, William Chan, Mohammad Norouzi, and David J.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Gritsenko, William Chan, Mohammad Norouzi, and David J

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.850704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.125562Z digest=sha256:448eab4f21bbaf07c01d58ca11233db96aa148eec16d48b8cf2672f25e0c07ef

Observation 22eb0cd0-27e1-4441-a6ff-efafc57f659a · outbound

This paper cites Lite- flownet: A lightweight convolutional neural network for op- tical flow estimation.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Lite- flownet: A lightweight convolutional neural network for op- tical flow estimation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.838500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.130447Z digest=sha256:078694410fd252a383b7e4fed74652ff5b099804b3d515c1bbe7bbc7a0c73378

Observation 303cfef8-470b-4e47-955a-e4b3ab7fedd8 · outbound

This paper cites Vmc: Video motion customization using temporal attention adap- tion for text-to-video diffusion models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Vmc: Video motion customization using temporal attention adap- tion for text-to-video diffusion models

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.825108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.135093Z digest=sha256:0409ac9e081a6be8218c3aff3266e727ba1e31c08f5a14f65b39748b74dfb043

Observation 2fbb744c-715f-4aa0-a16d-64183e423ae0 · outbound

This paper cites Imagic: Text-based real image editing with diffusion models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Imagic: Text-based real image editing with diffusion models

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.813063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.139775Z digest=sha256:b965e50a8d8841656b0d90dcca59c8054234a20c1d9ef3a8dcfa1a25581d89cc

Observation ba7ee9c9-93c2-4980-818b-b5d2eec3005a · outbound

This paper cites Dif- fusionclip: Text-guided diffusion models for robust image manipulation.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Dif- fusionclip: Text-guided diffusion models for robust image manipulation

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.800802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.144193Z digest=sha256:8047b21cb7187e856299f96a08d23f5c2bab96575c7330cefbe23d0fe667fc1d

Observation 21d1b036-6e13-49c2-9cfa-3c7771612fc2 · outbound

This paper cites Segment any- thing.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Segment any- thing

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.788808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.148464Z digest=sha256:812fd668b78b68a81f95a03573edddabe658ac98ec3660fafdf2b8b9e0bb108f

Observation 1c400dbe-334a-4ce3-bc64-2226a555fd0a · outbound

This paper cites Open-sora-plan, 2024.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Open-sora-plan, 2024

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.776350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.152779Z digest=sha256:54a612b6a40fc6892853969e8703db7bcaf0b584382baac8b299534e1507e465

Observation 28fe7f0e-30d1-4f9b-a908-377f16bc1f8a · outbound

This paper cites Learning blind video temporal consistency.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Learning blind video temporal consistency

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.764665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.157139Z digest=sha256:98b724edecce2e5a892093a26b69266c39ced8a30660ebca9546b05ab6f5acec

Observation 7b8feac7-bac0-47ea-910b-a825789673d9 · outbound

This paper cites Generative image dynamics.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Generative image dynamics

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.750869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.161189Z digest=sha256:a2cd2b3429e61a44a3739bf61cdd04068ffa71840b6dcb2539e71341a96b942c

Observation 6bf4ffda-2075-41f3-9724-4cd81b2aed1c · outbound

This paper cites Video-p2p: Video editing with cross-attention control.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Video-p2p: Video editing with cross-attention control

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.735746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.165082Z digest=sha256:62c445a0c147709413b9ef27c7a3731f95c240aa4d9bc1abcfd87f173645bcf4

Observation 351b6968-a739-4a8d-9c88-5f125db30d89 · outbound

This paper cites Null-text inversion for editing real images using guided diffusion models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Null-text inversion for editing real images using guided diffusion models

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.721659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.169506Z digest=sha256:aabe393b90a2fbd11fc07e11c807150220d9ebaaeaf2c6ec85cd8a63e39f28eb

Observation b7d3f3d1-ae61-48c9-8b46-e5f3b9286c63 · outbound

This paper cites Dreamix: Video Diffusion Models are General Video Editors.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Dreamix: Video Diffusion Models are General Video Editors

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.174350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.174350Z digest=sha256:3aa92c9e2d33165ce322caa81d6e74c95b416a840918bc6d3506e329faa23a8d

Observation ff936bca-0c05-49c4-b8f1-d60b135f20cb · outbound

This paper cites GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.178964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.178964Z digest=sha256:1c3f3aee782e6081f18242504ef170eb14b35a1ac163c95c5e1e1e3ba5d27776

Observation 517e7d7a-a1e1-41e9-8386-0a1cad6064ae · outbound

This paper cites The best free stock photos, royalty free images & videos shared by creators.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling The best free stock photos, royalty free images & videos shared by creators

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.707660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.183312Z digest=sha256:ebce05b2bf143efc395f2caaa7103d7777c361dd44be2214ebba231f6dcd5f2b

Observation cff524ce-59b4-409a-95cb-5f344c834a7f · outbound

This paper cites Fatezero: Fus- ing attentions for zero-shot text-based video editing.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Fatezero: Fus- ing attentions for zero-shot text-based video editing

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.694231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.186953Z digest=sha256:ad240888410a21effce8de5c5c09ce51dd48da6e8a90f90287c2731dacc13a92

Observation e44e7ee3-8a06-474e-841b-62de9cf95fcd · outbound

This paper cites High-resolution image syn- thesis with latent diffusion models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling High-resolution image syn- thesis with latent diffusion models

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.680598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.189983Z digest=sha256:4ea6468139f8b40ac2c2f7584ebb75fbc505c10009bace8459e8547a0c05e082

Observation f95c7446-3d04-448c-94d5-0893dc9c7a6a · outbound

This paper cites pytorch-fid: FID Score for PyTorch.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling pytorch-fid: FID Score for PyTorch

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.193229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.193229Z digest=sha256:e19a8707663d3cc7bf6fbad8761cd35cd812dc68d52d2282b7b4cf51f1e26bd5

Observation 318fe1bd-0d9b-4182-a75b-4daf35332f0c · outbound

This paper cites Motion-i2v: Consistent and controllable image-to-video generation with explicit motion modeling.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Motion-i2v: Consistent and controllable image-to-video generation with explicit motion modeling

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.658961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.197259Z digest=sha256:d6b8ed92fde3fe2c1d666f21f1523cd5e5b668093d623af643d882bc0ea92888

Observation 25c25d60-4a7e-4505-91ac-52ecb2aad6a5 · outbound

This paper cites Dragdiffusion: Harnessing diffusion models for interactive point-based image editing.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Dragdiffusion: Harnessing diffusion models for interactive point-based image editing

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.647253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.201069Z digest=sha256:8b816200624694619803d3f5f2ce5fb6e5235688afedee2aa4072182df5427fa

Observation 8536298f-97f5-47d7-8605-5557146a5510 · outbound

This paper cites Make-A-Video: Text-to-Video Generation without Text-Video Data.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Make-A-Video: Text-to-Video Generation without Text-Video Data

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.205165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.205165Z digest=sha256:7484d676bf847aca5d865ee932a4cf8a5073366860cbbc7d4673d2fd5576b450

Observation 083f91fd-af46-4d96-9ebc-9182b87255c9 · outbound

This paper cites Denoising Diffusion Implicit Models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Denoising Diffusion Implicit Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.210493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.210493Z digest=sha256:712e5c43b46784e90551fe5544d6ab22445ee72d832a77982652bebad2af84ad

Observation c352fc04-3141-449a-adb6-56113b1a92d3 · outbound

This paper cites Score-based generative modeling through stochastic differential equa- tions.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Score-based generative modeling through stochastic differential equa- tions

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.634223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.214836Z digest=sha256:050c116dcea66fa4c9a7ffee1f74e24f4101a73fc5bfa2db45d46d21c8abbcab

Observation 502916d0-8bbb-479d-8e24-5307fb3577bb · outbound

This paper cites Videomae: Masked autoencoders are data-efficient learners for self-supervised video pre-training.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Videomae: Masked autoencoders are data-efficient learners for self-supervised video pre-training

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.618168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.218661Z digest=sha256:f27f507956869d2b3c182443a1437accca045096c0aa1a6d3e7938b929f9d74f

Observation ab4bff79-767a-4a6e-b913-a29a6cfb6975 · outbound

This paper cites Zero-Shot Video Editing Using Off-The-Shelf Image Diffusion Models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Zero-Shot Video Editing Using Off-The-Shelf Image Diffusion Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.222710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.222710Z digest=sha256:01f290ac59d63f96dd74599487bf8164756cdbc9eefecd284a8e04eb7fea7949

Observation 70055fc3-79bc-432d-a0cc-067148524f7c · outbound

This paper cites Latent Image Animator: Learning to Animate Images via Latent Space Navigation.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Latent Image Animator: Learning to Animate Images via Latent Space Navigation

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.227332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.227332Z digest=sha256:53502092e76d4c67fa8079a1d399b1b285ffc5694dac69bbefc68704939e81a8

Observation 199495d0-fb7e-4021-9684-0400d74f479a · outbound

This paper cites Dynamicrafter: Animating open-domain images with video diffusion priors.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Dynamicrafter: Animating open-domain images with video diffusion priors

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.604875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.231637Z digest=sha256:f06cd80d71fc17f50e238123177ddb0fade49fd72ae2910cfd8cf2ff2c7b2c2b

Observation 1d8ae152-cffd-4991-a443-6e2d8181d4d4 · outbound

This paper cites Gmflow: Learning optical flow via global matching.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Gmflow: Learning optical flow via global matching

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.591772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.235517Z digest=sha256:c693a75cc0bf2f03e02318f6011cc965c19083a615da17b2240d8e551852cb57

Observation 9910b07b-a008-48bb-9a9e-ae9c377efd05 · outbound

This paper cites MagicProp: Diffusion-based Video Editing via Motion-aware Appearance Propagation.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling MagicProp: Diffusion-based Video Editing via Motion-aware Appearance Propagation

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-08-10T21:41:23.352369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.239654Z digest=sha256:4fece7acc6ed8819075676983990aa47fa47ea6c840944dbe0342bd7c832a411

Observation c0ed8248-855a-4756-8e54-53116e186248 · outbound

This paper cites Motion-Conditioned Image Animation for Video Editing.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Motion-Conditioned Image Animation for Video Editing

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.244051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.244051Z digest=sha256:72171d707fca120b0854f239c7c74f45af8d379b89306b2bea42def587d122ce

Observation b5c1d703-ac6c-4917-83d6-66ab981fce06 · outbound

This paper cites Paint by example: Exemplar-based image editing with diffusion models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Paint by example: Exemplar-based image editing with diffusion models

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.576761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.248392Z digest=sha256:8db8d48be0a6cefa91d88af71b9b4722b54743b2106a5228396eb1a309ee319b

Observation 2c8b40b6-e398-46f4-a5b6-572285736344 · outbound

This paper cites Depth anything: Unleashing the power of large-scale unlabeled data.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Depth anything: Unleashing the power of large-scale unlabeled data

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.561239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.252222Z digest=sha256:0fae777c63f067d7a4e639cf6922ee1dbb2204f06813cc201ffb9894534bf934

Observation d8a2c9ab-1301-4f81-9103-53f53e9856b8 · outbound

This paper cites Rerender a video: Zero-shot text-guided video-to-video translation.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Rerender a video: Zero-shot text-guided video-to-video translation

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.547123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.256073Z digest=sha256:36fd1010384c44c587506c0982268ef833cfd3d88b3914ff149aa6952590fe99

Observation c8588efb-512a-4e5b-b889-c58d44f1794e · outbound

This paper cites IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.259850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.259850Z digest=sha256:fa5b6bad119fdef8bd77f6ab4ad9ed93d56f6e5babb627465f3a5d9be8896eb6

Observation 0171d1bb-3607-4c92-97a5-a5bcbec45c7a · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Adding conditional control to text-to-image diffusion models

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.534380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.264348Z digest=sha256:a9aa88cfd4d767905c792466d8e117341b3a65c197c964e7b46e81e1560b88dd

Observation 9dd04a1e-5fb1-45db-877a-6230a2db62ba · outbound

This paper cites Sine: Single image editing with text-to-image diffusion models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Sine: Single image editing with text-to-image diffusion models

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.522591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.268609Z digest=sha256:40c74efdb88c748ce505113f45c4b011a49045675b2e2a4acf717569697cb7e0

Observation 9c9fb6a3-1622-4fcc-9732-5645b813b3f3 · outbound

This paper cites Avid: Any-length video inpainting with dif- fusion model.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Avid: Any-length video inpainting with dif- fusion model

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.509453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.273112Z digest=sha256:d4e7f84c30ac1cef035c2cb3905d4116f423f0cf0696f05f09525a8aad6ca22e

Observation 655872eb-92ff-4783-a0db-3fd0baa50417 · outbound

This paper cites Motiondirector: Motion customization of text-to-video diffusion models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Motiondirector: Motion customization of text-to-video diffusion models

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.494936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.277490Z digest=sha256:2fad42ec27ee820d9637c5fc1917902265143a9b36b5c784fbc1e11b69842739

Observation b79f31ca-b2ec-4a36-9ceb-54da19575f76 · outbound

This paper cites clip-score: CLIP Score for Py- Torch.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling clip-score: CLIP Score for Py- Torch

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.480291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T21:41:23.282342Z digest=sha256:13bf6ed4170400e4143333c9324c818abcf8fdcdfeb4c228327c19a8de6f1d13

Pith citing papers

No inbound Pith citation observations are available.