Pith. sign in

Paper Citation Record · LEDGER

Edit as You See: Image-guided Video Editing via Masked Motion Modeling

As of 20 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 0 inbound Pith citation observations for arXiv:2501.04325.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.04325 v1

Coverage vector

measured 54 of 54 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T21:41:23.282342Z

measured 54 of 54 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

54 of 54 outbound references displayed

  • verified exact1
  • verified fuzzy40
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 42b8a104-7d5a-4122-9737-66354e2c9150 · outbound

This paper cites Blended diffusion for text-driven editing of natural images.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Blended diffusion for text-driven editing of natural images

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.989202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.048302Z digest=sha256:5c95c773aabf62b1b09f3ecc31fb4baf6a76f3285f17ddaaa6e1286df5c7e0ca

Observation 9cb97f2d-8cf6-42a1-a1c2-d3eb9da35884 · outbound

This paper cites Text2live: Text-driven layered image and video editing.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Text2live: Text-driven layered image and video editing

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.978351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.053362Z digest=sha256:ee3d7fbe7662c9057df431ba6e61812f8b7b048c9f6b3b20dfdd027feb5e719d

Observation 0fcf37a5-6e79-4c3b-99c2-d52311add416 · outbound

This paper cites In- structpix2pix: Learning to follow image editing instructions.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling In- structpix2pix: Learning to follow image editing instructions

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.967675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.057864Z digest=sha256:db9db40a4f4190a01c9f4ea1cc0911b15071d05a868bba9e79821dda2c2abfcb

Observation 3e7be476-1869-486d-9b28-9808f15909bb · outbound

This paper cites Pix2video: Video editing using image diffusion.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Pix2video: Video editing using image diffusion

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.956733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.063303Z digest=sha256:ad41c4b43174c6f0894c1d03ec23ed5822af7e40332ac8e4738d29cc8e0242eb

Observation 3fa87c73-9e85-4344-8e4e-48911593c4de · outbound

This paper cites Stable- video: Text-driven consistency-aware diffusion video edit- ing.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Stable- video: Text-driven consistency-aware diffusion video edit- ing

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.944429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.069225Z digest=sha256:97277053e892f43692eafae6e45a5fe24d93f176606cc70924c8791740db8ec2

Observation a0f5fde4-86e9-4423-b57e-e194863c68c0 · outbound

This paper cites Zero-shot Image Editing with Reference Imitation.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Zero-shot Image Editing with Reference Imitation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.074303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.074303Z digest=sha256:d5fb788ed5d5c4a50ea7c2739fe903b3962dddcd8e9a9654dbf5f357e9f44a0b

Observation 0e6f196c-371a-4f18-81ea-501dc8838046 · outbound

This paper cites Xmem: Long- term video object segmentation with an atkinson-shiffrin memory model.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Xmem: Long- term video object segmentation with an atkinson-shiffrin memory model

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.931085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.079713Z digest=sha256:a51764922da9936f063b46467c0f21dfbad9000edf00ef012a57e9ced5e314ae

Observation cc5a1e3d-7f05-402e-9b34-ebf73f9ca1d0 · outbound

This paper cites Compvis/stable-diffusion: A latent text-to- image diffusion model.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Compvis/stable-diffusion: A latent text-to- image diffusion model

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.918275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.084951Z digest=sha256:f8ec99072b5430b388b64c0ed0097ce28a500dae3cf849f4e0ee9d9ad1ad12ad

Observation 392b5eee-b0e5-4756-9ac6-ae9d16393a0d · outbound

This paper cites DiffEdit: Diffusion-based semantic image editing with mask guidance.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling DiffEdit: Diffusion-based semantic image editing with mask guidance

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.090284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.090284Z digest=sha256:c936720ff0ddb00b3038767d6d370a09bd6ffc4e34c4da63d51feaea6689a80c

Observation 8e3650e1-78ba-4ee5-bcdf-3b5ca474cc2a · outbound

This paper cites Videdit: Zero-shot and spatially aware text-driven video editing.IEEE Trans.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Videdit: Zero-shot and spatially aware text-driven video editing.IEEE Trans

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.906371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.095632Z digest=sha256:2c98728482c766f2e9966bf9dee8d4f30527faea04f6a11e37fe4de20cce4a00

Observation 12e6d0ba-8e2b-4c86-b713-2e5465bd330a · outbound

This paper cites Diffusion models beat gans on image synthesis.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Diffusion models beat gans on image synthesis

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.894691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.100475Z digest=sha256:8ab33b61f95c53211c722891cc18c265309be9871bfeed25180ba40eaf3709c6

Observation b9ff7a3d-bf18-4496-8175-a42f42c738b5 · outbound

This paper cites Editanything: Empower- ing unparalleled flexibility in image editing and generation.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Editanything: Empower- ing unparalleled flexibility in image editing and generation

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.882900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.104235Z digest=sha256:6defdb13f0431cd080541e229617d0a9eecd1c9ffcafe54785ed87fd0507b97f

Observation 499e42d0-3b1c-4c68-be94-31de4bc362c1 · outbound

This paper cites TokenFlow: Consistent Diffusion Features for Consistent Video Editing.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling TokenFlow: Consistent Diffusion Features for Consistent Video Editing

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.107577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.107577Z digest=sha256:5cf43bb7522d8031a809f45255391666c2bf305415752ddecc30d3fe428315e1

Observation c61d8e86-da4b-49e6-a4c5-6cf38dbfbb06 · outbound

This paper cites AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.111674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.111674Z digest=sha256:4cb2aff2b186c64f52f49473f7c2169e62117daf4405167fc8228760f9b545ce

Observation f1c5f390-96dd-4313-b408-6294732e2a52 · outbound

This paper cites Masked autoencoders are scalable vision learners.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Masked autoencoders are scalable vision learners

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.871666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.117261Z digest=sha256:c3954f4b05ec51a3d53ae563a3d2e4cc88e22c56a4f967b205c9782ee27cad82

Observation 0bad9104-fd8f-40cb-b153-c7c5ddda50a0 · outbound

This paper cites Denoising dif- fusion probabilistic models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Denoising dif- fusion probabilistic models

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.861625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.121491Z digest=sha256:12d5efbe509f62b26d7fbe6f0dafc2781cce2e9825f766464a5f675f36bd8329

Observation 70106e01-8634-4d2a-a464-c6636c5f02f8 · outbound

This paper cites Gritsenko, William Chan, Mohammad Norouzi, and David J.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Gritsenko, William Chan, Mohammad Norouzi, and David J

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.850704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.125562Z digest=sha256:6d1234ec72832fc0584fb20cc2f8080c3ca9529c364087c4bb3619276b3b7e89

Observation 22eb0cd0-27e1-4441-a6ff-efafc57f659a · outbound

This paper cites Lite- flownet: A lightweight convolutional neural network for op- tical flow estimation.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Lite- flownet: A lightweight convolutional neural network for op- tical flow estimation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.838500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.130447Z digest=sha256:6b98ae58b74371c19071147f2b4b9f5b471ff589fb6aac3d43a8408bdbd44d7e

Observation 303cfef8-470b-4e47-955a-e4b3ab7fedd8 · outbound

This paper cites Vmc: Video motion customization using temporal attention adap- tion for text-to-video diffusion models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Vmc: Video motion customization using temporal attention adap- tion for text-to-video diffusion models

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.825108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.135093Z digest=sha256:30ccea9e8c4ce1a80b29817b3f0dd92e943df3f0d39df4b51a8a1ddcad963816

Observation 2fbb744c-715f-4aa0-a16d-64183e423ae0 · outbound

This paper cites Imagic: Text-based real image editing with diffusion models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Imagic: Text-based real image editing with diffusion models

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.813063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.139775Z digest=sha256:a17a651e27a84e72cce130223f8f5fea3fb8e57648ac365b297738b06c6932af

Observation ba7ee9c9-93c2-4980-818b-b5d2eec3005a · outbound

This paper cites Dif- fusionclip: Text-guided diffusion models for robust image manipulation.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Dif- fusionclip: Text-guided diffusion models for robust image manipulation

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.800802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.144193Z digest=sha256:c6fa200b25fc16348988efac914afbef3247674bcd3279c1d77b7165afa7d635

Observation 21d1b036-6e13-49c2-9cfa-3c7771612fc2 · outbound

This paper cites Segment any- thing.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Segment any- thing

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.788808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.148464Z digest=sha256:63c333b9d61fc5e64037d606a1809bc99ede22870f476268f23fdd167e9e535a

Observation 1c400dbe-334a-4ce3-bc64-2226a555fd0a · outbound

This paper cites Open-sora-plan, 2024.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Open-sora-plan, 2024

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.776350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.152779Z digest=sha256:79f5a13273d77da279a61bf7631b2276be859162f1fdf6ea996636d350a83015

Observation 28fe7f0e-30d1-4f9b-a908-377f16bc1f8a · outbound

This paper cites Learning blind video temporal consistency.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Learning blind video temporal consistency

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.764665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.157139Z digest=sha256:af65d0034ccefad71b5bc58d72ed57a25c95a189bd1d51daa887a49c2e5e8e8a

Observation 7b8feac7-bac0-47ea-910b-a825789673d9 · outbound

This paper cites Generative image dynamics.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Generative image dynamics

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.750869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.161189Z digest=sha256:8cb87f691c317cc8614930be22813cb04436fb04888c701591d91599e08dc328

Observation 6bf4ffda-2075-41f3-9724-4cd81b2aed1c · outbound

This paper cites Video-p2p: Video editing with cross-attention control.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Video-p2p: Video editing with cross-attention control

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.735746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.165082Z digest=sha256:aaef7c4744a5e80f88e63593b831ec147e665aaa0ce22c215951b40de6d570d3

Observation 351b6968-a739-4a8d-9c88-5f125db30d89 · outbound

This paper cites Null-text inversion for editing real images using guided diffusion models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Null-text inversion for editing real images using guided diffusion models

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.721659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.169506Z digest=sha256:e438fb73d87dfbdcd22463e76f96b19a6bf2eabb378b600c99c6f468b8cc2f51

Observation b7d3f3d1-ae61-48c9-8b46-e5f3b9286c63 · outbound

This paper cites Dreamix: Video Diffusion Models are General Video Editors.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Dreamix: Video Diffusion Models are General Video Editors

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.174350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.174350Z digest=sha256:acc14d9ec1c3ec490dbaaac28b3197baa139e5b83af25904390e4ea4175c2353

Observation ff936bca-0c05-49c4-b8f1-d60b135f20cb · outbound

This paper cites GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.178964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.178964Z digest=sha256:816458907c8afbb64972041eed1f46b76389f3921774c5c5869dfd21b9d3c2ac

Observation 517e7d7a-a1e1-41e9-8386-0a1cad6064ae · outbound

This paper cites The best free stock photos, royalty free images & videos shared by creators.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling The best free stock photos, royalty free images & videos shared by creators

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.707660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.183312Z digest=sha256:3efcbe923d2103180fe467960fff468a217d568e757240ecfd02d63d1fac2ba8

Observation cff524ce-59b4-409a-95cb-5f344c834a7f · outbound

This paper cites Fatezero: Fus- ing attentions for zero-shot text-based video editing.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Fatezero: Fus- ing attentions for zero-shot text-based video editing

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.694231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.186953Z digest=sha256:5fca57ed1c9349fbfed331aa42cdffaace263569873612c1cb6d130fe884884e

Observation e44e7ee3-8a06-474e-841b-62de9cf95fcd · outbound

This paper cites High-resolution image syn- thesis with latent diffusion models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling High-resolution image syn- thesis with latent diffusion models

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.680598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.189983Z digest=sha256:b92fd193c0974d032fa8a6caf132196a38eeb75adbfc9b521d9f8e9a18f75c39

Observation f95c7446-3d04-448c-94d5-0893dc9c7a6a · outbound

This paper cites pytorch-fid: FID Score for PyTorch.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling pytorch-fid: FID Score for PyTorch

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.193229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.193229Z digest=sha256:f8f1e52d2d8a8481cf55f90b0f32f1c0d1c42b09db9bf4f61d58318a09308901

Observation 318fe1bd-0d9b-4182-a75b-4daf35332f0c · outbound

This paper cites Motion-i2v: Consistent and controllable image-to-video generation with explicit motion modeling.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Motion-i2v: Consistent and controllable image-to-video generation with explicit motion modeling

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.658961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.197259Z digest=sha256:ba3f95da1a800a1b26ed0d191ac950296e14049d1a0b68f480998353a3ee3d7e

Observation 25c25d60-4a7e-4505-91ac-52ecb2aad6a5 · outbound

This paper cites Dragdiffusion: Harnessing diffusion models for interactive point-based image editing.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Dragdiffusion: Harnessing diffusion models for interactive point-based image editing

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.647253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.201069Z digest=sha256:f6dd6079f1fe0c74155543f325c1204483455ffbf4f85510196c5c2e88255f0e

Observation 8536298f-97f5-47d7-8605-5557146a5510 · outbound

This paper cites Make-A-Video: Text-to-Video Generation without Text-Video Data.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Make-A-Video: Text-to-Video Generation without Text-Video Data

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.205165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.205165Z digest=sha256:2f8a17d23fdf545f932b15d6071df2b6415eabe170a6eae756d8421e8e749ed0

Observation 083f91fd-af46-4d96-9ebc-9182b87255c9 · outbound

This paper cites Denoising Diffusion Implicit Models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Denoising Diffusion Implicit Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.210493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.210493Z digest=sha256:bd9b464ddd321ada4b82c1e9743ab23337cfe36af54411710aa0adeed353152f

Observation c352fc04-3141-449a-adb6-56113b1a92d3 · outbound

This paper cites Score-based generative modeling through stochastic differential equa- tions.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Score-based generative modeling through stochastic differential equa- tions

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.634223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.214836Z digest=sha256:fbee2eaedfaa0cc4d9d55387683f47f0fbb8b904b57b404cc3ed033ef2595db6

Observation 502916d0-8bbb-479d-8e24-5307fb3577bb · outbound

This paper cites Videomae: Masked autoencoders are data-efficient learners for self-supervised video pre-training.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Videomae: Masked autoencoders are data-efficient learners for self-supervised video pre-training

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.618168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.218661Z digest=sha256:55fdc5365c94126f9fc083dd6e5757a7ab22fd84debaa9edde5f37fbef57ac94

Observation ab4bff79-767a-4a6e-b913-a29a6cfb6975 · outbound

This paper cites Zero-Shot Video Editing Using Off-The-Shelf Image Diffusion Models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Zero-Shot Video Editing Using Off-The-Shelf Image Diffusion Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.222710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.222710Z digest=sha256:3bbb29f19ec9c6095a3a89b2620a5403a65541ad8db1bd5398787b1dbdbc96d8

Observation 70055fc3-79bc-432d-a0cc-067148524f7c · outbound

This paper cites Latent Image Animator: Learning to Animate Images via Latent Space Navigation.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Latent Image Animator: Learning to Animate Images via Latent Space Navigation

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.227332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.227332Z digest=sha256:01a2b04351c46560f642abd33ab2cb567cf91bdbd54f35185ed071bcb6d49127

Observation 199495d0-fb7e-4021-9684-0400d74f479a · outbound

This paper cites Dynamicrafter: Animating open-domain images with video diffusion priors.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Dynamicrafter: Animating open-domain images with video diffusion priors

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.604875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.231637Z digest=sha256:0ba621319d4d3e9079c814d993c9399304321c21d4cdd92c21751e0b282f3d18

Observation 1d8ae152-cffd-4991-a443-6e2d8181d4d4 · outbound

This paper cites Gmflow: Learning optical flow via global matching.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Gmflow: Learning optical flow via global matching

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.591772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.235517Z digest=sha256:145c19f31c5c7b7e10660cfe93a7e666a577525488c2dc4fd5c5a3eb42a3921e

Observation 9910b07b-a008-48bb-9a9e-ae9c377efd05 · outbound

This paper cites MagicProp: Diffusion-based Video Editing via Motion-aware Appearance Propagation.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling MagicProp: Diffusion-based Video Editing via Motion-aware Appearance Propagation

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-08-10T21:41:23.352369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.239654Z digest=sha256:ad38f0ad7c0935275fcc6d51941f0e4c2a5515c01105bf7615f381355717bb2b

Observation c0ed8248-855a-4756-8e54-53116e186248 · outbound

This paper cites Motion-Conditioned Image Animation for Video Editing.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Motion-Conditioned Image Animation for Video Editing

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.244051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.244051Z digest=sha256:914e29ebb379a91ea34f3a850eddcf90c80694d17b7f5e2dc41029eb7ac6749e

Observation b5c1d703-ac6c-4917-83d6-66ab981fce06 · outbound

This paper cites Paint by example: Exemplar-based image editing with diffusion models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Paint by example: Exemplar-based image editing with diffusion models

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.576761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.248392Z digest=sha256:072977f85143563b118fedad5b78b4c36d7d2f8ef2875375730a97e9c7720391

Observation 2c8b40b6-e398-46f4-a5b6-572285736344 · outbound

This paper cites Depth anything: Unleashing the power of large-scale unlabeled data.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Depth anything: Unleashing the power of large-scale unlabeled data

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.561239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.252222Z digest=sha256:f61f0bb65f5e67aecf13620ba815a739860bed8de501b60a76a4d3c098b0e899

Observation d8a2c9ab-1301-4f81-9103-53f53e9856b8 · outbound

This paper cites Rerender a video: Zero-shot text-guided video-to-video translation.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Rerender a video: Zero-shot text-guided video-to-video translation

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.547123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.256073Z digest=sha256:0b492bc58e339ceb93b01f6d0aefe09f6d6e3dd258bda4d7387d701bdd05439d

Observation c8588efb-512a-4e5b-b889-c58d44f1794e · outbound

This paper cites IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.259850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.259850Z digest=sha256:e8a1543c65d7a23fca4821186f7c759031aac0d7a621f8f5c3cb2549b1703ced

Observation 0171d1bb-3607-4c92-97a5-a5bcbec45c7a · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Adding conditional control to text-to-image diffusion models

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.534380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.264348Z digest=sha256:72b20da7488ab080b2ad9c5d8a71dcd5ec268a60798363711117a69dbd9b54f3

Observation 9dd04a1e-5fb1-45db-877a-6230a2db62ba · outbound

This paper cites Sine: Single image editing with text-to-image diffusion models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Sine: Single image editing with text-to-image diffusion models

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.522591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.268609Z digest=sha256:ed50335be9a13e798e55e5648bcd0726533fafac4470bc80c4aae8d9468e3da6

Observation 9c9fb6a3-1622-4fcc-9732-5645b813b3f3 · outbound

This paper cites Avid: Any-length video inpainting with dif- fusion model.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Avid: Any-length video inpainting with dif- fusion model

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.509453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.273112Z digest=sha256:588cc1fdd62f18b8438951abd5856c10b227cc07168ad7b263acb858b45e149b

Observation 655872eb-92ff-4783-a0db-3fd0baa50417 · outbound

This paper cites Motiondirector: Motion customization of text-to-video diffusion models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Motiondirector: Motion customization of text-to-video diffusion models

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.494936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.277490Z digest=sha256:4d29395d787b651c072a1fffb61a8cd8c7c6eaf7c3b2459c93f3fb2423cac855

Observation b79f31ca-b2ec-4a36-9ceb-54da19575f76 · outbound

This paper cites clip-score: CLIP Score for Py- Torch.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling clip-score: CLIP Score for Py- Torch

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.480291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T21:41:23.282342Z digest=sha256:ba0cad1097cb588f36bc0e3045ebf98b714fe75b47804e3ded5be09613391e30

Pith citing papers

No inbound Pith citation observations are available.