Pith. sign in

Paper Citation Record · LEDGER

Edit as You See: Image-guided Video Editing via Masked Motion Modeling

As of 11 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 0 inbound Pith citation observations for arXiv:2501.04325.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.04325 v1

Coverage vector

measured 54 of 54 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T21:41:23.282342Z

measured 54 of 54 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

54 of 54 outbound references displayed

  • verified exact1
  • verified fuzzy40
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 42b8a104-7d5a-4122-9737-66354e2c9150 · outbound

This paper cites Blended diffusion for text-driven editing of natural images.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Blended diffusion for text-driven editing of natural images

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.989202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.048302Z digest=sha256:e4becb0c8826e5464958acf627e3e06be5dec14c0d453fdbc00e798ac0102389

Observation 9cb97f2d-8cf6-42a1-a1c2-d3eb9da35884 · outbound

This paper cites Text2live: Text-driven layered image and video editing.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Text2live: Text-driven layered image and video editing

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.978351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.053362Z digest=sha256:79cf70e4dbd8a4cc3950c8f6b073fe2bd78f96dd25c3d2e1adb4b99ab2a0f2c9

Observation 0fcf37a5-6e79-4c3b-99c2-d52311add416 · outbound

This paper cites In- structpix2pix: Learning to follow image editing instructions.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling In- structpix2pix: Learning to follow image editing instructions

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.967675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.057864Z digest=sha256:7a04100a1c9a5ff7bcc23ebd67fdec02a3721fad313b66d3488ada4a2f6a1ecc

Observation 3e7be476-1869-486d-9b28-9808f15909bb · outbound

This paper cites Pix2video: Video editing using image diffusion.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Pix2video: Video editing using image diffusion

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.956733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.063303Z digest=sha256:4d658812539bdd76e1f065b5182226404e2faafb77d176ed500200e394fcd5b6

Observation 3fa87c73-9e85-4344-8e4e-48911593c4de · outbound

This paper cites Stable- video: Text-driven consistency-aware diffusion video edit- ing.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Stable- video: Text-driven consistency-aware diffusion video edit- ing

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.944429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.069225Z digest=sha256:55dcc8c2ef69f2a1d56c2f00ff08f2ff825603abeb5890e9a8409266f68d7173

Observation a0f5fde4-86e9-4423-b57e-e194863c68c0 · outbound

This paper cites Zero-shot Image Editing with Reference Imitation.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Zero-shot Image Editing with Reference Imitation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.074303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.074303Z digest=sha256:71dc86c74e2ff97cbf03c6f500a5cd5b92d8e534296d538a58aa557a1491a925

Observation 0e6f196c-371a-4f18-81ea-501dc8838046 · outbound

This paper cites Xmem: Long- term video object segmentation with an atkinson-shiffrin memory model.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Xmem: Long- term video object segmentation with an atkinson-shiffrin memory model

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.931085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.079713Z digest=sha256:010f97e72abfce7dc54dd6ada9757f1635d01703eec3b8fc716f502d52dee2f1

Observation cc5a1e3d-7f05-402e-9b34-ebf73f9ca1d0 · outbound

This paper cites Compvis/stable-diffusion: A latent text-to- image diffusion model.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Compvis/stable-diffusion: A latent text-to- image diffusion model

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.918275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.084951Z digest=sha256:05bacb77d3c587d61c80a5ef122524ee0f019ba6f7660977fc1df47cf21e9873

Observation 392b5eee-b0e5-4756-9ac6-ae9d16393a0d · outbound

This paper cites DiffEdit: Diffusion-based semantic image editing with mask guidance.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling DiffEdit: Diffusion-based semantic image editing with mask guidance

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.090284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.090284Z digest=sha256:c0c1412e37c735fca60dfba5ba407ef10e8aa0979d08c8d5b0a1dc1a6d167285

Observation 8e3650e1-78ba-4ee5-bcdf-3b5ca474cc2a · outbound

This paper cites Videdit: Zero-shot and spatially aware text-driven video editing.IEEE Trans.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Videdit: Zero-shot and spatially aware text-driven video editing.IEEE Trans

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.906371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.095632Z digest=sha256:a5cbab9e82452900c1c832686da7a0fa6f683860d178f04f6668369e2ebbbdb5

Observation 12e6d0ba-8e2b-4c86-b713-2e5465bd330a · outbound

This paper cites Diffusion models beat gans on image synthesis.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Diffusion models beat gans on image synthesis

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.894691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.100475Z digest=sha256:353f38061d847ab54396f4d8295c2726a4a3731c98b9732aadf0836d43ecebe0

Observation b9ff7a3d-bf18-4496-8175-a42f42c738b5 · outbound

This paper cites Editanything: Empower- ing unparalleled flexibility in image editing and generation.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Editanything: Empower- ing unparalleled flexibility in image editing and generation

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.882900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.104235Z digest=sha256:f13ed28a79c9716136c7224fb47865641878e8d729e633a340436fead3b04257

Observation 499e42d0-3b1c-4c68-be94-31de4bc362c1 · outbound

This paper cites TokenFlow: Consistent Diffusion Features for Consistent Video Editing.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling TokenFlow: Consistent Diffusion Features for Consistent Video Editing

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.107577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.107577Z digest=sha256:36d92c8ebf8ba15ca2b4ae0c20f4b955a629e69282122afef171c158f9c0f3c9

Observation c61d8e86-da4b-49e6-a4c5-6cf38dbfbb06 · outbound

This paper cites AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.111674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.111674Z digest=sha256:e453adaff27179c0504d8e965f2856d24e3717aea46b4bb17d474e1be2f26b81

Observation f1c5f390-96dd-4313-b408-6294732e2a52 · outbound

This paper cites Masked autoencoders are scalable vision learners.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Masked autoencoders are scalable vision learners

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.871666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.117261Z digest=sha256:a461e39558526fb2d298b781ec54e983f44e8c95344425e9d3c3b63e87de4d92

Observation 0bad9104-fd8f-40cb-b153-c7c5ddda50a0 · outbound

This paper cites Denoising dif- fusion probabilistic models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Denoising dif- fusion probabilistic models

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.861625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.121491Z digest=sha256:fb2fedc95cb27270cefc0ae7d278e5907113d6468cd9454dddf1b5eba36f1c7f

Observation 70106e01-8634-4d2a-a464-c6636c5f02f8 · outbound

This paper cites Gritsenko, William Chan, Mohammad Norouzi, and David J.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Gritsenko, William Chan, Mohammad Norouzi, and David J

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.850704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.125562Z digest=sha256:5d75a6c5c41550c991309cc4327120dafec3caa2e5d7cb6e030ee525fa68eacd

Observation 22eb0cd0-27e1-4441-a6ff-efafc57f659a · outbound

This paper cites Lite- flownet: A lightweight convolutional neural network for op- tical flow estimation.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Lite- flownet: A lightweight convolutional neural network for op- tical flow estimation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.838500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.130447Z digest=sha256:2ca189ff8f16f08f62d461857775f70ed37e4e030661aa0400424e762f724986

Observation 303cfef8-470b-4e47-955a-e4b3ab7fedd8 · outbound

This paper cites Vmc: Video motion customization using temporal attention adap- tion for text-to-video diffusion models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Vmc: Video motion customization using temporal attention adap- tion for text-to-video diffusion models

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.825108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.135093Z digest=sha256:7739ac9bc8e385cb58238d6aac359acb685768652d47f305268390af0117aedc

Observation 2fbb744c-715f-4aa0-a16d-64183e423ae0 · outbound

This paper cites Imagic: Text-based real image editing with diffusion models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Imagic: Text-based real image editing with diffusion models

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.813063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.139775Z digest=sha256:0bcae46c18ddd45db8c9e17378746853d2b840e14457f564ae99b93eeeabfb76

Observation ba7ee9c9-93c2-4980-818b-b5d2eec3005a · outbound

This paper cites Dif- fusionclip: Text-guided diffusion models for robust image manipulation.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Dif- fusionclip: Text-guided diffusion models for robust image manipulation

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.800802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.144193Z digest=sha256:b96015e6f831282a61fa7554286342299b2852d3ea1489c7922649180681fa14

Observation 21d1b036-6e13-49c2-9cfa-3c7771612fc2 · outbound

This paper cites Segment any- thing.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Segment any- thing

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.788808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.148464Z digest=sha256:3b5e461f84997a544d8944cc3b47c7e6677a2a885529353fba3a3a62d9e23bb6

Observation 1c400dbe-334a-4ce3-bc64-2226a555fd0a · outbound

This paper cites Open-sora-plan, 2024.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Open-sora-plan, 2024

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.776350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.152779Z digest=sha256:154579b656fedb4a5cb13da0768e90f0a16be2f49961843cc9078a1af16687ca

Observation 28fe7f0e-30d1-4f9b-a908-377f16bc1f8a · outbound

This paper cites Learning blind video temporal consistency.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Learning blind video temporal consistency

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.764665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.157139Z digest=sha256:b9770094a0838c9e1c563e4a6424e6727b3548cb293916ec59beb473ef5769bc

Observation 7b8feac7-bac0-47ea-910b-a825789673d9 · outbound

This paper cites Generative image dynamics.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Generative image dynamics

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.750869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.161189Z digest=sha256:b1d4d99f01b41fb9a25cc5e7c6a93e0eaf3ab40272809ff07998f7e383169866

Observation 6bf4ffda-2075-41f3-9724-4cd81b2aed1c · outbound

This paper cites Video-p2p: Video editing with cross-attention control.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Video-p2p: Video editing with cross-attention control

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.735746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.165082Z digest=sha256:acb18728438b49285059f74094f2d0aae5a31207321ca1e7a9a9bcd7375fc7b9

Observation 351b6968-a739-4a8d-9c88-5f125db30d89 · outbound

This paper cites Null-text inversion for editing real images using guided diffusion models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Null-text inversion for editing real images using guided diffusion models

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.721659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.169506Z digest=sha256:6d7ec1e0310b748d4c7c4bd439f75d4c4f4c93a034e73f4a0baccf71979f8c6a

Observation b7d3f3d1-ae61-48c9-8b46-e5f3b9286c63 · outbound

This paper cites Dreamix: Video Diffusion Models are General Video Editors.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Dreamix: Video Diffusion Models are General Video Editors

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.174350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.174350Z digest=sha256:3aa92c9e2d33165ce322caa81d6e74c95b416a840918bc6d3506e329faa23a8d

Observation ff936bca-0c05-49c4-b8f1-d60b135f20cb · outbound

This paper cites GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.178964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.178964Z digest=sha256:1c3f3aee782e6081f18242504ef170eb14b35a1ac163c95c5e1e1e3ba5d27776

Observation 517e7d7a-a1e1-41e9-8386-0a1cad6064ae · outbound

This paper cites The best free stock photos, royalty free images & videos shared by creators.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling The best free stock photos, royalty free images & videos shared by creators

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.707660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.183312Z digest=sha256:a41d1944f21b1175e107ffa363350167d2a9327766c90e3417f86fec602945c9

Observation cff524ce-59b4-409a-95cb-5f344c834a7f · outbound

This paper cites Fatezero: Fus- ing attentions for zero-shot text-based video editing.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Fatezero: Fus- ing attentions for zero-shot text-based video editing

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.694231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.186953Z digest=sha256:91e4d5b704843219329db4dee135f32f96f6107b94c5106857abd91aa12d5043

Observation e44e7ee3-8a06-474e-841b-62de9cf95fcd · outbound

This paper cites High-resolution image syn- thesis with latent diffusion models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling High-resolution image syn- thesis with latent diffusion models

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.680598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.189983Z digest=sha256:a25a85d9cc15c3cf176ac06fde9352a2a0ba96fa1236342387421e91bb1eb454

Observation f95c7446-3d04-448c-94d5-0893dc9c7a6a · outbound

This paper cites pytorch-fid: FID Score for PyTorch.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling pytorch-fid: FID Score for PyTorch

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.193229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.193229Z digest=sha256:e19a8707663d3cc7bf6fbad8761cd35cd812dc68d52d2282b7b4cf51f1e26bd5

Observation 318fe1bd-0d9b-4182-a75b-4daf35332f0c · outbound

This paper cites Motion-i2v: Consistent and controllable image-to-video generation with explicit motion modeling.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Motion-i2v: Consistent and controllable image-to-video generation with explicit motion modeling

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.658961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.197259Z digest=sha256:e632e226321b0b8e4f1b1a40fd77625a132d1e3928bda60068b729da03090c2d

Observation 25c25d60-4a7e-4505-91ac-52ecb2aad6a5 · outbound

This paper cites Dragdiffusion: Harnessing diffusion models for interactive point-based image editing.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Dragdiffusion: Harnessing diffusion models for interactive point-based image editing

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.647253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.201069Z digest=sha256:8b13708e5663cd9ceb5ab76f34372cfebed7e1fe62b2df83a20aa869915bb192

Observation 8536298f-97f5-47d7-8605-5557146a5510 · outbound

This paper cites Make-A-Video: Text-to-Video Generation without Text-Video Data.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Make-A-Video: Text-to-Video Generation without Text-Video Data

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.205165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.205165Z digest=sha256:7484d676bf847aca5d865ee932a4cf8a5073366860cbbc7d4673d2fd5576b450

Observation 083f91fd-af46-4d96-9ebc-9182b87255c9 · outbound

This paper cites Denoising Diffusion Implicit Models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Denoising Diffusion Implicit Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.210493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.210493Z digest=sha256:712e5c43b46784e90551fe5544d6ab22445ee72d832a77982652bebad2af84ad

Observation c352fc04-3141-449a-adb6-56113b1a92d3 · outbound

This paper cites Score-based generative modeling through stochastic differential equa- tions.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Score-based generative modeling through stochastic differential equa- tions

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.634223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.214836Z digest=sha256:3fc97f654b82220024b179a2a881a782864f7f0d19eb33a73569cb5f93e2fe3c

Observation 502916d0-8bbb-479d-8e24-5307fb3577bb · outbound

This paper cites Videomae: Masked autoencoders are data-efficient learners for self-supervised video pre-training.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Videomae: Masked autoencoders are data-efficient learners for self-supervised video pre-training

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.618168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.218661Z digest=sha256:5995ecd64f53754ac9b4b730a55ac4a3460ff28f3ef54895568d3a00802abf93

Observation ab4bff79-767a-4a6e-b913-a29a6cfb6975 · outbound

This paper cites Zero-Shot Video Editing Using Off-The-Shelf Image Diffusion Models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Zero-Shot Video Editing Using Off-The-Shelf Image Diffusion Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.222710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.222710Z digest=sha256:01f290ac59d63f96dd74599487bf8164756cdbc9eefecd284a8e04eb7fea7949

Observation 70055fc3-79bc-432d-a0cc-067148524f7c · outbound

This paper cites Latent Image Animator: Learning to Animate Images via Latent Space Navigation.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Latent Image Animator: Learning to Animate Images via Latent Space Navigation

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.227332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.227332Z digest=sha256:c7c6d8fb09d61ea4510a70eb6ee42f57348a024980c58f96094e86eb84875eb7

Observation 199495d0-fb7e-4021-9684-0400d74f479a · outbound

This paper cites Dynamicrafter: Animating open-domain images with video diffusion priors.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Dynamicrafter: Animating open-domain images with video diffusion priors

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.604875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.231637Z digest=sha256:845c1d88c0e452fc1b2a901c0f30990a92dfcdd0d3ea988d3f357892f1384a23

Observation 1d8ae152-cffd-4991-a443-6e2d8181d4d4 · outbound

This paper cites Gmflow: Learning optical flow via global matching.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Gmflow: Learning optical flow via global matching

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.591772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.235517Z digest=sha256:0df1e089955d0af1c03b63d137fd5f0d8ddb3fc96dc6a6872dbb61dfbfcff53d

Observation 9910b07b-a008-48bb-9a9e-ae9c377efd05 · outbound

This paper cites MagicProp: Diffusion-based Video Editing via Motion-aware Appearance Propagation.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling MagicProp: Diffusion-based Video Editing via Motion-aware Appearance Propagation

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-08-10T21:41:23.352369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.239654Z digest=sha256:b3dc347e98db0a505187c06e43f74a5893829db9e0041000d60e22a6f073659f

Observation c0ed8248-855a-4756-8e54-53116e186248 · outbound

This paper cites Motion-Conditioned Image Animation for Video Editing.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Motion-Conditioned Image Animation for Video Editing

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.244051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.244051Z digest=sha256:72171d707fca120b0854f239c7c74f45af8d379b89306b2bea42def587d122ce

Observation b5c1d703-ac6c-4917-83d6-66ab981fce06 · outbound

This paper cites Paint by example: Exemplar-based image editing with diffusion models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Paint by example: Exemplar-based image editing with diffusion models

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.576761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.248392Z digest=sha256:c2a7e79fcf4a7f9a82d6d594c2cf4e74357b9842c9265b37556990bb8f8b4808

Observation 2c8b40b6-e398-46f4-a5b6-572285736344 · outbound

This paper cites Depth anything: Unleashing the power of large-scale unlabeled data.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Depth anything: Unleashing the power of large-scale unlabeled data

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.561239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.252222Z digest=sha256:55944c66e4cb900b111368230beed37a0534c34cfc3375fcbe0bad2af241085d

Observation d8a2c9ab-1301-4f81-9103-53f53e9856b8 · outbound

This paper cites Rerender a video: Zero-shot text-guided video-to-video translation.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Rerender a video: Zero-shot text-guided video-to-video translation

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.547123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.256073Z digest=sha256:e4696a989c9cc670e78a7091647080665912e1207edf3aabf3cd3b52a26cf603

Observation c8588efb-512a-4e5b-b889-c58d44f1794e · outbound

This paper cites IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.259850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.259850Z digest=sha256:fa5b6bad119fdef8bd77f6ab4ad9ed93d56f6e5babb627465f3a5d9be8896eb6

Observation 0171d1bb-3607-4c92-97a5-a5bcbec45c7a · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Adding conditional control to text-to-image diffusion models

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.534380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.264348Z digest=sha256:01f140f2ee828b909ea459b9c28363365800fe2c9fe4cd8e774c1dfdc3fa3015

Observation 9dd04a1e-5fb1-45db-877a-6230a2db62ba · outbound

This paper cites Sine: Single image editing with text-to-image diffusion models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Sine: Single image editing with text-to-image diffusion models

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.522591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.268609Z digest=sha256:2d161fe02038088e3b0684868f7822b69e4abbff6c0c1bd4392feed71edbad56

Observation 9c9fb6a3-1622-4fcc-9732-5645b813b3f3 · outbound

This paper cites Avid: Any-length video inpainting with dif- fusion model.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Avid: Any-length video inpainting with dif- fusion model

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.509453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.273112Z digest=sha256:4a13dbd0d9f7afdf0f512acf05d37f0afc5b397c6857c41c6d786d97596b7625

Observation 655872eb-92ff-4783-a0db-3fd0baa50417 · outbound

This paper cites Motiondirector: Motion customization of text-to-video diffusion models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Motiondirector: Motion customization of text-to-video diffusion models

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.494936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.277490Z digest=sha256:887de743495826da8359759af5d0b3df2a75abc0279b3bb5f9c4a5f54ed8394c

Observation b79f31ca-b2ec-4a36-9ceb-54da19575f76 · outbound

This paper cites clip-score: CLIP Score for Py- Torch.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling clip-score: CLIP Score for Py- Torch

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.480291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:41:23.282342Z digest=sha256:aa6e1a8bae41d0bad4b9cddf6664224963cafb8e0bf4cd81d698170578f9bee5

Pith citing papers

No inbound Pith citation observations are available.