Pith. sign in

Paper Citation Record · LEDGER

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework

As of 15 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 1 inbound Pith citation observation for arXiv:2605.23891.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.23891 v1

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-25T04:30:20.593882Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-12T01:18:51.054590Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

43 of 43 outbound references displayed

  • verified exact17
  • verified fuzzy21
  • unresolved1
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ebfb3cf8-87be-40c9-9ed6-53d734b1552b · outbound

This paper cites Denoising diffusion proba- bilistic models.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework Denoising diffusion proba- bilistic models

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T13:30:53.994822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:6befa4e0ce6f3ab0ec5bec2b82a7e1fdc4f02f574cef91113f8103ffece9ebaa

Observation 3a096e2a-b362-4065-8a79-a946419accdf · outbound

This paper cites Scalable diffusion models with trans- formers.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework Scalable diffusion models with trans- formers

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T13:30:54.008191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:a618ff300803b3ae68ac349491e027c03d0c4820f49c681ce67d456483f68cb7

Observation 6353b522-f826-44d6-8481-1076223b41ed · outbound

This paper cites Visual autoregressive modeling: Scalable image generation via next- scale prediction.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework Visual autoregressive modeling: Scalable image generation via next- scale prediction

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T13:30:53.991022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:5af8d200b25861a4f58ca2dcc0e61e454aac70295ef013b1e0ac89f380dbc8b9

Observation 34b6c747-2ccb-42b9-8ab0-120463c95e79 · outbound

This paper cites A survey on video diffusion models.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework A survey on video diffusion models

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T13:30:54.000049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:21e51d4de3d069771a06e474425c6a3ee7c044b2fa86070ee68a5c89a2df933f

Observation 4589c1bb-0874-400d-b90c-d9ba4d273269 · outbound

This paper cites Omnigen: Unified image generation.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework Omnigen: Unified image generation

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T13:30:53.968647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:e3684adcc4e1c9ce9e7e13b76dc5d7101a1a2922664f046370bec95eecf27bb0

Observation 4c29aae5-5edd-4cf9-a800-c91ed76145a5 · outbound

This paper cites Open-Sora: Democratizing Efficient Video Production for All.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework Open-Sora: Democratizing Efficient Video Production for All

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-25T04:35:21.963369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:e969f6f56b014f37d6d2e691da3f074326bb4a3e31ef7835a226b57cdc22c55f

Observation 6d2b7383-01f6-4543-83f2-a0db1828593d · outbound

This paper cites Show-o: One Single Transformer to Unify Multimodal Understanding and Generation.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework Show-o: One Single Transformer to Unify Multimodal Understanding and Generation

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-25T04:35:21.959481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:9d63724c68a42a60a55137fd4d8882cc08e1e397e30201a9a985c225129b55db

Observation a2a30f7e-6a7c-4176-b983-fc77492c48be · outbound

This paper cites LTX-Video: Realtime Video Latent Diffusion.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework LTX-Video: Realtime Video Latent Diffusion

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-05-25T04:35:21.967731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:07febf9c4dddc663ab20a0fdcbb85488d64caa3ffadb904f406f86563bd02ed3

Observation 481bb6d3-2f17-4dfe-bcb0-90ac4652ab16 · outbound

This paper cites Dragvideo: Interactive drag-style video editing.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework Dragvideo: Interactive drag-style video editing

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T13:30:53.975636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:a10a27ba702f4c7cf0f2d9738cb8a1f30679329f3aea3de39ba25bf79dc7064a

Observation 3b6d7119-9113-48bd-8e79-718246b289a9 · outbound

This paper cites Direct-a-video: Customized video generation with user-directed camera movement and object motion.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework Direct-a-video: Customized video generation with user-directed camera movement and object motion

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T13:30:53.983707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:4326c3c68d1c8ff2af73163cc8cb8b5ea6151398233c28a174de8d5002b96ccf

Observation 9cf6a045-40fb-44df-a1a4-2d45f45b89fe · outbound

This paper cites Keyframe-guided creative video inpainting.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework Keyframe-guided creative video inpainting

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T13:30:53.979507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:d858b7329656e8e435f1d2897e46ce3b6fb7e39dc9e55670814310d4e92cf1bc

Observation df4578d0-7c9d-42b5-9069-7264c9ce97f8 · outbound

This paper cites Shape-for- motion: Precise and consistent video editing with 3d proxy.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework Shape-for- motion: Precise and consistent video editing with 3d proxy

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T13:30:53.987522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:1a5381ebdbcbf0a7d312dbb918937b2326ad8a0a1ee87ac539cdc9e9268c2057

Observation a45dbec6-9dc8-4418-bb2a-6a4723436eb1 · outbound

This paper cites Lora-edit: Controllable first-frame-guided video editing via mask-aware lora fine-tuning.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework Lora-edit: Controllable first-frame-guided video editing via mask-aware lora fine-tuning

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:35:21.971794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:42f08e712ea461b21f8ee18ba5e15f39b794aa2927868bac1f2d1761cc0f0ef3

Observation 46b0d7b4-0fe4-46ac-8dba-fce1c6cfbbe8 · outbound

This paper cites Videodirector: Precise video editing via text-to-video mod- els.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework Videodirector: Precise video editing via text-to-video mod- els

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T13:30:54.003262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:6851d3ff82555f5fd4ae41e902ff36b987c2139fc7909f18acb7c156fb6a40fa

Observation 6ec71593-e655-4d1c-94a2-cfaee7bee513 · outbound

This paper cites UNIC: Unified In-Context Video Editing.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework UNIC: Unified In-Context Video Editing

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:35:21.938133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:50e16a5937ff86bf94572b81cc766c7aeb573954bc60c88e3fe3327e98adc65f

Observation 522a8afb-394e-40f1-97c6-1b8c2f317044 · outbound

This paper cites Videograin: Modulat- ing space-time attention for multi-grained video editing.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework Videograin: Modulat- ing space-time attention for multi-grained video editing

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T13:30:54.014754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:925f8509ecc2bf19549d60396acfba4bf935474bfe24d77c710128a99880fe8c

Observation a2010f51-575b-4096-bd28-474704015af2 · outbound

This paper cites Videopainter: Any-length video inpainting and editing with plug-and-play context control.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework Videopainter: Any-length video inpainting and editing with plug-and-play context control

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T13:30:54.019040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:6ae47a87f31e8d752b8024f5327075fba9ea39257372804e37d2b73d17e2398f

Observation 35c7567e-a805-4257-b3c3-cdb421d93cb4 · outbound

This paper cites Uniedit: A unified tuning-free framework for video motion and appearance editing.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework Uniedit: A unified tuning-free framework for video motion and appearance editing

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T13:30:54.011020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:fd77f5b3a0b13248ad5d8d55d9ad54f1069389452f9bfdb4494e142b396a55cc

Observation d2b5f200-5652-4fb5-9289-ae3a93b9c4fa · outbound

This paper cites Moonshot: Towards controllable video generation and edit- ing with motion-aware multimodal conditions.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework Moonshot: Towards controllable video generation and edit- ing with motion-aware multimodal conditions

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T13:30:54.026737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:85fed80e3e70af80cd182aa63293e9b892864465552afdead509ed77ab12663d

Observation aa5d7ff5-6946-414f-bd52-4b0e42061d27 · outbound

This paper cites Diffusion as shader: 3d-aware video diffusion for versatile video generation control.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework Diffusion as shader: 3d-aware video diffusion for versatile video generation control

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T13:30:53.949471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:853d5089db454645261c30e93873d29dbcd10348eaa06a61b04418499a1a5df4

Observation 8dca4270-44d9-48ce-a287-09f4fe114fea · outbound

This paper cites Get In Video: Add Anything You Want to the Video.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework Get In Video: Add Anything You Want to the Video

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-05-25T04:35:21.955361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:df7f4cfeccf2ea45d8f1eacb9bff7dea1e9b90405b246758fb1852b7e1633e2d

Observation 069d5c31-df89-4eef-8fc3-dbd0ef9da9b2 · outbound

This paper cites Videoany- door: High-fidelity video object insertion with precise motion control.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework Videoany- door: High-fidelity video object insertion with precise motion control

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T13:30:53.952997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:8cf0c8e925fc7bdd7f95b82f21f040b096ca7775135a1126cb03994cdac9d0ae

Observation 04cd073c-de04-4dc7-b6ba-bf0ac6c9c6cd · outbound

This paper cites Anything in Any Scene: Photorealistic Video Object Insertion.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework Anything in Any Scene: Photorealistic Video Object Insertion

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:35:21.923064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:c3cd30fec489c3a00b5fa7d400bc8afad17bf6e030c64ed2fb3c4b467084b504

Observation a10b0d4f-02aa-469f-b9d1-dd3f9bc0e688 · outbound

This paper cites DreamInsert: Zero-Shot Image-to-Video Object Insertion from A Single Image.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework DreamInsert: Zero-Shot Image-to-Video Object Insertion from A Single Image

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:35:21.927961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:f241ed2d9353bc2ae5a9ee119adff1d368e1a4b1e8db103374c4165806d816c6

Observation 9e369578-4073-4d0a-9d50-6f07e4e56a32 · outbound

This paper cites InsertAnywhere: Geometrically Grounded and Optics-Aware Video Object Insertion.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework InsertAnywhere: Geometrically Grounded and Optics-Aware Video Object Insertion

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-06-30T03:17:25.995003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:dafe0e992e2f6044c7cb320fc89060250ab36bd0eb86895150a22c0ae6477426

Observation 9c30a616-6684-4f9b-a5e1-b8e1d6b9e2f9 · outbound

This paper cites Omniinsert: Mask-free video insertion of any reference via diffusion transformer models.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework Omniinsert: Mask-free video insertion of any reference via diffusion transformer models

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:35:21.918603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:11e43ad5706913fd0bc6942b7de38eaf541cc71a971b99936e12f228df0b070d

Observation 66a067ea-f1a5-4b27-9f52-812908c993ee · outbound

This paper cites UniVideo: Unified Understanding, Generation, and Editing for Videos.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework UniVideo: Unified Understanding, Generation, and Editing for Videos

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-07T03:17:13.018030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:6cc94f453e6c6ad0ef4d3f3c38a2d7c614bf3b1a709f4665a41568a0a3e18b7a

Observation aa7cf522-517c-492c-88b5-e1e5e82c8900 · outbound

This paper cites Tele-omni: a unified multimodal framework for video generation and editing.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework Tele-omni: a unified multimodal framework for video generation and editing

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:35:21.898720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:edd6191e54a7668740e4a2c579b4b534dcfda120d21c907b7a6665578c1cda1b

Observation 2051b2f6-5f29-466a-aa50-7791dd0c974b · outbound

This paper cites Vace: All-in-one video creation and editing.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework Vace: All-in-one video creation and editing

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T13:30:54.022877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:2c3a7fbc918779e5b487e8354d3ee20c5e4df2d00fc7ba9d3f5d3830ec29f32f

Observation 7b0e3336-454e-454b-9bb0-99acb4d42399 · outbound

This paper cites FullDiT: Multi-Task Video Generative Foundation Model with Full Attention.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework FullDiT: Multi-Task Video Generative Foundation Model with Full Attention

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:35:21.904497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:a3f6f851604c54a850cd735c82f7f62463d360f6b4d76fe2466095165e5965f9

Observation d688b603-ab28-4550-84a4-67e3f911945a · outbound

This paper cites Seedance 1.0: Exploring the Boundaries of Video Generation Models.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework Seedance 1.0: Exploring the Boundaries of Video Generation Models

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-05-25T04:35:21.913747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:f11d328e895548d6625179b77510e6ffd79f44e7e504be4e8a797732ec51696e

Observation e3bcd26b-4641-4bc1-a583-3084a798fba1 · outbound

This paper cites an unresolved cited work.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-05-25T13:30:53.939675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:43e38fe6097a33d1cffe7f7c24eac0ac35bc74f292b6002123122105573a92e5

Observation cadac9a3-cb6d-4ad6-bf7b-8cd27e5ed40f · outbound

This paper cites HunyuanVideo 1.5 Technical Report.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework HunyuanVideo 1.5 Technical Report

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-05-25T04:35:21.950304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:6e623265f58f8c3e38c2a4ec5def6e6cdae0b99e95037f5f2f7680753f5fab68

Observation 01f50e50-7474-467c-b172-490abda28146 · outbound

This paper cites CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

Reference 34

Resolution
verified exact
local_arxiv, observed 2026-05-25T04:35:21.888311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:6bc02e7f15b8c979eccabcef551456dc5e6239b000d7ee30c0feaf68d5064c15

Observation fd161cf9-f099-4143-9fee-3d6793b67fa6 · outbound

This paper cites OpenVE-3M: A Large-Scale High-Quality Dataset for Instruction-Guided Video Editing.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework OpenVE-3M: A Large-Scale High-Quality Dataset for Instruction-Guided Video Editing

Reference 35

Resolution
metadata mismatch
arxiv_id, observed 2026-08-13T02:21:48.005953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:d10d5e0e7f1a84e39b6ec0fd9f692f76462c52d8068dcfdb4bdf356b07004d6d

Observation f14a4594-bc96-46a0-9f4e-d9c46f181359 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework Gemini: A Family of Highly Capable Multimodal Models

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-05-25T04:35:21.883387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:93fe15f0d35af940c97712d4914d5f26f030f680d414ee56d840e13958d63e5b

Observation 7728fa12-dbec-4cb0-a2a9-a376d6901121 · outbound

This paper cites Langsam: Language-guided segment anything.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework Langsam: Language-guided segment anything

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T13:30:53.946155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:cb26295f15935151621c557d6a596d7b41a3eedd1770dbeb2cac15f4863c2908

Observation 38ce57cb-28e8-48ca-888d-f22857dbb240 · outbound

This paper cites MiniMax-Remover: Taming Bad Noise Helps Video Object Removal.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework MiniMax-Remover: Taming Bad Noise Helps Video Object Removal

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-05-25T04:35:21.946875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:77fae1ca4d23c26bc1f8f04f60a9955da3ca0779abaaa54a03544226c19f920e

Observation 6bb9afb9-5e13-45a1-8588-521c726e2234 · outbound

This paper cites Omnitransfer: All-in-one framework for spatio-temporal video transfer.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework Omnitransfer: All-in-one framework for spatio-temporal video transfer

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:35:21.942269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:23f7659bf0f4a7f876292997b167d3959252c76bfa23272b1d0b834be40b9cf3

Observation ab1998ea-4134-450d-aa1b-684e7f97ec7c · outbound

This paper cites Vision transformer with quadrangle attention.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework Vision transformer with quadrangle attention

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T13:30:53.943011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:dd8980748f1a2f3364cddc2fbe024d4d1a4a943b1547eeac4e8902078f2c41ee

Observation 4a44defa-d5ff-4516-b34e-13e86abb3e7a · outbound

This paper cites Chatgpt.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework Chatgpt

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T13:30:53.934099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:e1b9aec403e6dee8b2baf5af4bba306b6ae41fccf69815b893f7af112206e248

Observation c5a20eea-675b-49ba-a0ba-9fc3af1d71d0 · outbound

This paper cites Pytorch fsdp: Experiences on scaling fully sharded data parallel.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework Pytorch fsdp: Experiences on scaling fully sharded data parallel

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T13:30:53.937086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:73ccebeb7b7dcb6f60d1d534ec54a69df7f183284aefbc5d19c1d6e50a1e139e

Observation 214b3751-effb-4911-9c9c-79ed698b6e72 · outbound

This paper cites Qwen3-VL Technical Report.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework Qwen3-VL Technical Report

Reference 43

Resolution
malformed identifier
local_arxiv, observed 2026-05-25T04:35:21.879093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:14ee3b3aa007baed7348b048f2e1dbb08919fbc92bba3df9eebe0291671e9c7a

Pith citing papers

Observation 693454b5-0237-478f-836c-a642a0cfcafb · inbound

Token-Based Affordance Grounding with Large Vision-Language Models cites this paper.

Token-Based Affordance Grounding with Large Vision-Language Models Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-12T01:18:51.054590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:18:51.054590Z digest=sha256:320a442bca539eb00e7bbeec8306d3651810a82c38f06bb4e85e9eca04e0a7d9