Pith. sign in

Paper Citation Record · LEDGER

Learning Stochastic Bridges for Video Object Removal via Video-to-Video Translation

As of 11 August 2026, this Paper Citation Record lists 25 of 25 outbound references and 1 inbound Pith citation observation for arXiv:2601.12066.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2601.12066 v4

Coverage vector

measured 25 of 25 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-03T10:02:43.468347Z

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-29T08:06:27.191812Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-06-29T08:13:15.767173Z

Reference resolution

25 of 25 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved25
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a88afd3c-6177-48e4-b014-ff3f90b51277 · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

Learning Stochastic Bridges for Video Object Removal via Video-to-Video Translation HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T10:02:41.439147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:02:41.439147Z digest=sha256:342f2d8ed01be7e19dbbdb7873ecdccf095db70fab0e4bcd0f6ae3627918c437

Observation 0b52cd02-cc22-4f9b-a9ed-bc3518c2ab40 · outbound

This paper cites In practical implementation, time-dependent coefficients (e.g., ¯αtor at, bt, ct) are typically precomputed and retrieved from a schedule table for computational efficiency.

Learning Stochastic Bridges for Video Object Removal via Video-to-Video Translation In practical implementation, time-dependent coefficients (e.g., ¯αtor at, bt, ct) are typically precomputed and retrieved from a schedule table for computational efficiency

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T10:02:42.145678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:02:42.145678Z digest=sha256:90b505ac9a30754fd6945527bea5379152217cdf9b3f8bfeb5c709e5a502604c

Observation c406866d-1284-4a7b-a86b-9645b0d3d4cc · outbound

This paper cites Flow Matching for Generative Modeling.

Learning Stochastic Bridges for Video Object Removal via Video-to-Video Translation Flow Matching for Generative Modeling

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-03T10:02:41.603223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:02:41.603223Z digest=sha256:2c66000e2882a7da16c3f7df8d06ab7d8d0cdd1f5f6462262b7c1402b699d6af

Observation 9928931d-e477-4e1c-b8f7-91fecd1c2415 · outbound

This paper cites Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow.

Learning Stochastic Bridges for Video Object Removal via Video-to-Video Translation Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T10:02:41.684396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:02:41.684396Z digest=sha256:55f142914f7a5a9824006218860a1523cf9d80b58f4c2a2eb41f6df2aa28fe39

Observation 558e4ea7-882e-4bd0-a77e-b0e3eef68885 · outbound

This paper cites Non-Denoising Forward-Time Diffusions.

Learning Stochastic Bridges for Video Object Removal via Video-to-Video Translation Non-Denoising Forward-Time Diffusions

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T10:02:41.766710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:02:41.766710Z digest=sha256:9d4f751f0b71ab0ebd040ebba904a983855dfb36bd383f7f554014a508088ef2

Observation fe81f6bd-7389-4280-8ea1-36a7c090f41a · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

Learning Stochastic Bridges for Video Object Removal via Video-to-Video Translation Wan: Open and Advanced Large-Scale Video Generative Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T10:02:41.907536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:02:41.907536Z digest=sha256:890a5bbef7c4066d423f870e910348182dd3f067c8a1382c1574fde91124df8f

Observation d96ea9c5-c2d3-4698-9df0-993545473e18 · outbound

This paper cites Qwen3 Technical Report.

Learning Stochastic Bridges for Video Object Removal via Video-to-Video Translation Qwen3 Technical Report

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-03T10:02:41.984057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:02:41.984057Z digest=sha256:fe8009e49d45a4cbac60b3ecb0ca2d933e765b0bcc769e6826459316407c4f7f

Observation d2be03b6-092d-4af4-8837-256f21d96f04 · outbound

This paper cites MiniMax-Remover: Taming Bad Noise Helps Video Object Removal.

Learning Stochastic Bridges for Video Object Removal via Video-to-Video Translation MiniMax-Remover: Taming Bad Noise Helps Video Object Removal

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-03T10:02:42.036296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:02:42.036296Z digest=sha256:0b66f058ffac784b6531496c038e47ab87078e3252070ae1fbc96a2c53e8c510

Observation 856165b0-3c75-4333-8c40-d7c2b9d22af6 · outbound

This paper cites Pseudo Code for the Training and inference of BridgeRemoval.

Learning Stochastic Bridges for Video Object Removal via Video-to-Video Translation Pseudo Code for the Training and inference of BridgeRemoval

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-03T10:02:42.118215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:02:42.118215Z digest=sha256:ce1e7df8e21a225e27660d8fe54fe069757a03d1661732d370957f2d34b76494

Observation 1de171d0-3b7f-438e-94e4-4cdbc6042acf · outbound

This paper cites (18) Assuming the underlying process is a VP-SDE with Gaussian transition kernels: p(zt|z0) =N(z t;α tz0, σ2 t I), p(zT |zt) =N zT ; αT αt zt, σ2 T − α2 T α2 t σ2 t I.

Learning Stochastic Bridges for Video Object Removal via Video-to-Video Translation (18) Assuming the underlying process is a VP-SDE with Gaussian transition kernels: p(zt|z0) =N(z t;α tz0, σ2 t I), p(zT |zt) =N zT ; αT αt zt, σ2 T − α2 T α2 t σ2 t I

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-03T10:02:42.218330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:02:42.218330Z digest=sha256:5893be731b48c3320fce0f9d6615e7e622975b2c2acaa71805b30ad93b8c8398

Observation 3f8ea8f3-3617-4020-9325-411c2a2bc34a · outbound

This paper cites Thus, Eq.

Learning Stochastic Bridges for Video Object Removal via Video-to-Video Translation Thus, Eq

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-03T10:02:42.330347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:02:42.330347Z digest=sha256:4e19d9be2ef115547ec1b680cdc1298253c39b63653e7c34ce54c58da2fc1e2a

Observation 8cef6074-eaf1-45fc-846a-eda15f097f04 · outbound

This paper cites As the number of inference steps increases, CLIP-T gradually improves while PSNR steadily decreases.

Learning Stochastic Bridges for Video Object Removal via Video-to-Video Translation As the number of inference steps increases, CLIP-T gradually improves while PSNR steadily decreases

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T10:02:42.490318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:02:42.490318Z digest=sha256:e1916ac1e52b1267681faa0af5476143758c209cdb711e401fa217769d56ad38

Observation 85de7f2f-c165-4145-a9c4-4c3764b79459 · outbound

This paper cites an unresolved cited work.

Learning Stochastic Bridges for Video Object Removal via Video-to-Video Translation Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T10:02:42.610029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:02:42.610029Z digest=sha256:9e0b8dcce691ab68b4cf27c3031e25b6198e8c4f2dcf61f8e6972a64811ce4fb

Observation 737b5ba0-5e56-41cb-8264-5d0ab0d76e9d · outbound

This paper cites an unresolved cited work.

Learning Stochastic Bridges for Video Object Removal via Video-to-Video Translation Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-03T10:02:42.723042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:02:42.723042Z digest=sha256:a526f843c60ce7f45c59b3b6f7d648c31bc69b1b896a2e49f134fe64129404df

Observation 7f6182ff-8c4f-43e8-897d-9eb4fc683f4a · outbound

This paper cites an unresolved cited work.

Learning Stochastic Bridges for Video Object Removal via Video-to-Video Translation Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T10:02:42.845093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:02:42.845093Z digest=sha256:c4ce8eb139e36693f6be52b16c37078ef0de4724c0569b5cd5346326ae870090

Observation d22f5ec1-514d-495b-91c5-a709aa26c48d · outbound

This paper cites an unresolved cited work.

Learning Stochastic Bridges for Video Object Removal via Video-to-Video Translation Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T10:02:42.932641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:02:42.932641Z digest=sha256:bfe0c33119c3184b86248bada71c62047d239f4a4b127cd64bf220fa2e955fda

Observation 647848c0-fd8e-46e4-85f5-64ebad28a9a0 · outbound

This paper cites an unresolved cited work.

Learning Stochastic Bridges for Video Object Removal via Video-to-Video Translation Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-03T10:02:43.092942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:02:43.092942Z digest=sha256:8991f8720ecdd638ee8b8ef2ddd2e030e24a50c1a58f6da2b847f40c09a4e5b3

Observation f4d44413-8278-4d68-bb1f-3957edaf2c96 · outbound

This paper cites an unresolved cited work.

Learning Stochastic Bridges for Video Object Removal via Video-to-Video Translation Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-03T10:02:43.218821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:02:43.218821Z digest=sha256:07c17186c426b9499dd849067e2afb55d4f5554d8008694aed571cbffb0ecc18

Observation 3e202367-665a-4763-bbdb-85c20fd67ced · outbound

This paper cites This video is from the DA VIS dataset.

Learning Stochastic Bridges for Video Object Removal via Video-to-Video Translation This video is from the DA VIS dataset

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-03T10:02:43.306127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:02:43.306127Z digest=sha256:3b49940e9b6763cf3ff576e15afd5840fbe99758079148462d285f7a031f3e51

Observation d4ee4636-ac59-4ae9-8584-bf7a305c3e9e · outbound

This paper cites As shown in Tables 5 and 6, V ACE achieves relatively high scores on Temporal Flickering and Imaging Quality, yet this does not indicate effective object removal.

Learning Stochastic Bridges for Video Object Removal via Video-to-Video Translation As shown in Tables 5 and 6, V ACE achieves relatively high scores on Temporal Flickering and Imaging Quality, yet this does not indicate effective object removal

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-03T10:02:43.468347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:02:43.468347Z digest=sha256:af91e7390f6ec80d70c778d0f491f8c733e3a7265567a1600939f9a3af6c59b0

Observation 9f480668-ca63-477a-a49b-fdfc34ce8ffe · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

Learning Stochastic Bridges for Video Object Removal via Video-to-Video Translation SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-03T10:02:41.851372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:02:41.851372Z digest=sha256:22ef8a98a42f1233be57f29e906eb60295514f1715a92ca45385caa2afe175a7

Observation 5015c5b1-94ec-497f-800c-a158de67736b · outbound

This paper cites Schrodinger Bridges Beat Diffusion Models on Text-to-Speech Synthesis.

Learning Stochastic Bridges for Video Object Removal via Video-to-Video Translation Schrodinger Bridges Beat Diffusion Models on Text-to-Speech Synthesis

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-03T10:02:41.325605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:02:41.325605Z digest=sha256:2b1d3e747f0f2b5f9926407c1210f6824485eb8175604d219917b6ac7900b05f

Observation 4ebd48b7-d5d3-4ae8-b49e-c774297cc2e5 · outbound

This paper cites DiffuEraser: A Diffusion Model for Video Inpainting.

Learning Stochastic Bridges for Video Object Removal via Video-to-Video Translation DiffuEraser: A Diffusion Model for Video Inpainting

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-03T10:02:41.522899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:02:41.522899Z digest=sha256:4ba0db99c0f026565728a1f6fc27bfe2cb1cd9bdc4009f037404e5df9d5e7f91

Observation 9d0a7ea3-bd81-4f5b-8824-f73a841fcbfe · outbound

This paper cites LoRA-Edit: Controllable first-frame-guided video editing via mask-aware lora fine-tuning.arXiv preprint arXiv:2506.10082,.

Learning Stochastic Bridges for Video Object Removal via Video-to-Video Translation LoRA-Edit: Controllable first-frame-guided video editing via mask-aware lora fine-tuning.arXiv preprint arXiv:2506.10082,

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-03T10:02:41.388025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:02:41.388025Z digest=sha256:5a21f5e71c39c8ef68bd7e3e6bf41009419c260b8bcbe4afc3395bb863c5b760

Observation dc4b1de2-69cd-48cf-812f-4c5ebff73e78 · outbound

This paper cites L., Ma, S., Zeng, Y ., Liu, Z., Xu, Y ., Shen, Y ., and Chen, Q.

Learning Stochastic Bridges for Video Object Removal via Video-to-Video Translation L., Ma, S., Zeng, Y ., Liu, Z., Xu, Y ., Shen, Y ., and Chen, Q

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-03T10:02:41.265049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:02:41.265049Z digest=sha256:c584f745815ec17783e2bc963b4854438c1bf1c9430a543d9e2639cc979cc85d

Pith citing papers

Observation f64a1b89-4da8-42f6-a0ea-bdfd74fdcf4a · inbound

GenEraser: Generalizable Video Object Removal via Balanced Text-Mask Guidance and Decoupled Locator-Preserver cites this paper.

GenEraser: Generalizable Video Object Removal via Balanced Text-Mask Guidance and Decoupled Locator-Preserver Learning Stochastic Bridges for Video Object Removal via Video-to-Video Translation

Reference 16

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T08:13:15.768386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T08:06:27.191812Z digest=sha256:c4768e87218374d876b20707e9920a9c68fdedd6a8ad587a8bbad51edcf7855f