Pith. sign in

Paper Citation Record · LEDGER

OpenVE-3M: A Large-Scale High-Quality Dataset for Instruction-Guided Video Editing

As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2512.07826.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2512.07826 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 19 of 19 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-01T06:35:16.951243Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5b4764b9-d2bf-4e54-bbbd-843f1d49c663 · inbound

ImVideoEdit: Image-learning Video Editing via 2D Spatial Difference Attention Blocks cites this paper.

ImVideoEdit: Image-learning Video Editing via 2D Spatial Difference Attention Blocks OpenVE-3M: A Large-Scale High-Quality Dataset for Instruction-Guided Video Editing

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-08-13T02:21:48.005953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T17:49:33.822264Z digest=sha256:f6787ee9e2edfb731f0e48476110c5bc7b73c597ccd026608d33698becf32bea

Observation fad165c7-2943-465b-bc3f-e31ed0a41bf6 · inbound

InsEdit: Towards Instruction-based Visual Editing via Data-Efficient Video Diffusion Models Adaptation cites this paper.

InsEdit: Towards Instruction-based Visual Editing via Data-Efficient Video Diffusion Models Adaptation OpenVE-3M: A Large-Scale High-Quality Dataset for Instruction-Guided Video Editing

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-08-13T02:21:48.005953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T17:39:59.791758Z digest=sha256:8e1d6a4ead53b9116d4aa6b2439315a881b3abf90ad34abbf3ab77860bd092f2

Observation 2d61b83f-1478-4500-afd6-b66de4be8234 · inbound

VEFX-Bench: A Holistic Benchmark for Generic Video Editing and Visual Effects cites this paper.

VEFX-Bench: A Holistic Benchmark for Generic Video Editing and Visual Effects OpenVE-3M: A Large-Scale High-Quality Dataset for Instruction-Guided Video Editing

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-08-13T02:21:48.005953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T08:24:41.925800Z digest=sha256:4c1e29636ab041d4273c4af00794b2f0e1b40be2775b186e40472a3bc94d1123

Observation 163539ef-aa74-4de0-9825-9fedb467eb25 · inbound

LIVE: Leveraging Image Manipulation Priors for Instruction-based Video Editing cites this paper.

LIVE: Leveraging Image Manipulation Priors for Instruction-based Video Editing OpenVE-3M: A Large-Scale High-Quality Dataset for Instruction-Guided Video Editing

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-08-13T02:21:48.005953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T07:21:41.483427Z digest=sha256:bf84731e0a892d59616eb249e3b45c47710b7b87a025bbd779e73f33c49e709f

Observation 09227181-e68c-485b-b41f-72da79208551 · inbound

How Far Are Video Models from True Multimodal Reasoning? cites this paper.

How Far Are Video Models from True Multimodal Reasoning? OpenVE-3M: A Large-Scale High-Quality Dataset for Instruction-Guided Video Editing

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-08-13T02:21:48.005953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T02:44:52.920816Z digest=sha256:d3f12ea5a1024d7810a5a73338230d80497cab55721286018201148d8ad66127

Observation 90bc9793-19af-40b6-a9a6-bc09157d5a61 · inbound

Mamoda2.5: Enhancing Unified Multimodal Model with DiT-MoE cites this paper.

Mamoda2.5: Enhancing Unified Multimodal Model with DiT-MoE OpenVE-3M: A Large-Scale High-Quality Dataset for Instruction-Guided Video Editing

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-08-13T02:21:48.005953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-08T18:26:58.696936Z digest=sha256:64d7b806c2fc572c74074981ff322d24994710023ffb684de7256dc30c3f9884

Observation 89252370-c0c1-47fe-9454-4cd4f90611b4 · inbound

Sparkle: Realizing Lively Instruction-Guided Video Background Replacement via Decoupled Guidance cites this paper.

Sparkle: Realizing Lively Instruction-Guided Video Background Replacement via Decoupled Guidance OpenVE-3M: A Large-Scale High-Quality Dataset for Instruction-Guided Video Editing

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-08-13T02:21:48.005953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-08T12:34:34.315135Z digest=sha256:59196f3db62d38253dabcb342dd25c4efe93d2ac6bb9ef0a64a69dc2954de682

Observation 31a4e36b-58de-41fd-ab58-9bddadf5c2e5 · inbound

MiVE: Multiscale Vision-language features for reference-guided video Editing cites this paper.

MiVE: Multiscale Vision-language features for reference-guided video Editing OpenVE-3M: A Large-Scale High-Quality Dataset for Instruction-Guided Video Editing

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-08-13T02:21:48.005953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-05-15T05:23:20.179387Z digest=sha256:93286ff1ffc3d4f708b3e571c6ac2b3f90a2ff82c9c5a1981da6ba580a21cd3c

Observation 3fabf862-66be-4a5e-95ba-f7e18d49f3d4 · inbound

Aurora: Unified Video Editing with a Tool-Using Agent cites this paper.

Aurora: Unified Video Editing with a Tool-Using Agent OpenVE-3M: A Large-Scale High-Quality Dataset for Instruction-Guided Video Editing

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-08-13T02:21:48.005953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-20T10:47:05.038308Z digest=sha256:a267267d6e15b767684a2cd7d781002a2eedeb85c45680e2c1d5dd0d4d1662ab

Observation 3454d3fc-a24e-40b7-a483-c19beff2916e · inbound

Bernini: Latent Semantic Planning for Video Diffusion cites this paper.

Bernini: Latent Semantic Planning for Video Diffusion OpenVE-3M: A Large-Scale High-Quality Dataset for Instruction-Guided Video Editing

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-08-13T02:21:48.005953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-22T06:39:47.124605Z digest=sha256:e667bafb66afca1baf2e171415fb25038b98f1b3179fad5cc28767678f39ac1b

Observation ef47b832-69ec-45ab-a414-aafe309df72e · inbound

Occlusion-Aware Physics-Semantic Keyframe Selection for Robust Video Editing cites this paper.

Occlusion-Aware Physics-Semantic Keyframe Selection for Robust Video Editing OpenVE-3M: A Large-Scale High-Quality Dataset for Instruction-Guided Video Editing

Reference 113

Resolution
verified exact
arxiv_id, observed 2026-08-13T02:21:48.005953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-05-25T05:00:31.868653Z digest=sha256:1dabc5d3da49cba9d0862885ffb96a85039e7eae6a190a6a115c6841c9acbb6c

Observation fd161cf9-f099-4143-9fee-3d6793b67fa6 · inbound

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework cites this paper.

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework OpenVE-3M: A Large-Scale High-Quality Dataset for Instruction-Guided Video Editing

Reference 35

Resolution
metadata mismatch
arxiv_id, observed 2026-08-13T02:21:48.005953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T04:30:20.593882Z digest=sha256:291b0d1d791e3656651493a078f13d37344c41b7fcc524e327eb296a31b58c82

Observation e9949324-eddf-4d1a-b73b-e8cff83c90a0 · inbound

Reasoning to Align: Implicit Reasoning in Diffusion Transformers for Video Editing cites this paper.

Reasoning to Align: Implicit Reasoning in Diffusion Transformers for Video Editing OpenVE-3M: A Large-Scale High-Quality Dataset for Instruction-Guided Video Editing

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-08-13T02:21:48.005953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-30T13:37:09.892339Z digest=sha256:6e36f296ba1799542acf942ec32dfaab43737d859946f3f0f8ef5517c18bbde0

Observation 5261f483-5f19-4697-b315-9cd606b3758a · inbound

SpongeBob: Sync-Aware Harmonious Audio-Visual Generative Editing cites this paper.

SpongeBob: Sync-Aware Harmonious Audio-Visual Generative Editing OpenVE-3M: A Large-Scale High-Quality Dataset for Instruction-Guided Video Editing

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-08-13T02:21:48.005953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-30T12:07:44.914337Z digest=sha256:8b7e9d4630e32d6bb8fac90eeea0b919653dab144a2387ac2fd91d0ec0ede95b

Observation e919bf8e-c76f-44b7-bd97-a8db6fb94c7e · inbound

DualEraser: Joint Video Object and Effect Removal via Balanced Text-Mask Guidance and Decoupled Locator-Preserver cites this paper.

DualEraser: Joint Video Object and Effect Removal via Balanced Text-Mask Guidance and Decoupled Locator-Preserver OpenVE-3M: A Large-Scale High-Quality Dataset for Instruction-Guided Video Editing

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-08-13T02:21:48.005953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-29T08:06:27.191812Z digest=sha256:7f58104a92562b2c6302d3b561597d5dd1faea6a6f057e702651af160f3a5305

Observation 7e285cf4-547e-4eaa-815b-bb18befc6161 · inbound

SANA-Streaming: Real-time Streaming Video Editing with Hybrid Diffusion Transformer cites this paper.

SANA-Streaming: Real-time Streaming Video Editing with Hybrid Diffusion Transformer OpenVE-3M: A Large-Scale High-Quality Dataset for Instruction-Guided Video Editing

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-08-13T02:21:48.005953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-29T07:39:41.553623Z digest=sha256:dfe9b964eaae3304feca2ce61c0821e1af26fd73d64982a6093c97776b9336a8

Observation 18a0ee24-8692-4f78-8519-c5aa8200a789 · inbound

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation cites this paper.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation OpenVE-3M: A Large-Scale High-Quality Dataset for Instruction-Guided Video Editing

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-08-13T02:21:48.005953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:b688b3576375705881e6e6489db011471ab7a93c88c5e705f2b266bc31d8cc76

Observation 257a4673-cc77-47ee-9851-28518962325e · inbound

Goku: A Million-Scale Universal Dataset and Benchmark for Instruction-Based Video Editing cites this paper.

Goku: A Million-Scale Universal Dataset and Benchmark for Instruction-Based Video Editing OpenVE-3M: A Large-Scale High-Quality Dataset for Instruction-Guided Video Editing

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-08-13T02:21:48.005953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-30T06:00:24.487104Z digest=sha256:cf4df03def978ff7e3b84593807afa6691b6630ed8298c7b11a9749fcf43c259

Observation 6fb91da7-c493-407c-8b8f-a2eac79b91da · inbound

Goku: A Million-Scale Universal Dataset and Benchmark for Instruction-Based Video Editing cites this paper.

Goku: A Million-Scale Universal Dataset and Benchmark for Instruction-Based Video Editing OpenVE-3M: A Large-Scale High-Quality Dataset for Instruction-Guided Video Editing

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-08-13T02:21:48.005953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-01T06:35:16.951243Z digest=sha256:6541e297b6b1678b89c7e3597f5b3431534986039f8ef232b2ed50b709901eae