Pith. sign in

Paper Citation Record · LEDGER

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation

As of 23 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 0 inbound Pith citation observations for arXiv:2606.03168.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.03168 v1

Coverage vector

measured 39 of 39 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-28T11:04:49.366127Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

39 of 39 outbound references displayed

  • verified exact27
  • verified fuzzy0
  • unresolved8
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7eeccabf-dd94-47ca-8a6f-ab26a3f47d90 · outbound

This paper cites Insvie-1m: Effective instruction-based video editing with elaborate dataset construction.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation Insvie-1m: Effective instruction-based video editing with elaborate dataset construction

Reference 1

Resolution
unresolved
no resolver link, observed 2026-06-28T11:04:49.366127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:c6712afc245b878c5432ff128422603e959be6f839be18a6c9f9828e3464f958

Observation 18a0ee24-8692-4f78-8519-c5aa8200a789 · outbound

This paper cites OpenVE-3M: A Large-Scale High-Quality Dataset for Instruction-Guided Video Editing.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation OpenVE-3M: A Large-Scale High-Quality Dataset for Instruction-Guided Video Editing

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-08-13T02:21:48.005953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:2ad57b37fb412bdefc98aac89646b4424d9a5ba8b3457234fb7fadf39e7115a3

Observation f1e7ddff-77f5-44a3-9b6f-c9a8c5c9d552 · outbound

This paper cites arXiv preprint arXiv:2510.15742 , year=.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation arXiv preprint arXiv:2510.15742 , year=

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T02:16:26.745870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:27a6d81b715eb061e8ee46f60c2f8ca1b59c04989ddc3eefec0d13a7af13c7f6

Observation 5f227e3c-16a9-46de-8416-4b5504f19ecd · outbound

This paper cites Zero-Shot Audio-Visual Editing via Cross-Modal Delta Denoising.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation Zero-Shot Audio-Visual Editing via Cross-Modal Delta Denoising

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-02T02:16:26.764211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:89cd60e2efd62ec12073cabca0a29906d86152fea425bb50e85a49f42143bc61

Observation 512d2ac6-b1bb-4edd-839e-0009085e3bd9 · outbound

This paper cites Av-edit: Multimodal generative sound effect editing via audio-visual semantic joint control.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation Av-edit: Multimodal generative sound effect editing via audio-visual semantic joint control

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-28T11:04:49.366127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:94a23ba1e2c9209cb03d2623c377cd4fb6fcc8333586ffdb2e2dfa29a224daba

Observation a2f8037c-11a4-4a69-a147-76ccc93413ca · outbound

This paper cites AVI-Edit: Audio-sync Video Instance Editing with Granularity-Aware Mask Refiner.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation AVI-Edit: Audio-sync Video Instance Editing with Granularity-Aware Mask Refiner

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-07-02T02:16:26.759100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:86c9377dac87fd141d33c7bb12697190e4b82804b348e0e3a63ceaa19715fcd1

Observation dbcecc2b-82a8-4dd0-b96a-955cd28af8cb · outbound

This paper cites Openhumanvid: A large-scale high-quality dataset for enhancing human-centric video generation.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation Openhumanvid: A large-scale high-quality dataset for enhancing human-centric video generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-06-28T11:04:49.366127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:30a3c49699732fee4884fc8bcd5ace31bb1dbd8d2d0a1e932c5d733c1f73a0d6

Observation f9d4710d-4d57-43be-8766-893118d790e6 · outbound

This paper cites VidGen-1M: A Large-Scale Dataset for Text-to-video Generation.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation VidGen-1M: A Large-Scale Dataset for Text-to-video Generation

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-02T02:16:26.776236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:432e04653ee0cf0e62bc985ccb8d4d3d7cdb5fc6f697bd5aeeb49c34ec77ae5b

Observation 47715f08-aabd-4995-a65b-740a9166c7d4 · outbound

This paper cites VGGSound: A Large-scale Audio-Visual Dataset.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation VGGSound: A Large-scale Audio-Visual Dataset

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-02T02:16:26.809374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:746641165b560889c64be322bd0a4f78550ff1446cae88b19fdfa6c33a786855

Observation 12d17b61-ec44-445a-9033-1d22877cfc74 · outbound

This paper cites Koala-36m: A large-scale video dataset improving consistency between fine-grained conditions and video content.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation Koala-36m: A large-scale video dataset improving consistency between fine-grained conditions and video content

Reference 11

Resolution
unresolved
no resolver link, observed 2026-06-28T11:04:49.366127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:1a974cfc894c65ce832d7d91c943240e570a061f36d4ec0a7be0ad65b3916f51

Observation 733d0f7c-5215-441b-bd55-a469b9bdf976 · outbound

This paper cites Qwen3-Omni Technical Report.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation Qwen3-Omni Technical Report

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-07-02T02:16:26.750991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:89b51c2340269873f68c947aab80b8a80609392af1d06caae46b190e3eef5f14

Observation 5f42ee24-a374-4854-a97f-1a3ca3ed2eab · outbound

This paper cites Mel-RoFormer for Vocal Separation and Vocal Melody Transcription.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation Mel-RoFormer for Vocal Separation and Vocal Melody Transcription

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-02T02:16:26.773429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:fef22f332680035cbe82b92d0c841cc192e26961d17427f7aa0f76b43cd3d3db

Observation 3a3124cf-28ce-4e0f-9763-d40762330633 · outbound

This paper cites ZeroSep: Separate Anything in Audio with Zero Training.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation ZeroSep: Separate Anything in Audio with Zero Training

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-02T02:16:26.754001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:644388ec22e8c776e9950c128535601a37108553e1082d9289b2899fd5e944f3

Observation 3b7734ac-fe9d-442d-a325-c37dfa514464 · outbound

This paper cites Ruijie Tao, Zexu Pan, Rohan Kumar Das, Xinyuan Qian, Mike Zheng Shou, and Haizhou Li.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation Ruijie Tao, Zexu Pan, Rohan Kumar Das, Xinyuan Qian, Mike Zheng Shou, and Haizhou Li

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T02:16:26.732613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:d58672e66cf14cfca749d02a9d47aef03a0889f8453c1ab25940da1f4dd61ddf

Observation 935d5965-417a-4a9e-b7d6-f02832cef876 · outbound

This paper cites Qwen3 Technical Report.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation Qwen3 Technical Report

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-07-02T02:16:26.761375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:b88c28e61cfca592162220724be054d4b9c5e2e80dacc477f71992227a7ac9e8

Observation 80f263da-5825-4738-b7af-39bff0ff780e · outbound

This paper cites HunyuanImage 3.0 Technical Report.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation HunyuanImage 3.0 Technical Report

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-07-02T02:16:26.756423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:d0e0863a8bbb5a48a96bbb6eab3be55bb1d6145f9931b965250bc266bf4ed1dd

Observation 9ac78d33-4697-4e05-8e13-51fc775aab86 · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation Wan: Open and Advanced Large-Scale Video Generative Models

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-07-02T02:16:26.770215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:6395e775ae8ae85284de020ba47c9978c50bdec10b9b33cc113cbf3411e7b7e3

Observation d655ff30-74f8-4484-96db-181507b2e4d5 · outbound

This paper cites DreamVoice: Text-Guided Voice Conversion.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation DreamVoice: Text-Guided Voice Conversion

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-02T02:16:26.797282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:118376a79dec221835dbf658d4df8cb63ab822cc0605f853b35b08c11d53490b

Observation 965398af-e867-4dfa-9a4f-5e3aaec6a554 · outbound

This paper cites Ffp-300k: Scaling first-frame propagation for generalizable video editing.arXiv preprint arXiv:2601.01720, 2026.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation Ffp-300k: Scaling first-frame propagation for generalizable video editing.arXiv preprint arXiv:2601.01720, 2026

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-07-02T02:16:26.747725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:d02dea9dfe7b69fb78ca2346947fd958d168e70b0408442b083be68e4be91466

Observation aa346e1f-6b7a-4698-84f3-0395901c5310 · outbound

This paper cites HunyuanVideo-Foley: Multimodal Diffusion with Representation Alignment for High-Fidelity Foley Audio Generation.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation HunyuanVideo-Foley: Multimodal Diffusion with Representation Alignment for High-Fidelity Foley Audio Generation

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-02T02:16:26.744449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:a4356d89f07b2b06d33a1af8f2f5c51e81c5df67444bf28433477956bae7af7b

Observation 806f97f2-d01d-4411-a459-559b7a892109 · outbound

This paper cites MiniMax-Remover: Taming Bad Noise Helps Video Object Removal.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation MiniMax-Remover: Taming Bad Noise Helps Video Object Removal

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-07-02T02:16:26.812069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:64ef20f070b4d7ff7a66c19251afe0748b794e586758bc803f7b79f42d4618e0

Observation 2069844f-6987-4678-8437-c4648e3b40ad · outbound

This paper cites SAM 3: Segment Anything with Concepts.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation SAM 3: Segment Anything with Concepts

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-07-02T02:16:26.804202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:b31d7c5863ebbedf167c968e5f5711c665804d90d4e8414faa964ea6e4a278a2

Observation 0c992d99-9660-4583-94db-41731bc63462 · outbound

This paper cites Qwen3-TTS Technical Report.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation Qwen3-TTS Technical Report

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-07-02T02:16:26.741492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:f4258ea301d023b72e559c100ac6e509b1b9977e76ee0aa3a2d27d16fd521436

Observation ed43b2f1-d428-4817-bc4d-5b6c76f75484 · outbound

This paper cites Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-07-02T02:16:26.806856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:d895c511f27d32c8441f45b9e6c9ea8b0e74f271209252e887563ae8ae243eb8

Observation ec41fd86-9262-41f2-9019-b5c3cabf12bd · outbound

This paper cites LatentSync: Taming Audio-Conditioned Latent Diffusion Models for Lip Sync with SyncNet Supervision.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation LatentSync: Taming Audio-Conditioned Latent Diffusion Models for Lip Sync with SyncNet Supervision

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-07-02T02:16:26.791937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:9698e90233b824eca80f75e6eb1e82aca4e6a43e2369985093973b8b9fd690b3

Observation 47ff6187-ee6a-405d-a300-ccca218c79d6 · outbound

This paper cites The T05 System for The VoiceMOS Challenge 2024: Transfer Learning from Deep Image Classifier to Naturalness MOS Prediction of High-Quality Synthetic Speech.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation The T05 System for The VoiceMOS Challenge 2024: Transfer Learning from Deep Image Classifier to Naturalness MOS Prediction of High-Quality Synthetic Speech

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-07-02T02:16:26.789450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:27e8cb33bacf9f7eedfea9b9974bc9b46389d5450b3664a8142e5b5b374487ce

Observation 9c1beeb9-2b51-4a73-ab23-95bbac9f59d7 · outbound

This paper cites Kiwi-Edit: Versatile Video Editing via Instruction and Reference Guidance.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation Kiwi-Edit: Versatile Video Editing via Instruction and Reference Guidance

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-07-02T02:16:26.783808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:71c05a3f3d315d6f857aa28d7a74add34f9da97c00ca84bd2400ed8fd575d8c0

Observation 6f9db44e-0142-48b3-96b7-56b4b8944e60 · outbound

This paper cites Metaclaw: Just talk–an agent that meta-learns and evolves in the wild.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation Metaclaw: Just talk–an agent that meta-learns and evolves in the wild

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-07-02T02:16:26.778819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:1cb32a3ba3c68af574e1fb4931c1d71e658840605617bd75fa265de08f82d8e1

Observation 72ff8f0f-2504-47d9-a95b-02adaf134aa1 · outbound

This paper cites 2023 , keywords =.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation 2023 , keywords =

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T11:12:00.839746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:6dafcea6ca08f0bd7b9b1f1498964f88c1d9ed17c92068b3fb117eb510a03ed8

Observation 1ae7171d-17c2-4077-88e4-8db78d3feb49 · outbound

This paper cites Denoising diffusion probabilistic models.Advances in neural information processing systems, 33:6840–6851, 2020.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation Denoising diffusion probabilistic models.Advances in neural information processing systems, 33:6840–6851, 2020

Reference 33

Resolution
unresolved
no resolver link, observed 2026-06-28T11:04:49.366127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:65fa178bca3fc1c209d86960095287559f94229e9855687e3c5e18f6f175234d

Observation 9ccf29e2-36ea-42ba-a99e-cb35d05c7cee · outbound

This paper cites Scalable diffusion models with transformers.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation Scalable diffusion models with transformers

Reference 34

Resolution
unresolved
no resolver link, observed 2026-06-28T11:04:49.366127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:fe777c5de95ae7d4658a760901378f047bbee7c31419de0ba2f77e63c5df9b22

Observation fc659e16-7cc5-4acd-963a-3e5a40732356 · outbound

This paper cites arXiv preprint arXiv:2511.18822 (2025).

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation arXiv preprint arXiv:2511.18822 (2025)

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-07-02T02:16:26.786773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:de7150c366506fe8797e80a425022528a56feb3aa4de649ee5da369823d722da

Observation 3fc6c7ef-5ca3-43f1-9db4-0b6139758611 · outbound

This paper cites Classifier-Free Diffusion Guidance.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation Classifier-Free Diffusion Guidance

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-07-02T02:16:26.781353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:082ccef577826c00b86c72927acb3b01ec17a850c3bc2ce312350643024fda63

Observation 0e6a18fc-fd70-4e50-a77b-975544d28f03 · outbound

This paper cites Ragd: Regional-aware diffusion model for text-to-image generation.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation Ragd: Regional-aware diffusion model for text-to-image generation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-06-28T11:04:49.366127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:aff5fad55072c99ec3f6964a2ac7d3f5c5251435a2cb77dc01f42f9d72eb1bba

Observation c496c4b1-d363-43df-8fa6-283a2691de02 · outbound

This paper cites Score-Based Generative Modeling through Stochastic Differential Equations.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation Score-Based Generative Modeling through Stochastic Differential Equations

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-07-02T02:16:26.794697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:ae3756f9a7a34c3e01b0d95dc8a7d64b5aeeb63e93aca80f2bc79b0b1f20a1c3

Observation c26f449c-96fa-42ca-bdec-25df508e6230 · outbound

This paper cites L2P: Unlocking Latent Potential for Pixel Generation.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation L2P: Unlocking Latent Potential for Pixel Generation

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-07-02T02:16:26.799539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:e76d68ee31c7dd6d937058876d989077d782391897645627cd69cd86aaa3d251

Observation e6b72f66-a01c-4fde-9929-6efabc1cdd15 · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation SAM 2: Segment Anything in Images and Videos

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-07-02T02:16:26.766566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:84bdacc2c3c28d585d212d263b23db7c93f4bed71ef995e43453c0f037e34077

Observation e63be31f-b2b5-4d49-b4cd-f47b0ab46695 · outbound

This paper cites Ivebench: Modern benchmark suite for instruction-guided video editing assessment.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation Ivebench: Modern benchmark suite for instruction-guided video editing assessment

Reference 41

Resolution
unresolved
no resolver link, observed 2026-06-28T11:04:49.366127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:f45dafda0eca958828a6bc3b8db0e7cdd8d4df1cd0c3afcedeb8b509681eca61

Observation b2f06081-76d7-47ec-8aa8-c52e4c44a8d0 · outbound

This paper cites hard to distinguish.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation hard to distinguish

Reference 42

Resolution
malformed identifier
arxiv_id, observed 2026-07-02T02:16:26.801917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:6e16d887191e7eb3b9b9156ebb945784f214b396ded04d9fdda5fcd278bc1227

Pith citing papers

No inbound Pith citation observations are available.