Pith. sign in

Paper Citation Record · LEDGER

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis

As of 6 August 2026, this Paper Citation Record lists 79 of 79 outbound references and 0 inbound Pith citation observations for arXiv:2606.09098.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.09098 v1

Coverage vector

measured 79 of 79 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-27T15:20:14.761337Z

measured 79 of 79 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

79 of 79 outbound references displayed

  • verified exact25
  • verified fuzzy0
  • unresolved52
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 80d7d368-17f1-44c6-a4e4-bdaf4eec8890 · outbound

This paper cites LRS3-TED: a large-scale dataset for visual speech recognition.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis LRS3-TED: a large-scale dataset for visual speech recognition

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-07-03T03:27:35.617364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:3fd70a5d782edee1abbe53e139f8f65a03aa93e2beb46462301e63f8a7238532

Observation 54af6cf5-0305-4ebe-9253-89b2556e1416 · outbound

This paper cites MusicLM: Generating Music From Text.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis MusicLM: Generating Music From Text

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-07-03T03:27:35.619839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:377dc8d3ae2ffe8de74f48b82f082872cbd67b991b7f19af003770ea06a28e75

Observation 39d055ba-9a9b-49e1-b832-0e191ec52896 · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:2f5f87eef401e22bf5c5c20219a9699b251ea5e94013b5d58e943a690cbec864

Observation a172a72f-880c-4506-a03e-4f6720c7dea5 · outbound

This paper cites SoundStorm: Efficient Parallel Audio Generation.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis SoundStorm: Efficient Parallel Audio Generation

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:27:35.652278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:c169ccbe61dac1ec397e071cac82cd41d0853702a5e0260892ef607beb582b7e

Observation 5e9747ed-6d51-4057-9ad9-a6e5ea37dbb8 · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:f667af391b047ffa660d4746fef83c85d663d2435fb922a72f2ec11b26edd855

Observation cc90b911-8cc6-40ec-a3bd-dd0269f41924 · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:66e2f20d8562c18130cbb85cc2bbca62d68a5bebd941f855542414ee54edf660

Observation dc224f67-8baf-4294-a191-1d633688587c · outbound

This paper cites Vssflow: Unifying video-conditioned sound and speech generation via joint learning.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Vssflow: Unifying video-conditioned sound and speech generation via joint learning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:27:35.647749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:8f184aed2172ee0439854e6cad4981caa205bf47686e00ee5f9aa9463cc44490

Observation ec9a134d-c364-4ebd-acb3-df4619dc3699 · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:4c4292ec40ed2d6c6ec2fc56fb363c490803baf7bc97b42881176e6f3731d1fe

Observation e99de82d-a1ec-43bd-9c0d-4b5489ae2bb8 · outbound

This paper cites InProceedings of the 33rd ACM International Conference on Multimedia.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis InProceedings of the 33rd ACM International Conference on Multimedia

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:6ad083680de62a0039e7abfa7cb29a696c970ab28d6cd4f09cd5057b6d35fbe0

Observation 39fb70a7-3355-4d4f-a094-206115954f1e · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:8743071e1ce265959a676673cf0d3e57a26e57df205902042934ab0dda8ec258

Observation 963bf617-6d56-47f9-91f6-1cdcf50e158d · outbound

This paper cites VoxCeleb2: Deep Speaker Recognition.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis VoxCeleb2: Deep Speaker Recognition

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:27:35.655302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:f405656fc147e5f8dd8b9d4e6c09642ac38d491df352090666cfc0c64c946d46

Observation 52923315-3cf5-450e-9b30-d9aca60cfa28 · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:b3bbffd5ff37ad4305398aea0a830ab75c8cf1c58aabf1e715f900234ee633ff

Observation ff9a7f16-2c46-45c8-a08e-3796aaecd095 · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:d77880fde519e784dcc6bc5fcfc8ca0cc7d2526cb3a07578fce80236a98edb31

Observation bc3a717d-cfc4-4df3-b1d7-5fff953deedb · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:a7ee2f83e0f9cdf3c73e6109a8c906ad91e232500e54319afdabfb1282728dd9

Observation 170dee6d-5a24-418c-85a9-05405322760d · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:e9c232a145fda9ffb69a4b2a7d2b9fdf15966931429857ce1b221ac8e600ed47

Observation ed4ce7f5-05d2-4159-9a2d-09dac914e58a · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:744292ccd1070587720370afd770f1839d61b41781c8697b5b2c8c369777e67a

Observation 9ffb4329-3caf-40c1-bf50-5a7a6d1a41df · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:13fb93585e83be1bca8ec10ffcf56bd2bf2d9f2545cf3eac5885251606b1d2bf

Observation 2530444b-0ef2-47da-be39-2cb239eb691d · outbound

This paper cites High Fidelity Neural Audio Compression.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis High Fidelity Neural Audio Compression

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-07-03T03:27:35.675467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:6882ed2245985a6aa8f2a6e0f868f60d66e339272ea0ea6eb953e36330e82fa9

Observation 5c0a2463-f0e8-48e2-b568-851285f26dbf · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:c9426b3b6a38fe3e29258ae5f3896cf6c41e6e9b1dd7954313cdd16e29e00711

Observation 4b22dbed-fe8e-472d-bacb-74d0280e9bb0 · outbound

This paper cites CosyVoice: A Scalable Multilingual Zero-shot Text-to-speech Synthesizer based on Supervised Semantic Tokens.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis CosyVoice: A Scalable Multilingual Zero-shot Text-to-speech Synthesizer based on Supervised Semantic Tokens

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-07-03T03:27:35.620602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:34f573800f6124faf559a38dc2b32ffc6446a0e3d3ca02df525d5f9f2cdd6713

Observation f65fc66c-a3b5-4181-a64a-8d9872206e76 · outbound

This paper cites CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-07-03T03:27:35.625576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:f8c53915466b107d25a0251127b9f7d34d9566ccf445f7d8f5ed82c00b7dbd16

Observation 031e8125-6e54-46cf-8798-4b323ad01d04 · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:3644dda93c80dc3ee226229dd2ef17a5e7bf955bd35f284107ce5373f97e7279

Observation f9558494-816e-468e-88c9-f690ce419d9a · outbound

This paper cites InICASSP 2023-2023 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP).

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis InICASSP 2023-2023 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)

Reference 23

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:20b5bcbef5c2e396b82b65d0754843f818a9b1e0bbd54d4ba608175bedb1502a

Observation 958d2f8f-62e7-4f9f-ac72-822d7d5809dd · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:554fe7ad2ba4b6bd64c88a0c8d4c28f8b4b85e1a0fbe7731b044a60bb171aa94

Observation 651510cc-7438-417e-b605-ed5b15c3a581 · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:3095af112d69fcc35bb5bbc35ab4c8b76a0b084c899212e23ac5752d98cea50a

Observation 37d25c7f-a914-4692-a34a-ed230644fab4 · outbound

This paper cites InProceedings of the 31st ACM international conference on multimedia.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis InProceedings of the 31st ACM international conference on multimedia

Reference 26

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:428acee9777bb65ebb7046a7952f42d7a58efa65e21bec2ad47193009b5584a3

Observation 36d9ba5f-aff2-45ee-8782-99a93cecc3b9 · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:df30fd36c8375740603c04cecee47f2941cd77e2d6c47fa0c333c89d59e24b77

Observation ef0d6fc6-ee5e-44df-98a1-8c166c3faaba · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:4c63193440003bc2535b71b7a7341efba68d160fbb4e797fbecef722d24d601d

Observation 0f3a401c-6ffa-45d9-8b66-5b3376effd5b · outbound

This paper cites FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:27:35.634796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:6fbbc03c74f46abfe328178ce1a6356d489d5eca87bb9904846cd24f783a3371

Observation ebd7bb2f-a995-4097-a919-d73d46645e0d · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:67efc22a8e5d43f471771aaa4c2225031bc81b8f35297c278f2d66843c88f911

Observation a96a1e6e-25c5-4573-ad7d-03bba037baa1 · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:8873436e8226f3cfc7946d7c466efb981a15aa61a03f0680efcd8933a86e6e10

Observation cb84456e-ef8c-42ae-b36a-6bec764cb13e · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:f56de6f42d5690f79ba118fc4897ed7ab6bf648b077a07c0491ec18b734c1f3b

Observation 3e3859de-1679-4071-9137-11104745c44b · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:0596508f41eb8d18a85e8a43c5522ccf557d61f44256f6d2ec3777e1d790d695

Observation eedf0ef6-d58e-40b0-bac1-de25619142f4 · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:3df890d09e826999d636318745821d534643234d7cbb4984cacc97ef20331b5c

Observation 64bffb7d-7a08-4b7c-9fb8-020f74dd064c · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:d2069f438e72f3c2b81b4e99ccfc0d0f56b329ff63073853d56f67cfcc65dc99

Observation f9cb6270-4fc0-4fba-8be1-7451caaa2f12 · outbound

This paper cites AudioGen: Textually Guided Audio Generation.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis AudioGen: Textually Guided Audio Generation

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:27:35.639979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:9252fdbc0318bb7858186f36a2cf7ddb426a41b79fe20c5ebce106a49afd35ac

Observation c49cfb26-cbb0-4501-81ae-aab469efd8e2 · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:3bb77c840871559a69675596a2755d5991f75e65a4d1a8278659ee303f65f4e7

Observation c76ecad6-ae1a-47de-a6d2-55e93a7a2b2b · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:d05e67adbd9727f4ea053e10bfa3b3c72f082f6943c9f00b1be02ab0fb1bc48d

Observation fa2f1c33-5de6-46fb-be1a-0c610cd996dd · outbound

This paper cites MeanAudio: Fast and faithful text-to-audio generation with mean flows.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis MeanAudio: Fast and faithful text-to-audio generation with mean flows

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:27:35.645157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:fbe8e3b86e11531474bc228ed02c26a9121c018b234d1be9a63f019460d938e4

Observation 6b02e290-6baa-45d1-9e4d-b2fbb6d44354 · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:2df1e9819d81288db8611353d5e242b05a87c6837b59c2106f406a5fae851052

Observation 568a3b66-aa16-4c2a-8c93-b7ed9ed44457 · outbound

This paper cites Flow Matching for Generative Modeling.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Flow Matching for Generative Modeling

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-07-03T03:27:35.670895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:3b1a9ee8bce522ef811a6c6fbcfb16053d6a8fd8a389cd43530195b97d1d8fa7

Observation bc2620cd-f520-4e0a-9097-9ef173775668 · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:dcab33b7e632abbb51031cadcc8980046a308e33ee182ecd272631b97f2af192

Observation 32e0cc43-3cf8-4d55-b6d6-749757f4ad14 · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:b6019b4b795723432071f43bc10ff8e99f2a0f17dfded26920994bd45836081f

Observation e6286558-f158-4634-bc2b-dc9058a338d9 · outbound

This paper cites arXiv preprint arXiv:2601.14777 (2026).

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis arXiv preprint arXiv:2601.14777 (2026)

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:27:35.622489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:c45609c82c129c36cf7629463ba2610614890b0159558d037ba94169f558bd86

Observation 6a57ac6d-c5ad-41e0-949a-4c268b81d625 · outbound

This paper cites Autoregressive Diffusion Transformer for Text-to-Speech Synthesis.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Autoregressive Diffusion Transformer for Text-to-Speech Synthesis

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:27:35.628751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:f02943b775ae41b6620bf154e7c4ce6fe3ab7d43f6b62691ae70406862e2f78d

Observation 33303afa-dbfa-4d0f-9c66-4e5028a73c99 · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 46

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:1267c58e5170f7296e16db7381d8665eb7e2eb7f0cc1ad1f2d8f9b0020cf869f

Observation 566219a2-5266-490b-a9d0-ee3d916e6a76 · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 47

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:631d27dc4955bf1d6f365ea912a067e778903cf7ced21a4d14a766964dea6918

Observation a41d6c3a-10a6-4867-8d3d-ac6d25689172 · outbound

This paper cites In ICASSP 2024-2024 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP).

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis In ICASSP 2024-2024 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)

Reference 48

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:c95e8c3e9a0621f6355f70fcfc97387d1dfb066e36cd23ab7286ef31e4788683

Observation 27c66932-8c89-4ae4-b071-4143ca729f56 · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 49

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:124215c40a96e04b9fb4aa90ed60458f952102be3588db785090a45f26518962

Observation 879e34ea-0825-4489-a48e-94fa221b15fc · outbound

This paper cites Semantic-vae: Semantic-alignment latent representation for better speech synthesis.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Semantic-vae: Semantic-alignment latent representation for better speech synthesis

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:27:35.631726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:365eeb496e6a24b8d11fc356e702f1cccc580cd0f389b27d2a0dda146a526d84

Observation d9d511db-ffa5-4942-8152-aa390e640828 · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 51

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:76caa77d8e42da065133722a6dbb599c151ae82f3504a543645d7e70f78e990a

Observation 42c6e303-2cda-4745-8b19-b5db5210b1c0 · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 52

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:c466d6650a6bcce0dbc437e469d1a75b3609d97c242d0b2691062cf0a61dbaa8

Observation 76c794d5-9534-4c70-bbb8-f9cd85a51d13 · outbound

This paper cites VibeVoice Technical Report.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis VibeVoice Technical Report

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:27:35.637338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:4d74ce0bb0db41ce9170248784c78246701ab686da9b3a1050b8e7efda3e8674

Observation dff74fe8-1ec6-49e3-9e99-24ef303b3fc9 · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 54

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:ed455d6b5f0ed6b83a799234d47fc50ca41590601ded027e0ac72cb9d04d38a7

Observation e7e26847-0e30-4470-8303-65f5e17abd85 · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 55

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:77c9a648a8968677816564256a36b0353b712c4040312e5d4bbc8607d10486af

Observation f65a8580-983e-45e2-bc34-78cadef220c0 · outbound

This paper cites UTMOS: UTokyo-SaruLab System for VoiceMOS Challenge 2022.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis UTMOS: UTokyo-SaruLab System for VoiceMOS Challenge 2022

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:27:35.667011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:835fa2fdb74663175020095fa94ee6dc5af129135eb067df04c000b196336b60

Observation 6da4e840-e457-4dd1-a0d6-ed4167831369 · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 57

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:2f6a3fa797e8cbf24157e7fe1db03af5f9f82eef18ab14a55891512d3e340fd4

Observation 4dd24823-d855-4901-b87e-60553e6af052 · outbound

This paper cites In2018 IEEE international conference on acoustics, speech and signal processing (ICASSP).

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis In2018 IEEE international conference on acoustics, speech and signal processing (ICASSP)

Reference 58

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:c68062742ef334116eb412d7c36f2dc7d4e3488af84dad351f7a9f18f9e52f07

Observation 641792ed-feaf-437a-8aec-8387dd5eddc8 · outbound

This paper cites Learning Audio-Visual Speech Representation by Masked Multimodal Cluster Prediction.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Learning Audio-Visual Speech Representation by Masked Multimodal Cluster Prediction

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:27:35.642562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:1320776fd05c06a085d5c9a2d0481184db24e6563ca7060f60d05fd684752a06

Observation 9ade00a0-5574-4de2-b951-8a91c9a5c52e · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 60

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:61971afff4582893d9321ad13b79cb68d032f06ed0be5c7e86bce6a35459693b

Observation 2b82e9d5-75de-4e28-92e0-d426cd8dc933 · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 61

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:5dd38c0813f40015a2ca17376a06d7b8ef4657e48112ad22d13f29ffc587809b

Observation 82cdbb27-0be4-4605-910a-8ca5217ebd7e · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 62

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:ec4639dba9595d01782e28ee77d9fba41512560578b20e4586f9968b7d9b8e05

Observation 2c590f6f-ae5c-47a6-bbc3-0c9cbd2afdb7 · outbound

This paper cites Audiobox: Unified Audio Generation with Natural Language Prompts.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Audiobox: Unified Audio Generation with Natural Language Prompts

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:27:35.662075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:f28abf690dcc97499d8d4597f065c87f8cc55ddcde78d18c3658f424c7c614b1

Observation fb2b3e22-4bfe-4af9-a40f-7eb69dca2ec1 · outbound

This paper cites Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers

Reference 64

Resolution
verified exact
local_arxiv, observed 2026-07-03T03:27:35.617782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:44de3fc535ad8c9abcaf4b1b5d0f6bff01f59f8b4f94418afbf930b80c4fd823

Observation da81b2c8-3820-41d7-bae9-82a534c4321a · outbound

This paper cites Mel-Band RoFormer for Music Source Separation.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Mel-Band RoFormer for Music Source Separation

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:27:35.608385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:102deaceb03471547599be2b30b0686c62139e3c6edbde07db33f6d48b30cb21

Observation 1ae5565b-83bd-4351-8c3a-fa0453af18b8 · outbound

This paper cites AudioGen-Omni: A Unified Multimodal Diffusion Transformer for Video-Synchronized Audio, Speech, and Song Generation.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis AudioGen-Omni: A Unified Multimodal Diffusion Transformer for Video-Synchronized Audio, Speech, and Song Generation

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:27:35.615037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:3709fcac7cf82c8a37bc1107c23eb637b0f5558d2262209c43ddd723fbb62863

Observation 96739f5f-c1a2-40d6-8d11-3ecb64b70cfc · outbound

This paper cites Qwen3-Omni Technical Report.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Qwen3-Omni Technical Report

Reference 67

Resolution
verified exact
local_arxiv, observed 2026-07-03T03:27:35.605707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:c642a362320b8765837d70979fbb04b5be6dde5e5499b824f86497969da65984

Observation eb63cd64-c41f-4d99-8ea8-fcb1c16a4244 · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 68

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:62e89f7fad3b8721ce5456e144dbfd14b6adb27ec0339f3c4cb98e9b0a034e9a

Observation 94ca67c2-1a68-4ab6-a38a-4c7a2f737ac0 · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 69

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:831e7d4a72891134f08c29af6afbbdc7d0d5a9c8d9dc21eb106ee431e2a02e27

Observation 8cbd1fc1-8f43-4ab8-9b93-261530d4c1e7 · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 70

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:79ba695155874f7907e382a2f13ea918585db5d36ee73bf9b9d86efe6b17a5b1

Observation 22c41ad1-d414-458e-8d08-7e50d83c24fb · outbound

This paper cites DeepAudio-V1:Towards Multi-Modal Multi-Stage End-to-End Video to Speech and Audio Generation.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis DeepAudio-V1:Towards Multi-Modal Multi-Stage End-to-End Video to Speech and Audio Generation

Reference 71

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:27:35.611351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:fa7e4de12d5ac96fdaba3a5460cb2e148299efacfd77dfe6a7d210a84043ce85

Observation 6c16536b-5431-464c-9bc3-6b2d6df8cd58 · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 72

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:42b125d5306a8637d16c30c64babc31a03ac37e6f3bce880925a7513ca2f8b61

Observation 237041bc-8f08-43b5-ae17-01c574b6362e · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 73

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:32e34977f4dec05d3482f16cc2253b1abeea33554a18692b5182a63f5ef8eb3f

Observation 50491cdc-5be8-4f12-934c-2583d59cbf47 · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 74

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:31c376056225ebac18cdacce5226f369996919e46d1bff12751167ddb35d3740

Observation bf2988f5-115d-4032-b10a-de0aed39b4fb · outbound

This paper cites DeepDubber-V1: Towards High Quality and Dialogue, Narration, Monologue Adaptive Movie Dubbing Via Multi-Modal Chain-of-Thoughts Reasoning Guidance.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis DeepDubber-V1: Towards High Quality and Dialogue, Narration, Monologue Adaptive Movie Dubbing Via Multi-Modal Chain-of-Thoughts Reasoning Guidance

Reference 75

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:27:35.596113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:d1adadfd6c02c195548b0d7f8ab9580d84acac8a0e30f1dd7e27e030aefd38bd

Observation 9baea2ff-f776-4718-a006-148675ca13d3 · outbound

This paper cites Unspecified / Clean.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unspecified / Clean

Reference 76

Resolution
malformed identifier
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:8b9b9829cfddb698f17373e17dae730ac29fb1bbd6f6df814c669c4a0bee71bf

Observation b38a3556-0374-4765-879e-64385fe6ca58 · outbound

This paper cites an unresolved cited work.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Unresolved cited work

Reference 77

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:43e27a908b6c4c32d4a70508382943c92e4798cc787d0936b94a64d1b2189c94

Observation 420d76f0-8e53-4e78-a0ca-c623771c0458 · outbound

This paper cites Sound relationships: describe changes in loudness over time,layering of different sound sources, and any sense of spatial depth or distance.] Figure 4: Overview of Prompt Design.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Sound relationships: describe changes in loudness over time,layering of different sound sources, and any sense of spatial depth or distance.] Figure 4: Overview of Prompt Design

Reference 78

Resolution
unresolved
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:b101cc1531ac578cb02248c4580df4369152face1156ebf83185ff12c3bbb88f

Observation a17b5718-4d9d-470e-891b-836ca7c37364 · outbound

This paper cites As shown in Table 5, HoliDubber consistently outperforms the decoupled pipeline across nearly all metrics.

HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis As shown in Table 5, HoliDubber consistently outperforms the decoupled pipeline across nearly all metrics

Reference 79

Resolution
malformed identifier
no resolver link, observed 2026-06-27T15:20:14.761337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T15:20:14.761337Z digest=sha256:8bb846f96825016110c552f1cea8db0159669bacef672964c218179828d8be71

Pith citing papers

No inbound Pith citation observations are available.