Pith. sign in

Paper Citation Record · LEDGER

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing

As of 8 August 2026, this Paper Citation Record lists 33 of 33 outbound references and 1 inbound Pith citation observation for arXiv:2505.16279.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.16279 v1

Coverage vector

measured 33 of 33 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:07:10.491851Z

measured 34 of 34 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:07:07.961535Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T15:07:10.886803Z

Reference resolution

33 of 33 outbound references displayed

  • verified exact2
  • verified fuzzy20
  • unresolved10
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c74c4cec-69f9-42aa-861a-f36edf271657 · outbound

This paper cites Exist- ing dubbing methods can be categorized into two groups, each focusing on learning different styles of key prior information to generate high-quality voices.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Exist- ing dubbing methods can be categorized into two groups, each focusing on learning different styles of key prior information to generate high-quality voices

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:13.014975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:07:07.840682Z digest=sha256:3c4334e7fb72fb78f660667575cd2339f8bf470d3f3f3be2a88f8d129e3b1ebb

Observation 74e62b26-d73e-4307-998b-7010f449b1a2 · outbound

This paper cites MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing

Reference 2

Resolution
malformed identifier
local_arxiv, observed 2026-08-07T15:07:10.927698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:07:07.961535Z digest=sha256:f2a4efac0b2582b2ae846d8218fd6325bb95f46c1958767b104c8e5282fa7f7d

Observation 907a7eee-cea2-48df-a5a4-d728b32c66ec · outbound

This paper cites Datasets Emilia is a comprehensive multilingual speech generation dataset containing a total of 101,654 hours of speech data across six languages [21].

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Datasets Emilia is a comprehensive multilingual speech generation dataset containing a total of 101,654 hours of speech data across six languages [21]

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:12.775612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:07:08.031502Z digest=sha256:dacde5c25a488a089a857b559ce8fa4e0a18eca93c17677be656e10499bda825

Observation 375117c2-b77b-41b1-86d7-a80476fd686f · outbound

This paper cites To as- sess pronunciation accuracy, we use Word Error Rate (WER) with Whisper-V3[24] as the ASR model.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing To as- sess pronunciation accuracy, we use Word Error Rate (WER) with Whisper-V3[24] as the ASR model

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:12.637611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:07:08.114058Z digest=sha256:d2112126ecb36c5c275bae80c9afc4c41b1166d52062f7c109b27b0f56b072e5

Observation 498d685b-d814-4216-89ac-e165952d658e · outbound

This paper cites Additionally, we have de- veloped a movie dubbing dataset with multi-type annotations to enhance movie understanding and improve dubbing quality.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Additionally, we have de- veloped a movie dubbing dataset with multi-type annotations to enhance movie understanding and improve dubbing quality

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:12.543048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:07:08.198474Z digest=sha256:d9ddafa999fe65bac61cc3a642b8d7414fa8ff1f7b4901dc20389e923957157b

Observation 0aea5d8d-5aee-43db-9ce9-3407e0333e0f · outbound

This paper cites V2c: Vi- sual voice cloning,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing V2c: Vi- sual voice cloning,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:12.439554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:07:08.292579Z digest=sha256:4c182ffb3cd4556c2e3c06fe3b5c725a9cf6cd8973a019cb96c9f8a5d7b6338c

Observation fc319574-6ad6-4352-ad12-923d5cb6d196 · outbound

This paper cites More than words: In-the-wild visually-driven prosody for text-to-speech,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing More than words: In-the-wild visually-driven prosody for text-to-speech,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:12.336180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:07:08.360687Z digest=sha256:ee0ac035c275cae9d0e47f18e15c425bd1acdd872a4ff9b567beb32268e394e3

Observation 575950f6-6b37-4f81-b5a9-6886c6669eb4 · outbound

This paper cites Generalized end-to-end loss for speaker verification,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Generalized end-to-end loss for speaker verification,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T15:07:08.454604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:07:08.454604Z digest=sha256:630556f3f3ace516e9cd21b9b642393173f3aa66554f9e4b3a2fc2cadef08f95

Observation d6475ee3-f567-450d-b6ad-8bbabb0b15de · outbound

This paper cites Learning to dub movies via hierarchical prosody models,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Learning to dub movies via hierarchical prosody models,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:12.220221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:07:08.557957Z digest=sha256:8c7c7867fc00a05a1dcd19068220676fa5f169e835e51108a7ccdc9409f72d8c

Observation 52d242fa-daf6-4206-9957-d93433bf20b2 · outbound

This paper cites Neu- ral dubber: Dubbing for videos according to scripts,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Neu- ral dubber: Dubbing for videos according to scripts,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:12.046317Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:07:08.676630Z digest=sha256:afd1a2e05667f41f3c753409fdc15d9dd75f23527c8e91c6ce851f85e6bc5bba

Observation bc5635ba-34aa-4c19-942f-ce1a908c9175 · outbound

This paper cites Imaginary voice: Face- styled diffusion model for text-to-speech,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Imaginary voice: Face- styled diffusion model for text-to-speech,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:11.944032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:07:08.818516Z digest=sha256:716a01894a2aa5ccc169ce1e4d8c50219701b3cdb839320cc3569767c694f3e4

Observation 0912692c-d95d-4188-8855-7853c0c477ef · outbound

This paper cites Mcdubber: Multimodal context-aware expressive video dubbing,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Mcdubber: Multimodal context-aware expressive video dubbing,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:11.867601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:07:08.943533Z digest=sha256:01ee7e1decad83cf33ae319d28515a68ee7c1f796f650044b0a1219f9e259c1b

Observation 4c9175e7-1513-408d-8ac9-b6bb9cec84cd · outbound

This paper cites Audiopedia: Audio qa with knowledge,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Audiopedia: Audio qa with knowledge,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:11.766439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:07:09.054916Z digest=sha256:c72154e55e351836fac1caf3d9d92be5e994223742a9b8e13f02080931b79324

Observation e501b4aa-ab4b-48e9-bb12-9dc7b57355a7 · outbound

This paper cites Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T15:07:09.148574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:07:09.148574Z digest=sha256:1a33749ee970974e96b003b4ecdddf4377da2b11e7b92ac65ca554f403fc9164

Observation f8be3954-69be-438e-9ad5-4432531ef75d · outbound

This paper cites VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T15:07:09.258423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:07:09.258423Z digest=sha256:e929d9827d4a1effb479bcc4b5a067fb5b6401d63864ee2a303cf06cdd9041ff

Observation 6f78254c-f256-495b-aee1-8f76d4b0dadb · outbound

This paper cites Learning to dub movies via hierarchical prosody models,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Learning to dub movies via hierarchical prosody models,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:11.669537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:07:09.359041Z digest=sha256:3a978b799a30e79190a18e2ba1f59e95e5f8826c340b7251082573c441955a89

Observation 37ce9657-f384-4ed4-a705-5791653bab47 · outbound

This paper cites StyleDubber: Towards Multi-Scale Style Learning for Movie Dubbing.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing StyleDubber: Towards Multi-Scale Style Learning for Movie Dubbing

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T15:07:09.453123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:07:09.453123Z digest=sha256:a21dec0eb6d230630f21d7b95e4e8b2210540423f08705d327c959cd1c05fcba

Observation 3a0d21ad-481e-4ab7-aa36-f256148f54d8 · outbound

This paper cites From speaker to dubber: Movie dubbing with prosody and duration consistency learning,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing From speaker to dubber: Movie dubbing with prosody and duration consistency learning,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:11.589518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:07:09.527245Z digest=sha256:b9a30185a6a68f995cfca42352d82009a2edeaba747db4713f5eda433574c7d5

Observation a2ca5ff3-e452-482c-a4cc-9c0d122a62cd · outbound

This paper cites Visual instruction tuning,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Visual instruction tuning,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:11.508937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:07:09.645485Z digest=sha256:de9a3a257abfe5e453d9f998fdafbbc6b6c5ce825e800be277d0cfdb4262c630

Observation fb820020-97f2-4de1-9232-86c7858f292d · outbound

This paper cites LLaVA-CoT: Let Vision Language Models Reason Step-by-Step.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing LLaVA-CoT: Let Vision Language Models Reason Step-by-Step

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T15:07:09.737133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:07:09.737133Z digest=sha256:52fa1f8776a8276bce9541d8652300da3ba8715e5013d61907f1623c7e0b6c05

Observation b22e3ad7-cc45-48fe-b3c1-c8e53935a9a1 · outbound

This paper cites F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T15:07:09.853775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:07:09.853775Z digest=sha256:d66e604ea70e3cef6018cf057711266dad58195781a0c105e9ce6955b80dbb49

Observation 297f42d7-bd54-44b4-aed5-644aafba4cb8 · outbound

This paper cites Flow matching for generative modeling,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Flow matching for generative modeling,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:11.422727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:07:09.981897Z digest=sha256:cebfd33059c3ced3e1a5ca74758d1b3a5b293598b0290f2905860381000e7ad4

Observation e275e094-90a4-4712-9e20-fe0c7af3ff27 · outbound

This paper cites An audio-visual corpus for speech perception and automatic speech recognition,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing An audio-visual corpus for speech perception and automatic speech recognition,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T15:07:10.333160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:07:10.333160Z digest=sha256:42aae111ff7cd8139019a5e56e993dec60ff1a777c0f854eec328c54faeae5b4

Observation 3d6cc985-0e82-459d-850e-bb737c7a2e05 · outbound

This paper cites Exploring the limits of transfer learning with a unified text-to-text transformer,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Exploring the limits of transfer learning with a unified text-to-text transformer,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:11.331277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:07:10.067509Z digest=sha256:7285c07b903895a4c315b81db6d46c960472cb66f4ff1265d806218b643d8626

Observation 0cdbc5a6-b3fb-4b09-aebe-f4d579bae2a7 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Learning transferable visual models from natural language supervision,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:11.240478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:07:10.120717Z digest=sha256:a38ae025edddcc0de3678f88279fecedb325a8d5c14034f202cfad5834fd40d1

Observation 710d353d-232f-42a6-875b-bfde1a713ea6 · outbound

This paper cites DiVE: Dit-based video generation with enhanced control,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing DiVE: Dit-based video generation with enhanced control,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:11.170141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:07:10.174456Z digest=sha256:410718f4c14f9c2c28fc2ada110ff09398b3a49b549d9a1f6ed2ec9f3af30580

Observation f3fe19ef-6c80-4bbd-9835-209e6f266201 · outbound

This paper cites Emilia: An Extensive, Multilingual, and Diverse Speech Dataset for Large-Scale Speech Generation.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Emilia: An Extensive, Multilingual, and Diverse Speech Dataset for Large-Scale Speech Generation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T15:07:10.212017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:07:10.212017Z digest=sha256:70e57d5f520af47d310c54a2632ee4df70a627e6748354b690420b8dfb996f8f

Observation 5e30b8f5-a9a6-4481-8b01-4d2924fa4b96 · outbound

This paper cites V2C: Visual Voice Cloning.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing V2C: Visual Voice Cloning

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:07:10.775143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:07:10.279629Z digest=sha256:5b8b67fd5ca272dc45017c063203ba671476d5435474d8e80ee5623d94148db5

Observation 092bb8b5-818c-403a-9762-39b6ad937b9a · outbound

This paper cites Robust Speech Recognition via Large-Scale Weak Supervision.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Robust Speech Recognition via Large-Scale Weak Supervision

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T15:07:10.363576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:07:10.363576Z digest=sha256:c27a76d8314b8874779455188acdc14d23fdacd751a49b804d77f65fd1e50ce7

Observation fd3fd5fe-5aa3-4ab2-b700-eb822cd43a8a · outbound

This paper cites Location-Relative Attention Mechanisms For Robust Long-Form Speech Synthesis.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Location-Relative Attention Mechanisms For Robust Long-Form Speech Synthesis

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:07:10.612708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:07:10.404068Z digest=sha256:f45414b88b635279b0ebad7cbc1563c04a967fc07112b89c9d98f6583a2206ba

Observation 9edd3850-dac4-4e3a-a2ff-8b7a03a45d5b · outbound

This paper cites Tem- poral modeling matters: A novel temporal emotional modeling approach for speech emotion recognition,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Tem- poral modeling matters: A novel temporal emotional modeling approach for speech emotion recognition,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:11.100841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:07:10.445186Z digest=sha256:ed6fb53b4814e81f78394be9a6eb09672c66ff06c2827d9c73a68987dab28cd9

Observation 0c2c5a5c-a7cc-448a-aa58-3a770be52bba · outbound

This paper cites Out of time: automated lip sync in the wild,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Out of time: automated lip sync in the wild,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:11.031819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:07:10.491851Z digest=sha256:847cf646cb9cdf9cff7d79977212db3b6f5537489b828252968783f4ea5de2a1

Observation afeebc99-ff03-4cc8-8e88-c14dd6001225 · outbound

This paper cites Available: https://openreview.net/forum?id= PqvMRDCJT9t.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Available: https://openreview.net/forum?id= PqvMRDCJT9t

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T15:07:10.028412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:07:10.028412Z digest=sha256:c0f7d9b06cae92379ede228835ccefbafbc93cddd227aaf2e3c1cc7752ce4e22

Pith citing papers

Observation 74e62b26-d73e-4307-998b-7010f449b1a2 · inbound

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing cites this paper.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing

Reference 2

Resolution
malformed identifier
local_arxiv, observed 2026-08-07T15:07:10.927698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:07:07.961535Z digest=sha256:f2a4efac0b2582b2ae846d8218fd6325bb95f46c1958767b104c8e5282fa7f7d