Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-15T11:56:13.914121Z
Paper Citation Record · LEDGER
As of 4 August 2026, this Paper Citation Record lists 66 of 66 outbound references and 1 inbound Pith citation observation for arXiv:2603.14267.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-15T11:56:13.914121Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-10T15:46:58.010112Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-05-11T09:51:00.769273Z
66 of 66 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 9d5f1bba-507c-4995-b77c-a2db99d50867 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization wav2vec 2.0: A framework for self-supervised learning of speech representations
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1d678fc8-8c2e-4127-bdd9-d8e39793d0f3 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization V2c: Visual voice cloning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8a06e516-a4b9-40c4-ac2b-67d1c6f02cfb · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Vall-e 2: Neural codec language models are human parity zero-shot text to speech synthesizers
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0b9d66b4-0c1b-478f-8411-b3b0341ad95a · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Neural codec language models are zero-shot text to speech synthesizers
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9a249c86-d129-439a-99a1-9e4e73fb4754 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization F5-TTS: A fairytaler that fakes fluent and faithful speech with flow matching
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 78267e3a-602d-4195-a58e-580949862523 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Intelligible lip-to-speech synthesis with speech units
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 24c5ffae-93be-4c96-ac7c-116228609ef0 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Aligndit: Multimodal aligned dif- fusion transformer for synchronized speech generation
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3ff8c2c0-a1eb-4df4-95af-a7471d34c2dc · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Accelerating Diffusion- based Text-to-Speech Model Training with Dual Modality Alignment
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation aa7c9b87-bf90-4dc2-848d-dc3346e30468 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Reducing f0 frame error of f0 tracking algorithms under noisy conditions with an un- voiced/voiced classification frontend
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6113a4be-2e74-42ec-b59e-4a8f894d0367 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Out of time: au- tomated lip sync in the wild
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 65516fd5-6a0e-44db-99b9-b14de03b9c53 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Learning to dub movies via hierarchical prosody models
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ecbf7438-009c-4c50-9b91-65490977c615 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization StyleDubber: Towards multi- scale style learning for movie dubbing
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d0dd85ca-9e14-46ab-b4a9-62df08bcc87a · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Emodubber: Towards high quality and emotion con- trollable movie dubbing
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0b3d3851-40f9-40ce-b885-75240137e53c · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization An audio-visual corpus for speech perception and automatic speech recognition.The Journal of the Acoustical Society of America, 120(5):2421–2424
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 610db2db-cf00-4f9a-8785-5b95df970a71 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Sigmoid- weighted linear units for neural network function approx- imation in reinforcement learning.Neural networks, 107: 3–11
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation eded40d4-19d1-4aa7-8c13-607fbaed4df8 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization E2 tts: Embarrassingly easy fully non-autoregressive zero-shot tts
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 05cec627-2750-4cf7-b4ec-a65fbff14f5e · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization LLaMA-omni: Seamless speech interaction with large language models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7af4c8a5-b4c5-44e7-8cc5-2634a55c9f33 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Unresolved cited work
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9230ce63-01fe-45c3-a14b-00f7964073b3 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization V oiceflow: Efficient text-to-speech with rectified flow match- ing
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 972ee4d0-1303-4922-bdb3-2f0576d74bae · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization VALL-E R: Robust and Efficient Zero-Shot Text-to-Speech Synthesis via Monotonic Alignment
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation be3df557-aabd-4dfb-aa11-22010fe593a4 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Boosting large language model for speech synthesis: An empirical study
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ce3bc253-7c51-46bf-b025-60570bd53054 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Gaussian Error Linear Units (GELUs)
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation eda4591f-69d0-4a8a-a2a1-d51027a49c27 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization OZSpeech: One-step zero-shot speech synthesis with learned-prior-conditioned flow matching
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 83c8e916-1ff7-41bc-a4bf-56e67d60473b · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Neural dubber: Dubbing for videos according to scripts
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8d86afa4-e82d-4cd7-8dd9-fe2d248850cf · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Hunt and A.W
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b13465eb-c358-47ce-93d0-ffc7f18cfc66 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization MobileSpeech: A fast and high-fidelity frame- work for mobile zero-shot text-to-speech
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 33523617-7809-41ba-9b6a-61afe36c93b1 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization NaturalSpeech 3: Zero-shot speech synthesis with factor- ized codec and diffusion models
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0707335d-1c50-4199-bbf3-b7666a8f69f4 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Zet-speech: Zero-shot adaptive emotion-controllable text-to- speech synthesis with diffusion and style-based models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 91dd46d8-7084-42e2-888b-aab5cad29b4a · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Glow-tts: A generative flow for text-to-speech via monotonic alignment search
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 013720c2-a86e-4ca4-94dd-56a97aeed683 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Shih, Rohan Badlani, Joao Felipe Santos, Evelina Bakhturina, Mikyas T
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 92891c00-3e5c-433a-aee5-477f5062b300 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization V oicebox: Text- guided multilingual universal speech generation at scale
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation fbd66383-48ca-4d87-bbb1-ee6403c2eca6 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization DiTTo-TTS: Diffusion transformers for scalable text-to-speech without domain-specific factors
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5c43cf4b-cb38-4815-9f3b-e93c03285115 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Decoupled weight de- cay regularization
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0db26fd0-7396-452f-b1a8-3d1895eed41a · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Monotonic multihead attention
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 72858e84-e149-4e9f-a9f1-931307c2d965 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Montreal forced aligner: Trainable text-speech alignment using kaldi
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a251bfc1-96a7-4048-8339-e99a18161037 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Matcha-tts: A fast tts architecture with conditional flow matching
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5ae73557-32a8-4416-8004-af9aa69e2c3f · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Meng, and Furu Wei
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3c2b4ae7-6690-47b1-a932-130221a83ab8 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Mish: A Self Regularized Non-Monotonic Activation Function
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7d0661db-97be-4de2-8a1f-81cc01666265 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Unresolved cited work
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9e85afdf-2bbd-4be9-91e7-3fbad0df55e5 · outbound
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a51c1974-851f-4233-808c-5b6c7f5fc9d6 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Scalable diffusion models with transformers
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0aa3e1f0-945f-4ba2-873f-fcac30605741 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization VoiceCraft: Zero-shot speech editing and text-to-speech in the wild
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 22c663d5-c5e6-4bbf-aec6-6b26cf43b5e9 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Unresolved cited work
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b76406a8-6833-4157-8000-65cd8ea8ab7c · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Speech resynthesis from discrete disentangled self-supervised representations
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 269ba7b0-f5e4-4e7d-bc8a-63d08b6f3fc6 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Robust speech recog- nition via large-scale weak supervision
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1a56d666-ccce-4957-adde-0f7f44932a82 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Fastspeech: Fast, robust and con- trollable text to speech
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation feda58f0-07e5-433c-98eb-3c4618056172 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Fastspeech 2: Fast and high-quality 10 end-to-end text to speech
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation df84d74b-44f9-4720-93f7-a92cce382c1c · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Weiss, Mike Schuster, Navdeep Jaitly, Zongheng Yang, Zhifeng Chen, Yu Zhang, Yuxuan Wang, Rj Skerrv-Ryan, Rif A
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 040dcdfc-c036-4e61-8f75-a0c83d71980a · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Naturalspeech 2: Latent diffusion models are natural and zero-shot speech and singing synthesizers
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation be7bc2df-9edf-4a90-9cfa-a6f0aff4ea13 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Ella-v: Stable neural codec language modeling with alignment-guided sequence reordering
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6e2656ce-a6d0-493f-b076-b597206a1946 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Roformer: Enhanced transformer with rotary position embedding.Neurocomputing, 568:127063
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a6c6e768-8556-4d06-b5e7-30e87bdc5d7a · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization UTMOS: UTokyo-SaruLab System for V oiceMOS Challenge
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 28390b1a-4208-49ef-befb-ea2cd143fcc5 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Unresolved cited work
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d5fc9a45-8b02-43b6-939b-541e33fa06e6 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Spark-TTS: An Efficient LLM-Based Text-to-Speech Model with Single-Stream Decoupled Speech Tokens
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation def93211-0a50-454b-a80e-72cc5994cdcc · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Skerry-Ryan, Daisy Stanton, Yonghui Wu, Ron J
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 20a97d3e-33e2-4976-84e0-e16c9c3195ea · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization MaskGCT: Zero-shot text- to-speech with masked generative codec transformer
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation afa13a78-7126-4280-825c-a4ea4cffc492 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Con- vnext v2: Co-designing and scaling convnets with masked autoencoders
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5d465546-025b-4237-8d89-9df19158282e · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Lipvoicer: Generating speech from silent videos guided by lip reading
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation cd555d55-37e6-4687-9cce-3c7e2055ecb8 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Simultaneous mod- eling of spectrum, pitch and duration in hmm-based speech synthesis
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 63454965-20ff-4690-b961-e1a8fdbdebb8 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Weiss, Ye Jia, Zhifeng Chen, and Yonghui Wu
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 84b0e2ae-172b-42a0-a79f-bcfdf1f1c229 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4d7cd245-f93b-4952-a2d6-0d65447059cd · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization From speaker to dubber: Movie dubbing with prosody and duration consistency learning
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0899a3c2-894a-44ea-8591-f54b72650c29 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Prosody- enhanced acoustic pre-training and acoustic-disentangled prosody adapting for movie dubbing
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 271bd3ac-dcb3-4798-ad9f-81fdbb636d86 · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchro- nization
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 648d5ba8-dafb-47f4-ab5b-a6358189169e · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization Unresolved cited work
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 85d3807e-bfaa-495c-8b10-c15935c3ce5c · outbound
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization The light fusion network is implemented as a linear projection layer
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 886aca88-63d5-4480-afe1-cc6036298bfe · inbound
CoSyncDiT: Cognitive Synchronous Diffusion Transformer for Movie Dubbing DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.