Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T10:35:09.332508Z
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 1 inbound Pith citation observation for arXiv:2412.16530.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T10:35:09.332508Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-11T10:35:09.064293Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-11T10:35:09.599954Z
38 of 38 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 48f01f24-b522-42a5-a818-78d0f2e95868 · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation However, improving lip synchrony should not compromise translation quality and natural- ness [4, 5]
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 09b13862-6ead-4263-94e1-2b23d9679301 · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 1cc8f621-11e5-4b54-b384-6f20809688a0 · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation length pre- dictor
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation bd53b506-d9e8-459e-ad3c-e1cdb2554c6c · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation Dataset We leverage LRS3 [28] which is a large-scale video data consisting of thousands of spoken sentences collected from TED talks
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 0e87283f-f2f5-4279-a515-02381e8255ba · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation Baselines We use the latest work of A V2A V [6] as a strong baseline for this work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 8157bb37-63b4-4cf8-8d26-e21d8a365cc2 · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation not trading off lip-synchrony improvements over speech translation quality and naturalness
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation b3da65a4-4462-492d-972e-228dc9eb1f3f · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation Our A VS2S framework incorporates lip-synchrony and duration loss to enhance the alignment between speech and lip movements in audio-visual translation models
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 5b4d7f32-4bc9-47d3-ac65-249b21f5439f · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation Unresolved cited work
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation c2e5ce6e-1e54-4ae0-85fa-ba9292f71906 · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation Neural dubber: Dubbing for videos according to scripts,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation d87a7425-51f0-442f-b41f-85034527586e · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation Neural style-preserving visual dubbing,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 7bcbdedd-8edd-4c30-980d-4f51a097cae4 · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation An empirical take on the dubbing vs. subtitling debate: An eye movement study,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation db58254e-3b8d-41e7-af47-0ea539b880bb · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation Dub- bing in practice: A large scale study of human localization with insights for automatic dubbing,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation ae9b1e37-533a-463e-8761-e5a042b3c78b · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation Av2av: Direct audio-visual speech to audio-visual speech translation with unified audio-visual speech representation,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 5fc1aa79-cb70-4753-9aa0-ebeb9854336c · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation A lip sync expert is all you need for speech to lip generation in the wild,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation b30dd3ac-d018-4241-95a4-fc97368a5577 · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation Exposing lip-syncing deepfakes from mouth inconsistencies,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 4efb3245-a2e9-4dbc-b3bd-501912ea1623 · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation The DeepFake Detection Challenge (DFDC) Dataset
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d2f24ab-3b92-4199-98df-c3cefd8febbc · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation Faces of the future: How generative ai is redefining likeness and identity in the age of artificial intelligence,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 597d37e1-ed56-4bad-b3ce-5a10760bc06c · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation Face/off: Changing the face of movies with deepfakes,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 5ca9d8c6-0665-4be9-bb34-cad046ac5217 · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation Regu- lating deep fakes: legal and ethical considerations,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation e5449a1a-143f-490d-8929-bfc5e69fd43a · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation VASA-1: Lifelike Audio-Driven Talking Faces Generated in Real Time
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 968d1cab-5d56-4bec-bd12-ca9ff9c6e8d7 · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation AV-TranSpeech: Audio-Visual Robust Speech-to-Speech Translation
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 2bcf0595-4ea8-48f0-b2a0-60dea3d0c6bf · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation Mixspeech: Cross-modality self-learning with audio- visual stream mixup for visual speech translation and recogni- tion,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 7856dcaa-b336-4025-9c3f-d5ab29dd9ded · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation Isometric MT: Neural Machine Translation for Automatic Dubbing
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation a2b0c858-646d-44f2-85e3-a560cbd37672 · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation Duration modeling of neural tts for au- tomatic dubbing,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 5a59b774-35e0-47c9-9aef-d94b44097f21 · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation Prosodic alignment for off-screen automatic dubbing,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 8f9d8b29-748a-4cc4-9cd1-16e032190486 · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation Improving isochronous machine translation with target factors and auxil- iary counters,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 2d7bcfbd-78ec-4504-ad66-1ba22fd1f0b5 · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation Jointly optimizing translations and speech timing to improve isochrony in automatic dubbing,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 45d4c327-5039-489d-bbee-273c17c64754 · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation A lip sync expert is all you need for speech to lip generation in the wild,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 2f4f0e13-2e69-486b-b98c-53a90407e7b7 · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation Towards realistic visual dubbing with heterogeneous sources,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 10b682fe-da9f-4fb1-8504-5b4ebe74edad · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation Learning audio-visual speech representation by masked multimodal cluster prediction,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation d2425594-730b-45eb-bff1-6d52f5006d3c · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation Textless Unit-to-Unit training for Many-to-Many Multilingual Speech-to-Speech Translation
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c5933dd-6f8f-49d9-8433-a7b144d0d95d · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation Hifi-gan: Generative adversarial networks for efficient and high fidelity speech synthesis,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 2c217f1f-2490-4f0c-aaac-ef9188e734cd · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation Direct speech-to-speech translation with discrete units
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8162e89-b6ec-434d-870d-2ca6ec764b82 · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation Out of time: automated lip sync in the wild,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation c1020e8c-7ec4-4652-9238-bbd7e7c3dfbd · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation Deep audio-visual speech recognition,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 15901fcc-e929-4b2a-8533-f827d0fe3a59 · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation Seamlessm4t: Massively multilin- gual and multimodal machine translation,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 03b8eb3a-bd75-4700-9285-ad17a4e449cc · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation Bleu: a method for automatic evaluation of machine translation,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation c974c953-a91b-400b-86cd-7f22352d87eb · outbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation Decoupled weight decay regularization,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 09b13862-6ead-4263-94e1-2b23d9679301 · inbound
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.