Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:52:42.261811Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 0 inbound Pith citation observations for arXiv:2505.23406.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:52:42.261811Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
38 of 38 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation f1322c27-8c34-49bc-a3e2-d2277dc5e5a7 · outbound
Video Editing for Audio-Visual Dubbing What is the McGurk effect?Frontiers in psychology, 2014
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ed0023cb-ac21-4f26-ac0e-4f889831d3f4 · outbound
Video Editing for Audio-Visual Dubbing Denoising diffusion probabilistic models.Advances in neural information processing systems, 33:6840–6851, 2020
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4a2da30-0c30-432d-8fa3-26ccf99ee105 · outbound
Video Editing for Audio-Visual Dubbing FlowVQTalker: High-quality emotional talking face generation through normalizing flow and quantization
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 63cc3d09-49a8-4724-a0a7-1ade0748da22 · outbound
Video Editing for Audio-Visual Dubbing Diffused heads: Diffusion models beat GANs on talking-face generation
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4ac0ba71-6774-4098-9cb0-0719da40383c · outbound
Video Editing for Audio-Visual Dubbing EmoTalker: Emotionally editable talking face generation via diffusion model
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation debc9947-c133-4e07-925c-1c298f9f9325 · outbound
Video Editing for Audio-Visual Dubbing LatentSync: Taming Audio-Conditioned Latent Diffusion Models for Lip Sync with SyncNet Supervision
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2341ac32-84ad-4c94-8355-00b4dc29f8aa · outbound
Video Editing for Audio-Visual Dubbing A lip sync expert is all you need for speech to lip generation in the wild
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62472101-44d6-426d-8f8d-a4f920a5d12c · outbound
Video Editing for Audio-Visual Dubbing DiffDub: Person-generic visual dubbing using inpainting renderer with diffusion auto-encoder
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4ff86e3d-b9b7-4194-ac1b-24242dff3043 · outbound
Video Editing for Audio-Visual Dubbing Denoising Diffusion Implicit Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f4b7762-0bf8-4c58-ba33-bd1fda266a2a · outbound
Video Editing for Audio-Visual Dubbing Hubert: Self-supervised speech representation learning by masked prediction of hidden units.IEEE/ACM transactions on audio, speech, and language processing, 29:3451–3460, 2021
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 360c1290-b549-46de-9b16-2b5e828aab47 · outbound
Video Editing for Audio-Visual Dubbing Video diffusion models.Advances in Neural Information Processing Systems, 35:8633–8646, 2022
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f350542-197e-46a9-9c98-31c775b8c082 · outbound
Video Editing for Audio-Visual Dubbing Cascaded diffusion models for high fidelity image generation.Journal of Machine Learning Research, 23(47):1–33, 2022
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32a4212f-fed6-45b3-b3a0-678263211f86 · outbound
Video Editing for Audio-Visual Dubbing VoxCeleb2: Deep Speaker Recognition
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b42c7bd-5f60-4a3c-87e9-8465f7524021 · outbound
Video Editing for Audio-Visual Dubbing Stylesync: High-fidelity generalized and personalized lip sync in style-based generator
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a6f6325-ce77-4944-84fe-8795c8e8529e · outbound
Video Editing for Audio-Visual Dubbing Analyzing and improving the image quality of stylegan
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 717b0f2f-86bf-4bbe-b369-162ac42e22fc · outbound
Video Editing for Audio-Visual Dubbing Diff2lip: Audio conditioned diffusion models for lip-synchronization
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bdffe128-862e-4866-83fb-2843c342e7df · outbound
Video Editing for Audio-Visual Dubbing High- resolution image synthesis with latent diffusion models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e84d701-372d-474c-bd0e-be6047c1e44a · outbound
Video Editing for Audio-Visual Dubbing The LJ speech dataset
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation db5ab34d-fcdc-4e82-9ce2-490e37096c24 · outbound
Video Editing for Audio-Visual Dubbing Revise: Self-supervised speech resynthesis with visual input for universal and generalized speech regeneration
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d80aae8-7e19-4490-b57a-40d50b5a339c · outbound
Video Editing for Audio-Visual Dubbing Arbitrary style transfer in real-time with adaptive instance normalization
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b25e0712-8975-4adc-9e60-0ef52efd1ae3 · outbound
Video Editing for Audio-Visual Dubbing Classifier-Free Diffusion Guidance
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 874606e3-35a3-47dc-b410-a0c65bab477c · outbound
Video Editing for Audio-Visual Dubbing Multidiffusion: Fusing diffusion paths for controlled image generation
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ee029bc-1821-43a2-ac95-1c28bb61c2e4 · outbound
Video Editing for Audio-Visual Dubbing Lip reading sentences in the wild
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 731d8cc2-0dde-40e9-b0b1-095232f62973 · outbound
Video Editing for Audio-Visual Dubbing LRS3-TED: a large-scale dataset for visual speech recognition
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58b33779-ec12-4693-b66e-9b28acc92859 · outbound
Video Editing for Audio-Visual Dubbing Flow-guided one-shot talking face generation with a high-resolution audio-visual dataset
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1abfd6bf-0f26-4ee9-91b5-d4c1a43bcf23 · outbound
Video Editing for Audio-Visual Dubbing ModeFormer: Modality-preserving embed- ding for audio-video synchronization using transformers
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 311cda32-899a-4641-84e7-9799e5818fac · outbound
Video Editing for Audio-Visual Dubbing Learning transferable visual models from natural language supervision
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f242f5d-5fd3-4f4c-84e0-5c4e9edb49d1 · outbound
Video Editing for Audio-Visual Dubbing Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 92283cca-4813-4ac7-89c6-c4da7cc42a43 · outbound
Video Editing for Audio-Visual Dubbing facenet-pytorch.https://github.com/timesler/facenet-pytorch, 2022
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation deca0c92-4cf2-45b3-a204-adf4f34899d0 · outbound
Video Editing for Audio-Visual Dubbing Decoupled Weight Decay Regularization
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf48b885-2584-432c-a059-a91b939fceba · outbound
Video Editing for Audio-Visual Dubbing Wilcoxon signed-rank test.Encyclopedia of biostatistics, 8, 2005
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 75cd4042-aefe-4483-a308-bbc18b1a57fb · outbound
Video Editing for Audio-Visual Dubbing Diffusion models beat gans on image synthesis
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0549e04d-a162-4e05-9cd2-b88efe32227b · outbound
Video Editing for Audio-Visual Dubbing guided-diffusion.https://github.com/openai/guided-diffusion, 2021
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 98b48bc8-d009-440e-9268-119adc32404b · outbound
Video Editing for Audio-Visual Dubbing Improved denoising diffusion probabilistic models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a5ee8b1-41f4-493e-a827-2337c89698a3 · outbound
Video Editing for Audio-Visual Dubbing ns" represents non-significant results (p > 0.05),
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d311c2f6-5179-49aa-9ab4-95197024e9fa · outbound
Video Editing for Audio-Visual Dubbing We used a constant crop to ensure the corrupted video would not inadvertently synchronize with the audio
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0610c022-b4a9-4a17-bb58-392008b77913 · outbound
Video Editing for Audio-Visual Dubbing Unresolved cited work
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 16d05ea2-2420-457f-8298-3e4e46f44ab5 · outbound
Video Editing for Audio-Visual Dubbing facebook/hubert-large-ll60k
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
No inbound Pith citation observations are available.