Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T22:28:08.105440Z
Paper Citation Record · LEDGER
As of 23 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 1 inbound Pith citation observation for arXiv:2412.03430.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T22:28:08.105440Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T11:47:47.534590Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T11:47:52.259298Z
52 of 52 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c2a840d3-675c-4e10-978c-5398166d3d32 · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model A wavelet neural net- work conjunction model for groundwater level forecasting
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 94c28e0e-3bef-474f-b25c-0e4e374ad584 · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Unresolved cited work
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 7fcec83d-8225-44dd-9e13-a16c8290985d · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Speech enhancement in the STFT domain
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 930cc9c0-4c95-4225-a63a-f4414b1fdbbd · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Hierarchical cross-modal talking face generation with dynamic pixel-wise loss
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 4f77072e-d387-402c-a9f8-7de2d33cacef · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model EchoMimic: Lifelike Audio-Driven Portrait Animations through Editable Landmark Conditions
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7178f086-5cd4-4ca1-b208-e55a2d258e8a · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Audio surveillance: A systematic review
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 1ce5cb17-e64a-4fe8-a0be-c328b9684fbe · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Hallo2: Long-Duration and High-Resolution Audio-Driven Portrait Image Animation
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 459fc4d1-c1c5-435b-a121-6367282cb448 · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Diffusion models beat gans on image synthesis
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4fe5288-3f80-42fd-aadd-c5739330a0d3 · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Detection of the valvular split within the second heart sound using the reassigned smoothed pseudo wigner–ville distribution
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 05ca4531-a75f-4116-accc-c58e13dfa8f2 · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Wavelet multiresolution anal- ysis based speech emotion recognition system using 1d cnn lstm networks
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 357dddd9-0c18-4656-ac17-e3046727142f · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Sound quality evaluation of vehicle suspension shock absorber rattling noise based on the wigner–ville dis- tribution
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 81e3f114-36f8-4ba4-8bb3-3128c25ed44c · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Song2face: Synthesizing singing facial animation from audio
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation cf95e60e-54d2-4e75-a6fe-80824f2a2ad1 · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Audio-driven emotional video portraits
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 156a8704-9c3a-48a1-bbfe-4640e059ccac · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Exploring spatial- temporal multi-frequency analysis for high-fidelity and temporal-consistency video prediction
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 50864aed-67f9-4c80-8a75-e2262a578770 · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model A novel deep wavelet convolutional neural network for actual ecg signal denoising
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 6d2dcbbf-9cbe-4f77-9bca-4f37db4d9b3b · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Fre-GAN: Adversarial Frequency-consistent Audio Synthesis
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9910adbd-86d3-4e2f-b66b-3aebd915277f · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Lipsync3d: Data-efficient learning of per- sonalized 3d talking faces from video using pose and light- ing normalization
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f2d40c5-3527-4a4a-bab8-7f7120c844f8 · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Classification of audio signals using statis- tical features on time and wavelet transform domains
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 2c1c5fe7-6fe9-43a6-bc07-0c7ecf2a2e6d · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model On the generalization properties of diffusion models
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 8f0945c6-00df-447c-be26-a61ba405cdd7 · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Expressive talking head generation with granular audio-visual control
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 7d1dab60-0420-45ec-80ff-56ab632b9ee5 · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Rolling bearing fault diagnosis based on stft-deep learning and sound signals
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 38beb26a-07e3-4689-87c3-f2dd236ffd04 · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Music- face: Music-driven expressive singing face synthesis
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 755cf35e-f517-4877-8c38-ae4af9dc09e7 · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model The ryer- son audio-visual database of emotional speech and song (ravdess): A dynamic, multimodal set of facial and vocal expressions in north american english
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation dfc3292b-8baa-45a6-9cfa-349acaa1fa0e · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Formant estimation of speech and singing voice by combining wavelet with lpc and cepstrum techniques
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation b35ec0c2-4ede-4bbe-87a0-f29b2d2be03c · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model A no-reference im- age blur metric based on the cumulative probability of blur detection (cpbd)
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 5ee19b88-2cc9-4b50-8131-6062252ec535 · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Wnet: Audio-guided video object segmentation via wavelet-based cross-modal denoising networks
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation b0b82e77-f6b7-4987-9d0e-425f4b958b7f · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Wavelet diffusion models are fast and scalable image generators
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation c5885b82-da4b-4d2f-b768-a7a3a2f61b82 · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Diagnostic Biomedical Sig- nal and Image Processing Applications with Deep Learning Methods
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 9ae781a0-3a25-4c37-989d-7c1bfc398f29 · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Lung sound signal denoising us- ing discrete wavelet transform and artificial neural net- work
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation b31809e1-8be4-4af5-ac00-6a22b3615882 · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model A lip sync expert is all you need for speech to lip generation in the wild
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 2b6a15b4-fe1c-4a75-b609-56de6464e588 · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Difftalk: Crafting diffusion models for generalized audio-driven portraits animation
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 55083178-73b8-4813-8eef-09bcec344ace · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Bailando: 3d dance generation by actor-critic gpt with choreographic memory
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 408969d4-33b3-49ca-849e-26d00a75d5f7 · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Everybody’s talkin’: Let me talk as you want
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41490fa0-13b5-4922-92ab-dfdfcff36f88 · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Signal reconstruc- tion from stft magnitude: A state of the art
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 3d7f366d-dafc-4bde-9298-db42b2963cb3 · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Diffused heads: Diffusion models beat gans on talking-face genera- tion
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f540ddec-4eec-4104-ae47-3538868451d7 · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Audio anal- ysis using the discrete wavelet transform
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation bc266121-c6dd-4492-a31e-9ca3ce087adb · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Fvd: A new metric for video generation
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 092cd4b9-3e9a-48f8-8523-da2de40250cf · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Seeing what you said: Talking face gen- eration guided by a lip reading expert
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e54efc39-3ba1-4a82-93b9-1890e0c57b89 · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Audio2Head: Audio-driven One-shot Talking-head Generation with Natural Head Motion
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c16f206f-8548-4ab2-b3b7-41a838c01033 · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model One- shot talking face generation from single-speaker audio-visual correlation learning
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 3ffdca56-b916-4914-94cf-d752134a3cb8 · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Image quality assessment: from error visibility to structural similarity
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 9834dd9a-38ab-4947-b585-c9801ceea50b · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85ec3a90-b9b2-48a9-b13f-d49f99fb151a · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model SingingHead: A Large-scale 4D Dataset for Singing Head Animation
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13840db1-d430-42ca-afd3-3d1c4c990361 · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Facechain-imagineid: Freely crafting high- fidelity diverse talking faces from disentangled audio
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation ae3d08e6-8f1a-4775-be60-6eb94c765fac · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ebec744e-c540-4b68-9ae7-0fd5de71305f · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model One-shot domain adaptation for face generation
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation a08fd800-91f7-4138-af18-4d53073c8b08 · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Sadtalker: Learning realistic 3d motion coefficients for stylized audio- driven single image talking face animation
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation d1d351a7-3f33-4d31-946a-2e7787663330 · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model MuseTalk: Real-Time High-Fidelity Video Dubbing via Spatio-Temporal Sampling
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27a5750a-37f7-4539-824c-c316b8679a98 · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Identity- preserving talking face generation with landmark and ap- pearance priors
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c8c2218-00d4-4a81-b6e7-b0fc918adf45 · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Pose-controllable talking face generation by implicitly modularized audio-visual rep- resentation
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1993e18-87e0-41a3-a333-e7e0a786a237 · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Subj.”, “BGM
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation f1098c32-d98c-41b5-abe0-ac1ce88a85c7 · outbound
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model Unresolved cited work
Reference 2021
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation b1dae37c-e687-4d06-9f3e-5c93c144d6d0 · inbound
Think2Sing: Orchestrating Structured Motion Subtitles for Singing-Driven 3D Head Animation SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.