Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T00:27:32.622659Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 25 of 25 outbound references and 1 inbound Pith citation observation for arXiv:2511.01056.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T00:27:32.622659Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-04T00:27:30.285078Z
A source-named dated measurement, never combined with another source.
Source: cited_works
25 of 25 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 2033c47d-d8b4-407a-bf84-b43e0798d82e · outbound
WhisperVC: Decoupled Cross-Domain Alignment and Speech Generation for Low-Resource Whisper-to-Normal Conversion WhisperVC: Decoupled Cross-Domain Alignment and Speech Generation for Low-Resource Whisper-to-Normal Conversion
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6327563c-8ea8-4817-9ce9-72857de7159c · outbound
WhisperVC: Decoupled Cross-Domain Alignment and Speech Generation for Low-Resource Whisper-to-Normal Conversion Overview The proposed whisper-to-speech (W2S) framework comprises three stages, as illustrated in Fig
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4bb93aa6-22ec-457e-958d-9c59077066d8 · outbound
WhisperVC: Decoupled Cross-Domain Alignment and Speech Generation for Low-Resource Whisper-to-Normal Conversion Unresolved cited work
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83ea74e0-dfff-4670-89f7-f5e78a4d8f60 · outbound
WhisperVC: Decoupled Cross-Domain Alignment and Speech Generation for Low-Resource Whisper-to-Normal Conversion Objective evaluations show consistent gains over whispered inputs and performance approaching that of ground-truth recordings in terms of naturalness (DNSMOS 3.11, UTMOS 2.52vs
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93ae1124-b765-4d9a-b87f-b3b786023bad · outbound
WhisperVC: Decoupled Cross-Domain Alignment and Speech Generation for Low-Resource Whisper-to-Normal Conversion Attention-Guided Generative Adversarial Network for Whisper to Normal Speech Conversion
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 030bb5cc-66fb-4cd9-bbca-eee124b96177 · outbound
WhisperVC: Decoupled Cross-Domain Alignment and Speech Generation for Low-Resource Whisper-to-Normal Conversion A novel attention-guided generative ad- versarial network for whisper-to-normal speech conversion,
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a6215b3-b46f-4579-b6ce-c1914bed3cea · outbound
WhisperVC: Decoupled Cross-Domain Alignment and Speech Generation for Low-Resource Whisper-to-Normal Conversion End-to-End Whisper to Natural Speech Conversion using Modified Transformer Network
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2660c0f7-146d-46da-8c7c-55de5fa4a870 · outbound
WhisperVC: Decoupled Cross-Domain Alignment and Speech Generation for Low-Resource Whisper-to-Normal Conversion Gener- ative adversarial networks for whispered to voiced speech con- version: a comparative study,
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee10939d-fc61-4a92-bc9e-324173ed39a2 · outbound
WhisperVC: Decoupled Cross-Domain Alignment and Speech Generation for Low-Resource Whisper-to-Normal Conversion Maskcyclegan-based whisper to normal speech conversion,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e08590eb-faab-4165-879b-1d715e8a37cb · outbound
WhisperVC: Decoupled Cross-Domain Alignment and Speech Generation for Low-Resource Whisper-to-Normal Conversion V ocoder-free non-parallel conversion of whispered speech with masked cycle-consistent generative adversarial net- works,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd4b45c9-b9b1-4ea9-a473-2fb6394a360c · outbound
WhisperVC: Decoupled Cross-Domain Alignment and Speech Generation for Low-Resource Whisper-to-Normal Conversion Wesper: Zero-shot and realtime whisper to normal voice conversion for whisper-based speech interac- tions,
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01832f1e-6f80-4f34-b96a-89feb01029e7 · outbound
WhisperVC: Decoupled Cross-Domain Alignment and Speech Generation for Low-Resource Whisper-to-Normal Conversion Distillw2n: A lightweight one-shot whisper to normal voice conversion model using distillation of self- supervised features,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a2a0a42-ddca-4696-8201-fbc699ef57aa · outbound
WhisperVC: Decoupled Cross-Domain Alignment and Speech Generation for Low-Resource Whisper-to-Normal Conversion Improvement Speaker Similarity for Zero-Shot Any-to-Any Voice Conversion of Whispered and Regular Speech
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 894eef0e-9f7d-48dd-887c-61da2d95b4e9 · outbound
WhisperVC: Decoupled Cross-Domain Alignment and Speech Generation for Low-Resource Whisper-to-Normal Conversion Whis- pered speech conversion based on the inversion of mel fre- quency cepstral coefficient features,
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31dae4b7-00d6-4f15-b7a9-e23d9e3303e6 · outbound
WhisperVC: Decoupled Cross-Domain Alignment and Speech Generation for Low-Resource Whisper-to-Normal Conversion Glottal flow synthesis for whisper-to-speech conversion,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27428f30-a187-4aef-8e21-e3c211174b1b · outbound
WhisperVC: Decoupled Cross-Domain Alignment and Speech Generation for Low-Resource Whisper-to-Normal Conversion Robust speech recognition via large-scale weak supervision,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 467a1e4a-1c9b-4d7a-86b9-ec7f070fe0ce · outbound
WhisperVC: Decoupled Cross-Domain Alignment and Speech Generation for Low-Resource Whisper-to-Normal Conversion Aishell6-whisper: A chinese mandarin audio-visual whisper speech dataset with speech recognition baselines,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1aa43ace-e3cb-442d-840a-997ac9e62c34 · outbound
WhisperVC: Decoupled Cross-Domain Alignment and Speech Generation for Low-Resource Whisper-to-Normal Conversion Soft-dtw: a differentiable loss function for time-series,
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7dac97be-17ba-447e-ad96-257d4e781f18 · outbound
WhisperVC: Decoupled Cross-Domain Alignment and Speech Generation for Low-Resource Whisper-to-Normal Conversion FastSpeech 2: Fast and High-Quality End-to-End Text to Speech
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d8a2234-0ca5-4a03-b47e-b1c4c88de3c2 · outbound
WhisperVC: Decoupled Cross-Domain Alignment and Speech Generation for Low-Resource Whisper-to-Normal Conversion Wespeaker: A research and production oriented speaker embedding learning toolkit,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ea14fc2-563a-4b30-9675-108021d84dff · outbound
WhisperVC: Decoupled Cross-Domain Alignment and Speech Generation for Low-Resource Whisper-to-Normal Conversion VoxBlink2: A 100K+ Speaker Recognition Corpus and the Open-Set Speaker-Identification Benchmark
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3945a79a-ecfe-4eac-8e1c-4e45dd8250e5 · outbound
WhisperVC: Decoupled Cross-Domain Alignment and Speech Generation for Low-Resource Whisper-to-Normal Conversion VoxCeleb2: Deep Speaker Recognition
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de3b0585-e9fd-4f91-b6fe-3b758095457c · outbound
WhisperVC: Decoupled Cross-Domain Alignment and Speech Generation for Low-Resource Whisper-to-Normal Conversion Hifi-gan: Generative adversarial networks for efficient and high fidelity speech synthesis,
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5cbee1b4-df98-485e-891e-deffc5ba766a · outbound
WhisperVC: Decoupled Cross-Domain Alignment and Speech Generation for Low-Resource Whisper-to-Normal Conversion Dnsmos: A non-intrusive perceptual objective speech quality metric to evaluate noise suppressors,
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a38ff9c-55ca-4e97-9ee8-0db8bbd969e2 · outbound
WhisperVC: Decoupled Cross-Domain Alignment and Speech Generation for Low-Resource Whisper-to-Normal Conversion UTMOS: UTokyo-SaruLab System for VoiceMOS Challenge 2022
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2033c47d-d8b4-407a-bf84-b43e0798d82e · inbound
WhisperVC: Decoupled Cross-Domain Alignment and Speech Generation for Low-Resource Whisper-to-Normal Conversion WhisperVC: Decoupled Cross-Domain Alignment and Speech Generation for Low-Resource Whisper-to-Normal Conversion
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.