Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:52:46.527011Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 3 inbound Pith citation observations for arXiv:2505.21568.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:52:46.527011Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:52:43.353115Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T12:29:52.013717Z
31 of 31 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation ea8d6b6c-5020-4ef9-b1f5-c9b285b97071 · outbound
VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e7008c2-abaf-45ae-8b13-5e44351b3257 · outbound
VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents The RVQ model disentangles speaker-specific la- tents and reconstructs watermarked audio
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7c24cade-15fa-4c13-87c6-cd4e09f82949 · outbound
VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents Implementation Details For the RVQ model, we use the pretrained SpeechTokenizer *
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ab851a6b-e778-4f76-8afd-7f0d2b9de3ca · outbound
VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents Additionally, we incorporate VC-simulated augmenta- tions and V AD-based loss to further enhance the robustness of V oiceMark
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a051ce5e-b760-4763-bb4b-cbbb5527256d · outbound
VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 560a1257-afe8-4ed8-8c14-f2c1930dff7f · outbound
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a5a68e83-f819-4cbc-be9d-869bdade2d28 · outbound
VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents De- tecting voice cloning attacks via timbre watermarking,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 14a77150-a91c-43f9-b55a-f36d57b18d1f · outbound
VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents CosyVoice: A Scalable Multilingual Zero-shot Text-to-speech Synthesizer based on Supervised Semantic Tokens
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a21cbe0-c959-4c9d-a419-20ec1b584fe1 · outbound
VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf0e8148-8995-4008-a43f-61aa484e8827 · outbound
VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents Maskgct: Zero-shot text- to-speech with masked generative codec transformer,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 649eecf1-d788-43d8-bec7-b248dfbb6a4d · outbound
VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents Techniques for data hiding,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 62a4c0e3-aa0a-4d7c-b30e-95062793a035 · outbound
VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents Natural tts synthesis by conditioning wavenet on mel spectrogram pre- dictions,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d766219-fc65-47e1-af2d-db940381b093 · outbound
VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents WavMark: Watermarking for Audio Generation
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 679ef444-a4ce-4a87-87c4-17c4928f9f78 · outbound
VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents Proactive detection of voice cloning with localized watermarking,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 46b1198a-d9e1-4d0b-a99b-3ce91b446d1d · outbound
VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents Proactive audio authentication using speaker identity watermarking,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0078746f-1c65-45b0-893f-30698b2d6876 · outbound
VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents The model is trained for 30 epochs using Adam [21] optimizer with a learning rate of5e −5
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b931714d-b772-45de-8ccb-e266c0c7ce43 · outbound
VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents Conditional variational autoencoder with adversarial learning for end-to-end text-to-speech,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 976f6a41-62af-483b-981c-12e41c5378b2 · outbound
VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents Fastspeech 2: Fast and high-quality end-to-end text to speech,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3af694f5-0558-46a4-9409-bb198c8254e4 · outbound
VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents Paddlespeech: An easy-to-use all-in-one speech toolkit,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9c6a3a09-5827-421e-93d8-7158dcba402c · outbound
VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents V oice cloning app,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a749c905-0668-4cb1-baab-5bb29dd0c58a · outbound
VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents High Fidelity Neural Audio Compression
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 002776d9-c8c4-49d0-91e6-90e661347c88 · outbound
VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents Naturalspeech 3: Zero-shot speech syn- thesis with factorized codec and diffusion models,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9b249c78-daf0-46a8-8de0-3b860e34aaf8 · outbound
VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents SpeechTokenizer: Unified Speech Tokenizer for Speech Large Language Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bbd3d8e2-cbf4-4d6a-a3ec-91ec5318cc8d · outbound
VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents Hubert: Self-supervised speech represen- tation learning by masked prediction of hidden units,
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 700e9ca6-7118-4184-809e-16431506b768 · outbound
VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents Attention is all you need,
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67507c36-7537-48e0-8f02-75cd0581a9cf · outbound
VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents Unresolved cited work
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation da792311-bb83-416b-a441-ed529c7d588f · outbound
VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents Adam: A Method for Stochastic Optimization
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 596e068b-2eae-4c61-8935-c327a1f48739 · outbound
VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents English multi-speaker corpus for cstr voice cloning toolkit,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 721db919-dc4f-4f5c-b955-c534076d8de0 · outbound
VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents Lib- rispeech: an asr corpus based on public domain audio books,
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f46d6f39-d03b-47cc-8008-0261a0161b4b · outbound
VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents Perceptual evaluation of speech quality (pesq)-a new method for speech quality assessment of telephone networks and codecs,
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9258956e-4e59-4fcd-99ad-f12a4b3213c1 · outbound
VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents A short- time objective intelligibility measure for time-frequency weighted noisy speech,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ea8d6b6c-5020-4ef9-b1f5-c9b285b97071 · inbound
VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25ff0174-6c54-417b-a1ee-d4274bf0c2a7 · inbound
The Watermark Shortcut: How Provenance Marking Sabotages Audio Deepfake Detection VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6f14e012-b715-4550-9e8b-8e76118ffb03 · inbound
Investigating Codec-Internal Latent Audio Watermarking for Neural Codec Robustness VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.