Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:56:29.213609Z
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 1 inbound Pith citation observation for arXiv:2506.01020.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:56:29.213609Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T05:17:06.334514Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-06T05:17:07.028665Z
44 of 44 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c0158a4b-d980-41a5-91d1-22158098bec0 · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Fgp-gan: Fine- grained perception integrated generative adversarial network for expres- sive mandarin singing voice synthesis,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 012f4664-7b3e-46d6-82b4-8a58d7abfdee · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Fastspeech 2: Fast and high-quality end-to-end text to speech,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 01affd88-9573-44b8-b6be-f81a9d7be7a4 · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Multilingual speech-to-speech translation system for mobile consumer devices,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 2768803a-2e0b-43fd-bc37-077c1401840e · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Multi-speaker and multi-dialectal catalan tts models for video gaming,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 49379baa-ebad-4c6f-b6e1-b861fcc75433 · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Neural voice cloning with a few samples,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 54695d52-1902-4b97-ae2b-d9705a364e98 · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation AdaSpeech: Adaptive Text to Speech for Custom Voice
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59ca98f9-385f-4dcf-a332-4bd389d5b627 · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Rapid speaker adaptation in low resource text to speech systems using synthetic data and transfer learning,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 8468afde-e559-4552-9826-2348ea795914 · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Quantum target recognition enhancement algorithm for uav consumer applications,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 38df971a-f665-4f11-899b-126dece443b1 · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Dual channel based speech enhancement using novelty filter for robust speech recognition in automobile environ- ment,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 2fa65a51-a547-4c8f-bf03-7e1e6a88388b · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Transfer learning from speaker verification to multispeaker text-to-speech synthesis,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f58c910-faa2-4914-877a-f92329873ac3 · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation In- vestigating on incorporating pretrained and learnable speaker represen- tations for multi-speaker multi-style text-to-speech,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b838cdba-a696-4f64-a935-3a47f3e38f7b · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Mrmi-tts: Multi-reference audios and mutual information driven zero-shot voice cloning,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 659de971-e51a-43af-b5de-7068704a42dd · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Meta-stylespeech: Multi- speaker adaptive text-to-speech generation,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7333ae0-e69d-42ea-8692-bb4f34e4d6db · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Acfusion: Infrared and visible image fusion based on self-attention and convolution with enhanced information extraction,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation bfc1ebe1-7ada-408f-8c20-8b558ec68dc0 · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Multi-feature fusion-based convolutional neural networks for eeg epileptic seizure prediction in con- sumer internet of things,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 803e9747-1526-438f-a9f4-77f2c1cefae2 · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Styletts 2: Towards human-level text-to-speech through style diffusion and adversarial training with large speech language models,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 4b222e0b-a11f-495d-a694-d51626e6e81c · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Generspeech: Towards style transfer for generalizable out-of-domain text-to-speech,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 3277402f-9674-4962-9bc9-c028d761e197 · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation OpenVoice: Versatile Instant Voice Cloning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6601101-719e-4fd0-bb82-044aab76acee · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Towards end-to-end prosody transfer for expressive speech synthesis with tacotron,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation e2c2095d-4b6f-4572-9596-c8f1568d129e · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Arbitrary style transfer in real-time with adaptive instance normalization,
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1dbafcf-29d1-40c3-9372-de7635042253 · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation One-shot voice conversion by separating speaker and content representations with instance normalization,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 5cb7ec09-c5ca-49ed-b241-1ab8528c266e · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Yourtts: Towards zero-shot multi-speaker tts and zero-shot voice conversion for everyone,
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0dbcb50-1cfc-4f74-a214-b24f2add4081 · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Film: Visual reasoning with a general conditioning layer,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 60af8530-33fc-4522-a7ec-bf67ae9d6ee0 · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Dynamic neural networks: A survey,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 6a1dd9d0-1ef8-4662-9e56-43e48451c3aa · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Imagenet classification with deep convolutional neural networks,
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d30c6895-2136-41a5-8890-7a0ff3cc0d63 · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Very Deep Convolutional Networks for Large-Scale Image Recognition
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3135193-0559-453a-8050-9f402ef5ddee · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Going deeper with convolutions,
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40bd5bdf-f2ac-4109-9019-45ff90c43ca0 · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Bert: Pre-training of deep bidirectional transformers for language understanding,
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50369e01-2a93-42d3-a6ab-1f923d4c781b · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Language mod- els are few-shot learners,
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 074e8646-13bf-4a84-a571-c18d9456c3a8 · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Skipnet: Learning dynamic routing in convolutional networks,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 7d68e7f9-a106-4dea-918f-514c0e860226 · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Multi-scale dense networks for resource efficient image classification,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation d0fa8c20-3024-4eab-b738-0f81546160a0 · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Reso- lution adaptive networks for efficient inference,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 49e93ffd-1a53-41b8-86cd-7bcbe28befa5 · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Any-precision deep neural networks,
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52513412-4504-4517-b367-7580bd9b2935 · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Big/little deep neural network for ultra low power inference,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 3018709a-42d2-4868-b453-ac0de11c140c · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation A convolutional neural network cascade for face detection,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b3b7e980-8f8e-43ea-8e91-b6a2364c3e8f · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Changing Model Behavior at Test-Time Using Reinforcement Learning
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 1c21de8c-1bb8-47d3-96cc-bad795bef5f9 · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Dynamic deep neural networks: Optimizing accuracy-efficiency trade-offs by selective execution,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ea821dfb-05bd-4e37-ba10-fe69715eea86 · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Speech emotion recognition with co-attention based multi-level acoustic information,
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f851d4cb-ec79-40c1-abe9-fcbb5b0f818e · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Task-adaptive neural process for user cold-start recommendation,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 0830f5e2-6acb-47d6-9d10-d3866d1d8747 · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation LibriTTS: A Corpus Derived from LibriSpeech for Text-to-Speech
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 89eac614-048e-425e-a3e7-8f120e1986cf · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Melgan: Generative adversarial networks for conditional waveform synthesis,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 9a7b921a-32ce-4f59-9765-7ed4dbf538f6 · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d902b6a-eab8-4e7b-bdbf-ab210c7bf1d1 · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 571a9844-a593-4402-ac6d-3aebec911138 · outbound
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation XTTS: a Massively Multilingual Zero-Shot Text-to-Speech Model
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39a25de5-0fe5-4e27-b05d-6ca480cb2e41 · inbound
Marco-Voice Technical Report DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.