Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:50:27.922828Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 1 inbound Pith citation observation for arXiv:2506.07036.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:50:27.922828Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:50:27.774312Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-07T05:50:28.138136Z
36 of 36 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation ed6d504b-ac82-4492-ba4d-8b8377e6d331 · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 618533bc-c194-4a81-b5e9-327ed0ba2f61 · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion System Overview The proposed TES-VC model is trained on purely acoustic data (Figure 1(a)), and leverages text-guided control during infer- ence (Figure 1(b))
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 48cbb941-8c1f-4de7-8308-edf7658baac7 · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion w/o CLAP-timbre adapter
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3084a535-3219-4eec-b023-9e8548298b77 · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion Our systematic data con- struction methodology facilitates disentangled learning of con- tent preservation, environmental acoustics, and speaker char- acteristics
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a0052ca4-0a48-4862-befa-be64016bb46b · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fd4b639f-80ea-4197-8b1b-4f8999bca226 · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion VQVC+: One-Shot Voice Conversion by Vector Quantization and U-Net architecture
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 03f441c5-9272-4768-8e7c-f5276eea761a · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion From speaker to dubber: movie dubbing with prosody and duration consistency learning,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation da12b11f-3751-4cda-8e18-53c8fbf4ce2c · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion Diffdub: Person- generic visual dubbing using inpainting renderer with diffusion auto-encoder,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c91b86d5-674d-4609-b686-79f060a208b5 · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion (voick): Enhancing accessibility in audiobooks through voice cloning technology,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c4650323-b8e8-40f8-a62d-b59ddb9d6241 · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion Person- alized voice command systems in multi modal user interface,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a316880e-5cf8-4628-a8d2-59f35f16ea68 · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion Triaan- vc: Triple adaptive attention normalization for any-to-any voice conversion,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 452c982a-5f58-4b0e-9ee1-2f6db5dbaf41 · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion Converting Anyone's Emotion: Towards Speaker-Independent Emotional Voice Conversion
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83d8044b-d3a9-48d1-a304-99d98d629c03 · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion One-shot voice conversion by vector quantization,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c16a1228-2172-4aa0-9ce4-e3657be4de11 · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion VQMIVC: Vector Quantization and Mutual Information-Based Unsupervised Speech Representation Disentanglement for One-shot Voice Conversion
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e26d0bd-b304-4970-84f4-65a42fbd6529 · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion Leveraging Diverse Semantic-based Audio Pretrained Models for Singing Voice Conversion
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 79c730fa-d3d4-48e3-8383-72d2ee36d6cc · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion Styletts-vc: One-shot voice conversion by knowledge transfer from style-based tts models,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b4f63a61-1881-40ca-885c-39b0ced08d03 · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion Ace- vc: Adaptive and controllable voice conversion using explicitly disentangled self-supervised speech representations,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3836c19a-1375-4dcd-960d-f1e59a7a2acf · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion Incremental Disentanglement for Environment-Aware Zero-Shot Text-to-Speech Synthesis
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b6444f9b-2cb6-4349-9002-a8499bf57673 · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion Unsupervised End-to-End Learning of Discrete Linguistic Units for Voice Conversion
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fe6e7b29-8590-4365-9355-29198724fa7e · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion HybridVC: Efficient Voice Style Conversion with Text and Audio Prompts
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2952113-d869-4710-a3a2-a83e3d0c5956 · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion Towards general-purpose text-instruction-guided voice conversion,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a5280fb0-3e23-4c36-a394-0e05b4e1901e · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion Promptvc: Flexible stylistic voice conversion in latent space driven by natural language prompts,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 96fd65e0-3d24-474c-9732-ff7c45cc54c2 · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion Environment Aware Text-to-Speech Synthesis
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c8827523-b73b-449d-aaf4-c2bc8566e13f · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion Recent advancements in speech en- hancement,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f9669ce7-e361-4c04-b8ad-9df20ab4f70c · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion Unresolved cited work
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5164b042-7d47-4b7e-a8d2-93118912647f · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion V oiceldm: Text-to- speech with environmental context,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bb6e24f0-a5e0-44b5-9c2f-8a6ba4f54f11 · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion Clap learning audio concepts from natural language supervision,
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50490b61-a75d-45cc-a193-f9ac58bdf65c · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion Learning the unlearned: Mitigating feature suppression in con- trastive learning,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation cdee7974-125e-46dd-a51f-1729d9df69c3 · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion Prompttts: Control- lable text-to-speech with text descriptions,
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95accd3d-dbe4-4b30-abe1-c446db699210 · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion LibriTTS-P: A Corpus with Speaking Style and Speaker Identity Prompts for Text-to-Speech and Style Captioning
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74adbc61-b926-4cf6-9355-eb9286bd6796 · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion GLM-4-Voice: Towards Intelligent and Human-Like End-to-End Spoken Chatbot
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1019bbbb-a944-403e-a427-703a03b06966 · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion X-vectors: Robust dnn embeddings for speaker recognition,
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3764641e-6569-495a-a220-56d32fa87e82 · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion LibriTTS-R: A Restored Multi-Speaker Text-to-Speech Corpus
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff833e3e-4076-4bd8-a8d2-3758b2d4a262 · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion gpurir: A python library for room impulse response simulation with gpu acceler- ation,
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c369a375-9f8b-41e4-95a8-0cee80373675 · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion Freevc: Towards high-quality text-free one-shot voice conversion,
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3daa2ac8-e0a9-41e8-8508-b30a51a6870b · outbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion Robust speech recognition via large-scale weak supervision,
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed6d504b-ac82-4492-ba4d-8b8377e6d331 · inbound
In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion In This Environment, As That Speaker: A Text-Driven Framework for Multi-Attribute Speech Conversion
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.