Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:28:41.644243Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 2 inbound Pith citation observations for arXiv:2506.02414.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:28:41.644243Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:28:37.819963Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-11T11:26:01.472807Z
42 of 42 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 2df8c3c7-e932-4355-a298-55ec9b527926 · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0b4d78d-0dda-44f6-a87d-c6650ede0036 · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion System Architecture Our proposed StarVC framework is designed to jointly model speech conversion and text generation in an auto-regressive manner
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation dc1e7153-c55d-4ca3-902b-70bfe306697e · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation be2e4822-4938-407a-b34d-dcbfc2997b5b · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 97e42094-fa92-427b-9c45-7ef4b090d700 · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion Unresolved cited work
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0cd0cc7e-cfb5-4cac-8002-73ea19f1c9e0 · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion wav2vec: Unsupervised Pre-training for Speech Recognition
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 083bfd3a-15f4-40df-877c-9ff9bc513108 · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion Continuous probabilis- tic transform for voice conversion,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation cd16316b-df7e-472d-9fc1-abb2b2823e3f · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion Reimagining Speech: A Scoping Review of Deep Learning-Powered Voice Conversion
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fcd138b0-e68e-4169-8bd6-f62d62ee0058 · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion An overview of voice conversion and its challenges: From statistical modeling to deep learning,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 195e3a43-85d4-472c-a547-3a17620df70f · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion Wavlm: Large-scale self- supervised pre-training for full stack speech processing,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18398b93-4b38-4768-a86a-7eac87ab7345 · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion Hubert: Self-supervised speech represen- tation learning by masked prediction of hidden units,
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea4a4626-d600-476b-be27-4031ff45383f · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion OpenVoice: Versatile Instant Voice Cloning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 046b8f8f-619e-4b20-9a7f-093be92c6539 · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion wav2vec 2.0: A framework for self-supervised learning of speech repre- sentations,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54b065c9-800b-4167-9be1-c71ec560cc6a · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion Neural codec language mod- els for disentangled and textless voice conversion,
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ddda902-e377-42e3-8b8b-28ce480d2a1e · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion Sef-vc: Speaker embedding free zero-shot voice conversion with cross attention,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b8bb26c0-d798-4ad7-830a-f0d1443dacdc · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion Diffusion-Based Voice Conversion with Fast Maximum Likelihood Sampling Scheme
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b16057f-5c85-4021-8d28-792b3fb0e1d0 · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion StableVC: Style Controllable Zero-Shot Voice Conversion with Conditional Flow Matching
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08679d1f-dc81-4a25-aba9-1d489ffcb276 · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion Robust speech recognition via large-scale weak supervision,
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd1cb69f-cfcf-4442-9706-a202b435bc43 · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion Lm-vc: Zero- shot voice conversion via speech generation based on language models,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 36b70be6-0df8-4d65-bf04-5b5a7e66200c · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion DualVC 3: Leveraging Language Model Generated Pseudo Context for End-to-end Low Latency Streaming Voice Conversion
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation af841cf4-fb46-4464-a0d0-427ca241a113 · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion Moshi: a speech-text foundation model for real-time dialogue
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43b79292-ecd0-4ef7-8638-a3368fa496ff · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion Mini-Omni: Language Models Can Hear, Talk While Thinking in Streaming
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a38ced2f-e6bf-428d-9ab4-bc652ad434f4 · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion LLaMA-Omni: Seamless Speech Interaction with Large Language Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4cc7b23-095b-4950-b013-3e40c7124f70 · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion SNAC: Multi-Scale Neural Audio Codec
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 076302f5-652f-4c82-81b6-a0d019d2fe83 · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion An Enhanced Res2Net with Local and Global Feature Fusion for Speaker Verification
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8966055c-0600-48f5-842f-1b7441585887 · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion Simple and controllable music gen- eration,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5a2f2a56-bbbb-481e-ae2d-1f7952068cde · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion A learning algorithm for contin- ually running fully recurrent neural networks,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7ffb179f-d376-4940-a5e0-4b2a70c8daa0 · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion Roformer: Enhanced transformer with rotary position embedding,
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3af7947a-e4ec-4113-bfab-1b49e1d85185 · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion High Fidelity Neural Audio Compression
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d21c2aa6-c7f5-4489-9a50-e5509b91f13c · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion GigaSpeech covers diverse domains and acoustic conditions, while LibriTTS is cleaner and more consistent
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 79e7c961-e095-421a-9902-fe11c8cc0f37 · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion High-fidelity audio compression with improved rvqgan,
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b25b703-7d5c-4ef4-82f1-5e2b0b7e8614 · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion SemantiCodec: An Ultra Low Bitrate Semantic Audio Codec for General Sound
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea9bd571-082f-46ac-a8d4-4cb9ca20937e · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion Attention is all you need,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7c0f7bc1-a43d-481f-b55b-e24c4e539021 · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion Qwen Technical Report
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc33c64e-9538-4134-81bf-c2465e948553 · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion Qwen2.5 Technical Report
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d25116b-720e-4bed-b17b-d3100c804aa7 · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion GigaSpeech: An Evolving, Multi-domain ASR Corpus with 10,000 Hours of Transcribed Audio
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc77ec91-1540-4805-b166-5340f8844ec1 · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion LibriTTS: A Corpus Derived from LibriSpeech for Text-to-Speech
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a88d34e-8cf9-4b50-9f13-e91273c79b3e · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion Lib- rispeech: an asr corpus based on public domain audio books,
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be796505-843a-434d-988c-72a53d0e1f13 · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion CosyVoice: A Scalable Multilingual Zero-shot Text-to-speech Synthesizer based on Supervised Semantic Tokens
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f54052a-3d2c-46cd-a514-9554e24980f9 · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 148c0d86-6b60-40c9-842f-178b7be97ab7 · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion Yourtts: Towards zero-shot multi-speaker tts and zero-shot voice conversion for everyone,
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c2b4fcd-d666-4124-ba7a-1ca69d74738e · outbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion Triaan- vc: Triple adaptive attention normalization for any-to-any voice conversion,
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2df8c3c7-e932-4355-a298-55ec9b527926 · inbound
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cbaacff1-6a1a-4a1e-872c-bb26dfe00aea · inbound
MimicLM: Zero-Shot Voice Imitation through Autoregressive Modeling of Pseudo-Parallel Speech Corpora StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.