Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:49:13.241404Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 1 inbound Pith citation observation for arXiv:2505.20868.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:49:13.241404Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:49:08.825623Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-07T13:49:13.486015Z
37 of 37 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 2dfd6543-a213-4651-9f4d-04afb3ab856c · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech With recent advancements in deep learning technology [2, 3, 4], the naturalness of synthesized speech has improved significantly [5, 6]
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 37201e95-7427-4e53-a70f-ca8993b94ef7 · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4fb189b0-c7a9-45e0-9980-15690766f006 · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech Both are about the same distance
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation aeec44c5-4f37-4aae-8896-1e08c45796df · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech V oiced-aware style extraction considers the acoustic character- istics of different speech regions, enabling more detailed style extraction
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d96f3b0e-2712-4887-bdc8-68f6554cd658 · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech RS-2019-II190079), Artificial Intelligence Innova- tion Hub (No
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f5a2dd12-6165-4ae2-916e-8e79ce511c84 · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech Emosphere-tts: Emotional style and intensity modeling via spherical emotion vector for controllable emotional text-to- speech,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b5132ec1-c45e-40d4-a806-0af870f8d5b8 · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech Fastspeech 2: Fast and high-quality end-to-end text to speech,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29571275-8d64-43b2-9ff7-6b1c4801f5b2 · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech A new recurrent neural-network ar- chitecture for visual pattern recognition,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c62140b5-f2ed-4f10-a2e9-d7a13ac36582 · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech Multiresolution recognition of hand- written numerals with wavelet transform and multilayer cluster neural network,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 23c905f2-7275-49bc-b83c-26f720a333ff · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech Multilayer cluster neural network for totally un- constrained handwritten numeral recognition,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 492cd53c-4ab1-4a74-86e3-26f2be4dd38f · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech Hierspeech: Bridging the gap between text and speech by hierarchical variational inference using self-supervised represen- tations for speech synthesis,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a48a850a-8e74-4922-823f-2a7263d62987 · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech Neural dis- crete representation learning,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c48ee2b7-9b0b-4e3e-b5fd-9096f83f40b2 · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech Towards end-to-end prosody transfer for expressive speech synthesis with tacotron,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8448b2bf-8a24-430f-b812-f41398451042 · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech Style tokens: Unsu- pervised style modeling, control and transfer in end-to-end speech synthesis,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2e288fe4-32ca-4ff3-8ce1-20e4398137b7 · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech Meta-stylespeech : Multi-speaker adaptive text-to-speech generation,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 87bad46a-6ff1-4fcc-8ed0-00d8ecd8128b · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech Qi-tts: Questioning intonation control for emotional speech synthesis,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9c21da54-39e4-42aa-9ac8-76ae8847ed08 · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech Generspeech: Towards style transfer for generalizable out-of-domain text-to- speech,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 650cbabd-599a-43b2-9269-0a03b82b5f8d · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech Good helper is around you: Attention- driven masked image modeling,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f2af16f3-3e05-4ec7-84f7-77d67aa20ab5 · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech Furthermore, we introduce style direction adjustment, which adjusts the ex- tracted style by modifying its angle using content and prosody vectors in the embedding space
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 074771a2-cf4f-4efc-8812-f55a96c6d12b · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech Tsp-tts: Text-based style predictor with residual vector quantization for expressive text-to- speech,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0ea39445-2923-4323-b567-fd8a08025c1b · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech TCSinger: Zero-shot singing voice synthesis with style transfer and multi-level style control,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation dab69dc1-f56f-4ad0-8981-43d15e4c9865 · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech Estimating or Propagating Gradients Through Stochastic Neurons for Conditional Computation
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d9de12d-4f7d-41d8-b735-c0c6fe04e43b · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech Signal compression based on models of human perception,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0240fbeb-c308-44f8-b170-4b17c87a7776 · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech Not all image re- gions matter: Masked vector quantization for autoregressive im- age generation,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 97400640-9587-4f5a-a7f1-440d9a240ea3 · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech Restructuring vector quantization with the rotation trick,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 99691b28-48f4-4916-9eb4-a0ca2170a121 · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech Autoregres- sive image generation using residual quantization,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1fe9ced6-5848-4c35-96ab-7802f53d7094 · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech A convnet for the 2020s,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c1aa1b80-3cd3-4629-a678-65b6d0716cff · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech Cross-speaker emotion disentangling and transfer for end-to-end speech synthe- sis,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7a4cec8a-bac5-435d-a068-c5191f12d452 · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech Bigvgan: A universal neural vocoder with large-scale training,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9f0c0508-fd77-48f2-978e-a74fed2f6b3d · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech Syntaspeech: Syntax- aware generative adversarial text-to-speech,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7c89230c-00db-4818-a6cc-7e7873e10f0b · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech Emotional voice con- version: Theory, databases and esd,
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb99b960-9367-4744-a57c-e39fcc39f562 · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech Decoupled weight decay regulariza- tion,
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f854d94f-9b33-44f6-bdf1-248c9534ecb0 · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech Gaussian Error Linear Units (GELUs)
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 962890a7-b18b-4de6-bc54-d4d2fe32f05f · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech Attention is all you need,
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40934d75-70b6-4804-92c4-91e82b081f61 · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech Utmos: Utokyo-sarulab system for voicemos challenge 2022,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 28783661-d21c-4f50-ba61-08b5993018d3 · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech Robust speech recognition via large-scale weak su- pervision,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5fa609b9-6ec7-4f85-9a93-9a2989fae3c2 · outbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech Wavlm: Large-scale self- supervised pre-training for full stack speech processing,
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37201e95-7427-4e53-a70f-ca8993b94ef7 · inbound
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.