Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:01:43.372701Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 3 inbound Pith citation observations for arXiv:2506.00843.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:01:43.372701Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:01:43.249937Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-11T19:46:15.415472Z
37 of 37 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation db0b3957-29f6-4b6f-b8cd-57d4aaa09ccb · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b4929345-79bb-4128-8beb-68d305901ca9 · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement Unresolved cited work
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8bd3210c-2a8c-4609-9fee-b50842449518 · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eeba0d67-d4b5-4d19-b4af-2ae36ba30a39 · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement For acoustic training, we adopt the DAC framework [16], extracting random 5-second segments (vs
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 82947716-31aa-43ff-bbdf-3dc0d1a48e89 · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1e0b5dee-e1bc-4f0a-8c5a-9f05b98671d7 · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement Our approach effectively preserves semantic performance for ASR while achieving reconstruction quality comparable to state-of-the-art neural audio codecs like DAC
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0da11723-e0cd-4d2b-9188-999a55a0e457 · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement Comparing Discrete and Continuous Space LLMs for Speech Recognition
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa2e66cf-500f-461a-a0b3-adc34cf2a18f · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement AudioLM: A language modeling approach to audio generation,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2693840c-85e6-414a-85aa-b7687e230ab6 · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement A survey on speech large language models,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99b8b4a1-1f61-4892-aff3-7334e25a17f6 · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement SpeechGPT: Empowering large language models with in- trinsic cross-modal conversational abilities,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8267a6c4-ff08-4903-b453-4e3fed29d9db · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement SpeechTokenizer: Unified Speech Tokenizer for Speech Large Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f951cb4-542c-4b3a-b069-28b70dc236dd · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement Exploring speech recognition, translation, and understanding with discrete speech units: A comparative study,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a28f4ebe-6fac-4bd1-9202-ff7f9a27a104 · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement HuBERT: Self-supervised speech representation learning by masked prediction of hidden units,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 62870af1-70b1-4e6b-add9-e5cae4802663 · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement WavLM: Large-scale self- supervised pre-training for full stack speech processing,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 39eac272-4f3c-464f-ab0c-9e0b666e396d · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement W2v-BERT: Combining contrastive learning and masked language modeling for self-supervised speech pre- training,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b7bc03b7-2eb0-481f-9ebb-67c954e0e0d9 · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement Ex- ploration of efficient end-to-end ASR using discretized input from self-supervised learning,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 63762711-6ffb-4f12-9c90-beb58a970dd6 · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement Speech resynthesis from discrete disentangled self-supervised representations,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1debd9fd-98c2-4a19-af30-efeb8f8a53c2 · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement Ex- presso: A benchmark and analysis of discrete expressive speech resynthesis,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ad690667-0ceb-4340-948d-a5881e8ed5dc · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement SoundStream: An end-to-end neural audio codec,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 14e5435b-bbc6-4a45-84ca-5fc60c5ad807 · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement High Fidelity Neural Audio Compression
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1bec4980-6a30-4924-a5a5-417b5b962cdc · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement Moshi: a speech-text foundation model for real-time dialogue
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69d89b07-bf94-43c7-b1d9-5e91394365e4 · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement High-fidelity audio compression with improved rvqgan,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5b674017-f082-4303-b5d2-94a79dfa57b3 · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement Codec Does Matter: Exploring the Semantic Shortcoming of Codec for Audio Language Model
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c5d41c7-f072-4f97-8fc0-eb203deab3ff · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement QR-VC: Leveraging Quantization Residuals for Linear Disentanglement in Zero-Shot Voice Conversion
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac8b6874-1279-488e-ab19-fcd25f5d331f · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement MMM: Multi-layer multi-residual multi-stream discrete speech repre- sentation from self-supervised learning model,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9550b58b-0ce6-43f0-9933-47545d4125e1 · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement Towards universal speech discrete tokens: A case study for ASR and TTS,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5229250d-6dd1-4705-8258-566427df1879 · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement ReVISE: Self-supervised speech resynthesis with visual input for univer- sal and generalized speech regeneration,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c93f827d-aab7-45c8-ab08-d67d4ec6893a · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement Self-supervised disentan- gled representation learning for robust target speech extraction,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5457c13d-fa21-4af3-aece-434aff075c76 · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement A ConvNet for the 2020s,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1c0dc1e9-17c3-4a52-841d-b2ddf644724c · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement Self-supervised learning with random-projection quantizer for speech recogni- tion,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7cbdc036-e0b1-4139-b856-a6d2a7927a1a · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement Mel- GAN: Generative adversarial networks for conditional wave- form synthesis,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 46499834-a384-4020-954c-719ab49dde7d · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement HiFi-GAN: Generative adversarial networks for efficient and high fidelity speech synthesis,
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f28f306-a4f7-4dee-a2c4-9d2e8bdd0af8 · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement Lib- riSpeech: An ASR corpus based on public domain audio books,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 13b65055-9954-4a90-baa7-7cd5312b1987 · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement Lhotse: A speech data representation library for the modern deep learning ecosystem,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e29dcf14-9587-42e5-ac29-b1656da52e00 · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement Open Implementation and Study of BEST-RQ for Speech Processing
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a8a9fc1d-9240-4912-9dd7-2837dee109f3 · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement Con- nectionist temporal classification: Labelling unsegmented se- quence data with recurrent neural networks,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d64d0e91-ea06-4e9c-87b0-c57f1cf42165 · outbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement ECAPA- TDNN: Emphasized channel attention, propagation and aggre- gation in TDNN based speaker verification,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8bd3210c-2a8c-4609-9fee-b50842449518 · inbound
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da8f0ff8-c3eb-4a88-8fbf-6acdd6e5f6ab · inbound
Minimizing Modality Gap from the Input Side: Your Speech LLM Can Be a Prosody-Aware Text LLM HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 98a9678c-94cd-430a-88af-17b774c73b52 · inbound
Minimizing Modality Gap from the Input Side: Your Speech LLM Can Be a Prosody-Aware Text LLM HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.