Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:29:42.730075Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 2 inbound Pith citation observations for arXiv:2505.24314.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:29:42.730075Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:29:38.143308Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T12:58:08.378519Z
36 of 36 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 2fa164f2-80ec-4a8d-b819-b94dfee2e5c0 · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec A pivotal challenge is the transformation of continuous speech signals into interpretable representations suitable for inference and training within large language models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 76b3dc45-1f91-4739-b897-3356dcbba6c3 · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0c0a5dc-6434-47c6-b322-762abc0e66cd · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Dataset and Metrics We use LibriSpeech [27] to train the speech codec we proposed
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 50daa269-e8bb-427e-85fa-e338e3a0802c · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 135c561d-84a1-4b24-ad73-6fd213a3eb8b · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Unresolved cited work
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c93301bc-58d4-4e4a-895d-74c5133d163b · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec GPT-4 Technical Report
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df8710a3-de5f-4dc5-a521-9e6086b4050e · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec GPT-4o System Card
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66668e64-fd4e-4e02-80ef-9a0e7aedb0c6 · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d20e720-c46f-429d-940e-18ebc9537d45 · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Audiolm: a language modeling approach to audio gener- ation,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dd6f89eb-99a1-4a4f-b2c3-2541d03efd3c · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec AudioGen: Textually Guided Audio Generation
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52a05e72-d8c6-460b-8948-185223fca620 · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Slim- speech: Lightweight and efficient text-to-speech with slim recti- fied flow,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8aed8f60-4168-402d-833e-330fb7cca6b2 · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec wav2vec 2.0: A framework for self-supervised learning of speech repre- sentations,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7f091b9-a0e6-4a24-993f-f5a1d4d7890b · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec wav2vec: Unsupervised Pre-training for Speech Recognition
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7fbe07b5-6546-48b6-b90f-3772504bdf95 · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Hubert: Self-supervised speech represen- tation learning by masked prediction of hidden units,
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78a5e2e8-1abb-47ea-85e7-721a5bb96f9d · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Neural discrete represen- tation learning,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2523b863-a545-4f77-a382-86958b415664 · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Soundstream: An end-to-end neural audio codec,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6b6e291-50a2-419b-bda5-1d403f010a1e · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec A review of vector quantization tech- niques,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f05c427-1b02-471e-9010-8995781d9162 · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec High-fidelity audio compression with improved rvqgan,
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7be356b-fd57-472e-a7d5-42ef4f47fad2 · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec High Fidelity Neural Audio Compression
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25530ef0-39a0-44ea-99d3-c96a65232ecd · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be0e75f5-0f79-433d-a3df-d9d60d153304 · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec WavTokenizer: an Efficient Acoustic Discrete Codec Tokenizer for Audio Language Modeling
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18ce0122-83d8-410c-a52c-5e72d928fc6a · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec BigCodec: Pushing the Limits of Low-Bitrate Neural Speech Codec
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd699265-cf72-4969-8c57-dfa127b10c6c · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Single-Codec: Single-Codebook Speech Codec towards High-Performance Speech Generation
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9bd0c454-47d4-4e4c-a0f1-a5ed8524626a · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Addressing Index Collapse of Large-Codebook Speech Tokenizer with Dual-Decoding Product-Quantized Variational Auto-Encoder
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 446a5104-842b-467a-b0a4-d240d0b57b5b · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4799653-6b8b-4791-8dde-31578b37c5f6 · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Neural networks fail to learn periodic functions and how to fix it,
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88b08a74-0353-4fc9-bb31-d1c8306b6e77 · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec LLaMA: Open and Efficient Foundation Language Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea8bc6fa-13d4-4dc6-b9ed-2dc502f00316 · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Hifi-gan: Generative adversarial net- works for efficient and high fidelity speech synthesis,
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72a21425-ef2f-45e1-b045-18c894474aed · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Vector-quantized Image Modeling with Improved VQGAN
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7dfcef2-cb64-443f-964a-127c625c7164 · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Product quantization for nearest neighbor search,
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d360477-8cb8-4fd3-8c03-57363716f9aa · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Apcodec+: A spectrum-coding-based high-fidelity and high-compression-rate neural audio codec with staged training paradigm,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 78882b71-3869-46ee-be91-9f552fe4d393 · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Lib- rispeech: an asr corpus based on public domain audio books,
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7213156-5f1e-49bc-937f-5e62869b6ad6 · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Vocos: Closing the gap between time-domain and Fourier-based neural vocoders for high-quality audio synthesis
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e454bdb2-bee3-4c12-b1ac-760223fe8ed8 · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec UTMOS: UTokyo-SaruLab System for VoiceMOS Challenge 2022
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59040c31-6b56-4635-ba6f-3724689adf01 · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Perceptual evaluation of speech quality (pesq)-a new method for speech quality assessment of telephone networks and codecs,
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83dbc189-7a8e-4ccf-b0a1-1a46cc871e63 · outbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Fewer-token neural speech codec with time-invariant codes,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 76b3dc45-1f91-4739-b897-3356dcbba6c3 · inbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8fceee68-b136-429c-92a9-30fb9d09d7bd · inbound
SARA: A Dual-Stream VAE for High-Fidelity Speech Generation via Integrating Semantic and Acoustic Representations DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.