Pith. sign in

Paper Citation Record · LEDGER

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec

As of 7 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 2 inbound Pith citation observations for arXiv:2505.24314.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.24314 v1

Coverage vector

measured 36 of 36 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:29:42.730075Z

measured 38 of 38 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:29:38.143308Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T12:58:08.378519Z

Reference resolution

36 of 36 outbound references displayed

  • verified exact1
  • verified fuzzy7
  • unresolved28
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2fa164f2-80ec-4a8d-b819-b94dfee2e5c0 · outbound

This paper cites A pivotal challenge is the transformation of continuous speech signals into interpretable representations suitable for inference and training within large language models.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec A pivotal challenge is the transformation of continuous speech signals into interpretable representations suitable for inference and training within large language models

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:29:45.426994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:29:38.015315Z digest=sha256:32a9e264336421dda08e48abb0fd420a3cf24f1a94a92d720e1531f46d27ac1c

Observation 76b3dc45-1f91-4739-b897-3356dcbba6c3 · outbound

This paper cites DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:38.143308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:38.143308Z digest=sha256:a19a6f9c676ec3c84a40c9addc0069d504b270f2e3ab6c269e2896d4f4010ca6

Observation c0c0a5dc-6434-47c6-b322-762abc0e66cd · outbound

This paper cites Dataset and Metrics We use LibriSpeech [27] to train the speech codec we proposed.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Dataset and Metrics We use LibriSpeech [27] to train the speech codec we proposed

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:29:45.216566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:29:38.356758Z digest=sha256:8d6f69e383aa002e2c1f39cfdca6015cd45d6ee3e551dc4c0c1e00414e8c9b08

Observation 50daa269-e8bb-427e-85fa-e338e3a0802c · outbound

This paper cites an unresolved cited work.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:29:45.003225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:29:38.550836Z digest=sha256:a413c12a208bab59745f65ea4bbecb605a965ee59325af87e56d72c68b227254

Observation 135c561d-84a1-4b24-ad73-6fd213a3eb8b · outbound

This paper cites an unresolved cited work.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:38.724956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:38.724956Z digest=sha256:576daebaf086aa883b501f584732d48d9550f6e5a38e12454cc913ae402dca92

Observation c93301bc-58d4-4e4a-895d-74c5133d163b · outbound

This paper cites GPT-4 Technical Report.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec GPT-4 Technical Report

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:38.877708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:38.877708Z digest=sha256:969f78dbda72822615dbd28128672cfe312566b453d02638e26cef364ee6626f

Observation df8710a3-de5f-4dc5-a521-9e6086b4050e · outbound

This paper cites GPT-4o System Card.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec GPT-4o System Card

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:38.995350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:38.995350Z digest=sha256:b10a02f0ba417f22ffb8717b56dd2c0f8d403b1eb62922f3953b6459498be80f

Observation 66668e64-fd4e-4e02-80ef-9a0e7aedb0c6 · outbound

This paper cites Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:39.170430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:39.170430Z digest=sha256:facfc60f8232ed6f4d254f2675dbe3c3453d9f5c9c63800d32c7415aedb58d78

Observation 4d20e720-c46f-429d-940e-18ebc9537d45 · outbound

This paper cites Audiolm: a language modeling approach to audio gener- ation,.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Audiolm: a language modeling approach to audio gener- ation,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:29:44.799493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:29:39.311348Z digest=sha256:657d39c583d576638c7ca593d5308d78314b6a5320e66b31d647348f178e17f3

Observation dd6f89eb-99a1-4a4f-b2c3-2541d03efd3c · outbound

This paper cites AudioGen: Textually Guided Audio Generation.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec AudioGen: Textually Guided Audio Generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:39.487610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:39.487610Z digest=sha256:20baa83f7fdf6459348580619d123e819202e97a52f718bd380be80540332cc4

Observation 52a05e72-d8c6-460b-8948-185223fca620 · outbound

This paper cites Slim- speech: Lightweight and efficient text-to-speech with slim recti- fied flow,.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Slim- speech: Lightweight and efficient text-to-speech with slim recti- fied flow,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:29:44.553668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:29:39.605523Z digest=sha256:0c6bb37b34495daf7c05393fd7bbe2aab43cfda7ff642bec3db414e7ccf0c387

Observation 8aed8f60-4168-402d-833e-330fb7cca6b2 · outbound

This paper cites wav2vec 2.0: A framework for self-supervised learning of speech repre- sentations,.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec wav2vec 2.0: A framework for self-supervised learning of speech repre- sentations,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:39.703511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:39.703511Z digest=sha256:955c0429323e4fed568ca34f12063a6669e642a5f494a38a5fff964fe6b6a4e2

Observation e7f091b9-a0e6-4a24-993f-f5a1d4d7890b · outbound

This paper cites wav2vec: Unsupervised Pre-training for Speech Recognition.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec wav2vec: Unsupervised Pre-training for Speech Recognition

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:39.796888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:39.796888Z digest=sha256:4001203ff3feefff8a27549beebbe6584eed77060f103c044ccbc909bb531107

Observation 7fbe07b5-6546-48b6-b90f-3772504bdf95 · outbound

This paper cites Hubert: Self-supervised speech represen- tation learning by masked prediction of hidden units,.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Hubert: Self-supervised speech represen- tation learning by masked prediction of hidden units,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:39.892248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:39.892248Z digest=sha256:efa92775600ddca561cfb2b93c2ad6541d0da5dd0c1d3094bb4caf4c5452a239

Observation 78a5e2e8-1abb-47ea-85e7-721a5bb96f9d · outbound

This paper cites Neural discrete represen- tation learning,.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Neural discrete represen- tation learning,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:29:44.307824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:29:39.994883Z digest=sha256:ed77b6d859830a25dac8de298b5e0f890de73b68e69224cf418b47f33de99741

Observation 2523b863-a545-4f77-a382-86958b415664 · outbound

This paper cites Soundstream: An end-to-end neural audio codec,.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Soundstream: An end-to-end neural audio codec,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:40.070603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:40.070603Z digest=sha256:df1dea69cc78506cab8d49efe17095b3984671d15cc060ede2b4a905fc552293

Observation c6b6e291-50a2-419b-bda5-1d403f010a1e · outbound

This paper cites A review of vector quantization tech- niques,.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec A review of vector quantization tech- niques,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:40.192413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:40.192413Z digest=sha256:3dbd18ecd8a5354fe8f18c356ab95f938a6986bc84255e6ab393455a252e8f69

Observation 4f05c427-1b02-471e-9010-8995781d9162 · outbound

This paper cites High-fidelity audio compression with improved rvqgan,.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec High-fidelity audio compression with improved rvqgan,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:40.359010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:40.359010Z digest=sha256:a609d6562850192e8ca70989cb2bb6ac098a3dde3a451a4d410ac77d5bb5dbef

Observation e7be356b-fd57-472e-a7d5-42ef4f47fad2 · outbound

This paper cites High Fidelity Neural Audio Compression.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec High Fidelity Neural Audio Compression

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:40.426238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:40.426238Z digest=sha256:f9b3d41e19a76846291c6914a52579b22129ff59c95c5802dd930b1ae8617e89

Observation 25530ef0-39a0-44ea-99d3-c96a65232ecd · outbound

This paper cites HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:40.551229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:40.551229Z digest=sha256:5197cee34a047e83a7419c47afb6568dc1301bc6ffcb57f683d21832f53a65cf

Observation be0e75f5-0f79-433d-a3df-d9d60d153304 · outbound

This paper cites WavTokenizer: an Efficient Acoustic Discrete Codec Tokenizer for Audio Language Modeling.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec WavTokenizer: an Efficient Acoustic Discrete Codec Tokenizer for Audio Language Modeling

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:40.695004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:40.695004Z digest=sha256:6d2dc4a2dfcfe5dbb3a1caf89112cda31db0c7f3939253154fc5c99a0775cffe

Observation 18ce0122-83d8-410c-a52c-5e72d928fc6a · outbound

This paper cites BigCodec: Pushing the Limits of Low-Bitrate Neural Speech Codec.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec BigCodec: Pushing the Limits of Low-Bitrate Neural Speech Codec

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:40.791796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:40.791796Z digest=sha256:2f61d8cdac23cc9f1f39d4853343232db3c5b6992c98fee3de82adaa1ca7fd73

Observation dd699265-cf72-4969-8c57-dfa127b10c6c · outbound

This paper cites Single-Codec: Single-Codebook Speech Codec towards High-Performance Speech Generation.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Single-Codec: Single-Codebook Speech Codec towards High-Performance Speech Generation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:40.894154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:40.894154Z digest=sha256:32173172f3cb8cbb14ea6144b2d9ba183a1861af2d3aa05829ea7c1895f04828

Observation 9bd0c454-47d4-4e4c-a0f1-a5ed8524626a · outbound

This paper cites Addressing Index Collapse of Large-Codebook Speech Tokenizer with Dual-Decoding Product-Quantized Variational Auto-Encoder.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Addressing Index Collapse of Large-Codebook Speech Tokenizer with Dual-Decoding Product-Quantized Variational Auto-Encoder

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:29:43.140544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:29:41.008351Z digest=sha256:ab716331ba9009885acf54f1239f5582053db9a733625ca5bf2c90eec0222218

Observation 446a5104-842b-467a-b0a4-d240d0b57b5b · outbound

This paper cites NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:41.168284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:41.168284Z digest=sha256:a247d1b2a0d5f918456a96ce90c36b1e2067cfe0b9ce10ea49b0c0a58cbcbae8

Observation e4799653-6b8b-4791-8dde-31578b37c5f6 · outbound

This paper cites Neural networks fail to learn periodic functions and how to fix it,.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Neural networks fail to learn periodic functions and how to fix it,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:41.272611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:41.272611Z digest=sha256:56cb4fe31136c9b03d32ab0ed7cde3d48c5dba48e308c857af7cf05f5bd234b8

Observation 88b08a74-0353-4fc9-bb31-d1c8306b6e77 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec LLaMA: Open and Efficient Foundation Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:41.423921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:41.423921Z digest=sha256:18e0a8582eba5521b88184aec6d8cfe7c8e0ee1a8791532b33889475d488219b

Observation ea8bc6fa-13d4-4dc6-b9ed-2dc502f00316 · outbound

This paper cites Hifi-gan: Generative adversarial net- works for efficient and high fidelity speech synthesis,.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Hifi-gan: Generative adversarial net- works for efficient and high fidelity speech synthesis,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:41.569951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:41.569951Z digest=sha256:e889d8342d163db535b24d3db12aea640b07ab74b5d7f17e36be17845bca279d

Observation 72a21425-ef2f-45e1-b045-18c894474aed · outbound

This paper cites Vector-quantized Image Modeling with Improved VQGAN.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Vector-quantized Image Modeling with Improved VQGAN

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:41.764813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:41.764813Z digest=sha256:ffb96a88beae4ca68c9b3cfc0f2f8682e3a98d03ec78c4158a368373d012406e

Observation d7dfcef2-cb64-443f-964a-127c625c7164 · outbound

This paper cites Product quantization for nearest neighbor search,.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Product quantization for nearest neighbor search,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:41.908267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:41.908267Z digest=sha256:ba77157c7c8b2513d5c63488f7da84474fb6f3c0830a4eef1d0ede478eee1c8a

Observation 4d360477-8cb8-4fd3-8c03-57363716f9aa · outbound

This paper cites Apcodec+: A spectrum-coding-based high-fidelity and high-compression-rate neural audio codec with staged training paradigm,.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Apcodec+: A spectrum-coding-based high-fidelity and high-compression-rate neural audio codec with staged training paradigm,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:29:44.001741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:29:42.045168Z digest=sha256:b54a0c7edc225eef9f35bac294e53a0861ef45fee97a17aea91fe221e624308b

Observation 78882b71-3869-46ee-be91-9f552fe4d393 · outbound

This paper cites Lib- rispeech: an asr corpus based on public domain audio books,.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Lib- rispeech: an asr corpus based on public domain audio books,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:42.189677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:42.189677Z digest=sha256:d8880bdc9f00514766844838b0bd83b8ef15d76611f52096a4a15728b77a91b1

Observation f7213156-5f1e-49bc-937f-5e62869b6ad6 · outbound

This paper cites Vocos: Closing the gap between time-domain and Fourier-based neural vocoders for high-quality audio synthesis.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Vocos: Closing the gap between time-domain and Fourier-based neural vocoders for high-quality audio synthesis

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:42.327056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:42.327056Z digest=sha256:24c1b775560073c0fdf5f5c138d36f5dad5b3b73a0d6734d684ba3f461b2e126

Observation e454bdb2-bee3-4c12-b1ac-760223fe8ed8 · outbound

This paper cites UTMOS: UTokyo-SaruLab System for VoiceMOS Challenge 2022.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec UTMOS: UTokyo-SaruLab System for VoiceMOS Challenge 2022

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:42.472159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:42.472159Z digest=sha256:0d164bd3c5ff74135695cacd78996fa93a7ef4868999458e46e608d1927276a0

Observation 59040c31-6b56-4635-ba6f-3724689adf01 · outbound

This paper cites Perceptual evaluation of speech quality (pesq)-a new method for speech quality assessment of telephone networks and codecs,.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Perceptual evaluation of speech quality (pesq)-a new method for speech quality assessment of telephone networks and codecs,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:42.622160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:42.622160Z digest=sha256:55d1596669d5da91f8edc6ce2b95fb37c92dde6fbf708fac1ea753404fc89dbe

Observation 83dbc189-7a8e-4ccf-b0a1-1a46cc871e63 · outbound

This paper cites Fewer-token neural speech codec with time-invariant codes,.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec Fewer-token neural speech codec with time-invariant codes,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:29:43.717631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:29:42.730075Z digest=sha256:0d063dcf21be829b78fcd3111dedf64d28ed1304e9a57db78570da3775813a22

Pith citing papers

Observation 76b3dc45-1f91-4739-b897-3356dcbba6c3 · inbound

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec cites this paper.

DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:38.143308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:38.143308Z digest=sha256:a19a6f9c676ec3c84a40c9addc0069d504b270f2e3ab6c269e2896d4f4010ca6

Observation 8fceee68-b136-429c-92a9-30fb9d09d7bd · inbound

SARA: A Dual-Stream VAE for High-Fidelity Speech Generation via Integrating Semantic and Acoustic Representations cites this paper.

SARA: A Dual-Stream VAE for High-Fidelity Speech Generation via Integrating Semantic and Acoustic Representations DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-03T12:58:08.380228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T08:38:51.371500Z digest=sha256:92e4d01fa18f1591cf4bc336067f2e979ba21cc9dc607d867fb5f419e73d433c