Pith. sign in

Paper Citation Record · LEDGER

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners

As of 21 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 1 inbound Pith citation observation for arXiv:2509.09201.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.09201 v1

Coverage vector

measured 57 of 57 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T19:37:53.994131Z

measured 58 of 58 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T19:37:50.680596Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

57 of 57 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved55
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 418f8ce1-a402-44c1-90ce-25eb16f76ea9 · outbound

This paper cites DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:50.680596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:50.680596Z digest=sha256:cbed5425bf6f467d923fd3d181f68bb473192c60e7ab4bbbe8ebcd4e25289978

Observation 51ebac8a-437c-4215-9974-fe6f2ad2e220 · outbound

This paper cites an unresolved cited work.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:50.725994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:50.725994Z digest=sha256:060a0497e4ed9236299aca6f9b9e5fd7e6a081ed9dcb8ba8d8e80f460ec43b65

Observation d8b17411-afa8-4976-9d98-db076111a33b · outbound

This paper cites Based on the differences in the physical generation mechanisms, it can be assumed that speech signal and background sound signal are mutually independent [31].

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Based on the differences in the physical generation mechanisms, it can be assumed that speech signal and background sound signal are mutually independent [31]

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:50.757211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:50.757211Z digest=sha256:33c91c3f3a29050b11ff042fc965fc921c6f94343a294fe41f4ab2514e0045ec

Observation d759351c-a47f-49b0-ae0f-79cb43032097 · outbound

This paper cites an unresolved cited work.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Unresolved cited work

Reference 4

Resolution
malformed identifier
no resolver link, observed 2026-08-04T19:37:50.818276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:50.818276Z digest=sha256:aff3e7abbb9904a8ba3c4e45aa75bc0f63738965cd358634e32e1bf322ed43b8

Observation 72f25802-bf56-4d13-895e-b50624272a68 · outbound

This paper cites front-end separation + back-end processing.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners front-end separation + back-end processing

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:50.870202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:50.870202Z digest=sha256:318daf30417151e0ab09b3aa75e68f386ab77c4f5881c6ecf78527e39ce109ec

Observation 767b99ed-9000-425b-bb18-2a081141b7e3 · outbound

This paper cites an unresolved cited work.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Unresolved cited work

Reference 6

Resolution
malformed identifier
no resolver link, observed 2026-08-04T19:37:50.927592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:50.927592Z digest=sha256:a36191c07a2237ac7a0d9017b573e3baa1ac3897bb04788ab83a4cf26edadcdc

Observation d74fee62-8daa-43f8-866d-6477696a3bba · outbound

This paper cites The experimental results confirm that the representations are sufficiently disentangled, enabling controllable feature selection tailored to diverse downstream tasks.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners The experimental results confirm that the representations are sufficiently disentangled, enabling controllable feature selection tailored to diverse downstream tasks

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.015687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.015687Z digest=sha256:cf46aa6e1bcdcbbd6de72239d93aa0f989c1c8d2bed0f4d0635ee55a49156a9f

Observation 3e72c3ac-f75b-48d2-aed9-0dd507bb38bf · outbound

This paper cites AudioPaLM: A Large Language Model That Can Speak and Listen.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners AudioPaLM: A Large Language Model That Can Speak and Listen

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.101392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.101392Z digest=sha256:9b154f581265fb75733e0cd2da2977621f1dea245693901164493845f70e5616

Observation a000a9fd-0dd0-4ee8-9670-f95db715c479 · outbound

This paper cites Audit: Audio editing by follow- ing instructions with latent diffusion models,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Audit: Audio editing by follow- ing instructions with latent diffusion models,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.183385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.183385Z digest=sha256:54578578afcc2c26c1650c9db615db79ca81fa6e9abfd89969b99e74b10ba702

Observation 0093045f-152c-4769-ac0d-64635a00c715 · outbound

This paper cites Superm2m: Supervised and mixture-to-mixture co-learning for speech enhancement and noise-robust asr,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Superm2m: Supervised and mixture-to-mixture co-learning for speech enhancement and noise-robust asr,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.260311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.260311Z digest=sha256:c411a0a669a70c303056d60e73affdc3e826028c86a269bf1d7496768a9bd8cc

Observation 8c819585-5800-4c7e-8115-d2922fb1edba · outbound

This paper cites Preserving background sound in noise-robust voice conversion via multi-task learning,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Preserving background sound in noise-robust voice conversion via multi-task learning,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.322124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.322124Z digest=sha256:3bb10bd602d0c90428c180770382e19ed5ed710ca46a39b034cb9cf69ed14ed7

Observation c2c04314-f074-4cb2-bc94-9e63231eff77 · outbound

This paper cites Supervised speech separation based on deep learning: An overview,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Supervised speech separation based on deep learning: An overview,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.405886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.405886Z digest=sha256:cbe3e855a727070cb22b59916ecbc7e434ff4dd1911e664a7b1d128a42005c1f

Observation 675636fe-b5ad-4936-ae5f-1fda4a9b4368 · outbound

This paper cites Complex spectral mapping for single-and multi-channel speech enhancement and robust asr,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Complex spectral mapping for single-and multi-channel speech enhancement and robust asr,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.476825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.476825Z digest=sha256:fcda60a7ae8a16a8263b04bbbf6ec043d8c725d73641d1e746093ee0dbf8a89c

Observation 7636640a-1529-434f-8adf-e9212ee8de43 · outbound

This paper cites Speech enhancement with lstm recurrent neural networks and its application to noise- robust asr,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Speech enhancement with lstm recurrent neural networks and its application to noise- robust asr,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.519961Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.519961Z digest=sha256:58f9e37e97af88e4832201416468749c2bd78e19f7a62638267ec3ce4e3a2fca

Observation 0919e23d-e6ec-4f44-ba40-89530926a584 · outbound

This paper cites Deep neural networks for speech enhancement and speech recognition: A systematic review,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Deep neural networks for speech enhancement and speech recognition: A systematic review,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.571480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.571480Z digest=sha256:269219547c6c0bc5d69cbb90f94695f33d518a61bced2d0d0654c358852ab542

Observation 8e8b811f-e6a9-42c5-adea-5d923a969839 · outbound

This paper cites Attention-based latent features for jointly trained end-to-end automatic speech recognition with modified speech enhancement,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Attention-based latent features for jointly trained end-to-end automatic speech recognition with modified speech enhancement,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.650758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.650758Z digest=sha256:09ddb55ae43d270b14d90464829eba318fc7cd68a5ff2ceaadfd4b96fe8cbe34

Observation 0b3eaef0-d7cd-412e-89ea-c3cb7eabf06b · outbound

This paper cites Phonetic feature encoding in human superior temporal gyrus,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Phonetic feature encoding in human superior temporal gyrus,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.734805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.734805Z digest=sha256:71c8e3b091cbae413aa23ba33926f70130d12efbfa0234cade2ae7c526fa1a04

Observation 304a0721-9c83-40a0-acce-11fd71c31963 · outbound

This paper cites Representation of temporal sound features in the human auditory cortex,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Representation of temporal sound features in the human auditory cortex,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.791267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.791267Z digest=sha256:400802dbb5ba90a358a08eee5b5a48010d38981b8d2480880cc626bb178eca1f

Observation 288879da-e5ce-4acf-9ef9-f8d101a3c88a · outbound

This paper cites High Fidelity Neural Audio Compression.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners High Fidelity Neural Audio Compression

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.875056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.875056Z digest=sha256:9662e8684eb3bf5c2352408dd7ec87ae02c1b1e6aeba1e68a0fc7799df0427eb

Observation a70cc88a-458a-4e35-bd7a-f1b01a6c237e · outbound

This paper cites High-fidelity au- dio compression with improved rvqgan,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners High-fidelity au- dio compression with improved rvqgan,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.951394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.951394Z digest=sha256:f1b4fe016d9fd70fb8923e088aa6d3bbc6358eebaea72237bdcc1625219c7fcf

Observation 143f1e17-f16f-4f2d-b198-ec8a76fb5623 · outbound

This paper cites UniCodec: Unified Audio Codec with Single Domain-Adaptive Codebook.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners UniCodec: Unified Audio Codec with Single Domain-Adaptive Codebook

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.994459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.994459Z digest=sha256:73853fb8691059d9e7d8bfecf678a74dd26b90ca848814b778f705fd1ce5ff7e

Observation 7dfd02ad-5cab-496b-9e6d-33fe1fbad1e6 · outbound

This paper cites NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.044394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.044394Z digest=sha256:454c10e6c79b22a2c8e43945cb1f21fd7947ed063337a4c73911adbccfc3dac8

Observation 9d390e1d-3709-4354-a1b9-ec425eae07b5 · outbound

This paper cites SpeechTokenizer: Unified Speech Tokenizer for Speech Large Language Models.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners SpeechTokenizer: Unified Speech Tokenizer for Speech Large Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.109011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.109011Z digest=sha256:76a163e8e04c2886ec0e73a94a1b1de3e8dd36aa30bb7ecdd5e46a560a439913

Observation d2aa8661-24bb-4b08-835d-a182292dc5d0 · outbound

This paper cites Moshi: a speech-text foundation model for real-time dialogue.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Moshi: a speech-text foundation model for real-time dialogue

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.181884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.181884Z digest=sha256:c123a6b64afacd459c165053af58b81a3e79935a1b59c227b454f0eb82e4283c

Observation bfb54c71-31ca-47cd-bcd4-8373d8efd4c3 · outbound

This paper cites Dualcodec: A low-frame-rate, semantically-enhanced neural audio codec for speech generation,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Dualcodec: A low-frame-rate, semantically-enhanced neural audio codec for speech generation,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.245258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.245258Z digest=sha256:c8d2991bc0c0a2f2317679574e76c4064dfcf831ad31b1ee3eb2528765a4bd21

Observation 48c2eb87-df7c-4182-adba-3d8e8ad3215a · outbound

This paper cites Brain plasticity under early auditory deprivation: evidence from congenital hearing-impaired people,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Brain plasticity under early auditory deprivation: evidence from congenital hearing-impaired people,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.315157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.315157Z digest=sha256:b4f7c8073af254927ff4da8216d8d0eea4e7ace00e11eec1b307261dc3a55728

Observation f3acf3a7-62b6-4c28-9b30-641c07a98d23 · outbound

This paper cites A tutorial on mpeg/audio compression,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners A tutorial on mpeg/audio compression,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.370793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.370793Z digest=sha256:9a55801a92e121e9b75e7e59adb08b962f246ea64ff0c66304094fed4aca7578

Observation cf877525-a0bb-42e4-b250-0bafe52deed8 · outbound

This paper cites Soundstream: An end-to-end neural audio codec,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Soundstream: An end-to-end neural audio codec,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.454439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.454439Z digest=sha256:21c45372ea9349faf02cfc2e862a2305bc2b6d28a2245ab19a6f4890898ddc67

Observation eccc40f4-61df-481b-af0b-307de4dfe04a · outbound

This paper cites HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.501429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.501429Z digest=sha256:aa049fdb026c4390af6218c3f9ffc0964561b158397cae6a4f940861a35ec057

Observation 2d2e3a6a-f07b-4d3d-893a-fa92edcd6f4e · outbound

This paper cites Flow-vae vc: end-to-end flow framework with contrastive loss for zero-shot voice conversion,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Flow-vae vc: end-to-end flow framework with contrastive loss for zero-shot voice conversion,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.561461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.561461Z digest=sha256:fd1e80f70dcf4ce451fcf673f1e3fae2ad362668c21148648f04279ecef8cd5b

Observation bbc9e49c-ef69-47f3-9146-bd6ef4331b1a · outbound

This paper cites Hubert: Self-supervised speech rep- resentation learning by masked prediction of hidden units,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Hubert: Self-supervised speech rep- resentation learning by masked prediction of hidden units,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.614490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.614490Z digest=sha256:c7fb66605b1dd492c62de18bdeaef03f693ff6bd27e6974b75cc515446a64a8c

Observation 5f859916-789d-445a-8fa7-feb35227b639 · outbound

This paper cites wav2vec 2.0: A framework for self-supervised learning of speech representations,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners wav2vec 2.0: A framework for self-supervised learning of speech representations,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.643574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.643574Z digest=sha256:be3e7d214a627e90ce1d6e1f4cd36c04962960c11aff877b15e2ac7110b17657

Observation fdc269f3-687d-46b1-bfab-0830b3832565 · outbound

This paper cites Wavlm: Large-scale self-supervised pre-training for full stack speech processing,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Wavlm: Large-scale self-supervised pre-training for full stack speech processing,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.726141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.726141Z digest=sha256:c67a8c71e0500ec05d47a3b51eb26c60feaa8e7e5ea79a399ff75d10d4c65b99

Observation 987f8729-1321-4bef-bda1-a959f73e3152 · outbound

This paper cites SUPERB-SG: Enhanced Speech processing Universal PERformance Benchmark for Semantic and Generative Capabilities.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners SUPERB-SG: Enhanced Speech processing Universal PERformance Benchmark for Semantic and Generative Capabilities

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.768658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.768658Z digest=sha256:a99713a4ba695ee33d70240522685e32047be27853b6fde4aace98bd9302504c

Observation 308b81e9-2564-4b84-92a7-c3107e254bc9 · outbound

This paper cites RepCodec: A Speech Representation Codec for Speech Tokenization.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners RepCodec: A Speech Representation Codec for Speech Tokenization

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.824678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.824678Z digest=sha256:ea1b10270212192508a3a6dadc7760da9138ff3c40ac95ede1a9302c15544e72

Observation c4672b52-6489-4b5d-9fc1-6ca23d2aeb53 · outbound

This paper cites Funcodec: A fundamental, reproducible and integrable open-source toolkit for neural speech codec,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Funcodec: A fundamental, reproducible and integrable open-source toolkit for neural speech codec,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.878282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.878282Z digest=sha256:2e11d6221d8c360b698257a07126749065dc9f98bcaeb48826db1a5c0e6b48d0

Observation 5538d06b-21be-4960-9335-9ed3868c6660 · outbound

This paper cites Semanticodec: An ul- tra low bitrate semantic audio codec for general sound,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Semanticodec: An ul- tra low bitrate semantic audio codec for general sound,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.907602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.907602Z digest=sha256:1d9f202d28c6b76fc34630e5bd0b7e6d85dff952472c1376e53b73f45536a243

Observation fd8a5c3f-1518-47c4-a32f-e3ac2de11812 · outbound

This paper cites Sixty years of frequency-domain monaural speech en- hancement: From traditional to deep learning methods,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Sixty years of frequency-domain monaural speech en- hancement: From traditional to deep learning methods,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.960879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.960879Z digest=sha256:e48b77ab51bb6faedc8696d705d44e7f48752196334fd40872eadd40eb4133d3

Observation 0c42ae0b-25e4-466e-a677-449103f0395c · outbound

This paper cites A review of vector quantization techniques,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners A review of vector quantization techniques,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.007585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.007585Z digest=sha256:02c0982a052de53bc3038a2da4a2ede27aebc5bc578cd21a449b2b405e9dc392

Observation 3ac3e4db-175f-410b-80f3-72ee20d7fc32 · outbound

This paper cites an unresolved cited work.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.066979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.066979Z digest=sha256:1078aa812d0d812e8467438e5a8755fb3349705938ea836f6416f29c132d224e

Observation c7d849c3-24ba-4204-89b7-9c77de90223e · outbound

This paper cites Neural dis- crete representation learning,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Neural dis- crete representation learning,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.096021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.096021Z digest=sha256:e56256034992c2ab5ded284af1b77171fff0ecd19d99ef023923c4ccab8dff15

Observation 3fe8a6d4-6c3c-4f6e-a61e-b88ac4cb3a0d · outbound

This paper cites AISHELL-3: A Multi-speaker Mandarin TTS Corpus and the Baselines.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners AISHELL-3: A Multi-speaker Mandarin TTS Corpus and the Baselines

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.149135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.149135Z digest=sha256:71f6759b01e4deb8f5e3c51b12f7e7fe1a2f313c5cd357bd37d37fbb24a17356

Observation 0ff7203d-9946-4890-96b3-d0be7e7a1bb0 · outbound

This paper cites LibriTTS: A Corpus Derived from LibriSpeech for Text-to-Speech.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners LibriTTS: A Corpus Derived from LibriSpeech for Text-to-Speech

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.193305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.193305Z digest=sha256:e623bd3fa5a5e2054084cfbf26d591850b79076bc1a474fdffea6bf4adc50a7f

Observation b08ec887-b932-4c57-99aa-ded01908740d · outbound

This paper cites The voice bank corpus: Design, collection and data analysis of a large regional accent speech database,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners The voice bank corpus: Design, collection and data analysis of a large regional accent speech database,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.249521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.249521Z digest=sha256:3abca31b33a671d73f4e509fbfa500730205629fcdbca391eac1f8a606fdb112

Observation 0c2628a5-a575-43e7-ad54-c9f4e6319597 · outbound

This paper cites The design for the wall street journal- based,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners The design for the wall street journal- based,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.336254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.336254Z digest=sha256:6511989cdc53c2ba21bcbe5de7e058899ba62d0f3a224db6a0764018929dcc5b

Observation 2ade5bfd-d711-4079-85fb-10545fbdb348 · outbound

This paper cites Esc: Dataset for environmental sound classification,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Esc: Dataset for environmental sound classification,

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.415756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.415756Z digest=sha256:287745c3426152f9219ce316c05b3f1566e8f30b59e9850161feaa46f64e7005

Observation ed606094-3758-4290-9f32-7b616663d3b6 · outbound

This paper cites The INTERSPEECH 2020 Deep Noise Suppression Challenge: Datasets, Subjective Testing Framework, and Challenge Results.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners The INTERSPEECH 2020 Deep Noise Suppression Challenge: Datasets, Subjective Testing Framework, and Challenge Results

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.473571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.473571Z digest=sha256:6ca61691b6ca79fd77846523f7d9f3accc28fc3aa1dfb5558d69720eef2061ec

Observation 03832718-d907-4723-86a0-1af4f5511876 · outbound

This paper cites First stereo audio source separation evaluation campaign: data, algorithms and results,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners First stereo audio source separation evaluation campaign: data, algorithms and results,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.530768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.530768Z digest=sha256:8aa5a70e2479bee1dc88fbe613e4f8d5bc355d137707c0598c42e241c40a0675

Observation 9d4956ab-1ef7-4e02-b000-f4c7185a9688 · outbound

This paper cites Robust speech recognition via large-scale weak supervision,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Robust speech recognition via large-scale weak supervision,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.584010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.584010Z digest=sha256:28371ba93c178953d4bf2133e1bc78427d68c48872151d4c34eb41018c14dcd7

Observation 7f27f1f1-edb8-4ca8-b669-ec4ca13cbc3d · outbound

This paper cites Dnsmos: A non-intrusive perceptual objective speech quality metric to evaluate noise suppressors,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Dnsmos: A non-intrusive perceptual objective speech quality metric to evaluate noise suppressors,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.644016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.644016Z digest=sha256:b89bc1b40f0f022d5cdb75b8135f585e6df78be023fbadf4459a678f304da390

Observation 0df59a6a-045c-4892-b56a-b72933812da5 · outbound

This paper cites Inter-subnet: Speech enhancement with subband in- teraction,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Inter-subnet: Speech enhancement with subband in- teraction,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.669090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.669090Z digest=sha256:935a815ef9f767267fc3ae921298d61208db710046f20939e0613789bc684e7b

Observation ce48057c-768d-4a47-ac92-b8fc049ae857 · outbound

This paper cites Storm: A diffusion-based stochastic regeneration model for speech enhancement and dereverberation,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Storm: A diffusion-based stochastic regeneration model for speech enhancement and dereverberation,

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.722918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.722918Z digest=sha256:865597e8a5cd533e66c2fea53e09e552e7641bb44d686a89d6c85148c67bb0a2

Observation 4c826304-4843-4d75-8b72-968118a0e324 · outbound

This paper cites Selm: Speech enhancement using discrete tokens and language mod- els,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Selm: Speech enhancement using discrete tokens and language mod- els,

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.774367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.774367Z digest=sha256:407c3d18822619121e7be36cc3bf3b2f6ef35923b50e06050b31dbd420765a15

Observation 5ca5a13c-4aa2-47f2-825c-8cc2979366aa · outbound

This paper cites Attention is all you need,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Attention is all you need,

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.831514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.831514Z digest=sha256:7dd5949b24b016a846494515f4b6163b6b62c5f66fdcb158fac1d9f1992240b0

Observation 34b8bdc9-f242-4121-ac11-2ed83fb76510 · outbound

This paper cites PolySpeech: Exploring Unified Multitask Speech Models for Competitiveness with Single-task Models.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners PolySpeech: Exploring Unified Multitask Speech Models for Competitiveness with Single-task Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.886688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.886688Z digest=sha256:28b52ecd40418cba21ef69cebf2a859435eeb414d5280c15cdf67be5e0c07b9e

Observation f489a456-69c3-44f8-9163-9a4a535cf71a · outbound

This paper cites Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.944636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.944636Z digest=sha256:6e4e9048c5472c34283895e030c1c2dc3e8577341495142c20c5677ac74e95fa

Observation c6eee5b8-c5b0-4fd9-b5d5-aeff362fc939 · outbound

This paper cites Decoupled Weight Decay Regularization.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Decoupled Weight Decay Regularization

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.994131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.994131Z digest=sha256:18717d8c946dbd5b433c4f54f41877c97bce79b37afe0c6823180cf62d94a915

Pith citing papers

Observation 418f8ce1-a402-44c1-90ce-25eb16f76ea9 · inbound

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners cites this paper.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:50.680596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:50.680596Z digest=sha256:cbed5425bf6f467d923fd3d181f68bb473192c60e7ab4bbbe8ebcd4e25289978