Pith. sign in

Paper Citation Record · LEDGER

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners

As of 8 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 1 inbound Pith citation observation for arXiv:2509.09201.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.09201 v1

Coverage vector

measured 57 of 57 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T19:37:53.994131Z

measured 58 of 58 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T19:37:50.680596Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

57 of 57 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved55
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 418f8ce1-a402-44c1-90ce-25eb16f76ea9 · outbound

This paper cites DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:50.680596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:50.680596Z digest=sha256:4a5e1be06b53e43e10b948997c1ef174729e76732455fa0a8038ada15e377487

Observation 51ebac8a-437c-4215-9974-fe6f2ad2e220 · outbound

This paper cites an unresolved cited work.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:50.725994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:50.725994Z digest=sha256:5d235024e1f6f278b354ff66d517ce966ed083a6a657a290448699aae9cf47f8

Observation d8b17411-afa8-4976-9d98-db076111a33b · outbound

This paper cites Based on the differences in the physical generation mechanisms, it can be assumed that speech signal and background sound signal are mutually independent [31].

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Based on the differences in the physical generation mechanisms, it can be assumed that speech signal and background sound signal are mutually independent [31]

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:50.757211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:50.757211Z digest=sha256:7af77959d95cccabf89a84ea41d24d7448e1e54e98ab3e9b05d72a7285edfb8c

Observation d759351c-a47f-49b0-ae0f-79cb43032097 · outbound

This paper cites an unresolved cited work.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Unresolved cited work

Reference 4

Resolution
malformed identifier
no resolver link, observed 2026-08-04T19:37:50.818276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:50.818276Z digest=sha256:e6c79f3e9194ae3469a79a1923db82a65a285a6399b9457d0755ed27eb5d2a98

Observation 72f25802-bf56-4d13-895e-b50624272a68 · outbound

This paper cites front-end separation + back-end processing.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners front-end separation + back-end processing

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:50.870202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:50.870202Z digest=sha256:b0677f2884bdeb6f2bdce0463bd80cac458079815436b6d14ac41d3640efa2e3

Observation 767b99ed-9000-425b-bb18-2a081141b7e3 · outbound

This paper cites an unresolved cited work.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Unresolved cited work

Reference 6

Resolution
malformed identifier
no resolver link, observed 2026-08-04T19:37:50.927592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:50.927592Z digest=sha256:963f0b0f9e33d305566667ceba9aa9a2258cfa3de1a553bdebd7688982021a99

Observation d74fee62-8daa-43f8-866d-6477696a3bba · outbound

This paper cites The experimental results confirm that the representations are sufficiently disentangled, enabling controllable feature selection tailored to diverse downstream tasks.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners The experimental results confirm that the representations are sufficiently disentangled, enabling controllable feature selection tailored to diverse downstream tasks

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.015687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.015687Z digest=sha256:07466d1f581fbf7af9534018e97f8b650896702a325665d296f6b0675645787a

Observation 3e72c3ac-f75b-48d2-aed9-0dd507bb38bf · outbound

This paper cites AudioPaLM: A Large Language Model That Can Speak and Listen.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners AudioPaLM: A Large Language Model That Can Speak and Listen

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.101392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.101392Z digest=sha256:dfc9dd448da7d2c771c15ac8ec6fdc41fbc51a194c8256662a416923f406af06

Observation a000a9fd-0dd0-4ee8-9670-f95db715c479 · outbound

This paper cites Audit: Audio editing by follow- ing instructions with latent diffusion models,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Audit: Audio editing by follow- ing instructions with latent diffusion models,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.183385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.183385Z digest=sha256:b6e48ac4566ac8c8f45c106467cc986c83866bf770a22aeeb3a6f5dd4b02879a

Observation 0093045f-152c-4769-ac0d-64635a00c715 · outbound

This paper cites Superm2m: Supervised and mixture-to-mixture co-learning for speech enhancement and noise-robust asr,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Superm2m: Supervised and mixture-to-mixture co-learning for speech enhancement and noise-robust asr,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.260311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.260311Z digest=sha256:b2958bd3744388d2d20c929e45688d71acbe36784d59560270e0573244e1d5a1

Observation 8c819585-5800-4c7e-8115-d2922fb1edba · outbound

This paper cites Preserving background sound in noise-robust voice conversion via multi-task learning,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Preserving background sound in noise-robust voice conversion via multi-task learning,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.322124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.322124Z digest=sha256:fc047ce3cef96ede0829ed66554aded5650c554d1cb6f5e4a8d95ab9ae977158

Observation c2c04314-f074-4cb2-bc94-9e63231eff77 · outbound

This paper cites Supervised speech separation based on deep learning: An overview,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Supervised speech separation based on deep learning: An overview,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.405886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.405886Z digest=sha256:2502e935e35e93a247043b3ce60999af342578df97c73d44f30d8bd8980a823c

Observation 675636fe-b5ad-4936-ae5f-1fda4a9b4368 · outbound

This paper cites Complex spectral mapping for single-and multi-channel speech enhancement and robust asr,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Complex spectral mapping for single-and multi-channel speech enhancement and robust asr,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.476825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.476825Z digest=sha256:da51d28aa6b7b2533bc48c4f71a720485cd5cd7b38b84096813dac2bc52c05e5

Observation 7636640a-1529-434f-8adf-e9212ee8de43 · outbound

This paper cites Speech enhancement with lstm recurrent neural networks and its application to noise- robust asr,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Speech enhancement with lstm recurrent neural networks and its application to noise- robust asr,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.519961Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.519961Z digest=sha256:07151cb0d086ae8d8a4049f13c8bbdfa0e435c19f027cadec488576633311513

Observation 0919e23d-e6ec-4f44-ba40-89530926a584 · outbound

This paper cites Deep neural networks for speech enhancement and speech recognition: A systematic review,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Deep neural networks for speech enhancement and speech recognition: A systematic review,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.571480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.571480Z digest=sha256:be8adc310268427f25d2e68042eed7cbc7eef652ba8da6a69e446cb735f4ca54

Observation 8e8b811f-e6a9-42c5-adea-5d923a969839 · outbound

This paper cites Attention-based latent features for jointly trained end-to-end automatic speech recognition with modified speech enhancement,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Attention-based latent features for jointly trained end-to-end automatic speech recognition with modified speech enhancement,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.650758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.650758Z digest=sha256:0bb09a0d93998492144b47f7d23c1855762eee702331f55c1e6bc08db5c0321e

Observation 0b3eaef0-d7cd-412e-89ea-c3cb7eabf06b · outbound

This paper cites Phonetic feature encoding in human superior temporal gyrus,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Phonetic feature encoding in human superior temporal gyrus,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.734805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.734805Z digest=sha256:01e371174addd05c0fc8492cc2b8191ed7e8e068d00d4c3f30c73b12d99d6344

Observation 304a0721-9c83-40a0-acce-11fd71c31963 · outbound

This paper cites Representation of temporal sound features in the human auditory cortex,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Representation of temporal sound features in the human auditory cortex,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.791267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.791267Z digest=sha256:8682cabfcfa687357967305f282e15e44b0656e21472f89ac3684e83f9be48ec

Observation 288879da-e5ce-4acf-9ef9-f8d101a3c88a · outbound

This paper cites High Fidelity Neural Audio Compression.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners High Fidelity Neural Audio Compression

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.875056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.875056Z digest=sha256:d41e472bd20609ad8466cd71f5974b88f14dc6a087446f8d64458eff59b7b2aa

Observation a70cc88a-458a-4e35-bd7a-f1b01a6c237e · outbound

This paper cites High-fidelity au- dio compression with improved rvqgan,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners High-fidelity au- dio compression with improved rvqgan,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.951394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.951394Z digest=sha256:7b6a7a1d82b32db6704b6b206fb96b677ddc23751373ce5d60991e4d8c84a672

Observation 143f1e17-f16f-4f2d-b198-ec8a76fb5623 · outbound

This paper cites UniCodec: Unified Audio Codec with Single Domain-Adaptive Codebook.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners UniCodec: Unified Audio Codec with Single Domain-Adaptive Codebook

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:51.994459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:51.994459Z digest=sha256:8b891f1f1e7b45cef626f50d82dad7d5d16eb169077a8b31934e556c36081d1d

Observation 7dfd02ad-5cab-496b-9e6d-33fe1fbad1e6 · outbound

This paper cites NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.044394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.044394Z digest=sha256:6ebd96f8c4d37fe9ad8cc77dd9fa55e6a2c599af43e5f0d015d9931df55dde9f

Observation 9d390e1d-3709-4354-a1b9-ec425eae07b5 · outbound

This paper cites SpeechTokenizer: Unified Speech Tokenizer for Speech Large Language Models.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners SpeechTokenizer: Unified Speech Tokenizer for Speech Large Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.109011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.109011Z digest=sha256:b5a7935f552c311c880786a01596611809bdc9dcb29e7cccac34221d3887bf52

Observation d2aa8661-24bb-4b08-835d-a182292dc5d0 · outbound

This paper cites Moshi: a speech-text foundation model for real-time dialogue.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Moshi: a speech-text foundation model for real-time dialogue

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.181884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.181884Z digest=sha256:bbfc3db9a1be876b65046f249100be45b327bbd631ca1914233fa1180a4d304b

Observation bfb54c71-31ca-47cd-bcd4-8373d8efd4c3 · outbound

This paper cites Dualcodec: A low-frame-rate, semantically-enhanced neural audio codec for speech generation,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Dualcodec: A low-frame-rate, semantically-enhanced neural audio codec for speech generation,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.245258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.245258Z digest=sha256:68611ad01a5ebd5e08963ba0a2cf7feee76cb1f0a44dd6f7422d5eeb6d2921d9

Observation 48c2eb87-df7c-4182-adba-3d8e8ad3215a · outbound

This paper cites Brain plasticity under early auditory deprivation: evidence from congenital hearing-impaired people,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Brain plasticity under early auditory deprivation: evidence from congenital hearing-impaired people,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.315157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.315157Z digest=sha256:65d9624485c7856feb1b6b4ec3893da251af713f114288e154e45723fe26da9a

Observation f3acf3a7-62b6-4c28-9b30-641c07a98d23 · outbound

This paper cites A tutorial on mpeg/audio compression,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners A tutorial on mpeg/audio compression,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.370793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.370793Z digest=sha256:3b60fc6bfe46e19f4c25619cb73766baa198eb6ba1feea4331b69ab072d9ce74

Observation cf877525-a0bb-42e4-b250-0bafe52deed8 · outbound

This paper cites Soundstream: An end-to-end neural audio codec,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Soundstream: An end-to-end neural audio codec,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.454439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.454439Z digest=sha256:cfd4d62b62cf61c590e5139bcc7ddeed7c71410318fee512e5bf67ba607db5f4

Observation eccc40f4-61df-481b-af0b-307de4dfe04a · outbound

This paper cites HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.501429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.501429Z digest=sha256:dbf7a63b4b74f855c2ef1457ae66acb1c1b00e658d4c7a7a3a32d997a1f8afb4

Observation 2d2e3a6a-f07b-4d3d-893a-fa92edcd6f4e · outbound

This paper cites Flow-vae vc: end-to-end flow framework with contrastive loss for zero-shot voice conversion,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Flow-vae vc: end-to-end flow framework with contrastive loss for zero-shot voice conversion,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.561461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.561461Z digest=sha256:0617bea110e4d73118ccdd802a72d47efabf154c1fcf6a7d1fdc0e1040022a68

Observation bbc9e49c-ef69-47f3-9146-bd6ef4331b1a · outbound

This paper cites Hubert: Self-supervised speech rep- resentation learning by masked prediction of hidden units,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Hubert: Self-supervised speech rep- resentation learning by masked prediction of hidden units,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.614490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.614490Z digest=sha256:ffe8278fca00732ce363894aa010edaa5412cb2caf08de8d1ccde8725c80ded8

Observation 5f859916-789d-445a-8fa7-feb35227b639 · outbound

This paper cites wav2vec 2.0: A framework for self-supervised learning of speech representations,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners wav2vec 2.0: A framework for self-supervised learning of speech representations,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.643574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.643574Z digest=sha256:2cee0aaa63b3f64c5f138228e756ce764ef89c478301bb2f21cfd4da9d191c25

Observation fdc269f3-687d-46b1-bfab-0830b3832565 · outbound

This paper cites Wavlm: Large-scale self-supervised pre-training for full stack speech processing,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Wavlm: Large-scale self-supervised pre-training for full stack speech processing,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.726141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.726141Z digest=sha256:ea96e0b7c1c76f21b852d221e3f04316059dbfc2ed04ad42148e1138697a4620

Observation 987f8729-1321-4bef-bda1-a959f73e3152 · outbound

This paper cites SUPERB-SG: Enhanced Speech processing Universal PERformance Benchmark for Semantic and Generative Capabilities.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners SUPERB-SG: Enhanced Speech processing Universal PERformance Benchmark for Semantic and Generative Capabilities

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.768658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.768658Z digest=sha256:667d0d0aa30571684af27fbbcac6356133100728e30356901aa2f45f6feb5912

Observation 308b81e9-2564-4b84-92a7-c3107e254bc9 · outbound

This paper cites RepCodec: A Speech Representation Codec for Speech Tokenization.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners RepCodec: A Speech Representation Codec for Speech Tokenization

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.824678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.824678Z digest=sha256:0e7629bc55912c3e73f5e0d49595998d759b009eb730843d415bca4abd31444d

Observation c4672b52-6489-4b5d-9fc1-6ca23d2aeb53 · outbound

This paper cites Funcodec: A fundamental, reproducible and integrable open-source toolkit for neural speech codec,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Funcodec: A fundamental, reproducible and integrable open-source toolkit for neural speech codec,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.878282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.878282Z digest=sha256:60c563c5f041b2d8f3d3747e1ad610fa4f7426862e43296b8adc71ba3bc34cfa

Observation 5538d06b-21be-4960-9335-9ed3868c6660 · outbound

This paper cites Semanticodec: An ul- tra low bitrate semantic audio codec for general sound,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Semanticodec: An ul- tra low bitrate semantic audio codec for general sound,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.907602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.907602Z digest=sha256:25013e857a17b87c92f1d7954b44d8d9e84aa5530608ec124ff20e66a8df8aaa

Observation fd8a5c3f-1518-47c4-a32f-e3ac2de11812 · outbound

This paper cites Sixty years of frequency-domain monaural speech en- hancement: From traditional to deep learning methods,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Sixty years of frequency-domain monaural speech en- hancement: From traditional to deep learning methods,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:52.960879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:52.960879Z digest=sha256:faf85f1a9fe2f78cfbac00540dcf8e5c5bf3b1f1383fdb182ea606d2bee59941

Observation 0c42ae0b-25e4-466e-a677-449103f0395c · outbound

This paper cites A review of vector quantization techniques,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners A review of vector quantization techniques,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.007585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.007585Z digest=sha256:d51f0261df0fad2a4a7d75b29430affd52c42b13f3cdf4a85006c897ede6ea98

Observation 3ac3e4db-175f-410b-80f3-72ee20d7fc32 · outbound

This paper cites an unresolved cited work.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.066979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.066979Z digest=sha256:8a608b8ee3e90f225c16e9ee7b278a475067405ff3a9f3262c1b300b3423df4a

Observation c7d849c3-24ba-4204-89b7-9c77de90223e · outbound

This paper cites Neural dis- crete representation learning,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Neural dis- crete representation learning,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.096021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.096021Z digest=sha256:05dbb7fb08ee69ef168483a7e753ffc391c04398cfdef19227df85d216c344bc

Observation 3fe8a6d4-6c3c-4f6e-a61e-b88ac4cb3a0d · outbound

This paper cites AISHELL-3: A Multi-speaker Mandarin TTS Corpus and the Baselines.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners AISHELL-3: A Multi-speaker Mandarin TTS Corpus and the Baselines

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.149135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.149135Z digest=sha256:446fde46aea3c4533d668e0f509286925d5ff3f9b32bdae8caf237d9d40efb58

Observation 0ff7203d-9946-4890-96b3-d0be7e7a1bb0 · outbound

This paper cites LibriTTS: A Corpus Derived from LibriSpeech for Text-to-Speech.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners LibriTTS: A Corpus Derived from LibriSpeech for Text-to-Speech

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.193305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.193305Z digest=sha256:8b168161058de193593a723fbd1a6504f519e26bb2221dfade2e81ad884eb3ab

Observation b08ec887-b932-4c57-99aa-ded01908740d · outbound

This paper cites The voice bank corpus: Design, collection and data analysis of a large regional accent speech database,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners The voice bank corpus: Design, collection and data analysis of a large regional accent speech database,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.249521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.249521Z digest=sha256:46d7dea21861839486cdaa10f5d6bd1374595be5bc935cdc88d1ab0de6936491

Observation 0c2628a5-a575-43e7-ad54-c9f4e6319597 · outbound

This paper cites The design for the wall street journal- based,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners The design for the wall street journal- based,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.336254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.336254Z digest=sha256:3b3435172803057bdea8f033fb0c396c9ddf79ac429e89935abe5014cf8558d9

Observation 2ade5bfd-d711-4079-85fb-10545fbdb348 · outbound

This paper cites Esc: Dataset for environmental sound classification,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Esc: Dataset for environmental sound classification,

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.415756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.415756Z digest=sha256:def4493fe61b44e86fcda9f1b8aa7b31032196cd3324b60fecaf04169740d398

Observation ed606094-3758-4290-9f32-7b616663d3b6 · outbound

This paper cites The INTERSPEECH 2020 Deep Noise Suppression Challenge: Datasets, Subjective Testing Framework, and Challenge Results.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners The INTERSPEECH 2020 Deep Noise Suppression Challenge: Datasets, Subjective Testing Framework, and Challenge Results

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.473571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.473571Z digest=sha256:36925e0bca50cb2ea48d9cc479632673ec3cd74c70c156307f72e434cd2afbf1

Observation 03832718-d907-4723-86a0-1af4f5511876 · outbound

This paper cites First stereo audio source separation evaluation campaign: data, algorithms and results,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners First stereo audio source separation evaluation campaign: data, algorithms and results,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.530768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.530768Z digest=sha256:740bfebc858235f4ca035df54779eb71555a91d4039c03ba56a12bccde68e2ff

Observation 9d4956ab-1ef7-4e02-b000-f4c7185a9688 · outbound

This paper cites Robust speech recognition via large-scale weak supervision,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Robust speech recognition via large-scale weak supervision,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.584010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.584010Z digest=sha256:3f938ccba178e91139051ebe929d6ad02ec4e7b0c92abe1ad7733c01f3279155

Observation 7f27f1f1-edb8-4ca8-b669-ec4ca13cbc3d · outbound

This paper cites Dnsmos: A non-intrusive perceptual objective speech quality metric to evaluate noise suppressors,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Dnsmos: A non-intrusive perceptual objective speech quality metric to evaluate noise suppressors,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.644016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.644016Z digest=sha256:262ca8f7d1786388799c480a7dd79afc2fd3aaa61b5ee1332fceb79a99355101

Observation 0df59a6a-045c-4892-b56a-b72933812da5 · outbound

This paper cites Inter-subnet: Speech enhancement with subband in- teraction,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Inter-subnet: Speech enhancement with subband in- teraction,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.669090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.669090Z digest=sha256:647c98d8368e33390d3ddf90edb97bd85c7f5fde6f521897decd9e35ca906c3f

Observation ce48057c-768d-4a47-ac92-b8fc049ae857 · outbound

This paper cites Storm: A diffusion-based stochastic regeneration model for speech enhancement and dereverberation,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Storm: A diffusion-based stochastic regeneration model for speech enhancement and dereverberation,

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.722918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.722918Z digest=sha256:33b2986c9bd131e7eb9816fdaa1f4b662af1efc2277170c48fcab0f3676c7939

Observation 4c826304-4843-4d75-8b72-968118a0e324 · outbound

This paper cites Selm: Speech enhancement using discrete tokens and language mod- els,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Selm: Speech enhancement using discrete tokens and language mod- els,

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.774367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.774367Z digest=sha256:7b4ed39d5959954e66797a8cd40a6f767ee5f7c51770d3022c1145a2c8ff8b38

Observation 5ca5a13c-4aa2-47f2-825c-8cc2979366aa · outbound

This paper cites Attention is all you need,.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Attention is all you need,

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.831514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.831514Z digest=sha256:70444dcdb3ed0806c1b46c70d3135333ec812a892818d4fda2ee9928fc122e66

Observation 34b8bdc9-f242-4121-ac11-2ed83fb76510 · outbound

This paper cites PolySpeech: Exploring Unified Multitask Speech Models for Competitiveness with Single-task Models.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners PolySpeech: Exploring Unified Multitask Speech Models for Competitiveness with Single-task Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.886688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.886688Z digest=sha256:8e566ff6ff186a39f7c69f4bf56d90d1b5926f9f2d8d199366968b00a3c1d6e8

Observation f489a456-69c3-44f8-9163-9a4a535cf71a · outbound

This paper cites Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.944636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.944636Z digest=sha256:3ceaed2ea23b0cdfb69f0978b7ebddbb2b75f52e8e6698cc684708a0a366a205

Observation c6eee5b8-c5b0-4fd9-b5d5-aeff362fc939 · outbound

This paper cites Decoupled Weight Decay Regularization.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners Decoupled Weight Decay Regularization

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:53.994131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:53.994131Z digest=sha256:3807141b5caeac5d5a5f2209b2393e967a2ac05aeda88847fbb178880f9f1d3f

Pith citing papers

Observation 418f8ce1-a402-44c1-90ce-25eb16f76ea9 · inbound

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners cites this paper.

DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners DeCodec: Rethinking Audio Codecs as Universal Disentangled Representation Learners

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:50.680596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:50.680596Z digest=sha256:4a5e1be06b53e43e10b948997c1ef174729e76732455fa0a8038ada15e377487