Pith. sign in

Paper Citation Record · LEDGER

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec

As of 21 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 1 inbound Pith citation observation for arXiv:2608.12082.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.12082 v1

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T00:22:15.154844Z

measured 45 of 45 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T00:22:14.947985Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-16T00:22:15.394980Z

Reference resolution

44 of 44 outbound references displayed

  • verified exact1
  • verified fuzzy35
  • unresolved7
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e2548229-44f6-44ec-9f00-ec511af72113 · outbound

This paper cites an unresolved cited work.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-16T00:22:16.062546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:14.942116Z digest=sha256:9de2596f38bd6e1d994b27e383e84b21b9544a42a834ac1de2b11da11c5b2157

Observation ac0db359-6fad-4f44-92a9-7f49371b8aed · outbound

This paper cites Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-16T00:22:15.402867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:14.947985Z digest=sha256:b5d40c60d8a8dfd51ca64898d8d207a7fc5f766fb60e2f6aaf4cfdc8f0a2f18e

Observation 00ca68b4-a7de-40af-a5e4-c1d3bee9c0d7 · outbound

This paper cites Networks In the proposed SE models, the pretrained 16 kHzDAC[25] and WavLM[26] weights are adopted.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec Networks In the proposed SE models, the pretrained 16 kHzDAC[25] and WavLM[26] weights are adopted

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:16.023750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:14.958606Z digest=sha256:34d833801e20886e99b634121cb316f0b60d704c12ef86efaa6b118bfe78c912

Observation 8122321e-f33b-4445-bda3-da5b526a152e · outbound

This paper cites C... ”) are consistently outperforming their discrete-domain counterparts (methods named “D.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec C... ”) are consistently outperforming their discrete-domain counterparts (methods named “D

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:15.990634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:14.968671Z digest=sha256:16cd755a6dca96deff5a59ab589ffc55be2d71b1b1de1e17ee7b3bc6233d8bc5

Observation bd9652df-f5ba-4def-a2a0-eea0ab2b4965 · outbound

This paper cites We are the first to de- liver a comprehensive and fair synopsis of the paradigms by employ- ing both non-intrusive and intrusive metrics.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec We are the first to de- liver a comprehensive and fair synopsis of the paradigms by employ- ing both non-intrusive and intrusive metrics

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:15.974349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:14.973709Z digest=sha256:11db2215e496cfbea9d67a594aee81add9fccd5eb9c8f678f9877e1b38229c0f

Observation 5805f68a-0880-4c8f-b279-0a7ad01e96b5 · outbound

This paper cites an unresolved cited work.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-16T00:22:15.956646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:14.978330Z digest=sha256:b853c0aa1c9f959cf930d3979efe9171b9002f0ea2adf936ed592529702c46ce

Observation 6def8921-f3cf-4dc9-b242-b8a5b21f5389 · outbound

This paper cites Speech Enhancement Using Continuous Embeddings of Neural Audio Codec,.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec Speech Enhancement Using Continuous Embeddings of Neural Audio Codec,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:15.875410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:15.012970Z digest=sha256:b94f92904bf03d4aa91726ea9c954ea9477563b721f67479dec05f5396ffd8d9

Observation e5beb3b0-5980-42c6-a1cc-0e144eb5e2d6 · outbound

This paper cites AnyEnhance: A Unified Generative Model with Prompt-Guidance and Self-Critic for V oice En- hancement,.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec AnyEnhance: A Unified Generative Model with Prompt-Guidance and Self-Critic for V oice En- hancement,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T00:22:15.018061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:22:15.018061Z digest=sha256:e63949a9f7e57539c5f7852c8340e7e5b302c946b2b7fa38a0af5700a2670480

Observation 47592323-8a9b-4fdd-88bf-eced6bf8cdac · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec LLaMA: Open and Efficient Foundation Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T00:22:14.983206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:22:14.983206Z digest=sha256:d63ee51a33c1add4f585a6a9abe3271be7276380e36ef4ccef7fe16e246ac214

Observation b7b87fea-d3a7-4e44-acfc-16c87c9d6b14 · outbound

This paper cites GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling,.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:15.939941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:14.987662Z digest=sha256:de83a262c6c70c727514b1922dcec11921596fcb7ba13ef7f7869f40eb0b4068

Observation 9737c5ea-1406-45af-a768-3b463b472f2f · outbound

This paper cites LLaSE-G1: Incentivizing Generalization Capability for Llama-based Speech Enhancement,.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec LLaSE-G1: Incentivizing Generalization Capability for Llama-based Speech Enhancement,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:15.923043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:14.992217Z digest=sha256:8a86e77a09b2608f8b13a1d8aca1b54dc4a3d8270ab1752cd6420cb325619534

Observation bcf74c10-fa97-4f06-a083-916e468ca1d9 · outbound

This paper cites UniSE: A Unified Framework for Decoder-Only Autoregressive LM-Based Speech Enhancement.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec UniSE: A Unified Framework for Decoder-Only Autoregressive LM-Based Speech Enhancement

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T00:22:14.997112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:22:14.997112Z digest=sha256:3e02b138abd33105c86c7bfc213809a92083450b4607bd514992106012909803

Observation 5edebc01-c165-42a5-92b4-383163e90f1e · outbound

This paper cites Modeling Strategies for Speech En- hancement in the Latent Space of a Neural Audio Codec,.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec Modeling Strategies for Speech En- hancement in the Latent Space of a Neural Audio Codec,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:15.907608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:15.002641Z digest=sha256:b493e07a22aaf53eb0c9ce8249f082d3a87b8700f15e82bb8935186f54d82780

Observation 624a359c-9951-4bba-bc89-a6f2322753f2 · outbound

This paper cites an unresolved cited work.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-16T00:22:16.042114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:14.953402Z digest=sha256:2871dd742928bff9153941dea73ebf7aa57bc099a175b46233cc46d6c26594d8

Observation ba856785-120b-4c7d-a054-f33afb372a2e · outbound

This paper cites High-Fidelity Speech Enhance- ment via Discrete Audio Tokens,.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec High-Fidelity Speech Enhance- ment via Discrete Audio Tokens,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:15.891220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:15.007609Z digest=sha256:84ae83c764787673f8e7c19e6a9ce06687fcbf57970b5d77d65d4baae0a2efe3

Observation b2ea674e-b48a-43b1-9c28-9f1e58493520 · outbound

This paper cites Codec Does Matter: Exploring the Seman- tic Shortcoming of Codec for Audio Language Model,.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec Codec Does Matter: Exploring the Seman- tic Shortcoming of Codec for Audio Language Model,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:15.755519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:15.056628Z digest=sha256:2c5912a8f7da2be250dd26663d25e4cd397edebe1131bb162bfa260f63dc302f

Observation 906a7c7e-d675-4995-b2b8-84418b437623 · outbound

This paper cites Genhancer: High-Fidelity Speech Enhance- ment via Generative Modeling on Discrete Codec Tokens,.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec Genhancer: High-Fidelity Speech Enhance- ment via Generative Modeling on Discrete Codec Tokens,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:15.858802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:15.022765Z digest=sha256:60c64946e448a693f9c2696c432517a1fc7eb96eb963eb91a5f01028cdb449e5

Observation 58bef9d2-5649-4bbb-be02-b1aab53d4020 · outbound

This paper cites DisContSE: Single-Step Diffusion Speech Enhancement Based on Joint Discrete and Continuous Embed- dings,.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec DisContSE: Single-Step Diffusion Speech Enhancement Based on Joint Discrete and Continuous Embed- dings,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:15.842906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:15.027639Z digest=sha256:5f92ca8421a68d6b7e12136427c5254b5d18bf205a835468dcaa4c7d06d22a31

Observation 369a354c-3a12-4605-9891-81db50515939 · outbound

This paper cites Flow Matching for Generative Model- ing,.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec Flow Matching for Generative Model- ing,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:15.826540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:15.032659Z digest=sha256:e9ce9b6e95762768d3d3cdf391ae531b9197d5ce15cf4882cf54594a8220b663

Observation c34d593d-699d-40d9-9afc-dca70f70e02e · outbound

This paper cites FlowDec: A Flow-Based Full-Band General Audio Codec with High Perceptual Quality,.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec FlowDec: A Flow-Based Full-Band General Audio Codec with High Perceptual Quality,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:15.809078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:15.037522Z digest=sha256:e6e0e988624df1d24d84eddf108bcf03f100c60d86a1b92c4676af7a4418f057

Observation 894c424b-7db8-4cfb-8e7a-37b5e90e3464 · outbound

This paper cites High Fidelity Neural Audio Com- pression,.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec High Fidelity Neural Audio Com- pression,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:15.790923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:15.042074Z digest=sha256:017b7211981283c9e68a9116652e20c5020680df81ca28665d32d1fc9d9ca4ae

Observation 487d4ca8-2add-40de-acdf-e631c1b04246 · outbound

This paper cites High-Fidelity Audio Compression with Improved RVQGAN,.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec High-Fidelity Audio Compression with Improved RVQGAN,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:15.772457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:15.046748Z digest=sha256:3f3d356fcaf2195c441252f8bd06c5f5c0cc2dd23bd30300648a970dc39eb78e

Observation 1f4fc372-c793-4fc4-96fd-5ad5fa39434e · outbound

This paper cites Spark-TTS: An Efficient LLM-Based Text-to-Speech Model with Single-Stream Decoupled Speech Tokens.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec Spark-TTS: An Efficient LLM-Based Text-to-Speech Model with Single-Stream Decoupled Speech Tokens

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-16T00:22:15.051545Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:22:15.051545Z digest=sha256:66fd639bd6747774570f182e8b64ee3028df5d61546b844df69850ed610b3ddb

Observation 52049a1a-51e5-4b43-8fcb-c6f8572394bc · outbound

This paper cites Understanding Straight-Through Esti- mator in Training Activation Quantized Neural Nets,.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec Understanding Straight-Through Esti- mator in Training Activation Quantized Neural Nets,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:15.623089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:15.096307Z digest=sha256:2ea1e6758fa563656145c7d39a6138874c726d740d4721c3bdf5b4e5c058f7a0

Observation 204345bb-4593-4176-9ede-4020e25d6b0d · outbound

This paper cites Residual Quantization with Implicit Neural Codebooks,.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec Residual Quantization with Implicit Neural Codebooks,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:15.740047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:15.061410Z digest=sha256:8a93f772e0712871f6279d8d8e043cac3a8296d6e7ce55bc522bb8f169529043

Observation 784a394e-2b0d-410c-b458-64bafdaa2673 · outbound

This paper cites WavLM: Large-Scale Self-Supervised Pre-Training for Full Stack Speech Processing,.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec WavLM: Large-Scale Self-Supervised Pre-Training for Full Stack Speech Processing,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:15.724094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:15.066461Z digest=sha256:1029995f5e0bfc3ca57c05b377d81c28ae675d1385fb087193cc786abfc0c76f

Observation 06ec2a8e-080b-4a30-b149-35253cb7da8d · outbound

This paper cites MaskGIT: Masked Generative Image Transformer,.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec MaskGIT: Masked Generative Image Transformer,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:15.707831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:15.071584Z digest=sha256:4d746f1b7ad4a9b336bedbac205d8814203fdc20e74f875c1df473763a1225fc

Observation c56653f4-cf3e-4034-ad1b-0ee670d3a829 · outbound

This paper cites A Consolidated View of Loss Func- tions for Supervised Deep Learning-Based Speech Enhance- ment,.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec A Consolidated View of Loss Func- tions for Supervised Deep Learning-Based Speech Enhance- ment,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:15.690690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:15.076586Z digest=sha256:3a154c35172c757ad5f2503d8d1c141af6ad85a76164637721b201756dd67bdc

Observation 482c7594-a71e-48d2-afb2-a02f01812cdd · outbound

This paper cites an unresolved cited work.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec Unresolved cited work

Reference 29

Resolution
malformed identifier
raw_fallback, observed 2026-08-16T00:22:16.007900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:14.963425Z digest=sha256:1d29a17a34386450e20e3eae473c12bcc9b76d026851aed5b805ca39afe2c015

Observation 1cc15669-2248-48e0-b877-a56ac6c71d78 · outbound

This paper cites EffCRN: An Efficient Convolutional Re- current Network for High-Performance Speech Enhancement,.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec EffCRN: An Efficient Convolutional Re- current Network for High-Performance Speech Enhancement,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:15.674040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:15.081411Z digest=sha256:f4adb4989a0635bcc5e48bd14c1b0916843701fda7c6e2a24bd10518fa6c2e00

Observation bc3bb79a-19a5-4efb-a6ec-ee666c2aa898 · outbound

This paper cites Loss Function Inspired by the PESQ Score,.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec Loss Function Inspired by the PESQ Score,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:15.657046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:15.086320Z digest=sha256:a502a0086d257836a041713c1f1f2335f42efb6ef2c4fb8a6e6815d26f8a85f9

Observation c618ff79-e9b9-4841-9ee1-cd6ec182afc9 · outbound

This paper cites PyTorch Implementation of STOI ,.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec PyTorch Implementation of STOI ,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:15.640514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:15.091242Z digest=sha256:221ff9e8b2f317f14af05ac0dfb3cd586d78138a767530f6a1a1c78e7228880b

Observation 0d79e9ab-6245-4aad-b935-d81e2020c01a · outbound

This paper cites Descript Audio Codec (.dac): High-Fidelity Audio Compression with Improved RVQ- GAN,.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec Descript Audio Codec (.dac): High-Fidelity Audio Compression with Improved RVQ- GAN,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:15.607065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:15.101268Z digest=sha256:04945def4717d1e3390e16608346610308e8fecdaca8878b8bd8c46a98b8543a

Observation d5566bd3-5d2c-4994-aba5-54c565fd8ca8 · outbound

This paper cites WavLM-Large,.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec WavLM-Large,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:15.590387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:15.106054Z digest=sha256:b510139a65f663d43d95418fa90234925e9d7b2ea844c1096c22244c9d4e76a5

Observation 431f2714-505a-432e-bac7-9c346f6ee5a3 · outbound

This paper cites Attention Is All You Need,.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec Attention Is All You Need,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:15.573245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:15.111349Z digest=sha256:f64727193f19b5ee5dc9d5269596befa7e2f85b25472fb785014556a49878ec2

Observation 71241c32-39a6-4991-8cdb-eeec21fdced8 · outbound

This paper cites Interspeech 2025 URGENT Speech En- hancement Challenge,.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec Interspeech 2025 URGENT Speech En- hancement Challenge,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:15.556356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:15.116159Z digest=sha256:bb02cc1b38223149da447ce2d9beb063f9eabd8a5f135b6c9924ece72d4fa297

Observation 7bc8193f-34a9-445e-af79-a1e71569a62b · outbound

This paper cites Common V oice: A Massively- Multilingual Speech Corpus,.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec Common V oice: A Massively- Multilingual Speech Corpus,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:15.540227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:15.121044Z digest=sha256:411ac267c62acba5f5b8f65e99d6475341305aa0b8a08c5daf533bb06e178998

Observation 632c197f-80b8-481c-a11a-985af16d1fae · outbound

This paper cites P .56: Objective Measurement of Active Speech Level, International Telecommunication Union, Telecommu- nication Standardization Sector (ITU-T), Dec.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec P .56: Objective Measurement of Active Speech Level, International Telecommunication Union, Telecommu- nication Standardization Sector (ITU-T), Dec

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:15.524537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:15.125825Z digest=sha256:49c551097c7ab6d29cb607620aad10dbba28bdb3e74b1b6aae3c548c0ac19a90

Observation a33432d6-cf74-4312-9fb4-c4ff87d94223 · outbound

This paper cites Perceptual Evaluation of Speech Quality (PESQ)—A New Method for Speech Quality Assessment of Telephone Networks and Codecs,.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec Perceptual Evaluation of Speech Quality (PESQ)—A New Method for Speech Quality Assessment of Telephone Networks and Codecs,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:15.509281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:15.130624Z digest=sha256:8f027df35b153f62c910ea84d2e0c5897dfb35be62653a0df312cd67712347be

Observation d5bd8283-df88-46df-a8a6-3b2f3f01b669 · outbound

This paper cites An Algorithm for Predicting the Intel- ligibility of Speech Masked by Modulated Noise Maskers,.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec An Algorithm for Predicting the Intel- ligibility of Speech Masked by Modulated Noise Maskers,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:15.490807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:15.135625Z digest=sha256:f7ade8f86d4c6c6086b5d5306b19997d7a9f17033a42e58738b70fa0d6860eee

Observation 2902e11d-b382-46a9-9b4c-d7fa28b133a3 · outbound

This paper cites DNSMOS P. 835: A Non- Intrusive Perceptual Objective Speech Quality Metric to Eval- uate Noise Suppressors,.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec DNSMOS P. 835: A Non- Intrusive Perceptual Objective Speech Quality Metric to Eval- uate Noise Suppressors,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:15.473450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:15.140206Z digest=sha256:010949ac8d8b0b1ac4d02008c472990242a845359f41d70f1a1a25347276348d

Observation 519ef48a-0930-41bd-93eb-d53ccef1cf7f · outbound

This paper cites NISQA: A Deep CNN-Self-Attention Model for Multidimensional Speech Quality Prediction with Crowdsourced Datasets,.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec NISQA: A Deep CNN-Self-Attention Model for Multidimensional Speech Quality Prediction with Crowdsourced Datasets,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:15.456687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:15.144743Z digest=sha256:c976461f59f1fbb5932a0b3c245be94c9d7419ad3185429ee84aa4fffaf6c324

Observation 15b66ef5-562e-452b-87f9-328c3e9b3f80 · outbound

This paper cites UTMOS: UTokyo-SaruLab System for V oiceMOS Challenge 2022,.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec UTMOS: UTokyo-SaruLab System for V oiceMOS Challenge 2022,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:15.438156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:15.149709Z digest=sha256:b75dc281b810a6242b795b24938d253b9a70d6305ba6ae10689a973fded8c8c0

Observation 7d2d90cd-a032-468f-9817-23d8c452b7c5 · outbound

This paper cites Evaluation Metrics for Generative Speech Enhancement Methods: Issues and Perspectives,.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec Evaluation Metrics for Generative Speech Enhancement Methods: Issues and Perspectives,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:22:15.420878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:15.154844Z digest=sha256:4e59f456472738d931c24cc862a60665cdc57ec78eefa04d48b77454210580e2

Pith citing papers

Observation ac0db359-6fad-4f44-92a9-7f49371b8aed · inbound

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec cites this paper.

Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec Rethinking Language Model-Based Generative Speech Enhancement in the Latent Space of a Neural Audio Codec

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-16T00:22:15.402867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T00:22:14.947985Z digest=sha256:b5d40c60d8a8dfd51ca64898d8d207a7fc5f766fb60e2f6aaf4cfdc8f0a2f18e