Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T10:37:03.781195Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 4 inbound Pith citation observations for arXiv:2502.02942.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T10:37:03.781195Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-03T14:20:49.180965Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-19T13:12:18.207183Z
48 of 48 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 06c49604-3c9a-4dd1-8f00-73114a9b30ea · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling APCodec: A Neural Audio Codec with Parallel Amplitude and Phase Spectrum Encoding and Decoding
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f9c9a7a-a4bb-4285-bb43-bec73dc56985 · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling Unresolved cited work
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 31249799-37a2-4892-b4d0-c234593b4b53 · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling Unsupervised Cross-lingual Representation Learning for Speech Recognition
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fbe90fbe-7b1b-4277-9412-1d6f0bcfd610 · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling High Fidelity Neural Audio Compression
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92ae16bd-9663-40d1-8f3b-627c34399576 · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling PolyVoice: Language Models for Speech to Speech Translation
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e841061e-bef3-4a40-b3aa-58a2eca8e5f3 · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 48a5a75d-e41e-476f-9180-c9c63600d22c · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling Variational autoencoder for speech enhancement with a noise-aware encoder
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 12139086-bc1d-4c2b-a568-7ebed5a690f8 · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling Fullsubnet: A full-band and sub-band fusion model for real-time single-channel speech enhancement
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f2468bc9-e5a0-4d90-89f6-93ce67188133 · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling Hubert: Self-supervised speech representation learning by masked prediction of hidden units
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9e98d6b6-d2a6-4112-9438-43f1561609be · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling ReVISE: Self-Supervised Speech Resynthesis with Visual Input for Universal and Generalized Speech Enhancement
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d5073d6b-362b-4ea3-9cd6-5d5a8f4957a7 · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling Language-Codec: Bridging Discrete Codec Representations and Speech Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69673469-7183-4116-9255-789da0f976fc · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling Libri- light: A benchmark for asr with limited or no supervision
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 22716251-6203-4a62-82f4-731b6ca1ba61 · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling A study on data augmentation of reverberant speech for robust speech recognition
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 949cdf6a-c968-43eb-a060-482e25189ce8 · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling Lan- guage models as controlled natural language semantic parsers for knowledge graph question an- swering
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 12f57c02-ea07-424d-9f4b-e8d8571dd0f1 · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling Single-Codec: Single-Codebook Speech Codec towards High-Performance Speech Generation
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 873e4451-21d3-4a4d-a091-16da579da7c4 · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling Noise Tokens: Learning Neural Noise Templates for Environment-Aware Speech Enhancement
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 76e3ce1d-53e6-45d8-9dea-2b7133e179f0 · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling VoiceFixer: Toward General Speech Restoration with Neural Vocoder
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b44fc7fe-1069-487b-b416-2d2cd9e14006 · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling SemantiCodec: An Ultra Low Bitrate Semantic Audio Codec for General Sound
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21fbec8c-b87a-481f-b3a2-7836e1134c6f · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling Speaker independence of neural vocoders and their effect on parametric resynthesis speech enhancement
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 17a3139c-8abc-47f6-aac1-38c28d6113b0 · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling Finite Scalar Quantization: VQ-VAE Made Simple
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6acf754b-dbc7-405d-9b67-22473d9c8e3c · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling Dnsmos p
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 55d42621-87a4-4fd1-9fda-f5cc23ccb46d · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling Fewer-token neural speech codec with time-invariant codes
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fe002d0a-cc8c-41e1-a026-f26f74bb9db9 · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling UTMOS: UTokyo-SaruLab System for VoiceMOS Challenge 2022
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c76a641e-ac99-424b-84f4-67a3546f93d4 · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling Universal Score-based Speech Enhancement with High Content Preservation
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation acc77a20-f55a-4746-86b1-74de8fce3ff3 · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling Vocos: Closing the gap between time-domain and Fourier-based neural vocoders for high-quality audio synthesis
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b30ab931-29a9-4a06-b78d-1c13f78ceb4b · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling The voice bank corpus: Design, collection and data analysis of a large regional accent speech database
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 464655a4-19bf-40e4-abcb-5fcbbed33fcb · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling WHAM!: Extending Speech Separation to Noisy Environments
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afa3289d-9183-4eda-a8c8-8ba17ff32094 · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling Audiodec: An open-source streaming high-fidelity neural audio codec
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1322981d-a653-44ed-a58f-ded95262096f · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9203ab6-3161-42cd-8002-275405161514 · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling LibriTTS: A Corpus Derived from LibriSpeech for Text-to-Speech
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0c6c1ff-b73b-4952-93c3-ec6c44925d76 · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling Speaking in Wavelet Domain: A Simple and Efficient Approach to Speed up Speech Diffusion Model
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5cc50158-7a0e-45a4-ac35-9bf2cd84f3ad · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling SpeechTokenizer: Unified Speech Tokenizer for Speech Large Language Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ef73369-47ff-47a4-8da5-f840c90aaa53 · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling We discuss several common questions for the design of GenSE
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation abf7d858-0d37-46c7-ad03-e6066347135e · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling On the other hand, we chose XLSR as the semantic extractor due to the key advantage of its multilin- gual speech representation capabilities
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c74d7a16-07f5-49f5-bded-6d167a258413 · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling To address this limitation, we also conducted an ABX test to assess the perceptual quality of the enhanced speech compared to clean speech, as shown in Figure
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3d73fd3b-356c-4181-a7f4-f86cd57c86b2 · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling No Preference
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 36e4c198-c62a-4e1c-a0e0-5a05278cb404 · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling We also observe that finite scalar quantization (FSQ) (Mentzer et al., 2023a) demonstrates lower reconstruction quality
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 12ea5c2a-2b9b-40a2-965f-1b5b29361e4c · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling We employ the AdamW optimizer with a learning rate of 1e-4 to optimize the codec model
Reference 128
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e5545a6d-5a73-46d4-8d01-d56cb98ebb8f · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers
Reference 2013
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a752c229-8f33-415f-8e9c-fcfa022b4eb1 · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling Interspeech 2021 Deep Noise Suppression Challenge
Reference 2015
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df688006-3bb2-455e-8e55-212d5ed86908 · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling Funcodec: A fundamental, reproducible and integrable open-source toolkit for neural speech codec
Reference 2016
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation dc161ac4-b595-45a0-b79e-f979fe65ab21 · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling NU-GAN: High resolution neural upsampling with GAN
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c4c018e-0e23-4898-976b-24d6cfb1cec1 · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling Self-Supervised Speech Quality Estimation and Enhancement Using Only Clean Speech
Reference 2019
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9c711ee6-7141-46e4-a003-909bbdfd5f41 · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling Fish-Speech: Leveraging Large Language Models for Advanced Multilingual Text-to-Speech Synthesis
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32685557-cfe0-42df-8a48-0100d4aa5dea · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling Metricgan: Generative adversarial net- works based black-box metric scores optimization for speech enhancement
Reference 2021
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 48941f34-dddc-46bc-9f8d-d1e66e99fbc8 · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling SoundStorm: Efficient Parallel Audio Generation
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72b5b59b-50fd-4783-9f12-2e1ab27da007 · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31880e8c-2126-42e5-8443-ce43a3992637 · outbound
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling Unsupervised speech en- hancement using dynamical variational autoencoders.IEEE/ACM Transactions on Audio, Speech, and Language Processing, 30:2993–3007,
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0934fbbf-1b5d-4e48-a3c9-4affd5419835 · inbound
SwitchCodec: A High-Fidelity Nerual Audio Codec With Sparse Quantization GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9fd7e768-dd26-4c45-b1b6-e75b77c43489 · inbound
GenTSE: Enhancing Target Speaker Extraction via a Coarse-to-Fine Generative Language Model GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32b38a54-4b1a-46c1-afd9-a606c2a2009d · inbound
UniPASE: A Generative Model for Universal Speech Enhancement with High Fidelity and Low Hallucinations GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0edcb4b6-3f4a-40e9-b27e-001503530b53 · inbound
UniPASE: A Generative Model for Universal Speech Enhancement with High Fidelity and Low Hallucinations GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.