Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T23:11:24.534728Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 7 inbound Pith citation observations for arXiv:2506.19398.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T23:11:24.534728Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T23:11:24.362189Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T13:08:08.743449Z
45 of 45 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 69262a98-01cc-4375-8be0-ce2617879210 · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment While crucial for these applications, ac- curately processing speech is challenged by the often degraded quality of real-world audio
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 714fd3b6-e2d1-45c8-b151-2d020951eb2c · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffbe1fee-e561-42d0-b535-2900b9f5f465 · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment Training strategies 3.1.1
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 72b8ed4b-36ef-4577-8a86-85b840ff6436 · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment Beyond the presented evaluations, ClearerV oice- Studio is available for live demos on HuggingFace and Mod- elScope, enabling users to experiment with real-world record- ings
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 641f26e5-0e0d-4eb6-aae4-80b3166b30ec · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment Deep learning for audio signal processing,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 84304d9c-cd5d-423d-9c93-b1b942667db1 · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment Mamba in Speech: Towards an Alternative to Self-Attention
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6082d4d3-b631-4c69-ad1b-0d83a76af01a · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment DeepMMSE: A deep learning approach to mmse-based noise power spectral density estimation,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2a35cea7-d23a-4042-93f2-bdc8a40e17dd · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment SpeechBrain: A general-purpose speech toolkit,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 651b4301-b4e2-413f-8fbd-95a5ae5587c8 · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment AudioSR: Versatile audio super-resolution at scale,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0a8da161-6948-4863-9ea9-090d76f982a9 · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment ESPnet: End-to-end speech processing toolkit,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d4baf6c7-b9da-4399-9bab-d6526be4a8df · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment Summary on the multimodal information-based speech processing 2023 challenge,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 51bb2732-9407-4c47-a76c-6aebd1a7ea79 · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment As- teroid: the PyTorch-based audio source separation toolkit for re- searchers,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 95a7cd92-8c58-4079-bed2-62ae14bd37ca · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment DeepFilterNet: Perceptually motivated real-time speech en- hancement,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cc184939-26b3-4313-8d7f-57dec752ac13 · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment Hifi-SR: A unified generative transformer-convolutional adversarial net- work for high-fidelity speech super-resolution,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b444fbaa-fe93-4bfe-8501-0343628453ef · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment FlowA VSE: Ef- ficient audio-visual speech enhancement with conditional flow matching,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 60e16eda-03ee-4f93-af38-77890af77535 · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment FRCRN: Boosting feature representation using frequency recurrence for monaural speech enhancement,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f9a92ae9-a8d1-46e3-9c7c-c7226a044498 · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment MossFormer2: Combin- ing transformer and rnn-free recurrent network for enhanced time- domain monaural speech separation,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3f76367b-dfd3-42af-8818-5c8d59f4bbf4 · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment MossFormer: Pushing the performance limit of monaural speech separation using gated single-head trans- former with convolution-augmented joint self-attentions,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5658879f-5b6d-49b8-a6d9-67f991f513c3 · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment NeuroHeed: Neuro-Steered Speaker Extraction using EEG Signals
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation de92d52c-4bf6-4860-8822-ba96854e3fe5 · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment HiFi-GAN: Generative adversarial networks for efficient and high fidelity speech synthesis,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation aabc8ee1-2c97-4e64-aa86-a8da807dc826 · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment Scenario-aware audio-visual TF- Gridnet for target speech extraction,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d8c239ae-1c8d-4c32-859d-5764628568ce · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment Speaker extraction with co-speech gestures cue,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bac71848-182d-4296-86e5-a0f824999631 · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment SpEx+: A complete time domain speaker extraction network,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ec631b81-503b-41f1-8a50-b5ab29dfa231 · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment DCCRN+: Channel-wise subband dccrn with snr estimation for speech enhancement,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 56a28c20-a421-4a64-9fef-8351e699a4d6 · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment The interspeech 2020 deep noise suppression challenge: Datasets, subjective testing framework, and challenge results,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3f12de00-69f9-4ec6-9155-361df8278134 · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment CSTR VCTK Corpus: English multi-speaker corpus for cstr voice cloning toolkit (version 0.92),
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 659d3a06-5962-4ebf-b2e5-03c71bfa0038 · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment Audio set: An ontology and human-labeled dataset for audio events,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 95fdaf25-0bc1-40ec-a534-9a7e5ab7d920 · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment DEMAND: a collection of multi-channel recordings of acoustic noise in diverse environments,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 84ac29bb-be00-47d6-9960-e826474b4b6e · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment Phase- sensitive and recognition-boosted speech separation using deep recurrent neural networks,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ae34a824-1169-40b3-8f35-5e3e3c074c7b · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment A mask free neural network for monaural speech enhancement,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 33a9f56d-411a-468a-9246-0760cb623b87 · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment TridentSE: Guiding speech enhancement with 32 global tokens,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5ab46bd6-5543-4693-94c5-762f92a30bea · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment LibriTTS: A corpus derived from librispeech for text- to-speech,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 37ab05c7-7c56-42e3-8b87-94489e6e4ff0 · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment Explor- ing strategies for training deep neural networks,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1a55352d-409e-441e-8e99-b07c8b85bb82 · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment An efficient encoder-decoder archi- tecture with top-down attention for speech separation,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a3042b7b-1c2e-4e49-9fcb-38ad66742e36 · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment CMGAN: Conformer-based metric gan for speech enhancement,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d3c53e10-c4f2-4b40-953d-c1a1122527eb · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment Permutation invari- ant training of deep models for speaker-independent multi-talker speech separation,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9f24950c-a1cd-4283-9cd1-8ab001d8e2c4 · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment Dual-Path RNN: Efficient long sequence modeling for time-domain single-channel speech sepa- ration,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d376de3c-6851-490f-a8d7-be8a5d857958 · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment Attention is all you need in speech separation,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 22a3c68e-0b4a-41ee-a4f2-14383c8ead30 · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment Selective listening by synchronizing speech with lips,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 07b9b447-c93c-4a44-9e37-c8d9f9885845 · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment TF-GridNet: Making time-frequency domain models great again for monaural speaker separation,
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 100526fc-2a2d-40f2-a4e1-ed08c42b5971 · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment SPMamba: State-space model is all you need in speech separation
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9606353-6c7f-47a4-a8e3-39ca7e2241b3 · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment Time domain audio visual speech separation,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f5854460-642f-4afb-9029-0d21a5d977fe · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment MuSE: Multi-modal target speaker extraction with visual cues,
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ea871c47-8135-4ac8-a3a6-cdb06e9d4e2a · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment USEV: Universal speaker extraction with visual cue,
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3ae2ef09-51e3-4364-9d36-43eabf7fb5d6 · outbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment SpeechBrain: A General-Purpose Speech Toolkit
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 714fd3b6-e2d1-45c8-b151-2d020951eb2c · inbound
ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02546d07-92d4-4255-a078-45e13b8b6517 · inbound
CueNet: Robust Audio-Visual Speaker Extraction through Cross-Modal Cue Mining and Interaction ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54563704-4c79-430a-bd33-33beede6abb6 · inbound
Beyond Monologue: Interactive Talking-Listening Avatar Generation with Conversational Audio Context-Aware Kernels ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation deafe58a-e056-440e-8781-a9f84efcd48a · inbound
Hierarchical Codec Diffusion for Video-to-Speech Generation ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c8e55c73-53a0-4cd6-8931-5862900542ea · inbound
Can Large Audio Language Models Ignore Multilingual Distractors? An Evaluation of Their Selective Auditory Attention Capabilities ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6c4486fa-1b94-4290-bd04-2b0d6369226b · inbound
Feature-Aligned Speech Watermarking for Robustness to Reconstruction Distortions ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f08e5579-001a-4a99-96e0-c7ceb8cf0d7d · inbound
Where Speech Enhancement Hurts Recognition: An Inference Time Polar Projection Diagnosis ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.