Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:22:07.475291Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 80 of 80 outbound references and 3 inbound Pith citation observations for arXiv:2505.19314.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:22:07.475291Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-03T14:20:49.132229Z
A source-named dated measurement, never combined with another source.
Source: cited_works
80 of 80 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation d86250cd-c963-4ac1-9256-ddbe38fad0ee · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline The cocktail-party problem revisited: early pro- cessing and selection of multi-talker speech,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d24a1ce8-4e6c-4d93-9999-a8c2e4cf756a · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Neural target speech extraction: An overview,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a5f4c12f-df00-4fbe-9813-b06d63e71704 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Neural spatial filter: Target speaker speech separation assisted with directional information,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8b003eec-47d2-42bc-b634-cecfc738d80d · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Far-field location guided target speech extraction using end-to-end speech recognition objectives,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1eedf57c-0046-4639-a424-985e93779fbf · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Looking to listen at the cocktail party: a speaker-independent audio-visual model for speech separation,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 27460b18-8024-4c28-a89d-4479dbe0a0ef · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Conceptbeam: Concept driven target speech extraction,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 693bb2b9-cf52-4e47-8d02-945d341d136c · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline V oicefilter: Targeted voice separation by speaker-conditioned spectrogram masking,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 930acadf-25cf-4460-b1bf-70371d4bf6ca · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Speakerbeam: Speaker aware neural network for target speaker extraction in speech mixtures,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 33604da3-1634-4163-ac55-ae8543afb007 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Wesep: A scalable and flexible toolkit towards generalizable target speaker extraction,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 460ed315-7fe9-41ff-b469-ddcbc96234dd · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Target confusion in end-to-end speaker extraction: Analysis and approaches,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f22eb8bc-4f60-4f6c-803d-194c16e497aa · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Dpccn: Densely-connected pyramid complex convolutional network for robust speech separation 11 and extraction,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a5f74896-a48a-44e2-a4d5-4a405cdc5bdc · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Improving target sound extraction with timestamp information,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 610f6f0c-164e-427f-ad8d-55bce01922a0 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline WeSep: A Scalable and Flexible Toolkit Towards Generalizable Target Speaker Extraction
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 611ca561-de03-4463-8427-7df5c5ba0e9c · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Spex: Multi-scale time domain speaker extraction network,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3858436a-b8b9-43ec-9a1f-78e582e3d462 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Spex+: A complete time domain speaker extraction network,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 85db493b-472e-4c86-895d-e4730dc4985c · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline X-SEPFORMER: end-to- end speaker extraction network with explicit optimization on speaker confusion,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8da9cead-1b4d-4c0c-877a-787dddb08cab · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline X-tf-gridnet: A time-frequency domain target speaker extraction network with adaptive speaker embedding fusion,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1de9266a-258a-4493-b63b-e2782b33b22b · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline USEF-TSE: Universal Speaker Embedding Free Target Speaker Extraction
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ed61ae57-6b30-475d-b20a-a71d244102f2 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Target speech extraction with conditional diffusion model,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3f2ee616-3578-4865-9a29-f62fd92bdbb1 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Noise-robust Speech Separation with Fast Generative Correction
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9f531854-3c29-43d1-bf74-867007a691f8 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Speech enhancement and dereverberation with diffusion-based generative models,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1c4af11a-89ad-4505-9ea5-a24441951168 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Diffusion-based generative speech source separation,
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63742595-d3a9-4838-bd75-3a4f27419613 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Generative pre-training for speech with flow matching,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ffb056b2-5e5a-4357-a714-17287e0219e7 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Metis: A Foundation Speech Generation Model with Masked Generative Pre-training
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af0f0304-7d6e-4afa-91ce-5577d88b2cbd · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline SoloAudio: Target Sound Extraction with Language-oriented Audio Diffusion Transformer
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf62dc7e-f961-46a4-be01-5f8568be862f · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Separate And Diffuse: Using a Pretrained Diffusion Model for Improving Source Separation
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eadbdd15-2a50-446f-88e2-1527b08990b5 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Attention is all you need,
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17e636f9-cabc-4df8-b41f-2cbca8dbe101 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Large language model based generative error correction: A challenge and baselines for speech recognition, speaker tagging, and emotion recognition,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 58f778a2-4cd6-4848-b207-4291a63e5542 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline SepALM: Audio Language Models Are Error Correctors for Robust Speech Separation
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ff992483-89bd-4c28-8710-1079063a1ac4 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline LibriMix: An Open-Source Dataset for Generalizable Speech Separation
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 335f9c72-14f9-4d0b-9d3d-7ba3e47b1cba · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Target speech extraction with conditional diffusion model,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e99c4cc7-f709-4a9a-a5f6-b9cb4bf63e49 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Dpm-tse: A diffusion probabilistic model for target sound extraction,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 73ea1afc-061b-4f60-a9f9-019697bca4fb · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Diffusion- based generative speech source separation,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 246e030c-ca12-4386-8f20-fdd375bd3a4e · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Generative Speech Foundation Model Pretraining for High-Quality Speech Extraction and Restoration
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation aa515a27-f2d8-4c0c-99c9-7d5b713dd5b7 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Generation- based target speech extraction with speech discretization and vocoder,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 25dbd4bf-1922-457e-abd1-dd3165775919 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 72098b94-b6c8-4c80-a0a5-18864715c414 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Diffusion-based signal refiner for speech separation,
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4842be6c-412c-46b7-83ea-3ef0264395a8 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Storm: A diffusion-based stochastic regeneration model for speech enhancement and dereverberation,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 22f71e89-d785-415d-b410-649fa43d864b · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Ddtse: Discriminative diffusion model for target speech extraction,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4c8713c5-9933-4c50-ba86-8f2066bbe735 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Speaker-aware neural network based beamformer for speaker extraction in speech mixtures,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 52a51c3a-7be6-4adb-80a6-4269e55219eb · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline X- vectors: Robust DNN embeddings for speaker recognition,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4f347263-8123-49eb-a228-bd5c259de03c · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Probing self-supervised learning models with target speech extraction,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 70347c62-4802-4526-b4b1-884eaeda3a8c · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Target speech extraction with pre-trained self-supervised learning models,
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a6131e0a-85aa-4405-8e71-6099872bfdb6 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Smma-net: An audio clue-based target speaker extraction network with spectrogram matching and mutual attention,
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dc6868b9-d78f-42f0-848f-bb5f43732e07 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Target speaker extraction by directly exploiting contextual information in the time-frequency domain,
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f0f6c734-18fe-4623-9839-4327c25098b4 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Target speaker extraction with ultra-short reference speech by VE-VE framework,
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cfcbd0d9-aa6b-4853-b2a3-7dd89ec2c885 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Sef-net: Speaker embedding free target speaker extraction network,
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6dc8bc8b-9d24-4d81-9c76-b40ddc55d071 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Common diffusion noise schedules and sample steps are flawed,
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 23f2da4c-8b3e-435f-ad39-abb56ef48d70 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Progressive distillation for fast sampling of diffusion models,
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f157c4e6-1f2e-4093-a974-7438ab5c6e9d · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline High- fidelity audio compression with improved RVQGAN,
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 010a6aa9-89f3-47cb-9fbb-8b2e41a8ab8f · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Stable Audio Open
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96e4134f-1562-4fc0-96fe-b16912434989 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline EzAudio: Enhancing Text-to-Audio Generation with Efficient Diffusion Transformer
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15bc5cc7-3c30-4004-a002-2d6a7760ec0b · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Tf- gridnet: Integrating full- and sub-band modeling for speech separation,
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e6a27d0f-83b0-47a4-b831-3b0e46ed4003 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline SPMamba: State-space model is all you need in speech separation
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11054119-9270-47de-9613-91804548ef4b · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Complex ratio masking for monaural speech separation,
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation aadc30e0-17e9-4e6e-ba89-151dfc7f0c01 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline auraloss: Audio focused loss functions in pytorch,
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54a7bc8f-5781-4c69-b7e1-ad5fe4ce80f4 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline High fidelity neural audio compression,
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation de2e9680-5b1a-41a2-9c2f-441ef16b5157 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Scalable diffusion models with transformers,
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 31cfd854-1be9-4e3c-804e-60f026a25267 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline An image is worth 16x16 words: Transformers for image recognition at scale,
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a0bfd558-bae4-425f-91e7-49c5aed8eebf · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Sigmoid-weighted linear units for neural network function approximation in reinforcement learning,
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f748628-177d-4a38-b1bd-f2f82efc59c0 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Roformer: Enhanced transformer with rotary position embedding,
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0edd31f9-8f3b-4bb4-9518-44c1cbd8072d · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Single- channel multi-speaker separation using deep clustering,
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 75bdcf37-727d-4132-8a8d-0786387e5952 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Conv-tasnet: Surpassing ideal time-frequency magnitude masking for speech separation,
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4918245f-567b-4278-be56-986fac214ce5 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Wham!: Extending speech separation to noisy environments,
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 79627c97-0cab-4265-9437-c08bde14e8bb · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Librispeech: An ASR corpus based on public domain audio books,
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 73d7aa18-fb73-4828-97c6-8dd42e67a898 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Improving speaker discrimination of target speech extraction with time-domain speakerbeam,
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7deb9100-49d1-4a44-81eb-7d1dac4d1144 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline MUSAN: A Music, Speech, and Noise Corpus
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5dd377b-9cfe-4767-b06e-9e8dc831bfbf · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Multichannel audio database in various acoustic environments,
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6be34bc1-4554-44a6-bbce-60ee64db0d1c · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline The fifth ’chime’ speech separation and recognition challenge: Dataset, task and baselines,
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b89ca30e-01c2-46bf-b953-d0904c8f36e3 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline SonicSim: A customizable simulation platform for speech processing in moving sound source scenarios
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35abf0fb-854c-4bf8-ae44-0209e167cd0d · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Perceptual evaluation of speech quality (pesq)-a new method for speech quality assessment of telephone networks and codecs,
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e24fe973-e9ee-44e8-88a7-9561bd4b5938 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline An algorithm for predicting the intelligibility of speech masked by modulated noise maskers,
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9ca04559-6a14-4eb9-95b6-5a04d202f30a · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Dnsmos P.835: A non- intrusive perceptual objective speech quality metric to evaluate noise suppressors,
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a04353fe-156d-442a-ba94-d5b2718ed1be · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Robust speech recognition via large-scale weak supervision,
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7f382b3e-dce3-42ac-8325-5ba18d26efb8 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Wavlm: Large-scale self-supervised pre- training for full stack speech processing,
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5c1acbc-6245-400f-8509-f8751ee5b120 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Improving Generalization of Speech Separation in Real-World Scenarios: Strategies in Simulation, Optimization, and Evaluation
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 31812576-3bd2-4f26-836a-81dda307c00c · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Classifier-Free Diffusion Guidance
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2da558ad-64b7-47c4-b910-1665ea66ae00 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Photorealistic text-to-image diffusion models with deep language understanding,
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6f599b9a-c3e8-4b82-9d3d-e8ebea2852e9 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Available: https://openreview.net/forum?id=08Yk-n5l2Al
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f9ec7a85-4eca-4bf2-b7a6-cc964b864693 · outbound
SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline Unresolved cited work
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc329c18-7386-4acd-8846-95a8f1672c13 · inbound
GenTSE: Enhancing Target Speaker Extraction via a Coarse-to-Fine Generative Language Model SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0293e581-3ff6-4913-8d5f-4960ebfa64ed · inbound
Enroll-on-Wakeup: A First Comparative Study of Target Speech Extraction for Seamless Interaction in Real Noisy Human-Machine Dialogue Scenarios SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29ebe771-cc57-4139-8478-42c13570334e · inbound
Beyond Acoustic Prefixes: Persistent Grounding in Serialized Acoustic Memory for LLM-Based Multi-Talker Speech Recognition SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.