Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T17:58:41.063893Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 1 inbound Pith citation observation for arXiv:2507.09510.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T17:58:41.063893Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T17:58:34.654304Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-06T17:58:42.865534Z
51 of 51 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 870e7914-99ff-4e70-b38b-98f092d64188 · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling The chal- lenge of isolating the target speech while ignoring other inter- ferences is known as the cocktail party problem [1]
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 460796ea-b9b2-49bc-a537-08fdb67af52d · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5eb2f227-2ea8-410d-b8b2-2201a0991e5a · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling The proposed methods enhance overall performance, achieving state-of-the-art results
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cab9f041-ad09-4942-af0c-63b64cb3edf0 · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Dataset We use the clean Libri2Mix [26] dataset with two-speaker mix- tures
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c075f8a8-3d53-4810-bb23-37fc9b388278 · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 337e1f65-c57a-4676-86b4-feffcd802435 · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Comparative studies with proposed methods Table 1 presents the performance of BSRNN with proposed methods on Libri2Mix
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation db43320f-9074-446f-bf5f-854a9124fb86 · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling and other metrics in all scenarios
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 822d6794-86b9-46b5-8b5b-4d94271d6249 · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling For ResNet34 in pretrained mode, SI-SDR improved by 0.43 dB, accuracy by 1.56%, and Sim
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 28f2b8ea-0ebd-44eb-9479-22d2da22a519 · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling However, a slight decrease in the Sim
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f050ae84-64f6-4b44-bd57-609f7e050a6b · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Unresolved cited work
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 57c6b5e7-099a-4b66-b1c0-2b9c376c295c · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Spex+: A complete time domain speaker extraction network,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4e26e3b0-37c9-44a3-aacb-3604a1ade830 · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Some further experiments upon the recognition of speech, with one and with two ears,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 861f1f0d-d5cb-473c-8324-3503423679c5 · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling The cocktail-party problem revisited: early process- ing and selection of multi-talker speech,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7d027329-0848-4d74-a6a2-05b65b530330 · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Speakerbeam: Speaker aware neural network for target speaker extraction in speech mixtures,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 22189dbc-c058-4ea1-abee-5f7698904856 · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Adaptive blind audio source extraction supervised by dom- inant speaker identification using x-vectors,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 79c0e0c6-74b9-4def-8d3a-9f331e82fed9 · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Looking to listen at the cocktail party: A speaker-independent audio-visual model for speech separation,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6b3d310-16e3-40f0-a688-f132e61e3109 · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Audio-visual sound separation via hidden markov models,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation af15a8ea-4823-4eef-a39e-c93bd86e7267 · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Neural spatial filter: Target speaker speech separation assisted with directional information,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0f880b43-d2f2-41c2-8626-efe03a73eef1 · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling V oice filter: Few- shot text-to-speech speaker adaptation using voice conversion as a post- processing module,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5c1b3112-7e5e-434f-95f4-fe3f2eedf8d1 · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Single-channel speech extraction using speaker inventory and attention network,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0565405d-e347-4ded-91e0-f928a88263ed · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Target confusion in end- to-end speaker extraction: Analysis and approaches,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f9a99aaa-3293-48a6-ae46-d0c273b5b740 · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Sc-glowtts: an efficient zero-shot multi-speaker text-to-speech model,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6c39300c-79b4-4bd9-81f3-510e8762c787 · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling X-sepformer: End-to-end speaker extraction network with explicit optimization on speaker confusion,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5b01aa03-ecc7-4709-9135-d444296e4cfd · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Speaker ex- traction with detection of presence and absence of target speakers,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a27b6499-477a-4aa7-9b33-2cb633290a4c · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling On the effectiveness of enrollment speech augmentation for target speaker extraction,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 65129ac3-5d07-47d4-a9ac-157856c1a130 · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Selective hubert: Self- supervised pre-training for target speaker in clean and mixture speech,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7ec856e6-7044-4601-8c3e-eae9808aedd0 · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Contrastive learning for target speaker extraction with attention-based fusion,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5ec61414-d81e-4b3f-9672-5da76eb3834c · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Real-time personalised speech enhancement transformers with dynamic cross-attended speaker representations,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dc186a4c-5358-4844-a292-4e3437f233aa · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Smma-net: An audio clue- based target speaker extraction network with spectrogram matching and mu- tual attention,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fbb47c09-a1a4-4457-b8f6-55b72c625c02 · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Target speaker extraction by di- rectly exploiting contextual information in the time-frequency domain,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4d180dee-aad7-483b-85c5-dc9bd8c3c1cb · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Music source separation with band-split rnn,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c76a0f6f-7617-463d-b028-365e09664def · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling An algorithm for intelligibility prediction of time–frequency weighted noisy speech,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f235ec62-893c-47f2-8c67-67648fa1cd8b · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling In defence of metric learning for speaker recognition,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ec33ac18-6273-44d0-a7e4-c91995132715 · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Centroid estimation with transformer-based speaker embedder for robust target speaker extraction,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7553ccb9-467f-4907-8590-877601927c26 · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Sdr–half-baked or well done?
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 31d26293-98a0-43d4-9c27-cc536da3f3d7 · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Lib- rimix: An open-source dataset for generalizable speech separation,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 49af23a8-a315-4150-a6f8-c19823114eaf · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling But system description to voxceleb speaker recognition challenge 2019,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f58f258f-bf1c-4927-9cfe-01972083417b · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Ecapa-tdnn : Emphasized channel attention, propagation and aggregation in tdnn based speaker verification,
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8624695-9fa7-4885-92e4-b63d42c05b9f · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Wespeaker: A research and production oriented speaker embed- ding learning toolkit,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d861da22-9724-4e98-9b00-3e6cf57d6b19 · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling V oxceleb2: Deep speaker recognition,
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9beb7c1c-b899-43c8-a306-d54f488fcbd1 · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling WeSep: A Scalable and Flexible Toolkit Towards Generalizable Target Speaker Extraction
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 19611a75-f91d-49d7-82f2-05ece02de0c5 · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Perceptual evaluation of speech quality (pesq): An objective method for end-to-end speech quality assessment of narrow-band telephone networks and speech codecs,
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2b029e47-5f9f-4c78-a18c-4ff84e909f84 · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Xtts: a massively mul- tilingual zero-shot text-to-speech model,
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 08b27bca-02f0-49a5-9cbb-97cca54c4648 · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Ecapa2: A hybrid neural network archi- tecture and training strategy for robust speaker embeddings,
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4cb35800-a4d5-4da0-8aa0-d612e50c931e · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Multi-Level Speaker Representation for Target Speaker Extraction
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 053c27d9-27b9-4d9a-a5eb-f6d36f47f9b4 · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Dpccn: Densely-connected pyramid complex convolutional network for robust speech separation and extraction,
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9850ff13-8b76-4491-a351-cabbcf7af160 · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Mc-spex: Towards effective speaker extraction with multi-scale interfu- sion and conditional speaker modulation,
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation de256083-5227-41d8-b2ef-3a28a4986f1a · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Tar- get speech extraction with pre-trained self-supervised learning models,
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 531299cb-327f-401b-804b-7061af17d7fb · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Tf- gridnet: Making time-frequency domain models great again for monaural speaker separation,
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8e319a68-b513-4e42-8018-03d56c53fe1a · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling V oxceleb: A large-scale speaker identification dataset,
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4202cf03-9f11-4a00-80b4-ede8b5b33f86 · outbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Unresolved cited work
Reference 128
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 460796ea-b9b2-49bc-a537-08fdb67af52d · inbound
Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.