Pith. sign in

Paper Citation Record · LEDGER

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling

As of 8 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 1 inbound Pith citation observation for arXiv:2507.09510.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.09510 v3

Coverage vector

measured 51 of 51 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:58:41.063893Z

measured 52 of 52 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:58:34.654304Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T17:58:42.865534Z

Reference resolution

51 of 51 outbound references displayed

  • verified exact6
  • verified fuzzy38
  • unresolved5
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 870e7914-99ff-4e70-b38b-98f092d64188 · outbound

This paper cites The chal- lenge of isolating the target speech while ignoring other inter- ferences is known as the cocktail party problem [1].

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling The chal- lenge of isolating the target speech while ignoring other inter- ferences is known as the cocktail party problem [1]

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:53.038346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:34.470611Z digest=sha256:45e7a63b1960ac915d9e94257c56add9cf7507925523a72d9b4ae431f66b6238

Observation 460796ea-b9b2-49bc-a537-08fdb67af52d · outbound

This paper cites Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-06T17:58:42.987162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:34.654304Z digest=sha256:93c78a0f591799b905d777b44ba6d0b8c364b98ef3a36325bd66d433adce5f89

Observation 5eb2f227-2ea8-410d-b8b2-2201a0991e5a · outbound

This paper cites The proposed methods enhance overall performance, achieving state-of-the-art results.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling The proposed methods enhance overall performance, achieving state-of-the-art results

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:52.741434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:34.551888Z digest=sha256:711b7f65a00aad35578c9c72c62e3b9ff91f21ffa270c1a5150651c3b8211004

Observation cab9f041-ad09-4942-af0c-63b64cb3edf0 · outbound

This paper cites Dataset We use the clean Libri2Mix [26] dataset with two-speaker mix- tures.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Dataset We use the clean Libri2Mix [26] dataset with two-speaker mix- tures

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:52.458594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:34.793706Z digest=sha256:180d1b02d422f80233e1a64af36354613fdd3ee1e741e624d3520a2e0768676b

Observation c075f8a8-3d53-4810-bb23-37fc9b388278 · outbound

This paper cites an unresolved cited work.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:58:50.898381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:35.711552Z digest=sha256:e459238d10702b7f479957832812a998a05e88c7816caa5727c34762eafa4c18

Observation 337e1f65-c57a-4676-86b4-feffcd802435 · outbound

This paper cites Comparative studies with proposed methods Table 1 presents the performance of BSRNN with proposed methods on Libri2Mix.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Comparative studies with proposed methods Table 1 presents the performance of BSRNN with proposed methods on Libri2Mix

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:51.893049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:35.033551Z digest=sha256:07d7b3aaecda134ec67d01de8a6eb0c7fa64e9196754a302b53c5fdc91e9554a

Observation db43320f-9074-446f-bf5f-854a9124fb86 · outbound

This paper cites and other metrics in all scenarios.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling and other metrics in all scenarios

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:51.611443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:35.199613Z digest=sha256:04909562b50e35854f838143edcccdfd48e48e4527501b32df2b1d9ad8f7d247

Observation 822d6794-86b9-46b5-8b5b-4d94271d6249 · outbound

This paper cites For ResNet34 in pretrained mode, SI-SDR improved by 0.43 dB, accuracy by 1.56%, and Sim.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling For ResNet34 in pretrained mode, SI-SDR improved by 0.43 dB, accuracy by 1.56%, and Sim

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:51.397031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:35.317717Z digest=sha256:230c7560e989c2a2d04eb78ad829f0ff229171eed9399cfc5bfbc0b9b419459f

Observation 28f2b8ea-0ebd-44eb-9479-22d2da22a519 · outbound

This paper cites However, a slight decrease in the Sim.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling However, a slight decrease in the Sim

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:51.118859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:35.445487Z digest=sha256:15d919c12043a829e6fd59486c823e39ec94afac88bfc3735e93c67e03df35c8

Observation f050ae84-64f6-4b44-bd57-609f7e050a6b · outbound

This paper cites an unresolved cited work.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Unresolved cited work

Reference 10

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T17:58:42.746251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:35.564318Z digest=sha256:926feab2dfe92f5d8ab0284cd700edd2c9001563681f8583e48e1f38d4b60413

Observation 57c6b5e7-099a-4b66-b1c0-2b9c376c295c · outbound

This paper cites Spex+: A complete time domain speaker extraction network,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Spex+: A complete time domain speaker extraction network,

Reference 11

Resolution
verified exact
doi, observed 2026-08-06T17:58:41.640664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:37.093240Z digest=sha256:37b1383418736ab78db3a9a1e36629c9616c27fb6fc81301ad713b9fc5c01ee6

Observation 4e26e3b0-37c9-44a3-aacb-3604a1ade830 · outbound

This paper cites Some further experiments upon the recognition of speech, with one and with two ears,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Some further experiments upon the recognition of speech, with one and with two ears,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:50.695497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:35.827574Z digest=sha256:d2930010a1951467b8ee1b9d5b57412ef65bd489b1f48f51e5b81b1e58c7a2a6

Observation 861f1f0d-d5cb-473c-8324-3503423679c5 · outbound

This paper cites The cocktail-party problem revisited: early process- ing and selection of multi-talker speech,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling The cocktail-party problem revisited: early process- ing and selection of multi-talker speech,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:50.464346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:35.953626Z digest=sha256:b0c38ecdf3934b945b01469360d04337d6253af0c411effecbe40832fb12edc3

Observation 7d027329-0848-4d74-a6a2-05b65b530330 · outbound

This paper cites Speakerbeam: Speaker aware neural network for target speaker extraction in speech mixtures,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Speakerbeam: Speaker aware neural network for target speaker extraction in speech mixtures,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:50.159472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:36.061089Z digest=sha256:b4ca8d7060bbc458f2154ba981c3fb6b9cb663a7bcde31171f1d985c2cf0cd50

Observation 22189dbc-c058-4ea1-abee-5f7698904856 · outbound

This paper cites Adaptive blind audio source extraction supervised by dom- inant speaker identification using x-vectors,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Adaptive blind audio source extraction supervised by dom- inant speaker identification using x-vectors,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:49.916815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:36.196820Z digest=sha256:fd85032d842165ee5aec89bc2f2f47b6072b69484bb6a28cb1c0289a730cceb6

Observation 79c0e0c6-74b9-4def-8d3a-9f331e82fed9 · outbound

This paper cites Looking to listen at the cocktail party: A speaker-independent audio-visual model for speech separation,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Looking to listen at the cocktail party: A speaker-independent audio-visual model for speech separation,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T17:58:36.334856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:58:36.334856Z digest=sha256:7f6599e6a301e2bb49a483939d4a4e366d0406d6396c0204761be57c20c02f43

Observation f6b3d310-16e3-40f0-a688-f132e61e3109 · outbound

This paper cites Audio-visual sound separation via hidden markov models,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Audio-visual sound separation via hidden markov models,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:49.657251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:36.449562Z digest=sha256:0ed9f9024f3e7244ada7159139f9566f3cb7bd2b5f5f2bfd62ecdb1a4997b80a

Observation af15a8ea-4823-4eef-a39e-c93bd86e7267 · outbound

This paper cites Neural spatial filter: Target speaker speech separation assisted with directional information,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Neural spatial filter: Target speaker speech separation assisted with directional information,

Reference 18

Resolution
verified exact
doi, observed 2026-08-06T17:58:41.908685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:36.579776Z digest=sha256:8a1a8e72b8fc9a8ccaa44d937cd508de03a76e7ce3894cd8af0745b09f0a6fe1

Observation 0f880b43-d2f2-41c2-8626-efe03a73eef1 · outbound

This paper cites V oice filter: Few- shot text-to-speech speaker adaptation using voice conversion as a post- processing module,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling V oice filter: Few- shot text-to-speech speaker adaptation using voice conversion as a post- processing module,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:49.408598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:36.727248Z digest=sha256:aff16d88a1fc758fb2ec3076debc5804788dda87df64e6eb1f44f8ac158d0672

Observation 5c1b3112-7e5e-434f-95f4-fe3f2eedf8d1 · outbound

This paper cites Single-channel speech extraction using speaker inventory and attention network,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Single-channel speech extraction using speaker inventory and attention network,

Reference 20

Resolution
malformed identifier
no resolver link, observed 2026-08-06T17:58:36.849729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:58:36.849729Z digest=sha256:585e1abcde3727eb61d1fdf2d51e53cd693b0d4dbb3773106498d47a3c7199b1

Observation 0565405d-e347-4ded-91e0-f928a88263ed · outbound

This paper cites Target confusion in end- to-end speaker extraction: Analysis and approaches,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Target confusion in end- to-end speaker extraction: Analysis and approaches,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:49.075349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:36.966592Z digest=sha256:2bf8f1d079b7bc5fcdbe12d91bc94805a2987974abec7005250714791f55b3d9

Observation f9a99aaa-3293-48a6-ae46-d0c273b5b740 · outbound

This paper cites Sc-glowtts: an efficient zero-shot multi-speaker text-to-speech model,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Sc-glowtts: an efficient zero-shot multi-speaker text-to-speech model,

Reference 22

Resolution
verified exact
doi, observed 2026-08-06T17:58:41.332281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:38.645691Z digest=sha256:fa98f047b140cc0ae72362fcb95647c578e114d0e9499c6a4edbf950296b0ec0

Observation 6c39300c-79b4-4bd9-81f3-510e8762c787 · outbound

This paper cites X-sepformer: End-to-end speaker extraction network with explicit optimization on speaker confusion,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling X-sepformer: End-to-end speaker extraction network with explicit optimization on speaker confusion,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:48.848806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:37.191473Z digest=sha256:7950b87318d4453bc9f683d84118a8b8b4e05fe4bf188bd2b69e01e90551c485

Observation 5b01aa03-ecc7-4709-9135-d444296e4cfd · outbound

This paper cites Speaker ex- traction with detection of presence and absence of target speakers,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Speaker ex- traction with detection of presence and absence of target speakers,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:48.588117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:37.334445Z digest=sha256:ffdbe9607e773355c295263b657c64d165f935eb72d54cc00a000ab0d6a48982

Observation a27b6499-477a-4aa7-9b33-2cb633290a4c · outbound

This paper cites On the effectiveness of enrollment speech augmentation for target speaker extraction,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling On the effectiveness of enrollment speech augmentation for target speaker extraction,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:48.321486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:37.493824Z digest=sha256:ff20981a22fa9d9ee4ff7f677c8445b1043d4859315febb1ddc94cea28bffa1e

Observation 65129ac3-5d07-47d4-a9ac-157856c1a130 · outbound

This paper cites Selective hubert: Self- supervised pre-training for target speaker in clean and mixture speech,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Selective hubert: Self- supervised pre-training for target speaker in clean and mixture speech,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:47.974497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:37.612121Z digest=sha256:85b5651ec6d14ab65c8c891a887ff046fba18a2893ae68f5fd164c793c901077

Observation 7ec856e6-7044-4601-8c3e-eae9808aedd0 · outbound

This paper cites Contrastive learning for target speaker extraction with attention-based fusion,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Contrastive learning for target speaker extraction with attention-based fusion,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:47.452568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:37.916014Z digest=sha256:910a81b1998622c9eb04221cbb2008b20c7a1c6d5ffb64c60c769d6191d77167

Observation 5ec61414-d81e-4b3f-9672-5da76eb3834c · outbound

This paper cites Real-time personalised speech enhancement transformers with dynamic cross-attended speaker representations,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Real-time personalised speech enhancement transformers with dynamic cross-attended speaker representations,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:47.148722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:38.081178Z digest=sha256:35ac5548b0a1a9531d77ef1773ecbc0a5000dee4ec7c128b72906c315b83328c

Observation dc186a4c-5358-4844-a292-4e3437f233aa · outbound

This paper cites Smma-net: An audio clue- based target speaker extraction network with spectrogram matching and mu- tual attention,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Smma-net: An audio clue- based target speaker extraction network with spectrogram matching and mu- tual attention,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:46.806035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:38.291163Z digest=sha256:4ba9bf1fd0789edbc356042515d423546ceac5db4a46ed61533997ffd3d82d23

Observation fbb47c09-a1a4-4457-b8f6-55b72c625c02 · outbound

This paper cites Target speaker extraction by di- rectly exploiting contextual information in the time-frequency domain,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Target speaker extraction by di- rectly exploiting contextual information in the time-frequency domain,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:46.475563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:38.420322Z digest=sha256:0e9e324cc567320ad1c399c7d7dcbb45f2aa8966e653a4d2214fd1e35174a2c2

Observation 4d180dee-aad7-483b-85c5-dc9bd8c3c1cb · outbound

This paper cites Music source separation with band-split rnn,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Music source separation with band-split rnn,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:46.199355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:38.536836Z digest=sha256:a618c5c8ce5287c76d163927cccd0671e66472853096f9138df4a5b0eaf45e03

Observation c76a0f6f-7617-463d-b028-365e09664def · outbound

This paper cites An algorithm for intelligibility prediction of time–frequency weighted noisy speech,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling An algorithm for intelligibility prediction of time–frequency weighted noisy speech,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:44.836454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:40.037370Z digest=sha256:15385938684ecfea5aa5dbd840385525b74cf5195a3e4f01fa7842ebf2bd2193

Observation f235ec62-893c-47f2-8c67-67648fa1cd8b · outbound

This paper cites In defence of metric learning for speaker recognition,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling In defence of metric learning for speaker recognition,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:45.995819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:38.754873Z digest=sha256:d063b9c7efc319a0fa3d304ff2a3490785d36de9c4fb4c076b7e6f96254878cf

Observation ec33ac18-6273-44d0-a7e4-c91995132715 · outbound

This paper cites Centroid estimation with transformer-based speaker embedder for robust target speaker extraction,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Centroid estimation with transformer-based speaker embedder for robust target speaker extraction,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:45.802159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:38.870085Z digest=sha256:9e969d0760afcc37ab73f71df84df26e58516072c4ff8cd1243633eed11aaec5

Observation 7553ccb9-467f-4907-8590-877601927c26 · outbound

This paper cites Sdr–half-baked or well done?.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Sdr–half-baked or well done?

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:45.638351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:39.010501Z digest=sha256:ffc87102f6e49a029adfc7f9ac4311291edef324407e915eae41dfc4e7befefc

Observation 31d26293-98a0-43d4-9c27-cc536da3f3d7 · outbound

This paper cites Lib- rimix: An open-source dataset for generalizable speech separation,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Lib- rimix: An open-source dataset for generalizable speech separation,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:45.473589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:39.151681Z digest=sha256:07c4477656749b2329aa9d675fc228433c34287c79df141e0f94558b72ccecbd

Observation 49af23a8-a315-4150-a6f8-c19823114eaf · outbound

This paper cites But system description to voxceleb speaker recognition challenge 2019,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling But system description to voxceleb speaker recognition challenge 2019,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:45.304615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:39.253114Z digest=sha256:6928afd41d30f8908b30cb69c55e6208ebf28bec272e0d003d7c5d7f8eb7df0c

Observation f58f258f-bf1c-4927-9cfe-01972083417b · outbound

This paper cites Ecapa-tdnn : Emphasized channel attention, propagation and aggregation in tdnn based speaker verification,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Ecapa-tdnn : Emphasized channel attention, propagation and aggregation in tdnn based speaker verification,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T17:58:39.400949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:58:39.400949Z digest=sha256:f70f9041407b58eed6bf3117938b5d74fabb554fbe89a08290c3432b56fe4240

Observation b8624695-9fa7-4885-92e4-b63d42c05b9f · outbound

This paper cites Wespeaker: A research and production oriented speaker embed- ding learning toolkit,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Wespeaker: A research and production oriented speaker embed- ding learning toolkit,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:45.136229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:39.509318Z digest=sha256:6c48b952a713020a98b33a7df102e39b31a12666cc9bb2d29d9e65fafbcea357

Observation d861da22-9724-4e98-9b00-3e6cf57d6b19 · outbound

This paper cites V oxceleb2: Deep speaker recognition,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling V oxceleb2: Deep speaker recognition,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T17:58:39.616004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:58:39.616004Z digest=sha256:fc8b4654f12c6a4f232e07cb5010c5349a850e757aa31fb948e4b94eb2cc4aa1

Observation 9beb7c1c-b899-43c8-a306-d54f488fcbd1 · outbound

This paper cites WeSep: A Scalable and Flexible Toolkit Towards Generalizable Target Speaker Extraction.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling WeSep: A Scalable and Flexible Toolkit Towards Generalizable Target Speaker Extraction

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-08-06T17:58:42.352929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:39.732096Z digest=sha256:1839159a99813698dc74f88b405f9f3cb20db37c2a5d06e9eabefb5489e129f8

Observation 19611a75-f91d-49d7-82f2-05ece02de0c5 · outbound

This paper cites Perceptual evaluation of speech quality (pesq): An objective method for end-to-end speech quality assessment of narrow-band telephone networks and speech codecs,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Perceptual evaluation of speech quality (pesq): An objective method for end-to-end speech quality assessment of narrow-band telephone networks and speech codecs,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:44.962435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:39.867589Z digest=sha256:ea77f5410c14e3285dc0b69f02065da7f36d96c537983b1d408e738a2ab2c53d

Observation 2b029e47-5f9f-4c78-a18c-4ff84e909f84 · outbound

This paper cites Xtts: a massively mul- tilingual zero-shot text-to-speech model,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Xtts: a massively mul- tilingual zero-shot text-to-speech model,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:44.620467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:40.186751Z digest=sha256:19f8c5d4fc74b7435f263b953f450764847a908d25d95d4d10a825c56e326c10

Observation 08b27bca-02f0-49a5-9cbb-97cca54c4648 · outbound

This paper cites Ecapa2: A hybrid neural network archi- tecture and training strategy for robust speaker embeddings,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Ecapa2: A hybrid neural network archi- tecture and training strategy for robust speaker embeddings,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:44.343895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:40.303176Z digest=sha256:ee0a5dadb155274275d75f3149bc682591097828845593beaeca75680e40f3df

Observation 4cb35800-a4d5-4da0-8aa0-d612e50c931e · outbound

This paper cites Multi-Level Speaker Representation for Target Speaker Extraction.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Multi-Level Speaker Representation for Target Speaker Extraction

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-08-06T17:58:42.148745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:40.451824Z digest=sha256:948d91ed04416735fcb4fb4954f9e674aa8a9cc1d52897d1d52768ba2bfc858e

Observation 053c27d9-27b9-4d9a-a5eb-f6d36f47f9b4 · outbound

This paper cites Dpccn: Densely-connected pyramid complex convolutional network for robust speech separation and extraction,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Dpccn: Densely-connected pyramid complex convolutional network for robust speech separation and extraction,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:44.054911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:40.578466Z digest=sha256:29dab6f9c11013b8bbd1e70d785a1dec5df8611fa46a93d055753d8e24041d1a

Observation 9850ff13-8b76-4491-a351-cabbcf7af160 · outbound

This paper cites Mc-spex: Towards effective speaker extraction with multi-scale interfu- sion and conditional speaker modulation,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Mc-spex: Towards effective speaker extraction with multi-scale interfu- sion and conditional speaker modulation,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:43.814757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:40.670334Z digest=sha256:358ae976319ffbfcf2134ae1cf88d32f77f10d18c5a45c7b07a2d2013e096737

Observation de256083-5227-41d8-b2ef-3a28a4986f1a · outbound

This paper cites Tar- get speech extraction with pre-trained self-supervised learning models,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Tar- get speech extraction with pre-trained self-supervised learning models,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:47.697564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:40.797574Z digest=sha256:f31b9bd94d0d065b01b2d8964454d6727906fa7a124ed9661b34e8a491f72a60

Observation 531299cb-327f-401b-804b-7061af17d7fb · outbound

This paper cites Tf- gridnet: Making time-frequency domain models great again for monaural speaker separation,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Tf- gridnet: Making time-frequency domain models great again for monaural speaker separation,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:43.595248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:40.942430Z digest=sha256:0a540d4428f0d284e348aca1485daded19e0c62091d5ddefa295d2a326c1cd89

Observation 8e319a68-b513-4e42-8018-03d56c53fe1a · outbound

This paper cites V oxceleb: A large-scale speaker identification dataset,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling V oxceleb: A large-scale speaker identification dataset,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:43.348936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:41.063893Z digest=sha256:26ac802a59008837bd43731df5137e12652ec61252e8942f6b58de726e7bace8

Observation 4202cf03-9f11-4a00-80b4-ede8b5b33f86 · outbound

This paper cites an unresolved cited work.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Unresolved cited work

Reference 128

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:58:52.159552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:34.919879Z digest=sha256:b293a3c6d36ed255fe2779bb4d44719da90a9ccf577a98b4f2c0407e26563028

Pith citing papers

Observation 460796ea-b9b2-49bc-a537-08fdb67af52d · inbound

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling cites this paper.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-06T17:58:42.987162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:58:34.654304Z digest=sha256:93c78a0f591799b905d777b44ba6d0b8c364b98ef3a36325bd66d433adce5f89