Pith. sign in

Paper Citation Record · LEDGER

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling

As of 8 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 1 inbound Pith citation observation for arXiv:2507.09510.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.09510 v3

Coverage vector

measured 51 of 51 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:58:41.063893Z

measured 52 of 52 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:58:34.654304Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T17:58:42.865534Z

Reference resolution

51 of 51 outbound references displayed

  • verified exact6
  • verified fuzzy38
  • unresolved5
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 870e7914-99ff-4e70-b38b-98f092d64188 · outbound

This paper cites The chal- lenge of isolating the target speech while ignoring other inter- ferences is known as the cocktail party problem [1].

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling The chal- lenge of isolating the target speech while ignoring other inter- ferences is known as the cocktail party problem [1]

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:53.038346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:34.470611Z digest=sha256:97b900c3d70054c9bcffc206bee0eb1b0ff88842d36735b015ba6e1761f77e08

Observation 460796ea-b9b2-49bc-a537-08fdb67af52d · outbound

This paper cites Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-06T17:58:42.987162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:34.654304Z digest=sha256:7eb1821475b3e56e86db655861a90c084be2ab1c78b7e62e8584755cd1e7fe7e

Observation 5eb2f227-2ea8-410d-b8b2-2201a0991e5a · outbound

This paper cites The proposed methods enhance overall performance, achieving state-of-the-art results.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling The proposed methods enhance overall performance, achieving state-of-the-art results

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:52.741434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:34.551888Z digest=sha256:a3aaa03d9a1ecaffdd2a45f388d63d9a03d2c74528d2397b7e2c027159dba735

Observation cab9f041-ad09-4942-af0c-63b64cb3edf0 · outbound

This paper cites Dataset We use the clean Libri2Mix [26] dataset with two-speaker mix- tures.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Dataset We use the clean Libri2Mix [26] dataset with two-speaker mix- tures

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:52.458594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:34.793706Z digest=sha256:9c523669527477ec1d287cfaa19242ce7ddd41a9e658212525e91f6723d5555a

Observation c075f8a8-3d53-4810-bb23-37fc9b388278 · outbound

This paper cites an unresolved cited work.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:58:50.898381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:35.711552Z digest=sha256:5c7147fde37d4829dfa71fde3121ef690ac5ed57e17a76b1531a29ac753089a5

Observation 337e1f65-c57a-4676-86b4-feffcd802435 · outbound

This paper cites Comparative studies with proposed methods Table 1 presents the performance of BSRNN with proposed methods on Libri2Mix.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Comparative studies with proposed methods Table 1 presents the performance of BSRNN with proposed methods on Libri2Mix

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:51.893049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:35.033551Z digest=sha256:54fbe42d5123d1bb9225a08231b353171160d3f969a5b783bbb194237b1e34c6

Observation db43320f-9074-446f-bf5f-854a9124fb86 · outbound

This paper cites and other metrics in all scenarios.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling and other metrics in all scenarios

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:51.611443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:35.199613Z digest=sha256:2de26a3bd83d15bb810c9ac65c710b3155c25c195cd31b932bdba256e03dcabe

Observation 822d6794-86b9-46b5-8b5b-4d94271d6249 · outbound

This paper cites For ResNet34 in pretrained mode, SI-SDR improved by 0.43 dB, accuracy by 1.56%, and Sim.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling For ResNet34 in pretrained mode, SI-SDR improved by 0.43 dB, accuracy by 1.56%, and Sim

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:51.397031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:35.317717Z digest=sha256:484feb4c3abbb6ac8354a0f086390da4c24a7def7df3ae1bb13761f81bb69c63

Observation 28f2b8ea-0ebd-44eb-9479-22d2da22a519 · outbound

This paper cites However, a slight decrease in the Sim.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling However, a slight decrease in the Sim

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:51.118859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:35.445487Z digest=sha256:755103996c2162798d6e9092219e2f45e779ecd41ccf4bce642363e747610b10

Observation f050ae84-64f6-4b44-bd57-609f7e050a6b · outbound

This paper cites an unresolved cited work.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Unresolved cited work

Reference 10

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T17:58:42.746251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:35.564318Z digest=sha256:f6fe9ef9d2ce454a3c7c7370ff46d0ea598c320488ef0138848b9e59ab56cbc3

Observation 57c6b5e7-099a-4b66-b1c0-2b9c376c295c · outbound

This paper cites Spex+: A complete time domain speaker extraction network,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Spex+: A complete time domain speaker extraction network,

Reference 11

Resolution
verified exact
doi, observed 2026-08-06T17:58:41.640664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:37.093240Z digest=sha256:974fb869a334e724ef445a149d135b68a0a123368d137a2bf9b17bc16de0daa7

Observation 4e26e3b0-37c9-44a3-aacb-3604a1ade830 · outbound

This paper cites Some further experiments upon the recognition of speech, with one and with two ears,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Some further experiments upon the recognition of speech, with one and with two ears,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:50.695497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:35.827574Z digest=sha256:79dbdc7eea4be7a3821858c5aee18ef088fcb28c5b3f3567b0a5b2fc759cbf91

Observation 861f1f0d-d5cb-473c-8324-3503423679c5 · outbound

This paper cites The cocktail-party problem revisited: early process- ing and selection of multi-talker speech,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling The cocktail-party problem revisited: early process- ing and selection of multi-talker speech,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:50.464346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:35.953626Z digest=sha256:682b110fdf8453f6005ffe367b77b4f2f16b7042e42b41d555953594331ef120

Observation 7d027329-0848-4d74-a6a2-05b65b530330 · outbound

This paper cites Speakerbeam: Speaker aware neural network for target speaker extraction in speech mixtures,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Speakerbeam: Speaker aware neural network for target speaker extraction in speech mixtures,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:50.159472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:36.061089Z digest=sha256:7f2efc08bdf43d1acec75edcae9c9e35cf91c1539c85524c91fb6e3fb44dfa28

Observation 22189dbc-c058-4ea1-abee-5f7698904856 · outbound

This paper cites Adaptive blind audio source extraction supervised by dom- inant speaker identification using x-vectors,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Adaptive blind audio source extraction supervised by dom- inant speaker identification using x-vectors,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:49.916815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:36.196820Z digest=sha256:558d21aadb3a25c820f7141fc8f3d4fd22bd6c5d3861d6d4d9fa92b87a6fa983

Observation 79c0e0c6-74b9-4def-8d3a-9f331e82fed9 · outbound

This paper cites Looking to listen at the cocktail party: A speaker-independent audio-visual model for speech separation,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Looking to listen at the cocktail party: A speaker-independent audio-visual model for speech separation,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T17:58:36.334856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:58:36.334856Z digest=sha256:7f6599e6a301e2bb49a483939d4a4e366d0406d6396c0204761be57c20c02f43

Observation f6b3d310-16e3-40f0-a688-f132e61e3109 · outbound

This paper cites Audio-visual sound separation via hidden markov models,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Audio-visual sound separation via hidden markov models,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:49.657251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:36.449562Z digest=sha256:8c9d5c0c8cfff1edc1cab92075d9234ed1ae89c6b211a97e9c3ebba9b93ca91b

Observation af15a8ea-4823-4eef-a39e-c93bd86e7267 · outbound

This paper cites Neural spatial filter: Target speaker speech separation assisted with directional information,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Neural spatial filter: Target speaker speech separation assisted with directional information,

Reference 18

Resolution
verified exact
doi, observed 2026-08-06T17:58:41.908685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:36.579776Z digest=sha256:73bdf9c6051df70d2524b017b65314dba5266fcdda08d4efdb1703748d0ffa39

Observation 0f880b43-d2f2-41c2-8626-efe03a73eef1 · outbound

This paper cites V oice filter: Few- shot text-to-speech speaker adaptation using voice conversion as a post- processing module,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling V oice filter: Few- shot text-to-speech speaker adaptation using voice conversion as a post- processing module,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:49.408598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:36.727248Z digest=sha256:f8c9daa9618044c7cac22e6580ba188d1a61f89623f5bfac3773f5e5350c2a91

Observation 5c1b3112-7e5e-434f-95f4-fe3f2eedf8d1 · outbound

This paper cites Single-channel speech extraction using speaker inventory and attention network,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Single-channel speech extraction using speaker inventory and attention network,

Reference 20

Resolution
malformed identifier
no resolver link, observed 2026-08-06T17:58:36.849729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:58:36.849729Z digest=sha256:585e1abcde3727eb61d1fdf2d51e53cd693b0d4dbb3773106498d47a3c7199b1

Observation 0565405d-e347-4ded-91e0-f928a88263ed · outbound

This paper cites Target confusion in end- to-end speaker extraction: Analysis and approaches,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Target confusion in end- to-end speaker extraction: Analysis and approaches,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:49.075349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:36.966592Z digest=sha256:33fa7220e17f180922dd2da8c9caf928f524a9b178cd964d71c8e4c4652573ba

Observation f9a99aaa-3293-48a6-ae46-d0c273b5b740 · outbound

This paper cites Sc-glowtts: an efficient zero-shot multi-speaker text-to-speech model,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Sc-glowtts: an efficient zero-shot multi-speaker text-to-speech model,

Reference 22

Resolution
verified exact
doi, observed 2026-08-06T17:58:41.332281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:38.645691Z digest=sha256:c41ec93ee4026185615a34fb08ec69c9f55b540e374cd68b10dc6c1ce57dadd9

Observation 6c39300c-79b4-4bd9-81f3-510e8762c787 · outbound

This paper cites X-sepformer: End-to-end speaker extraction network with explicit optimization on speaker confusion,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling X-sepformer: End-to-end speaker extraction network with explicit optimization on speaker confusion,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:48.848806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:37.191473Z digest=sha256:be707896d0d9d293da44f4985b234d4ba487a4c73153c21c04fd690e816e9fbb

Observation 5b01aa03-ecc7-4709-9135-d444296e4cfd · outbound

This paper cites Speaker ex- traction with detection of presence and absence of target speakers,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Speaker ex- traction with detection of presence and absence of target speakers,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:48.588117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:37.334445Z digest=sha256:5960b71b3a1fcb653cf1c4647cf464f5b54976fccc530ba967a72890ce780c1a

Observation a27b6499-477a-4aa7-9b33-2cb633290a4c · outbound

This paper cites On the effectiveness of enrollment speech augmentation for target speaker extraction,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling On the effectiveness of enrollment speech augmentation for target speaker extraction,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:48.321486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:37.493824Z digest=sha256:a619d729a55f736fa3b8df19ceb01cafba16c1fc363589cf7cdfb20a40687bd5

Observation 65129ac3-5d07-47d4-a9ac-157856c1a130 · outbound

This paper cites Selective hubert: Self- supervised pre-training for target speaker in clean and mixture speech,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Selective hubert: Self- supervised pre-training for target speaker in clean and mixture speech,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:47.974497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:37.612121Z digest=sha256:07ff6ac140eaf7c6d2075312708eeee3d3179e9e22af9af528a43384cfc13b18

Observation 7ec856e6-7044-4601-8c3e-eae9808aedd0 · outbound

This paper cites Contrastive learning for target speaker extraction with attention-based fusion,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Contrastive learning for target speaker extraction with attention-based fusion,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:47.452568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:37.916014Z digest=sha256:341efcecf49f4b8d56d7ff6a753a55335da61b2f6304595cb4da5f42d9866c45

Observation 5ec61414-d81e-4b3f-9672-5da76eb3834c · outbound

This paper cites Real-time personalised speech enhancement transformers with dynamic cross-attended speaker representations,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Real-time personalised speech enhancement transformers with dynamic cross-attended speaker representations,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:47.148722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:38.081178Z digest=sha256:a3e3127e8a946bd173c9343b10107e28ba209e760944a765227f115260e5fed0

Observation dc186a4c-5358-4844-a292-4e3437f233aa · outbound

This paper cites Smma-net: An audio clue- based target speaker extraction network with spectrogram matching and mu- tual attention,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Smma-net: An audio clue- based target speaker extraction network with spectrogram matching and mu- tual attention,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:46.806035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:38.291163Z digest=sha256:9061125d883db4496cbd35571b401e2cdaf69b560a8cb558f2fe5f14911dc177

Observation fbb47c09-a1a4-4457-b8f6-55b72c625c02 · outbound

This paper cites Target speaker extraction by di- rectly exploiting contextual information in the time-frequency domain,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Target speaker extraction by di- rectly exploiting contextual information in the time-frequency domain,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:46.475563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:38.420322Z digest=sha256:437ca121895b0f126bfab892aff9413739f6567587f44d735a161ad31d875c5e

Observation 4d180dee-aad7-483b-85c5-dc9bd8c3c1cb · outbound

This paper cites Music source separation with band-split rnn,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Music source separation with band-split rnn,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:46.199355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:38.536836Z digest=sha256:a3556e1762ff96727a3deb7b630bef8b41e910d3386839250008eb31f6199a16

Observation c76a0f6f-7617-463d-b028-365e09664def · outbound

This paper cites An algorithm for intelligibility prediction of time–frequency weighted noisy speech,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling An algorithm for intelligibility prediction of time–frequency weighted noisy speech,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:44.836454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:40.037370Z digest=sha256:44f7ad3f61aa4e7e2cb6096a658621b2a90d522d79a986c18a259fcea19ed47a

Observation f235ec62-893c-47f2-8c67-67648fa1cd8b · outbound

This paper cites In defence of metric learning for speaker recognition,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling In defence of metric learning for speaker recognition,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:45.995819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:38.754873Z digest=sha256:c0a097a409ec1286286f588455c72a062a4c71c7bf542011896374daa2e44694

Observation ec33ac18-6273-44d0-a7e4-c91995132715 · outbound

This paper cites Centroid estimation with transformer-based speaker embedder for robust target speaker extraction,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Centroid estimation with transformer-based speaker embedder for robust target speaker extraction,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:45.802159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:38.870085Z digest=sha256:616d8c18f6669284c4312b9eea5df1045bce6c1f1444a8fde602283ee90eb025

Observation 7553ccb9-467f-4907-8590-877601927c26 · outbound

This paper cites Sdr–half-baked or well done?.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Sdr–half-baked or well done?

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:45.638351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:39.010501Z digest=sha256:953bc09d51fdf3e43beb54136e89ea92e0a0799f2d95da6e81f2e0608c217c35

Observation 31d26293-98a0-43d4-9c27-cc536da3f3d7 · outbound

This paper cites Lib- rimix: An open-source dataset for generalizable speech separation,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Lib- rimix: An open-source dataset for generalizable speech separation,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:45.473589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:39.151681Z digest=sha256:a1971f77c3ced50f7e3dd7f625131f4a07848617a824f3e04ab1be5ae8a3a4cd

Observation 49af23a8-a315-4150-a6f8-c19823114eaf · outbound

This paper cites But system description to voxceleb speaker recognition challenge 2019,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling But system description to voxceleb speaker recognition challenge 2019,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:45.304615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:39.253114Z digest=sha256:a6944d28db15ca2c7da726f5b2fdbe32133eded5971c239529e66136dd17537d

Observation f58f258f-bf1c-4927-9cfe-01972083417b · outbound

This paper cites Ecapa-tdnn : Emphasized channel attention, propagation and aggregation in tdnn based speaker verification,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Ecapa-tdnn : Emphasized channel attention, propagation and aggregation in tdnn based speaker verification,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T17:58:39.400949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:58:39.400949Z digest=sha256:f70f9041407b58eed6bf3117938b5d74fabb554fbe89a08290c3432b56fe4240

Observation b8624695-9fa7-4885-92e4-b63d42c05b9f · outbound

This paper cites Wespeaker: A research and production oriented speaker embed- ding learning toolkit,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Wespeaker: A research and production oriented speaker embed- ding learning toolkit,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:45.136229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:39.509318Z digest=sha256:1278f01a5d7c634b098829747035e9ba5a19da13f5ceed2648f808e2eb89aa9e

Observation d861da22-9724-4e98-9b00-3e6cf57d6b19 · outbound

This paper cites V oxceleb2: Deep speaker recognition,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling V oxceleb2: Deep speaker recognition,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T17:58:39.616004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:58:39.616004Z digest=sha256:fc8b4654f12c6a4f232e07cb5010c5349a850e757aa31fb948e4b94eb2cc4aa1

Observation 9beb7c1c-b899-43c8-a306-d54f488fcbd1 · outbound

This paper cites WeSep: A Scalable and Flexible Toolkit Towards Generalizable Target Speaker Extraction.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling WeSep: A Scalable and Flexible Toolkit Towards Generalizable Target Speaker Extraction

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-08-06T17:58:42.352929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:39.732096Z digest=sha256:ca0c75912751797d127c7221b0ae8717813c00cc06364adb4d25f6a799008197

Observation 19611a75-f91d-49d7-82f2-05ece02de0c5 · outbound

This paper cites Perceptual evaluation of speech quality (pesq): An objective method for end-to-end speech quality assessment of narrow-band telephone networks and speech codecs,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Perceptual evaluation of speech quality (pesq): An objective method for end-to-end speech quality assessment of narrow-band telephone networks and speech codecs,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:44.962435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:39.867589Z digest=sha256:34ec0f2f86eb1d8c3dd07043caf6e028c73fc64f928af9f644351247de8c8a6b

Observation 2b029e47-5f9f-4c78-a18c-4ff84e909f84 · outbound

This paper cites Xtts: a massively mul- tilingual zero-shot text-to-speech model,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Xtts: a massively mul- tilingual zero-shot text-to-speech model,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:44.620467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:40.186751Z digest=sha256:ad49a17c7fa8fdaed5584e9435a70817807539d60a56c1f5466263c4e3f16b18

Observation 08b27bca-02f0-49a5-9cbb-97cca54c4648 · outbound

This paper cites Ecapa2: A hybrid neural network archi- tecture and training strategy for robust speaker embeddings,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Ecapa2: A hybrid neural network archi- tecture and training strategy for robust speaker embeddings,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:44.343895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:40.303176Z digest=sha256:7f29916ca84d85cbaba8d810124ede95f7fbbf3c4e7e8683ead40815d627eab9

Observation 4cb35800-a4d5-4da0-8aa0-d612e50c931e · outbound

This paper cites Multi-Level Speaker Representation for Target Speaker Extraction.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Multi-Level Speaker Representation for Target Speaker Extraction

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-08-06T17:58:42.148745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:40.451824Z digest=sha256:784f60bb4abb97d9fe46d889fd56ff3daa8121edce44a6f4d6de674acc142de4

Observation 053c27d9-27b9-4d9a-a5eb-f6d36f47f9b4 · outbound

This paper cites Dpccn: Densely-connected pyramid complex convolutional network for robust speech separation and extraction,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Dpccn: Densely-connected pyramid complex convolutional network for robust speech separation and extraction,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:44.054911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:40.578466Z digest=sha256:74761a22ab7e86e5007d750bd8d7705c513e03e07f21d71b3c67da723db6c945

Observation 9850ff13-8b76-4491-a351-cabbcf7af160 · outbound

This paper cites Mc-spex: Towards effective speaker extraction with multi-scale interfu- sion and conditional speaker modulation,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Mc-spex: Towards effective speaker extraction with multi-scale interfu- sion and conditional speaker modulation,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:43.814757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:40.670334Z digest=sha256:184206da3c9d839b0362afa2eb07f7a6658d5bd466ff44fe559e6c413c20eecd

Observation de256083-5227-41d8-b2ef-3a28a4986f1a · outbound

This paper cites Tar- get speech extraction with pre-trained self-supervised learning models,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Tar- get speech extraction with pre-trained self-supervised learning models,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:47.697564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:40.797574Z digest=sha256:fd67ed6572734137ab39192efc546a14ef24995bd41011f4145c557c897e26f3

Observation 531299cb-327f-401b-804b-7061af17d7fb · outbound

This paper cites Tf- gridnet: Making time-frequency domain models great again for monaural speaker separation,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Tf- gridnet: Making time-frequency domain models great again for monaural speaker separation,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:43.595248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:40.942430Z digest=sha256:960c1f5dbd838f32c7f71104520a68874ec12cb342c05647a79f981d6d968119

Observation 8e319a68-b513-4e42-8018-03d56c53fe1a · outbound

This paper cites V oxceleb: A large-scale speaker identification dataset,.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling V oxceleb: A large-scale speaker identification dataset,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:58:43.348936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:41.063893Z digest=sha256:431868c7ded8f09f904bcf4a14be85aaf5d16d68700c88ba9b35c8bd2d646bf5

Observation 4202cf03-9f11-4a00-80b4-ede8b5b33f86 · outbound

This paper cites an unresolved cited work.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Unresolved cited work

Reference 128

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:58:52.159552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:34.919879Z digest=sha256:08b0c229e531f224835f3bd0305832260f7335679cf01cac47e1b4339835c223

Pith citing papers

Observation 460796ea-b9b2-49bc-a537-08fdb67af52d · inbound

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling cites this paper.

Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling Enhancing Target Speaker Extraction with Explicit Speaker Consistency Modeling

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-06T17:58:42.987162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T17:58:34.654304Z digest=sha256:7eb1821475b3e56e86db655861a90c084be2ab1c78b7e62e8584755cd1e7fe7e