Pith. sign in

Paper Citation Record · LEDGER

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions

As of 23 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 0 inbound Pith citation observations for arXiv:2502.08191.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.08191 v1

Coverage vector

measured 36 of 36 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T10:06:36.061223Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

36 of 36 outbound references displayed

  • verified exact3
  • verified fuzzy25
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3747b99a-2b8c-4218-974f-f3bf502692f7 · outbound

This paper cites Some experiments on the recognition of speech, with one and with two ears,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Some experiments on the recognition of speech, with one and with two ears,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T10:06:35.923599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:06:35.923599Z digest=sha256:25edd7d6dcd61586e1dab323d8abb0010522f9402e78b8fa36c0c8e439372fa7

Observation 089cb7e3-f9e3-4167-9c3a-f6ab29ef4689 · outbound

This paper cites The cocktail party phenomenon revisited: The importance of working memory capacity,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions The cocktail party phenomenon revisited: The importance of working memory capacity,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.571334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-08T10:06:35.928089Z digest=sha256:a5fd0ab6f3d13f98b81d227aef3c6681736b269dc35c4f85af8ecca8077d2e5a

Observation 2bea3240-35d6-4ad0-8511-c50204933188 · outbound

This paper cites An event-related potential study of selective auditory attention in children and adults,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions An event-related potential study of selective auditory attention in children and adults,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.556553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-08T10:06:35.932253Z digest=sha256:c463f1fc239925c4d373c6c6bfabb4e87120ff4cb9cb849a0e3381700532bb4e

Observation 5857d84c-5ee4-4e63-abce-94cf6a016643 · outbound

This paper cites Selective cortical representation of attended speaker in multi-talker speech perception,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Selective cortical representation of attended speaker in multi-talker speech perception,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.543202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-08T10:06:35.937068Z digest=sha256:7084ebf6f30234e95e74c2568f3892f831fd64d55b07a6155978504b4100fd55

Observation 2e61c42f-8f5d-4c2d-a605-fcd7010b3ed8 · outbound

This paper cites Supervised speech separation based on deep learning: An overview,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Supervised speech separation based on deep learning: An overview,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-08T10:06:35.941317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:06:35.941317Z digest=sha256:be3caab5d93e69c9f5a9fbc979d246eaff51eedd5c0191bccc018675ad67dec5

Observation eef047c9-031f-447f-88bf-8363d0e36557 · outbound

This paper cites Dual-path rnn: efficient long sequence modeling for time-domain single- channel speech separation,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Dual-path rnn: efficient long sequence modeling for time-domain single- channel speech separation,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.520393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-08T10:06:35.946093Z digest=sha256:491bd27cf5a447f53c14c7b71774d7a5e6d30c2fccf9d4e9d351798fe199aeb6

Observation c6eeedb6-77db-4bcb-9040-bc30aa612f7f · outbound

This paper cites Dual-Path Transformer Network: Direct Context-Aware Modeling for End-to-End Monaural Speech Separation.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Dual-Path Transformer Network: Direct Context-Aware Modeling for End-to-End Monaural Speech Separation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T10:06:35.950643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:06:35.950643Z digest=sha256:b8013956e0b0c32f17a403c76ba77fbfce44b364efbb43403c9d4c42e35059e7

Observation 92dc1bb7-45be-4719-9d83-723f08d71608 · outbound

This paper cites Tf-gridnet: Making time-frequency domain models great again for monaural speaker separation,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Tf-gridnet: Making time-frequency domain models great again for monaural speaker separation,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.507409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-08T10:06:35.955445Z digest=sha256:77b1b8470431b29c743e73f671d2a2ad9e235b869bc4364fdef622d9c04fb516

Observation 221bdceb-83a8-44c7-a5f5-b735fd482ccb · outbound

This paper cites CrossNet: Leveraging Global, Cross-Band, Narrow-Band, and Positional Encoding for Single- and Multi-Channel Speaker Separation.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions CrossNet: Leveraging Global, Cross-Band, Narrow-Band, and Positional Encoding for Single- and Multi-Channel Speaker Separation

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-08T10:06:36.189043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-08T10:06:35.959087Z digest=sha256:bcafbfb8f9014f1285b524c3edf8aacdfddb3e15d9ec3a30371c9f60d0ef0ff9

Observation b94022f0-c4fe-48e1-8e4b-33f3e08fa950 · outbound

This paper cites Neural target speech extraction: An overview,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Neural target speech extraction: An overview,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.493809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-08T10:06:35.963133Z digest=sha256:b6140970236e6141113bba8bdf49c1067b477a7e9afa4d86801a392870f3a6a0

Observation 2210c1a4-6f72-4a10-a827-355990ffd65a · outbound

This paper cites VoiceFilter: Targeted Voice Separation by Speaker-Conditioned Spectrogram Masking.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions VoiceFilter: Targeted Voice Separation by Speaker-Conditioned Spectrogram Masking

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T10:06:35.967023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:06:35.967023Z digest=sha256:4fa055e4bd84a94d412c39d0f5d6b62fe6b7abfc088397bd28be52e83e2a16d4

Observation 48c25eb6-4f8b-47b3-b2fb-195fdaa8e166 · outbound

This paper cites Time-domain speaker extraction network,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Time-domain speaker extraction network,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.480225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-08T10:06:35.971084Z digest=sha256:cc74f2ad316e0dc1400ec05ff1418a725423a099f762d4af42dc8460824bdf33

Observation 6651f7ab-cec5-42a8-b726-e770ba0fb542 · outbound

This paper cites Spex: Multi-scale time domain speaker extraction network,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Spex: Multi-scale time domain speaker extraction network,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.467488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-08T10:06:35.975088Z digest=sha256:032e4907c060130c03bf6dfbdf0b78d3b4a3bd97784fd303b8f4375669d89ee9

Observation 6b72a265-deb3-4807-b0da-30af340934d6 · outbound

This paper cites SpEx+: A Complete Time Domain Speaker Extraction Network.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions SpEx+: A Complete Time Domain Speaker Extraction Network

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T10:06:35.978757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:06:35.978757Z digest=sha256:67f785458f89e8e744f68ec472ee66b42d052a7130f3d144356a18fc4a82e5f1

Observation c9959fb2-8b07-40a2-adbd-4e8b258d9cac · outbound

This paper cites Adaptive-spex: Local and global perceptual modeling with speaker adaptation for target speaker extraction,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Adaptive-spex: Local and global perceptual modeling with speaker adaptation for target speaker extraction,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.454231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-08T10:06:35.982669Z digest=sha256:15a4825674f665907e9bd3c9d234de9d7f1bbea1aa8d11ba6b81a98d6dc89fde

Observation 4682cbad-a6bd-4abd-a463-b0d7ff93328a · outbound

This paper cites Multi-stage speaker extraction with utterance and frame-level reference signals,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Multi-stage speaker extraction with utterance and frame-level reference signals,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.441070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-08T10:06:35.986358Z digest=sha256:8c606559506914a9c45948f6d77fc239bd8b623f79437a3feb72c5cfb188566d

Observation 57a39665-19a3-422c-8372-87685bc5b386 · outbound

This paper cites Neural speaker extraction with speaker-speech cross-attention network.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Neural speaker extraction with speaker-speech cross-attention network

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.427846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-08T10:06:35.990204Z digest=sha256:e7d88f3cbcc35a3b0151732566fa65d474af37b818d8f70bd6ce0ea59c4599cc

Observation 1e563c58-4a7d-4181-b7a9-78217b2db8e5 · outbound

This paper cites Robust Speaker Extraction Network Based on Iterative Refined Adaptation.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Robust Speaker Extraction Network Based on Iterative Refined Adaptation

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-08T10:06:36.140824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-08T10:06:35.993768Z digest=sha256:5a76bdecafdf8a75696f50d5afd3ec236a09d25b9e18d8e582e3f79db30e61d0

Observation 0ace271d-d7ea-45ab-8126-78bae64fccb3 · outbound

This paper cites Target speaker extraction with ultra-short reference speech by ve-ve framework,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Target speaker extraction with ultra-short reference speech by ve-ve framework,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.414639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-08T10:06:35.997632Z digest=sha256:f3c4a5a792e2c55fd50bb23b6a5d002b2297194d48dd0a4dd67c9741d9be05c1

Observation 2cb27952-e0e2-48a1-b9ac-741ab2cd8a21 · outbound

This paper cites X-sepformer: End-to-end speaker extraction network with explicit optimization on speaker confusion,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions X-sepformer: End-to-end speaker extraction network with explicit optimization on speaker confusion,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.400779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-08T10:06:36.001194Z digest=sha256:0fef4c7d49e5a5c14e33429e36b26ba3ecf54e22a57ed331fa826625761a9d42

Observation 1bc63093-491e-4ee0-90c1-082d190792a7 · outbound

This paper cites Target speech extraction with pre-trained self-supervised learning models,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Target speech extraction with pre-trained self-supervised learning models,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.387373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-08T10:06:36.005004Z digest=sha256:fddd06757da07503bf8a577e1f190c02a990c23422fac8afa05571a653b832fc

Observation 49c8b380-dd29-4ecb-b1eb-3cadf93c1fc4 · outbound

This paper cites Speakerbeam: Speaker aware neural network for target speaker extraction in speech mixtures,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Speakerbeam: Speaker aware neural network for target speaker extraction in speech mixtures,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.372641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-08T10:06:36.008561Z digest=sha256:06685211461321b2cf36456b35aa34e8f65e94e640e697c6280e0e95075db6da

Observation 437363be-9fa3-4d84-9def-bbffade4d16b · outbound

This paper cites X-tf-gridnet: A time–frequency domain target speaker extraction network with adaptive speaker embedding fusion,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions X-tf-gridnet: A time–frequency domain target speaker extraction network with adaptive speaker embedding fusion,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.357442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-08T10:06:36.012214Z digest=sha256:eadbc848132d8325717715c57d41664f1d11994797b1a469cbb6177d451398e6

Observation 008e86cc-b5c9-4fb0-a838-e68530ce945e · outbound

This paper cites Single-channel speech extraction using speaker inventory and attention network,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Single-channel speech extraction using speaker inventory and attention network,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.342905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-08T10:06:36.015841Z digest=sha256:e16aada4fa37c9885c1a5fd293d4d0c667753eef7c044aa22cd86b0b13d6d6a4

Observation 0348d1ef-8c9e-472a-8489-c33df89c7751 · outbound

This paper cites Sef-net: Speaker embedding free target speaker extraction network,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Sef-net: Speaker embedding free target speaker extraction network,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.329135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-08T10:06:36.019330Z digest=sha256:7edbfcbb54ee1a04ea3decfa4c6fe242a239fbd8126a198952d6a4cd8d3091fb

Observation cf045db9-2cf4-43e4-8829-9899f3e07ab9 · outbound

This paper cites Target speaker extraction by directly exploiting contextual information in the time-frequency domain,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Target speaker extraction by directly exploiting contextual information in the time-frequency domain,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.316022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-08T10:06:36.023295Z digest=sha256:ad509801d0b454cfdf88b5e778565552440fc8827d35d67bdc2c725db4d12b88

Observation 8450e850-d73e-4009-bd77-31f799426d3b · outbound

This paper cites On the importance of power compression and phase estimation in monaural speech dereverberation,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions On the importance of power compression and phase estimation in monaural speech dereverberation,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.302169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-08T10:06:36.026941Z digest=sha256:a38f73b78ba2c82520ea1edd368a665826348d7a04d0f4d1331222a17c7a3f26

Observation 3d76b5a3-25e9-42c8-b7a8-6ea956d8929b · outbound

This paper cites Root mean square layer normalization,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Root mean square layer normalization,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-08T10:06:36.030834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:06:36.030834Z digest=sha256:76263514ddff125e0f24683ab3964248479c36b5f315067b4160ecb8bf934f1c

Observation f21e3ffd-f1ea-4a4b-a201-08815658a7e1 · outbound

This paper cites Single image reflection separation via component synergy,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Single image reflection separation via component synergy,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.279476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-08T10:06:36.034676Z digest=sha256:fdba3d96b50b04ec2b729b1ecfebe8ac448de656af672fafa355f94eb387d511

Observation 833b93e9-6b0d-45eb-b5b9-fdddd2923c9d · outbound

This paper cites Squeeze-and-excitation networks,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Squeeze-and-excitation networks,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-08T10:06:36.038354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:06:36.038354Z digest=sha256:610a43f30c3b0e5ff2899957a100b837b2069b15b5e990397cbd1ef291945047

Observation 988293ba-2ded-4372-8ebd-54de16006b57 · outbound

This paper cites Sdr–half-baked or well done?.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Sdr–half-baked or well done?

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.257046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-08T10:06:36.041955Z digest=sha256:4752f356e10aecdc00a405a95b60abb120e2258f9ec3478f375e63b5bdda78be

Observation ecfaf012-9d1a-4967-b870-84dc9e0d61fb · outbound

This paper cites Csr-i (wsj0) complete ldc93s6a,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Csr-i (wsj0) complete ldc93s6a,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.243969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-08T10:06:36.045599Z digest=sha256:bc13790753f0d7cf9663c982235d109e8c21675a94f3371753a13128887051c1

Observation fe23e7e1-d0a2-492b-a0ef-d66c5a0e5c28 · outbound

This paper cites WHAM!: Extending Speech Separation to Noisy Environments.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions WHAM!: Extending Speech Separation to Noisy Environments

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-08T10:06:36.049963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:06:36.049963Z digest=sha256:74e37ed5bd3cf39d776c5689985689165c952ed43cdddf2e5842dab5a7541fe5

Observation a15ad033-fe13-4d55-851e-d81c5112c7ad · outbound

This paper cites Whamr!: Noisy and reverberant single-channel speech separation,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Whamr!: Noisy and reverberant single-channel speech separation,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.230796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-08T10:06:36.053896Z digest=sha256:aa5fb16af4446e7df4ec6276cfbc24ae14a3c7fd2ae7f124aca4facd713a4369

Observation 0a2cad35-1e31-4976-8d1a-f57a28e0ed81 · outbound

This paper cites Target Confusion in End-to-end Speaker Extraction: Analysis and Approaches.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Target Confusion in End-to-end Speaker Extraction: Analysis and Approaches

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-08-08T10:06:36.105012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-08T10:06:36.057455Z digest=sha256:fed6105f3d09cc7d3cedf8b26c116b9246965d2f2bd6ed887fa274be35011233

Observation 4d426059-eea5-4d34-b37b-064985ae43e2 · outbound

This paper cites Attention is all you need,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Attention is all you need,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.217503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-08T10:06:36.061223Z digest=sha256:c2d81a055822a785074e1436a79ef0f05504d4249a4043448e3b6f581e8dd49a

Pith citing papers

No inbound Pith citation observations are available.