Pith. sign in

Paper Citation Record · LEDGER

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions

As of 11 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 0 inbound Pith citation observations for arXiv:2502.08191.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.08191 v1

Coverage vector

measured 36 of 36 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T10:06:36.061223Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

36 of 36 outbound references displayed

  • verified exact3
  • verified fuzzy25
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3747b99a-2b8c-4218-974f-f3bf502692f7 · outbound

This paper cites Some experiments on the recognition of speech, with one and with two ears,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Some experiments on the recognition of speech, with one and with two ears,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T10:06:35.923599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:06:35.923599Z digest=sha256:05af6c9bcb5215279793da9308bf7669a4eb114d33e3a55a265315c4adc61712

Observation 089cb7e3-f9e3-4167-9c3a-f6ab29ef4689 · outbound

This paper cites The cocktail party phenomenon revisited: The importance of working memory capacity,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions The cocktail party phenomenon revisited: The importance of working memory capacity,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.571334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T10:06:35.928089Z digest=sha256:5bcd5a4bba3d7f5f5347085ce445240f089279dc09196f2d9d47f2cdea1c9866

Observation 2bea3240-35d6-4ad0-8511-c50204933188 · outbound

This paper cites An event-related potential study of selective auditory attention in children and adults,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions An event-related potential study of selective auditory attention in children and adults,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.556553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T10:06:35.932253Z digest=sha256:ebffc60e2c34bb0360824c79acc1ba4b96a4997091f45630f99a297f3574c90f

Observation 5857d84c-5ee4-4e63-abce-94cf6a016643 · outbound

This paper cites Selective cortical representation of attended speaker in multi-talker speech perception,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Selective cortical representation of attended speaker in multi-talker speech perception,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.543202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T10:06:35.937068Z digest=sha256:95e51ea308de06b12196512c768e344ac4944cad2648b84b681d9846be270641

Observation 2e61c42f-8f5d-4c2d-a605-fcd7010b3ed8 · outbound

This paper cites Supervised speech separation based on deep learning: An overview,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Supervised speech separation based on deep learning: An overview,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-08T10:06:35.941317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:06:35.941317Z digest=sha256:97ea588838772e3568882462b857077a3124b2bbf062ab3483d34b7bdb200416

Observation eef047c9-031f-447f-88bf-8363d0e36557 · outbound

This paper cites Dual-path rnn: efficient long sequence modeling for time-domain single- channel speech separation,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Dual-path rnn: efficient long sequence modeling for time-domain single- channel speech separation,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.520393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T10:06:35.946093Z digest=sha256:8a35cd979e30a01fd9228ef822277c632d4ec7cdbd9b5fcc88f38f7c3f0d8473

Observation c6eeedb6-77db-4bcb-9040-bc30aa612f7f · outbound

This paper cites Dual-Path Transformer Network: Direct Context-Aware Modeling for End-to-End Monaural Speech Separation.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Dual-Path Transformer Network: Direct Context-Aware Modeling for End-to-End Monaural Speech Separation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T10:06:35.950643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:06:35.950643Z digest=sha256:724fa26d96cbac08a373d51c77cf6c821bd0c9a0030f272d6d9d47fd80422093

Observation 92dc1bb7-45be-4719-9d83-723f08d71608 · outbound

This paper cites Tf-gridnet: Making time-frequency domain models great again for monaural speaker separation,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Tf-gridnet: Making time-frequency domain models great again for monaural speaker separation,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.507409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T10:06:35.955445Z digest=sha256:0807ff7b888f651a0605099abb663b63f64b3c1a23e515bdc1ec412e2de54400

Observation 221bdceb-83a8-44c7-a5f5-b735fd482ccb · outbound

This paper cites CrossNet: Leveraging Global, Cross-Band, Narrow-Band, and Positional Encoding for Single- and Multi-Channel Speaker Separation.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions CrossNet: Leveraging Global, Cross-Band, Narrow-Band, and Positional Encoding for Single- and Multi-Channel Speaker Separation

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-08T10:06:36.189043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T10:06:35.959087Z digest=sha256:6c87c7eebddef5bd3307f11720cd9e4bbce30db7e75de7084e3a68b679ffdd7d

Observation b94022f0-c4fe-48e1-8e4b-33f3e08fa950 · outbound

This paper cites Neural target speech extraction: An overview,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Neural target speech extraction: An overview,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.493809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T10:06:35.963133Z digest=sha256:39214fbae7f7cbafe26230eaef35d6cc7c7e2bc1cd50b2bc35060088afa54cd4

Observation 2210c1a4-6f72-4a10-a827-355990ffd65a · outbound

This paper cites VoiceFilter: Targeted Voice Separation by Speaker-Conditioned Spectrogram Masking.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions VoiceFilter: Targeted Voice Separation by Speaker-Conditioned Spectrogram Masking

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T10:06:35.967023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:06:35.967023Z digest=sha256:154bdc2348bb8ba26ea4cced76086ff019406262b8679e1e4e0b538632afe7bb

Observation 48c25eb6-4f8b-47b3-b2fb-195fdaa8e166 · outbound

This paper cites Time-domain speaker extraction network,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Time-domain speaker extraction network,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.480225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T10:06:35.971084Z digest=sha256:f71090647406a1768d32e4b89aa5bc70cc849e1ea01fca8f44346e91fea675c6

Observation 6651f7ab-cec5-42a8-b726-e770ba0fb542 · outbound

This paper cites Spex: Multi-scale time domain speaker extraction network,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Spex: Multi-scale time domain speaker extraction network,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.467488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T10:06:35.975088Z digest=sha256:2f51df0d689a1bd32151f52a79b4ed1f7dc56886b1e68463b3a1304b1b68cf3e

Observation 6b72a265-deb3-4807-b0da-30af340934d6 · outbound

This paper cites SpEx+: A Complete Time Domain Speaker Extraction Network.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions SpEx+: A Complete Time Domain Speaker Extraction Network

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T10:06:35.978757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:06:35.978757Z digest=sha256:e65b25cff8de13f673aca9c73e91ebf2207e7ab0ca01a5059cf6b94cb6eb9955

Observation c9959fb2-8b07-40a2-adbd-4e8b258d9cac · outbound

This paper cites Adaptive-spex: Local and global perceptual modeling with speaker adaptation for target speaker extraction,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Adaptive-spex: Local and global perceptual modeling with speaker adaptation for target speaker extraction,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.454231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T10:06:35.982669Z digest=sha256:3528c1960ff489ac2e64e552374db6942c263d0ef80b26415e93e1a8a762f84d

Observation 4682cbad-a6bd-4abd-a463-b0d7ff93328a · outbound

This paper cites Multi-stage speaker extraction with utterance and frame-level reference signals,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Multi-stage speaker extraction with utterance and frame-level reference signals,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.441070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T10:06:35.986358Z digest=sha256:359177a71742d1e2bf1917af77e549532dd701b8937c5ad7a5c7b53f8a5d4f3c

Observation 57a39665-19a3-422c-8372-87685bc5b386 · outbound

This paper cites Neural speaker extraction with speaker-speech cross-attention network.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Neural speaker extraction with speaker-speech cross-attention network

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.427846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T10:06:35.990204Z digest=sha256:5a5a6863dbd86f34e1d2aec3d1f67ec6bad19d3e61215e51a080514c440ae129

Observation 1e563c58-4a7d-4181-b7a9-78217b2db8e5 · outbound

This paper cites Robust Speaker Extraction Network Based on Iterative Refined Adaptation.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Robust Speaker Extraction Network Based on Iterative Refined Adaptation

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-08T10:06:36.140824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T10:06:35.993768Z digest=sha256:e097ae268c72ce83ece2ed1d6bdd7c4ebe1c778a71d32cd910805f46dc30eabb

Observation 0ace271d-d7ea-45ab-8126-78bae64fccb3 · outbound

This paper cites Target speaker extraction with ultra-short reference speech by ve-ve framework,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Target speaker extraction with ultra-short reference speech by ve-ve framework,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.414639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T10:06:35.997632Z digest=sha256:f68d220e0dc89ec8e929d1afefeb9266b4e987e59e37f997e34baeb87a965d0c

Observation 2cb27952-e0e2-48a1-b9ac-741ab2cd8a21 · outbound

This paper cites X-sepformer: End-to-end speaker extraction network with explicit optimization on speaker confusion,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions X-sepformer: End-to-end speaker extraction network with explicit optimization on speaker confusion,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.400779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T10:06:36.001194Z digest=sha256:af3ccdf4413e19ef28887f74c33c20ec6f837b3239b4e3e558b157c20f270159

Observation 1bc63093-491e-4ee0-90c1-082d190792a7 · outbound

This paper cites Target speech extraction with pre-trained self-supervised learning models,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Target speech extraction with pre-trained self-supervised learning models,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.387373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T10:06:36.005004Z digest=sha256:fdf186492d0edaba52dd11ba1f781d4f5a2f5b1ab898100c55dc6e8585b7d7eb

Observation 49c8b380-dd29-4ecb-b1eb-3cadf93c1fc4 · outbound

This paper cites Speakerbeam: Speaker aware neural network for target speaker extraction in speech mixtures,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Speakerbeam: Speaker aware neural network for target speaker extraction in speech mixtures,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.372641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T10:06:36.008561Z digest=sha256:5faa33ecf29508dbc56e27335ff12075c59ec8e58110f7fd3d0e13295993049d

Observation 437363be-9fa3-4d84-9def-bbffade4d16b · outbound

This paper cites X-tf-gridnet: A time–frequency domain target speaker extraction network with adaptive speaker embedding fusion,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions X-tf-gridnet: A time–frequency domain target speaker extraction network with adaptive speaker embedding fusion,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.357442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T10:06:36.012214Z digest=sha256:d03ea6741a23173fa7de92723700a268106cfc8f8068353dbaeb13705a283374

Observation 008e86cc-b5c9-4fb0-a838-e68530ce945e · outbound

This paper cites Single-channel speech extraction using speaker inventory and attention network,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Single-channel speech extraction using speaker inventory and attention network,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.342905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T10:06:36.015841Z digest=sha256:43c3137968f826218b2f07b2929cfd3a70eca7dd8f93370f4997840284e2b055

Observation 0348d1ef-8c9e-472a-8489-c33df89c7751 · outbound

This paper cites Sef-net: Speaker embedding free target speaker extraction network,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Sef-net: Speaker embedding free target speaker extraction network,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.329135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T10:06:36.019330Z digest=sha256:05fd30a16dab6b4ee2ad890001d28d99da0ad4e4bd10f6956f0006acbe12fa23

Observation cf045db9-2cf4-43e4-8829-9899f3e07ab9 · outbound

This paper cites Target speaker extraction by directly exploiting contextual information in the time-frequency domain,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Target speaker extraction by directly exploiting contextual information in the time-frequency domain,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.316022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T10:06:36.023295Z digest=sha256:36af6ae3dbcad2024e700e388f07f3f96b0557641a0df1b0a45660aced7a6551

Observation 8450e850-d73e-4009-bd77-31f799426d3b · outbound

This paper cites On the importance of power compression and phase estimation in monaural speech dereverberation,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions On the importance of power compression and phase estimation in monaural speech dereverberation,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.302169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T10:06:36.026941Z digest=sha256:f74aa17e75a5d946c4e42b4ee4a716763b2a6038c80bb86d21f105045e9ec3f5

Observation 3d76b5a3-25e9-42c8-b7a8-6ea956d8929b · outbound

This paper cites Root mean square layer normalization,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Root mean square layer normalization,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-08T10:06:36.030834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:06:36.030834Z digest=sha256:66e2477f59be4e0e016a5dd97f49e9c4eef27eccc6bf9bdfe0844f62db836581

Observation f21e3ffd-f1ea-4a4b-a201-08815658a7e1 · outbound

This paper cites Single image reflection separation via component synergy,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Single image reflection separation via component synergy,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.279476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T10:06:36.034676Z digest=sha256:e43640fa0784dbd28313ae3e3015d078c473cb527713d6bfb0b76a7b1f32f094

Observation 833b93e9-6b0d-45eb-b5b9-fdddd2923c9d · outbound

This paper cites Squeeze-and-excitation networks,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Squeeze-and-excitation networks,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-08T10:06:36.038354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:06:36.038354Z digest=sha256:442df7ab09484b7985b26be5f1ef06787c3ccd274ba80ca6989ead8dad89d977

Observation 988293ba-2ded-4372-8ebd-54de16006b57 · outbound

This paper cites Sdr–half-baked or well done?.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Sdr–half-baked or well done?

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.257046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T10:06:36.041955Z digest=sha256:4c12e725a3cd55a76ac98162bbbda6a539b7b61d4e322842497ab35a3f42e29c

Observation ecfaf012-9d1a-4967-b870-84dc9e0d61fb · outbound

This paper cites Csr-i (wsj0) complete ldc93s6a,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Csr-i (wsj0) complete ldc93s6a,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.243969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T10:06:36.045599Z digest=sha256:0ce1946ae9a67e7442aee5a32774ac45e402c7c136e09ef3b00feef3dfe25b13

Observation fe23e7e1-d0a2-492b-a0ef-d66c5a0e5c28 · outbound

This paper cites WHAM!: Extending Speech Separation to Noisy Environments.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions WHAM!: Extending Speech Separation to Noisy Environments

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-08T10:06:36.049963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:06:36.049963Z digest=sha256:b7c65e9126e8ed640627e4820907e1d23f1cc47db9deee34012f05b238417998

Observation a15ad033-fe13-4d55-851e-d81c5112c7ad · outbound

This paper cites Whamr!: Noisy and reverberant single-channel speech separation,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Whamr!: Noisy and reverberant single-channel speech separation,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.230796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T10:06:36.053896Z digest=sha256:78432af3b60266cea959c514bd8a003b0d7acf95e404e3e10e0421fc3d15fea1

Observation 0a2cad35-1e31-4976-8d1a-f57a28e0ed81 · outbound

This paper cites Target Confusion in End-to-end Speaker Extraction: Analysis and Approaches.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Target Confusion in End-to-end Speaker Extraction: Analysis and Approaches

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-08-08T10:06:36.105012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T10:06:36.057455Z digest=sha256:ac2158b1679364617958cab772ac3f2d4404673bd2bde6204efe80dd4c040427

Observation 4d426059-eea5-4d34-b37b-064985ae43e2 · outbound

This paper cites Attention is all you need,.

DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions Attention is all you need,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:06:36.217503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T10:06:36.061223Z digest=sha256:7e54794f5f28c2666c6b6d920934df81a7b9360e09c88b08f3ffeea0e03e7879

Pith citing papers

No inbound Pith citation observations are available.