Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:08:17.836014Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 2 inbound Pith citation observations for arXiv:2506.00466.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:08:17.836014Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T18:41:16.865462Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-06T18:41:18.737055Z
38 of 38 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation bc195b6e-d5d0-4f63-a9d0-04ac866c9326 · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction Electrophysiological correlates of semantic dissimilarity reflect the comprehension of natural, narra- tive speech.Current Biology, 28(5):803–809,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8f307be1-86c4-49ac-8956-6f9759ca9460 · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction Improved Feature Extraction Network for Neuro-Oriented Target Speaker Extraction
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 13523b52-39f6-431a-b171-19ca9d351e34 · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction L-spex: Localized target speaker extraction
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e2974a40-42d1-4de0-8110-e608a5038a5f · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction The cocktail party problem.Neural computation, 17(9):1875– 1902,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation df764bc2-ffa9-4846-8928-5a6b8931a09f · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction Speaker-independent brain enhanced speech denoising
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3c2dcafd-2de0-422b-af07-0b1cda3acfb3 · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction Cross-modal global interaction and local alignment for audio-visual speech recognition
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5fa6fbf7-6719-4168-b5c3-cc96e9102fa3 · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction [Le Rouxet al., 2019 ] Jonathan Le Roux, Scott Wisdom, Hakan Erdogan, and John R Hershey
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation efeb095b-13d3-4666-b5b3-3425874acfd2 · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction Align before fuse: Vision and language representation learning with momentum distilla- tion.Advances in neural information processing systems, 34:9694–9705,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2efd5f25-00e7-4d3f-8874-eb720ed719a3 · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction Audio-visual active speaker extraction for sparsely overlapped multi-talker speech
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 51088153-38b9-44cf-9d43-d86685d3973a · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction Av- sepformer: Cross-attention sepformer for audio-visual tar- get speaker extraction
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bafb6e57-f5e0-45ae-8f6c-cb2825084cc9 · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction Development of the audi- tory system.Handbook of clinical neurology, 129:55–72,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5bd4c207-95ca-4561-bd42-5119b2fd12e1 · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction Dual-path rnn: efficient long sequence modeling for time-domain single-channel speech separation
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1556424d-5a68-4f54-a3d0-172c329858fa · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction Dbpnet: Dual- branch parallel network with temporal-frequency fusion for auditory attention detection
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 37e4535f-4d26-4e5c-9eeb-a1b405a74090 · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction Attentional selection in a cocktail party environment can be decoded from single- trial eeg.Cerebral cortex, 25(7):1697–1706,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6046fe9f-d1f6-4977-8a76-b8d3c9047317 · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction Neural decoding of at- tentional selection in multi-speaker environments without access to clean sources.Journal of neural engineering, 14(5):056001,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b3e6997e-0ba2-4970-ac42-73566d32e77a · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction Neu- roheed+: Improving neuro-steered speaker extraction with joint auditory attention detection
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 20c99bcf-0126-4174-a528-1a75a08b5db6 · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction Tf-nsse: A time–frequency domain neuro-steered speaker extractor.Applied Acous- tics, 211:109519,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6acfd61d-a7aa-40dc-a6c3-b881a3eb6a0e · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction Phase space graph convolutional network for chaotic time series learning.IEEE Transactions on Industrial Informat- ics,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4beca81b-dc3b-475c-aa27-943dd8141bc5 · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction Perceptual evaluation of speech quality (pesq)-a new method for speech quality assessment of telephone networks and codecs
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8e2a4f39-bf0f-4ec0-ae23-f64f8f361103 · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction An algorithm for intelligibil- ity prediction of time–frequency weighted noisy speech
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3cb4f928-39ea-4f02-acbd-1f4a5fbd31e0 · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction A study of multichannel spatiotemporal features and knowledge distillation on robust target speaker extraction
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ff249c62-b548-44b8-aa3d-a2595cb51d0a · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction Spex: Multi-scale time domain speaker extraction network.IEEE/ACM transactions on audio, speech, and language processing, 28:1370–1384,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3ad76a10-e7c1-4e4d-8d48-72455b9e78c3 · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction DARNet: Dual Attention Refinement Network with Spatiotemporal Construction for Auditory Attention Detection
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bc8e5ec2-d71f-41e9-8950-eb2003d6de73 · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction Basen: Time-domain brain-assisted speech enhancement network with convolutional cross at- tention in multi-talker conditions.Interspeech 2023,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4c96733a-00b6-4f10-8b93-6354c97f32d8 · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction Based on audio-video evoked auditory attention detection electroencephalogram dataset.Journal of Tsinghua University (Science and Tech- nology), 64(11):1919–1926,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 54d544ed-48a9-4f5b-b627-dfc15aff88a2 · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction Neural target speech extraction: An overview.IEEE Signal Processing Magazine, 40(3):8–29, 2023
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cb6d813a-1da5-4dd1-82c2-343b6f83e2c9 · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction GroupMamba: Efficient Group-Based Visual State Space Model
Reference 2001
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4fc0c39-1540-4c36-83cd-f0c5369100b8 · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction Centroid estimation with transformer-based speaker embedder for robust target speaker extraction
Reference 2005
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cc40db08-4309-4add-a24d-7547d36f37ff · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction Speech intelligibility predicted from neural entrainment of the speech envelope.Journal of the Association for Re- search in Otolaryngology, 19:181–191,
Reference 2011
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 187771db-3a3e-4b65-a4cd-ec87fac5dad5 · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction Conv-tasnet: Surpassing ideal time–frequency magnitude masking for speech separation.IEEE/ACM transactions on audio, speech, and language processing, 27(8):1256– 1266,
Reference 2015
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 493b7a6c-6905-47f2-88dd-dccd89182d9c · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction Brain-informed speech separation (biss) for enhancement of target speaker in multitalker speech perception.NeuroImage, 223:117282,
Reference 2018
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d0c08969-5790-4ea4-b3f3-29de1790de9c · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction Typing to Listen at the Cocktail Party: Text-Guided Target Speaker Extraction
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0cd41983-8063-42d2-adc1-94a39ac04e1c · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction NeuroSpex: Neuro-Guided Speaker Extraction with Cross-Modal Attention
Reference 2020
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f3f7cfbf-b1ee-46ac-8306-2586aa49a79c · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction End-to-end brain-driven speech enhance- ment in multi-talker conditions.IEEE/ACM Transactions on Audio, Speech, and Language Processing, 30:1718– 1733,
Reference 2021
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 65930c91-b637-452e-9df8-32bf922a2c89 · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction Speaker-independent auditory attention decoding with- out access to clean speech sources.Science advances, 5(5):eaav6134,
Reference 2022
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f78d3635-634d-4020-9633-6f9d2e4ad98c · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction X-tf-gridnet: A time–frequency domain target speaker extraction network with adaptive speaker embed- ding fusion.Information Fusion, 112:102550,
Reference 2023
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e8195b1e-72b8-4d7c-aa33-933a0806b837 · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction Msfnet: Multi-scale fusion net- work for brain-controlled speaker extraction
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b3fb19f1-d6c2-4a3c-ad58-871fcf6c22aa · outbound
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction SpEx+: A Complete Time Domain Speaker Extraction Network
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8de67a4d-dd98-4d1f-9a7e-c8cfb58efa2a · inbound
DMF2Mel: A Dynamic Multiscale Fusion Network for EEG-Driven Mel Spectrogram Reconstruction M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 36a42793-abee-433f-8157-6c307e1077df · inbound
Decoding Speech Envelopes from Electroencephalogram with a Contrastive Pearson Correlation Coefficient Loss M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.