Pith. sign in

Paper Citation Record · LEDGER

LibriMix: An Open-Source Dataset for Generalizable Speech Separation

As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 51 inbound Pith citation observations for arXiv:2005.11262.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2005.11262 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 51 of 51 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 51 of 51 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T00:00:56.674087Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T12:57:07.566849Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f06a28ed-bbef-477f-a65c-e419aab2f783 · inbound

SALMONN: Towards Generic Hearing Abilities for Large Language Models cites this paper.

SALMONN: Towards Generic Hearing Abilities for Large Language Models LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-18T02:29:46.414801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-18T02:29:46.242983Z digest=sha256:f38a1ca6c07f5c0ee7304eeafaf9659c2d67987eeeb612ab8638daad05806a67

Observation 5332612a-c249-475b-b45f-46194f927f85 · inbound

DASB - Discrete Audio and Speech Benchmark cites this paper.

DASB - Discrete Audio and Speech Benchmark LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-05-24T00:28:39.512180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-24T00:26:57.419537Z digest=sha256:412f45688e9dceccb5783c571fc96a52af64a48854494c12d0e6e873b69ac25a

Observation 22292930-4f3e-4ae9-bc5d-9aee646b3d76 · inbound

Developing an Effective Training Dataset to Enhance the Performance of AI-based Speaker Separation Systems cites this paper.

Developing an Effective Training Dataset to Enhance the Performance of AI-based Speaker Separation Systems LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T21:44:06.821525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:44:06.821525Z digest=sha256:f0f33dad55295052f4c800c6bd91463d60b4387ab2ec7dc199f6d935a38739d0

Observation 7619ca74-a52c-444c-9310-ace1ec8b0838 · inbound

GhostRNN: Reducing State Redundancy in RNN with Cheap Operations cites this paper.

GhostRNN: Reducing State Redundancy in RNN with Cheap Operations LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T16:47:25.997599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:47:25.997599Z digest=sha256:08dc74cb2c998aa779d568f4f04bbeb86418815801bb33a45f7b6fa3cefdc75b

Observation 1c55b051-41f8-4025-b770-15c63041707e · inbound

Disentangled-Transformer: An Explainable End-to-End Automatic Speech Recognition Model with Speech Content-Context Separation cites this paper.

Disentangled-Transformer: An Explainable End-to-End Automatic Speech Recognition Model with Speech Content-Context Separation LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T11:53:09.715229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:53:09.715229Z digest=sha256:8184e561267ce452f0242668b7e431320b4d9c73b4a7f9bf3b8a68d09d401016

Observation 06a52b33-ee20-4967-bd55-5f50280b74fd · inbound

Multiple Choice Learning for Efficient Speech Separation with Many Speakers cites this paper.

Multiple Choice Learning for Efficient Speech Separation with Many Speakers LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T11:12:18.051999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:12:18.051999Z digest=sha256:934d702b8ce2b46cb22188963bb8f9111e374231968e4e0f5403852a79f87559

Observation 076ef45a-78c9-4184-bbc0-9457f5f1c265 · inbound

SQ-Whisper: Speaker-Querying based Whisper Model for Target-Speaker ASR cites this paper.

SQ-Whisper: Speaker-Querying based Whisper Model for Target-Speaker ASR LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-11T20:38:17.114106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:38:17.114106Z digest=sha256:42236db2a32e7cf0faeeb809523e1e7b04fc2d24258d4808e3484612434e55b2

Observation da9cba9c-7934-4220-bc05-bfe2a49c4a1f · inbound

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data cites this paper.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T14:05:25.840587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:05:25.840587Z digest=sha256:0efc0439185b7747cb2107f37fb358567cf6f5813208157a10af9c8e7a3f8c7e

Observation 9b206c6c-41e4-4e62-9b8a-9e0cee88fcb7 · inbound

Scale This, Not That: Investigating Key Dataset Attributes for Efficient Speech Enhancement Scaling cites this paper.

Scale This, Not That: Investigating Key Dataset Attributes for Efficient Speech Enhancement Scaling LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T11:52:02.153325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:52:02.153325Z digest=sha256:e75359acc730a1309bd2b5dc87f39ce726c45174e2da82c7ab49f2efe5186f5a

Observation 2f270ae9-7702-4761-bae5-017a66649232 · inbound

AnCoGen: Analysis, Control and Generation of Speech with a Masked Autoencoder cites this paper.

AnCoGen: Analysis, Control and Generation of Speech with a Masked Autoencoder LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T21:18:15.115417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:18:15.115417Z digest=sha256:f969f81dd3fc8efac630983f81036eecce643e0cd194907a36aebd526c1689d9

Observation 2335e1c5-3911-4214-a303-5b7847321c5f · inbound

Beyond Speaker Identity: Text Guided Target Speech Extraction cites this paper.

Beyond Speaker Identity: Text Guided Target Speech Extraction LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T20:14:00.691235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:14:00.691235Z digest=sha256:551c4f17f1fa87a8044c6743a81cc5ccd03804825a08868aed4cc5761feb2828

Observation 8a1bfd48-365e-4c1f-a900-2edd009a2537 · inbound

30+ Years of Source Separation Research: Achievements and Future Challenges cites this paper.

30+ Years of Source Separation Research: Achievements and Future Challenges LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-10T17:51:53.013194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:51:53.013194Z digest=sha256:2eac6b36d0845ad9e7a288a6e82a8cbba7daccadbcecaf1ea9d7bb4d7d7d5e88

Observation 81187307-ea8b-4b5f-8e4b-937abbb286a9 · inbound

Metis: A Foundation Speech Generation Model with Masked Generative Pre-training cites this paper.

Metis: A Foundation Speech Generation Model with Masked Generative Pre-training LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-09T05:54:18.471270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T05:54:18.471270Z digest=sha256:7a5954fe65f1c1918f0f29ebb0528313ce4d690bf0a9692369c6f82bf1a2c06e

Observation 5722bfab-470e-4d5b-a8d8-f126a78667b0 · inbound

SepALM: Audio Language Models Are Error Correctors for Robust Speech Separation cites this paper.

SepALM: Audio Language Models Are Error Correctors for Robust Speech Separation LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-16T00:00:56.674087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:00:56.674087Z digest=sha256:707c0b0d9ccabd6487bca0e76c9ebed8958d3b20a1ef0768f99cadcc59c17a11

Observation 8ef09ace-19ae-419c-9c2c-3c6142c2b8d0 · inbound

TS-SUPERB: A Target Speech Processing Benchmark for Speech Self-Supervised Learning Models cites this paper.

TS-SUPERB: A Target Speech Processing Benchmark for Speech Self-Supervised Learning Models LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T22:40:24.353352Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:40:24.353352Z digest=sha256:444a1af4129f40f71f1b80d333d283879658f65e2389a2d70e5996b7386cb9a3

Observation 80573499-0b52-463e-9876-63d12850ee2c · inbound

Survey of End-to-End Multi-Speaker Automatic Speech Recognition for Monaural Audio cites this paper.

Survey of End-to-End Multi-Speaker Automatic Speech Recognition for Monaural Audio LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T21:03:17.999940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:03:17.999940Z digest=sha256:ac9800ea3576966fe904e8973b9961cf942fd4defb255fd6a4bebf53671f9583

Observation 28cd140d-fb11-42d0-9430-6f8f5f4c2bbc · inbound

SepPrune: Structured Pruning for Efficient Deep Speech Separation cites this paper.

SepPrune: Structured Pruning for Efficient Deep Speech Separation LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T20:46:07.504201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:46:07.504201Z digest=sha256:9d6cac6f8fc8a512587b8cee339bbf8a96fd3f48e3e36e0c3d15c3215d1155b2

Observation 6771830a-7bab-4547-bc47-dae05fc4360c · inbound

Unified Architecture and Unsupervised Speech Disentanglement for Speaker Embedding-Free Enrollment in Personalized Speech Enhancement cites this paper.

Unified Architecture and Unsupervised Speech Disentanglement for Speaker Embedding-Free Enrollment in Personalized Speech Enhancement LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-15T20:43:27.359520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:43:27.359520Z digest=sha256:77c802d0a1c9b3f6505d8305cff5608972b9eee43b286427995eeb1ad257d04b

Observation 991ebbc0-4aa7-47a7-8982-912d1144ade2 · inbound

Time-Frequency-Based Attention Cache Memory Model for Real-Time Speech Separation cites this paper.

Time-Frequency-Based Attention Cache Memory Model for Real-Time Speech Separation LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T20:23:17.354936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:23:17.354936Z digest=sha256:40fe60b31b59b9f2fdb668a7a22303e4b400fe82098f197fbe6292040be4f6cb

Observation 3b9ba10c-8545-4541-a349-57e46c218f8d · inbound

Steering Deep Non-Linear Spatially Selective Filters for Weakly Guided Extraction of Moving Speakers in Dynamic Scenarios cites this paper.

Steering Deep Non-Linear Spatially Selective Filters for Weakly Guided Extraction of Moving Speakers in Dynamic Scenarios LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T15:38:39.245802Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:38:39.245802Z digest=sha256:e528d5dbf954fc979e34ec1800021b1b11d8f157b26d04d4d62008ee13055b0f

Observation ff992483-89bd-4c28-8710-1079063a1ac4 · inbound

SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline cites this paper.

SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T14:22:07.278378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:22:07.278378Z digest=sha256:383eb5a294786b2c778edd3699fccf84c4cc5c481e0b467ab799704e8dd0a68e

Observation 62add7c9-64bb-4ba3-98ad-d7324cb2cdc9 · inbound

An Investigation on Speaker Augmentation for End-to-End Speaker Extraction cites this paper.

An Investigation on Speaker Augmentation for End-to-End Speaker Extraction LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T13:28:29.445402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:28:29.445402Z digest=sha256:855aaa3b0a2bbd91efcb91915b7061a7e407d97d249761f7e8a05aae91c55634

Observation fdbc0bb6-2d51-466e-966c-b009ac9d16e8 · inbound

AISHELL-5: The First Open-Source In-Car Multi-Channel Multi-Speaker Speech Dataset for Automatic Speech Diarization and Recognition cites this paper.

AISHELL-5: The First Open-Source In-Car Multi-Channel Multi-Speaker Speech Dataset for Automatic Speech Diarization and Recognition LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T13:02:17.367614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:02:17.367614Z digest=sha256:7a2c826200d27197a6f5cbafcafe9bc9c6ec734d26da308f1014a23625fdf6bc

Observation c9f657fd-e3a5-423d-b32a-b959443a57b8 · inbound

CMT-LLM: Contextual Multi-Talker ASR Utilizing Large Language Models cites this paper.

CMT-LLM: Contextual Multi-Talker ASR Utilizing Large Language Models LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T12:09:31.011992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:09:31.011992Z digest=sha256:dbc186f7f98a8e1fac23ee7c46563e5e74ad11242976471aafbc029ea255504f

Observation 0123293d-e901-4c82-8787-7cce00d2debe · inbound

SC-SOT: Conditioning the Decoder on Diarized Speaker Information for End-to-End Overlapped Speech Recognition cites this paper.

SC-SOT: Conditioning the Decoder on Diarized Speaker Information for End-to-End Overlapped Speech Recognition LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T00:50:17.251656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:50:17.251656Z digest=sha256:5ab4603ae0a389357c6fb84a34d279a8fb59feef6caf285efdac25464478e903

Observation e503fbd7-3814-4625-8716-bed84fe69c16 · inbound

SpeechRefiner: Towards Perceptual Quality Refinement for Front-End Algorithms cites this paper.

SpeechRefiner: Towards Perceptual Quality Refinement for Front-End Algorithms LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T00:30:35.412898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:30:35.412898Z digest=sha256:3037a47e9409114ed0753c93bbf7892e8b0e53642ef4d2778bbc9c9fa61c112e

Observation 077f8369-e5bf-44e9-827e-6f8abc59379a · inbound

Self-Steering Deep Non-Linear Spatially Selective Filters for Efficient Extraction of Moving Speakers under Weak Guidance cites this paper.

Self-Steering Deep Non-Linear Spatially Selective Filters for Efficient Extraction of Moving Speakers under Weak Guidance LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-06T20:28:04.990163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:28:04.990163Z digest=sha256:4d01553690e6aa55679ad60130900014c65348104f80489c428b67ac2a369881

Observation 203a6940-1875-4995-8852-f38d58673943 · inbound

Autoregressive Speech Enhancement via Acoustic Tokens cites this paper.

Autoregressive Speech Enhancement via Acoustic Tokens LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T16:41:24.033479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:41:24.033479Z digest=sha256:85486dd00b33074e2031082d702caf197148153d76c549678dccacc32cb6af14

Observation 9db2a1b1-70dc-40ce-a056-72781458b97b · inbound

CodecBench: A Comprehensive Benchmark for Acoustic and Semantic Evaluation cites this paper.

CodecBench: A Comprehensive Benchmark for Acoustic and Semantic Evaluation LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-05T15:00:13.980353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:00:13.980353Z digest=sha256:8263c68a841bda7b8af6f65f193a7ead086da2dfefbacc850fbc296b87f12296

Observation adcecf9c-1444-4687-98b4-4c2002d9ab09 · inbound

Serialized Output Prompting for Large Language Model-based Multi-Talker Speech Recognition cites this paper.

Serialized Output Prompting for Large Language Model-based Multi-Talker Speech Recognition LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T12:59:06.244600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:59:06.244600Z digest=sha256:832bb8502a4941ec2fea805cb22399a5621369289e175d683bd776992fb78da9

Observation 1ff0967b-a56b-49f1-bc60-e249b3c5b830 · inbound

GenTSE: Enhancing Target Speaker Extraction via a Coarse-to-Fine Generative Language Model cites this paper.

GenTSE: Enhancing Target Speaker Extraction via a Coarse-to-Fine Generative Language Model LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-03T14:20:49.173085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T14:20:49.173085Z digest=sha256:5536f3b58ffab6051311fb2a064cf6550cec0e23db38642c6f2546e32d177165

Observation ae1d01cb-8ea0-452f-b3a7-5ec4ae79accc · inbound

Discriminative-Generative Target Speaker Extraction with Decoder-Only Language Models cites this paper.

Discriminative-Generative Target Speaker Extraction with Decoder-Only Language Models LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 80

Resolution
verified exact
arxiv_id, observed 2026-05-21T15:44:14.744479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-21T15:42:57.766341Z digest=sha256:47279a0d6ed049d126d246fe32bc1edb98c7e64c8b26e29cd52b087d3bb30c93

Observation 1842cc52-26a7-4b0a-a006-3a66ab9697b2 · inbound

Detect, Attend and Extract: Keyword Guided Target Speaker Extraction cites this paper.

Detect, Attend and Extract: Keyword Guided Target Speaker Extraction LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-03T03:30:39.487256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:30:39.487256Z digest=sha256:b0ef7b8c6a805c3bf91a5b1ee6a3b92c8f80d1358b3e38db491c9a70a72e82aa

Observation d6923ac0-1294-4bbd-abdd-8ff6941e1c84 · inbound

Enroll-on-Wakeup: A First Comparative Study of Target Speech Extraction for Seamless Interaction in Real Noisy Human-Machine Dialogue Scenarios cites this paper.

Enroll-on-Wakeup: A First Comparative Study of Target Speech Extraction for Seamless Interaction in Real Noisy Human-Machine Dialogue Scenarios LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T22:51:20.971040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:51:20.971040Z digest=sha256:fa2ebb5165eb9a9cc3755e9e5d33b9ffcbfc3397f1296d4e927b399821d97392

Observation 3b04f370-561f-4dec-bdea-985215ab7392 · inbound

Autoregressive Guidance of Deep Spatially Selective Filters using Bayesian Tracking for Efficient Extraction of Moving Speakers cites this paper.

Autoregressive Guidance of Deep Spatially Selective Filters using Bayesian Tracking for Efficient Extraction of Moving Speakers LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-02T17:38:16.051420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T17:38:16.051420Z digest=sha256:bb25c84d9c5518523daa98a685d54c85504088f680377ab292fd564c0557b5fd

Observation 0c96491e-fe24-43bd-b84e-8ec4658d6f12 · inbound

Beyond Acoustic Prefixes: Persistent Grounding in Serialized Acoustic Memory for LLM-Based Multi-Talker Speech Recognition cites this paper.

Beyond Acoustic Prefixes: Persistent Grounding in Serialized Acoustic Memory for LLM-Based Multi-Talker Speech Recognition LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-13T17:09:15.281290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T17:09:15.281290Z digest=sha256:c43472f5670b118ec199910c59a1a6599986b8c685b92f898d9c9c4b6eb2f1f6

Observation 3981a470-337e-43cd-8a41-68deb62186a7 · inbound

Ring Mixing with Auxiliary Signal-to-Consistency-Error Ratio Loss for Unsupervised Denoising in Speech Separation cites this paper.

Ring Mixing with Auxiliary Signal-to-Consistency-Error Ratio Loss for Unsupervised Denoising in Speech Separation LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:20:58.492645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T17:13:12.239842Z digest=sha256:2f4ecca0338b2180777f2429d01fdae1cf701e2dff003a560afe1266fbd35734

Observation 721891be-e055-430e-bd64-185e1f06766a · inbound

BUT System Description for CHiME-9 MCoRec Challenge cites this paper.

BUT System Description for CHiME-9 MCoRec Challenge LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:36:27.164475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-07T10:13:38.938473Z digest=sha256:9f394a513b3aa2cb913d9b150bd40e0dd79170a6e34e1417ab4dec9290d8cf26

Observation 0aa1a2e2-b0ae-4882-a625-91c5e647e3b5 · inbound

Exploring Token-Space Manipulation in Latent Audio Tokenizers cites this paper.

Exploring Token-Space Manipulation in Latent Audio Tokenizers LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:42:08.303081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-13T02:40:05.812426Z digest=sha256:c555c1ed1add558c9d51268e135ec83d9f3e13e681224f020f56cc0b87217264

Observation 9f55dbf0-6b5c-417d-ada3-f299b4fc2c05 · inbound

Mind the Gap: Impact of Synthetic Conversational Data on Multi-Talker ASR and Speaker Diarization cites this paper.

Mind the Gap: Impact of Synthetic Conversational Data on Multi-Talker ASR and Speaker Diarization LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-19T14:42:37.472254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-19T14:39:27.382652Z digest=sha256:0221d357994dc983b3c5c901004c2092bde2005dcce9a7ada81c6971743893e6

Observation 2da88f1b-d49b-40b8-b43a-b83f9f52e565 · inbound

Can Large Audio Language Models Ignore Multilingual Distractors? An Evaluation of Their Selective Auditory Attention Capabilities cites this paper.

Can Large Audio Language Models Ignore Multilingual Distractors? An Evaluation of Their Selective Auditory Attention Capabilities LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T23:17:57.489331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-19T23:17:08.124240Z digest=sha256:7d4b0d3508509871114993c687db9aafe4851ded296e3e4b81e96fb730d36200

Observation 2d52bf4d-6972-4fdd-bfa8-7298ca335a7e · inbound

Uncertainty-based Debiasing and Unlearning for Decontamination cites this paper.

Uncertainty-based Debiasing and Unlearning for Decontamination LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 66

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T12:49:52.401615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-26T05:57:44.576411Z digest=sha256:4433def3c1fca63d08451261aa80d9c5367ae771af6088cb91bcd750a8f74891

Observation 65e970a6-19c2-4c65-a180-31bf8fa1ff0e · inbound

Don't Listen to Me: A Lightweight, Low-Latency Model for Own-Voice Cancellation in Far-Field Speech Enhancement cites this paper.

Don't Listen to Me: A Lightweight, Low-Latency Model for Own-Voice Cancellation in Far-Field Speech Enhancement LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-07-04T12:29:51.684420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-26T06:49:46.425849Z digest=sha256:87d26ab7ca573e5caa77df1e30f26ef63959278b4d765395276e66f05adfa9a2

Observation d84843c3-9065-4235-b6a2-c2ece42db3f2 · inbound

Child-Centric Voice Anonymization in Single and Multi-Speaker Speech via Domain-Adapted SSL Models cites this paper.

Child-Centric Voice Anonymization in Single and Multi-Speaker Speech via Domain-Adapted SSL Models LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-06-30T14:24:45.590969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-30T05:26:46.541637Z digest=sha256:b8aee14f85febd4a3b0865078f78307576483cfed021c2e4267db95729f6e583

Observation fb7266b1-5393-4087-bbb9-ed013a370ef5 · inbound

Flow Matching-Based Speech Source Separation with Best-of-N Biometric Sampling cites this paper.

Flow Matching-Based Speech Source Separation with Best-of-N Biometric Sampling LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-07-08T17:15:09.327729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-08T17:05:57.548636Z digest=sha256:1961f72529562de8d57c167c9f485feb78dbb45743f924bb56e5c8990abec672

Observation d613abed-a7c3-418e-9e39-ea4a1dba0775 · inbound

PS4: Proxy-Supervised Joint Training for Real Target Speaker Extraction cites this paper.

PS4: Proxy-Supervised Joint Training for Real Target Speaker Extraction LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-07-10T12:57:07.568338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-10T12:54:17.490105Z digest=sha256:88a10a092fe899551921d0ca341749e554a6ec601eac1f5041961d95b4107019

Observation 45de677f-dd71-40c4-8ab4-6bc719a5a567 · inbound

Technical Report for MERL's Real-TSE Challenge Submission cites this paper.

Technical Report for MERL's Real-TSE Challenge Submission LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-13T00:46:07.473150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T00:46:07.473150Z digest=sha256:8b54b2705cfc13b266302d0da6d3b844cc4336e2983dcf7f45946d9e96253d97

Observation 069e1799-3731-4254-ac16-51700c65bbc3 · inbound

SLT 2026 REAL-TSE Challenge: Real-world Target Speaker Extraction from Conversational Recordings cites this paper.

SLT 2026 REAL-TSE Challenge: Real-world Target Speaker Extraction from Conversational Recordings LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T23:55:29.954391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:55:29.954391Z digest=sha256:8c299aa97a5818694b6889e256e544fa788afc65e5b5fd2b2ea69d0c9a337f8b

Observation 432e0dbd-1837-496a-b54b-a55e0dda9d31 · inbound

Listen, Do Not Copy: Internalizing Audio-Grounded Scaffold Context for Robust Omni-Model Speech Understanding cites this paper.

Listen, Do Not Copy: Internalizing Audio-Grounded Scaffold Context for Robust Omni-Model Speech Understanding LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-01T06:19:31.674770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:19:31.674770Z digest=sha256:f733473e4e19499d05bc2c5b054b566df34aaf2ecf45f4751a86721b40bf693f

Observation cb5b584b-f626-4d57-9d08-740ad6673e3d · inbound

Cocktail-Talker: Multi-Speaker Dialog Modeling in Noisy Social Environments with Turn Action GRPO cites this paper.

Cocktail-Talker: Multi-Speaker Dialog Modeling in Noisy Social Environments with Turn Action GRPO LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-01T01:56:13.878197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:56:13.878197Z digest=sha256:268051b05de89ecc4cd1fc3956b318aa720eec5b736f611ea47bd8844289a5dd

Observation d71e7252-bb08-46c4-9f06-f0a42d6c960f · inbound

Hear, Invoke, and Understand: A Skill-Calling Multimodal Agent for Large Audio Language Models cites this paper.

Hear, Invoke, and Understand: A Skill-Calling Multimodal Agent for Large Audio Language Models LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-04T19:02:40.730450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:02:40.730450Z digest=sha256:c1552d8b11df08a4276bd05b8ed6359477b694bc3bcb253cbfc55b6391be82d6