Pith. sign in

Paper Citation Record · LEDGER

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data

As of 21 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 2 inbound Pith citation observations for arXiv:2412.12512.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.12512 v1

Coverage vector

measured 55 of 55 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T14:05:25.940672Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T16:58:02.481574Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T14:24:45.592114Z

Reference resolution

55 of 55 outbound references displayed

  • verified exact1
  • verified fuzzy41
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 19b250f4-bbc9-49f6-8971-e1667a6e15f7 · outbound

This paper cites Neural target speech extraction: An overview,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data Neural target speech extraction: An overview,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.751756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.702675Z digest=sha256:1997d9808b97a88f342fdc11a6f1db4fa14491bca2979c0502a8e88100bd9f62

Observation be20176c-dc09-4550-a0b7-0238337c6f35 · outbound

This paper cites Speech separation with pretrained frontend to minimize domain mismatch,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data Speech separation with pretrained frontend to minimize domain mismatch,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.738211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.707581Z digest=sha256:dd3b9cc5976b702e6ebe1c78d0d7c450fa28af655cd957bc370717f86798f5c2

Observation 3961b256-2393-4ed4-9444-dc071ccb6b22 · outbound

This paper cites Spex: Multi-scale time domain speaker extraction network,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data Spex: Multi-scale time domain speaker extraction network,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.725395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.711986Z digest=sha256:2091bd934e50b8b2fe2b2eb957ad920ed4a037db90fe773120d0c8820e315f22

Observation f491bc91-3d51-4ea3-bbaa-75fec09f2183 · outbound

This paper cites Spex+: A complete time domain speaker extraction network,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data Spex+: A complete time domain speaker extraction network,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.712160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.716431Z digest=sha256:9d67d5844135c81c544c8a98ff30e0fe933c30727367892f8f13dd0e29cd87e6

Observation 09d2fa82-48e0-443b-934c-266b32585615 · outbound

This paper cites Target speaker verification with se- lective auditory attention for single and multi-talker speech,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data Target speaker verification with se- lective auditory attention for single and multi-talker speech,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.698216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.720684Z digest=sha256:e5b6ccaab3b75cf02eb6fff289b8a70ab2055e7453c5ed4b29cc9c7abfe185b9

Observation be92ded4-772c-4814-8355-db5d3b23a9e6 · outbound

This paper cites V oxCeleb2: Deep speaker recognition,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data V oxCeleb2: Deep speaker recognition,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.684196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.725194Z digest=sha256:8788eb87276dfdef40859d802fa66be17ff767d09dfa6f98d8ba77b549b6fa2a

Observation 5402311e-22b6-4b1c-a716-8762526b6352 · outbound

This paper cites LibriTTS: A corpus derived from librispeech for text-to- speech,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data LibriTTS: A corpus derived from librispeech for text-to- speech,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.669640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.730017Z digest=sha256:dc7ff2462f86387d6a887465731b7fb8a7d1aa9707eba29a7c126982b03a04ff

Observation efc50805-be03-43fb-91f4-e11d6ea6c955 · outbound

This paper cites LibriSpeech: an asr corpus based on public domain audio books,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data LibriSpeech: an asr corpus based on public domain audio books,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.655365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.733998Z digest=sha256:819ee251962b76bba0200ed2b5f6d299246c4db84b41dc16e33f8a55a312db79

Observation e86aba8e-439f-454c-9f28-f12d5591fe2d · outbound

This paper cites Machine Learning for Synthetic Data Generation: A Review.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data Machine Learning for Synthetic Data Generation: A Review

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T14:05:25.737993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:05:25.737993Z digest=sha256:055fb438af6717b78b664232a548b0a207d4b5a76e8235f8335484fe004a157e

Observation 8985dcd2-17ca-41a7-a31e-d7633fc9338e · outbound

This paper cites an unresolved cited work.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-11T14:05:26.640832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.742359Z digest=sha256:4eab8b714c5f80df653f682726a50ecc20546df818322d6b2cef201a1835a5bb

Observation f699387c-056d-46b3-9eca-131c26961bd9 · outbound

This paper cites What makes good synthetic training data for learning dis- parity and optical flow estimation?.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data What makes good synthetic training data for learning dis- parity and optical flow estimation?

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.626515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.746171Z digest=sha256:a0b5c1d6b53200ef9c042429dfa2c23af872b5f904fbf8dfb75a05ceb561f9d5

Observation 39b4c628-81e1-48bd-8bbb-7d4820bc7724 · outbound

This paper cites Deep learning-enabled medical computer vision,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data Deep learning-enabled medical computer vision,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.607128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.750681Z digest=sha256:1f27f2549d9fe95c82c9240ec86b5f40692f409305209f22596cc73beae865b4

Observation cf96305f-aa52-49fa-9d75-1d3cc51da5f3 · outbound

This paper cites Data augmentation for low- resource neural machine translation,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data Data augmentation for low- resource neural machine translation,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.593384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.755219Z digest=sha256:723aed6e20a179f0c9983bed49e053e9fe056e73c92a763a3b1408fb4f17952a

Observation 42e86da4-6763-4692-8cee-6ba6485242e4 · outbound

This paper cites A survey of data augmentation approaches for nlp,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data A survey of data augmentation approaches for nlp,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.579956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.759531Z digest=sha256:3139ff352794dbd02b647abcf6632975d94d6f1dd8c60b555123d9cf96cb458e

Observation 1012f1fd-2c6e-4cb2-8a2b-4440191e488d · outbound

This paper cites A Survey on Data Synthesis and Augmentation for Large Language Models.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data A Survey on Data Synthesis and Augmentation for Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T14:05:25.763909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:05:25.763909Z digest=sha256:aa99bc73cf2527fb8f5cf709e8acd5a5161cc960b5b620f52b33373ac761f205

Observation c71a945b-80a6-4d9d-9a53-c52991619646 · outbound

This paper cites SYNT++: Utilizing imperfect synthetic data to improve speech recognition,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data SYNT++: Utilizing imperfect synthetic data to improve speech recognition,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.566353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.768611Z digest=sha256:e51d3417ff2620bb059c7d1c68f006803a599dd7c644f28c6a7fb367eded0783

Observation a8b35c9a-e69a-4c40-bde6-33bf290bf0c3 · outbound

This paper cites SynthASR: Unlocking synthetic data for speech recogni- tion,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data SynthASR: Unlocking synthetic data for speech recogni- tion,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.551974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.773119Z digest=sha256:2ad4a7ed6bc1943ee4e9021d906167c98c1191885abf748006c9cd9df3fb0f8d

Observation f20d2363-edaa-44b1-9dae-93e634e7a4b3 · outbound

This paper cites Effective data augmentation methods for neural text-to-speech systems,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data Effective data augmentation methods for neural text-to-speech systems,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.536732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.777524Z digest=sha256:ef8d28f0f851e18600de3437afad25de9820f99c45f965526bdd44287c2d8af4

Observation 4f623136-4b83-4772-98ec-083d16a362ef · outbound

This paper cites TTS-by-TTS 2: Data-selective augmentation for neural speech synthesis using ranking support vector machine with variational autoencoder,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data TTS-by-TTS 2: Data-selective augmentation for neural speech synthesis using ranking support vector machine with variational autoencoder,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.520665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.781952Z digest=sha256:8a6a6e4674a3f616eedaf3527749e8e521fd3e65bff718beb8f9fc8f49e3ec2a

Observation 3e0fbc96-5f59-4500-a78d-89731dd42161 · outbound

This paper cites Speaker augmentation for low resource speech recognition,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data Speaker augmentation for low resource speech recognition,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.505299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.786348Z digest=sha256:603b49b42fd807bc54d9f777cb515fee3b027460417c130b9dc617c7c85cb46a

Observation 871c20cc-f190-43e9-b436-ab8a5947669c · outbound

This paper cites Overcoming data scarcity in speaker identification: Dataset augmentation with synthetic mfccs via character-level rnn,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data Overcoming data scarcity in speaker identification: Dataset augmentation with synthetic mfccs via character-level rnn,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.489895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.790558Z digest=sha256:10d0879062a4f479e05184a879a5c641af5c642db82d91a3e7d103945654c7e0

Observation eac30b13-f307-4f99-b3c5-6717e31c8146 · outbound

This paper cites Target speaker extraction with curriculum learning,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data Target speaker extraction with curriculum learning,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.474379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.795645Z digest=sha256:98549f449e51a75b2c26000e8b9c4ff9f19736ecf13df41078d60d02976966d6

Observation c5f08fdc-bb07-4df0-8766-8e36006b7d81 · outbound

This paper cites Curriculum learning: A survey,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data Curriculum learning: A survey,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T14:05:25.800072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:05:25.800072Z digest=sha256:54b331e3fcac744cac69de87d999f153ce77e9fd40838d4faef468b85d0eed13

Observation 1617ea84-6a51-4321-97ab-468889d7cd89 · outbound

This paper cites Improving curriculum learning for target speaker extraction with synthetic speakers.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data Improving curriculum learning for target speaker extraction with synthetic speakers

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-08-11T14:05:26.063213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.803945Z digest=sha256:30a535668b07d536089b5d3eb65486b5aadc2d0e186a647ef1644444ddf5ef62

Observation 9316f2b1-420c-4572-aab6-6dd903560241 · outbound

This paper cites SALT: Distinguish- able speaker anonymization through latent space transformation,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data SALT: Distinguish- able speaker anonymization through latent space transformation,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.450059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.808186Z digest=sha256:10f7a665ebf76e8459374d94facc18e28ca38d5c09a7b24f2428f59ff8112b38

Observation 804a8539-7d82-43dc-8fa8-3bfdbb3591e6 · outbound

This paper cites SpeakerBeam: Speaker aware neural network for target speaker extraction in speech mixtures,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data SpeakerBeam: Speaker aware neural network for target speaker extraction in speech mixtures,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.435098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.812182Z digest=sha256:2903044428238f5399015a4a726f476f7beeae977ee814b9b9e40574ba7aa622

Observation f6bf5473-da86-4929-a7b5-b5e6f429f65f · outbound

This paper cites V oiceFilter: Targeted voice separation by speaker-conditioned spectrogram masking,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data V oiceFilter: Targeted voice separation by speaker-conditioned spectrogram masking,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.420690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.816184Z digest=sha256:61980945118f731eb34911630f7e4ba34c4f5462e9c775fb961f89c700d892ab

Observation d2ea3ce5-c068-443c-980b-9d92cda77b16 · outbound

This paper cites Deep neural networks for small footprint text-dependent speaker verification,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data Deep neural networks for small footprint text-dependent speaker verification,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.405758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.820412Z digest=sha256:21f4a616d3fda1dd397cd4ac28a40de5ac5f407108dc92b503535e7df1870b0d

Observation a3b5fd32-94fd-4060-9e11-bb5a4d452b45 · outbound

This paper cites Conformer: Convolution- JOURNAL OF LATEX CLASS FILES, VOL. XX, NO. X, XXX 2022 12 augmented transformer for speech recognition,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data Conformer: Convolution- JOURNAL OF LATEX CLASS FILES, VOL. XX, NO. X, XXX 2022 12 augmented transformer for speech recognition,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.390600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.824385Z digest=sha256:f5ee689081654c37944ce5aff7ab4f1e5e6841f6b5a3e55ba2bbdfb920b0bd05

Observation 99968acd-29bc-4dae-9a59-acc260fb9d00 · outbound

This paper cites ECAPA-TDNN: Emphasized channel attention, propagation and aggregation in TDNN based speaker verification,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data ECAPA-TDNN: Emphasized channel attention, propagation and aggregation in TDNN based speaker verification,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.375718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.828421Z digest=sha256:352780db67d3878ab0a266682c754e81e1808e4b8533968e2ae1bbbd7f6f82e7

Observation 19c705e8-7ddb-458f-91f3-89305faf7973 · outbound

This paper cites Complex ratio masking for monaural speech separation,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data Complex ratio masking for monaural speech separation,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.361821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.832434Z digest=sha256:294d7555896a820f1b538b50cab85b3a5eeef7bda3b6e72c6a4bd0be7e7aa94d

Observation 288dd581-62b1-4ac1-8f38-77c3b136e5bd · outbound

This paper cites CSR-I (WSJ0) Complete,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data CSR-I (WSJ0) Complete,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.348431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.836392Z digest=sha256:e1dc449965962a57969c1f8eb203ac350182c0958490d9059bc16725fa568db5

Observation da9cba9c-7934-4220-bc05-bfe2a49c4a1f · outbound

This paper cites LibriMix: An Open-Source Dataset for Generalizable Speech Separation.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data LibriMix: An Open-Source Dataset for Generalizable Speech Separation

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T14:05:25.840587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:05:25.840587Z digest=sha256:0efc0439185b7747cb2107f37fb358567cf6f5813208157a10af9c8e7a3f8c7e

Observation 2635c539-2ee4-45c7-aa1f-f127e5be13ad · outbound

This paper cites Data augmentation for speech separation,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data Data augmentation for speech separation,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.333658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.844693Z digest=sha256:834eb1ffa05e36df3e7bd0f224076fddf5e0ce7a4841bde537eddfb07986a3e5

Observation d2e71f56-a5fe-4421-9d43-0ffb6f9d0b29 · outbound

This paper cites Employing real training data for deep noise suppression,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data Employing real training data for deep noise suppression,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.318647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.848730Z digest=sha256:a83d6ffef9f1b1afa011de9f69f5f060d888e411b2264df74a38b483344395f3

Observation 527d4c46-f7f3-4386-ac5c-57653f5451c3 · outbound

This paper cites Perceptual evaluation of speech quality (pesq)—a new method for speech quality assessment of telephone networks and codecs,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data Perceptual evaluation of speech quality (pesq)—a new method for speech quality assessment of telephone networks and codecs,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.304912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.852847Z digest=sha256:e96d5b22a17557b4c24bc69a6fe61b8c587e0d8a763599d4d5039bf4fbc3b8a6

Observation 854c88b4-ee8a-4a76-a158-6de42a5078fc · outbound

This paper cites NaturalSpeech 3: Zero-shot speech synthesis with factorized codec and diffusion models,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data NaturalSpeech 3: Zero-shot speech synthesis with factorized codec and diffusion models,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.289956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.856726Z digest=sha256:c5fd37fea2108f87dc16d64df03c7f6808c325329c73d161fbbf23b2a4cd1128

Observation 5e94b42c-f0ea-454e-979c-0521f4e787d6 · outbound

This paper cites Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T14:05:25.865952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:05:25.865952Z digest=sha256:221e815b8a8e000dc821fae5b8790b5bdf2f6e54d8d7d2437e45616b334054e2

Observation 2ebe90da-d97e-4d12-a128-be2ba4309274 · outbound

This paper cites FastSpeech 2: Fast and High-Quality End-to-End Text to Speech.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data FastSpeech 2: Fast and High-Quality End-to-End Text to Speech

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T14:05:25.870697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:05:25.870697Z digest=sha256:e6c63460b29884e25a1de303fd04d61866d2a5ae5fb845bbfcaaa4bb30d00adc

Observation cf41df21-2d85-4d03-ad71-8eb0d50cda16 · outbound

This paper cites Synvox2: Towards a privacy-friendly V oxCeleb2 dataset,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data Synvox2: Towards a privacy-friendly V oxCeleb2 dataset,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.259664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.875729Z digest=sha256:a59c55cdfe4c1d5f6d2cf344cb7114211c11daa8d0e28eb7751d1dba4e306882

Observation 52217345-bdb1-499b-955b-a8c234d5b174 · outbound

This paper cites WavLM: Large-scale self-supervised pre- training for full stack speech processing,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data WavLM: Large-scale self-supervised pre- training for full stack speech processing,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-11T14:05:25.880173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:05:25.880173Z digest=sha256:eab6213354f25bd3ab0676aefc2329df43b2faa5a36c0d28d0977151db1c0aa9

Observation 096296db-06f7-4ded-94e9-73174d971e7e · outbound

This paper cites Nearest neighbor pattern classification,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data Nearest neighbor pattern classification,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-11T14:05:25.884477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:05:25.884477Z digest=sha256:3bf44a504f0247839c451d490cdf43b5d8f0abc43523394ab219ff0d9982f1e0

Observation cdbb77f7-6cb8-436f-bb4e-a134358a7efd · outbound

This paper cites HiFi-GAN: Generative adversarial net- works for efficient and high fidelity speech synthesis,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data HiFi-GAN: Generative adversarial net- works for efficient and high fidelity speech synthesis,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.226267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.889434Z digest=sha256:bf681f9bada6cfe3603f14d1b697c77cc5b8e79dcc91a1b31cace86eca599f35

Observation e892e776-5dd9-47b3-8235-e48426ad68c9 · outbound

This paper cites Speaker anonymization using orthogonal householder neural network,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data Speaker anonymization using orthogonal householder neural network,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.211189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.894042Z digest=sha256:9680bad51689f064ad47a6d4847613bafe94779c73da31b7e646556f24e06376

Observation 4c8eb7fe-090e-4898-b47f-7fde47838429 · outbound

This paper cites Yet another algorithm for pitch tracking (Y AAPT),.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data Yet another algorithm for pitch tracking (Y AAPT),

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.196025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.898539Z digest=sha256:72c8a4ca00ca5960daf31d405425c890ee94a06abc3264f71e2449135e3c7059

Observation 7e4ed4c6-8c3e-483b-b8f3-2b0f1bad3473 · outbound

This paper cites HuBERT: self-supervised speech representation learning by masked prediction of hidden units,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data HuBERT: self-supervised speech representation learning by masked prediction of hidden units,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.181280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.902947Z digest=sha256:f739ed1a79b8162b35571147140f922fb2ea76ed3d793645403d443326ccdac3

Observation d164a870-fcd7-4d8e-9aeb-6b2f8413bd49 · outbound

This paper cites Attentive statistics pooling for deep speaker embedding,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data Attentive statistics pooling for deep speaker embedding,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-11T14:05:25.907289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:05:25.907289Z digest=sha256:bac9743613822ebc22676742c1a89864c831bafb1fd5ad894d0526c014373bfe

Observation 0bfad4d8-744c-4803-b2a0-e44cd4113727 · outbound

This paper cites Additive margin softmax for face verification,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data Additive margin softmax for face verification,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-11T14:05:25.911910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:05:25.911910Z digest=sha256:59908935c9cccae8a290b40a413ea0b16009771030c9cc00b72d38d9eb5b3ad3

Observation 8224650f-5820-4953-82e6-ab950dfccf9e · outbound

This paper cites Open-Source Conversational AI with SpeechBrain 1.0.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data Open-Source Conversational AI with SpeechBrain 1.0

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-11T14:05:25.916256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:05:25.916256Z digest=sha256:fbc9a5c7ec9d942a7e24f48d268933c5cde4d739d1fe6c24b99e9f1a9c9221e4

Observation c75973e2-3d35-4d1b-b005-75f4b6b5d6a0 · outbound

This paper cites V oxCeleb: a large- scale speaker identification dataset,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data V oxCeleb: a large- scale speaker identification dataset,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.148483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.921052Z digest=sha256:6786aaac8bda2bd4d3d0b8314df0d2da753256fb4e282494d2cfb08ed50df818

Observation fb754191-333a-4408-9b30-a31a132ed685 · outbound

This paper cites CN-Celeb: multi-genre speaker recognition,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data CN-Celeb: multi-genre speaker recognition,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.134549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.926135Z digest=sha256:aff2715c2a749e48831daecebed8bbc7d474101fd9ce1d1c5a9068156e640ae8

Observation 44c5f1b1-47d2-47bc-bed1-0978a1be1f0d · outbound

This paper cites TorchMetrics - measuring reproducibility in pytorch,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data TorchMetrics - measuring reproducibility in pytorch,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.120482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.931173Z digest=sha256:f2a9472b2f5e6e2c88c558ad846c328afb9bb2a850d1ccd1ecbbd9b64ad12326

Observation 31b2afeb-3331-40cb-992a-ee4dd3bd7a85 · outbound

This paper cites ICASSP 2021 deep noise suppression challenge,.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data ICASSP 2021 deep noise suppression challenge,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.106950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.936181Z digest=sha256:4d292e1059b0bb493f3b8b6a2db539c14ead987198a2a2fe2a73db9e9c55dbed

Observation 57f5eb54-bea9-4e99-a9ad-6897ee97b23f · outbound

This paper cites Towards a Theoretical Understanding of Synthetic Data in LLM Post-Training: A Reverse-Bottleneck Perspective.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data Towards a Theoretical Understanding of Synthetic Data in LLM Post-Training: A Reverse-Bottleneck Perspective

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-11T14:05:25.940672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:05:25.940672Z digest=sha256:f381c6cb452ba0b2e5e195c1940849af7bff7dc7c02de7f2d9fe2d28315e7cf1

Observation b0f1d672-b665-4b7c-bfdb-d58aeb88c18f · outbound

This paper cites 22 605–22 623.

Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data 22 605–22 623

Reference 235

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:05:26.273992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:05:25.861418Z digest=sha256:8f15d0c0649e70cc7a4b09afc22e0008069618ba6774ff5706dd8933fd41cd61

Pith citing papers

Observation ff96ee8b-2756-4927-82cc-6de689893135 · inbound

Interpolating Speaker Identities in Embedding Space for Data Expansion cites this paper.

Interpolating Speaker Identities in Embedding Space for Data Expansion Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T16:58:02.481574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:58:02.481574Z digest=sha256:10093d0d0a78de717cb0deac83ffaada4e72d499b45d3f2ea2d967e1c2dfca09

Observation 8503687a-c68d-471e-8c66-def7496542be · inbound

Child-Centric Voice Anonymization in Single and Multi-Speaker Speech via Domain-Adapted SSL Models cites this paper.

Child-Centric Voice Anonymization in Single and Multi-Speaker Speech via Domain-Adapted SSL Models Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-06-30T14:24:45.594078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-30T05:26:46.541637Z digest=sha256:cc074d85bc2a0ab788baeaba6aa52254ca3593a31bf1ce95d9007a204237890f