Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-14T13:38:44.464676Z
Paper Citation Record · LEDGER
As of 16 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 1 inbound Pith citation observation for arXiv:1908.04737.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-14T13:38:44.464676Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-14T13:38:43.983210Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-14T13:38:44.647594Z
45 of 45 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 651a7a66-fbd6-481f-9e72-b12532fa685c · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning Overlapped speech – well known in a more general context as the cocktail party problem – remains, however, to be a largely unsolved problem
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 1679087c-a2fe-45f6-899e-da05fe52473a · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation eb5e50e7-9421-40c9-ae83-1019dba0aa1b · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning Datasets We evaluate our models on the widely used mixed speech datasets wsj0-2mix and wsj0-3mix [9, 10]
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation c3b48b80-dcf2-4563-bc6d-ed4d7d4e1351 · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning Speaker embeddings inclusion strategies The first set of experiments aims to determine the best strat- egy for inclusion of speaker embeddings in the model’s in- put
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation ce66aa74-d728-418d-80f4-049bc957192a · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 3fa5607e-9c79-466f-acb3-474e4d332e38 · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning Single-channel speech sepa- ration using sparse non-negative matrix factorization,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 1d7798b8-b707-41d0-b106-b4bb392411ff · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning Similarly to speech recognition, speech separation methods have also made major progress with the help of deep learning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 80b02dcd-5bfe-4388-800c-ba63f749f3ff · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning Learning spectral clustering, with application to speech separation,
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90517952-21c9-4dda-89e1-2432d8ff8d9b · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning Deep Neural Networks for Acoustic Modeling in Speech Recognition,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 50045803-b5ae-4dfc-b450-03dd4ea5bcd4 · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning Context-Dependent Pre-trained Deep Neural Networks for Large V ocabulary Speech Recognition,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 32128274-1284-4060-bd96-ba02c6d14182 · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning The Microsoft 2016 Conversational Speech Recognition System,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 14693f19-1e94-4cbf-ac15-eb8d6047ba35 · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning We evaluate our proposed framework on overlapped speech datasets with two and three overlapped speakers, within and across set- tings
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 2cfcf485-9698-4ef0-bb1a-83ddd39297de · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning Purely Sequence-Trained Neural Networks for ASR Based on Lattice-Free MMI,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 02cf7ebe-f578-4965-bddf-7b003df9831d · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning Wang and G
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 343adcf4-5d1d-4057-bde8-52943c1d8180 · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning Super-human multi-talker speech recognition: A graphical mod- eling approach,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 76f58950-e03b-4a51-93e3-d9256fa023d7 · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning Listen, attend and spell: A neural network for large vocabulary conversational speech recognition,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9291abe4-8011-4a7f-90f0-7bb8606d21b2 · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning Deep clus- tering: Discriminative embeddings for segmentation and separa- tion,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 48258b02-6caa-4026-8f76-1cd99ad347d6 · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning Single-Channel Multi-Speaker Separation Using Deep Cluster- ing,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 3e709503-00f4-4fb0-9cae-dad1cbb09baf · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning Alternative Objective Functions for Deep Clustering,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 45f9cf6b-e8a3-4aa0-9f84-112abc83ea01 · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning VoiceFilter: Targeted Voice Separation by Speaker-Conditioned Spectrogram Masking
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ddd76945-b2a4-4759-84cb-d106895b4453 · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning Deep Speech: Scaling up end-to-end speech recognition
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5289f2c-c191-4c08-90fd-3884db4a0288 · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning EESEN: End-to-end speech recognition using deep RNN models and WFST-based decoding,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation f2eb5c52-38b1-4e8f-a77d-244942965096 · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning End-to-end attention-based large vocabulary speech recog- nition,
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d55b2c1-1420-40b0-b4cd-1aec4d8a367f · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning Transfer learning from speaker verification to multispeaker text-to-speech synthesis,
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f08052f1-9b32-4763-90af-b65645c2e978 · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning Hybrid CTC/attention architecture for end-to-end speech recog- nition,
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aff3f211-a2ca-4d01-b38e-4a522225d479 · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning End-to-end multi-speaker speech recognition,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 51ee53ba-e313-4c8a-8171-0bd8a61dfbbf · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning A Purely End-to-End System for Multi-speaker Speech Recog- nition,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 32da70e1-b1b3-47f0-ba1d-4a034ef7e7c4 · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning Progressive joint mod- eling in unsupervised single-channel overlapped speech recogni- tion,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 85ad0927-995c-4d25-bc29-cdba43df20e1 · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning Speaker-Aware Neural Network Based Beam- former for Speaker Extraction in Speech Mixtures,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 4a422157-4728-4c9a-b4cc-3d7ecd6c267c · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning Deep Extractor Network for Target Speaker Recovery from Sin- gle Channel Speech Mixtures,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation ef0aa0f8-d6d3-458d-b841-7b7f035c8ad7 · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning Speaker diarization with LSTM,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 17d809fd-63a3-4a3d-81a1-231379a0e2e3 · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning VoxCeleb: A Large- Scale Speaker Identification Dataset,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation d59d8051-2bae-443b-bdba-b7a353cdfe5a · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning Multilingual acoustic models using dis- tributed deep neural networks,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 667993ef-e747-441b-a793-78a5a8f6be98 · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning The input features of x-vector extractor are 30-dimensional MFCCs without cepstral truncation with a frame length of 25 ms and shift of 10 ms
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 43a8af3a-23de-49a6-972e-5ef8595165f3 · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning Investigation of transfer learning for ASR using LF-MMI trained neural networks,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 89d25c21-6373-4c68-bf68-cc5bee4a5c46 · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning Lib- rispeech: an ASR corpus based on public domain audio books,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 589ec6d4-28d2-47a6-a474-f06a72dd5d66 · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning ESP- net: End-to-End Speech Processing Toolkit,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation dda78843-61c7-4ccc-912f-d631e6d2a375 · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning The Kaldi speech recognition toolkit,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 866d7ca0-0a36-4bb8-b333-808f1f91cec1 · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning ADADELTA: An Adaptive Learning Rate Method
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7eebc545-9de0-4bcf-ad0f-c7f33c2f7108 · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning End-to-end Speech Recognition with Word-based RNN Language Models
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 55fecded-c0c2-407f-a1b9-08a4ebf75a24 · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning VoxCeleb2: Deep Speaker Recognition,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 17b0b9ba-0761-475d-a354-e0d94cd488ab · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning X-vectors: Robust DNN embeddings for speaker recogni- tion,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 29c5e611-c86a-4eae-be53-ae1808fa4eb0 · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning The Speak- ers in the Wild (SITW) Speaker Recognition Database,
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 78a550f5-856e-40f6-9652-85e183c6b9e8 · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning Single-channel multi-talker speech recognition with permutation invariant training,
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 76f8073d-ad2f-45ae-aac8-28317a70026e · outbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning End-to-End Monaural Multi-speaker ASR System without Pretraining
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 1679087c-a2fe-45f6-899e-da05fe52473a · inbound
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.