Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-18T21:20:00.286888Z
Paper Citation Record · LEDGER
As of 12 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 3 inbound Pith citation observations for arXiv:2508.20474.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-18T21:20:00.286888Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-14T18:22:25.235866Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-02T13:06:59.588978Z
50 of 50 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation cc1e516e-daf8-414e-affa-3b6d5cbd1200 · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder A review of speaker diarization: Recent advances with deep learning
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 1abf1080-e2d4-4d74-ab43-8d8e1c77d440 · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder Encoder-decoder based attractors for end-to-end neural diarization
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a60e7931-d5e1-418e-8b2b-0d2664e07d68 · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder Powerset multi-class cross entropy loss for neural speaker diarization
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b39f67f9-37f3-4af8-8a4b-384bd4efc6f9 · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder Supervised speech separation based on deep learning: An overview
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f67f2e83-9d22-4a3e-b71e-2f14b0e089eb · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder Conv-TasNet: Surpassing ideal time–frequency magnitude masking for speech separation
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 0c6da9fd-a07e-4e07-82b8-c2fb332aabaa · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder TF-GRIDNET: Making time- frequency domain models great again for monaural speaker separation
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 4c02c8ad-f4f3-419d-bdd3-221c81e49f65 · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder Single-channel multi-talker speech recognition with permutation invariant training
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 45c099f3-3f00-47d2-bc4f-d66815e2e024 · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder A purely end-to-end system for multi-speaker speech recognition
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation ed1b6dfe-6b36-4689-8d4b-50422846b25a · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder End-to-end multi-speaker speech recognition with transformer
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c52fda73-b706-4475-84dc-9bfab6b708ac · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder Serialized output training for end-to-end overlapped speech recognition
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 722ac973-82dc-4934-a935-e7be7630ad27 · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder Integration of speech separation, diarization, and recog- nition for multi-speaker meetings: System description, comparison, and analysis
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 1ec9cd09-1348-4340-aceb-c9a52307fb26 · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder Continuous speech separation: Dataset and analysis
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 67da9005-dfa6-4c8b-b8b9-5b543763bf9a · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder CHiME-6 Challenge: Tackling multispeaker speech recognition for unsegmented recordings
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation bc4b6450-3be1-4b7e-8761-49fc6e2debab · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder Tandem multitask training of speaker diarisation and speech recognition for meeting transcription
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 762a0fee-3104-4434-9b49-a94fb67e76c3 · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder pyannote.audio 2.1 speaker diarization pipeline: principle, benchmark, and recipe
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation df07468f-a45d-4196-8c01-690fdcf01e29 · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder TS-SEP: Joint di- arization and separation conditioned on estimated speaker embeddings
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 50689e6e-e261-4bb3-a9b4-77eb36952239 · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder PixIT: Joint training of speaker diarization and speech separation from real-world multi-speaker recordings
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f1bb6545-f1ad-499e-93bd-24ef735f654e · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder Adapting multi-lingual asr models for handling multiple talkers
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 960184b7-bad8-465d-8ef3-12012ca61b0a · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder Speech recog- nition and multi-speaker diarization of long conversations
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c2e1733f-1936-42c1-bd33-400d120e93d1 · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder One model to rule them all ? towards end-to-end joint speaker diarization and speech recognition
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 76cb90c1-5a35-40e0-ba7b-2d067aaae32f · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder Streaming speaker-attributed ASR with token-level speaker embeddings
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 236d8abc-7075-4136-87f8-49b22738ab0d · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder MIMO-Speech: End-to-end multi-channel multi- speaker speech recognition
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f8614a95-9226-4d32-9ae6-05d74efd289c · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder Multi-talker ASR for an unknown number of sources: Joint training of source counting, separation and ASR
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation cf027948-31a7-4b95-997e-cd92bcc796b3 · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder All-neural online source separation, counting, and diarization for meeting analysis
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 1f10084c-135a-4091-a0c4-fc7efeb7c4ff · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder Neural blind source separa- tion and diarization for distant speech recognition
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 7cab4107-4143-4bef-a3cc-505c4e95f314 · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder Stcon system for the chime-8 challenge
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 0bb1b868-855a-4ca0-92f0-b9e0733d44bb · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder BUT/JHU system description for CHiME-8 NOTSOFAR-1 challenge
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 606b2f6a-6520-46ee-b716-b15d4e638f9b · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder The USTC-NERCSLIP systems for the CHiME-8 NOTSOFAR-1 challenge
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f7bdd670-a1ac-4663-a1fa-9f2266ec0e9f · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder NTT multi-speaker asr system for the DASR task of CHiME-8 challenge
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation ab504eb1-b0bf-4581-a620-6f8fad373239 · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder wav2vec 2.0: a framework for self-supervised learning of speech representations
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 343e67f3-df0d-41cb-b46d-05d4a14b0da1 · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder HuBERT: Self-supervised speech representation learning by masked prediction of hidden units
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 4858d18b-79bb-4bdc-964d-49e7c49d7d8c · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder WavLM: Large-scale self-supervised pre-training for full stack speech processing
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 20417b24-cf3d-49ab-bbbb-f36cea69aadc · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder Robust speech recognition via large-scale weak supervision
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 05ade4e2-ba05-4679-b0c2-5cca8698f5e5 · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder OWSM-CTC: An open encoder-only speech foundation model for speech recognition, translation, and language identification
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a4e2f819-fca3-42b5-b06a-7089112f9394 · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder SUPERB: Speech processing universal performance benchmark
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation e0559b20-3460-4a95-9f9c-c98b6852082c · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder OWSM v3.1: Better and faster open whisper-style speech models based on e-branchformer
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a776efa8-d7a7-4d1d-ab6e-595abfbe18a7 · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder LibriMix: An open-source dataset for generalizable speech separation
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation e8d22c95-1384-41b7-9533-7b1c6676ca72 · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder E-Branchformer: Branchformer with enhanced merging for speech recognition
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 1aa5f4f6-74fd-42a4-aafb-56df01fc1f70 · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder End-to-end training of time domain audio separation and recognition
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation bb4fc01a-c0d9-48fc-92eb-a61912afbfeb · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder The AMI meeting corpus: A pre-announcement
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 836dae3d-0366-450f-be64-6d9cd2690d65 · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder The Hitachi-JHU DIHARD III System: Competitive end-to-end neural diarization and x-vector clustering systems combined by dover-lap
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 36842d5e-d6fc-45d0-9273-c58a1a0edbf1 · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder The rich transcription 2006 spring meeting recognition evaluation
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a05a1210-cd17-4742-9c24-84c1aa0bf6e8 · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder A short- time objective intelligibility measure for time-frequency weighted noisy speech
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f2cf1b82-09e7-42b0-b476-1ec39cfd2ef8 · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder Performance measurement in blind audio source separation
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 14982b27-c6e9-44cc-b12f-ea1681c0252c · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder Streaming end-to-end multi-talker speech recognition
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 680035f5-dd40-42cb-b75e-5581cbbcf8d5 · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder End-to-end Speaker-Attributed ASR with transformer
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b2996ebe-c4bb-49ca-8f35-5bd1017e589c · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder Empowering whisper as a joint multi- talker and target-talker speech recognition system
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 8ad9e094-13de-4499-87bc-facd18ada5c5 · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder ESPnet: End-to-end speech processing toolkit
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 28581f29-c463-401d-a942-75821bf5a48a · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder The power of the weighted sum scalarization for approximating multiobjective optimization problems
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f23eaa7c-20d3-47d0-8c89-3111cd76295b · outbound
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder Joint beam search integrating CTC, attention, and trans- ducer decoders
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f5605c3b-0e95-4a0b-9f50-bdb78413ef8a · inbound
Beyond Acoustic Prefixes: Persistent Grounding in Serialized Acoustic Memory for LLM-Based Multi-Talker Speech Recognition Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8dff840-eb26-4084-ae2e-2bdd1e2ebf5e · inbound
Multi-task Learning is Not Enough: Representational Entanglement in Dual-output Second Language Speech Recognition Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 21835f3a-9d1d-4932-b954-8e1d145f9110 · inbound
Multi-task Learning is Not Enough: Representational Entanglement in Dual-output Second Language Speech Recognition Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.