Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 48 inbound Pith citation observations for arXiv:2106.04624.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T15:52:34.585187Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-07T20:34:09.733636Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 03ff4e08-5398-4fa5-8d4f-0ee3c7af1055 · inbound
DASB - Discrete Audio and Speech Benchmark SpeechBrain: A General-Purpose Speech Toolkit
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ab84775f-5968-4bab-96fd-a80e98d4b669 · inbound
MECAT: A Multi-Experts Constructed Benchmark for Fine-Grained Audio Understanding Tasks SpeechBrain: A General-Purpose Speech Toolkit
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4b6c3108-644b-42f8-982b-fcde1c237a11 · inbound
Benchmarking Large Pretrained Multilingual Models on Qu\'ebec French Speech Recognition SpeechBrain: A General-Purpose Speech Toolkit
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa2b14c0-3adf-407a-8f9c-8fcc845c9604 · inbound
Speaker-Conditioned Phrase Break Prediction for Text-to-Speech with Phoneme-Level Pre-trained Language Model SpeechBrain: A General-Purpose Speech Toolkit
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8795520f-d940-4ae1-8b92-aa907a3fd6ca · inbound
Analysis of Speaker Verification Performance Trade-offs with Neural Audio Codec Transmission SpeechBrain: A General-Purpose Speech Toolkit
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7815c48d-833d-4958-98a9-37ff63980bb0 · inbound
DarkStream: real-time speech anonymization with low latency SpeechBrain: A General-Purpose Speech Toolkit
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ba76ccf-b79e-42c4-a5ee-9e40c665cd94 · inbound
Efficient Trie-based Biasing using K-step Prediction for Rare Word Recognition SpeechBrain: A General-Purpose Speech Toolkit
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c449d607-5843-4983-9e02-a7dfbb3cc2d1 · inbound
Improving Synthetic Data Training for Contextual Biasing Models with a Keyword-Aware Cost Function SpeechBrain: A General-Purpose Speech Toolkit
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1bdf1c66-9918-442e-9e88-34e8c11aa282 · inbound
FAC-FACodec: Controllable Zero-Shot Foreign Accent Conversion with Factorized Speech Codec SpeechBrain: A General-Purpose Speech Toolkit
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation efe285bf-bfab-4787-a307-57fb14c1594f · inbound
A Study of Data Selection Strategies for Pre-training Self-Supervised Speech Models SpeechBrain: A General-Purpose Speech Toolkit
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d4d3ea68-d443-4092-9532-321246a79cd2 · inbound
Precision-Varying Prediction (PVP): Robustifying ASR systems against adversarial attacks SpeechBrain: A General-Purpose Speech Toolkit
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 398b8ad4-9035-4187-9ac8-b4c826dd6fc3 · inbound
DeepFense: A Unified, Modular, and Extensible Framework for Robust Deepfake Audio Detection SpeechBrain: A General-Purpose Speech Toolkit
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a2f19927-ea26-4e36-a4bc-0ddffeaa8bfe · inbound
Contextual Biasing for ASR in Speech LLM with Common Word Cues and Bias Word Position Prediction SpeechBrain: A General-Purpose Speech Toolkit
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f7a336ad-8814-4316-b4a5-d50dc6e9a686 · inbound
SEDTalker: Emotion-Aware 3D Facial Animation Using Frame-Level Speech Emotion Diarization SpeechBrain: A General-Purpose Speech Toolkit
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2210251c-a48b-4864-8bbb-3f5aa96affb6 · inbound
Hierarchical Codec Diffusion for Video-to-Speech Generation SpeechBrain: A General-Purpose Speech Toolkit
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a20f70c0-7572-414a-8a3b-02936d9d4387 · inbound
Learning Posterior Predictive Distributions for Node Classification from Synthetic Graph Priors SpeechBrain: A General-Purpose Speech Toolkit
Reference 224
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 356dd397-1fc8-48cd-93e2-f80740ceef87 · inbound
Voice of India: A Large-Scale Benchmark for Real-World Speech Recognition in India SpeechBrain: A General-Purpose Speech Toolkit
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 9682e527-aec6-4734-a843-d164e549f27b · inbound
Text-To-Speech with Chain-of-Details: modeling temporal dynamics in speech generation SpeechBrain: A General-Purpose Speech Toolkit
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 19ea0026-1b6e-4764-916d-f7c5a07eac0a · inbound
Enhancing Speaker Verification with Whispered Speech via Post-Processing SpeechBrain: A General-Purpose Speech Toolkit
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ae3991df-2278-4d19-9f79-fa697650c224 · inbound
A Toolkit for Detecting Spurious Correlations in Speech Datasets SpeechBrain: A General-Purpose Speech Toolkit
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f62378f7-5d2d-48fe-8df3-4a476d1e214d · inbound
HATS: An Open data set Integrating Human Perception Applied to the Evaluation of Automatic Speech Recognition Metrics SpeechBrain: A General-Purpose Speech Toolkit
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 337f6e71-ec55-4aa5-877a-c2bace00e8b2 · inbound
A Paradigm for Interpreting Metrics and Identifying Critical Errors in Automatic Speech Recognition SpeechBrain: A General-Purpose Speech Toolkit
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 702ef76d-b5a0-4a21-9237-3a00d4fd7144 · inbound
Evaluating voice anonymisation using similarity rank disclosure SpeechBrain: A General-Purpose Speech Toolkit
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 88532c32-d578-47f6-80fb-34aa42460307 · inbound
Mechanisms of Misgeneralization in Physical Sequence Modeling SpeechBrain: A General-Purpose Speech Toolkit
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b20d2510-fb91-4fed-84da-c5a819cf4ae3 · inbound
Phonetic Modeling of Dialectal Variation in Vietnamese Speech SpeechBrain: A General-Purpose Speech Toolkit
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6b9e549b-3f59-4f58-a5b8-cf140b364a33 · inbound
PashtoTTS-Bench: automated screening for low-resource non-Latin-script text-to-speech SpeechBrain: A General-Purpose Speech Toolkit
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d4e9fddf-723c-4e89-b04a-2551a4a4a1dd · inbound
Syllabic-Structure Decoder for Automatic Speech Recognition in Vietnamese SpeechBrain: A General-Purpose Speech Toolkit
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3d596dbb-32e0-41bb-b76d-1342e3e3cee9 · inbound
Spiking and Event-driven Neuromorphic Mamba Models for Efficient Speech Recognition SpeechBrain: A General-Purpose Speech Toolkit
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d6bee9e5-98fd-412b-b6d1-96bf8548f102 · inbound
KIT's Submission to Cross-Lingual Voice Cloning in IWSLT 2026 SpeechBrain: A General-Purpose Speech Toolkit
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e01fcc82-81d9-4ba4-82fc-1eff38a70c69 · inbound
SpeechDx: A Multi-Task Benchmark for Clinical Speech AI SpeechBrain: A General-Purpose Speech Toolkit
Reference 115
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 8f53b2f8-a5e8-44c8-b12a-dd8716ec0836 · inbound
Montreal Forced Aligner and the state of speech-to-text alignment in 2026 SpeechBrain: A General-Purpose Speech Toolkit
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 141b9598-06ad-49f2-af8d-e8f0ba02889d · inbound
Beyond ROC-AUC: Operating-Point Performance Reporting for Biometric Verification SpeechBrain: A General-Purpose Speech Toolkit
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 546f19b6-79d6-4f92-92af-af6e65361b18 · inbound
LISE : Listenable Interpretable Speaker Embeddings SpeechBrain: A General-Purpose Speech Toolkit
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f657bdcd-718b-4de8-8622-20664f22c5e1 · inbound
ESPnet3: Infrastructure for Scalable Speech and Audio Research in the Foundation Model Era SpeechBrain: A General-Purpose Speech Toolkit
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 15644453-f388-4335-a515-62b9e3980227 · inbound
What Counts as an Error? Dual-Reference Benchmarking for Atypical ASR SpeechBrain: A General-Purpose Speech Toolkit
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5beeb03a-ed44-42de-afd5-42e41bcd8edc · inbound
LuxEmo: Expressive Text-to-Speech Corpus for Luxembourgish SpeechBrain: A General-Purpose Speech Toolkit
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 015a7276-ef5e-4851-a4d3-0633fc1ed2f8 · inbound
LuxEmo: Expressive Text-to-Speech Corpus for Luxembourgish SpeechBrain: A General-Purpose Speech Toolkit
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 448946c4-88cc-401e-9a9b-9e68b73e0e91 · inbound
From Monolingual to Multilingual: Evaluating Mamba for ASR in South African Languages SpeechBrain: A General-Purpose Speech Toolkit
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f1a3da29-8750-4303-98f5-833a79ab54ee · inbound
Conversational Human Audio-visual Talking Dialogue Generation SpeechBrain: A General-Purpose Speech Toolkit
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d832546d-4020-4aff-9a35-390c916bd3ea · inbound
ProPS: Prompted Profile Synthesis for Natural Language-Conditioned Speaker Embedding Distributions SpeechBrain: A General-Purpose Speech Toolkit
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 02c95dd8-0ba1-4995-b20c-5122bb44e192 · inbound
Gender Gap Analysis in News and Talk Online Radio Broadcast SpeechBrain: A General-Purpose Speech Toolkit
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24215b28-f5b4-4b50-b952-df5106a09a86 · inbound
AMECxSV: Adaptive Metadata-Driven Embedding-Fusion Calibration for X-Lingual Speaker Verification SpeechBrain: A General-Purpose Speech Toolkit
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 836c1f18-2ec3-4ce6-b6b3-f52efceae5e4 · inbound
From Read Speech to Spoken Digits: A Task-Specific Evaluation of Speech Privacy With Informed Attackers SpeechBrain: A General-Purpose Speech Toolkit
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43581c37-f2a4-4509-babf-e337a3a127d3 · inbound
Leveraging Gradient Reversal Loss and Multitask Learning for Datasets-Aware Audio Deepfake Detection SpeechBrain: A General-Purpose Speech Toolkit
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5a6f97a-d528-47aa-9b47-c52229f6fb69 · inbound
Face and Voice Cross-modal Association with Learning Convex Feature Embedding SpeechBrain: A General-Purpose Speech Toolkit
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21b7b645-460b-4af0-bb54-72671c7b5029 · inbound
Leveraging Beam Search Information for Confidence Estimation in E2E ASR SpeechBrain: A General-Purpose Speech Toolkit
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7040812-3e7f-4077-a96f-54d202f54a14 · inbound
Multi-Backbone Self-Supervised Ensembles for Audio Deepfake Detection and a Cross-Track Analysis of Generation-Detection Asymmetry SpeechBrain: A General-Purpose Speech Toolkit
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bcaad713-9f58-4df8-a603-773df50a24d9 · inbound
Speaker Verification Under Real Classroom Conditions for English Speech SpeechBrain: A General-Purpose Speech Toolkit
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.