Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 47 inbound Pith citation observations for arXiv:2111.09296.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:43:08.358475Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation df060ada-c158-40dd-a58e-c297ca8c74c8 · inbound
Towards Generalized Source Tracing for Codec-Based Deepfake Speech XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97d95895-7355-4b31-83b6-d50e6d76d3c4 · inbound
Joint ASR and Speaker Role Tagging with Serialized Output Training XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e892ec6f-7acc-49c8-a6fa-fde74123eb15 · inbound
From Sharpness to Better Generalization for Speech Deepfake Detection XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff93de16-16f8-419b-9e1a-54ae05522f3d · inbound
Pushing the Performance of Synthetic Speech Detection with Kolmogorov-Arnold Networks and Self-Supervised Learning Models XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ba1af66-387e-4ddc-bbca-aa4840297d0d · inbound
Breaking the Transcription Bottleneck: Fine-tuning ASR Models for Extremely Low-Resource Fieldwork Languages XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bacce3da-fc5d-4cf5-bbbf-bfeef09b4d93 · inbound
Data Quality Issues in Multilingual Speech Datasets: The Need for Sociolinguistic Awareness and Proactive Language Planning XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 045ac136-f4e3-4f8b-8dce-fd38e6707afd · inbound
Thinking Beyond Tokens: From Brain-Inspired Intelligence to Cognitive Foundations for Artificial General Intelligence and its Societal Impact XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 159
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dad648f9-9aa1-4934-b17b-0668ee6d121e · inbound
DiceHuBERT: Distilling HuBERT with a Self-Supervised Learning Objective XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 845ae894-6d3c-49bf-8e4e-53e6d3f9e739 · inbound
Word stress in self-supervised speech models: A cross-linguistic comparison XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6910d54b-696a-4824-9f91-cfc860254baf · inbound
RepeaTTS: Towards Feature Discovery through Repeated Fine-Tuning XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1991a958-5633-48f9-ac24-634cc4e0fab1 · inbound
On Barriers to Archival Audio Processing XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1c21fd1-9842-447b-8a1c-14db27d29faa · inbound
SpeechFake: A Large-Scale Multilingual Speech Deepfake Dataset Incorporating Cutting-Edge Generation Methods XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ddfdcf68-6bbb-418c-92ca-b7e5ef60aa85 · inbound
CAM\~OES: A Comprehensive Automatic Speech Recognition Benchmark for European Portuguese XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1e5f998-c1c8-4785-91e5-dd8b66b957c6 · inbound
Generalizable Audio Spoofing Detection using Non-Semantic Representations XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b423e264-8805-408e-8884-c868e8f5ad10 · inbound
Forensic Similarity for Speech Deepfakes XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7ea9d4ba-0fda-47c3-ada1-5f2e3c64aae3 · inbound
SONAR: Spectral-Contrastive Audio Residuals for Generalizable Deepfake Detection XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98b8ee19-d213-48f5-9aa8-f181363ba2b3 · inbound
A SUPERB-Style Benchmark of Self-Supervised Speech Models for Audio Deepfake Detection XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation be6fd803-7145-4683-a19c-101dd423dd21 · inbound
Benchmarking Multilingual Speech Models on Pashto: Zero-Shot ASR, Script Failure, and Cross-Domain Evaluation XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation db07feee-6407-4425-8621-09c452858f2c · inbound
Giving Voice to the Constitution: Low-Resource Text-to-Speech for Quechua and Spanish Using a Bilingual Legal Corpus XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dbb3df5e-639c-4e03-a9f6-4c0a3e50176d · inbound
Similarity Choice and Negative Scaling in Supervised Contrastive Learning for Deepfake Audio Detection XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e98638b2-ad8c-42d9-8f32-79e8ac4438e6 · inbound
Profiling the Voice: Speaker-Specific Phoneme Fingerprinting for Speech Deepfake Detection XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cce6bf73-5148-4a1d-b5a5-29cb6cbcb3d4 · inbound
MixFake: Benchmarking and Enhancing Audio Deepfake Detection in Diverse Real-world Mixed Audio XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ddaf9454-c96e-42c0-b5fe-c0ab99677897 · inbound
Escaping the Linearity Trap: Manifold Detours for Black-Box Adversarial Attacks on Singing Audio Deepfake Detection XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 92c547c6-ca7c-4952-967d-fb9b26bfd232 · inbound
Dual-Branch Gated Fusion for Open-Set Audio Deepfake Source Tracing XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f4260727-f010-4480-8756-d5f2a3dee562 · inbound
Pretrained self-supervised speech models can recognize unseen consonants XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cbe71276-4f07-4fdb-8176-4a33c86cc761 · inbound
SpAArSIST: Sparsified AASIST for Efficient and Reliable Anti-Spoofing XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 18ed6e2d-fcd1-4e0b-9178-354240c9df74 · inbound
Responsible ASR: Overcoming Challenges of Foundational Models in Narrow-Band and Low-Resource Settings XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dd75c079-7ec4-4819-9099-12cd136d5123 · inbound
Light-weight Pronunciation Assessment via Discrete Speech Token Surprisal XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dedb92c1-9f15-4ce5-8432-63b950b9a2fe · inbound
Beyond Speaker Independence: Evaluating Cross-Lingual Acoustic-to-Articulatory Inversion Across Finnish and Russian XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e82d9116-99fc-4ed9-b1a5-7b70cbc47ce5 · inbound
Impact Analysis of Speech Representation Learning Models for Acoustic Side-Channel Attack XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fbff2c09-d503-4660-b16d-78a7dabb35c4 · inbound
Impact Analysis of Speech Representation Learning Models for Acoustic Side-Channel Attack XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 61adb2a8-dc23-40af-a1f2-68848c557b19 · inbound
From Speech to Text Corpora: Evaluating ASR-Based Data Acquisition for Low-Resource Fongbe and Hausa XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5d125a69-5efe-41d1-b65a-f8de9b2f395c · inbound
What Do Deepfake Benchmarks Measure? An Audit Using Frozen Self-Supervised Representations XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2a86893c-fa6b-413d-9c74-b2607ed4d8f7 · inbound
Syntactic Belief Update as the Driver of Garden Path Processing Difficulty XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 295
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5e46740a-119f-4c26-ac54-9e04ccec34c6 · inbound
Detecting Audio Deepfakes on the Edge:Lightweight SSL-Based Detection in a Browser Plugin XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation abbf0d63-cf32-495b-a9cb-570350d25ac4 · inbound
From Objectives to Applications: Aligning Architectural Biases in Audio Self-Supervised Learning XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 93
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5a231095-8343-43cd-9b28-7b84916812e4 · inbound
Speaker-Aware Temporal Aggregation Strategies on Segment Representations for Depression Detection in Dyadic Interaction: A Benchmark Study XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59a919b8-fb23-4828-adb2-3d3995836e7c · inbound
An Intervention-Based Framework for Shortcut Diagnosis in Spoofing Countermeasures XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5265bd2a-3338-4e30-b7ee-b6840c656f53 · inbound
Towards Digital Preservation of Efik: TTS for a Low-Resource African Language XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83eed9c1-9c22-4fb8-b782-3ff5b7d21eb3 · inbound
Evaluating the Effect of Linguistic Relatedness on Cross-Lingual Transfer in Large Multilingual Automatic Speech Recognition XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 087998ef-6c92-4171-8258-2dacc3719058 · inbound
Evaluating the Effect of Linguistic Relatedness on Cross-Lingual Transfer in Large Multilingual Automatic Speech Recognition XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d8c6bc4-790e-437b-acb7-3688e45872b3 · inbound
GigaAM Multilingual: Foundation Model for Underrepresented Languages XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b197b28-1724-4dbe-9a52-d2703304a7d9 · inbound
Unified Gradient Projection: Language-Balanced Continual Learning for Multilingual Low-Resource ASR XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 397ebc1a-c572-4365-b947-30a876530d83 · inbound
DONDO: Open w2v-BERT Speech-Recognition Base Models for African Languages XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af26c9ad-3b23-4deb-be16-701e9507960f · inbound
Leveraging Gradient Reversal Loss and Multitask Learning for Datasets-Aware Audio Deepfake Detection XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46196a50-20d5-4b09-b76d-b2dd78425e91 · inbound
Teffic-Audio: Tell Fact from Fiction XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d97e4775-c32e-4e8e-a2a1-0ac095920b20 · inbound
Multi-Backbone Self-Supervised Ensembles for Audio Deepfake Detection and a Cross-Track Analysis of Generation-Detection Asymmetry XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.