Pith. sign in

Paper Citation Record · LEDGER

XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 47 inbound Pith citation observations for arXiv:2111.09296.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2111.09296 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 47 of 47 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 47 of 47 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:43:08.358475Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation df060ada-c158-40dd-a58e-c297ca8c74c8 · inbound

Towards Generalized Source Tracing for Codec-Based Deepfake Speech cites this paper.

Towards Generalized Source Tracing for Codec-Based Deepfake Speech XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T05:43:08.358475Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:43:08.358475Z digest=sha256:d625a3eb49df4ffcf75740f9520813d813d1e9a2d8b7aa74516c7488da761b58

Observation 97d95895-7355-4b31-83b6-d50e6d76d3c4 · inbound

Joint ASR and Speaker Role Tagging with Serialized Output Training cites this paper.

Joint ASR and Speaker Role Tagging with Serialized Output Training XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T04:32:30.227311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:32:30.227311Z digest=sha256:eeb5c8b40635e5e26e7e0ed96d74e5e8e2ec8c2f6fdce46abd65d112be047a6a

Observation e892ec6f-7acc-49c8-a6fa-fde74123eb15 · inbound

From Sharpness to Better Generalization for Speech Deepfake Detection cites this paper.

From Sharpness to Better Generalization for Speech Deepfake Detection XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T04:08:06.077547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:08:06.077547Z digest=sha256:f3ea582f453d5e77962b51b55f0b1a7b72bb38adedad21c88ee84dbfd2aa5c17

Observation ff93de16-16f8-419b-9e1a-54ae05522f3d · inbound

Pushing the Performance of Synthetic Speech Detection with Kolmogorov-Arnold Networks and Self-Supervised Learning Models cites this paper.

Pushing the Performance of Synthetic Speech Detection with Kolmogorov-Arnold Networks and Self-Supervised Learning Models XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:26.715238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:26.715238Z digest=sha256:7924a41cd42a1ab405b9a97048bc093549255ea7e4a63f2803cd282da4bbf5fc

Observation 4ba1af66-387e-4ddc-bbca-aa4840297d0d · inbound

Breaking the Transcription Bottleneck: Fine-tuning ASR Models for Extremely Low-Resource Fieldwork Languages cites this paper.

Breaking the Transcription Bottleneck: Fine-tuning ASR Models for Extremely Low-Resource Fieldwork Languages XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T23:35:28.171349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:35:28.171349Z digest=sha256:8cbd71d17590882099574c57a94707763bfaec2bede189c226c76ef6caf8198b

Observation bacce3da-fc5d-4cf5-bbbf-bfeef09b4d93 · inbound

Data Quality Issues in Multilingual Speech Datasets: The Need for Sociolinguistic Awareness and Proactive Language Planning cites this paper.

Data Quality Issues in Multilingual Speech Datasets: The Need for Sociolinguistic Awareness and Proactive Language Planning XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:49.009463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:34:49.009463Z digest=sha256:26d0f8d6f2a59ef52fed4f26067f5a3db70b62257d507bbb8a0882f0204103e4

Observation 045ac136-f4e3-4f8b-8dce-fd38e6707afd · inbound

Thinking Beyond Tokens: From Brain-Inspired Intelligence to Cognitive Foundations for Artificial General Intelligence and its Societal Impact cites this paper.

Thinking Beyond Tokens: From Brain-Inspired Intelligence to Cognitive Foundations for Artificial General Intelligence and its Societal Impact XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 159

Resolution
unresolved
no resolver link, observed 2026-08-06T21:07:11.272341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:07:11.272341Z digest=sha256:004164716b8c7e96422dc321fd3d37d074d3fbc4dd81cdf8eefd01dc2caed579

Observation dad648f9-9aa1-4934-b17b-0668ee6d121e · inbound

DiceHuBERT: Distilling HuBERT with a Self-Supervised Learning Objective cites this paper.

DiceHuBERT: Distilling HuBERT with a Self-Supervised Learning Objective XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T23:02:14.927874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:02:14.927874Z digest=sha256:94d40be70bc463c1708bcb7dcd32a41f647081926e8cb4892b8927655acea96b

Observation 845ae894-6d3c-49bf-8e4e-53e6d3f9e739 · inbound

Word stress in self-supervised speech models: A cross-linguistic comparison cites this paper.

Word stress in self-supervised speech models: A cross-linguistic comparison XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T19:44:26.885657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:44:26.885657Z digest=sha256:1bbb92cd115ea4cd6835f0ba912b4e214de6d90ee1b2ebe5dac4ee3216fdd9ec

Observation 6910d54b-696a-4824-9f91-cfc860254baf · inbound

RepeaTTS: Towards Feature Discovery through Repeated Fine-Tuning cites this paper.

RepeaTTS: Towards Feature Discovery through Repeated Fine-Tuning XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T20:01:35.492150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:01:35.492150Z digest=sha256:8d9c190b3e8c2243342394d9c79efbdc417b5e751212cf823b6e89ff1583cd88

Observation 1991a958-5633-48f9-ac24-634cc4e0fab1 · inbound

On Barriers to Archival Audio Processing cites this paper.

On Barriers to Archival Audio Processing XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T18:14:08.779769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:14:08.779769Z digest=sha256:adbc2365af7a12a67210211c218bcdbbe271398d2aa2c86831dd7a0e61c9a64f

Observation f1c21fd1-9842-447b-8a1c-14db27d29faa · inbound

SpeechFake: A Large-Scale Multilingual Speech Deepfake Dataset Incorporating Cutting-Edge Generation Methods cites this paper.

SpeechFake: A Large-Scale Multilingual Speech Deepfake Dataset Incorporating Cutting-Edge Generation Methods XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T12:49:16.380954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:49:16.380954Z digest=sha256:89e4bb68dcc3386f7b704b6cb0cf4e8ef0bab03e74f37d898ab7986b01d47e6f

Observation ddfdcf68-6bbb-418c-92ca-b7e5ef60aa85 · inbound

CAM\~OES: A Comprehensive Automatic Speech Recognition Benchmark for European Portuguese cites this paper.

CAM\~OES: A Comprehensive Automatic Speech Recognition Benchmark for European Portuguese XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-05T15:39:21.516370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:39:21.516370Z digest=sha256:144e3e0ad03413fca3c938e274340eb958752ca9c6ad66a6a79dc3962b2b1775

Observation b1e5f998-c1c8-4785-91e5-dd8b66b957c6 · inbound

Generalizable Audio Spoofing Detection using Non-Semantic Representations cites this paper.

Generalizable Audio Spoofing Detection using Non-Semantic Representations XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T13:55:18.277810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:55:18.277810Z digest=sha256:b45bbdea75879385787750bf9de1dcc8e1fddbb280c1fe41351c5e19cdeb2c95

Observation b423e264-8805-408e-8884-c868e8f5ad10 · inbound

Forensic Similarity for Speech Deepfakes cites this paper.

Forensic Similarity for Speech Deepfakes XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-18T10:31:14.902664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-18T10:28:12.467606Z digest=sha256:bd05ac7b1499614c26ef0c895f56b9e151c58df5a2e95f18d4e1e4bf533829a0

Observation 7ea9d4ba-0fda-47c3-ada1-5f2e3c64aae3 · inbound

SONAR: Spectral-Contrastive Audio Residuals for Generalizable Deepfake Detection cites this paper.

SONAR: Spectral-Contrastive Audio Residuals for Generalizable Deepfake Detection XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-03T20:05:29.992627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T20:05:29.992627Z digest=sha256:03bc4268262b3848a5eef0dcd0653aeed5ed70dae5ec6e9f7c9e81a012ba238a

Observation 98b8ee19-d213-48f5-9aa8-f181363ba2b3 · inbound

A SUPERB-Style Benchmark of Self-Supervised Speech Models for Audio Deepfake Detection cites this paper.

A SUPERB-Style Benchmark of Self-Supervised Speech Models for Audio Deepfake Detection XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:26:22.179791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T17:25:27.137423Z digest=sha256:169413f9db086d16ae27de86d5c1c74babe51271e31af080f61a9b40f932be18

Observation be6fd803-7145-4683-a19c-101dd423dd21 · inbound

Benchmarking Multilingual Speech Models on Pashto: Zero-Shot ASR, Script Failure, and Cross-Domain Evaluation cites this paper.

Benchmarking Multilingual Speech Models on Pashto: Zero-Shot ASR, Script Failure, and Cross-Domain Evaluation XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:35:48.700833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T19:44:30.762851Z digest=sha256:fbc82a38fdc437785076a6602784dc0927a2e8499f0d8f3a622e19f5c0e3e4cb

Observation db07feee-6407-4425-8621-09c452858f2c · inbound

Giving Voice to the Constitution: Low-Resource Text-to-Speech for Quechua and Spanish Using a Bilingual Legal Corpus cites this paper.

Giving Voice to the Constitution: Low-Resource Text-to-Speech for Quechua and Spanish Using a Bilingual Legal Corpus XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T11:21:00.993257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T15:01:20.902986Z digest=sha256:99ea0dfcf63ed09e4874272d3b70e104aaac7465ef611adc7f8b7299c403869b

Observation dbb3df5e-639c-4e03-a9f6-4c0a3e50176d · inbound

Similarity Choice and Negative Scaling in Supervised Contrastive Learning for Deepfake Audio Detection cites this paper.

Similarity Choice and Negative Scaling in Supervised Contrastive Learning for Deepfake Audio Detection XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:46:25.862041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-07T13:50:53.553261Z digest=sha256:7afdbfb048d144ba9b5ab036d4ded54d67d88c12e21e605525a953c2f9d3e51e

Observation e98638b2-ad8c-42d9-8f32-79e8ac4438e6 · inbound

Profiling the Voice: Speaker-Specific Phoneme Fingerprinting for Speech Deepfake Detection cites this paper.

Profiling the Voice: Speaker-Specific Phoneme Fingerprinting for Speech Deepfake Detection XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-20T01:22:55.817247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T01:20:04.609575Z digest=sha256:2d345c627199fe61066effacd369daecaab3a8698de8edcbdc34d6ddb84f0640

Observation cce6bf73-5148-4a1d-b5a5-29cb6cbcb3d4 · inbound

MixFake: Benchmarking and Enhancing Audio Deepfake Detection in Diverse Real-world Mixed Audio cites this paper.

MixFake: Benchmarking and Enhancing Audio Deepfake Detection in Diverse Real-world Mixed Audio XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-25T03:15:17.023264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T03:14:08.880072Z digest=sha256:37f4542ed7eccfb87dd91c87e4831291b8041d8dabb8f3396dca32dff8914d54

Observation ddaf9454-c96e-42c0-b5fe-c0ab99677897 · inbound

Escaping the Linearity Trap: Manifold Detours for Black-Box Adversarial Attacks on Singing Audio Deepfake Detection cites this paper.

Escaping the Linearity Trap: Manifold Detours for Black-Box Adversarial Attacks on Singing Audio Deepfake Detection XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-06-30T19:05:01.112227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T18:58:31.318883Z digest=sha256:18a01b6280c13fa4110b836e9b47f673b16ba97dac6b220bca65a68742b528d9

Observation 92c547c6-ca7c-4952-967d-fb9b26bfd232 · inbound

Dual-Branch Gated Fusion for Open-Set Audio Deepfake Source Tracing cites this paper.

Dual-Branch Gated Fusion for Open-Set Audio Deepfake Source Tracing XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:47:35.572305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T14:45:45.359616Z digest=sha256:c578262978e3948d69a4f18667b93e98992556c840265a2e18ba521c6c38b4d9

Observation f4260727-f010-4480-8756-d5f2a3dee562 · inbound

Pretrained self-supervised speech models can recognize unseen consonants cites this paper.

Pretrained self-supervised speech models can recognize unseen consonants XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:07:56.070600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T10:14:47.932613Z digest=sha256:9ce2148bf977aefa83486a0ca9352798a5eb64fbb5baf1988b46be3faa6a929f

Observation cbe71276-4f07-4fdb-8176-4a33c86cc761 · inbound

SpAArSIST: Sparsified AASIST for Efficient and Reliable Anti-Spoofing cites this paper.

SpAArSIST: Sparsified AASIST for Efficient and Reliable Anti-Spoofing XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-03T12:58:08.668281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T08:35:24.321197Z digest=sha256:9138866b625012a73d28cbddb525041c9b72368172689aaf3405ca73c12ba215

Observation 18ed6e2d-fcd1-4e0b-9178-354240c9df74 · inbound

Responsible ASR: Overcoming Challenges of Foundational Models in Narrow-Band and Low-Resource Settings cites this paper.

Responsible ASR: Overcoming Challenges of Foundational Models in Narrow-Band and Low-Resource Settings XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-04T01:59:26.658181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T20:03:28.491546Z digest=sha256:7f05409cb748ac805ed06af1271e57f85874f67bbd816268bdfddcd915db14e0

Observation dd75c079-7ec4-4819-9099-12cd136d5123 · inbound

Light-weight Pronunciation Assessment via Discrete Speech Token Surprisal cites this paper.

Light-weight Pronunciation Assessment via Discrete Speech Token Surprisal XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-07-04T03:49:30.231851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T17:37:25.043607Z digest=sha256:0c5a9ddb63655256239be81d0bcf02b79ab46770cda185d75518de767b904464

Observation dedb92c1-9f15-4ce5-8432-63b950b9a2fe · inbound

Beyond Speaker Independence: Evaluating Cross-Lingual Acoustic-to-Articulatory Inversion Across Finnish and Russian cites this paper.

Beyond Speaker Independence: Evaluating Cross-Lingual Acoustic-to-Articulatory Inversion Across Finnish and Russian XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-07-04T05:49:37.763673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T15:24:04.594993Z digest=sha256:d52c0075e03ecd88a4b81ee36a22051a976d9966587a91695d0e5d3f443d50bb

Observation e82d9116-99fc-4ed9-b1a5-7b70cbc47ce5 · inbound

Impact Analysis of Speech Representation Learning Models for Acoustic Side-Channel Attack cites this paper.

Impact Analysis of Speech Representation Learning Models for Acoustic Side-Channel Attack XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-07-04T06:39:37.331175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T14:19:59.573391Z digest=sha256:28427fc9f0e04964fc54f815169325d3e6bdd5948fa4b0bddfc0f55e37240501

Observation fbff2c09-d503-4660-b16d-78a7dabb35c4 · inbound

Impact Analysis of Speech Representation Learning Models for Acoustic Side-Channel Attack cites this paper.

Impact Analysis of Speech Representation Learning Models for Acoustic Side-Channel Attack XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-06-29T19:33:54.406475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T04:44:52.537405Z digest=sha256:a45a1536cb8c23391487a03ca0af2075505c897d31c968edb7edb1593659f428

Observation 61adb2a8-dc23-40af-a1f2-68848c557b19 · inbound

From Speech to Text Corpora: Evaluating ASR-Based Data Acquisition for Low-Resource Fongbe and Hausa cites this paper.

From Speech to Text Corpora: Evaluating ASR-Based Data Acquisition for Low-Resource Fongbe and Hausa XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-04T08:29:42.395446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T11:31:09.747113Z digest=sha256:48989d8d40504823074b31048ab5cc14f68f3cd95140d37d002df2009841a10f

Observation 5d125a69-5efe-41d1-b65a-f8de9b2f395c · inbound

What Do Deepfake Benchmarks Measure? An Audit Using Frozen Self-Supervised Representations cites this paper.

What Do Deepfake Benchmarks Measure? An Audit Using Frozen Self-Supervised Representations XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-04T15:49:57.200451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T01:24:46.200846Z digest=sha256:3b9aa0347f6ce83fc439823c99bb378cba3530b318a4de18a1fa19579bbbaf59

Observation 2a86893c-fa6b-413d-9c74-b2607ed4d8f7 · inbound

Syntactic Belief Update as the Driver of Garden Path Processing Difficulty cites this paper.

Syntactic Belief Update as the Driver of Garden Path Processing Difficulty XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 295

Resolution
metadata mismatch
arxiv_id, observed 2026-06-26T04:38:58.354104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-26T04:38:01.183423Z digest=sha256:34dce59135cc3039126cb9f0cfdc1958e147ca5e5fbe221d64c3eaeb92ddedd2

Observation 5e46740a-119f-4c26-ac54-9e04ccec34c6 · inbound

Detecting Audio Deepfakes on the Edge:Lightweight SSL-Based Detection in a Browser Plugin cites this paper.

Detecting Audio Deepfakes on the Edge:Lightweight SSL-Based Detection in a Browser Plugin XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-01T12:45:44.602950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-01T01:47:41.684335Z digest=sha256:be6989b943ebe3e462c001ecd01328e838e9af96ee064317a5fd220c88500cd6

Observation abbf0d63-cf32-495b-a9cb-570350d25ac4 · inbound

From Objectives to Applications: Aligning Architectural Biases in Audio Self-Supervised Learning cites this paper.

From Objectives to Applications: Aligning Architectural Biases in Audio Self-Supervised Learning XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 93

Resolution
verified exact
arxiv_id, observed 2026-07-02T05:56:39.846458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-02T05:52:55.818877Z digest=sha256:e2b0cf8882f5bc697a67276eae4287d47a3a60edfc22424920172302032b44f6

Observation 5a231095-8343-43cd-9b28-7b84916812e4 · inbound

Speaker-Aware Temporal Aggregation Strategies on Segment Representations for Depression Detection in Dyadic Interaction: A Benchmark Study cites this paper.

Speaker-Aware Temporal Aggregation Strategies on Segment Representations for Depression Detection in Dyadic Interaction: A Benchmark Study XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-12T06:14:42.321767Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:14:42.321767Z digest=sha256:3458fb796131fcafd2f65f571b0b9a126368c7482137adce13293da3252ee6ba

Observation 59a919b8-fb23-4828-adb2-3d3995836e7c · inbound

An Intervention-Based Framework for Shortcut Diagnosis in Spoofing Countermeasures cites this paper.

An Intervention-Based Framework for Shortcut Diagnosis in Spoofing Countermeasures XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-12T04:31:11.771594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T04:31:11.771594Z digest=sha256:1aaf78e173ad23318a6dfef13904fc6128bdab659ea60f5072908027a486053a

Observation 5265bd2a-3338-4e30-b7ee-b6840c656f53 · inbound

Towards Digital Preservation of Efik: TTS for a Low-Resource African Language cites this paper.

Towards Digital Preservation of Efik: TTS for a Low-Resource African Language XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-11T18:11:36.626189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:11:36.626189Z digest=sha256:b4bbb51efa81940ae4c3255d4a6d7a1f698a106acbceb2f3e9a53d51333ccd5c

Observation 83eed9c1-9c22-4fb8-b782-3ff5b7d21eb3 · inbound

Evaluating the Effect of Linguistic Relatedness on Cross-Lingual Transfer in Large Multilingual Automatic Speech Recognition cites this paper.

Evaluating the Effect of Linguistic Relatedness on Cross-Lingual Transfer in Large Multilingual Automatic Speech Recognition XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-11T13:07:13.175134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T13:07:13.175134Z digest=sha256:81a05eafc65fd7829fb044578f127fa7898b0917227b97142343bbdcdcad659d

Observation 087998ef-6c92-4171-8258-2dacc3719058 · inbound

Evaluating the Effect of Linguistic Relatedness on Cross-Lingual Transfer in Large Multilingual Automatic Speech Recognition cites this paper.

Evaluating the Effect of Linguistic Relatedness on Cross-Lingual Transfer in Large Multilingual Automatic Speech Recognition XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T08:34:55.605107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T08:34:55.605107Z digest=sha256:ecee23613a10c9938623ebe5e152e2635698d6aa5edaf24be6b056493c652661

Observation 2d8c6bc4-790e-437b-acb7-3688e45872b3 · inbound

GigaAM Multilingual: Foundation Model for Underrepresented Languages cites this paper.

GigaAM Multilingual: Foundation Model for Underrepresented Languages XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-14T12:15:06.470439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:15:06.470439Z digest=sha256:0a3848ceb97dc5c1dce8188a765a945ce09179a23df1f8fa065586673d57326a

Observation 0b197b28-1724-4dbe-9a52-d2703304a7d9 · inbound

Unified Gradient Projection: Language-Balanced Continual Learning for Multilingual Low-Resource ASR cites this paper.

Unified Gradient Projection: Language-Balanced Continual Learning for Multilingual Low-Resource ASR XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-14T06:34:55.089754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:34:55.089754Z digest=sha256:a19ed681d48c7083fbeb0dd45ab69cdcd675902a52ee883448e16c7a5da30c15

Observation 397ebc1a-c572-4365-b947-30a876530d83 · inbound

DONDO: Open w2v-BERT Speech-Recognition Base Models for African Languages cites this paper.

DONDO: Open w2v-BERT Speech-Recognition Base Models for African Languages XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T07:10:39.837818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T07:10:39.837818Z digest=sha256:58123987cbbd5b79f03934f4a9beb933a9fe4d49222fe3aab144d89f9c96754c

Observation af26c9ad-3b23-4deb-be16-701e9507960f · inbound

Leveraging Gradient Reversal Loss and Multitask Learning for Datasets-Aware Audio Deepfake Detection cites this paper.

Leveraging Gradient Reversal Loss and Multitask Learning for Datasets-Aware Audio Deepfake Detection XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-31T23:27:02.078077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:27:02.078077Z digest=sha256:250acc0ef0f02e11c180d79dce00e0e080c0802f887d9fc14134b82b7b87b41f

Observation 46196a50-20d5-4b09-b76d-b2dd78425e91 · inbound

Teffic-Audio: Tell Fact from Fiction cites this paper.

Teffic-Audio: Tell Fact from Fiction XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-31T10:28:22.087327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T10:28:22.087327Z digest=sha256:08033b10a45c2afbd2e98c0b56277ed83dc50f58f24c5a6c224c996325124372

Observation d97e4775-c32e-4e8e-a2a1-0ac095920b20 · inbound

Multi-Backbone Self-Supervised Ensembles for Audio Deepfake Detection and a Cross-Track Analysis of Generation-Detection Asymmetry cites this paper.

Multi-Backbone Self-Supervised Ensembles for Audio Deepfake Detection and a Cross-Track Analysis of Generation-Detection Asymmetry XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T20:44:55.686214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:44:55.686214Z digest=sha256:9a077b223c27919274f1c13e76d0371622c2914e70fe9dbc4a288c73a8277b0f