Pith. sign in

Paper Citation Record · LEDGER

Unsupervised Cross-lingual Representation Learning for Speech Recognition

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2006.13979.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2006.13979 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:34:49.820015Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T08:09:41.969238Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f1b31344-c420-416c-9a96-ebf90e51f097 · inbound

Data Quality Issues in Multilingual Speech Datasets: The Need for Sociolinguistic Awareness and Proactive Language Planning cites this paper.

Data Quality Issues in Multilingual Speech Datasets: The Need for Sociolinguistic Awareness and Proactive Language Planning Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:49.820015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:34:49.820015Z digest=sha256:1c5fc1149b63ca77bd6195fc598ec01a50ab643d96ce4097a98d55c995013835

Observation 88cb9a4f-b88e-4145-8c66-caeb5886c12d · inbound

Leveraging Large Language Models for Spontaneous Speech-Based Suicide Risk Detection cites this paper.

Leveraging Large Language Models for Spontaneous Speech-Based Suicide Risk Detection Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T21:13:14.129554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:13:14.129554Z digest=sha256:6a514cbfadf908903f5c5b7602d0a4070a76c1ee3dd1eabf9ac977cc438a94ff

Observation c0d15c12-3187-4410-894e-48984a8915a1 · inbound

Hybrid Decoding: Rapid Pass and Selective Detailed Correction for Sequence Models cites this paper.

Hybrid Decoding: Rapid Pass and Selective Detailed Correction for Sequence Models Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T15:39:22.322031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:39:22.322031Z digest=sha256:6e27eb1b7ab301fea659966059c1e3099eed73f044aaa836411ba1c24cfaada1

Observation 8830ce8b-09bc-4675-af92-97d446d65ef9 · inbound

TokenVerse++: Towards Flexible Multitask Learning with Dynamic Task Activation cites this paper.

TokenVerse++: Towards Flexible Multitask Learning with Dynamic Task Activation Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T15:27:29.308951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:27:29.308951Z digest=sha256:2c78d496c7ed99d7f9d3dcd8abb42ceafc49a6d5cd59a866d160307acab0cd18

Observation 41b657f4-015d-434b-b217-61c7c4443fe8 · inbound

Bona fide Cross Testing Reveals Weak Spot in Audio Deepfake Detection Systems cites this paper.

Bona fide Cross Testing Reveals Weak Spot in Audio Deepfake Detection Systems Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:20.032674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:20.032674Z digest=sha256:ca7f8ed9d872492e82e1de29d62b02da4a431bb820c0401af800cbdfa12193e9

Observation 28d5c376-de19-43ae-80aa-0fbd64522cbc · inbound

Prominence-aware automatic speech recognition for conversational speech cites this paper.

Prominence-aware automatic speech recognition for conversational speech Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T18:11:15.997955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:11:15.997955Z digest=sha256:208de21e125e04be7aea852e8dd71e53aa82f369c71a95bd5eb19d55c1aba834

Observation 721520c4-c719-475d-9685-3c36740ee176 · inbound

UniVoice: Unifying Autoregressive ASR and Flow-Matching based TTS with Large Language Models cites this paper.

UniVoice: Unifying Autoregressive ASR and Flow-Matching based TTS with Large Language Models Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T11:29:31.132461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:29:31.132461Z digest=sha256:790f1f6f389d0f78933a0bbd92b557b46ee736bfeb50a46c1a8f250e154ef7c3

Observation 37726469-f05b-4102-ad28-7c08f9ba63ff · inbound

FAC-FACodec: Controllable Zero-Shot Foreign Accent Conversion with Factorized Speech Codec cites this paper.

FAC-FACodec: Controllable Zero-Shot Foreign Accent Conversion with Factorized Speech Codec Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-18T08:01:06.824442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-18T07:56:46.302918Z digest=sha256:f40ce02a0f9b9950c406e350864bd7e3a843379f91ea16d67db31eddd213849e

Observation 56f4377f-3a06-45be-ac84-90a15a47fd14 · inbound

ENEC: A Lossless AI Model Compression Method Enabling Fast Inference on Ascend NPUs cites this paper.

ENEC: A Lossless AI Model Compression Method Enabling Fast Inference on Ascend NPUs Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:58:03.011495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:56:38.684104Z digest=sha256:7b09afc12f875bd41cbb01669323f1c608567147ee4f146b1b54022279c88dee

Observation ea73a65e-fc5b-49f6-9e0d-fcaea6b83d8e · inbound

Virtual Speech Therapist: A Clinician-in-the-Loop AI Speech Therapy Agent for Personalized and Supervised Therapy cites this paper.

Virtual Speech Therapist: A Clinician-in-the-Loop AI Speech Therapy Agent for Personalized and Supervised Therapy Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 47

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:47:24.686187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-09T19:11:40.161390Z digest=sha256:7bfc48b8d027252ddd4c1a21aa60ff3570305e5afd4a95c851e2371632ba9ced

Observation 3d33b2a8-1c9f-4932-af46-b461b42c97c4 · inbound

A study on the impact of region specific data on the performance of Indic ASR cites this paper.

A study on the impact of region specific data on the performance of Indic ASR Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:37:35.604071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T15:06:19.198758Z digest=sha256:a7d96353dd795b8c6dfd78444dbc674e6fe4ac8b3d9cf18ae1bdec02eee7561b

Observation 8c183895-9eba-4a4d-a3e4-56483739f3b4 · inbound

Overcoming Decoder Inconsistencies in Whisper for Dravidian and Low-Resource Languages cites this paper.

Overcoming Decoder Inconsistencies in Whisper for Dravidian and Low-Resource Languages Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-03T01:47:31.183833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T16:21:39.960715Z digest=sha256:d0e7c696d384389484b49f3c9cc5d183749e15ca13a734de2f9e48619fb95456

Observation 1ea0e690-d1bd-42d6-a64e-af5e414ccedf · inbound

Pretrained self-supervised speech models can recognize unseen consonants cites this paper.

Pretrained self-supervised speech models can recognize unseen consonants Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:07:56.057986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T10:14:47.932613Z digest=sha256:eabd548e7c5b388fb7e0333bfd2b794dd76fdfea04de190e5a700f2d69ce3c1f

Observation decfd35a-bd0a-432d-9585-356787c27d38 · inbound

Adding Robust Code-Switching Capabilities to High Performance Multilingual ASR cites this paper.

Adding Robust Code-Switching Capabilities to High Performance Multilingual ASR Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-07-04T08:09:41.971092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-26T12:01:57.319491Z digest=sha256:bae1f9fb87064ef9dd48f736a3967a729b2720eae5faff07581a4cd40aed1caf

Observation a6500f5d-adf0-422f-b698-46425134df1c · inbound

Which Languages Transfer Best to Warlpiri? A Similarity-Based Study for Low-Resource ASR cites this paper.

Which Languages Transfer Best to Warlpiri? A Similarity-Based Study for Low-Resource ASR Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-14T13:08:11.026186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T13:08:11.026186Z digest=sha256:6dc112c38ed5615b64870fae23a0c9c8bbdf9bd043c9b35c208452209caf0ff2

Observation c0e44047-2705-449b-9c3c-749894f32257 · inbound

Less is More: Modality-Decoupling for General AIGC Audio-Video Detection cites this paper.

Less is More: Modality-Decoupling for General AIGC Audio-Video Detection Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-01T02:09:49.283745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:09:49.283745Z digest=sha256:83aef0b4b669abf8f7e186222b87d394df952ee21aa59af452ea87f8429cf703