Pith. sign in

Paper Citation Record · LEDGER

Scaling Speech Technology to 1,000+ Languages

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2305.13516.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2305.13516 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:48:39.174752Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

116
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e3d61ba3-faed-4c9c-a3ef-41e280caaed2 · inbound

MLAAD: The Multi-Language Audio Anti-Spoofing Dataset cites this paper.

MLAAD: The Multi-Language Audio Anti-Spoofing Dataset Scaling Speech Technology to 1,000+ Languages

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-24T04:36:00.889425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-24T04:34:59.205522Z digest=sha256:90ecfca2e6c0a10854dc220717c2c40f58ee82face80a73409b34cb092a0e531

Observation 307f9388-67b8-41fa-9068-7d207fb2a396 · inbound

Double Entendre: Robust Audio-Based AI-Generated Lyrics Detection via Multi-View Fusion cites this paper.

Double Entendre: Robust Audio-Based AI-Generated Lyrics Detection via Multi-View Fusion Scaling Speech Technology to 1,000+ Languages

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T23:48:39.174752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:48:39.174752Z digest=sha256:54494524e1bb6e38e867b694e771d261a49f773b904b094272f65d99d9c4a0e6

Observation 5fab6ba7-56d8-410e-b917-55ded70f0929 · inbound

On Barriers to Archival Audio Processing cites this paper.

On Barriers to Archival Audio Processing Scaling Speech Technology to 1,000+ Languages

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T18:14:08.841029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:14:08.841029Z digest=sha256:fc0069183ab65f482d21b3666c9bf00a1e8b998ef22f337b53546907eb7c5a54

Observation ab8d4adc-37bd-45eb-920c-f7370d130374 · inbound

A Hybrid Machine Learning Framework for Optimizing Crop Selection via Agronomic and Economic Forecasting cites this paper.

A Hybrid Machine Learning Framework for Optimizing Crop Selection via Agronomic and Economic Forecasting Scaling Speech Technology to 1,000+ Languages

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T19:56:20.540912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:56:20.540912Z digest=sha256:0e6d2ca2a8bcacd6ef3d2078c46809b85d27ee0d08272a639a1d81faf9af29b1

Observation c7afc37d-fa66-4172-ae08-6e15342c902e · inbound

Coherence in the brain unfolds across separable temporal regimes cites this paper.

Coherence in the brain unfolds across separable temporal regimes Scaling Speech Technology to 1,000+ Languages

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:13:22.862525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T20:11:29.476625Z digest=sha256:3b813b4d14965abc59de7ee93ebe7b0bae928cfe4fe34f2838654fb3413ea68d

Observation ddd22543-ba0a-49e6-a203-fe708de2b2d7 · inbound

Evaluating Generalization and Robustness in Russian Anti-Spoofing: The RuASD Initiative cites this paper.

Evaluating Generalization and Robustness in Russian Anti-Spoofing: The RuASD Initiative Scaling Speech Technology to 1,000+ Languages

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:46:12.487399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T02:27:05.776290Z digest=sha256:551831b28894a4760ffe5519ed3076f662d304f78ebb7df5d7efece4307eab44

Observation 60d89d46-3473-489f-b9c7-f81ed640d2c2 · inbound

Benchmarking Multilingual Speech Models on Pashto: Zero-Shot ASR, Script Failure, and Cross-Domain Evaluation cites this paper.

Benchmarking Multilingual Speech Models on Pashto: Zero-Shot ASR, Script Failure, and Cross-Domain Evaluation Scaling Speech Technology to 1,000+ Languages

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:35:48.754248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T19:44:30.762851Z digest=sha256:dcd2ef9a1d89d392cc0740c1a441872f95d5a4af3d36fbb82908eb8358d45354

Observation 5a19eb96-1e3a-4e65-864c-e643fa3a51a4 · inbound

Training-Free Cross-Lingual Dysarthria Severity Assessment via Phonological Subspace Analysis in Self-Supervised Speech Representations cites this paper.

Training-Free Cross-Lingual Dysarthria Severity Assessment via Phonological Subspace Analysis in Self-Supervised Speech Representations Scaling Speech Technology to 1,000+ Languages

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:05:58.892265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T16:16:39.890339Z digest=sha256:098b8fdf61cc456100d81bc8f1a4b03fed51ef6eb485f76e66d6444f4f15ddff

Observation 5e08dc38-63c2-469d-8de7-cc5d58a180b1 · inbound

BlasBench: An Open Benchmark for Irish Speech Recognition cites this paper.

BlasBench: An Open Benchmark for Irish Speech Recognition Scaling Speech Technology to 1,000+ Languages

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:41:01.558355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-10T15:53:54.092426Z digest=sha256:7670563b518d53d5803150ce6e0103d5e124d8c96e2656b1448166eb243630c3

Observation d41f838d-44b4-43e7-a11a-56f15e06e0f4 · inbound

Tadabur: A Large-Scale Quran Audio Dataset cites this paper.

Tadabur: A Large-Scale Quran Audio Dataset Scaling Speech Technology to 1,000+ Languages

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:11:03.683889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T02:16:09.216191Z digest=sha256:58f28418d9d4c995e41fdb151d254209b8feef37b55b82f8005147b4b425deef

Observation 1d0a1619-11f1-432b-8a93-44166d0ee087 · inbound

A framework for analyzing concept representations in neural models cites this paper.

A framework for analyzing concept representations in neural models Scaling Speech Technology to 1,000+ Languages

Reference 189

Resolution
verified exact
arxiv_id, observed 2026-05-09T22:18:59.078659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-09T14:49:22.776209Z digest=sha256:3e42914e0cd6505b648ccd3e344475e8150a81e66c8ac9dae40cf045858288d0

Observation f7355c90-ca9c-4bb5-971f-fb6fc32b7c38 · inbound

A Comparative Study of Pre-trained Speech Encoders and Training Objectives for Large-Scale Indic Spoken Language Identification cites this paper.

A Comparative Study of Pre-trained Speech Encoders and Training Objectives for Large-Scale Indic Spoken Language Identification Scaling Speech Technology to 1,000+ Languages

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:37:35.361386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T15:12:45.261874Z digest=sha256:3ac61653d3103c999534db025274a26eb00a67fa08a3a8a9f37ca5dffabe187d

Observation 41a4aab9-8e7e-4b78-90f1-602b7d0e7f87 · inbound

Pretrained self-supervised speech models can recognize unseen consonants cites this paper.

Pretrained self-supervised speech models can recognize unseen consonants Scaling Speech Technology to 1,000+ Languages

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:07:56.073107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T10:14:47.932613Z digest=sha256:6dc8da9b4b17a0d3f76f7b2b3b2901826d353d77cb320ff3fe4f4176c2aae8aa

Observation c7c44471-50e5-4550-bed9-a73c4eed6376 · inbound

Closing the Quality Gap in Low-Resource Text-to-Speech: LoRA Fine-Tuning of VoxCPM2 for Khmer and Korean cites this paper.

Closing the Quality Gap in Low-Resource Text-to-Speech: LoRA Fine-Tuning of VoxCPM2 for Khmer and Korean Scaling Speech Technology to 1,000+ Languages

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:19:50.277522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T05:24:03.268025Z digest=sha256:58b8b06a33ffa3757e0c97fe84f307b483f1eae2ef6685f339eda1485337f7f0

Observation 10db3fde-f1e0-4ce7-a1ea-aa0756afa508 · inbound

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling cites this paper.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Scaling Speech Technology to 1,000+ Languages

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:f3cb8068031ae7e9d77753eb51cb3b8a0f38c57447533dbe550fbb5681274fcb

Observation ba5de31c-da83-4527-8f35-47a42e0f611b · inbound

Towards Digital Preservation of Efik: TTS for a Low-Resource African Language cites this paper.

Towards Digital Preservation of Efik: TTS for a Low-Resource African Language Scaling Speech Technology to 1,000+ Languages

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-11T18:11:36.626189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:11:36.626189Z digest=sha256:2698b6ff4aabf82b1b794b43d7dabd35779dd8d5091d8a29a9fbbfcb9b0a41e8

Observation 69f96c9a-58e0-4b13-9b50-9cae96658e6c · inbound

DONDO: Open w2v-BERT Speech-Recognition Base Models for African Languages cites this paper.

DONDO: Open w2v-BERT Speech-Recognition Base Models for African Languages Scaling Speech Technology to 1,000+ Languages

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T07:10:39.841258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T07:10:39.841258Z digest=sha256:da7f90ed01aa3be4c6b4e9c5de59b0f35a26bfd337662c90762a404c3d252cf2