Pith. sign in

Paper Citation Record · LEDGER

Scaling Speech Technology to 1,000+ Languages

As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2305.13516.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2305.13516 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:56:20.540912Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

116
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e3d61ba3-faed-4c9c-a3ef-41e280caaed2 · inbound

MLAAD: The Multi-Language Audio Anti-Spoofing Dataset cites this paper.

MLAAD: The Multi-Language Audio Anti-Spoofing Dataset Scaling Speech Technology to 1,000+ Languages

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-24T04:36:00.889425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-24T04:34:59.205522Z digest=sha256:aecd77c98b2b0235fcf7d5271f8bfe29c5ce75fafa919822e2d464079744b441

Observation 5fab6ba7-56d8-410e-b917-55ded70f0929 · inbound

On Barriers to Archival Audio Processing cites this paper.

On Barriers to Archival Audio Processing Scaling Speech Technology to 1,000+ Languages

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T18:14:08.841029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:14:08.841029Z digest=sha256:fc0069183ab65f482d21b3666c9bf00a1e8b998ef22f337b53546907eb7c5a54

Observation ab8d4adc-37bd-45eb-920c-f7370d130374 · inbound

A Hybrid Machine Learning Framework for Optimizing Crop Selection via Agronomic and Economic Forecasting cites this paper.

A Hybrid Machine Learning Framework for Optimizing Crop Selection via Agronomic and Economic Forecasting Scaling Speech Technology to 1,000+ Languages

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T19:56:20.540912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:56:20.540912Z digest=sha256:0e6d2ca2a8bcacd6ef3d2078c46809b85d27ee0d08272a639a1d81faf9af29b1

Observation c7afc37d-fa66-4172-ae08-6e15342c902e · inbound

Coherence in the brain unfolds across separable temporal regimes cites this paper.

Coherence in the brain unfolds across separable temporal regimes Scaling Speech Technology to 1,000+ Languages

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:13:22.862525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T20:11:29.476625Z digest=sha256:b21c34482f24d3ca9474cf1ce762b9769df7856af2cbc65e3bf059c5007b736f

Observation ddd22543-ba0a-49e6-a203-fe708de2b2d7 · inbound

Evaluating Generalization and Robustness in Russian Anti-Spoofing: The RuASD Initiative cites this paper.

Evaluating Generalization and Robustness in Russian Anti-Spoofing: The RuASD Initiative Scaling Speech Technology to 1,000+ Languages

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:46:12.487399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T02:27:05.776290Z digest=sha256:475bf493647a59298cf7df3bfa535c049bf683509b87779da966c9dc8c3b5d5c

Observation 60d89d46-3473-489f-b9c7-f81ed640d2c2 · inbound

Benchmarking Multilingual Speech Models on Pashto: Zero-Shot ASR, Script Failure, and Cross-Domain Evaluation cites this paper.

Benchmarking Multilingual Speech Models on Pashto: Zero-Shot ASR, Script Failure, and Cross-Domain Evaluation Scaling Speech Technology to 1,000+ Languages

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:35:48.754248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T19:44:30.762851Z digest=sha256:86a99170933e80d83e62c391e570a69f19e5413dc1fe04c9b3a3775ed6d3665a

Observation 5a19eb96-1e3a-4e65-864c-e643fa3a51a4 · inbound

Training-Free Cross-Lingual Dysarthria Severity Assessment via Phonological Subspace Analysis in Self-Supervised Speech Representations cites this paper.

Training-Free Cross-Lingual Dysarthria Severity Assessment via Phonological Subspace Analysis in Self-Supervised Speech Representations Scaling Speech Technology to 1,000+ Languages

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:05:58.892265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T16:16:39.890339Z digest=sha256:7077ac9d8515ce4e5d1cd58ccc190088749dd56b072ddfc60d552bd09ac3fe16

Observation 5e08dc38-63c2-469d-8de7-cc5d58a180b1 · inbound

BlasBench: An Open Benchmark for Irish Speech Recognition cites this paper.

BlasBench: An Open Benchmark for Irish Speech Recognition Scaling Speech Technology to 1,000+ Languages

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:41:01.558355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-10T15:53:54.092426Z digest=sha256:f6100b4feee42c39bc97e6b9447c0a0c4b1b20f2eb8509bc2870c42670d31dcf

Observation d41f838d-44b4-43e7-a11a-56f15e06e0f4 · inbound

Tadabur: A Large-Scale Quran Audio Dataset cites this paper.

Tadabur: A Large-Scale Quran Audio Dataset Scaling Speech Technology to 1,000+ Languages

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:11:03.683889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T02:16:09.216191Z digest=sha256:4ece5aa49a8d1c1c1b390223910a3d81fa3f0b4f793f3df30bca8cca2c0442dc

Observation 1d0a1619-11f1-432b-8a93-44166d0ee087 · inbound

A framework for analyzing concept representations in neural models cites this paper.

A framework for analyzing concept representations in neural models Scaling Speech Technology to 1,000+ Languages

Reference 189

Resolution
verified exact
arxiv_id, observed 2026-05-09T22:18:59.078659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-09T14:49:22.776209Z digest=sha256:e30adcd786a5433743d4bba55f3063b1b85da098fb2428e7401e352b0f0bf8a1

Observation f7355c90-ca9c-4bb5-971f-fb6fc32b7c38 · inbound

A Comparative Study of Pre-trained Speech Encoders and Training Objectives for Large-Scale Indic Spoken Language Identification cites this paper.

A Comparative Study of Pre-trained Speech Encoders and Training Objectives for Large-Scale Indic Spoken Language Identification Scaling Speech Technology to 1,000+ Languages

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:37:35.361386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T15:12:45.261874Z digest=sha256:8d2859599935e61179eab14d2e6823c45880ae444cef7b8fb02b4e07fa3b4591

Observation 41a4aab9-8e7e-4b78-90f1-602b7d0e7f87 · inbound

Pretrained self-supervised speech models can recognize unseen consonants cites this paper.

Pretrained self-supervised speech models can recognize unseen consonants Scaling Speech Technology to 1,000+ Languages

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:07:56.073107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T10:14:47.932613Z digest=sha256:8a216823ab1dbdb5dd0dcba248e9d322fac45ec0c756027224533600ae9e8a9b

Observation c7c44471-50e5-4550-bed9-a73c4eed6376 · inbound

Closing the Quality Gap in Low-Resource Text-to-Speech: LoRA Fine-Tuning of VoxCPM2 for Khmer and Korean cites this paper.

Closing the Quality Gap in Low-Resource Text-to-Speech: LoRA Fine-Tuning of VoxCPM2 for Khmer and Korean Scaling Speech Technology to 1,000+ Languages

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:19:50.277522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-26T05:24:03.268025Z digest=sha256:6c1be65a6d4abdf2622f7ea348b97155c0b107956cd27f226089ab5465c94334

Observation 10db3fde-f1e0-4ce7-a1ea-aa0756afa508 · inbound

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling cites this paper.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Scaling Speech Technology to 1,000+ Languages

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:f3cb8068031ae7e9d77753eb51cb3b8a0f38c57447533dbe550fbb5681274fcb

Observation ba5de31c-da83-4527-8f35-47a42e0f611b · inbound

Towards Digital Preservation of Efik: TTS for a Low-Resource African Language cites this paper.

Towards Digital Preservation of Efik: TTS for a Low-Resource African Language Scaling Speech Technology to 1,000+ Languages

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-11T18:11:36.626189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:11:36.626189Z digest=sha256:2698b6ff4aabf82b1b794b43d7dabd35779dd8d5091d8a29a9fbbfcb9b0a41e8

Observation 69f96c9a-58e0-4b13-9b50-9cae96658e6c · inbound

DONDO: Open w2v-BERT Speech-Recognition Base Models for African Languages cites this paper.

DONDO: Open w2v-BERT Speech-Recognition Base Models for African Languages Scaling Speech Technology to 1,000+ Languages

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T07:10:39.841258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T07:10:39.841258Z digest=sha256:8b5c2f4f38e2c446b4bbc5695701251404d9e14e5eae20830153238c04130fb1