Pith. sign in

Paper Citation Record · LEDGER

Multilingual Speech Recognition Using Discrete Tokens with a Two-step Training Strategy

As of 15 August 2026, this Paper Citation Record lists 16 of 16 outbound references and 0 inbound Pith citation observations for arXiv:2509.01900.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.01900 v1

Coverage vector

measured 16 of 16 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T12:08:09.011558Z

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

16 of 16 outbound references displayed

  • verified exact0
  • verified fuzzy6
  • unresolved8
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 46bd497d-7276-4b6b-8f3d-5db437760bad · outbound

This paper cites IEEE Journal of Selected Topics in Signal Processing 11(8), 1240–1253 (2017) 5 https://huggingface.co/microsoft/wavlm-large 8 Z.

Multilingual Speech Recognition Using Discrete Tokens with a Two-step Training Strategy IEEE Journal of Selected Topics in Signal Processing 11(8), 1240–1253 (2017) 5 https://huggingface.co/microsoft/wavlm-large 8 Z

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T12:08:09.259385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T12:08:08.960604Z digest=sha256:fcfd2317d5eb33b97d40531b23a97ffd573fb52eec6e37c078152184346d2bcd

Observation f7b6ef47-c1e3-4799-89a5-052c7c47c45a · outbound

This paper cites In: Proc.

Multilingual Speech Recognition Using Discrete Tokens with a Two-step Training Strategy In: Proc

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T12:08:08.964201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:08:08.964201Z digest=sha256:b2e313c37c988a20ec0f3e9de6c29e5fed9e653dd8347518c080f86a2e891466

Observation 5da9d4ec-2121-4beb-b11e-7bc1eae123f5 · outbound

This paper cites In: Proc.

Multilingual Speech Recognition Using Discrete Tokens with a Two-step Training Strategy In: Proc

Reference 3

Resolution
malformed identifier
no resolver link, observed 2026-08-05T12:08:08.967802Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:08:08.967802Z digest=sha256:2b9e6a9021e77a5827f8d0dce41b9e1a2da614721fa2f8b25c0f4b800f1d1944

Observation f5b3bc3e-0f68-47fc-bd61-041949817533 · outbound

This paper cites an unresolved cited work.

Multilingual Speech Recognition Using Discrete Tokens with a Two-step Training Strategy Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-05T12:08:09.248561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T12:08:08.971122Z digest=sha256:20e59ed956061c06298e6bf961ea6b774ead5b5b7d48c873767c22b1f87ff54f

Observation fd3ec435-e5ec-4973-b708-5bf7db519b34 · outbound

This paper cites Advances in Neural Information Processing Systems 33, 12449–12460 (2020).

Multilingual Speech Recognition Using Discrete Tokens with a Two-step Training Strategy Advances in Neural Information Processing Systems 33, 12449–12460 (2020)

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T12:08:09.237912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T12:08:08.974857Z digest=sha256:2fd41de2b2e4b7ec49f2274240ffce55ca7f0e1d115998e154cda7f20370a695

Observation d0b67b0a-1839-4a7a-9bf6-1f10ba157ae0 · outbound

This paper cites IEEE/ACM Transactions on Audio, Speech, and Language Process- ing 29, 3451–3460 (2021).

Multilingual Speech Recognition Using Discrete Tokens with a Two-step Training Strategy IEEE/ACM Transactions on Audio, Speech, and Language Process- ing 29, 3451–3460 (2021)

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T12:08:09.227113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T12:08:08.978098Z digest=sha256:da863ff0579129112ec893597150bdbee4a513696743c7a38b23ce71a1eeed4f

Observation 54e0254d-cd4c-446f-9683-d03c8fa47e1c · outbound

This paper cites In: Interna- tional Conference on Machine Learning.

Multilingual Speech Recognition Using Discrete Tokens with a Two-step Training Strategy In: Interna- tional Conference on Machine Learning

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T12:08:09.217089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T12:08:08.981623Z digest=sha256:4770387d1f52bb88ce14d1b75cffc0320c27dc1b5239ad07c29d48c8975dc6a9

Observation d7f7bd33-9691-4cad-807b-5c57656fd400 · outbound

This paper cites Towards Universal Speech Discrete Tokens: A Case Study for ASR and TTS.

Multilingual Speech Recognition Using Discrete Tokens with a Two-step Training Strategy Towards Universal Speech Discrete Tokens: A Case Study for ASR and TTS

Reference 8

Resolution
metadata mismatch
local_arxiv, observed 2026-08-05T12:08:09.176710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T12:08:08.984761Z digest=sha256:872183e9800264a7a25697cb0d03a5ee089e59b839dcb95e1d0e5030ab98d611

Observation c6d2b590-1a40-4552-9036-50c9be9a64e6 · outbound

This paper cites Textless Speech-to-Speech Translation on Real Data.

Multilingual Speech Recognition Using Discrete Tokens with a Two-step Training Strategy Textless Speech-to-Speech Translation on Real Data

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T12:08:08.988069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:08:08.988069Z digest=sha256:b47c9b2d23a6a8eea733e12955508ccc04b4d2a945db66b4191511f061b211b7

Observation 58ece04e-a0b7-403f-9331-96d0bd98817f · outbound

This paper cites SpeechGPT: Empowering Large Language Models with Intrinsic Cross-Modal Conversational Abilities.

Multilingual Speech Recognition Using Discrete Tokens with a Two-step Training Strategy SpeechGPT: Empowering Large Language Models with Intrinsic Cross-Modal Conversational Abilities

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T12:08:08.991738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:08:08.991738Z digest=sha256:611a94494c8ed8474c653ffeed4b2fec735d19424bd21bf83e4c7bdb2074cc06

Observation 7f3ad701-dcea-4c3b-bad3-7a246672d405 · outbound

This paper cites VioLA: Unified Codec Language Models for Speech Recognition, Synthesis, and Translation.

Multilingual Speech Recognition Using Discrete Tokens with a Two-step Training Strategy VioLA: Unified Codec Language Models for Speech Recognition, Synthesis, and Translation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T12:08:08.995130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:08:08.995130Z digest=sha256:d3af6ec862c3dc7e7c7395dff4220dc076caa65088c408286e4c98a0ded9ca80

Observation d126038c-dfc4-4d38-8df3-5602c5a08959 · outbound

This paper cites In: ICASSP 2023-2023 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP).

Multilingual Speech Recognition Using Discrete Tokens with a Two-step Training Strategy In: ICASSP 2023-2023 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T12:08:09.207394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T12:08:08.998466Z digest=sha256:167059d2d992f4dafedf00522323ab3d45a5b3c8170302c921b482007d9cfa49

Observation bb859c70-94e9-4f38-8d10-c783ab15536b · outbound

This paper cites In: 2021 IEEE Automatic Speech Recognition and Under- standing Workshop (ASRU).

Multilingual Speech Recognition Using Discrete Tokens with a Two-step Training Strategy In: 2021 IEEE Automatic Speech Recognition and Under- standing Workshop (ASRU)

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T12:08:09.197545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T12:08:09.001744Z digest=sha256:a31765171c91f036e09c45b0a3f2d3978245fa873ef872ec98084b732b97146c

Observation 1194f928-b990-4663-b947-f6561b72b815 · outbound

This paper cites In: Proc.

Multilingual Speech Recognition Using Discrete Tokens with a Two-step Training Strategy In: Proc

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T12:08:09.004839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:08:09.004839Z digest=sha256:382e1fb88af34dcd2f819dbae1018c6463c3fb412357ae3e338d77b56213d348

Observation 95b30e7a-da77-49fb-98bf-c33d544c107c · outbound

This paper cites an unresolved cited work.

Multilingual Speech Recognition Using Discrete Tokens with a Two-step Training Strategy Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-05T12:08:09.187467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T12:08:09.008142Z digest=sha256:6aa371fdf21c09fccacd8971b7bf3eec7eace140162ec1ef44dcc363602dbf52

Observation 2ec08c07-70a4-41b3-8da3-62cf385ce270 · outbound

This paper cites In: 2022 IEEE Spoken Language Technology Workshop (SLT).

Multilingual Speech Recognition Using Discrete Tokens with a Two-step Training Strategy In: 2022 IEEE Spoken Language Technology Workshop (SLT)

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T12:08:09.011558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:08:09.011558Z digest=sha256:fa2c7cc231a89c40089b3a6fec23cca6823500d88b5319df86e68d045c8c692f

Pith citing papers

No inbound Pith citation observations are available.