Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:35:38.605780Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 2 inbound Pith citation observations for arXiv:2506.01845.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:35:38.605780Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:35:33.841116Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-18T22:21:53.596123Z
50 of 50 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 3b3040d7-d133-4d99-99b7-a3186741704a · outbound
On-device Streaming Discrete Speech Units On-device Streaming Discrete Speech Units
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd7be342-39df-4441-b9a5-af3cb23e0db8 · outbound
On-device Streaming Discrete Speech Units To evaluate the predicted DSUs, we focus on the discrete ASR system
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b7b9d3b1-8a94-428f-8a22-75b492385509 · outbound
On-device Streaming Discrete Speech Units However, this issue can be mitigated by limiting the future window size of the S2U module
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b195da30-838c-4d25-b2e7-ccdc6181ee2c · outbound
On-device Streaming Discrete Speech Units Since the number of layers has a linear relationship with computational cost, reduc- ing them benefits resource-constrained on-device applications
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9ee2a7ed-9fc7-4c27-98c5-c9e0f55ab937 · outbound
On-device Streaming Discrete Speech Units Also, by applying such, we produce a Pareto optimal curve that represents the trade-off between the computational over- head and the downstream performance
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ca88d02a-26d2-4804-a677-28f69b6e1438 · outbound
On-device Streaming Discrete Speech Units Knowledge distillation (KD) is of- ten used to reduce the size of S3Ms, such as DistilHuBERT
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6a63d6b9-9db6-4ab7-bc41-edc36e66e719 · outbound
On-device Streaming Discrete Speech Units How- ever, current methods for generating DSUs rely on full speech input and computationally heavy S3Ms
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1b6685b7-021a-4d33-ad1e-93762220c499 · outbound
On-device Streaming Discrete Speech Units Unresolved cited work
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 883af9d9-3cfa-4aed-a84b-952631c2eb17 · outbound
On-device Streaming Discrete Speech Units AudioPaLM: A Large Language Model That Can Speak and Listen
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb8e25f6-e0ee-4362-a085-1085116cd7e9 · outbound
On-device Streaming Discrete Speech Units Self-supervised speech representation learning: A review,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d0349d48-e10b-46b9-b2b3-95aa9672037f · outbound
On-device Streaming Discrete Speech Units wav2vec 2.0: A framework for self-supervised learning of speech representa- tions,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b1a38999-104f-4bd6-b490-f4c4186463eb · outbound
On-device Streaming Discrete Speech Units WavLM: Large-scale self- supervised pre-training for full stack speech processing,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a2fefbab-f650-4fac-b5c1-b5778469abb2 · outbound
On-device Streaming Discrete Speech Units Exploration of efficient end- to-end asr using discretized input from self-supervised learning,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8addd65c-ff24-4e71-a487-140226bb057b · outbound
On-device Streaming Discrete Speech Units Exploring speech recognition, translation, and understanding with discrete speech units: A com- parative study,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2659ddbf-a1be-4096-9366-39b9c74fa523 · outbound
On-device Streaming Discrete Speech Units The Interspeech 2024 Challenge on Speech Processing Using Discrete Units,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4a07b900-ead9-493f-b5b7-aeb5743b252e · outbound
On-device Streaming Discrete Speech Units AudioLM: a language modeling approach to audio generation,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e1038a45-d8ee-476a-8752-ae87d5c6a854 · outbound
On-device Streaming Discrete Speech Units SpeechGPT: Empowering large language models with intrinsic cross-modal conversational abili- ties,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 916bc0c0-d73e-4d5f-b63d-516cf1bcf51e · outbound
On-device Streaming Discrete Speech Units PARP: Prune, adjust and re-prune for self-supervised speech recognition,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0d9a634d-c28b-4afb-b631-224e242920ca · outbound
On-device Streaming Discrete Speech Units Discrete speech unit ex- traction via independent component analysis,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d30ae050-9d65-48c5-a60a-99dd72168338 · outbound
On-device Streaming Discrete Speech Units AnyGPT: Unified multimodal LLM with discrete sequence modeling,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 919ad50c-538a-4515-8b99-dd899e645a30 · outbound
On-device Streaming Discrete Speech Units Anonymizing dysarthric speech: Investigating the effects of voice conversion on pathological information preservation,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation abc90dcf-0751-421d-90f5-3809f408c2a0 · outbound
On-device Streaming Discrete Speech Units Self-supervised speech representations are more phonetic than semantic,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cc2f1b35-f9d0-476f-8d31-9ac1e0d3da60 · outbound
On-device Streaming Discrete Speech Units Comparative layer-wise analy- sis of self-supervised speech models,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 85d66725-ab20-470e-abfb-8c5fc79aa576 · outbound
On-device Streaming Discrete Speech Units Leveraging Allophony in Self- Supervised Speech Models for Atypical Pronunciation Assess- ment,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 59c1d868-181c-4fe9-88c8-cba105d943c4 · outbound
On-device Streaming Discrete Speech Units Understanding probe be- haviors through variational bounds of mutual information,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 93cfc809-ae91-446a-be94-c1dee10b9a23 · outbound
On-device Streaming Discrete Speech Units HuBERT: Self- supervised speech representation learning by masked prediction of hidden units,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cad73ed6-e25f-47ee-a5fa-70c4061fbdc0 · outbound
On-device Streaming Discrete Speech Units Streaming automatic speech recog- nition with the transformer model,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c597c76b-7e9a-4baa-819f-0f44d0b3d551 · outbound
On-device Streaming Discrete Speech Units Structured pruning of self- supervised pre-trained models for speech recognition and under- standing,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1c848c50-5e5d-4087-b73b-1cd5c6ed27eb · outbound
On-device Streaming Discrete Speech Units Lib- rispeech: an asr corpus based on public domain audio books,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f2c8a963-77d6-4eea-a8c2-d15d2b146235 · outbound
On-device Streaming Discrete Speech Units Techniques like pruning [18, 19] and quantization [32] is also used
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6836055f-0aac-4321-89eb-00f044bc6033 · outbound
On-device Streaming Discrete Speech Units ML-SUPERB: Multilin- gual Speech Universal PERformance Benchmark,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation df10f0e3-a7f8-4700-bfb4-17008181fa40 · outbound
On-device Streaming Discrete Speech Units calflops: a FLOPs and params calculate tool for neu- ral networks,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d5b472c5-242c-40d1-a0e6-e305f3aa9582 · outbound
On-device Streaming Discrete Speech Units E-branchformer: Branchformer with enhanced merging for speech recognition,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 81c4a3de-5ea5-46e4-90bc-14a2bacc25d9 · outbound
On-device Streaming Discrete Speech Units Attention is all you need,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e2e3b421-4941-4942-99f8-6c13fe1b1fba · outbound
On-device Streaming Discrete Speech Units Decoupled weight decay regulariza- tion,
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f22b8d0-8eaf-49b9-a87a-e16a0c9da93e · outbound
On-device Streaming Discrete Speech Units A time-restricted self- attention layer for asr,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5de08d89-b382-4417-b5a3-8275f762d20c · outbound
On-device Streaming Discrete Speech Units SUPERB: Speech Processing Universal PERformance Benchmark,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4e696492-252e-4194-bdfa-935cdea04107 · outbound
On-device Streaming Discrete Speech Units wav2vec-S: Adapting pre-trained speech models for streaming,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1a26bfc3-6eba-4726-8fad-12891fd93bcf · outbound
On-device Streaming Discrete Speech Units DistilHuBERT: Speech Representation Learning by Layer-wise Distillation of Hidden- unit BERT,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation aeba3496-4a71-412e-a4af-ef70cab8972a · outbound
On-device Streaming Discrete Speech Units FitHuBERT: Going thinner and deeper for knowledge distillation of speech self-supervised learn- ing,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3a370e26-a96a-4a7e-858e-ce90a52634d1 · outbound
On-device Streaming Discrete Speech Units Adaptive compression of supervised and self-supervised models for green speech recognition,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5eceb7eb-0847-4ffc-a6f0-8c450b93715f · outbound
On-device Streaming Discrete Speech Units Soundstream: An end- to-end neural audio codec,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 050abf2a-59eb-45ec-82d8-03ace57cc6e9 · outbound
On-device Streaming Discrete Speech Units High fidelity neural audio compression,
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ea671ed-c414-4979-ac6d-3f853db954b4 · outbound
On-device Streaming Discrete Speech Units ESPnet-Codec: Comprehensive train- ing and evaluation of neural codecs for audio, music, and speech,
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8154b962-78ad-4326-af30-5f2a6dab1602 · outbound
On-device Streaming Discrete Speech Units Neural discrete representa- tion learning,
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6eaa95f1-0deb-4bcf-a909-f9a53f1c2f9b · outbound
On-device Streaming Discrete Speech Units Distilling hubert with lstms via decoupled knowledge distillation,
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 70bdf51d-a647-4d23-a17d-dd84c784539e · outbound
On-device Streaming Discrete Speech Units Knowledge distillation from self- supervised representation learning model with discrete speech units for any-to-any streaming voice conversion,
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d1f994b4-9f6f-420d-be44-2b1665970710 · outbound
On-device Streaming Discrete Speech Units Speechtokenizer: Unified speech tokenizer for speech language models,
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation de6e815b-53b2-4c54-8b57-1be42abe5d1f · outbound
On-device Streaming Discrete Speech Units Moshi: a speech-text foundation model for real-time dialogue
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82907e2b-c766-4208-846a-79b7f705f2a1 · outbound
On-device Streaming Discrete Speech Units We denote various attention window configurations as the number of left, center, and right frames, i.e., [l, c = 1 , r]
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3b3040d7-d133-4d99-99b7-a3186741704a · inbound
On-device Streaming Discrete Speech Units On-device Streaming Discrete Speech Units
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 908b712e-f978-4af9-8ab2-a9e54b3f0ad2 · inbound
WhisperRT -- Turning Whisper into a Causal Streaming Model On-device Streaming Discrete Speech Units
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.