Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:36:46.248344Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 2 inbound Pith citation observations for arXiv:2505.14470.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:36:46.248344Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:36:41.638006Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T12:48:12.260808Z
48 of 48 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 6d1d0618-b4ee-44a9-8e73-da79dbbe5c50 · outbound
PAST: Phonetic-Acoustic Speech Tokenizer These models usually operate over acoustic tokens or phonetic speech tokens (also known as semantic tokens)
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 24be95c4-b49a-460b-b670-fa9d1c2ba676 · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Unresolved cited work
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8e225647-b7d9-4597-9a81-c995d783fc67 · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2073e341-ac4e-420d-a8c5-69758a47bf76 · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3bd6c800-6965-417d-a7ae-ccea8cbdb256 · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fd002efa-0613-4dbf-80fe-94327ad589b8 · outbound
PAST: Phonetic-Acoustic Speech Tokenizer PAST: Phonetic-Acoustic Speech Tokenizer
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a9ed943-e7c3-4768-ad3c-9a6b1bb4955c · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Problem Setup Our model is composed of three main components: Encoder, Quantizer, and Decoder
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 51300dc2-8046-4718-b918-da187b26b91a · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Soundstream: An end-to-end neural audio codec,
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92fd0231-a65c-4b49-8483-1dddd9a223bd · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Data We use all training subsets of LibriSpeech [29] and TIMIT [30] for our training set, yielding a total of965hours of raw audio
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a4d1ca86-875f-44c2-b182-6aeb909450aa · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Baseline Comparison We compare PAST with two baseline hybrid models, Speech- Tokenizer and X-Codec, on both reconstruction and phonetic information metrics
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6e8e1277-a058-4f74-b933-60c5cf533c7b · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c69020ab-9cb6-4e28-ae47-4ad8073d2623 · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Generative spoken language model based on continuous word-sized audio tokens,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 91225a23-a38d-4679-a578-fe8950aea761 · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Textless acoustic model with self-supervised distillation for noise-robust expressive speech-to-speech transla- tion,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2b67d8cf-cf05-4e28-820a-7ef18474dfc2 · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Audiolm: a language modeling approach to au- dio generation,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c14e6a54-8a6b-4e38-ac5d-5e1881a85c0b · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Text-free prosody-aware generative spoken language modeling,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9ae9b94b-4759-4cd1-bf42-0a0da3dece75 · outbound
PAST: Phonetic-Acoustic Speech Tokenizer On generative spoken language modeling from raw audio,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c37f05e4-a93e-4275-bdd8-d56991b118f6 · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Textually pretrained speech language mod- els,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 093fc44b-671b-450c-b501-926022514b33 · outbound
PAST: Phonetic-Acoustic Speech Tokenizer High Fidelity Neural Audio Compression
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ad809f5-5aec-45ed-9b29-cc3bc56237f8 · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Pyramidcodec: Hierarchical codec for long-form music generation in audio domain,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 41a88ea8-e634-44c9-a273-a8d22177ed8b · outbound
PAST: Phonetic-Acoustic Speech Tokenizer wav2vec 2.0: A framework for self-supervised learning of speech repre- sentations,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6552382-832e-4c72-bdb3-c74f93108585 · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Hubert: Self-supervised speech represen- tation learning by masked prediction of hidden units,
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4541b15-f6fc-48f2-a527-dd324dc16ca4 · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Analysing discrete self supervised speech representation for spoken language modeling,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fce9e7a6-b01d-4b23-badc-691f7bf72840 · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Hifi-gan: Generative adversarial net- works for efficient and high fidelity speech synthesis,
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ea204e1-eb42-41a0-a676-61cebaea42d3 · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Speechtok- enizer: Unified speech tokenizer for speech language models,
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65286334-7429-4af7-a48b-ad9146096970 · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Codec Does Matter: Exploring the Semantic Shortcoming of Codec for Audio Language Model
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bbf55102-cca8-4a26-b575-518bf7a68b9d · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Moshi: a speech-text foundation model for real-time dialogue,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 356517eb-c842-450f-85a4-0c0e666dcdab · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Wavlm: Large-scale self-supervised pre-training for full stack speech processing,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 467667de-8c01-4e10-ac1f-b910263e95df · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ae2a569d-ba4b-42a9-a956-bcf254094187 · outbound
PAST: Phonetic-Acoustic Speech Tokenizer HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7c73625-ae79-45d1-9199-36b2bc83944a · outbound
PAST: Phonetic-Acoustic Speech Tokenizer High fidelity neural audio compression,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 722f6378-a58a-4deb-81f4-32f4f5f10418 · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Audiodec: An open-source streaming high- fidelity neural audio codec,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3945ee03-0863-472a-9fb8-29c097240400 · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Funcodec: A funda- mental, reproducible and integrable open-source toolkit for neural speech codec,
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e8865d9-12f2-4ef8-872f-9f91f8d12791 · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Wavtokenizer: an efficient acoustic discrete codec tokenizer for audio language modeling,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 35126764-2f60-4e9d-8e58-84d1b7e60eeb · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Scaling Speech-Text Pre-training with Synthetic Interleaved Data
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a62e195-0f55-469f-a96c-04642d3a81fe · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Robust speech recognition via large-scale weak supervision,
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b62e96f-ebfd-4928-8e66-23090e9c5428 · outbound
PAST: Phonetic-Acoustic Speech Tokenizer LAST: Language Model Aware Speech Tokenization
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f90e1a3-4ff1-4f91-9f46-27076370343f · outbound
PAST: Phonetic-Acoustic Speech Tokenizer NAST: Noise Aware Speech Tokenization for Speech Language Models
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9620259b-327c-4042-8f3c-652536fdf4b1 · outbound
PAST: Phonetic-Acoustic Speech Tokenizer A systematic compar- ison of phonetic aware techniques for speech enhancement,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation cc04afe2-3599-4af3-8073-c2bfb689ab90 · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Con- nectionist temporal classification: labelling unsegmented se- quence data with recurrent neural networks,
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef3f6159-93f0-405c-8250-5d7ade4b97f2 · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Lib- rispeech: An asr corpus based on public domain audio books,
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 839a24ed-933d-4929-8277-66023614f503 · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Timit acoustic-phonetic continuous speech corpus,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 154029ea-b6bf-4d81-9854-c7ae7695950e · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Visqol v3: An open source production ready objec- tive speech and audio metric,
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb65a8c7-94e0-44ae-9425-442ed6a4016d · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Perceptual eval- uation of speech quality (pesq)-a new method for speech quality assessment of telephone networks and codecs,
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01192191-8782-4dca-8efc-a5b90cce4364 · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Evaluating speech features with the minimal- pair abx task: analysis of the classical mfc/plp pipeline,
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1be950b0-83c3-403c-b17f-6ace2183e66e · outbound
PAST: Phonetic-Acoustic Speech Tokenizer DASB - Discrete Audio and Speech Benchmark
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a95b9ff-de8a-497d-8d58-def212839c4e · outbound
PAST: Phonetic-Acoustic Speech Tokenizer AudioGen: Textually Guided Audio Generation
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1909cf34-ad72-475b-b8b3-427cf0d212b2 · outbound
PAST: Phonetic-Acoustic Speech Tokenizer Simple and controllable music gen- eration,
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f5e42fd6-873a-4829-94b9-d8e497fbe07f · outbound
PAST: Phonetic-Acoustic Speech Tokenizer The zero resource speech benchmark 2021: Metrics and baselines for unsupervised spoken language model- ing,
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fd002efa-0613-4dbf-80fe-94327ad589b8 · inbound
PAST: Phonetic-Acoustic Speech Tokenizer PAST: Phonetic-Acoustic Speech Tokenizer
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af391476-d310-4819-81f2-fea65bdf6141 · inbound
Benchmarking Neural Speech Compression from a Rate-Distortion Perspective PAST: Phonetic-Acoustic Speech Tokenizer
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.