Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:43:07.274864Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 1 inbound Pith citation observation for arXiv:2505.17076.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:43:07.274864Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-03T10:13:57.724078Z
A source-named dated measurement, never combined with another source.
Source: cited_works
22 of 22 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 21a846d2-cdd2-4376-9909-a4cf5715e0bd · outbound
Impact of Frame Rates on Speech Tokenizer: A Case Study on Mandarin and English Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98c1dfae-e363-4fb1-98d3-790361e9d7d3 · outbound
Impact of Frame Rates on Speech Tokenizer: A Case Study on Mandarin and English CosyVoice: A Scalable Multilingual Zero-shot Text-to-speech Synthesizer based on Supervised Semantic Tokens
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79864f0f-c893-49f5-802e-3d94cdd45593 · outbound
Impact of Frame Rates on Speech Tokenizer: A Case Study on Mandarin and English Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d402c68d-3c50-4cf1-a39d-17d51ce622ae · outbound
Impact of Frame Rates on Speech Tokenizer: A Case Study on Mandarin and English Moshi: a speech-text foundation model for real-time dialogue
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9344925-fe3d-47f6-9de6-681c54e09ec3 · outbound
Impact of Frame Rates on Speech Tokenizer: A Case Study on Mandarin and English Not All Votes Count! Programs as Verifiers Improve Self-Consistency of Language Models for Math Reasoning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 140e13e9-533c-4fda-9fae-710e5bec802b · outbound
Impact of Frame Rates on Speech Tokenizer: A Case Study on Mandarin and English GenSE: A Series of Audio Generation Models,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 44c14ae2-aab0-4e07-b60d-3d12d1b66ed7 · outbound
Impact of Frame Rates on Speech Tokenizer: A Case Study on Mandarin and English BEATs: Audio Pre-Training with Acoustic Tokenizers
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4435bc8-5bbf-4d90-8269-6431442dd77c · outbound
Impact of Frame Rates on Speech Tokenizer: A Case Study on Mandarin and English SpeechTokenizer: Unified Speech Tokenizer for Speech Large Language Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78e6ee6c-90dc-4461-9a1c-c0fdbf86a18d · outbound
Impact of Frame Rates on Speech Tokenizer: A Case Study on Mandarin and English WavTokenizer: an Efficient Acoustic Discrete Codec Tokenizer for Audio Language Modeling
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56e0ad82-5987-481e-82ac-daea9cbf5b87 · outbound
Impact of Frame Rates on Speech Tokenizer: A Case Study on Mandarin and English SMTL: A Stratified Logic for Expressive Multi-Level Temporal Specifications
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 89177396-e7e2-43c4-932e-ec319e1bbae8 · outbound
Impact of Frame Rates on Speech Tokenizer: A Case Study on Mandarin and English High Fidelity Neural Audio Compression
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd29d5e4-fd7e-42a2-b6cb-6238da7c5199 · outbound
Impact of Frame Rates on Speech Tokenizer: A Case Study on Mandarin and English SoundStream: An End-to-End Neural Audio Codec,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 28d695fd-43d7-46ff-8a3d-090e9a4fe551 · outbound
Impact of Frame Rates on Speech Tokenizer: A Case Study on Mandarin and English Attention is all you need,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 77b97ef8-a72e-4b57-bb9d-3e6fa480b3ad · outbound
Impact of Frame Rates on Speech Tokenizer: A Case Study on Mandarin and English Neural discrete representation learning,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bd8c011f-56fb-40f6-8a36-6a36de94f288 · outbound
Impact of Frame Rates on Speech Tokenizer: A Case Study on Mandarin and English Connectionist temporal classification: labelling unseg- mented sequence data with recurrent neural networks,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7d4b39c0-8e81-496f-92b7-eccef56ee622 · outbound
Impact of Frame Rates on Speech Tokenizer: A Case Study on Mandarin and English AISHELL-2: Transforming Mandarin ASR Research Into Industrial Scale
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2eb5f9e9-09aa-4949-8b99-a6e9bb932388 · outbound
Impact of Frame Rates on Speech Tokenizer: A Case Study on Mandarin and English Librispeech: an ASR corpus based on public do- main audio books,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d8aafd0a-cd3c-4a4b-85cd-d208cb1617ae · outbound
Impact of Frame Rates on Speech Tokenizer: A Case Study on Mandarin and English Robust speech recognition via large-scale weak supervision,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ba1d2e57-1294-46e5-82ce-83a13c71c556 · outbound
Impact of Frame Rates on Speech Tokenizer: A Case Study on Mandarin and English StepAudio: A framework for pre-trained au- dio models training and reasoning,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b997a9c4-edb4-44f9-80d3-942ff55a5ebb · outbound
Impact of Frame Rates on Speech Tokenizer: A Case Study on Mandarin and English Characterizing Polkadot's Transactions Ecosystem: methodology, tools, and insights
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 33c999dd-1326-467f-86e1-1f932a255f1e · outbound
Impact of Frame Rates on Speech Tokenizer: A Case Study on Mandarin and English A three-layered model for expressive speech perception,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 63460d8c-dfcf-4b75-813b-993897525a1b · outbound
Impact of Frame Rates on Speech Tokenizer: A Case Study on Mandarin and English Tone recognition in Mandarin Chinese using convolutional neural networks,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d6e5ce31-033b-4fde-86cd-ac72e05ad0a5 · inbound
ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition Impact of Frame Rates on Speech Tokenizer: A Case Study on Mandarin and English
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.