Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T15:00:14.179533Z
Paper Citation Record · LEDGER
As of 6 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2508.20660.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T15:00:14.179533Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
28 of 28 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation e9045120-c6bd-4e02-a8ef-0072e4086007 · outbound
CodecBench: A Comprehensive Benchmark for Acoustic and Semantic Evaluation Sdr–half-baked or well done? In ICASSP 2019-2019 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), pp
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0c97b505-afe9-45f2-9ef1-596b508117c7 · outbound
CodecBench: A Comprehensive Benchmark for Acoustic and Semantic Evaluation Librispeech: an asr corpus based on public domain audio books
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 8f3a2d7c-3c5c-4a82-ad09-b6f4e4ee3e4e · outbound
CodecBench: A Comprehensive Benchmark for Acoustic and Semantic Evaluation MELD: A Multimodal Multi-Party Dataset for Emotion Recognition in Conversations
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49e0a2e3-dd50-461a-a6b3-fc6499b242ff · outbound
CodecBench: A Comprehensive Benchmark for Acoustic and Semantic Evaluation A short-time objective intelligibility measure for time-frequency weighted noisy speech
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 92059eaf-e1d3-4fd6-a32d-58e5de8d0d8c · outbound
CodecBench: A Comprehensive Benchmark for Acoustic and Semantic Evaluation Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd815554-0fb7-4877-a0b9-f2406c475e7b · outbound
CodecBench: A Comprehensive Benchmark for Acoustic and Semantic Evaluation MaskGCT: Zero-Shot Text-to-Speech with Masked Generative Codec Transformer
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 379d0d46-c6e6-48b0-b389-e657e0d22e09 · outbound
CodecBench: A Comprehensive Benchmark for Acoustic and Semantic Evaluation FlowDec: A flow-based full-band general audio codec with high perceptual quality
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd009f05-b7c0-49a8-bca3-cdaa2670d4a1 · outbound
CodecBench: A Comprehensive Benchmark for Acoustic and Semantic Evaluation Codec-SUPERB: An In-Depth Analysis of Sound Codec Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef9e0fda-ef88-4633-b2cc-8b8da6e8ec72 · outbound
CodecBench: A Comprehensive Benchmark for Acoustic and Semantic Evaluation Laughter Synthesis using Pseudo Phonetic Tokens with a Large-scale In-the-wild Laughter Corpus
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d06df0a6-9d79-4011-803e-22139be9f85e · outbound
CodecBench: A Comprehensive Benchmark for Acoustic and Semantic Evaluation BigCodec: Pushing the Limits of Low-Bitrate Neural Speech Codec
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ccf870b0-a18d-4c93-abe1-94eff0dcee0c · outbound
CodecBench: A Comprehensive Benchmark for Acoustic and Semantic Evaluation Qwen2.5 Technical Report
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31ade6f1-8007-4a2f-84af-848fe68826cc · outbound
CodecBench: A Comprehensive Benchmark for Acoustic and Semantic Evaluation Llasa: Scaling Train-Time and Inference-Time Compute for Llama-based Speech Synthesis
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c53f719a-b5e5-4271-b8d6-e98932066d8e · outbound
CodecBench: A Comprehensive Benchmark for Acoustic and Semantic Evaluation SpeechGPT: Empowering Large Language Models with Intrinsic Cross-Modal Conversational Abilities
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93203482-b7d1-49f0-8ed7-7fd2812cbe20 · outbound
CodecBench: A Comprehensive Benchmark for Acoustic and Semantic Evaluation SpeechTokenizer: Unified Speech Tokenizer for Speech Large Language Models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e2992c9-eb93-4c69-8f52-ea94bc1765ae · outbound
CodecBench: A Comprehensive Benchmark for Acoustic and Semantic Evaluation The clips within this dataset are manually selected from public field recordings compiled by the Freesound.org project
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5445decb-3782-43f6-ab4d-934660afc49d · outbound
CodecBench: A Comprehensive Benchmark for Acoustic and Semantic Evaluation Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2fb3cf35-e48e-468f-83b3-8bc50d89a366 · outbound
CodecBench: A Comprehensive Benchmark for Acoustic and Semantic Evaluation VERSA: A Versatile Evaluation Toolkit for Speech, Audio, and Music
Reference 2001
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71afe80c-116d-42cc-b965-6cffa81d88dc · outbound
CodecBench: A Comprehensive Benchmark for Acoustic and Semantic Evaluation Audio set: An ontology and human-labeled dataset for audio events
Reference 2005
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e4178831-f504-4363-8fbf-12f1d8439c64 · outbound
CodecBench: A Comprehensive Benchmark for Acoustic and Semantic Evaluation Visqol v3: An open source production ready objective speech and audio metric
Reference 2014
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a1fb8b1a-4fd4-46d0-a271-3b7e7aeec344 · outbound
CodecBench: A Comprehensive Benchmark for Acoustic and Semantic Evaluation Scaling Transformers for Low-Bitrate High-Quality Speech Coding
Reference 2015
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3190eddb-5666-4dde-8754-2cc8bf56d42b · outbound
CodecBench: A Comprehensive Benchmark for Acoustic and Semantic Evaluation 1632–1636,
Reference 2016
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ef66f28f-a91b-48ba-9091-de9ec5c49eb6 · outbound
CodecBench: A Comprehensive Benchmark for Acoustic and Semantic Evaluation V ocalsound: A dataset for improving human vocal sounds recognition
Reference 2017
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 40b1610e-2ed8-40ec-b53c-be989790466a · outbound
CodecBench: A Comprehensive Benchmark for Acoustic and Semantic Evaluation Baichuan-Audio: A Unified Framework for End-to-End Speech Interaction
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9db2a1b1-70dc-40ce-a056-72781458b97b · outbound
CodecBench: A Comprehensive Benchmark for Acoustic and Semantic Evaluation LibriMix: An Open-Source Dataset for Generalizable Speech Separation
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69c6417a-e655-405f-ad0b-77ab70ec46da · outbound
CodecBench: A Comprehensive Benchmark for Acoustic and Semantic Evaluation Robust Speech Recognition via Large-Scale Weak Supervision
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05f7c0ff-f8eb-47aa-8c09-462361b66ef4 · outbound
CodecBench: A Comprehensive Benchmark for Acoustic and Semantic Evaluation Benchmarking representations for speech, music, and acoustic events
Reference 2023
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d8fd45dc-73ae-4d33-b7fc-36a93b950a4c · outbound
CodecBench: A Comprehensive Benchmark for Acoustic and Semantic Evaluation Moshi: a speech-text foundation model for real-time dialogue
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2a2d476-99a5-47f2-b5ce-84675ebd4f29 · outbound
CodecBench: A Comprehensive Benchmark for Acoustic and Semantic Evaluation Clotho- aqa: A crowdsourced dataset for audio question answering
Reference 2025
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
No inbound Pith citation observations are available.