Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T12:41:28.171047Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 5 inbound Pith citation observations for arXiv:2509.01336.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T12:41:28.171047Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-03T09:53:16.690620Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
52 of 52 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation 71747265-4239-494d-9b26-0d3a57761db5 · outbound
The AudioMOS Challenge 2025 The V oiceMOS Challenge 2022,
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75c08112-cd56-493a-bcd9-95ed609c0cf4 · outbound
The AudioMOS Challenge 2025 The V oiceMOS Challenge 2023: Zero-Shot Subjective Speech Quality Prediction for Multiple Domains,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 9f15f70c-f59e-4775-97da-fec479570f35 · outbound
The AudioMOS Challenge 2025 The V oiceMOS Challenge 2024: Beyond Speech Quality Prediction,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation bfbc1959-0a0c-4d00-8ae3-09eb3a872dc6 · outbound
The AudioMOS Challenge 2025 How do voices from past speech synthesis challenges compare today?
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 9fee816d-8cca-43a4-89c3-636f263db927 · outbound
The AudioMOS Challenge 2025 Fr ´echet Audio Distance: A Reference-Free Metric for Evaluating Music Enhancement Algorithms,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 7d788b0d-7c87-4918-9c0b-ce177cd587fc · outbound
The AudioMOS Challenge 2025 Make-An-Audio: Text-To-Audio Generation with Prompt-Enhanced Diffusion Models,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 8cf02930-b725-4b01-84d8-8d5c8ef8c35a · outbound
The AudioMOS Challenge 2025 Evaluating generative audio systems and their metrics,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 444c3028-1fea-4c8b-ad67-517eb4e10b73 · outbound
The AudioMOS Challenge 2025 Correlation of Fr ´echet Audio Distance With Human Perception of Environmental Audio Is Embedding Dependent,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation afbd36a3-ac50-4275-8770-a90feaf8348b · outbound
The AudioMOS Challenge 2025 Meta Audiobox Aesthetics: Unified Automatic Quality Assessment for Speech, Music, and Sound
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 076c667e-3f55-40e0-b9fd-ef547b50bbd3 · outbound
The AudioMOS Challenge 2025 MusicEval: A Generative Music Dataset with Expert Ratings for Automatic Text-to-Music Evaluation,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 8a45dfc5-e190-4e62-8d36-03f44096efd7 · outbound
The AudioMOS Challenge 2025 LibriTTS-P: A Corpus with Speaking Style and Speaker Identity Prompts for Text-to-Speech and Style Captioning,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 3c74dd48-e844-46c3-a5f3-b34603c82a54 · outbound
The AudioMOS Challenge 2025 AudioCaps: Generating captions for audios in the wild,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 6257b821-5018-4178-97e7-df8164efd9b9 · outbound
The AudioMOS Challenge 2025 MusicLM: Generating Music From Text
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 094207ec-0676-478b-b7b7-cc10dc9b3636 · outbound
The AudioMOS Challenge 2025 LibriTTS-R: A Restored Multi-Speaker Text-to-Speech Corpus,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 8a3ff8ec-1567-4005-a81d-116bb2d934fc · outbound
The AudioMOS Challenge 2025 Hi-Fi-CAPTAIN: High-fidelity and high-capacity conversational speech synthesis corpus developed by NICT,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 35552078-7a96-431f-94bc-d8529290e268 · outbound
The AudioMOS Challenge 2025 World: a vocoder-based high-quality speech synthesis system for real-time applications,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c59f6e4b-80de-434f-b4ad-d173d91396c7 · outbound
The AudioMOS Challenge 2025 Fast Neural Speech Waveform Generative Models With Fully-Connected Layer-Based Upsampling,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 1716e873-33db-4514-9b46-f6a1c648661a · outbound
The AudioMOS Challenge 2025 Speech masking system based on spatially separated multiple TTS maskers with a compact circular loudspeaker array,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a4f16271-e27a-455a-bb07-aa84247e09b7 · outbound
The AudioMOS Challenge 2025 AudioSR: Versatile audio super-resolution at scale,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e0184f3c-dd9c-48ff-9dc4-9d0c847682df · outbound
The AudioMOS Challenge 2025 pyloudnorm: A simple yet flexible loudness meter in python,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 7056e9ba-8ce3-49e2-a2c5-c4e50924387a · outbound
The AudioMOS Challenge 2025 Large-scale Contrastive Language-Audio Pretraining with Feature Fusion and Keyword-to-Caption Augmentation,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ed998dd8-45e3-4d88-a9e8-13eed8541ef6 · outbound
The AudioMOS Challenge 2025 HTS-AT: A Hierarchical Token-Semantic Audio Transformer for Sound Classification and Detection,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 0a6684ce-f2dc-4ee8-ad7d-d7ddcc3aa8c1 · outbound
The AudioMOS Challenge 2025 Wavlm: Large-scale self-supervised pre- training for full stack speech processing,
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 172d95ca-1a10-4e60-866c-eebf56c26229 · outbound
The AudioMOS Challenge 2025 Attention is all you need,
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17a12fbd-41de-4400-bc07-6938649d4dfb · outbound
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9500da5a-ebd8-4855-b507-6863ad5ae1ad · outbound
The AudioMOS Challenge 2025 Gaussian Error Linear Units (GELUs)
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a227ea61-bb4b-4bb2-ae57-34891ddc50c2 · outbound
The AudioMOS Challenge 2025 Generalization ability of MOS prediction networks,
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b1348ed-832e-45ef-9242-286d6553fe4c · outbound
The AudioMOS Challenge 2025 PAM: Prompting Audio-Language Models for Audio Quality Assessment,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 818ff978-f69d-4b09-9641-8cef1b9b56b2 · outbound
The AudioMOS Challenge 2025 EARS: An Anechoic Fullband Speech Dataset Benchmarked for Speech Enhancement and Dereverberation,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 4cb3aca0-b399-434a-b226-15d309c40397 · outbound
The AudioMOS Challenge 2025 Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26bc3c32-a67e-4203-96cb-347d0b747098 · outbound
The AudioMOS Challenge 2025 BEATs: Audio Pre-Training with Acoustic Tokenizers,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 67cf87db-4185-409f-a2b8-e18bf6b0b21e · outbound
The AudioMOS Challenge 2025 Masked Modeling Duo: Towards a Universal Audio Pre-training Frame- work,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 7eee63f3-1fd0-41d5-aff8-9dea27ca73bf · outbound
The AudioMOS Challenge 2025 High Fidelity Neural Audio Compression,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation fbd19bd8-0b75-4f1e-8863-3654bf650e33 · outbound
The AudioMOS Challenge 2025 Scaling up masked audio encoder learning for general audio classification,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 2c3b9206-cf74-418d-a760-716d816c152c · outbound
The AudioMOS Challenge 2025 MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 9ea7b282-98af-4e89-a446-fe71f12d221b · outbound
The AudioMOS Challenge 2025 MuQ: Self-Supervised Music Representation Learning with Mel Residual Vector Quantization
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1808c671-6225-400b-8cb7-30eecc0d2c3a · outbound
The AudioMOS Challenge 2025 BERT: Pre-training of deep bidirectional transformers for language understanding,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 371f0a5f-9676-4638-98b1-7420db71a21b · outbound
The AudioMOS Challenge 2025 RoBERTa: A Robustly Optimized BERT Pretraining Approach
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 717265ef-56f9-4984-845e-912e8ccf3430 · outbound
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation decdce1d-cc57-4be4-9034-0b74cec2acf4 · outbound
The AudioMOS Challenge 2025 Robust Speech Recognition via Large-Scale Weak Super- vision,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 242f1acb-fb3b-437f-ac5d-02f21b625cce · outbound
The AudioMOS Challenge 2025 HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 44ff0cea-57f8-4d50-ab61-f525cc04c564 · outbound
The AudioMOS Challenge 2025 Scaling Speech Technology to 1,000+ Languages,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 30883e1f-8453-4819-b884-c147b9987d9a · outbound
The AudioMOS Challenge 2025 EAT: Self- Supervised Pre-Training with Efficient Audio Transformer,
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 66f7ec39-c620-4676-9128-c3b5ad480448 · outbound
The AudioMOS Challenge 2025 LDNet: Unified Listener Dependent Modeling in MOS Prediction for Synthetic Speech,
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a147eebb-9fd7-4379-9024-91c8fc04aa03 · outbound
The AudioMOS Challenge 2025 KAN: Kolmogorov–Arnold networks,
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 9688a5e9-93bd-485b-a80f-a456901d0779 · outbound
The AudioMOS Challenge 2025 Transformers are SSMs: Generalized Models and Efficient Algorithms Through Structured State Space Duality,
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 5555346e-949e-49d5-b1bb-a279437c8bdf · outbound
The AudioMOS Challenge 2025 Sampling- Frequency-Independent Convolutional Layer and its Application to Audio Source Separation,
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 68362bec-02a2-44b1-b7d7-fc29fbe516ec · outbound
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a58e257f-295f-4f00-9adf-156dd1676430 · outbound
The AudioMOS Challenge 2025 VERSA: A Versatile Evaluation Toolkit for Speech, Audio, and Music,
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation fbbb3225-c21b-4a2a-afb7-d2c1f8254618 · outbound
The AudioMOS Challenge 2025 XGBoost: A Scalable Tree Boosting Sys- tem,
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 37341450-6e5d-4fa2-9674-b6724494b15c · outbound
The AudioMOS Challenge 2025 Pseudo Label Is Better Than Human Label,
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 635d2866-4a97-49a0-9603-f6d928cb71e9 · outbound
The AudioMOS Challenge 2025 Unresolved cited work
Reference 150
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 08935185-338f-4cbd-a06d-1242a073ac96 · inbound
Song Aesthetics Evaluation with Multi-Stem Attention and Hierarchical Uncertainty Modeling The AudioMOS Challenge 2025
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e76adabc-464e-4244-bdcb-bf4f0b2cb540 · inbound
JASTIN: Aligning LLMs for Zero-Shot Audio and Speech Evaluation via Natural Language Instructions The AudioMOS Challenge 2025
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a724a6a2-505c-4287-aab0-d21125a9daa1 · inbound
Afrispeech Semantics: Evaluating Audio Semantic Reasoning in Spoken Language Models Across Domains and Accents The AudioMOS Challenge 2025
Reference 151
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 92cc6aa3-c70b-470b-8915-a73b1bfc21c0 · inbound
Evaluating SSL and ViViT Architectures for Cross-Corpus Audio MOS Prediction via LODO Validation The AudioMOS Challenge 2025
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e086a207-662d-4702-8c06-a55293b673a3 · inbound
Is One Score Enough? Assessing Singing Quality of Songs with Temporal Score Curves The AudioMOS Challenge 2025
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.