Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T14:27:09.170151Z
Paper Citation Record · LEDGER
As of 12 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 1 inbound Pith citation observation for arXiv:2501.15302.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T14:27:09.170151Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T16:13:20.213020Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-06T16:13:20.931322Z
35 of 35 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation d5742c54-78ce-439f-acd9-f4858c731762 · outbound
The ICME 2025 Audio Encoder Capability Challenge Neural discrete representation learning,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 483913b3-fe05-46f2-a617-e6a317fdf08d · outbound
The ICME 2025 Audio Encoder Capability Challenge Finite Scalar Quantization: VQ-VAE Made Simple
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b299c320-1e63-46ff-8682-f3701a879f4d · outbound
The ICME 2025 Audio Encoder Capability Challenge High-fidelity audio compression with improved rvqgan,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 1c30ab33-fb6e-44df-9604-748d76e9988f · outbound
The ICME 2025 Audio Encoder Capability Challenge SNAC: Multi-Scale Neural Audio Codec
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ecdc0d7a-ce59-4dab-8e7a-cf9320779054 · outbound
The ICME 2025 Audio Encoder Capability Challenge Moshi: a speech-text foundation model for real-time dialogue
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf075be4-d900-495b-b85f-80164947a406 · outbound
The ICME 2025 Audio Encoder Capability Challenge A Comparative Study of Discrete Speech Tokens for Semantic-Related Tasks with Large Language Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc7dbac0-f5eb-4f43-a2e6-060590170c4a · outbound
The ICME 2025 Audio Encoder Capability Challenge Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4297e96e-970e-4e22-817e-13e50c59578a · outbound
The ICME 2025 Audio Encoder Capability Challenge Mini-Omni: Language Models Can Hear, Talk While Thinking in Streaming
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14e6393b-9283-46b5-b13f-bf5f6b465aad · outbound
The ICME 2025 Audio Encoder Capability Challenge SALMONN-omni: A Codec-free LLM for Full-duplex Speech Understanding and Generation
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea4c6619-997a-4f60-90dc-9ae8e2f9b9c8 · outbound
The ICME 2025 Audio Encoder Capability Challenge wav2vec 2.0: A fr amework for self-supervised learning of speech representations,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c48c1a4a-613b-485f-b7d1-b28ec87a461e · outbound
The ICME 2025 Audio Encoder Capability Challenge Data2 vec: A general framework for self- supervised learning in speech, vision and language,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 6d92c814-716c-4c51-9d5c-ee968146577a · outbound
The ICME 2025 Audio Encoder Capability Challenge Sc aling up masked audio encoder learning for general audio classification,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b609f98e-8786-4cfa-9002-31e146495105 · outbound
The ICME 2025 Audio Encoder Capability Challenge HEAR: Holistic evaluation of audio representations,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f699b613-476e-400b-81ac-938993fe15d6 · outbound
The ICME 2025 Audio Encoder Capability Challenge SUPERB: Speech processing universal performance benchmar k,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation eac09b7b-2251-4600-8aea-229613780c1a · outbound
The ICME 2025 Audio Encoder Capability Challenge DASB - Discrete Audio and Speech Benchmark
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c95c848-6575-4f2f-b5a5-d8a8e783cf0f · outbound
The ICME 2025 Audio Encoder Capability Challenge Speech Commands: A Dataset for Limited-Vocabulary Speech Recognition
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70681049-0a50-4aa7-bd9c-6ebbec8e85f1 · outbound
The ICME 2025 Audio Encoder Capability Challenge Lib ricount, a dataset for speaker count estima- tion,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 60ec8637-d5dc-464a-829a-f0eb4faa7e4b · outbound
The ICME 2025 Audio Encoder Capability Challenge Voxlingua107: a dataset for spoken lan guage recognition,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d4049d2a-85e2-4917-b84b-02feaa2f07a7 · outbound
The ICME 2025 Audio Encoder Capability Challenge Voxceleb: L arge-scale speaker verification in the wild,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 4bfa7132-234c-4247-bc62-793d83ec73ea · outbound
The ICME 2025 Audio Encoder Capability Challenge Librisp eech: an asr corpus based on public domain audio books,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 08bbda33-2a61-404b-9116-d7ce1ba5c64b · outbound
The ICME 2025 Audio Encoder Capability Challenge Speech Model Pre-training for End-to-End Spoken Language Understanding
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a89e309-af28-4ec2-b079-a918ad385d20 · outbound
The ICME 2025 Audio Encoder Capability Challenge Vocalsound: A dataset for impro ving human vocal sounds recognition,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a1f097c8-add5-43fd-81b5-2c4f1967fcb2 · outbound
The ICME 2025 Audio Encoder Capability Challenge Crema-d: Crowd- sourced emotional multimodal actors dataset,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b199aac7-ee97-4346-ac42-95719c2825ad · outbound
The ICME 2025 Audio Encoder Capability Challenge spee- chocean762: An open-source non-native english speech corpus f or pronunciation assessment,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f72d2a44-6b37-4c5e-a149-b16d19d43ef7 · outbound
The ICME 2025 Audio Encoder Capability Challenge Autom atic speaker verification spoofing and countermeasures challenge (asvspoof 2015) database,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 60ee67fa-45d4-45b6-b882-cf47dccb327d · outbound
The ICME 2025 Audio Encoder Capability Challenge Esc: Dataset for environmental sound classific ation,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 647aa645-ec38-4b94-a432-7f7004b675eb · outbound
The ICME 2025 Audio Encoder Capability Challenge Fsd 50k: an open dataset of human-labeled sound events,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 734b0a95-283c-4511-8f8e-c31e0a3e67a5 · outbound
The ICME 2025 Audio Encoder Capability Challenge A dataset and taxonom y for urban sound research,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 0a8327d8-ad6c-4706-b9b9-af6b735cb42c · outbound
The ICME 2025 Audio Encoder Capability Challenge Sound eve nt detection in domestic environments with weakly labeled data and soundscape synthesis,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation e0152afd-8583-423a-bc21-81ae392ac2b9 · outbound
The ICME 2025 Audio Encoder Capability Challenge General-purpose Tagging of Freesound Audio with AudioSet Labels: Task Description, Dataset, and Baseline
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4dd00642-a5e3-4629-bdde-1c92b4c6867d · outbound
The ICME 2025 Audio Encoder Capability Challenge Clotho: An audio capt ioning dataset,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 99504da6-da4d-46c3-a1b1-a46ae6aebb4f · outbound
The ICME 2025 Audio Encoder Capability Challenge Enabling factorized piano music modeling and generation with the MAESTRO dataset,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c7b1239c-9c2c-4c2b-b75e-5b5ea524e3e7 · outbound
The ICME 2025 Audio Encoder Capability Challenge The GTZAN dataset: Its contents, its faults, their effects on evaluation, and its future use
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28403eb4-d70b-49f0-9d15-1e0ec430e1a5 · outbound
The ICME 2025 Audio Encoder Capability Challenge Neural audio synthesis of musical notes with wavenet autoencoders,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f443ba02-0d14-4007-91a0-ad1005d83722 · outbound
The ICME 2025 Audio Encoder Capability Challenge FMA: A Dataset For Music Analysis
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee2596fe-f80a-466a-af9c-fdd960cdf5e5 · inbound
OpenBEATs: A Fully Open-Source General-Purpose Audio Encoder The ICME 2025 Audio Encoder Capability Challenge
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.