Pith. sign in

Paper Citation Record · LEDGER

The ICME 2025 Audio Encoder Capability Challenge

As of 12 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 1 inbound Pith citation observation for arXiv:2501.15302.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.15302 v1

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T14:27:09.170151Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T16:13:20.213020Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T16:13:20.931322Z

Reference resolution

35 of 35 outbound references displayed

  • verified exact0
  • verified fuzzy22
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d5742c54-78ce-439f-acd9-f4858c731762 · outbound

This paper cites Neural discrete representation learning,.

The ICME 2025 Audio Encoder Capability Challenge Neural discrete representation learning,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.727024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:27:09.007976Z digest=sha256:9b15ae7b643e2bfe90f0b8e915434cfb97420dd0fa5db239d27e57888b839ea0

Observation 483913b3-fe05-46f2-a617-e6a317fdf08d · outbound

This paper cites Finite Scalar Quantization: VQ-VAE Made Simple.

The ICME 2025 Audio Encoder Capability Challenge Finite Scalar Quantization: VQ-VAE Made Simple

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T14:27:09.014071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:27:09.014071Z digest=sha256:dcd20bf6d28333a77d3be3bdadd2faabc45b4d6a0b93062f645d4fc350a05e4a

Observation b299c320-1e63-46ff-8682-f3701a879f4d · outbound

This paper cites High-fidelity audio compression with improved rvqgan,.

The ICME 2025 Audio Encoder Capability Challenge High-fidelity audio compression with improved rvqgan,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.712327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:27:09.019677Z digest=sha256:1c8bfd825405eea94e050fb2b4586ce99a20c4bede3c458109cfda0c6b000b31

Observation 1c30ab33-fb6e-44df-9604-748d76e9988f · outbound

This paper cites SNAC: Multi-Scale Neural Audio Codec.

The ICME 2025 Audio Encoder Capability Challenge SNAC: Multi-Scale Neural Audio Codec

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T14:27:09.024591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:27:09.024591Z digest=sha256:c35e2c4bbaec481c3cfe21191cb115060dc35af2502a8e87c41158e68a771b59

Observation ecdc0d7a-ce59-4dab-8e7a-cf9320779054 · outbound

This paper cites Moshi: a speech-text foundation model for real-time dialogue.

The ICME 2025 Audio Encoder Capability Challenge Moshi: a speech-text foundation model for real-time dialogue

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T14:27:09.030098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:27:09.030098Z digest=sha256:d04b64436f61fc4a5fc2aecbd9bcf2a51373d5f854462070f77944255079c07b

Observation cf075be4-d900-495b-b85f-80164947a406 · outbound

This paper cites A Comparative Study of Discrete Speech Tokens for Semantic-Related Tasks with Large Language Models.

The ICME 2025 Audio Encoder Capability Challenge A Comparative Study of Discrete Speech Tokens for Semantic-Related Tasks with Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T14:27:09.035252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:27:09.035252Z digest=sha256:8339ffc2e5898bc43c014537c6d0ab34354b7e4a453aa7b9b9e1fd7e931edce8

Observation dc7dbac0-f5eb-4f43-a2e6-060590170c4a · outbound

This paper cites Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models.

The ICME 2025 Audio Encoder Capability Challenge Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T14:27:09.041296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:27:09.041296Z digest=sha256:5351badb722fa06401ec7cfe38a701308710978d21c6f1b536bd297f17ce7b86

Observation 4297e96e-970e-4e22-817e-13e50c59578a · outbound

This paper cites Mini-Omni: Language Models Can Hear, Talk While Thinking in Streaming.

The ICME 2025 Audio Encoder Capability Challenge Mini-Omni: Language Models Can Hear, Talk While Thinking in Streaming

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T14:27:09.046367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:27:09.046367Z digest=sha256:4bde0c01e29c9bc79f270e237872b1c0d6b41de4fed2265302f5c98357d66bd8

Observation 14e6393b-9283-46b5-b13f-bf5f6b465aad · outbound

This paper cites SALMONN-omni: A Codec-free LLM for Full-duplex Speech Understanding and Generation.

The ICME 2025 Audio Encoder Capability Challenge SALMONN-omni: A Codec-free LLM for Full-duplex Speech Understanding and Generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T14:27:09.051470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:27:09.051470Z digest=sha256:bb132124a84d28663480775b140e0afadb3a5094e23d882a3e3c0d1ca1eb5a2a

Observation ea4c6619-997a-4f60-90dc-9ae8e2f9b9c8 · outbound

This paper cites wav2vec 2.0: A fr amework for self-supervised learning of speech representations,.

The ICME 2025 Audio Encoder Capability Challenge wav2vec 2.0: A fr amework for self-supervised learning of speech representations,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.698860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:27:09.056593Z digest=sha256:57028c50773bc5b7e54ecf38650bb69c3a194d567a6a072171e6aedc18ec4252

Observation c48c1a4a-613b-485f-b7d1-b28ec87a461e · outbound

This paper cites Data2 vec: A general framework for self- supervised learning in speech, vision and language,.

The ICME 2025 Audio Encoder Capability Challenge Data2 vec: A general framework for self- supervised learning in speech, vision and language,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.685160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:27:09.061310Z digest=sha256:0f6f5e6d88a3bf9fa1a0a0f652c25d24ec45a96696cf6cbda72234724562166c

Observation 6d92c814-716c-4c51-9d5c-ee968146577a · outbound

This paper cites Sc aling up masked audio encoder learning for general audio classification,.

The ICME 2025 Audio Encoder Capability Challenge Sc aling up masked audio encoder learning for general audio classification,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.670843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:27:09.066268Z digest=sha256:3fe69dede6352ad813ec4b313bc557dd12560248276d2b3778a6010c9faaecd0

Observation b609f98e-8786-4cfa-9002-31e146495105 · outbound

This paper cites HEAR: Holistic evaluation of audio representations,.

The ICME 2025 Audio Encoder Capability Challenge HEAR: Holistic evaluation of audio representations,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.656081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:27:09.071436Z digest=sha256:610188ebf31aba7c770dd2a67be5521dde78ecd8bc83f9994e05d2dc87b06a0f

Observation f699b613-476e-400b-81ac-938993fe15d6 · outbound

This paper cites SUPERB: Speech processing universal performance benchmar k,.

The ICME 2025 Audio Encoder Capability Challenge SUPERB: Speech processing universal performance benchmar k,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.641396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:27:09.076302Z digest=sha256:cabaa45b521f6279627a3518654df431db40e940ed0d41e1b2f6e7ee68fc2d76

Observation eac09b7b-2251-4600-8aea-229613780c1a · outbound

This paper cites DASB - Discrete Audio and Speech Benchmark.

The ICME 2025 Audio Encoder Capability Challenge DASB - Discrete Audio and Speech Benchmark

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T14:27:09.081166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:27:09.081166Z digest=sha256:4eb6836eb0435c1527b9e3ef195a453793da77c2d8a88be5aa89903e388b7975

Observation 2c95c848-6575-4f2f-b5a5-d8a8e783cf0f · outbound

This paper cites Speech Commands: A Dataset for Limited-Vocabulary Speech Recognition.

The ICME 2025 Audio Encoder Capability Challenge Speech Commands: A Dataset for Limited-Vocabulary Speech Recognition

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T14:27:09.086029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:27:09.086029Z digest=sha256:5085a83fb58527b9f01ac8970d594398bdc37ebd67d2619c46a89e8292e0d9c2

Observation 70681049-0a50-4aa7-bd9c-6ebbec8e85f1 · outbound

This paper cites Lib ricount, a dataset for speaker count estima- tion,.

The ICME 2025 Audio Encoder Capability Challenge Lib ricount, a dataset for speaker count estima- tion,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.626585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:27:09.090811Z digest=sha256:c01fec9aa7ab3963edd26a69a8b9101e1ada82beb4016657002700bcef20298f

Observation 60ec8637-d5dc-464a-829a-f0eb4faa7e4b · outbound

This paper cites Voxlingua107: a dataset for spoken lan guage recognition,.

The ICME 2025 Audio Encoder Capability Challenge Voxlingua107: a dataset for spoken lan guage recognition,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.611650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:27:09.094633Z digest=sha256:4665d77ed1c427ea515769987d37ab57f5525a44508ce15c62129afc4b0fa0c8

Observation d4049d2a-85e2-4917-b84b-02feaa2f07a7 · outbound

This paper cites Voxceleb: L arge-scale speaker verification in the wild,.

The ICME 2025 Audio Encoder Capability Challenge Voxceleb: L arge-scale speaker verification in the wild,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.596567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:27:09.098417Z digest=sha256:d15cefe868af45f6918a99ee547baf0cbcf28d369dd6c08b45ab3ee04996ec32

Observation 4bfa7132-234c-4247-bc62-793d83ec73ea · outbound

This paper cites Librisp eech: an asr corpus based on public domain audio books,.

The ICME 2025 Audio Encoder Capability Challenge Librisp eech: an asr corpus based on public domain audio books,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.581761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:27:09.102379Z digest=sha256:29f9cdd87462995d4d92c1d51edbf209df8a7a5138b8a57d660b05917a84b86f

Observation 08bbda33-2a61-404b-9116-d7ce1ba5c64b · outbound

This paper cites Speech Model Pre-training for End-to-End Spoken Language Understanding.

The ICME 2025 Audio Encoder Capability Challenge Speech Model Pre-training for End-to-End Spoken Language Understanding

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T14:27:09.106687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:27:09.106687Z digest=sha256:fb99c2a73d458ffb6d36d65e43fef32bb4a3f034046cb1505af39619da320f0a

Observation 9a89e309-af28-4ec2-b079-a918ad385d20 · outbound

This paper cites Vocalsound: A dataset for impro ving human vocal sounds recognition,.

The ICME 2025 Audio Encoder Capability Challenge Vocalsound: A dataset for impro ving human vocal sounds recognition,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.567012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:27:09.111099Z digest=sha256:fee87cda5821edc7596640fff684e9c389a380bdc77c457918417acc3b3325f0

Observation a1f097c8-add5-43fd-81b5-2c4f1967fcb2 · outbound

This paper cites Crema-d: Crowd- sourced emotional multimodal actors dataset,.

The ICME 2025 Audio Encoder Capability Challenge Crema-d: Crowd- sourced emotional multimodal actors dataset,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.551485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:27:09.115119Z digest=sha256:580c7239052edf04e51d6309fc337406d25641d0d56dac7c06447786e788f43f

Observation b199aac7-ee97-4346-ac42-95719c2825ad · outbound

This paper cites spee- chocean762: An open-source non-native english speech corpus f or pronunciation assessment,.

The ICME 2025 Audio Encoder Capability Challenge spee- chocean762: An open-source non-native english speech corpus f or pronunciation assessment,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.535715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:27:09.118950Z digest=sha256:7d69841a2a4ba9bcdaaed87cddb5fd89fd9b4f4cf35d2de2fdc9d78a5e9f847f

Observation f72d2a44-6b37-4c5e-a149-b16d19d43ef7 · outbound

This paper cites Autom atic speaker verification spoofing and countermeasures challenge (asvspoof 2015) database,.

The ICME 2025 Audio Encoder Capability Challenge Autom atic speaker verification spoofing and countermeasures challenge (asvspoof 2015) database,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.521154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:27:09.123480Z digest=sha256:05a981c46fc8671825770f8cdbfaae046edd3226bcd881ce8aaf5692abc5252d

Observation 60ee67fa-45d4-45b6-b882-cf47dccb327d · outbound

This paper cites Esc: Dataset for environmental sound classific ation,.

The ICME 2025 Audio Encoder Capability Challenge Esc: Dataset for environmental sound classific ation,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.506592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:27:09.128144Z digest=sha256:e7daadc4eb3048350681f8764baf7423314f6c5d3e8d0f43cfe1e788b367dd83

Observation 647aa645-ec38-4b94-a432-7f7004b675eb · outbound

This paper cites Fsd 50k: an open dataset of human-labeled sound events,.

The ICME 2025 Audio Encoder Capability Challenge Fsd 50k: an open dataset of human-labeled sound events,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.491604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:27:09.132952Z digest=sha256:07bb58f1b79f31b74d1f6021305e8511e99bb1eea62fd9abfd4786d95c7611a5

Observation 734b0a95-283c-4511-8f8e-c31e0a3e67a5 · outbound

This paper cites A dataset and taxonom y for urban sound research,.

The ICME 2025 Audio Encoder Capability Challenge A dataset and taxonom y for urban sound research,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.475397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:27:09.137432Z digest=sha256:cf490c9997855d16f43ead56327c5e5b86b3122dce5533498e60b4e79f716e73

Observation 0a8327d8-ad6c-4706-b9b9-af6b735cb42c · outbound

This paper cites Sound eve nt detection in domestic environments with weakly labeled data and soundscape synthesis,.

The ICME 2025 Audio Encoder Capability Challenge Sound eve nt detection in domestic environments with weakly labeled data and soundscape synthesis,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.459336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:27:09.142173Z digest=sha256:2ac41eb3e8fbdce229f4027f7599c8380af2ca80b4aeb6a33eb45671c7f77f49

Observation e0152afd-8583-423a-bc21-81ae392ac2b9 · outbound

This paper cites General-purpose Tagging of Freesound Audio with AudioSet Labels: Task Description, Dataset, and Baseline.

The ICME 2025 Audio Encoder Capability Challenge General-purpose Tagging of Freesound Audio with AudioSet Labels: Task Description, Dataset, and Baseline

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T14:27:09.146676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:27:09.146676Z digest=sha256:c80b559247c879911745472ef77eb21e5779a98b9bd88553b2bfe43e636a16e3

Observation 4dd00642-a5e3-4629-bdde-1c92b4c6867d · outbound

This paper cites Clotho: An audio capt ioning dataset,.

The ICME 2025 Audio Encoder Capability Challenge Clotho: An audio capt ioning dataset,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.443976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:27:09.151378Z digest=sha256:4f340cbd98bc87fdb2757654eccf4f2e379e70d67de87fdfbbe22aa27cc7eac9

Observation 99504da6-da4d-46c3-a1b1-a46ae6aebb4f · outbound

This paper cites Enabling factorized piano music modeling and generation with the MAESTRO dataset,.

The ICME 2025 Audio Encoder Capability Challenge Enabling factorized piano music modeling and generation with the MAESTRO dataset,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.427914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:27:09.155999Z digest=sha256:fc74503bbe6cc50bc54cdd7f973d52d089f6d65f4a98ae785582e351eee49011

Observation c7b1239c-9c2c-4c2b-b75e-5b5ea524e3e7 · outbound

This paper cites The GTZAN dataset: Its contents, its faults, their effects on evaluation, and its future use.

The ICME 2025 Audio Encoder Capability Challenge The GTZAN dataset: Its contents, its faults, their effects on evaluation, and its future use

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T14:27:09.160670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:27:09.160670Z digest=sha256:96a59f874224e70aca9d386680678fc0c87d3e8604187fcfc92da6fedadcc7e5

Observation 28403eb4-d70b-49f0-9d15-1e0ec430e1a5 · outbound

This paper cites Neural audio synthesis of musical notes with wavenet autoencoders,.

The ICME 2025 Audio Encoder Capability Challenge Neural audio synthesis of musical notes with wavenet autoencoders,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.411329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:27:09.165586Z digest=sha256:69a9058e6a2ae6733dec16070ea6c8046bf15fc0104bb3e01182ade1adc2dda9

Observation f443ba02-0d14-4007-91a0-ad1005d83722 · outbound

This paper cites FMA: A Dataset For Music Analysis.

The ICME 2025 Audio Encoder Capability Challenge FMA: A Dataset For Music Analysis

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T14:27:09.170151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:27:09.170151Z digest=sha256:0aa00dc2fcc49d0dfa597dc4a6d26e7d07cedcad19db850af2936b6114346ece

Pith citing papers

Observation ee2596fe-f80a-466a-af9c-fdd960cdf5e5 · inbound

OpenBEATs: A Fully Open-Source General-Purpose Audio Encoder cites this paper.

OpenBEATs: A Fully Open-Source General-Purpose Audio Encoder The ICME 2025 Audio Encoder Capability Challenge

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-08-06T16:13:20.989094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T16:13:20.213020Z digest=sha256:16529f00aad1d6b596ebc1ce576e8431ae9f4ae0f6b0a4027aae63ce7c04a292