Pith. sign in

Paper Citation Record · LEDGER

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models

As of 8 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 4 inbound Pith citation observations for arXiv:2505.16306.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.16306 v1

Coverage vector

measured 40 of 40 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:05:56.973279Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:05:55.280579Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T20:50:12.622669Z

Reference resolution

40 of 40 outbound references displayed

  • verified exact1
  • verified fuzzy32
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9efc2332-6bf7-45d7-aff9-f81f0e0fd96e · outbound

This paper cites One of the most significant advantages of SSL mod- els is their ability to leverage unlabeled audio data, enabling the possibility of training with large-scale datasets.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models One of the most significant advantages of SSL mod- els is their ability to leverage unlabeled audio data, enabling the possibility of training with large-scale datasets

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:06:01.921801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:54.872351Z digest=sha256:cbc7816d3a90cf15d87efa9773ed1d2fa72759945c0f9372eb9d251f9a8f3438

Observation 5432a338-3e0a-41d7-958d-62c4f8c5292e · outbound

This paper cites an unresolved cited work.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:06:01.723689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:54.927510Z digest=sha256:218e428d1ce8ebee8a1b5bf785249ffc820e36e3132a6e9378e513feb9d4450c

Observation 4e428a11-788d-4cad-b699-519f384a2c3d · outbound

This paper cites Work performed during an internship at Tencent AI Lab.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Work performed during an internship at Tencent AI Lab

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:06:01.505537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:55.009689Z digest=sha256:e5da17cbcf3c945c12d5aa1908db3ebf49ece5261b747fcf2ea682662b7554d6

Observation 2257135f-cad6-4523-8a82-668e65d577b7 · outbound

This paper cites However, re- search in this area has been hindered by the limitations in data access [6].

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models However, re- search in this area has been hindered by the limitations in data access [6]

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:06:01.329849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:55.074949Z digest=sha256:88e336fe013ce81887853644bd156f2948f9e9f8f817279eb616611876ab9ed0

Observation 46bac54d-5dd4-4305-8af0-d80f5213af3d · outbound

This paper cites Layer-wise analysis provides us with a detailed perspective on the relationship between the performance of each specific layer and each specific task.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Layer-wise analysis provides us with a detailed perspective on the relationship between the performance of each specific layer and each specific task

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:06:01.046361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:55.453571Z digest=sha256:1523016fc11b106da7eb7013023877f2bff10bbc10b901511103db2f139c0355

Observation 39a9a0a6-60e4-4052-bea7-1b5c7ae6983e · outbound

This paper cites Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T15:05:55.280579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:05:55.280579Z digest=sha256:7223acfcd768af3686e62cc64b5c30d2693b5cc929198305b12f076c0067400d

Observation 6c7e2d90-dd72-4f85-8010-59217d9f5447 · outbound

This paper cites SSL outperforms low-level features.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models SSL outperforms low-level features

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:06:01.148931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:55.370470Z digest=sha256:1105d3f90272f8c39651332ebce3aee951c1df12f7510814e4a167fd5fb86c83

Observation 2a2c04ad-e596-4459-afc8-422aafedc336 · outbound

This paper cites wav2vec 2.0: A framework for self-supervised learning of speech representations.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models wav2vec 2.0: A framework for self-supervised learning of speech representations

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:06:00.112543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:55.885908Z digest=sha256:9450c8e9194018f8081eb04772599c32a58b0fc3fcbeaf388b03bee7ad71f0fc

Observation 85d6cc51-2b55-4f3b-885c-427bfc96bd02 · outbound

This paper cites Mert: Acoustic music un- derstanding model with large-scale self-supervised training, 2023.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Mert: Acoustic music un- derstanding model with large-scale self-supervised training, 2023

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:06:00.892246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:55.539934Z digest=sha256:7b659d171d7f0925add132428c3bd3b52cbc6a9cbbc6838913af4aacf8d09a76

Observation d8f84f0a-0a37-42db-a3e9-68fc1698e0a4 · outbound

This paper cites an unresolved cited work.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:06:01.246021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:55.178803Z digest=sha256:410396676aa677f4a247496adc018d2f77904273e56c43df9df120af53254f23

Observation f1945d7b-1e6b-47e4-bd1b-aac6b4dfd18a · outbound

This paper cites A Foundation Model for Music Informatics.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models A Foundation Model for Music Informatics

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T15:05:55.604317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:05:55.604317Z digest=sha256:1bc11b3a6a8c4486f836b8adf69320c11d6d1389d4f7df2f973779f1fc792389

Observation d70d69ec-f262-4e6d-9acb-7913b0d923e4 · outbound

This paper cites Contrastive learn- ing of musical representations.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Contrastive learn- ing of musical representations

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:06:00.736849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:55.644568Z digest=sha256:8dee197fe2c35e2ff248272d1198f7f106973ec341a3e3560969d689da595694

Observation d9a493ef-32e2-400e-9217-934118d4d34b · outbound

This paper cites Codified audio language modeling learns useful representations for music information retrieval.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Codified audio language modeling learns useful representations for music information retrieval

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:06:00.556330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:55.674618Z digest=sha256:78f16766aababddb2dad1ac6c597958cd542b267aa94894499fdaf6997ed8d8f

Observation b138a273-f183-4747-8d7d-8b2a8f188898 · outbound

This paper cites Supervised and Unsupervised Learning of Audio Representations for Music Understanding.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Supervised and Unsupervised Learning of Audio Representations for Music Understanding

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T15:05:55.715841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:05:55.715841Z digest=sha256:388a024dfe482e75be4d7e62930deb1ac06291bcff62a0211e0b1752e17fa9a0

Observation c2bf2af4-3915-4dc3-91b7-ecdb7f6e5993 · outbound

This paper cites Freeman, Jessie Wang, Sherry Cai, and KatherineM.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Freeman, Jessie Wang, Sherry Cai, and KatherineM

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:06:00.431173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:55.773738Z digest=sha256:58a65ed9c0e96e47bb5eada428fe65ede2a4c54b90c84488b051d6baef7cc238

Observation 315b912d-4b93-4224-aa5c-b09749dee9b4 · outbound

This paper cites wav2vec: Unsupervised pre-training for speech recognition.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models wav2vec: Unsupervised pre-training for speech recognition

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:06:00.230779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:55.832610Z digest=sha256:72ea49a5bde1fd76bc3f3018d7384ea2a25ab63e7a31207a48bc441ceb98c249

Observation 25d1bd04-b69d-4344-a163-75ea9161965a · outbound

This paper cites Hubert: Self-supervised speech representation learning by masked prediction of hidden units.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Hubert: Self-supervised speech representation learning by masked prediction of hidden units

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:59.928652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:55.922222Z digest=sha256:36a4445039a8c26a026545cc751b94923ba4a8fa1780f2470f7a846ad3fc6efa

Observation 218c925e-a4aa-4497-ba42-15e85c138bfd · outbound

This paper cites MuQ: Self-Supervised Music Representation Learning with Mel Residual Vector Quantization.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models MuQ: Self-Supervised Music Representation Learning with Mel Residual Vector Quantization

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T15:05:55.961928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:05:55.961928Z digest=sha256:13979902dd83470f4d0acbdbfff713f47cec9ff1f77b0b1ca86143e33b97587d

Observation 93e2fabd-2865-48a5-bb8b-88f70466276e · outbound

This paper cites Self-supervised learning with random-projection quantizer for speech recognition.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Self-supervised learning with random-projection quantizer for speech recognition

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:59.685321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:56.033041Z digest=sha256:3047a34e139f22b7940bba2030e466a5a970552a76a370b4dc3854112ddc4be2

Observation b6d2a11f-30f8-4932-8eb3-a91a6585aefb · outbound

This paper cites Why does Self-Supervised Learning for Speech Recognition Benefit Speaker Recognition?.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Why does Self-Supervised Learning for Speech Recognition Benefit Speaker Recognition?

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:05:57.096033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:56.086578Z digest=sha256:6499c9c1c22c501e253be1aa05c005fcb2148f204cb2d1cad68d523c80952f01

Observation ad8b857a-d35a-4139-9563-4b7cb6a970f6 · outbound

This paper cites Layer-wise analysis of a self-supervised speech representation model.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Layer-wise analysis of a self-supervised speech representation model

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:59.394216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:56.135270Z digest=sha256:bdc6a0c04c2c4ce4f1bb5da6c4760bc7aab42d1f3cfe6230dd10742417866878

Observation c035fea0-94a9-4f72-bed2-aef1cf739023 · outbound

This paper cites Liu, Cheng- I Lai, Haibin Wu, Jiatong Shi, Xuankai Chang, Hsiang-Sheng Tsai, Wen-Chin Huang, Tzu hsun Feng, Po-Han Chi, Yist Y.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Liu, Cheng- I Lai, Haibin Wu, Jiatong Shi, Xuankai Chang, Hsiang-Sheng Tsai, Wen-Chin Huang, Tzu hsun Feng, Po-Han Chi, Yist Y

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:59.144765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:56.183033Z digest=sha256:c0c2f9ff7d891b79740b23083a0a1b62fbf8b95b53f2dd4d684bcb36758668f0

Observation 77e055ca-dc0f-400b-b57b-2b65d10e45a9 · outbound

This paper cites MARBLE: Music Audio Representation Benchmark for Universal Evaluation.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models MARBLE: Music Audio Representation Benchmark for Universal Evaluation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T15:05:56.227445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:05:56.227445Z digest=sha256:4d115a068cd9ad535e3b82226e32b05cf6e62aab0bcb8bc3dffc20535ac59f19

Observation e56e578f-8b3a-4696-b2f1-43f5bb32b03e · outbound

This paper cites Bert: Pre-training of deep bidirectional transformers for language understanding.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Bert: Pre-training of deep bidirectional transformers for language understanding

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:59.057122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:56.275271Z digest=sha256:99b68049b3e1f28dbd1988b477898df13bc08829c23f8db238dfe5e067416d46

Observation 377bcae6-c52f-4742-8655-64623ff324cc · outbound

This paper cites Advances in residual vector quantization: A review.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Advances in residual vector quantization: A review

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:58.877723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:56.329474Z digest=sha256:f9aa5582ae95a90451137f6d0d2a26b791800ac052d34f3bf0e92f0c02acc9e5

Observation 6cf9e074-cb8e-4a1e-8941-39bdc97b284b · outbound

This paper cites The million song dataset.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models The million song dataset

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:58.692091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:56.373393Z digest=sha256:b183479a19a5be97fa7a08ded5c167d4108f9436f1548a0c801222b632d0eefc

Observation fa72f0d1-d06c-43f9-92cd-631e3e0d6125 · outbound

This paper cites Relations between two sets of variates.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Relations between two sets of variates

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:58.557201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:56.406095Z digest=sha256:1dbed412f8a5587e3c10f9a5f58ac0539c9827b3c19e02b55d7848e807a42b49

Observation cd1447f1-e243-4277-b2f7-4ec82ae1aeaa · outbound

This paper cites Svcca: Singular vector canonical correlation analysis for deep learning dynamics and interpretability.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Svcca: Singular vector canonical correlation analysis for deep learning dynamics and interpretability

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:58.385874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:56.464439Z digest=sha256:1f4831450fc52915e4d057858c1ddbb2d41748364a1de15e13ad6c492029e98c

Observation a980500c-e345-4267-828b-e9528db167d4 · outbound

This paper cites A neural network that finds a naturalistic so- lution for the production of muscle activity.Nature Neuroscience, page 1025–1033, Jul 2015.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models A neural network that finds a naturalistic so- lution for the production of muscle activity.Nature Neuroscience, page 1025–1033, Jul 2015

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:58.196899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:56.520896Z digest=sha256:07e1e53b54ac7427a9c1a026625b77df3d4427a206490e58cabca7668e4563d2

Observation 6c0c51e5-f56c-471a-a48f-0438729b505a · outbound

This paper cites Morcos, Maithra Raghu, and Samy Bengio.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Morcos, Maithra Raghu, and Samy Bengio

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:58.054768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:56.566040Z digest=sha256:3f1a2fdf7000d0f9ba76cfe331b8bbc486e6f6e1ff2fc0301db434d46ce30f40

Observation 5e80005a-8637-412d-9ac7-f8a0d9613f3f · outbound

This paper cites Neu- ral audio synthesis of musical notes with wavenet autoencoders.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Neu- ral audio synthesis of musical notes with wavenet autoencoders

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:57.981388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:56.603367Z digest=sha256:ec6a279a43b99bc38bc4b9f9248ddbafa3db360593684eceb903cbe6496eae57

Observation 9826e2c9-f844-4b75-9456-a9b9e296db94 · outbound

This paper cites V ocalset: A singing voice dataset.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models V ocalset: A singing voice dataset

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:57.911137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:56.648332Z digest=sha256:9f765df73336f406e898c634fc81d139e2fb60b50f1728e75682d237fc5bac4b

Observation 621ffcf2-1421-48d8-850f-a471c50c7a3e · outbound

This paper cites Tzanetakis and P.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Tzanetakis and P

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:57.834666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:56.691957Z digest=sha256:59945b99adc0458d1c66a63747ff585122674e3e597a8a23d43e73f95b9b8f90

Observation c6b6f3c0-e917-40c0-9004-3552bfc45adc · outbound

This paper cites Caro, Erik M.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Caro, Erik M

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:57.754059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:56.727428Z digest=sha256:144f559fa75581f6491a20efd228c6ee0711a5b2701eb29a95ce79145dbcda7f

Observation ab32d531-7d2a-45b1-b59b-7745c758fc49 · outbound

This paper cites The harmonix set: Beats, downbeats, and functional segment annotations of western popular music.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models The harmonix set: Beats, downbeats, and functional segment annotations of western popular music

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:57.668921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:56.770281Z digest=sha256:3af57b0abd2d04698ec4cbe9fe28b745a35484e06d3ae1a650a8de14e9019f22

Observation 60082b96-8d3c-42eb-a152-06b10dfa428b · outbound

This paper cites Evaluation of algorithms using games: The case of mu- sic tagging.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Evaluation of algorithms using games: The case of mu- sic tagging

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:57.609063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:56.818004Z digest=sha256:20ccb75951f21f70e18f178caadfe947bd3e3c3776898b5a24a62a0c84e31907

Observation 341419e3-4c77-4e32-a151-59477eb26d50 · outbound

This paper cites Two data sets for tempo estimation and key detection in electronic dance music annotated from user corrections.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Two data sets for tempo estimation and key detection in electronic dance music annotated from user corrections

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:57.512753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:56.855205Z digest=sha256:e0e2ef0f8673690554616560097efd287a8bb720b66ffb251d46d827dc662b04

Observation 4cc1e258-4270-4c12-bc7a-01dcf23f582b · outbound

This paper cites Mir eval: A transparent implementation of common mir metrics.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Mir eval: A transparent implementation of common mir metrics

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:57.419662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:56.902772Z digest=sha256:bc8493a5eccee06973f2b1fd80984a3ddd7dee698febf850b51dba92550913e8

Observation 30a79f16-94e4-41ef-91c5-4ab9dd9e0b5e · outbound

This paper cites Deep contextualized word representations.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Deep contextualized word representations

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:57.335343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:56.938557Z digest=sha256:1c6e4bdb037298810430929e8bb8ca98684dc2e983963d56150cc1affb757385

Observation 999f0691-12ec-4619-915f-d4b0badc3a51 · outbound

This paper cites Comparative layer- wise analysis of self-supervised speech models.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Comparative layer- wise analysis of self-supervised speech models

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:57.247631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:05:56.973279Z digest=sha256:147393ce1037567cfb02a7400204c47da11189529d2da5d4eb1ff3f5fd39b134

Pith citing papers

Observation 39a9a0a6-60e4-4052-bea7-1b5c7ae6983e · inbound

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models cites this paper.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T15:05:55.280579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:05:55.280579Z digest=sha256:7223acfcd768af3686e62cc64b5c30d2693b5cc929198305b12f076c0067400d

Observation 1cd830bc-e233-4431-a554-d7fca4c1f6c8 · inbound

Revisiting Content-Based Music Recommendation: Efficient Feature Aggregation from Large-Scale Music Models cites this paper.

Revisiting Content-Based Music Recommendation: Efficient Feature Aggregation from Large-Scale Music Models Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-16T02:40:30.828825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T02:39:46.434831Z digest=sha256:ef05ad9722ab4df280aa8163bcd1f48b0112425d289dfa334e5d7f5967722efa

Observation f8a44b3f-8f2f-44f0-a274-b5067cebb103 · inbound

Frequency-Aware Self-Supervised Music Representation Learning cites this paper.

Frequency-Aware Self-Supervised Music Representation Learning Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-07-04T20:50:12.625036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-25T19:25:54.322923Z digest=sha256:dc9b333e3b19406cd05cd506bc42cef6a98c057b01ac0ff54c49beeea6170f0e

Observation 8830d88e-5624-44f1-ab3c-0209d364f729 · inbound

Frequency-Aware Self-Supervised Music Representation Learning cites this paper.

Frequency-Aware Self-Supervised Music Representation Learning Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-06-30T10:04:35.631723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T10:04:11.233115Z digest=sha256:3dbf741bbc96627e8fa612f4819c79b9a46ce749722fcb41f030e6e3d1cee125