Pith. sign in

Paper Citation Record · LEDGER

Learning Music Audio Representations With Limited Data

As of 22 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 0 inbound Pith citation observations for arXiv:2505.06042.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.06042 v1

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T22:52:42.452591Z

measured 30 of 30 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

30 of 30 outbound references displayed

  • verified exact0
  • verified fuzzy29
  • unresolved0
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e673b856-4eec-48cf-adb4-9942ec6a3298 · outbound

This paper cites A software framework for musical data augmentation,.

Learning Music Audio Representations With Limited Data A software framework for musical data augmentation,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:52:42.881907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:52:42.328469Z digest=sha256:469cc568e7bc2183a0fb18c67fc263b867a73fa38a6b46c8a0f3020b0a647e08

Observation f80f1afd-9555-4ad8-b283-d14c0b9985a3 · outbound

This paper cites Training Neural Audio Classifiers with Few Data,.

Learning Music Audio Representations With Limited Data Training Neural Audio Classifiers with Few Data,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:52:42.868410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:52:42.333653Z digest=sha256:3422de237427945b718d7c3e62c0bed9e03d762deb0a82f78df7fe8438d9d45d

Observation ea30be88-7e5d-4bbb-9625-218db9ebc540 · outbound

This paper cites Music cold-start and long-tail recommendation: bias in deep representations,.

Learning Music Audio Representations With Limited Data Music cold-start and long-tail recommendation: bias in deep representations,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:52:42.855038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:52:42.338293Z digest=sha256:cab18d4dca1c88e26a3a4f1e559bb5b958401d8feb10d90c3a1fca02c60a362a

Observation 71aabe2b-7ff4-4c5a-8ae4-9f4ad9b11760 · outbound

This paper cites Randomly Weighted CNNs for (Music) Audio Classification,.

Learning Music Audio Representations With Limited Data Randomly Weighted CNNs for (Music) Audio Classification,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:52:42.841188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:52:42.342728Z digest=sha256:44d39c7f5b2e9c6adbab030ce6863bac9f45ad36125f319ead51b3acbab62d82

Observation be5f1c25-ade1-4868-8ed1-17eeb74ec610 · outbound

This paper cites Deep Convolutional Neural Net- works and Data Augmentation for Environmental Sound Classification,.

Learning Music Audio Representations With Limited Data Deep Convolutional Neural Net- works and Data Augmentation for Environmental Sound Classification,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:52:42.827565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:52:42.347110Z digest=sha256:ba5911a9ae12835fcb6d1b084fb2186b0c5b44026b3cfd9ac3ccb81e3b847459

Observation 76512e8d-1006-4524-9de0-7bef9e0fb14c · outbound

This paper cites Environmental sound classification with convolutional neural networks,.

Learning Music Audio Representations With Limited Data Environmental sound classification with convolutional neural networks,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:52:42.813763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:52:42.351739Z digest=sha256:ac5ea719faadf15acc495a340c27a535f735cba4ee2ccddde35db5820b61d940

Observation 16deab34-71cc-497a-b07c-30f09767441d · outbound

This paper cites Generative Adversarial Network Based Acoustic Scene Training Set Augmentation and Selection Using SVM Hyper-Plane,.

Learning Music Audio Representations With Limited Data Generative Adversarial Network Based Acoustic Scene Training Set Augmentation and Selection Using SVM Hyper-Plane,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:52:42.800396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:52:42.356361Z digest=sha256:74fde8c50b9c7a6e575ce2c63793c8fcb989d0d8f304763fcb7979ef60efa3a0

Observation a3dcb4ec-348d-4f60-aadd-714acdc6b5c2 · outbound

This paper cites An ensemble of deep transfer learning models for handwritten music symbol recognition,.

Learning Music Audio Representations With Limited Data An ensemble of deep transfer learning models for handwritten music symbol recognition,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:52:42.787069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:52:42.360612Z digest=sha256:545b86778b1549c9124995c72b345dc9e87d8373a66fdd2a36e2f727ad463198

Observation 0b9a71b2-c5dd-4b21-8ee8-22b76256183f · outbound

This paper cites Transfer learning from speech to music: towards language-sensitive emotion recognition models,.

Learning Music Audio Representations With Limited Data Transfer learning from speech to music: towards language-sensitive emotion recognition models,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:52:42.773979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:52:42.364884Z digest=sha256:cc724bdd73ef48d0d1fb79660600a90bb9563741506ba869efb0573586f8967e

Observation d0a6c030-f9f6-4d1b-ad20-9f721101ce40 · outbound

This paper cites From West to East: Who can understand the music of the others better?,.

Learning Music Audio Representations With Limited Data From West to East: Who can understand the music of the others better?,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:52:42.760687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:52:42.369026Z digest=sha256:b6e48f1507d91d2be800450b713833d1a68f914c97ecf512e7c6a869b4900526

Observation 43b1ee8f-af81-4391-bd57-05034c8d206d · outbound

This paper cites Leveraging Hierarchical Structures for Few-Shot Musical Instrument Recognition,.

Learning Music Audio Representations With Limited Data Leveraging Hierarchical Structures for Few-Shot Musical Instrument Recognition,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:52:42.747041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:52:42.373246Z digest=sha256:41b8d034982d03ed9ffed3e9b979e32ce99d1129a97e8dd0779c13ef4eb266fa

Observation 11e5ff51-6f9a-47cf-98fd-3806e84cda0a · outbound

This paper cites Few- Shot Musical Source Separation,.

Learning Music Audio Representations With Limited Data Few- Shot Musical Source Separation,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:52:42.733439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:52:42.377256Z digest=sha256:06317bc78144b0788cf36427b60be149b87ef9acf73243d108b915c5a7ccdd98

Observation 844151d0-46c6-4730-b6de-ed93a844f44c · outbound

This paper cites Very Deep Convolutional Networks for Large-Scale Image Recognition,.

Learning Music Audio Representations With Limited Data Very Deep Convolutional Networks for Large-Scale Image Recognition,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:52:42.720028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:52:42.381366Z digest=sha256:a42881c8abb683cc8c00cdc943ed2ab76680e32a868c6e2fea28687359a52657

Observation 04d4de80-a51a-4078-9e9c-bd3961741dde · outbound

This paper cites MusiCNN: Pre-trained convolutional neural networks for music audio tagging,.

Learning Music Audio Representations With Limited Data MusiCNN: Pre-trained convolutional neural networks for music audio tagging,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:52:42.706198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:52:42.385765Z digest=sha256:4965cefaf92b51ba1466e4f1c017bae802f681b2037ccb8e36884732d6210bd9

Observation 51fcbfa3-6da0-4ea1-bf02-3fe85be025c1 · outbound

This paper cites AST: Audio Spectrogram Transformer,.

Learning Music Audio Representations With Limited Data AST: Audio Spectrogram Transformer,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:52:42.692449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:52:42.389907Z digest=sha256:d6d62ccb518783a129ae5ead7c02b0db075c472a1746f3e8c6247463230e2d6b

Observation 91367d50-38a7-490e-bbf3-4bf37c170696 · outbound

This paper cites Contrastive learning of musical representations,.

Learning Music Audio Representations With Limited Data Contrastive learning of musical representations,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:52:42.679485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:52:42.394251Z digest=sha256:f4b4b8bb241cf4f7b87fb3fc2e86411e17970c1501a22a64a61a9ef572cf5a82

Observation 54f4b755-2891-4b2c-8702-f05c7a49d9b2 · outbound

This paper cites Modality-Agnostic Self-Supervised Learning with Meta-Learned Masked Auto-Encoder,.

Learning Music Audio Representations With Limited Data Modality-Agnostic Self-Supervised Learning with Meta-Learned Masked Auto-Encoder,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:52:42.665313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:52:42.398338Z digest=sha256:bfbf83918e0f7dbbc8771b7c5599fdcb2d1186957e06ef97529b695d712cac28

Observation 9aa7a5ad-46c6-463a-99d2-24d915fed981 · outbound

This paper cites Eval- uation of cnn-based automatic music tagging models,.

Learning Music Audio Representations With Limited Data Eval- uation of cnn-based automatic music tagging models,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:52:42.651720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:52:42.402370Z digest=sha256:8a1bbd747bcb6568fed59c797a01e301ab75cd33cbb811b34c360db0a0557644

Observation 071ed08c-0a9e-4825-bf62-7ed87102920e · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale,.

Learning Music Audio Representations With Limited Data An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:52:42.638352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:52:42.406500Z digest=sha256:4574c2a3265fac522895a4a7e6674e690f073dd29b505491925e33b14ee8fca7

Observation 3ab3b913-f030-4de8-a95a-87c9e0d82464 · outbound

This paper cites A simple framework for contrastive learning of visual representations,.

Learning Music Audio Representations With Limited Data A simple framework for contrastive learning of visual representations,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:52:42.624529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:52:42.410712Z digest=sha256:3f579c4205f7907e5da027fde6659fd1596b010dd24d32a35327318f4033430c

Observation 9e077617-b770-410d-8851-cc996e76cc45 · outbound

This paper cites Sample-level cnn ar- chitectures for music auto-tagging using raw waveforms,.

Learning Music Audio Representations With Limited Data Sample-level cnn ar- chitectures for music auto-tagging using raw waveforms,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:52:42.611044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:52:42.414789Z digest=sha256:bccc77d17e34e789bfe7eed86aa3bafe80108c54caaf27fe4743192f0efff405

Observation 5dfb4f22-f45f-4016-accf-54997e131bec · outbound

This paper cites Evaluation of algorithms using games: The case of music tagging,.

Learning Music Audio Representations With Limited Data Evaluation of algorithms using games: The case of music tagging,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:52:42.597603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:52:42.419533Z digest=sha256:052c9a3d73395681b348856dc7f14200cb7f08d13d22e90da252b219b56aa18a

Observation b018bcb0-4a97-4a07-9924-7623e6c4fad9 · outbound

This paper cites mir ref: A Representation Evaluation Framework for Music Informa- tion Retrieval Tasks,.

Learning Music Audio Representations With Limited Data mir ref: A Representation Evaluation Framework for Music Informa- tion Retrieval Tasks,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:52:42.584725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:52:42.423389Z digest=sha256:30e6628b2a8d24542d135536c44bc447f0669e4ce1497a12d1ac426444cade55

Observation c106c726-377d-4535-bcec-3e3ca4ea57c8 · outbound

This paper cites Mir eval: A transparent implementation of common mir metrics,.

Learning Music Audio Representations With Limited Data Mir eval: A transparent implementation of common mir metrics,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:52:42.570948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:52:42.427500Z digest=sha256:6f6d30309dd0443adaf73615415a08f9d48264bb645332c340aa166ff71d8356

Observation 8b38782f-5c0f-4631-9205-b338ac624fa7 · outbound

This paper cites OrchideaSOL: a dataset of extended instrumental techniques for computer-aided orchestration,.

Learning Music Audio Representations With Limited Data OrchideaSOL: a dataset of extended instrumental techniques for computer-aided orchestration,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:52:42.555698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:52:42.431534Z digest=sha256:e9597ca7eda2db65ea5bf71025fa980cba3532a43436c1afeeb7537be7acc00b

Observation b84b1c43-4436-4a69-9b13-48941277cd23 · outbound

This paper cites Beatport EDM Key Dataset,.

Learning Music Audio Representations With Limited Data Beatport EDM Key Dataset,

Reference 26

Resolution
malformed identifier
no resolver link, observed 2026-08-15T22:52:42.435579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:52:42.435579Z digest=sha256:b1449c2f817f3f042f8042ededdea1420dfb1ffed3b441337fe0ada3ce593f55

Observation 57334717-26b4-4de7-acb8-90b0ca26b1e7 · outbound

This paper cites thesis, Universitat Pompeu Fabra, 2017.

Learning Music Audio Representations With Limited Data thesis, Universitat Pompeu Fabra, 2017

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:52:42.541448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:52:42.439816Z digest=sha256:36f2552477e1a61b8e91dde42257264a8215dd3436b10f3799df81ca74cf19e3

Observation 953cffe8-dcc4-4a7f-a3dd-b09118709309 · outbound

This paper cites An Empirical Study of Training Self-Supervised Vision Transformers,.

Learning Music Audio Representations With Limited Data An Empirical Study of Training Self-Supervised Vision Transformers,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:52:42.527165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:52:42.443843Z digest=sha256:bc06ba449a7acd066faea8ca239cfbbe198b6d573a4189a8b2990d3bac327d7c

Observation 20970267-b364-4ca8-9c5b-98c0d2657a68 · outbound

This paper cites A Survey on Reservoir Computing and its Interdisciplinary Applications Beyond Traditional Machine Learning,.

Learning Music Audio Representations With Limited Data A Survey on Reservoir Computing and its Interdisciplinary Applications Beyond Traditional Machine Learning,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:52:42.512484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:52:42.448197Z digest=sha256:0bda56612df7ff3976f2b855dd31abdf7ca72c0eab0140ed2d7cf54840d73cf4

Observation b9b93e66-f05f-4aaa-be21-fd4337cb0060 · outbound

This paper cites Reservoir Transformers,.

Learning Music Audio Representations With Limited Data Reservoir Transformers,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:52:42.497248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:52:42.452591Z digest=sha256:b0595790b49132d13d36d2780543bb11ed7fb8d21a70f4488ba9965a34f31523

Pith citing papers

No inbound Pith citation observations are available.