Pith. sign in

Paper Citation Record · LEDGER

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation

As of 22 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 1 inbound Pith citation observation for arXiv:2505.18984.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.18984 v1

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:26:15.094318Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:26:13.455801Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T14:26:15.140016Z

Reference resolution

35 of 35 outbound references displayed

  • verified exact0
  • verified fuzzy29
  • unresolved5
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 48b46bf4-b2c4-416a-ba41-4483c963d721 · outbound

This paper cites Self-supervised learning method using multiple sampling strategies for general-purpose audio representation.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Self-supervised learning method using multiple sampling strategies for general-purpose audio representation

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T14:26:15.146229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:26:13.455801Z digest=sha256:4d196f996ca01802162ba8acc447dbfd366f85da1b9e447c18e71e30355dbd9d

Observation 1d365e53-61e9-47cd-a079-47686ed747c0 · outbound

This paper cites an unresolved cited work.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:26:15.473825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:26:13.522458Z digest=sha256:577d44a469cdb71263818513acdd559c8db398da43356dd16f3ed00eb3611958

Observation ecf4b75a-3130-4d09-9856-f5da78ec815b · outbound

This paper cites an unresolved cited work.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:26:15.464221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:26:13.596456Z digest=sha256:09f1f6836e12739d130005c8cfd192fb2b8bd6de08e3fb26d7863ce73a42a310

Observation 7d23d873-1395-4f19-95d9-d6c6627e9ec1 · outbound

This paper cites an unresolved cited work.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:26:15.455145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:26:13.717712Z digest=sha256:45d0a98f7301243db4c47bfe3d1d00505f122db58a8740cff9b99a8d8bfe5248

Observation 1f11e76a-f560-4804-a45b-3c62de208ae4 · outbound

This paper cites Our method improves the performance of all tasks compared to existing methods in the experiment of the subset of Audioset.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Our method improves the performance of all tasks compared to existing methods in the experiment of the subset of Audioset

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.444845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:26:13.856526Z digest=sha256:b57303d05eb7a953deb194ad2b3eda1da6ef4b8d788d5c89d276dd5a8e181349

Observation 6e6c43ae-0e3b-42bf-90e3-daab6052f574 · outbound

This paper cites Zhang, J.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Zhang, J

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.434624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:26:13.964205Z digest=sha256:9dccc355dba4b03085c2d3b4767dda80952645d47cde4634a6bfb5ac29a05d72

Observation 8cf1eabf-32f1-4f57-87bd-db48d3abf6f8 · outbound

This paper cites V oxCeleb2: Deep Speaker Recognition,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation V oxCeleb2: Deep Speaker Recognition,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.424791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:26:14.125586Z digest=sha256:f66fb39f7d5c0becdc68992b9018ecc074d9db581339b0b509dfc56e69adbbce

Observation b6cc71bd-6a73-4439-9b19-e67a166e88ab · outbound

This paper cites Broadcasted Resid- ual Learning for Efficient Keyword Spotting,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Broadcasted Resid- ual Learning for Efficient Keyword Spotting,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.414577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:26:14.205583Z digest=sha256:fc83d16e29c1ccb06391359f98002e87c17093ae88a078991dcd385496793612

Observation cc1bcaec-6077-4705-9f60-e47444a973e4 · outbound

This paper cites Acous- tic Event Detection Method Using Semi-Supervised Non- Negative Matrix Factorization with Mixtures of Local Dictio- naries,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Acous- tic Event Detection Method Using Semi-Supervised Non- Negative Matrix Factorization with Mixtures of Local Dictio- naries,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.404451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:26:14.292479Z digest=sha256:ae5e54b4e2621abfe1d73c5b015deca3ae258b927c0e63bc15ab171f420424cd

Observation 1d5e790f-2b66-47be-9af5-b4792fcac3de · outbound

This paper cites Crepe: A Convolutional Representation for Pitch Estimation,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Crepe: A Convolutional Representation for Pitch Estimation,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.394428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:26:14.421892Z digest=sha256:6889ba167cd05648591718e7b1b6aa804122c462165f9c51cf48823d50441d0d

Observation 271aff89-e090-484e-9276-2133fc6a1935 · outbound

This paper cites PANNs: Large-scale pre- trained audio neural networks for audio pattern recognition,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation PANNs: Large-scale pre- trained audio neural networks for audio pattern recognition,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.383661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:26:14.452385Z digest=sha256:f96b919e703172d3d5b8d5720c60c59071227201561b4287a683187c6a08c67c

Observation eb5f3bcf-9037-43ff-9499-4981d1a64270 · outbound

This paper cites Audio set: An ontology and human-labeled dataset for audio events,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Audio set: An ontology and human-labeled dataset for audio events,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.374236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:26:14.463942Z digest=sha256:3adc03f659b7417777117c552c95a68b9e9c97da63f8839c33ddec3510358e41

Observation fa39727f-bf1f-453e-8145-6cec55c245f3 · outbound

This paper cites Language models are few-shot learners,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Language models are few-shot learners,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.364561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:26:14.499846Z digest=sha256:7428c77a9a3af7fd95220f66a8a60a91e15558b0c2d95684200c69524b0aa4fd

Observation 2a765ba9-705f-44e3-8df0-30025229487f · outbound

This paper cites BERT: Pre-training of deep bidirectional transformers for language understanding,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation BERT: Pre-training of deep bidirectional transformers for language understanding,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.353992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:26:14.538877Z digest=sha256:5ec0277f440c0897c9f844998ce40107f645daf7bfb19cc87082462ecc0578c5

Observation 9dfdb037-9d11-4739-9a84-b2bfcb93c3d6 · outbound

This paper cites van den Oord, Y.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation van den Oord, Y

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.343929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:26:14.576176Z digest=sha256:7ba4ff5b595b59b6850936f54de6c0af497c7685c9f6d011a5860d9bdaa7dee7

Observation e785da60-2e2f-4746-a0d1-1ecf54df335f · outbound

This paper cites Spatiotemporal con- trastive video representation learning,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Spatiotemporal con- trastive video representation learning,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.335105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:26:14.611759Z digest=sha256:de814fcf7dd3e5671d2803e2f6ed246d72b1831718c84f0175b251488014912f

Observation 875a203f-6db4-4d08-a62a-c8a8a6123d2b · outbound

This paper cites Momentum contrast for unsupervised visual representation learning,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Momentum contrast for unsupervised visual representation learning,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.325640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:26:14.684273Z digest=sha256:f63a1cb8a8e03a013c02fbcb90015c6580ebbb0a931e9f70a293e7832ea696d2

Observation b4135544-6834-4e57-aa62-4f5d9e78f94d · outbound

This paper cites CURL: Contrastive Unsupervised Representations for Reinforcement Learning.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation CURL: Contrastive Unsupervised Representations for Reinforcement Learning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:26:14.769986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:26:14.769986Z digest=sha256:b9df7a28c1db9a7ca4f5c86ef480ce3f073d74c9c4dde9286e2732e1cad3a96c

Observation b692c067-3335-431a-87f7-a97e44c6312b · outbound

This paper cites Data aug- menting contrastive learning of speech representations in the time domain,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Data aug- menting contrastive learning of speech representations in the time domain,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.315626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:26:14.842785Z digest=sha256:da4d165cb98d94b585c78e76ddb80ca8f39986392d58b71ed4734e65c0c8424b

Observation 59b0cce2-ceb8-43a8-91c5-3eee57250880 · outbound

This paper cites Un- supervised pretraining transfers well across languages,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Un- supervised pretraining transfers well across languages,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.305950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:26:14.922554Z digest=sha256:15c50ef75dfdbb76e8dd7ec611b99770130400d2c4c713e9b719b11b2c437259

Observation d515b184-3187-4d3d-b367-45ffb719a087 · outbound

This paper cites Vq-wav2vec: Self- supervised learning of discrete speech representations,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Vq-wav2vec: Self- supervised learning of discrete speech representations,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.295581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:26:14.983591Z digest=sha256:5cca8fc080c544db0df085b2c028cd4bedd5463adf8552403499d1d1c4e4b8ab

Observation 378785f9-1a3f-429e-9821-9f05fcd1a878 · outbound

This paper cites Towards Learning a Uni- versal Non-Semantic Representation of Speech,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Towards Learning a Uni- versal Non-Semantic Representation of Speech,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.285790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:26:15.019483Z digest=sha256:239e308d1d4cedd2d29b790bd49bcf9dd218d2fb88e41e4bf3b52196e90e6dbb

Observation 7fdff210-619f-4579-9761-6314246933fe · outbound

This paper cites Unsupervised learn- ing of semantic audio representations,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Unsupervised learn- ing of semantic audio representations,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.275794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:26:15.032551Z digest=sha256:40f3174ff333fd83239f885b4efd101572bff5321796df1aa5e08efd5775dd7f

Observation 1c684355-4e1b-416c-98a5-c0a4e876d0fc · outbound

This paper cites Contrastive learning of general-purpose audio representations,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Contrastive learning of general-purpose audio representations,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.266366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:26:15.038803Z digest=sha256:33bfb3c617f32e37c9d29daa9947043d150ddc55ea87f4ab701f6c8e0d63f13f

Observation a1f249f7-88e6-4ee0-ad61-eeb72e973e28 · outbound

This paper cites Speech Commands: A Dataset for Limited- V ocabulary Speech Recognition,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Speech Commands: A Dataset for Limited- V ocabulary Speech Recognition,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.256554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:26:15.045522Z digest=sha256:659962d342adf35f784fabcb74dada6d0c95a95ddec178ed30ad6282c84983a2

Observation 2bde6d46-b343-4d3f-9d5f-97ad3700ed27 · outbound

This paper cites Detection and classification of acoustic scenes and events: Outcome of the DCASE 2016 challenge,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Detection and classification of acoustic scenes and events: Outcome of the DCASE 2016 challenge,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.246276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:26:15.050539Z digest=sha256:d127301d30d66da01748425ab4715c0c79362ded0994ac13ddc11eb9ff8c18b7

Observation c18c0708-47a8-4b94-8ee3-6b515452dd27 · outbound

This paper cites Sound event de- tection in synthetic audio: Analysis of the DCASE 2016 task results,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Sound event de- tection in synthetic audio: Analysis of the DCASE 2016 task results,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.236043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:26:15.058282Z digest=sha256:83f7af2b5650423adff5a1d21ffe61890350714cebb8c6535b3dfea9cbe5d2c1

Observation 85cd9a4a-1503-44f9-929f-a47320bc853d · outbound

This paper cites Neural audio syn- thesis of musical notes with wavenet autoencoders,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Neural audio syn- thesis of musical notes with wavenet autoencoders,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.225323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:26:15.064430Z digest=sha256:90f993100c05dc3b776817413c2555864661ef1206c749253288bc365e9116d8

Observation 2ab1526b-ed67-48f7-b827-42d9e438671a · outbound

This paper cites Crepe: A convo- lutional representation for pitch estimation,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Crepe: A convo- lutional representation for pitch estimation,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.214614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:26:15.071478Z digest=sha256:867b90f38f026ffa508539f5b846d08739cb22e58743cf5f758b1cc41d1747c3

Observation fef3c815-c64d-4fda-8226-fe579a862e91 · outbound

This paper cites EfficientNet: Rethinking model scaling for convolutional neural networks,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation EfficientNet: Rethinking model scaling for convolutional neural networks,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.204775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:26:15.075322Z digest=sha256:0a9d8fa6accd411d0f70ea9c7aa02618d654d5a4b92aefde0a1d6bc83740ee80

Observation ee623c89-baf4-4df3-afc2-a2cc622df993 · outbound

This paper cites Layer Normalization.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Layer Normalization

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:26:15.079404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:26:15.079404Z digest=sha256:0a148d2ca3c104b9d2c9a61beec85408ce40bd66a96f8da5a59ca7ce4ee3e33e

Observation 1a08c4ff-ea39-4cda-afa8-d243f55b969a · outbound

This paper cites Adam: A method for stochastic op- timization,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Adam: A method for stochastic op- timization,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.193151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:26:15.083555Z digest=sha256:35891df3a8a5cf449f399f821c845a0a949dd0eec68d4eee176cea15afb20a94

Observation d8909262-716a-4fa9-8160-8cc1eeb391d2 · outbound

This paper cites Tagliasacchi, B.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Tagliasacchi, B

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.181985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:26:15.087529Z digest=sha256:473d58176bef5478359f16e39582a68d1919a241c6fc4568f0aceaa8e29bd7a2

Observation 9ec2a0ce-34a3-4a42-91b9-06c3e843fc90 · outbound

This paper cites A simple framework for contrastive learning of visual representations,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation A simple framework for contrastive learning of visual representations,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.170421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:26:15.091212Z digest=sha256:049fc5767222a642c47f1f0f1ab49b4e7dadb6e0a031344751ee6a9a1b9250c9

Observation 0a1819fe-7103-42a9-a291-f07153c04876 · outbound

This paper cites Seeing voices and hearing voices: Learning discriminative embeddings us- ing cross-modal self-supervision,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Seeing voices and hearing voices: Learning discriminative embeddings us- ing cross-modal self-supervision,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.157936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:26:15.094318Z digest=sha256:ef0af3cf1409b3ec69459a592413b1d6e1059babd6ec910f1a3354120d6cca3d

Pith citing papers

Observation 48b46bf4-b2c4-416a-ba41-4483c963d721 · inbound

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation cites this paper.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Self-supervised learning method using multiple sampling strategies for general-purpose audio representation

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T14:26:15.146229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:26:13.455801Z digest=sha256:4d196f996ca01802162ba8acc447dbfd366f85da1b9e447c18e71e30355dbd9d