Pith. sign in

Paper Citation Record · LEDGER

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation

As of 8 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 1 inbound Pith citation observation for arXiv:2505.18984.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.18984 v1

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:26:15.094318Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:26:13.455801Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T14:26:15.140016Z

Reference resolution

35 of 35 outbound references displayed

  • verified exact0
  • verified fuzzy29
  • unresolved5
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 48b46bf4-b2c4-416a-ba41-4483c963d721 · outbound

This paper cites Self-supervised learning method using multiple sampling strategies for general-purpose audio representation.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Self-supervised learning method using multiple sampling strategies for general-purpose audio representation

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T14:26:15.146229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:26:13.455801Z digest=sha256:0a402f95b9bd126784dd9912fd4df46723902c91427271fce3b254ab7823280f

Observation 1d365e53-61e9-47cd-a079-47686ed747c0 · outbound

This paper cites an unresolved cited work.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:26:15.473825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:26:13.522458Z digest=sha256:543810bf3ac384e061c3b1bc1510439a77643265da924b9f85ab2ca7602c39f8

Observation ecf4b75a-3130-4d09-9856-f5da78ec815b · outbound

This paper cites an unresolved cited work.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:26:15.464221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:26:13.596456Z digest=sha256:becf612a1ff714672ebce769710da2c790c9ec31b4a095cb8780201688c210cf

Observation 7d23d873-1395-4f19-95d9-d6c6627e9ec1 · outbound

This paper cites an unresolved cited work.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:26:15.455145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:26:13.717712Z digest=sha256:f5ca2deb44db02eab37d765e5fff94e53cf7d097bc085281318f0d5d01d25a16

Observation 1f11e76a-f560-4804-a45b-3c62de208ae4 · outbound

This paper cites Our method improves the performance of all tasks compared to existing methods in the experiment of the subset of Audioset.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Our method improves the performance of all tasks compared to existing methods in the experiment of the subset of Audioset

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.444845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:26:13.856526Z digest=sha256:e8b0869c55119914db904e965ad63e29182188173aa50d3261e1839c61b24e20

Observation 6e6c43ae-0e3b-42bf-90e3-daab6052f574 · outbound

This paper cites Zhang, J.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Zhang, J

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.434624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:26:13.964205Z digest=sha256:b756f6f82bfd0328c6fad3a510ef7eb3bdcb9230d202724a5bd6721131234214

Observation 8cf1eabf-32f1-4f57-87bd-db48d3abf6f8 · outbound

This paper cites V oxCeleb2: Deep Speaker Recognition,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation V oxCeleb2: Deep Speaker Recognition,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.424791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:26:14.125586Z digest=sha256:76f0b4a3ce34206e499e97f09a447ed59609cd18c2f3940bdf480bbbf15cc53a

Observation b6cc71bd-6a73-4439-9b19-e67a166e88ab · outbound

This paper cites Broadcasted Resid- ual Learning for Efficient Keyword Spotting,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Broadcasted Resid- ual Learning for Efficient Keyword Spotting,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.414577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:26:14.205583Z digest=sha256:f838d5e894a3d3e8729e9559ff609992a63bad4656a1aa0884457d0f447092f9

Observation cc1bcaec-6077-4705-9f60-e47444a973e4 · outbound

This paper cites Acous- tic Event Detection Method Using Semi-Supervised Non- Negative Matrix Factorization with Mixtures of Local Dictio- naries,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Acous- tic Event Detection Method Using Semi-Supervised Non- Negative Matrix Factorization with Mixtures of Local Dictio- naries,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.404451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:26:14.292479Z digest=sha256:b5e2c71f4e72472ebc370459c9ab6ece523e3f6f9f1a81c5825251d1ac7dd533

Observation 1d5e790f-2b66-47be-9af5-b4792fcac3de · outbound

This paper cites Crepe: A Convolutional Representation for Pitch Estimation,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Crepe: A Convolutional Representation for Pitch Estimation,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.394428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:26:14.421892Z digest=sha256:e0121240332aebb1396ba751de7339d1b699b7ac3552fc5af2ab3da40372a44a

Observation 271aff89-e090-484e-9276-2133fc6a1935 · outbound

This paper cites PANNs: Large-scale pre- trained audio neural networks for audio pattern recognition,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation PANNs: Large-scale pre- trained audio neural networks for audio pattern recognition,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.383661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:26:14.452385Z digest=sha256:fdcf9259b131101ab0715f8be43ed2733c0bff53b75c4cb124e3f1e2ea0942ca

Observation eb5f3bcf-9037-43ff-9499-4981d1a64270 · outbound

This paper cites Audio set: An ontology and human-labeled dataset for audio events,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Audio set: An ontology and human-labeled dataset for audio events,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.374236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:26:14.463942Z digest=sha256:386a96659f06d7f085bddd756900782d91796e4deaeeb73294c0f9000281aab0

Observation fa39727f-bf1f-453e-8145-6cec55c245f3 · outbound

This paper cites Language models are few-shot learners,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Language models are few-shot learners,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.364561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:26:14.499846Z digest=sha256:001502e9763c53c2a17d53d0d9202a3d57eaa09ae5aa1fb3b234357796190c9c

Observation 2a765ba9-705f-44e3-8df0-30025229487f · outbound

This paper cites BERT: Pre-training of deep bidirectional transformers for language understanding,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation BERT: Pre-training of deep bidirectional transformers for language understanding,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.353992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:26:14.538877Z digest=sha256:5d9651c07996571491f295633967421f79ef9d16792fc35349d4ba5089d8fb52

Observation 9dfdb037-9d11-4739-9a84-b2bfcb93c3d6 · outbound

This paper cites van den Oord, Y.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation van den Oord, Y

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.343929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:26:14.576176Z digest=sha256:f1891f788a5f3a5074aeeac15ab9223f498aedd3a55e94fb92b6294bdea3d9f5

Observation e785da60-2e2f-4746-a0d1-1ecf54df335f · outbound

This paper cites Spatiotemporal con- trastive video representation learning,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Spatiotemporal con- trastive video representation learning,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.335105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:26:14.611759Z digest=sha256:17eba92459b8040a3250ffd0ad5c69affd7b7df15a20eb8b54e15bd6fe8d8d07

Observation 875a203f-6db4-4d08-a62a-c8a8a6123d2b · outbound

This paper cites Momentum contrast for unsupervised visual representation learning,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Momentum contrast for unsupervised visual representation learning,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.325640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:26:14.684273Z digest=sha256:e9ebfece3ff696b9dc7cec6aebef5b588a214c9641380c31a8748a1c232111f3

Observation b4135544-6834-4e57-aa62-4f5d9e78f94d · outbound

This paper cites CURL: Contrastive Unsupervised Representations for Reinforcement Learning.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation CURL: Contrastive Unsupervised Representations for Reinforcement Learning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:26:14.769986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:26:14.769986Z digest=sha256:b4b3107f35df0f9e872d6ff6cf47372d13a818fc7670d70d415fb2a2ca3c1db2

Observation b692c067-3335-431a-87f7-a97e44c6312b · outbound

This paper cites Data aug- menting contrastive learning of speech representations in the time domain,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Data aug- menting contrastive learning of speech representations in the time domain,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.315626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:26:14.842785Z digest=sha256:281ddae3f098522245d7b8f67c3317f232bdde39fed2414528f936ee97d2c058

Observation 59b0cce2-ceb8-43a8-91c5-3eee57250880 · outbound

This paper cites Un- supervised pretraining transfers well across languages,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Un- supervised pretraining transfers well across languages,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.305950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:26:14.922554Z digest=sha256:20f175c393a2ea261ef63220a05f68a16d86dbbbb5ef8d22f7eb3e03b8b02b29

Observation d515b184-3187-4d3d-b367-45ffb719a087 · outbound

This paper cites Vq-wav2vec: Self- supervised learning of discrete speech representations,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Vq-wav2vec: Self- supervised learning of discrete speech representations,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.295581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:26:14.983591Z digest=sha256:61d72713d65db8d9a587ff16bf1906fdd83b46aaac8561a79a140441f7fc72f4

Observation 378785f9-1a3f-429e-9821-9f05fcd1a878 · outbound

This paper cites Towards Learning a Uni- versal Non-Semantic Representation of Speech,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Towards Learning a Uni- versal Non-Semantic Representation of Speech,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.285790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:26:15.019483Z digest=sha256:54801d6777e19397240a2204451596d94c45138f6aa8b15d87f57c6d6cfc7494

Observation 7fdff210-619f-4579-9761-6314246933fe · outbound

This paper cites Unsupervised learn- ing of semantic audio representations,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Unsupervised learn- ing of semantic audio representations,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.275794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:26:15.032551Z digest=sha256:db2d303d27b07b52ef72be7cee16f039009480c73a3717278d76dab769a74a4c

Observation 1c684355-4e1b-416c-98a5-c0a4e876d0fc · outbound

This paper cites Contrastive learning of general-purpose audio representations,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Contrastive learning of general-purpose audio representations,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.266366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:26:15.038803Z digest=sha256:434c4265a1bf88c031f099b9b5991ab394d53ece36ec54ca6ec99b6f64e7e9fe

Observation a1f249f7-88e6-4ee0-ad61-eeb72e973e28 · outbound

This paper cites Speech Commands: A Dataset for Limited- V ocabulary Speech Recognition,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Speech Commands: A Dataset for Limited- V ocabulary Speech Recognition,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.256554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:26:15.045522Z digest=sha256:74f10722bad7e5db3a8b3316ad88874680ce890013e792e391c3dc050c37125f

Observation 2bde6d46-b343-4d3f-9d5f-97ad3700ed27 · outbound

This paper cites Detection and classification of acoustic scenes and events: Outcome of the DCASE 2016 challenge,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Detection and classification of acoustic scenes and events: Outcome of the DCASE 2016 challenge,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.246276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:26:15.050539Z digest=sha256:dc9da3cd02d09a18a3ff8116382784cb67fa5dfc8382a904738f2b25ea4a9289

Observation c18c0708-47a8-4b94-8ee3-6b515452dd27 · outbound

This paper cites Sound event de- tection in synthetic audio: Analysis of the DCASE 2016 task results,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Sound event de- tection in synthetic audio: Analysis of the DCASE 2016 task results,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.236043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:26:15.058282Z digest=sha256:94078fb7071353e73fcb5fa61e78b83e4795169a6a1c502a57345f7335103e38

Observation 85cd9a4a-1503-44f9-929f-a47320bc853d · outbound

This paper cites Neural audio syn- thesis of musical notes with wavenet autoencoders,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Neural audio syn- thesis of musical notes with wavenet autoencoders,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.225323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:26:15.064430Z digest=sha256:9f78ae6ba1c84b9529eeaf83e18a23ee1b30d795b08721267ae8a221e6f279d6

Observation 2ab1526b-ed67-48f7-b827-42d9e438671a · outbound

This paper cites Crepe: A convo- lutional representation for pitch estimation,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Crepe: A convo- lutional representation for pitch estimation,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.214614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:26:15.071478Z digest=sha256:e208e260d06d4a1d4706e0f7833b82fcce3dc69c44247fe0022c5ca99ef8be67

Observation fef3c815-c64d-4fda-8226-fe579a862e91 · outbound

This paper cites EfficientNet: Rethinking model scaling for convolutional neural networks,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation EfficientNet: Rethinking model scaling for convolutional neural networks,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.204775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:26:15.075322Z digest=sha256:2d377ed319a85fb9d9b1257313f434de3547dd4acedc7d358271ef54c009b682

Observation ee623c89-baf4-4df3-afc2-a2cc622df993 · outbound

This paper cites Layer Normalization.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Layer Normalization

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:26:15.079404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:26:15.079404Z digest=sha256:1a7818042edbd65322b929268019519d1f115b2abb1b6952ba2bfb561591e7f6

Observation 1a08c4ff-ea39-4cda-afa8-d243f55b969a · outbound

This paper cites Adam: A method for stochastic op- timization,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Adam: A method for stochastic op- timization,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.193151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:26:15.083555Z digest=sha256:8eaf65f808d4e817743db9024f6673039b06d18ffac39f381b35bdc60667d9eb

Observation d8909262-716a-4fa9-8160-8cc1eeb391d2 · outbound

This paper cites Tagliasacchi, B.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Tagliasacchi, B

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.181985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:26:15.087529Z digest=sha256:9ce784cae1d1645e167890f838befb62f490ee26b51f0b95a15283dfcaa78928

Observation 9ec2a0ce-34a3-4a42-91b9-06c3e843fc90 · outbound

This paper cites A simple framework for contrastive learning of visual representations,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation A simple framework for contrastive learning of visual representations,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.170421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:26:15.091212Z digest=sha256:3535c07617bf494f777a25e6e08c5d9e4b5698bdb00ae75982191b91507bf54c

Observation 0a1819fe-7103-42a9-a291-f07153c04876 · outbound

This paper cites Seeing voices and hearing voices: Learning discriminative embeddings us- ing cross-modal self-supervision,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Seeing voices and hearing voices: Learning discriminative embeddings us- ing cross-modal self-supervision,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.157936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:26:15.094318Z digest=sha256:db8d7e4f08cd2607ece012d0c080c6c8fe72e84c5cb4278a5c135043b3805c7d

Pith citing papers

Observation 48b46bf4-b2c4-416a-ba41-4483c963d721 · inbound

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation cites this paper.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Self-supervised learning method using multiple sampling strategies for general-purpose audio representation

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T14:26:15.146229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:26:13.455801Z digest=sha256:0a402f95b9bd126784dd9912fd4df46723902c91427271fce3b254ab7823280f