Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T04:29:30.849620Z
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2412.01356.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T04:29:30.849620Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
28 of 28 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation d54ccc7c-78ef-43de-891d-e2760bac62ad · outbound
Text-based Audio Retrieval by Learning from Similarities between Audio Captions Language-Based Audio Retrieval Task in DCASE 2022 Challenge,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation f62bbf35-64f5-40d8-8267-bee3aae09633 · outbound
Text-based Audio Retrieval by Learning from Similarities between Audio Captions Improving Natural-Language-Based Audio Retrieval with Transfer Learning and Audio & Text Augmentations,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation d3d6a855-df62-41aa-a998-ef08512a331f · outbound
Text-based Audio Retrieval by Learning from Similarities between Audio Captions Matching Text and Audio Embeddings: Exploring Transfer-Learning Strategies for Language-Based Audio Retrieval,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 932fcd97-c9a4-4002-baf7-e022f3b1eb62 · outbound
Text-based Audio Retrieval by Learning from Similarities between Audio Captions Large-Scale Contrastive Language-Audio Pretraining with Feature Fusion and Keyword-to-Caption Augmentation,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation dd3dbe2f-db55-4696-9997-d2104e6f41f1 · outbound
Text-based Audio Retrieval by Learning from Similarities between Audio Captions WavCaps: A ChatGPT-Assisted Weakly-Labelled Audio Captioning Dataset for Audio-Language Multimodal Research,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation cffba290-5e42-4089-9ee6-0a04c523689f · outbound
Text-based Audio Retrieval by Learning from Similarities between Audio Captions Auto-ACD: A Large-scale Dataset for Audio-Language Representation Learning,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 8fb02410-cb4e-4d2d-954d-174335f52f97 · outbound
Text-based Audio Retrieval by Learning from Similarities between Audio Captions Advancing Natural-Language Based Audio Retrieval with Passt and Large Audio-Caption Data Sets,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation b2fab24d-c5ba-4752-a01a-a134a45c827c · outbound
Text-based Audio Retrieval by Learning from Similarities between Audio Captions Language-based Audio Retrieval in DCASE 2023 Challenge,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 4f6547c7-2fa3-4bb2-bc93-b781b1605063 · outbound
Text-based Audio Retrieval by Learning from Similarities between Audio Captions AudioCaps: Generating Cap- tions for Audios in The Wild,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 2a84f74e-62bf-48c7-a417-791813cf2f8a · outbound
Text-based Audio Retrieval by Learning from Similarities between Audio Captions Clotho: an Audio Captioning Dataset,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 73cdbc5b-87b8-47e1-88fd-b13e44179738 · outbound
Text-based Audio Retrieval by Learning from Similarities between Audio Captions Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 293b2e9f-5f5c-4ff6-a492-e3cd7cd04e0a · outbound
Text-based Audio Retrieval by Learning from Similarities between Audio Captions Graded Relevance,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation f75ef85d-f33b-4c95-ab54-88d4ceac6072 · outbound
Text-based Audio Retrieval by Learning from Similarities between Audio Captions On the effect of relevance scales in crowdsourcing relevance assessments for Information Retrieval evaluation,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 5416fdc1-b405-4654-90bb-878c841f2856 · outbound
Text-based Audio Retrieval by Learning from Similarities between Audio Captions Crowdsourcing and Evaluating Text-Based Audio Retrieval Relevances,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation b8bb1445-9b22-41b0-9c59-ca365e320b1c · outbound
Text-based Audio Retrieval by Learning from Similarities between Audio Captions Integrating Continuous and Binary Relevances in Audio-Text Relevance Learning
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 3e037dc8-4f0e-47e6-a502-fa7717086dbd · outbound
Text-based Audio Retrieval by Learning from Similarities between Audio Captions Estimated Audio-Caption Correspondences Improve Language-Based Audio Retrieval
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation dd0aa89d-c1d4-4730-9ad4-80e70c453f0b · outbound
Text-based Audio Retrieval by Learning from Similarities between Audio Captions Learning to Rank: From Pairwise Approach to Listwise Approach,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation f718f14e-6ca7-4604-a238-c06c5766a496 · outbound
Text-based Audio Retrieval by Learning from Similarities between Audio Captions Image-Text Retrieval with Binary and Continuous Label Supervision
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 6ac29d83-9077-4ee8-be3d-b052e91e8680 · outbound
Text-based Audio Retrieval by Learning from Similarities between Audio Captions Integrating Listwise Ranking into Pairwise-based Image-Text Retrieval,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 9d1228eb-5555-42f0-bd00-00002d6f57de · outbound
Text-based Audio Retrieval by Learning from Similarities between Audio Captions Freesound Technical Demo,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 63af5136-c1ae-4ed2-81f3-692e514dc221 · outbound
Text-based Audio Retrieval by Learning from Similarities between Audio Captions BBC Sound Effects,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 4cf472e2-8dfe-4c29-8ac7-a9f3dafcd981 · outbound
Text-based Audio Retrieval by Learning from Similarities between Audio Captions SoundBible,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 0c1be653-1ff1-4e4f-a461-af16978f09c4 · outbound
Text-based Audio Retrieval by Learning from Similarities between Audio Captions The Benefit of Temporally-Strong Labels in Audio Event Classification,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 3921b18b-826e-4cc7-bdf9-bd3c10dfc271 · outbound
Text-based Audio Retrieval by Learning from Similarities between Audio Captions Language-based Audio Retrieval in DCASE 2024 Challenge,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation a1ff3e74-046f-4aea-a084-2ef2fc42d95a · outbound
Text-based Audio Retrieval by Learning from Similarities between Audio Captions Efficient Training of Audio Transformers with Patchout,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 98763c99-d3e5-469e-a47e-4cb60febaaf5 · outbound
Text-based Audio Retrieval by Learning from Similarities between Audio Captions RoBERTa: A Robustly Optimized BERT Pretraining Approach
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e1ff558-e9ed-4098-aea2-61ba95c9401a · outbound
Text-based Audio Retrieval by Learning from Similarities between Audio Captions Representation Learning with Contrastive Predictive Coding
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6a3a6c6-771b-40d9-8e12-631211197e55 · outbound
Text-based Audio Retrieval by Learning from Similarities between Audio Captions SGDR: Stochastic Gradient Descent with Warm Restarts,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
No inbound Pith citation observations are available.