Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T20:21:45.248236Z
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 0 inbound Pith citation observations for arXiv:2412.05831.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T20:21:45.248236Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
24 of 24 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation bb449eaf-2301-4268-9367-95989e861735 · outbound
Semi-Supervised Contrastive Learning for Controllable Video-to-Music Retrieval Cbvmr: Content-based video-music retrieval using soft intra-modal structure constraint,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f5f4363e-4e16-439a-801a-a63c9e3840f3 · outbound
Semi-Supervised Contrastive Learning for Controllable Video-to-Music Retrieval Cross-modal music-video recommendation: A study of design choices,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c6af6665-5c07-4a42-82ee-5b0b736a97d4 · outbound
Semi-Supervised Contrastive Learning for Controllable Video-to-Music Retrieval Is there a
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 27449000-beef-427e-9282-e4b6e83e1ee6 · outbound
Semi-Supervised Contrastive Learning for Controllable Video-to-Music Retrieval It’s time for artistic correspondence in music and video,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 38a289ed-55e0-4c7a-adfc-7045d6dcdabc · outbound
Semi-Supervised Contrastive Learning for Controllable Video-to-Music Retrieval Attention is all you need,
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87c840bb-2d31-4861-a70e-572cd4012f7f · outbound
Semi-Supervised Contrastive Learning for Controllable Video-to-Music Retrieval Language-guided music recommendation for video via prompt analogies,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 9cd9c6ce-5e02-4100-a3df-2f65224f4608 · outbound
Semi-Supervised Contrastive Learning for Controllable Video-to-Music Retrieval Mulan: A joint embedding of music audio and natural language,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 839b6c45-13b0-48c1-aa9c-b6ca7a82b65d · outbound
Semi-Supervised Contrastive Learning for Controllable Video-to-Music Retrieval Contrastive audio- language learning for music,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 46f0e205-ad4a-4668-a746-d986377aff5c · outbound
Semi-Supervised Contrastive Learning for Controllable Video-to-Music Retrieval Large-scale contrastive language-audio pretraining with feature fusion and keyword-to-caption augmentation,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation ce8764b0-7571-4890-b2c9-5126d060ba8b · outbound
Semi-Supervised Contrastive Learning for Controllable Video-to-Music Retrieval Textless speech-to-music retrieval using emotion similarity,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 246ff8b7-8352-43c2-a5d1-7b7e745d4d3e · outbound
Semi-Supervised Contrastive Learning for Controllable Video-to-Music Retrieval Emotion-aligned contrastive learning between images and music,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation aa6e237d-28f0-4c98-b99e-ea898419faf1 · outbound
Semi-Supervised Contrastive Learning for Controllable Video-to-Music Retrieval Bridging high-quality audio and video via language for sound effects retrieval from visual queries,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 87f43bf2-7912-433e-92ed-3b47121c36dd · outbound
Semi-Supervised Contrastive Learning for Controllable Video-to-Music Retrieval Query by video: Cross-modal music retrieval,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 8a1d4109-64f7-419f-a3c5-fae3900c02f8 · outbound
Semi-Supervised Contrastive Learning for Controllable Video-to-Music Retrieval Mert: Acoustic music understanding model with large-scale self-supervised training,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 5315cc00-07e9-4ec6-bf7b-31b9f8fee674 · outbound
Semi-Supervised Contrastive Learning for Controllable Video-to-Music Retrieval Learning transferable visual models from natural language supervision,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 9a90eb1a-149d-486b-b3c9-d111a5cbc16b · outbound
Semi-Supervised Contrastive Learning for Controllable Video-to-Music Retrieval Representation Learning with Contrastive Predictive Coding
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c4248c1-2fea-4f18-a459-441c48197366 · outbound
Semi-Supervised Contrastive Learning for Controllable Video-to-Music Retrieval Supervised contrastive learn- ing,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation dc021c04-d0be-413c-92b5-28d481b10248 · outbound
Semi-Supervised Contrastive Learning for Controllable Video-to-Music Retrieval Audio set: An ontology and human-labeled dataset for audio events,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e1bb36fb-4eb3-4685-8bea-62196cbe658e · outbound
Semi-Supervised Contrastive Learning for Controllable Video-to-Music Retrieval YouTube-8M: A Large-Scale Video Classification Benchmark
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d33b1025-da5c-431d-bb30-330bd096f8ac · outbound
Semi-Supervised Contrastive Learning for Controllable Video-to-Music Retrieval Decoupled weight decay regularization,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation cebd0518-2118-4e0b-907d-5f280b6fcbe6 · outbound
Semi-Supervised Contrastive Learning for Controllable Video-to-Music Retrieval Wav2clip: Learning robust audio representations from clip,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e8927204-5729-4347-aa63-2d6b62883056 · outbound
Semi-Supervised Contrastive Learning for Controllable Video-to-Music Retrieval Audioclip: Extending clip to image, text and audio,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 056ff467-4796-400f-9baa-23b753fcbc4d · outbound
Semi-Supervised Contrastive Learning for Controllable Video-to-Music Retrieval Emotion embedding spaces for matching music to stories,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 27f1c0d7-b391-4ae3-8296-0644829dff74 · outbound
Semi-Supervised Contrastive Learning for Controllable Video-to-Music Retrieval Vggsound: A large- scale audio-visual dataset,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
No inbound Pith citation observations are available.