Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-08T19:18:56.334254Z
Paper Citation Record · LEDGER
As of 23 August 2026, this Paper Citation Record lists 18 of 18 outbound references and 0 inbound Pith citation observations for arXiv:2607.05971.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-08T19:18:56.334254Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
18 of 18 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 259b61b3-d44e-4231-bb7a-c2add02cbdb5 · outbound
Multimodal Video-to-Music Recommendation via Semantic Retrieval and Temporal Reranking Qwen3-VL Technical Report
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 7ebe0395-9509-48d8-92db-6cd7385b7699 · outbound
Multimodal Video-to-Music Recommendation via Semantic Retrieval and Temporal Reranking PianoBind: A Multimodal Joint Embedding Model for Pop-piano Music
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation e97f020d-7ff6-49d1-bfc2-b5c8bb65b8ab · outbound
Multimodal Video-to-Music Recommendation via Semantic Retrieval and Temporal Reranking MMTrail: A Multimodal Trailer Video Dataset with Language and Music Descriptions
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 0f43de9d-3a28-49f7-bcda-b161e1b38bc6 · outbound
Multimodal Video-to-Music Recommendation via Semantic Retrieval and Temporal Reranking Listen, Read, and Identify: Multimodal Singing Language Identification of Music
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 62e5769b-1db1-4242-b4a5-304f1d3eec2a · outbound
Multimodal Video-to-Music Recommendation via Semantic Retrieval and Temporal Reranking LP-MusicCaps: LLM-Based Pseudo Music Captioning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation df24de1f-78d8-4daf-97ae-00ce9be4a859 · outbound
Multimodal Video-to-Music Recommendation via Semantic Retrieval and Temporal Reranking TALKPLAY: Multimodal Music Recommendation with Large Language Models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 106044c0-c970-42c3-b7b0-55d5f5e293d5 · outbound
Multimodal Video-to-Music Recommendation via Semantic Retrieval and Temporal Reranking Music Flamingo: Scaling music understanding in audio language models
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 11b251ca-0c62-4049-bc12-f094ec43e4bc · outbound
Multimodal Video-to-Music Recommendation via Semantic Retrieval and Temporal Reranking Audioclip: Extending clip to image, text and audio
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 63d8c160-3af7-4b67-911a-0382a29a9226 · outbound
Multimodal Video-to-Music Recommendation via Semantic Retrieval and Temporal Reranking Poly-encoders: Transformer Architectures and Pre-training Strategies for Fast and Accurate Multi-sentence Scoring
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 24cca5ac-b8cd-433b-a52c-17b234aa06ae · outbound
Multimodal Video-to-Music Recommendation via Semantic Retrieval and Temporal Reranking Qwen3-ASR Technical Report
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 54100690-59be-4a42-be60-535360c2eac1 · outbound
Multimodal Video-to-Music Recommendation via Semantic Retrieval and Temporal Reranking Meta Audiobox Aesthetics: Unified Automatic Quality Assessment for Speech, Music, and Sound
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 6446a29d-988b-4712-9e70-91a565430759 · outbound
Multimodal Video-to-Music Recommendation via Semantic Retrieval and Temporal Reranking SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation f3493831-43e7-471f-b24c-1c4ba103c559 · outbound
Multimodal Video-to-Music Recommendation via Semantic Retrieval and Temporal Reranking Pushing the frontier of audiovisual perception with large-scale multimodal correspondence learning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 93fc22e5-8c88-4a8e-8711-d2b5fcc2d0e9 · outbound
Multimodal Video-to-Music Recommendation via Semantic Retrieval and Temporal Reranking Unresolved cited work
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 40586d8e-e555-4f52-864b-102e45a0dc0a · outbound
Multimodal Video-to-Music Recommendation via Semantic Retrieval and Temporal Reranking CLaMP 3: Universal Music Information Retrieval Across Unaligned Modalities and Unseen Languages
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation c1a56a38-f603-44a6-a0c4-d078f2299d2f · outbound
Multimodal Video-to-Music Recommendation via Semantic Retrieval and Temporal Reranking Qwen3 Technical Report
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation afb22485-52a9-41b0-8329-d4f9cdbb7685 · outbound
Multimodal Video-to-Music Recommendation via Semantic Retrieval and Temporal Reranking C-Pack: Packed Resources For General Chinese Embeddings
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 00597ec6-990a-441e-9ce9-3dc155acaa75 · outbound
Multimodal Video-to-Music Recommendation via Semantic Retrieval and Temporal Reranking Languagebind: Extending video-language pretraining to n-modality by language-based semantic alignment
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
No inbound Pith citation observations are available.