Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T01:07:11.155102Z
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 20 of 20 outbound references and 0 inbound Pith citation observations for arXiv:2505.02096.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T01:07:11.155102Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
20 of 20 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation a7dd5de0-2dab-41e4-a835-14a405e210d3 · outbound
TeMTG: Text-Enhanced Multi-Hop Temporal Graph Modeling for Audio-Visual Video Parsing Unified multisensory perception: Weakly-supervised audio-visual video parsing
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 021b902c-6d23-4de5-b802-8a5810a5ba1f · outbound
TeMTG: Text-Enhanced Multi-Hop Temporal Graph Modeling for Audio-Visual Video Parsing Anchor-aware Deep Metric Learning for Audio-visual Retrieval
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation a3492174-f6e7-4e7d-9f35-a25b1b3c0052 · outbound
TeMTG: Text-Enhanced Multi-Hop Temporal Graph Modeling for Audio-Visual Video Parsing Boosting Audio Visual Question Answer- ing via Key Semantic-Aware Cues
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation add480ef-be21-4bf4-9aa9-760545df8d74 · outbound
TeMTG: Text-Enhanced Multi-Hop Temporal Graph Modeling for Audio-Visual Video Parsing Open-Vocabulary Audio-Visual Semantic Segmentation
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 58e56d40-8fb5-4b30-a1ea-3aa846ee1c41 · outbound
TeMTG: Text-Enhanced Multi-Hop Temporal Graph Modeling for Audio-Visual Video Parsing Drcnet: Dynamic image restoration contrastive network
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e0ab1285-db4f-42fb-80d3-4c5d3154ca63 · outbound
TeMTG: Text-Enhanced Multi-Hop Temporal Graph Modeling for Audio-Visual Video Parsing Collecting cross-modal presence-absence evidence for weakly-supervised audio-visual event percep- tion
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation dae141fa-5161-48f0-ac2f-5ad330576e3f · outbound
TeMTG: Text-Enhanced Multi-Hop Temporal Graph Modeling for Audio-Visual Video Parsing ColeaF: A Contrastive-Collaborative Learning Framework for Weakly Supervised Audio- Visual Video Parsing
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b7ea8ed5-813f-4743-a7c7-4aaa4fd7f381 · outbound
TeMTG: Text-Enhanced Multi-Hop Temporal Graph Modeling for Audio-Visual Video Parsing Modality-independent teachers meet weakly-supervised audio-visual event parser
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 3046f933-f1a6-488f-89cc-edad067283ab · outbound
TeMTG: Text-Enhanced Multi-Hop Temporal Graph Modeling for Audio-Visual Video Parsing Large-scale contrastive language-audio pretraining with feature fusion and keyword-to-caption augmentation
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c576de82-eff6-4850-ac0a-16b5c962a0ed · outbound
TeMTG: Text-Enhanced Multi-Hop Temporal Graph Modeling for Audio-Visual Video Parsing Learning transferable visual models from natural language supervision
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc383dae-d85b-413c-9e92-fd23c178ec86 · outbound
TeMTG: Text-Enhanced Multi-Hop Temporal Graph Modeling for Audio-Visual Video Parsing Label-anticipated event disentanglement for audio-visual video parsing
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b2387bf3-1c95-4779-86e2-6a98eb59c3f4 · outbound
TeMTG: Text-Enhanced Multi-Hop Temporal Graph Modeling for Audio-Visual Video Parsing Revisit weakly-supervised audio- visual video parsing from the language perspective
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c0f97708-68f5-42ec-9bf9-e3c403287f8f · outbound
TeMTG: Text-Enhanced Multi-Hop Temporal Graph Modeling for Audio-Visual Video Parsing Multi-modal grouping network for weakly- supervised audio-visual video parsing
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation a522badc-0009-4868-8061-94284f83e6c2 · outbound
TeMTG: Text-Enhanced Multi-Hop Temporal Graph Modeling for Audio-Visual Video Parsing CM-PIE: Cross-modal perception for interactive-enhanced audio- visual video parsing
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 205d1c3f-f0ce-4a7c-a39f-29915a5f5e2a · outbound
TeMTG: Text-Enhanced Multi-Hop Temporal Graph Modeling for Audio-Visual Video Parsing Advancing Weakly- Supervised Audio-Visual Video Parsing via Segment-Wise Pseudo Labeling
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 6f958f28-13b4-4bd8-a62b-2ed47f291f51 · outbound
TeMTG: Text-Enhanced Multi-Hop Temporal Graph Modeling for Audio-Visual Video Parsing Resisting Noise in Pseudo Labels: Audible Video Event Parsing With Evidential Learning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 5ece17e2-c5a1-4539-b62f-01df159f5a7a · outbound
TeMTG: Text-Enhanced Multi-Hop Temporal Graph Modeling for Audio-Visual Video Parsing LINK: Adaptive Modality Interaction for Audio-Visual Video Parsing
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 5270ffcd-4953-4a08-bf54-75c0b06c962e · outbound
TeMTG: Text-Enhanced Multi-Hop Temporal Graph Modeling for Audio-Visual Video Parsing Multilayer perceptron (MLP)
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation ee8ed6a6-d325-4e62-8c34-e0eae99be691 · outbound
TeMTG: Text-Enhanced Multi-Hop Temporal Graph Modeling for Audio-Visual Video Parsing Graph Attention Networks
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 7a03950c-6d74-458c-bc6f-f95aa5fb23ae · outbound
TeMTG: Text-Enhanced Multi-Hop Temporal Graph Modeling for Audio-Visual Video Parsing Exploring heterogeneous clues for weakly-supervised audio- visual video parsing
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
No inbound Pith citation observations are available.