Pith. sign in

Paper Citation Record · LEDGER

SAVEn-Vid: Synergistic Audio-Visual Integration for Enhanced Understanding in Long Video Context

As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 2 inbound Pith citation observations for arXiv:2411.16213.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.16213 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 2 of 2 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T16:56:06.220677Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T05:49:54.210292Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 8ad76a10-5bb5-430a-a9fb-20d3772ac61f · inbound

InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System for Long-term Streaming Video and Audio Interactions cites this paper.

InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System for Long-term Streaming Video and Audio Interactions SAVEn-Vid: Synergistic Audio-Visual Integration for Enhanced Understanding in Long Video Context

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-11T16:56:06.220677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:56:06.220677Z digest=sha256:2d05b534cc364fdf234721c312493fb233bbee5ff31da6a1b62573f256319bf1

Observation 74a42b6f-062f-4581-8e7b-74b0436b721c · inbound

MAGNET: A Multi-agent Framework for Finding Audio-Visual Needles by Reasoning over Multi-Video Haystacks cites this paper.

MAGNET: A Multi-agent Framework for Finding Audio-Visual Needles by Reasoning over Multi-Video Haystacks SAVEn-Vid: Synergistic Audio-Visual Integration for Enhanced Understanding in Long Video Context

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:49:54.215338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:49:53.299116Z digest=sha256:2ec956d1c5499362d0016809d7e8fb8e5fcb61f064afcb49da1b4406482c4d0d