Pith. sign in

Paper Citation Record · LEDGER

AdaReTaKe: Adaptive Redundancy Reduction to Perceive Longer for Video-language Understanding

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2503.12559.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.12559 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:59:09.407693Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 4a88b79f-f459-423e-9304-afec24aed3d6 · inbound

LVBench: An Extreme Long Video Understanding Benchmark cites this paper.

LVBench: An Extreme Long Video Understanding Benchmark AdaReTaKe: Adaptive Redundancy Reduction to Perceive Longer for Video-language Understanding

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:55:30.118112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-19T11:55:30.048525Z digest=sha256:482b6991efb55e14e705973d730e412e398366030dcd3a6f98634612e9beba97

Observation 59d56e0e-84dd-4318-9897-74785c7466c4 · inbound

FlexSelect: Flexible Token Selection for Efficient Long Video Understanding cites this paper.

FlexSelect: Flexible Token Selection for Efficient Long Video Understanding AdaReTaKe: Adaptive Redundancy Reduction to Perceive Longer for Video-language Understanding

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T11:59:09.407693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:59:09.407693Z digest=sha256:a7c9d0cf4aa047dbd140247fcee041c3db5b2830826879793456ab542c84302e

Observation ca1a1c3b-6481-4ff9-a327-12cf066bf2a2 · inbound

CoTasks: Chain-of-Thought based Video Instruction Tuning Tasks cites this paper.

CoTasks: Chain-of-Thought based Video Instruction Tuning Tasks AdaReTaKe: Adaptive Redundancy Reduction to Perceive Longer for Video-language Understanding

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T16:24:41.591387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:24:41.591387Z digest=sha256:9cd7d6f6cdaf1543d477a0e568a2229be16b473a2907cabe23f38b5492e61a91

Observation da5f9203-6ff9-487b-ba43-122aec4747f1 · inbound

HumanPCR: Probing MLLM Capabilities in Diverse Human-Centric Scenes cites this paper.

HumanPCR: Probing MLLM Capabilities in Diverse Human-Centric Scenes AdaReTaKe: Adaptive Redundancy Reduction to Perceive Longer for Video-language Understanding

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-05T19:03:07.167305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:03:07.167305Z digest=sha256:fb31b0b72252b2e5f6ce9e5ab068b3f40eecccb602a0494c9fffdba9c31eef2d

Observation c2c22972-8ae0-4bfd-8dab-8065d9bb208b · inbound

DATE: Dynamic Absolute Time Enhancement for Long Video Understanding cites this paper.

DATE: Dynamic Absolute Time Enhancement for Long Video Understanding AdaReTaKe: Adaptive Redundancy Reduction to Perceive Longer for Video-language Understanding

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-04T19:28:34.417628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:28:34.417628Z digest=sha256:5e8f8ac4ab715007706a0691d815140850176cc7c3dd18aba468d527767b77ce

Observation 80731edd-f564-4938-8d2b-753f8092a61a · inbound

Mosaic: Cross-Modal Clustering for Efficient Video Understanding cites this paper.

Mosaic: Cross-Modal Clustering for Efficient Video Understanding AdaReTaKe: Adaptive Redundancy Reduction to Perceive Longer for Video-language Understanding

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-10T16:10:34.407672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T16:07:36.133404Z digest=sha256:2fe7f204d5efe3784766152fba351f9ffeeb27c035dded623c70525230cd11af

Observation 659a6dfb-9521-4811-ab31-a4ca266faee6 · inbound

Searching Videos as Trees: Self-Correcting Agents for Grounded Long Video QA cites this paper.

Searching Videos as Trees: Self-Correcting Agents for Grounded Long Video QA AdaReTaKe: Adaptive Redundancy Reduction to Perceive Longer for Video-language Understanding

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-01T21:09:44.952695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T21:09:44.952695Z digest=sha256:5fa9ed4184c555ce1ce3e26d98d3d03470d8498b60dea6696405594b07bd0178