Pith. sign in

Paper Citation Record · LEDGER

CoS: Chain-of-Shot Prompting for Long Video Understanding

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2502.06428.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.06428 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:20:57.091933Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T07:54:22.417988Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3beebf5b-6499-41ad-95b5-4d720e775111 · inbound

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey cites this paper.

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey CoS: Chain-of-Shot Prompting for Long Video Understanding

Reference 112

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T17:18:53.141676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T17:18:52.996467Z digest=sha256:9ec15ec84985c5c147077b9a016dab452808122e2fca7f2d102a7d70d504d013

Observation db2ab293-089b-4639-b534-f9246a05f5a8 · inbound

ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning cites this paper.

ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning CoS: Chain-of-Shot Prompting for Long Video Understanding

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T15:20:57.091933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:20:57.091933Z digest=sha256:d9507253f26bed6b5b9672e51c3a350bbd3f537a9b7dd06e94ec56ddb77fd006

Observation 86481275-9056-4f32-9893-441e6d675a85 · inbound

CyberV: Cybernetics for Test-time Scaling in Video Understanding cites this paper.

CyberV: Cybernetics for Test-time Scaling in Video Understanding CoS: Chain-of-Shot Prompting for Long Video Understanding

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T05:26:47.111365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:26:47.111365Z digest=sha256:6671cc2767b21d47b13a811620adee7be87f317e00c3c91bc36d2b63c37e4df1

Observation 2608cca2-ea4e-443e-af2f-9c56d48093d5 · inbound

ViTCoT: Video-Text Interleaved Chain-of-Thought for Boosting Video Understanding in Large Language Models cites this paper.

ViTCoT: Video-Text Interleaved Chain-of-Thought for Boosting Video Understanding in Large Language Models CoS: Chain-of-Shot Prompting for Long Video Understanding

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T17:49:34.796119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:49:34.796119Z digest=sha256:f013e94327ec58f7c9c61e09ea2841af1c5ed7da1d4bac4fe13546a8468094f4

Observation 9419f8c6-e79c-4b8a-87e9-0244b78e9076 · inbound

Uncertainty-quantified Rollout Policy Adaptation for Unlabelled Cross-domain Temporal Grounding cites this paper.

Uncertainty-quantified Rollout Policy Adaptation for Unlabelled Cross-domain Temporal Grounding CoS: Chain-of-Shot Prompting for Long Video Understanding

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T22:54:32.335081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:54:32.335081Z digest=sha256:27908ae60b341c49643213ca2ae2566036f09caa3a9fa3b3829bc0e8faefb348

Observation 43fd0233-742b-426d-9870-781e792e941d · inbound

Empowering Multimodal LLMs with External Tools: A Comprehensive Survey cites this paper.

Empowering Multimodal LLMs with External Tools: A Comprehensive Survey CoS: Chain-of-Shot Prompting for Long Video Understanding

Reference 160

Resolution
unresolved
no resolver link, observed 2026-08-05T20:28:58.207978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:28:58.207978Z digest=sha256:9f6509eeab2b67bddc6741be68891d8dde92018a1c7f9f1714d5a888f042040c

Observation 514acf4c-28ec-443f-9c54-f96477d49094 · inbound

AdsQA: Towards Advertisement Video Understanding cites this paper.

AdsQA: Towards Advertisement Video Understanding CoS: Chain-of-Shot Prompting for Long Video Understanding

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T20:20:36.731548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:20:36.731548Z digest=sha256:54d016a643cfdb550998777184dcf8037061af0263fe25c01fd1209882f09968

Observation 5985101b-e687-4c51-ad81-83573ef7fe84 · inbound

Act2See: Emergent Active Visual Perception for Video Reasoning cites this paper.

Act2See: Emergent Active Visual Perception for Video Reasoning CoS: Chain-of-Shot Prompting for Long Video Understanding

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T05:45:23.124557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T19:34:53.683729Z digest=sha256:887ff50d37f3972ef919f7209b2d3b15c0f9425d3f727c093e26df508d73cd34

Observation 5ed05f66-4101-4cf6-910a-c0112a66b507 · inbound

PointLLM-R: Enhancing 3D Point Cloud Reasoning via Chain-of-Thought cites this paper.

PointLLM-R: Enhancing 3D Point Cloud Reasoning via Chain-of-Thought CoS: Chain-of-Shot Prompting for Long Video Understanding

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:31:14.211557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T07:26:17.850942Z digest=sha256:86400738c2a3abb481870de9e593dd3de59a5468b6531d7d1217356318897e03

Observation 27bcad07-30e0-45fa-b29d-75c5862742ff · inbound

Swift Sampling: Selecting Temporal Surprises via Taylor Series cites this paper.

Swift Sampling: Selecting Temporal Surprises via Taylor Series CoS: Chain-of-Shot Prompting for Long Video Understanding

Reference 46

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T05:56:07.910455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T05:55:23.479344Z digest=sha256:39707ea5a51d84b35954e612e0b1b684bcbec07f6cb00f5acd67cac45f23ecad

Observation a3bf1934-716b-4c9b-a41b-98b311482ab2 · inbound

Bridging VideoQA and Video-Guided Agentic Tasks via Generalized Keyframe Extraction cites this paper.

Bridging VideoQA and Video-Guided Agentic Tasks via Generalized Keyframe Extraction CoS: Chain-of-Shot Prompting for Long Video Understanding

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T07:54:22.419385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T07:48:01.719339Z digest=sha256:2174d1e3e4c4ef26bc5428485a69c730bedf3e28def6c58fa1862f1ca82e8d22

Observation 023e1e81-9023-4cee-9360-1e5ca636e1a5 · inbound

Efficient Frame Selection for Long Videos at Test Time with Attention-Based MLLM Selectors cites this paper.

Efficient Frame Selection for Long Videos at Test Time with Attention-Based MLLM Selectors CoS: Chain-of-Shot Prompting for Long Video Understanding

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T22:38:42.688411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T22:38:42.688411Z digest=sha256:8b485fa6749f90d6a7a5208dd2cdb4cef25b877e3416e67aa04148d156985bc5