Pith. sign in

Paper Citation Record · LEDGER

Comparing Learning Paradigms for Egocentric Video Summarization

As of 19 August 2026, this Paper Citation Record lists 19 of 19 outbound references and 0 inbound Pith citation observations for arXiv:2506.21785.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.21785 v1

Coverage vector

measured 19 of 19 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:23:43.243274Z

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

19 of 19 outbound references displayed

  • verified exact4
  • verified fuzzy1
  • unresolved14
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9c8b10c0-3408-4d37-b3b5-e2fca4dcb349 · outbound

This paper cites UniVTG: Towards Unified Video-Language Temporal Grounding.

Comparing Learning Paradigms for Egocentric Video Summarization UniVTG: Towards Unified Video-Language Temporal Grounding

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-08-06T22:23:44.366891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:23:41.171290Z digest=sha256:e5aad1af86eeccffe2011e247cecc954388dd368fff9834a844ddf2602319064

Observation 90c0f3b1-5916-4edf-8bfe-b087c15b290a · outbound

This paper cites VideoLLM-online: Online Video Large Language Model for Streaming Video.

Comparing Learning Paradigms for Egocentric Video Summarization VideoLLM-online: Online Video Large Language Model for Streaming Video

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T22:23:41.274752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:23:41.274752Z digest=sha256:fb242869970de601fa662424cbd1d0047e50db5cfccb7f50526e61f68ee4c6f5

Observation 60fd24f9-7f8e-4d8d-bade-a4d8a55cc9a8 · outbound

This paper cites VideoMamba: State Space Model for Efficient Video Understanding.

Comparing Learning Paradigms for Egocentric Video Summarization VideoMamba: State Space Model for Efficient Video Understanding

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T22:23:41.396361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:23:41.396361Z digest=sha256:303276dba64dd58a23808699e1b03a8018e49d53c53d8931224f6581537b83bd

Observation bdf647d9-3f6f-4ac8-bb9d-c1c7908b92a9 · outbound

This paper cites Mamba: Linear-Time Sequence Modeling with Selective State Spaces.

Comparing Learning Paradigms for Egocentric Video Summarization Mamba: Linear-Time Sequence Modeling with Selective State Spaces

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T22:23:41.495758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:23:41.495758Z digest=sha256:7651ce54f326f70ac92eb40c2dbd312e39e88093545040820443c1172d86f4f9

Observation bd7e80c6-f1a6-4542-be12-6e0b683c42a7 · outbound

This paper cites Is Space-Time Attention All You Need for Video Understanding?.

Comparing Learning Paradigms for Egocentric Video Summarization Is Space-Time Attention All You Need for Video Understanding?

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T22:23:41.597659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:23:41.597659Z digest=sha256:e6c67b6d867f3dae2d4374ae34eeef5e53c55206e1f374870770db802d68fe7f

Observation 3ae6917d-ea86-4819-bad3-0222acd15ed4 · outbound

This paper cites Ego-Exo4D: Understanding Skilled Human Activity from First- and Third-Person Perspectives.

Comparing Learning Paradigms for Egocentric Video Summarization Ego-Exo4D: Understanding Skilled Human Activity from First- and Third-Person Perspectives

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T22:23:41.740633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:23:41.740633Z digest=sha256:c3644228a9ef9df79293b467872785e5064e52aa1d419390b7bde3ac6bed0d84

Observation 248d10b3-b2d1-4c0d-b58e-6c6220c682b4 · outbound

This paper cites Shotluck Holmes: A Family of Efficient Small-Scale Large Language Vision Models For Video Captioning and Summarization.

Comparing Learning Paradigms for Egocentric Video Summarization Shotluck Holmes: A Family of Efficient Small-Scale Large Language Vision Models For Video Captioning and Summarization

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-08-06T22:23:44.044371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:23:41.864051Z digest=sha256:c9307c2c1d3c170b0347541844d3313b23560700111d00977d6f07dae8ca5b7e

Observation e5aed2b2-69a5-42fe-87bd-b4dce7efe8ad · outbound

This paper cites Shot2Story: A New Benchmark for Comprehensive Understanding of Multi-shot Videos.

Comparing Learning Paradigms for Egocentric Video Summarization Shot2Story: A New Benchmark for Comprehensive Understanding of Multi-shot Videos

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T22:23:41.969901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:23:41.969901Z digest=sha256:d33f7fc495cf64e924854e230635d24514c22479332de50082f9a6d4df33bedb

Observation 5e87f939-0cdb-4a1c-bef7-4e88cf736c47 · outbound

This paper cites Advancing High-Resolution Video-Language Representation with Large-Scale Video Transcriptions.

Comparing Learning Paradigms for Egocentric Video Summarization Advancing High-Resolution Video-Language Representation with Large-Scale Video Transcriptions

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-06T22:23:43.860557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:23:42.126947Z digest=sha256:10ecf94156fb8dd3cf39f595943c912f2e0ff0458f70451e79447ec3857b489d

Observation 504e0a5d-bbdc-46a6-9602-16acb08fd1e5 · outbound

This paper cites TransNet V2: An effective deep network architecture for fast shot transition detection.

Comparing Learning Paradigms for Egocentric Video Summarization TransNet V2: An effective deep network architecture for fast shot transition detection

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T22:23:42.246470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:23:42.246470Z digest=sha256:e911745c90952b453bba85bf4b737e71f2978baa63d59f927fa471ef280d7d20

Observation 0335ffe1-ba69-4b56-90af-e15f5577fc2e · outbound

This paper cites Ego4D: Around the World in 3,000 Hours of Egocentric Video.

Comparing Learning Paradigms for Egocentric Video Summarization Ego4D: Around the World in 3,000 Hours of Egocentric Video

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T22:23:42.360607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:23:42.360607Z digest=sha256:ce1c8ea2e3b5a3f95b845394df6c70c6d2aba5304dc3fdbf1af94fe0ac04a960

Observation 2f91c3c5-a537-41c3-a653-86e22eebc81f · outbound

This paper cites From Sparse to Dense: GPT-4 Summarization with Chain of Density Prompting.

Comparing Learning Paradigms for Egocentric Video Summarization From Sparse to Dense: GPT-4 Summarization with Chain of Density Prompting

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-08-06T22:23:43.644497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:23:42.496883Z digest=sha256:6bcd8b22efa34434208426e1f0745691c4fea9498b1b573f1a30422c68d9163e

Observation ece7af2f-5675-4b0a-8fb9-4d57d673dfd8 · outbound

This paper cites GPT-4 Technical Report.

Comparing Learning Paradigms for Egocentric Video Summarization GPT-4 Technical Report

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T22:23:42.657182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:23:42.657182Z digest=sha256:02067295d047c136ddb4eef4ff00c15848eefdf8dd71ad7d8afdc77e78b7823c

Observation 998b9a53-fd36-4a54-a714-cae86b08c33a · outbound

This paper cites Cluster-based Video Summarization with Temporal Context Awareness.

Comparing Learning Paradigms for Egocentric Video Summarization Cluster-based Video Summarization with Temporal Context Awareness

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T22:23:42.816049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:23:42.816049Z digest=sha256:6431ed671cc017c32b9481e27cefdc20c17ddfd9785483c8c2c62f4938b1ec83

Observation a2d84fed-fae2-4413-bc87-bb203fad4ccf · outbound

This paper cites Creating Summaries from User Videos.

Comparing Learning Paradigms for Egocentric Video Summarization Creating Summaries from User Videos

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T22:23:42.917546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:23:42.917546Z digest=sha256:47a0bda8ca72f7e64b3a06b051bf87e48a9fa1b67b5fbe8dbfbedab6cabb834c

Observation 9d10f3e3-ba32-4d81-88d6-3cf103de00d9 · outbound

This paper cites Sigmoid Loss for Language Image Pre-Training.

Comparing Learning Paradigms for Egocentric Video Summarization Sigmoid Loss for Language Image Pre-Training

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T22:23:42.985153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:23:42.985153Z digest=sha256:2780a0c7b87c1041539e14bcd4a2569a627614ba3683044e09591953e0d4ae39

Observation e64a774c-1eea-46ea-987e-18ebdce0fd69 · outbound

This paper cites BIRCH: An Efficient Data Clustering Method for Very Large Databases.

Comparing Learning Paradigms for Egocentric Video Summarization BIRCH: An Efficient Data Clustering Method for Very Large Databases

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T22:23:43.101087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:23:43.101087Z digest=sha256:e708b95a10ab6570fd700ecf82bdd06219fbe8787646b2d3256221782895cc98

Observation 46b74914-c675-4384-9cba-1e2f60511c7d · outbound

This paper cites Chain-of-Thought Prompting Elicits Reasoning in Large Language Models.

Comparing Learning Paradigms for Egocentric Video Summarization Chain-of-Thought Prompting Elicits Reasoning in Large Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T22:23:43.159750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:23:43.159750Z digest=sha256:554501c9b5106cd46f28686d4afa61ec5bb3399c58a2811262bae60219c9e9c8

Observation 6b7f6f11-5185-4657-9eca-d3c950b1ec78 · outbound

This paper cites Nov 4, 2024, https://build.nvidia.com/nvidia/video-search-and-summarization 13.

Comparing Learning Paradigms for Egocentric Video Summarization Nov 4, 2024, https://build.nvidia.com/nvidia/video-search-and-summarization 13

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:23:44.649217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:23:43.243274Z digest=sha256:5c0f16d3d23ecb0812d978efc4e061fa9ed6112a1c66ee4c4ad0e07f74e90d49

Pith citing papers

No inbound Pith citation observations are available.