Pith. sign in

Paper Citation Record · LEDGER

A Comprehensive Study of Deep Video Action Recognition

As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2012.06567.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2012.06567 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T12:18:37.818453Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T18:23:51.204442Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 735d3440-32fe-49d3-b768-2f8d059bb947 · inbound

Low-Latency Video Anonymization for Crowd Anomaly Detection: Privacy Versus Performance cites this paper.

Low-Latency Video Anonymization for Crowd Anomaly Detection: Privacy Versus Performance A Comprehensive Study of Deep Video Action Recognition

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-23T18:28:19.001700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-23T18:25:44.799112Z digest=sha256:bc53026ec17605896b62145ce6b23c0ba9b16483d4a75f6849c5dbded4b369a4

Observation 2d69d7a3-54c0-452e-a900-87ece0af7b7f · inbound

Bridging the Data Provenance Gap Across Text, Speech and Video cites this paper.

Bridging the Data Provenance Gap Across Text, Speech and Video A Comprehensive Study of Deep Video Action Recognition

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:37.818453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:37.818453Z digest=sha256:b9058604ed4d64221bc6f49921fd9374ff90785725ca692bb7a0559ac51664a5

Observation 1dc5a034-c241-4ee0-b834-9ed607815125 · inbound

Dynamic Scene Understanding from Vision-Language Representations cites this paper.

Dynamic Scene Understanding from Vision-Language Representations A Comprehensive Study of Deep Video Action Recognition

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-10T18:04:29.251417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:04:29.251417Z digest=sha256:c559315eb1cc9d91b854346e82594c6bc706895d588e96b519107a54978ae819

Observation 151329ed-cf5c-4c17-bd2e-de575de3502d · inbound

Can masking background and object reduce static bias for zero-shot action recognition? cites this paper.

Can masking background and object reduce static bias for zero-shot action recognition? A Comprehensive Study of Deep Video Action Recognition

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-10T16:57:36.918200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T16:57:36.918200Z digest=sha256:bd9d17a48e2ee984b2ec694581b62339f6c3308cd00b0947fcdaa245a3eb239f

Observation 05b208c8-697a-4256-8236-10283c538c9c · inbound

SMART-Vision: Survey of Modern Action Recognition Techniques in Vision cites this paper.

SMART-Vision: Survey of Modern Action Recognition Techniques in Vision A Comprehensive Study of Deep Video Action Recognition

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T16:32:14.884011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T16:32:14.884011Z digest=sha256:a826db5fc23f6264144946f6a0b158ed9b1f632abcfdf731671469d288a0aad5

Observation a61ab16e-617c-4862-a10d-3ac3da5b4032 · inbound

What to Do Next? Memorizing skills from Egocentric Instructional Video cites this paper.

What to Do Next? Memorizing skills from Egocentric Instructional Video A Comprehensive Study of Deep Video Action Recognition

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T21:03:08.926468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:03:08.926468Z digest=sha256:972444e214f2a560a74c3e9067a94a95ecae0ab0437bd5686b4f7d530d121540

Observation a1115ff8-dc7e-4c5d-b44e-ff31a2438b41 · inbound

A Survey on Video Temporal Grounding with Multimodal Large Language Model cites this paper.

A Survey on Video Temporal Grounding with Multimodal Large Language Model A Comprehensive Study of Deep Video Action Recognition

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-05T23:32:17.622367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:32:17.622367Z digest=sha256:b7d9642ef02d10bc12321dba86448cf6cac408499d6225d9018ea9fb4ffcad32

Observation 168a7747-f835-4501-906a-29f09152926e · inbound

Leveraging Vision-Language Large Models for Interpretable Video Action Recognition with Semantic Tokenization cites this paper.

Leveraging Vision-Language Large Models for Interpretable Video Action Recognition with Semantic Tokenization A Comprehensive Study of Deep Video Action Recognition

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T05:12:54.460881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:12:54.460881Z digest=sha256:02946174b254d3041b83dcbefcf07addf088e6255524a8ae0579c84b15e25619

Observation 9d01e54a-3ce2-40d0-aff1-53a055647e03 · inbound

AVATAAR: Agentic Video Answering via Temporal Adaptive Alignment and Reasoning cites this paper.

AVATAAR: Agentic Video Answering via Temporal Adaptive Alignment and Reasoning A Comprehensive Study of Deep Video Action Recognition

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-17T20:20:11.813025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-17T20:18:19.580156Z digest=sha256:d67e685a0d5e0f8375686166e91d656566dbb2aa9b2e891a61280699061e3d2d

Observation d9a1d30a-67d0-4183-bb3c-e1b2329784ec · inbound

TRUST: Efficient Abdominal Trauma Recognition via Image-to-Ultrasound-Video Transfer Learning cites this paper.

TRUST: Efficient Abdominal Trauma Recognition via Image-to-Ultrasound-Video Transfer Learning A Comprehensive Study of Deep Video Action Recognition

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T18:23:51.205822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T05:08:02.242341Z digest=sha256:6d6d6971869b10c869e1722ab7a904ebc35f0cb07ddefec5f6935c01a37473b6

Observation 8c62c160-30cc-4b2b-b040-a835b7ec13d9 · inbound

Efficient Video Dataset Distillation via Cluster-Guided Prototype Blending cites this paper.

Efficient Video Dataset Distillation via Cluster-Guided Prototype Blending A Comprehensive Study of Deep Video Action Recognition

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-05T22:10:41.654548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:10:41.654548Z digest=sha256:0c52647c59d12aef73d4b68b6c5f1443ca319329302834ba735a2ef33b7cb8a4