Pith. sign in

Paper Citation Record · LEDGER

Co-training Transformer with Videos and Images Improves Action Recognition

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2112.07175.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2112.07175 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T04:37:09.846651Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-19T17:52:42.907729Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b43a2758-3a6a-4865-bc22-4468f6c5be92 · inbound

CoCa: Contrastive Captioners are Image-Text Foundation Models cites this paper.

CoCa: Contrastive Captioners are Image-Text Foundation Models Co-training Transformer with Videos and Images Improves Action Recognition

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T10:53:08.374529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T10:53:08.292063Z digest=sha256:cb32c919c8c5ca61811cafb668ad4ff87001399ac12f68a2beb2cc63338d0526

Observation 8ff8d952-8172-4892-8a76-e4246122f55e · inbound

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners cites this paper.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Co-training Transformer with Videos and Images Improves Action Recognition

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-09T04:37:09.846651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:37:09.846651Z digest=sha256:ea5b35ccb31a83eb33a7c5b0c439ca505816ebf4e50800c9656d972d4310a40e

Observation b4f2f0f4-d201-45e4-acc2-a84f132cb09b · inbound

CrossVideoMAE: Self-Supervised Image-Video Representation Learning with Masked Autoencoders cites this paper.

CrossVideoMAE: Self-Supervised Image-Video Representation Learning with Masked Autoencoders Co-training Transformer with Videos and Images Improves Action Recognition

Reference 109

Resolution
unresolved
no resolver link, observed 2026-08-08T19:17:39.868111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:17:39.868111Z digest=sha256:1a5d6ec368f45f9e8ce5675d348024f6a9989dd27f50f3cf4d9099a3144fc715

Observation 059d974a-5f2f-4223-8d06-e0144d9e2906 · inbound

SURGE: Surrogate Gradient Adaptation in Binary Neural Networks cites this paper.

SURGE: Surrogate Gradient Adaptation in Binary Neural Networks Co-training Transformer with Videos and Images Improves Action Recognition

Reference 82

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:37:26.611087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-13T06:37:12.626356Z digest=sha256:d966cd4df6487c960fd2a874218097447323f872908559011fa63c6f34fbd33e

Observation 4eda177d-367c-4758-8483-990625629244 · inbound

SURGE: Surrogate Gradient Adaptation in Binary Neural Networks cites this paper.

SURGE: Surrogate Gradient Adaptation in Binary Neural Networks Co-training Transformer with Videos and Images Improves Action Recognition

Reference 82

Resolution
verified exact
arxiv_id, observed 2026-05-19T17:52:42.910291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-19T17:49:19.281712Z digest=sha256:34f110e4e4db5e83e61c34d21e5ca65f25bfcec8a662caaf25279736b849a52e