Pith. sign in

Paper Citation Record · LEDGER

LAVENDER: Unifying Video-Language Understanding as Masked Language Modeling

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 3 inbound Pith citation observations for arXiv:2206.07160.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2206.07160 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 3 of 3 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-17T00:36:53.235740Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-17T00:36:53.350076Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 43d26c3f-fb83-487a-9878-515c56343167 · inbound

InternVideo: General Video Foundation Models via Generative and Discriminative Learning cites this paper.

InternVideo: General Video Foundation Models via Generative and Discriminative Learning LAVENDER: Unifying Video-Language Understanding as Masked Language Modeling

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-17T00:36:53.351832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-17T00:36:53.235740Z digest=sha256:52f4f4a6566a22b34fa2f43ff1020652748a1ab1c945b95eb897fd8383ee1a85

Observation 95867805-c515-4620-b73c-0fb3e2bf0c0a · inbound

VideoChat: Chat-Centric Video Understanding cites this paper.

VideoChat: Chat-Centric Video Understanding LAVENDER: Unifying Video-Language Understanding as Masked Language Modeling

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:30:00.578691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-13T23:30:00.457974Z digest=sha256:b167fe0b3e007fe478f2e1b415d32bd0b0a1eb99996de9ccdfbc3d221d33a44e

Observation 10fb84ee-7504-4aef-be19-cc3458e42a42 · inbound

InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation cites this paper.

InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation LAVENDER: Unifying Video-Language Understanding as Masked Language Modeling

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-15T06:30:22.608373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-15T06:30:22.431538Z digest=sha256:b021ff51c491959523997632ceaec531ef1af14d48b3e6c16606816dd86058f6