Pith. sign in

Paper Citation Record · LEDGER

EVA: Exploring the Limits of Masked Visual Representation Learning at Scale

As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2211.07636.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2211.07636 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T13:29:32.095801Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-09T22:06:35.172946Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 27003d71-68cf-480d-ab33-7be43511df90 · inbound

BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models cites this paper.

BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models EVA: Exploring the Limits of Masked Visual Representation Learning at Scale

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-12T00:10:49.251971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-12T00:10:48.610351Z digest=sha256:a841cb9bd45977e6da4c95b1d151653b78b92d97a31fc0d16896f7211eb611b1

Observation 4b240e77-6838-4723-a1a2-7770350f0f94 · inbound

EVA-CLIP: Improved Training Techniques for CLIP at Scale cites this paper.

EVA-CLIP: Improved Training Techniques for CLIP at Scale EVA: Exploring the Limits of Masked Visual Representation Learning at Scale

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:54:21.991369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T01:54:21.943160Z digest=sha256:182658f7a766532c212778be75f4f4f3e423811fe56d89ac6e98c02d58d1dc51

Observation c1ca183a-1d75-4fe3-a795-072176faa68d · inbound

MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models cites this paper.

MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models EVA: Exploring the Limits of Masked Visual Representation Learning at Scale

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-10T20:37:01.845612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T20:37:01.617345Z digest=sha256:7662d47f1923c8b646735532a794a2d31da4248ec131db7fa9a64c5f579641cf

Observation 0e23fad6-3b2b-4281-a6cf-84f65abd0ec1 · inbound

InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning cites this paper.

InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning EVA: Exploring the Limits of Masked Visual Representation Learning at Scale

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:13:52.188951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T02:13:52.097263Z digest=sha256:15531e1f5be09d9eebfe61940c5020ed851fb6a49e425c5442088000533a70c3

Observation 2855931f-9a6d-4f73-b179-bd16a7ef92b6 · inbound

MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning cites this paper.

MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning EVA: Exploring the Limits of Masked Visual Representation Learning at Scale

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-16T07:13:08.959578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T07:13:08.867745Z digest=sha256:9ecbc60abe3c5e518b20b007be3d0c188b58e3c20c0a687b2a03dde1179bfb12

Observation a5d990b2-0c36-404d-8687-62e407cd3fe9 · inbound

InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks cites this paper.

InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks EVA: Exploring the Limits of Masked Visual Representation Learning at Scale

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:46:09.882830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T22:46:09.693156Z digest=sha256:93cd314bd890bd7ed2c1c84ff593c59788f7f720cd56fd5f786572d45bfa5ceb

Observation 471bc513-cb40-4322-97c5-bb209aa56152 · inbound

XAMI -- A Benchmark Dataset for Artefact Detection in XMM-Newton Optical Images cites this paper.

XAMI -- A Benchmark Dataset for Artefact Detection in XMM-Newton Optical Images EVA: Exploring the Limits of Masked Visual Representation Learning at Scale

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-24T00:33:40.089255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T00:29:09.172759Z digest=sha256:22ff1b53bf2aa6276367e72358ccf684dcf1c8b4ca476f69e6f4b4da9ef6cd2c

Observation 17fdfc44-6a79-40eb-b09d-b6045a270c36 · inbound

Are Large Pre-trained Vision Language Models Effective Construction Safety Inspectors? cites this paper.

Are Large Pre-trained Vision Language Models Effective Construction Safety Inspectors? EVA: Exploring the Limits of Masked Visual Representation Learning at Scale

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-18T22:32:52.373550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T22:32:11.401201Z digest=sha256:ebaa69c23d039a3b25dd1245abf5d15bb4bed97225949432d7b77b8c0215a684

Observation 95528180-a457-46b2-8221-353ca9a0be23 · inbound

ER-LoRA: Effective-Rank Guided Adaptation for Weather-Generalized Depth Estimation cites this paper.

ER-LoRA: Effective-Rank Guided Adaptation for Weather-Generalized Depth Estimation EVA: Exploring the Limits of Masked Visual Representation Learning at Scale

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T13:29:32.095801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:29:32.095801Z digest=sha256:790f82aa58ae4c1349c8e384ac8d38ced888916dce029a355c023c579a5b68fb

Observation ee792cb6-d05a-4e66-84a6-7cf69ac44089 · inbound

Ego-Human Motion Prediction with 3D-Aware LLM cites this paper.

Ego-Human Motion Prediction with 3D-Aware LLM EVA: Exploring the Limits of Masked Visual Representation Learning at Scale

Reference 21

Resolution
metadata mismatch
local_arxiv, observed 2026-07-09T22:06:35.174307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-09T22:04:11.433376Z digest=sha256:9a3e4516cab457ca7a8af156a873529699cbc73739e37845c74a53936593c009