Pith. sign in

Paper Citation Record · LEDGER

Vision-and-Language Pretrained Models: A Survey

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2204.07356.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2204.07356 v5

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T16:18:38.615236Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

7
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1cea652e-fc8f-42b5-8cf3-7d9fb02d8294 · inbound

One Fits All: General Mobility Trajectory Modeling via Masked Conditional Diffusion cites this paper.

One Fits All: General Mobility Trajectory Modeling via Masked Conditional Diffusion Vision-and-Language Pretrained Models: A Survey

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T16:18:38.615236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T16:18:38.615236Z digest=sha256:e2376774f1f3e7666e0f11c32882b603b7b2e0a9bb8b3a778710fb4a88aa0be9

Observation 58c82d19-f1b6-4c88-ab74-476bc7be3be1 · inbound

Measuring and Mitigating Hallucinations in Vision-Language Dataset Generation for Remote Sensing cites this paper.

Measuring and Mitigating Hallucinations in Vision-Language Dataset Generation for Remote Sensing Vision-and-Language Pretrained Models: A Survey

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T14:52:22.143739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T14:52:22.143739Z digest=sha256:224464a33ee0de6a9a7c5dbd2ff117c28a193aef6a355bec21a8bb2cba5f59f2

Observation 959687f3-5e8b-4b36-8cad-c7bfa06da886 · inbound

One Head Eight Arms: Block Matrix based Low Rank Adaptation for CLIP-based Few-Shot Learning cites this paper.

One Head Eight Arms: Block Matrix based Low Rank Adaptation for CLIP-based Few-Shot Learning Vision-and-Language Pretrained Models: A Survey

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T11:16:55.799715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T11:16:55.799715Z digest=sha256:c0c0061d8552e22c9af1550afe5e8c4cd811e140df90d2726d50d03cd57297c4

Observation ce3d1c0b-f348-4bdc-a973-23c90907feb9 · inbound

CF-VLM:CounterFactual Vision-Language Fine-tuning cites this paper.

CF-VLM:CounterFactual Vision-Language Fine-tuning Vision-and-Language Pretrained Models: A Survey

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:07.195795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:07.195795Z digest=sha256:6e0c7f6e54bd38f07bcac465b0520a34ced2b2896ea6c802b9e52926476b0455

Observation 0deb1c5c-9151-42e6-8570-5675dbf7853b · inbound

Unmasking LAION-5B: Age, Gender, Race, and Emotion Biases in Large-Scale Image Datasets cites this paper.

Unmasking LAION-5B: Age, Gender, Race, and Emotion Biases in Large-Scale Image Datasets Vision-and-Language Pretrained Models: A Survey

Reference 277

Resolution
verified exact
arxiv_id, observed 2026-06-26T09:19:16.946994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-26T09:12:19.873337Z digest=sha256:948f6aef51300d959ba50eecbe871d59cfd0f1ab34843be697814cc012ff0075