Pith. sign in

Paper Citation Record · LEDGER

Revisit Large-Scale Image-Caption Data in Pre-training Multimodal Foundation Models

As of 13 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:2410.02740.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.02740 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 4 of 4 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T15:17:22.609197Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T10:36:05.116399Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a833fbd6-1add-4ddf-8647-8e803f910412 · inbound

Multimodal Autoregressive Pre-training of Large Vision Encoders cites this paper.

Multimodal Autoregressive Pre-training of Large Vision Encoders Revisit Large-Scale Image-Caption Data in Pre-training Multimodal Foundation Models

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-12T15:17:22.609197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:17:22.609197Z digest=sha256:74a0e6da013965d9607b44c0f27c3ff02a99003a46b3db0e6edaa3e6170d76c1

Observation 01a791ac-b095-4a8e-b28d-7419fd44d011 · inbound

Scaling Inference-Time Search with Vision Value Model for Improved Visual Comprehension cites this paper.

Scaling Inference-Time Search with Vision Value Model for Improved Visual Comprehension Revisit Large-Scale Image-Caption Data in Pre-training Multimodal Foundation Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T22:18:01.471331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:18:01.471331Z digest=sha256:d0c6285a3984cb0b4d786b138b5b9ca44097ab38f93c350b16d6cde5be7d76cd

Observation 690ccd34-cd30-4a63-bdf3-caf85380e7c7 · inbound

Concrete Jungle: Towards Concreteness Paved Contrastive Negative Mining for Compositional Understanding cites this paper.

Concrete Jungle: Towards Concreteness Paved Contrastive Negative Mining for Compositional Understanding Revisit Large-Scale Image-Caption Data in Pre-training Multimodal Foundation Models

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:36:05.119696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T15:24:57.169737Z digest=sha256:e6eaba675a1e570eeaf0b2982ed55b50b5fe2d8ac9794766169eaa7cf9c4cf1c

Observation 1592e1d7-4a1c-4987-96b0-93cab0612cbe · inbound

A Reconstruction-Based Framework for Caption Evaluation Beyond Reference Captions cites this paper.

A Reconstruction-Based Framework for Caption Evaluation Beyond Reference Captions Revisit Large-Scale Image-Caption Data in Pre-training Multimodal Foundation Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T00:03:54.285204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:03:54.285204Z digest=sha256:128dd2e4ff394b359bf43d2c09a58be2c9af15ca2aa018bac49cc57ffb99bea7