Pith. sign in

Paper Citation Record · LEDGER

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries

As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2501.01282.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.01282 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T04:32:12.656115Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T20:16:12.001782Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 56f58a89-accb-4087-82a8-e4c9fb06670e · inbound

Large Language Model Agent: A Survey on Methodology, Applications and Challenges cites this paper.

Large Language Model Agent: A Survey on Methodology, Applications and Challenges CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries

Reference 248

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:52:10.327270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T21:51:34.309870Z digest=sha256:c52a9e298c07f5e2175c50d303cd3c40f077fdf03f3164563faf3e3694443f51

Observation 6274f23e-47bc-47da-96b8-c796c8ca706b · inbound

Appear2Meaning: A Cross-Cultural Benchmark for Structured Cultural Metadata Inference from Images cites this paper.

Appear2Meaning: A Cross-Cultural Benchmark for Structured Cultural Metadata Inference from Images CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries

Reference 36

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T06:05:55.519801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:49:25.227568Z digest=sha256:0036569c18e774062334b100c7e2f6f1b376a6b617fdec0f561ebd4ad1cac09b

Observation bcf456a2-4aa1-49a2-a39a-e0dff11e04d2 · inbound

BhashaSutra: A Task-Centric Unified Survey of Indian NLP Datasets, Corpora, and Resources cites this paper.

BhashaSutra: A Task-Centric Unified Survey of Indian NLP Datasets, Corpora, and Resources CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T12:01:02.085493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T04:22:34.046014Z digest=sha256:d7ec61640e31b46cab41cd981e57b1b4e42690207242b3b48cb7aeb5802fd76b

Observation bb620059-7136-4826-96eb-ad34ac4750f1 · inbound

When Cultures Move: Measuring and Improving Multicultural Text-to-Video Generation cites this paper.

When Cultures Move: Measuring and Improving Multicultural Text-to-Video Generation CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T22:02:49.023392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T21:59:44.030267Z digest=sha256:016b498099be5c29db2fc55d3e59d4e5bcddc5479abf7ecc40e28d4d022fa960

Observation 2cf76168-110d-4e89-85ec-9599d1807bef · inbound

When Cultures Move: Measuring and Improving Multicultural Text-to-Video Generation cites this paper.

When Cultures Move: Measuring and Improving Multicultural Text-to-Video Generation CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T19:45:01.074023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T19:36:43.654439Z digest=sha256:ffbcce5d716b440d08e3bb101de5fd97bbb79d4d8f87894b274941eeea346842

Observation 528c4356-9aa1-47be-8658-c2530d7ab751 · inbound

When Cultures Move: Measuring and Improving Multicultural Text-to-Video Generation cites this paper.

When Cultures Move: Measuring and Improving Multicultural Text-to-Video Generation CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T13:52:47.550786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:52:47.550786Z digest=sha256:1838b6fcb50ae8117938c27abeb594c1ad1bc3f7fc24c9f3f76ca045613fc485

Observation 6ab89fc6-1078-4da3-9d4e-e256b772702b · inbound

Computer-Aided Tagging on Wikimedia Commons: Designing for Human-AI Collaboration in Open Knowledge Work cites this paper.

Computer-Aided Tagging on Wikimedia Commons: Designing for Human-AI Collaboration in Open Knowledge Work CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-01T20:16:12.003355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T21:23:51.431142Z digest=sha256:d6fc942542600333bf58bd1aaf1a32de6b2d8393d07f60e1883efe5a4d39b3fa

Observation d1674364-8765-43a4-b7d2-96e63035fddd · inbound

Failing to See or Failing to Know? Attributing Errors in Vision-Language Models cites this paper.

Failing to See or Failing to Know? Attributing Errors in Vision-Language Models CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-11T15:14:25.526254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T15:14:25.526254Z digest=sha256:c3c547d185a539bec15c25efb3aa2b808fe98de3b8065b5f17700ca5a94b36ab

Observation 6c6ff0e7-d50f-4986-a70c-519d5ecad198 · inbound

Failing to See or Failing to Know? Attributing Errors in Vision-Language Models cites this paper.

Failing to See or Failing to Know? Attributing Errors in Vision-Language Models CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-13T06:57:34.446788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T06:57:34.446788Z digest=sha256:964cbfa5f38a5f449a34b114c4b6da72dbdd6720b0b738f2de957234f65b0177

Observation dfd18ba2-00b5-49c3-a827-a3932a04772f · inbound

Failing to See or Failing to Know? Attributing Errors in Vision-Language Models cites this paper.

Failing to See or Failing to Know? Attributing Errors in Vision-Language Models CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-04T04:32:12.656115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:32:12.656115Z digest=sha256:91d57616830044142159ee7b0bdafdf0e23db3b124e89253aa8c940638e24f7a