Pith. sign in

Paper Citation Record · LEDGER

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation

As of 8 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 0 inbound Pith citation observations for arXiv:2508.16762.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.16762 v1

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T17:15:49.702336Z

measured 35 of 35 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

35 of 35 outbound references displayed

  • verified exact0
  • verified fuzzy26
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b2c894af-23a1-4d9e-874a-f1d50757802f · outbound

This paper cites To- wards measuring and modeling “culture” in LLMs: A sur- vey.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation To- wards measuring and modeling “culture” in LLMs: A sur- vey

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.507144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T17:15:49.487041Z digest=sha256:7158f11cc5c7b0be7c4236366de97c370a131777c9ad233967730b59b37fad40

Observation e6385693-90d8-416f-9ec7-ffe9cbb1395b · outbound

This paper cites Investigating Cultural Alignment of Large Language Models.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Investigating Cultural Alignment of Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T17:15:49.496287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:15:49.496287Z digest=sha256:402aebe64f4f1115fd7cde5472274f048f1b7f47719a2fd8642d85e55e91efd2

Observation e26cacf3-710f-4600-b4af-5286c44cab1c · outbound

This paper cites Probing Pre-Trained Language Models for Cross-Cultural Differences in Values.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Probing Pre-Trained Language Models for Cross-Cultural Differences in Values

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T17:15:49.502065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:15:49.502065Z digest=sha256:bd9f8c6e8effeed52e7efe9e68f91eb0398bde5d496ecb49593d38a28b7fb6db

Observation eeff2b56-5420-42a3-a3ee-a0d9ff2c354e · outbound

This paper cites Probing pre-trained language models for cross-cultural dif- ferences in values.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Probing pre-trained language models for cross-cultural dif- ferences in values

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.488957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T17:15:49.511930Z digest=sha256:4098d0e0e4ffd4f8bc9a1b42faf6afdaa0927f8b63b3fc333384b02d9414953f

Observation 66a2cf49-afc4-4e96-b9ad-aeec69601ceb · outbound

This paper cites Qwen2.5-vl technical report, 2025.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Qwen2.5-vl technical report, 2025

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.470713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T17:15:49.518870Z digest=sha256:fc38af5e0c004492e0bb3f5994fcfb77e289fc9ecb038f6cc992ddf9ddbf256f

Observation 4f13b1f1-b300-4206-8e75-309ffe398716 · outbound

This paper cites Venkatesh Babu, and Danish Pruthi.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Venkatesh Babu, and Danish Pruthi

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.451398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T17:15:49.526751Z digest=sha256:49abf1a570c067853ea31e9cda4b79766feca96cd81c83d9fa9ba2a165d5ad68

Observation 91957dbb-4d85-4206-aead-d19ce5c8ec25 · outbound

This paper cites From local concepts to univer- sals: Evaluating the multicultural understanding of vision- language models.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation From local concepts to univer- sals: Evaluating the multicultural understanding of vision- language models

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.434223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T17:15:49.534440Z digest=sha256:e8f72d2ffc89d4f9ee94c4e4de56be8852d71a1da286953000be550ba3d7b2fa

Observation 97432fa6-e243-42be-9367-9868606a8645 · outbound

This paper cites Extrinsic evaluation of cultural competence in large language models.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Extrinsic evaluation of cultural competence in large language models

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.419391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T17:15:49.540424Z digest=sha256:ea68683551cd495a4bc8ea34ae028b5e2a529cddce8e21f03fb370651e7ee069

Observation ce25a871-4a72-407e-9cf4-76306ed76a63 · outbound

This paper cites Language (technology) is power: A critical sur- vey of “bias” in NLP.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Language (technology) is power: A critical sur- vey of “bias” in NLP

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.400342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T17:15:49.545812Z digest=sha256:23ee1072a85646c605c30ce20852fc5c902f1942c55ff16ce7dddfe95183af82

Observation b419d52b-e392-4a3a-9732-fd291315fdf8 · outbound

This paper cites Assessing Cross-Cultural Alignment between ChatGPT and Human Societies: An Empirical Study.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Assessing Cross-Cultural Alignment between ChatGPT and Human Societies: An Empirical Study

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T17:15:49.550399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:15:49.550399Z digest=sha256:4c01ee0a3d4d18e250e1979b4902eafd074ef74a816e9639c267b615c77d8a63

Observation 9f93418e-cef3-412c-af0f-cd8c33a085aa · outbound

This paper cites MaXM: Towards multilingual visual question an- swering.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation MaXM: Towards multilingual visual question an- swering

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.244100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T17:15:49.556238Z digest=sha256:1fc56681b27c899487077cd36cbb8b797010d2dae8231078d4f66eb76dea3b69

Observation bd15319a-ccc5-41a6-b263-edc9fac94030 · outbound

This paper cites The SAGE handbook of intercultural competence.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation The SAGE handbook of intercultural competence

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.217411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T17:15:49.566255Z digest=sha256:62bc40005f6f875fc1c4a52420f9ccd5e3f4f73b5ccb39299f2cbfae008746d2

Observation 44af451b-f7bc-47fb-ac35-773a8126fcff · outbound

This paper cites Towards Measuring the Representation of Subjective Global Opinions in Language Models.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Towards Measuring the Representation of Subjective Global Opinions in Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T17:15:49.570715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:15:49.570715Z digest=sha256:753846d108251ed4af97ec7a825f9778f98070b9a9b803f9e738fd362e3347c9

Observation f5c7ff0f-a9b3-41d4-9292-2a671d4991b2 · outbound

This paper cites ”i wouldn’t say offensive but...”: Disability-centered perspectives on large language models.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation ”i wouldn’t say offensive but...”: Disability-centered perspectives on large language models

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.201890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T17:15:49.575968Z digest=sha256:8c98035e72832b646afd08ce0e0ee414171500d2679e86a38acfdb7a37d1acb0

Observation 36a82315-2055-48ba-be2e-18e18cb78857 · outbound

This paper cites World values survey: Round seven - country-pooled datafile version 5.0, 2022.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation World values survey: Round seven - country-pooled datafile version 5.0, 2022

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.183490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T17:15:49.581681Z digest=sha256:3775bde980418704297f20774a42d3950bfb92101613683f1700f1f9e8f00d8d

Observation 6a98888b-6b7f-41f2-9525-77fb27da3e41 · outbound

This paper cites Challenges and strategies in cross- cultural NLP.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Challenges and strategies in cross- cultural NLP

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.161813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T17:15:49.586036Z digest=sha256:ce70857ffa90367d8db92975a4a0fde8cfb56481f93987f02d7a56f2a1cdf79a

Observation ab7e2c3b-6d81-492f-8621-97fa4e3f144a · outbound

This paper cites Clipscore: A reference-free evaluation met- ric for image captioning, 2022.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Clipscore: A reference-free evaluation met- ric for image captioning, 2022

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.147569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T17:15:49.591603Z digest=sha256:e06bdfd7de3ce10708a665bc17b9575a266e8b4830e9789861d67d6b33020cfd

Observation 8a2bcde3-1a01-4517-bc7d-f93ac753c396 · outbound

This paper cites Dimensionalizing cultures: The hofstede model in context.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Dimensionalizing cultures: The hofstede model in context

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.133979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T17:15:49.601101Z digest=sha256:b9e36fc3ad940c08e9c172c6a693de221c07ead198c868856aabd682f4b2b17f

Observation 07b5b54a-1e3c-40a4-961c-e0fe3acb6855 · outbound

This paper cites The PRISM Alignment Dataset: What Participatory, Representative and Individualised Human Feedback Reveals About the Subjective and Multicultural Alignment of Large Language Models.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation The PRISM Alignment Dataset: What Participatory, Representative and Individualised Human Feedback Reveals About the Subjective and Multicultural Alignment of Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T17:15:49.605724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:15:49.605724Z digest=sha256:c30da6074d705044ef1d0639e78620ed204da96a010d16f3bfe707db939deeb5

Observation 51157317-dd22-42c3-8893-0d9745f895a1 · outbound

This paper cites Visually 9 grounded reasoning across languages and cultures.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Visually 9 grounded reasoning across languages and cultures

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.120240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T17:15:49.613174Z digest=sha256:77452b48dd31b8c7b67aa1fd04a93d9944988e6ffcc52aef493d032d2245df62

Observation 7b1f3961-56ea-4081-9caa-e6ffaeb19ea3 · outbound

This paper cites Wong, Qing- song Wen, Lichao Sun, Haipeng Chen, Xing Xie, and Jin- dong Wang.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Wong, Qing- song Wen, Lichao Sun, Haipeng Chen, Xing Xie, and Jin- dong Wang

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.106486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T17:15:49.619857Z digest=sha256:b9a117dc47f63aa681c49224387ae973ea3ffd90ad5451727bf67ceb4a18035b

Observation feb9ebcd-dc92-4475-8b1b-1c7fc41cec29 · outbound

This paper cites Smolvlm: Re- defining small and efficient multimodal models, 2025.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Smolvlm: Re- defining small and efficient multimodal models, 2025

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.089084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T17:15:49.624563Z digest=sha256:c0422ec3e0ca6d2c0b24fbb508c2732bcffcaace97334a237cc68ab2953033da

Observation 382ea78f-496d-486c-b6c1-0567dd6a0531 · outbound

This paper cites Cultural Alignment in Large Language Models: An Explanatory Analysis Based on Hofstede's Cultural Dimensions.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Cultural Alignment in Large Language Models: An Explanatory Analysis Based on Hofstede's Cultural Dimensions

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T17:15:49.632259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:15:49.632259Z digest=sha256:8c0e65bb2f9d27e97e9e7a5dcc6545a84801df05041ccd2b94a020e3af245565

Observation b035525d-fed5-4b4d-b39f-9d007f0fa8ef · outbound

This paper cites Assessing demographic bias in named entity recognition, 2020.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Assessing demographic bias in named entity recognition, 2020

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.071548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T17:15:49.638211Z digest=sha256:d24547cc251c4db76aaf3ac486bf5729e9acde18df214e82699124c271b9836d

Observation e4b1d5bd-0ba2-4192-bea4-f45810b8508b · outbound

This paper cites Benchmarking vision language models for cultural understanding.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Benchmarking vision language models for cultural understanding

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.053582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T17:15:49.642962Z digest=sha256:43a0b67eabdae732d8618e538a8b384c971627c27cf0e4b521d5219e07f46a4c

Observation 56c09c94-b480-49b8-b306-f218cdf8f666 · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Bleu: a method for automatic evaluation of machine translation

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.034075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T17:15:49.647539Z digest=sha256:2eccd1d33b05db9c95775ebdefa755e6caa4fd89e7c4226ee1055590ad626f46

Observation 6f975014-75c2-46a3-9868-a7711e77cba0 · outbound

This paper cites Knowledge of cultural moral norms in large language models.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Knowledge of cultural moral norms in large language models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T17:15:49.655207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:15:49.655207Z digest=sha256:49ddce7a9875d7229b8637a26955f4503899d7f127e0814d7be30f424104f879

Observation 1694d9d3-b235-4ed8-beb3-5dc8c4667c2c · outbound

This paper cites Normad: A benchmark for measuring the cultural adaptability of large language mod- els.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Normad: A benchmark for measuring the cultural adaptability of large language mod- els

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.007915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T17:15:49.664060Z digest=sha256:44f8fc48e2cb56c14abe9289f312ef413ddf96232ecaa5a7630fcb9804fcedbe

Observation bcbf7f3f-335a-4d07-8dfa-752d6484d6a3 · outbound

This paper cites Cvqa: culturally-diverse multilingual visual question answering benchmark.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Cvqa: culturally-diverse multilingual visual question answering benchmark

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:49.988507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T17:15:49.669594Z digest=sha256:71c0051f932c83e717e546a7de276a8398460e1c042b2b2e88e401b6e9ab5462

Observation 4caac69a-eeaa-4e52-9984-a26eca4641b5 · outbound

This paper cites Geographical erasure in language generation.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Geographical erasure in language generation

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:49.969879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T17:15:49.676360Z digest=sha256:07a08c8f71b0a293739a2b3b4fb97f542b38d90daa612a6a0634b0a90f5639c2

Observation 8ea07621-8980-4273-b82c-8f15f33284f4 · outbound

This paper cites Cultural bias and cultural alignment of large language mod- els.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Cultural bias and cultural alignment of large language mod- els

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:49.953680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T17:15:49.683290Z digest=sha256:604ad1f7be3e9aca44e8f9622da3bc600f8adcdaa38e1d1bf013ee5e7c1ff6f6

Observation 79701dd4-3670-4960-82c9-139fbec99289 · outbound

This paper cites an unresolved cited work.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-05T17:15:49.933681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T17:15:49.689074Z digest=sha256:c06cecdcd04afdfd22ee14d7d0c5dad5c5c4b90354549ec265e5e901886833ed

Observation cc5f8c52-4426-499d-8f33-852afcc446fe · outbound

This paper cites Broaden the vision: Geo-diverse visual com- monsense reasoning.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Broaden the vision: Geo-diverse visual com- monsense reasoning

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:49.914244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T17:15:49.694750Z digest=sha256:7d7909aa5f2f3042e3e30eeacaa64641dea056ba52558e3d8fd285f2c9c20ac9

Observation 76a60ec6-aed8-4285-85f8-955bc4c04e0c · outbound

This paper cites Internvl3: Exploring advanced training and test-time recipes for open-source multimodal models, 2025.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Internvl3: Exploring advanced training and test-time recipes for open-source multimodal models, 2025

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:49.897339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T17:15:49.702336Z digest=sha256:b8300901a80c82b9c0e146bdca26f7c08dbb90ac52eecc7bb077f708f46d0a5d

Observation 0e39aa5f-683e-468c-932e-56a15a210dc5 · outbound

This paper cites an unresolved cited work.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Unresolved cited work

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-05T17:15:49.561069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:15:49.561069Z digest=sha256:76f8ffdf5b6098396e0b960f56b66bfed9aefe1a012bfebbf7e0be3409d787bd

Pith citing papers

No inbound Pith citation observations are available.