Pith. sign in

Paper Citation Record · LEDGER

KRETA: A Benchmark for Korean Reading and Reasoning in Text-Rich VQA Attuned to Diverse Visual Contexts

As of 19 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 0 inbound Pith citation observations for arXiv:2508.19944.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.19944 v2

Coverage vector

measured 21 of 21 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T15:23:54.203192Z

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

21 of 21 outbound references displayed

  • verified exact0
  • verified fuzzy3
  • unresolved14
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch4

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a7834192-50b9-4a2c-a1b0-fab71059e897 · outbound

This paper cites Distilling System 2 into System 1.

KRETA: A Benchmark for Korean Reading and Reasoning in Text-Rich VQA Attuned to Diverse Visual Contexts Distilling System 2 into System 1

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T15:23:54.147498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:23:54.147498Z digest=sha256:17c6cd8fd3c618f8a7c8aab070c2628854f0cf619b1dfeae05b3e33bc4b92faf

Observation 332a2505-24d6-4798-b43d-9f2c31962986 · outbound

This paper cites Evaluating Visual and Cultural Interpretation: The K-Viscuit Benchmark with Human-VLM Collaboration.

KRETA: A Benchmark for Korean Reading and Reasoning in Text-Rich VQA Attuned to Diverse Visual Contexts Evaluating Visual and Cultural Interpretation: The K-Viscuit Benchmark with Human-VLM Collaboration

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-08-05T15:23:54.843244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T15:23:52.390750Z digest=sha256:b04e1e3362377d7f10bb9bce9a0ac4457761943f169673a5bbecbea6f906d577

Observation 562a8238-04d2-4831-b6c0-67ea53962bb3 · outbound

This paper cites target-language in- structions for multilingual LLMs.

KRETA: A Benchmark for Korean Reading and Reasoning in Text-Rich VQA Attuned to Diverse Visual Contexts target-language in- structions for multilingual LLMs

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:23:55.619383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T15:23:52.635515Z digest=sha256:b67b28773b9108da8eb133c92341fffecd3d562a4cedbbea33812a1effe3c57f

Observation f00b02e9-d38b-42fe-8d8d-75f5b9d9d396 · outbound

This paper cites QGEval: Benchmarking Multi-dimensional Evaluation for Question Generation.

KRETA: A Benchmark for Korean Reading and Reasoning in Text-Rich VQA Attuned to Diverse Visual Contexts QGEval: Benchmarking Multi-dimensional Evaluation for Question Generation

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-08-05T15:23:54.730802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T15:23:52.739805Z digest=sha256:24b38163003dd65bd863a4882ff706fa3eabe34ca46f66bc9ec47d4c4c6ba782

Observation f86f6a0d-31ae-4ba1-9bb2-869119a11ede · outbound

This paper cites Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark.

KRETA: A Benchmark for Korean Reading and Reasoning in Text-Rich VQA Attuned to Diverse Visual Contexts Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T15:23:52.822183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:23:52.822183Z digest=sha256:8545cb54774d8fe8bd224990ae12d47150b5f434f77a89d42adaa19dbe3a0c59

Observation 42f4af09-c14f-415e-a1e6-575f581e615f · outbound

This paper cites VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models.

KRETA: A Benchmark for Korean Reading and Reasoning in Text-Rich VQA Attuned to Diverse Visual Contexts VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T15:23:53.069460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:23:53.069460Z digest=sha256:40dcab02b8f220db1ef728c1a036d32af9bf429f0dc279eea668dd0bd892ae58

Observation ba2e8402-f650-4103-be53-4d5e95c4e695 · outbound

This paper cites KOFFVQA: An Objectively Evaluated Free-form VQA Benchmark for Large Vision-Language Models in the Korean Language.

KRETA: A Benchmark for Korean Reading and Reasoning in Text-Rich VQA Attuned to Diverse Visual Contexts KOFFVQA: An Objectively Evaluated Free-form VQA Benchmark for Large Vision-Language Models in the Korean Language

Reference 9

Resolution
metadata mismatch
local_arxiv, observed 2026-08-05T15:23:54.587026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T15:23:53.207102Z digest=sha256:e55cc6f8c61aa033e79c1acd6b851588296b1f41a6e328ad25183348e1017220

Observation afef7787-2222-4467-83d4-49502986deca · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

KRETA: A Benchmark for Korean Reading and Reasoning in Text-Rich VQA Attuned to Diverse Visual Contexts LLaVA-OneVision: Easy Visual Task Transfer

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T15:23:53.312371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:23:53.312371Z digest=sha256:93ae7341ea6b6ccfc3db3e7b75ae120ae407308811450c1d35b39a9a704a8890

Observation adb73585-05b2-45da-8192-0ca009ca97cf · outbound

This paper cites Ovis: Structural Embedding Alignment for Multimodal Large Language Model.

KRETA: A Benchmark for Korean Reading and Reasoning in Text-Rich VQA Attuned to Diverse Visual Contexts Ovis: Structural Embedding Alignment for Multimodal Large Language Model

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T15:23:53.395808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:23:53.395808Z digest=sha256:5c4626789b6560636857960b9ba1fd35bceae7c69666b84ba3611ceeaf9fb79b

Observation d20a17e8-5e5a-480f-b5e7-34c77897b53b · outbound

This paper cites ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning.

KRETA: A Benchmark for Korean Reading and Reasoning in Text-Rich VQA Attuned to Diverse Visual Contexts ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T15:23:53.484435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:23:53.484435Z digest=sha256:c859d93b77e8ad79040a23c0f81b8a26a583ef2138a26f041d1e7e50dec63584

Observation 3987501e-c357-48f3-a73e-e59064896d08 · outbound

This paper cites In Findings of the Association for Computational Linguistics: ACL 2022 , pages 2497–.

KRETA: A Benchmark for Korean Reading and Reasoning in Text-Rich VQA Attuned to Diverse Visual Contexts In Findings of the Association for Computational Linguistics: ACL 2022 , pages 2497–

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:23:55.266273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T15:23:53.649279Z digest=sha256:ee5f286414eb33e1c1afbcdde1ee577eeb2c2969d70503da4aed0ac1073b74da

Observation f73cfec0-2720-4807-9af4-74bc2f9b15f9 · outbound

This paper cites Parrot: Multilingual Visual Instruction Tuning.

KRETA: A Benchmark for Korean Reading and Reasoning in Text-Rich VQA Attuned to Diverse Visual Contexts Parrot: Multilingual Visual Instruction Tuning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T15:23:53.733257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:23:53.733257Z digest=sha256:e87967b53391fa84c133bc7edb4e8f09d92db988d54a8f08a7f67f9cb08fb45d

Observation 3a81e72e-319b-4f15-86f8-1740168afd13 · outbound

This paper cites MUST-VQA: MUltilingual Scene-text VQA.

KRETA: A Benchmark for Korean Reading and Reasoning in Text-Rich VQA Attuned to Diverse Visual Contexts MUST-VQA: MUltilingual Scene-text VQA

Reference 16

Resolution
metadata mismatch
local_arxiv, observed 2026-08-05T15:23:54.395283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T15:23:53.836153Z digest=sha256:2c8d82dc9e68f0f887a1445756f073b96f3e485a953ce6083b683de415314dc6

Observation c492fa4c-5f08-4133-8496-7f968711638b · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

KRETA: A Benchmark for Korean Reading and Reasoning in Text-Rich VQA Attuned to Diverse Visual Contexts Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T15:23:53.923746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:23:53.923746Z digest=sha256:a4e886c7219b4b4f16bfc91c719da99c2f6ffd6d1c505b9af943e9fa47f5cbea

Observation 6546014c-ec27-47c5-b35f-4c61ae4e52bf · outbound

This paper cites DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding.

KRETA: A Benchmark for Korean Reading and Reasoning in Text-Rich VQA Attuned to Diverse Visual Contexts DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T15:23:54.018200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:23:54.018200Z digest=sha256:ea6f4a7411e4ed3ff413b22fb9b7599ed40b2e8c148fbefbb2e9231d3b3dbaf7

Observation ee559cbe-d8a1-40b9-afcc-1aefb63e9d0f · outbound

This paper cites MiniCPM-V: A GPT-4V Level MLLM on Your Phone.

KRETA: A Benchmark for Korean Reading and Reasoning in Text-Rich VQA Attuned to Diverse Visual Contexts MiniCPM-V: A GPT-4V Level MLLM on Your Phone

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T15:23:54.087145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:23:54.087145Z digest=sha256:67884022d90e6a1463c62b3f2247255a227bdd642deb14bb7d2d4ab7e753e602

Observation de6250fe-b1c2-463e-b2d0-349481c18986 · outbound

This paper cites an unresolved cited work.

KRETA: A Benchmark for Korean Reading and Reasoning in Text-Rich VQA Attuned to Diverse Visual Contexts Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:23:55.051104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T15:23:54.203192Z digest=sha256:1490400251333f5092b1d4a30828d3da9d61de4e6436bb00f0676fdfb2f3d6b9

Observation d1a26ad7-6b7d-4a65-8591-e0701589cbe3 · outbound

This paper cites ScreenQA: Large-Scale Question-Answer Pairs over Mobile App Screenshots.

KRETA: A Benchmark for Korean Reading and Reasoning in Text-Rich VQA Attuned to Diverse Visual Contexts ScreenQA: Large-Scale Question-Answer Pairs over Mobile App Screenshots

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-05T15:23:52.935815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:23:52.935815Z digest=sha256:ffe2afda3d31b9b1ef4499a78858dec85d97cfe3eecfed024ddfca3e76a0e267

Observation 5d9d3a6a-90e8-4631-96a9-f0f8dbe03073 · outbound

This paper cites In Proceedings of ACL.

KRETA: A Benchmark for Korean Reading and Reasoning in Text-Rich VQA Attuned to Diverse Visual Contexts In Proceedings of ACL

Reference 2023

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:23:55.427249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T15:23:53.538099Z digest=sha256:60ee67a1d8a434c6e348565e71938ea4182687e953ffce5340dc7c01b9d8212e

Observation cd2966e3-c467-46af-80c3-6a334040aac3 · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

KRETA: A Benchmark for Korean Reading and Reasoning in Text-Rich VQA Attuned to Diverse Visual Contexts Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-05T15:23:52.283285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:23:52.283285Z digest=sha256:b2dc729801b0af48194023d6112fa942a99f3eef2e1ff7f227a1d9b69c7ee15e

Observation 5cda32ee-1c9f-4d96-b821-c1e6a60c4538 · outbound

This paper cites Qwen2.5-VL Technical Report.

KRETA: A Benchmark for Korean Reading and Reasoning in Text-Rich VQA Attuned to Diverse Visual Contexts Qwen2.5-VL Technical Report

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-05T15:23:52.538063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:23:52.538063Z digest=sha256:2111526ba893b45bb853bea494e9a592bacd0f476c5969bdeebd77d789c45433

Pith citing papers

No inbound Pith citation observations are available.