Pith. sign in

Paper Citation Record · LEDGER

VisText: A Benchmark for Semantically Rich Chart Captioning

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2307.05356.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2307.05356 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T20:13:47.670655Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T17:53:46.746830Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 7a7cd7fd-8a48-4bd7-9b16-11f97db346c7 · inbound

Otter: A Multi-Modal Model with In-Context Instruction Tuning cites this paper.

Otter: A Multi-Modal Model with In-Context Instruction Tuning VisText: A Benchmark for Semantically Rich Chart Captioning

Reference 80

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T02:43:47.859889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-15T02:43:47.775691Z digest=sha256:e00decc7f11ee9a4458baaa3923bb4b5c003a99827fb73b17b2f87af6549eb2f

Observation 410ddfd0-f1b3-4ba7-a7d1-86983ebf87a1 · inbound

Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning cites this paper.

Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning VisText: A Benchmark for Semantically Rich Chart Captioning

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T17:34:56.974820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-14T17:34:56.836034Z digest=sha256:2701f9947517f6f14d82b4fc527e336ff8b5191df9313044155f2846115a5462

Observation 2fd067ba-9792-4da9-8622-ec2eb13d51dc · inbound

Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs cites this paper.

Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs VisText: A Benchmark for Semantically Rich Chart Captioning

Reference 125

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T00:05:03.720092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-17T00:05:03.547664Z digest=sha256:85037fadb729fc750371fe1609d503bed7d6b67b282047b89f84d3d36cb3ff22

Observation 02c29023-905d-4a25-ac3c-be39f194f06a · inbound

Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling cites this paper.

Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling VisText: A Benchmark for Semantically Rich Chart Captioning

Reference 227

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T13:23:58.032587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-10T13:23:57.588851Z digest=sha256:3df6c8f9a7dec3c667cff794342377fbf4bbed7f0c6e8ad93860fdb14f04d03a

Observation f40e9638-e5d1-4fa9-889c-cb1e56e1b3a9 · inbound

Chimera: Improving Generalist Model with Domain-Specific Experts cites this paper.

Chimera: Improving Generalist Model with Domain-Specific Experts VisText: A Benchmark for Semantically Rich Chart Captioning

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-11T20:13:47.670655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:13:47.670655Z digest=sha256:c95eb65b1974921b5f83fdfbf45a5dce04f817861c098f034fe0e9fe888ea73e

Observation 8516a13f-ee44-4b43-9ecb-9822d129ab35 · inbound

MetaMorph: Multimodal Understanding and Generation via Instruction Tuning cites this paper.

MetaMorph: Multimodal Understanding and Generation via Instruction Tuning VisText: A Benchmark for Semantically Rich Chart Captioning

Reference 139

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T07:51:13.239657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-17T07:51:12.953777Z digest=sha256:949770e609ff3fe2f20831031dc9b6c8773c37bb5a19e3e0e34a63ef1fee4f14

Observation 47b3c0ef-c431-4302-9533-68fdf6dc9809 · inbound

Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models cites this paper.

Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models VisText: A Benchmark for Semantically Rich Chart Captioning

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-10T18:04:34.166684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:04:34.166684Z digest=sha256:7cbb1a8e2c427549b7b672453f4002f9d0f65026c08d81d7cfb428961506876b

Observation c72e89a8-60a6-4c19-b4c0-9cade7122bff · inbound

Pluto: Authoring Semantically Aligned Text and Charts for Data-Driven Communication cites this paper.

Pluto: Authoring Semantically Aligned Text and Charts for Data-Driven Communication VisText: A Benchmark for Semantically Rich Chart Captioning

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-08T11:47:45.080370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:47:45.080370Z digest=sha256:39d6d388dff942dfa718456433579c5b0ad84ab67e4c21e6360c27e6e704e0d8

Observation b15b11c2-bddf-4362-8445-2ccf572f18de · inbound

FLARE: Fully Integration of Vision-Language Representations for Deep Cross-Modal Understanding cites this paper.

FLARE: Fully Integration of Vision-Language Representations for Deep Cross-Modal Understanding VisText: A Benchmark for Semantically Rich Chart Captioning

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-22T19:52:01.842766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-22T19:49:00.961388Z digest=sha256:4823a5a527202676db3549eddde90f303b40849937977b15a7d868276188b525

Observation fee02811-32cc-4cbf-8c58-30ea7bd237dd · inbound

Chart-to-Experience: Benchmarking Multimodal LLMs for Predicting Experiential Impact of Charts cites this paper.

Chart-to-Experience: Benchmarking Multimodal LLMs for Predicting Experiential Impact of Charts VisText: A Benchmark for Semantically Rich Chart Captioning

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:53:09.700402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:53:09.700402Z digest=sha256:31298ba74b4817ee95b56e010d5291082c1051533cfda9fa05c3a65e35776407

Observation bf171219-4f17-444c-862b-8a273c2bd491 · inbound

ChartLens: Fine-grained Visual Attribution in Charts cites this paper.

ChartLens: Fine-grained Visual Attribution in Charts VisText: A Benchmark for Semantically Rich Chart Captioning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T14:19:24.218373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:19:24.218373Z digest=sha256:73065b16ba707576a802e2951851b389cde6ad94e4b40887986b42124a4b111f

Observation 40e0ab61-176a-48d7-ad58-facc1c2cd120 · inbound

Boosting Document Parsing Efficiency and Performance with Coarse-to-Fine Visual Processing cites this paper.

Boosting Document Parsing Efficiency and Performance with Coarse-to-Fine Visual Processing VisText: A Benchmark for Semantically Rich Chart Captioning

Reference 48

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T00:28:23.540411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-15T00:25:19.782732Z digest=sha256:7128819a969c5f38c7f64bf4299b018ed24570a816d0c920aa44ddad3476fb73

Observation dfc42b26-6ab1-4da9-90d6-ff1ccf5656b0 · inbound

LocateAnything: Fast and High-Quality Vision-Language Grounding with Parallel Box Decoding cites this paper.

LocateAnything: Fast and High-Quality Vision-Language Grounding with Parallel Box Decoding VisText: A Benchmark for Semantically Rich Chart Captioning

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T17:53:46.748360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-29T17:52:59.346637Z digest=sha256:2edee674a39f7892ba1222f5c0d759ed5a980d00139e32e198dd3ac7f10203cb

Observation 5f77c327-dc9a-4794-a710-ae630298d438 · inbound

ABot-OCR Technical Report cites this paper.

ABot-OCR Technical Report VisText: A Benchmark for Semantically Rich Chart Captioning

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:33:28.135686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-29T13:29:17.221676Z digest=sha256:d153098e33281552ff1b934407841918d2ad1d61d6fe51df44212f273879aa5b