Pith. sign in

Paper Citation Record · LEDGER

Beyond Visual Understanding: Introducing PARROT-360V for Vision Language Model Benchmarking

As of 14 August 2026, this Paper Citation Record lists 19 of 19 outbound references and 0 inbound Pith citation observations for arXiv:2411.15201.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.15201 v1

Coverage vector

measured 19 of 19 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T17:05:29.568947Z

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

19 of 19 outbound references displayed

  • verified exact2
  • verified fuzzy0
  • unresolved17
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a50b815b-ee9a-401f-ad57-623a48141d41 · outbound

This paper cites online" 'onlinestring :=.

Beyond Visual Understanding: Introducing PARROT-360V for Vision Language Model Benchmarking online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T17:05:29.514856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:05:29.514856Z digest=sha256:0fa1a15b93e1545325e60b5c4606e55ae6e49afb270bdcc563dd55e19783633d

Observation 72867fcf-136c-4c32-95c1-bd0478f264a2 · outbound

This paper cites write newline.

Beyond Visual Understanding: Introducing PARROT-360V for Vision Language Model Benchmarking write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T17:05:29.518129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:05:29.518129Z digest=sha256:90cacddfc2e6190925fd56365780c3eda4939f98572073bce0574695dd68632b

Observation 7943c39f-5f6e-4d50-927a-b77c4caff949 · outbound

This paper cites How Susceptible are LLMs to Influence in Prompts?.

Beyond Visual Understanding: Introducing PARROT-360V for Vision Language Model Benchmarking How Susceptible are LLMs to Influence in Prompts?

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T17:05:29.520919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:05:29.520919Z digest=sha256:0c8f01a131d73f0094c41b2050a09222f7f53c4f9f857a5f662f2e64ced728d2

Observation 38d3f405-07bd-4c5c-831e-4e1407805306 · outbound

This paper cites an unresolved cited work.

Beyond Visual Understanding: Introducing PARROT-360V for Vision Language Model Benchmarking Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T17:05:29.523863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:05:29.523863Z digest=sha256:81f19141dd72ef60ca03e78f9bbfaf906c65a55801c8fe280458a2e8e30ea475

Observation fa7b33ce-7e6b-46f1-a249-0d61e2c457b3 · outbound

This paper cites an unresolved cited work.

Beyond Visual Understanding: Introducing PARROT-360V for Vision Language Model Benchmarking Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T17:05:29.527084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:05:29.527084Z digest=sha256:942fd1ef4c4a3003c30e9fc9bfb6ee2d6d1ac0f95aed8a0fdbe3c43c61fdbb87

Observation 1ceac050-f95e-4418-9e91-f931a51fc0b4 · outbound

This paper cites A Diagram Is Worth A Dozen Images.

Beyond Visual Understanding: Introducing PARROT-360V for Vision Language Model Benchmarking A Diagram Is Worth A Dozen Images

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T17:05:29.530090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:05:29.530090Z digest=sha256:b68f51f7776c7cae16c4eea2f4dae5da761d6ed9716da89905618675cb147737

Observation ead30a38-8a77-4778-a5dd-5a47d4bee969 · outbound

This paper cites ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning.

Beyond Visual Understanding: Introducing PARROT-360V for Vision Language Model Benchmarking ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T17:05:29.533185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:05:29.533185Z digest=sha256:b4e2d62bf1f0c1594542b61566f75368f2a6d70c635c74d609da616e51147046

Observation c9f8cc8f-4ce3-4813-8ec7-8e7910634cdb · outbound

This paper cites an unresolved cited work.

Beyond Visual Understanding: Introducing PARROT-360V for Vision Language Model Benchmarking Unresolved cited work

Reference 8

Resolution
verified exact
raw_fallback, observed 2026-08-12T17:05:29.713825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-12T17:05:29.536469Z digest=sha256:e66d123d7f09f2f8ea6a15f331017c6a595d78e9437f0277ebaddf3a83efb356

Observation cbb32688-9ecc-4d4e-83fa-077b4d7dbba2 · outbound

This paper cites an unresolved cited work.

Beyond Visual Understanding: Introducing PARROT-360V for Vision Language Model Benchmarking Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-12T17:05:29.748519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-12T17:05:29.539558Z digest=sha256:e4e4f070b6c801010cc2ad45222b3c4a6873b88b6f293a0825a732b853db890f

Observation a86f3c66-d3e3-47e7-947b-75ebb55d69f5 · outbound

This paper cites Towards Data Contamination Detection for Modern Large Language Models: Limitations, Inconsistencies, and Oracle Challenges.

Beyond Visual Understanding: Introducing PARROT-360V for Vision Language Model Benchmarking Towards Data Contamination Detection for Modern Large Language Models: Limitations, Inconsistencies, and Oracle Challenges

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T17:05:29.543359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:05:29.543359Z digest=sha256:35a163990a1324e944a166de9a2ef7cfe7540fce775117394bdc155aa210f5d9

Observation 56057d6d-cec6-496d-8f46-8be555afd0c7 · outbound

This paper cites Enhancing Trust in LLM-Based AI Automation Agents: New Considerations and Future Challenges.

Beyond Visual Understanding: Introducing PARROT-360V for Vision Language Model Benchmarking Enhancing Trust in LLM-Based AI Automation Agents: New Considerations and Future Challenges

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T17:05:29.546552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:05:29.546552Z digest=sha256:e5557dd0f1d823b43eaed5e5f5e9680e2b953ede69889ac19fd38defcbe490a4

Observation d8c64b50-2f5c-4cca-9319-3864444ff4ca · outbound

This paper cites ODE: Open-Set Evaluation of Hallucinations in Multimodal Large Language Models.

Beyond Visual Understanding: Introducing PARROT-360V for Vision Language Model Benchmarking ODE: Open-Set Evaluation of Hallucinations in Multimodal Large Language Models

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-08-12T17:05:29.643242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-12T17:05:29.550308Z digest=sha256:522b80bb338a4ae3e28e1c0bb26ddb2106c173bad879e56c8c838eacff2e6cd4

Observation e7ed00f1-719c-48bf-a3db-6ade57f9bc00 · outbound

This paper cites Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning.

Beyond Visual Understanding: Introducing PARROT-360V for Vision Language Model Benchmarking Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T17:05:29.553692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:05:29.553692Z digest=sha256:7ca36313a6fc3e62bff33701e0d7e72cf5bd53155de23ef67010585f5a306311

Observation fac9e23d-29b5-4e23-a62a-1a1a1ab71a30 · outbound

This paper cites Chain-of-Thought Prompting Elicits Reasoning in Large Language Models.

Beyond Visual Understanding: Introducing PARROT-360V for Vision Language Model Benchmarking Chain-of-Thought Prompting Elicits Reasoning in Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T17:05:29.556277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:05:29.556277Z digest=sha256:01e741062bf567f00e604b6480537a04dacc237bd363cb91a360ebb2dbe841d1

Observation 54ba0670-dd8d-4bf3-91fa-110e3655d6af · outbound

This paper cites an unresolved cited work.

Beyond Visual Understanding: Introducing PARROT-360V for Vision Language Model Benchmarking Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T17:05:29.559049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:05:29.559049Z digest=sha256:4c5b6e81ed71c8ea4cd99830822efc2afa6993ddd7c3c5b0d7269785542ac0a3

Observation 65fce983-7336-4619-a91d-254b60dda55a · outbound

This paper cites On the Vulnerability of LLM/VLM-Controlled Robotics.

Beyond Visual Understanding: Introducing PARROT-360V for Vision Language Model Benchmarking On the Vulnerability of LLM/VLM-Controlled Robotics

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T17:05:29.561341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:05:29.561341Z digest=sha256:6d4fe9383886e21d047c27205a1f1834e7a892108723833f34fc80d6ad4cbe64

Observation e9f988b2-2ba3-4580-945b-28b375dd7fd5 · outbound

This paper cites Model Merging in LLMs, MLLMs, and Beyond: Methods, Theories, Applications and Opportunities.

Beyond Visual Understanding: Introducing PARROT-360V for Vision Language Model Benchmarking Model Merging in LLMs, MLLMs, and Beyond: Methods, Theories, Applications and Opportunities

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T17:05:29.563777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:05:29.563777Z digest=sha256:1639d20a239fc6f40c5b8cbb969168a7bcbd368c65a7e1f8e86a7b4c26068801

Observation 34d4364f-2f49-40c9-afab-56daad0c10d6 · outbound

This paper cites Decompose and Compare Consistency: Measuring VLMs' Answer Reliability via Task-Decomposition Consistency Comparison.

Beyond Visual Understanding: Introducing PARROT-360V for Vision Language Model Benchmarking Decompose and Compare Consistency: Measuring VLMs' Answer Reliability via Task-Decomposition Consistency Comparison

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T17:05:29.566320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:05:29.566320Z digest=sha256:c44626094cb66075ae2ebe9590d1beac53122ef353c2bef086863ef59369c972

Observation 267bbf16-f600-4915-8948-e7f80b7c6986 · outbound

This paper cites MMMU: A Massive Multi-discipline Multimodal Understanding and Reasoning Benchmark for Expert AGI.

Beyond Visual Understanding: Introducing PARROT-360V for Vision Language Model Benchmarking MMMU: A Massive Multi-discipline Multimodal Understanding and Reasoning Benchmark for Expert AGI

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T17:05:29.568947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:05:29.568947Z digest=sha256:d14daafac36a9eb23643dd315e2f1ce0cfdbec14522f403389d56c56733b8596

Pith citing papers

No inbound Pith citation observations are available.