Pith. sign in

Paper Citation Record · LEDGER

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning

As of 21 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 2 inbound Pith citation observations for arXiv:2412.13540.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.13540 v3

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T13:06:05.361783Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T13:59:28.143143Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T11:40:58.082760Z

Reference resolution

42 of 42 outbound references displayed

  • verified exact4
  • verified fuzzy1
  • unresolved37
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f0b64821-21f1-4a2d-bf4d-4975c4bd173b · outbound

This paper cites online" 'onlinestring :=.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T13:06:05.155478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:06:05.155478Z digest=sha256:4dcb4445a182fc27f8a3e309680ed73dc682024657ecb653d8434b4ef12155d8

Observation 80ece96f-74fe-4b5e-a5ac-72e4c9f7d444 · outbound

This paper cites write newline.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T13:06:05.160717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:06:05.160717Z digest=sha256:5186e9a89d04b4c577d429bbe53f3988f76a7c329987e227fea06b25a093e3ef

Observation fb8ad8b8-1490-467b-bef8-74891ae02f8f · outbound

This paper cites an unresolved cited work.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning Unresolved cited work

Reference 3

Resolution
verified exact
doi, observed 2026-08-11T13:06:05.531000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T13:06:05.165543Z digest=sha256:87c7ec20b22fbb3980b17b82211b912e0492c889f9032c36392f7a6516ce6bf4

Observation c415a1b8-6f9c-40a0-84e6-dedff9162f6e · outbound

This paper cites an unresolved cited work.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:06:06.349536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T13:06:05.170648Z digest=sha256:f8eee083f6c47d139bfbace2a89ccf14b367548d824993c97f45b29d1069fce3

Observation fefaee8a-5f13-42cf-ab50-615b69fca54e · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T13:06:05.175354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:06:05.175354Z digest=sha256:cb75b4a52890574b2b4ce7badc08c664153d894e02c0efa451b30dcc2672325b

Observation 9a2ee92d-90b2-49a7-9391-1e2576ff6f4a · outbound

This paper cites an unresolved cited work.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T13:06:05.180596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:06:05.180596Z digest=sha256:abcc8417a64ef2dc8b9ec197e514261555cffea5b0e06a492ebcccb061743f58

Observation 6ae90d98-5b94-4ef3-a6e8-90c9bfb42584 · outbound

This paper cites an unresolved cited work.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning Unresolved cited work

Reference 7

Resolution
verified exact
doi, observed 2026-08-11T13:06:05.516215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T13:06:05.184838Z digest=sha256:558900c64c8ab88763d64a6a2a3c2c203e8fae7fb16330fb3fcbfae63aba4b37

Observation 1f55244a-1dff-42fa-b3c7-a57eccd79f9f · outbound

This paper cites MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T13:06:05.189248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:06:05.189248Z digest=sha256:a404e5c7c133a1cd37ad944dea5928fde500d2489e290f7dd12392f251ec0811

Observation 39e6e16e-12d2-4d17-9884-d8abe65c7f36 · outbound

This paper cites an unresolved cited work.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:06:06.321229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T13:06:05.193686Z digest=sha256:eb65d6c75bbbc09d511bde0a8cc4c83d0c5ff4a70f1264f23cb721b07909da20

Observation e0330989-3435-4a80-a926-a7b7c3736ff7 · outbound

This paper cites an unresolved cited work.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:06:06.299286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T13:06:05.197558Z digest=sha256:5ed83f8b29c163c20dceca325ba1b6211b9a30cf61c988b56c3ab6e5e45d75cb

Observation a9a96c77-80a8-4aee-96b1-dcc14f997ac1 · outbound

This paper cites an unresolved cited work.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:06:06.279814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T13:06:05.201706Z digest=sha256:6b40cda9c5046d43ab24b82406324d0b06a5c2d0bde39f1a29b7f09128d36958

Observation 851fd589-97b2-4be4-bfee-9dd000def835 · outbound

This paper cites an unresolved cited work.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:06:06.261796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T13:06:05.207828Z digest=sha256:35cc15673fe7fcf29f6f7654c9414b3888f0c300a5ff4c1f0ada1d959f3a7a11

Observation 0c5c3514-b454-4ea0-ace4-e3758bafacfb · outbound

This paper cites Test of Time: A Benchmark for Evaluating LLMs on Temporal Reasoning.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning Test of Time: A Benchmark for Evaluating LLMs on Temporal Reasoning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T13:06:05.211951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:06:05.211951Z digest=sha256:beaa39adf44a50142ab8aa08425873cda12fd8999755b1a5e1a4015de1077f4a

Observation 09ef2d7f-81e4-4bb3-91cd-26284a77bc71 · outbound

This paper cites an unresolved cited work.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:06:06.249344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T13:06:05.217593Z digest=sha256:dc93b14b74947fc233e37db3d06f4312e099a3ae6ab25862c30170ad9ec5cd90

Observation eb42abdf-69bf-4d4f-867e-8e99739e837f · outbound

This paper cites an unresolved cited work.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T13:06:05.221688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:06:05.221688Z digest=sha256:876a1da88b1e0c6b1a55e0bf48f29920f946aee11fd1fd6595a30da96f006114

Observation 3b593d02-2620-4ee1-a3df-a9daceebbb07 · outbound

This paper cites an unresolved cited work.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T13:06:05.226205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:06:05.226205Z digest=sha256:8953d6a920705b7f1618162b295daa3f275f87a3b6b9f0a7ccc4eef4f70ffedc

Observation 0951bcb4-2430-4566-abdb-e89a7166458f · outbound

This paper cites Large Language Models Are Cross-Lingual Knowledge-Free Reasoners.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning Large Language Models Are Cross-Lingual Knowledge-Free Reasoners

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T13:06:05.231030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:06:05.231030Z digest=sha256:5fb2116817de41934af76d86ea7671bd319fcc72d969f588cff06fd9aa7f7e5f

Observation c4640a8d-8198-4f88-9bf7-2dd7da615c9d · outbound

This paper cites an unresolved cited work.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T13:06:05.236301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:06:05.236301Z digest=sha256:420004ec4dbca29120b75e5c15b30b1bff9c29a220c67853faf9f1fe64b8e09a

Observation b3ca1c2a-cd24-4645-9bbf-26070bdd2d39 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning LLaVA-OneVision: Easy Visual Task Transfer

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T13:06:05.241791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:06:05.241791Z digest=sha256:fa38b495d941b1d4daaa04c5fd08a5aadafee07ff5aa178b63f2aa9cf2e9cda9

Observation 812a7e48-71d0-4a0e-bf51-5f28f034c6bc · outbound

This paper cites LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T13:06:05.246904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:06:05.246904Z digest=sha256:c2f94acc9a41b2c0dd4d2d7b938af52642f7a9c54a346e695f594c8978cf602e

Observation aa658bcf-0e45-431d-8514-d1244ea9801d · outbound

This paper cites an unresolved cited work.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T13:06:05.252051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:06:05.252051Z digest=sha256:55ec5222c8056ddab8ccaf6da376ca18d67eadbffafde0e63db8d7a5a4a47657

Observation c146ea4f-718c-47c8-9bbc-c0fa0dbf233e · outbound

This paper cites an unresolved cited work.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:06:06.218265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T13:06:05.255865Z digest=sha256:b9dd4c9142e17551db4a891d4bba8f99a0607e3caa84c47a8ea1ef7840aa9f82

Observation 12899ee2-1640-473e-9f13-4ccdf48c2a12 · outbound

This paper cites an unresolved cited work.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:06:06.196011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T13:06:05.259791Z digest=sha256:5f07290af550e7b3ef685780998378e046f82919aaddd48ddb3839e098123cc0

Observation eb1825ca-8701-42f8-9d45-0bf9e431713f · outbound

This paper cites SPHINX: The Joint Mixing of Weights, Tasks, and Visual Embeddings for Multi-modal Large Language Models.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning SPHINX: The Joint Mixing of Weights, Tasks, and Visual Embeddings for Multi-modal Large Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T13:06:05.263535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:06:05.263535Z digest=sha256:7d1f18556deed93328fa350bd771e0a65c20923f606e96c763f81076ca035f0c

Observation 70892786-e5c1-412f-afac-12e2f45151f3 · outbound

This paper cites an unresolved cited work.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T13:06:05.267381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:06:05.267381Z digest=sha256:5c6eb5f04fb24059a5dd4f4160624ca8bcdf5accd841c014a67ce41a29d8b652

Observation d80c0aea-f929-415d-9b21-e5395a06b6a8 · outbound

This paper cites an unresolved cited work.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T13:06:05.270951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:06:05.270951Z digest=sha256:f8a0c9ecf2fafe4fb63fe10661aafce5d86e74383d79a0c3bf4e85aae1cf8a20

Observation 2ea82d55-e2f0-4eea-8b79-50459516c335 · outbound

This paper cites Melo, Ana Paiva, and Danica Kragic.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning Melo, Ana Paiva, and Danica Kragic

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:06:06.143187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T13:06:05.280439Z digest=sha256:7ce277ff459b3a6cadb5890d9a71a42fa7f4c258511fe95ee100b49f19ee7d01

Observation 943a7003-d8b5-47d1-853d-b3f414e5a7ca · outbound

This paper cites an unresolved cited work.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T13:06:05.286136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:06:05.286136Z digest=sha256:b19e88a2dec3d0a6f4e8d4d01ade9d0c45f0b2fa770fc0ecccf9b67fc57c3362

Observation 0e7e5b7e-3d35-4e45-b1c1-9693a27de23a · outbound

This paper cites Velimsky, Robert Els\" a sser, and Bernhard C.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning Velimsky, Robert Els\" a sser, and Bernhard C

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T13:06:05.290444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:06:05.290444Z digest=sha256:3cee0a7907224536cf2ca50e03f85468915684982891004c2c3280574a543ec3

Observation 595a7311-5b37-4be5-8db3-26f8462dbb3e · outbound

This paper cites an unresolved cited work.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T13:06:05.295274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:06:05.295274Z digest=sha256:57c42a0701eb80b4bc1b1bd4084ac607d5b547d279883a67b22070abf0306f71

Observation 4d33372e-7b58-482f-9de6-46c27e7dd07f · outbound

This paper cites an unresolved cited work.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning Unresolved cited work

Reference 31

Resolution
verified exact
doi, observed 2026-08-11T13:06:05.464607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T13:06:05.300885Z digest=sha256:bc9d7a19b398e370f7238113f44faab1c758f72405859592964ef875dbd12e0f

Observation 8222af83-de49-4666-815e-ef16cd5c4412 · outbound

This paper cites GraphArena: Evaluating and Exploring Large Language Models on Graph Computation.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning GraphArena: Evaluating and Exploring Large Language Models on Graph Computation

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T13:06:05.307592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:06:05.307592Z digest=sha256:bae1ed2f099adca6f6ab9b9f519fffbb020809ed89a9160718233a6d7de45010

Observation d9de79e4-64ae-4a1d-8fe8-e9ae5cf89fe1 · outbound

This paper cites an unresolved cited work.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T13:06:05.315672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:06:05.315672Z digest=sha256:3ab330b443746e507b05e2661c95a3595bf004f10974cd2d3c2d6ded81874831

Observation 5fdf3641-f15a-4bb3-a5ed-e777628e2971 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T13:06:05.320717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:06:05.320717Z digest=sha256:6ba8a258737e91258e8ab276df6be250532d14a144e3b7bea5a85390db6b1532

Observation c32babc7-5d74-4559-a164-2be010aad900 · outbound

This paper cites an unresolved cited work.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:06:06.115490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T13:06:05.326634Z digest=sha256:ff29cf8f3e867124fffbb3a68e357a79092c12316fca9f99ea8e155e22711a21

Observation 0d145824-ebe5-457b-9316-df8d59d70ef1 · outbound

This paper cites mPLUG-Owl3: Towards Long Image-Sequence Understanding in Multi-Modal Large Language Models.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning mPLUG-Owl3: Towards Long Image-Sequence Understanding in Multi-Modal Large Language Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T13:06:05.333604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:06:05.333604Z digest=sha256:550bc3e7bdc60b81505b86ac403c9f59bd3a734baee29cd42d642da7395ad17b

Observation 334e7b8c-73f2-40db-a409-b57e0a09cd4e · outbound

This paper cites InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T13:06:05.339490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:06:05.339490Z digest=sha256:94c05aac9b7380cab253f27cee121d33f40eb154af7cef1bb354af47eb100b1f

Observation 41312f96-79bf-41d7-aa10-9299f6b2ed98 · outbound

This paper cites an unresolved cited work.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T13:06:05.343947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:06:05.343947Z digest=sha256:a6262553fb89e4d70f3c7b2843d81da1654e34d52c5f965c800c60c434b73b13

Observation 8fea1a0c-a3e6-44a7-a5fe-185f406db8e6 · outbound

This paper cites an unresolved cited work.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T13:06:05.348632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:06:05.348632Z digest=sha256:21024ccf64af8894751765154adb77dd415866dc7e36c90a4dea2d3be96ceb1b

Observation 7b4d83ab-fedc-4149-92c0-71035cce7095 · outbound

This paper cites an unresolved cited work.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning Unresolved cited work

Reference 40

Resolution
verified exact
doi, observed 2026-08-11T13:06:05.415668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T13:06:05.352958Z digest=sha256:19772c152bd5c309561f27545153f77e8cddbc58106a1dd17233e7deff8946ff

Observation d54f8523-9fd2-4c2e-9688-cfe3e8c2a10c · outbound

This paper cites an unresolved cited work.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-11T13:06:05.357095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:06:05.357095Z digest=sha256:1ca195e732173e1c48aae7d1e5834be771c110beb6360a950f6ef1cdf43b6adb

Observation 5eb30745-0155-4c61-890b-54bfe2f39bc8 · outbound

This paper cites an unresolved cited work.

Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-11T13:06:05.361783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:06:05.361783Z digest=sha256:1309114d0d9616c33d9dfcce3b6ab81d9be3ee8d957e35f5b1328b637be2d020

Pith citing papers

Observation 099b569b-d471-4dd7-98eb-d28046c78c52 · inbound

Make Imagination Clearer! Stable Diffusion-based Visual Imagination for Multimodal Machine Translation cites this paper.

Make Imagination Clearer! Stable Diffusion-based Visual Imagination for Multimodal Machine Translation Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning

Reference 2697

Resolution
unresolved
no resolver link, observed 2026-08-11T13:59:28.143143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:59:28.143143Z digest=sha256:ef745080f0c8bf7f723e02fa0180fb8c13fdfc86583c99e05457337aa907b21e

Observation 450120ed-bfc4-49b7-80f7-3b98e6bdb5e7 · inbound

Thinking in Character: Advancing Role-Playing Agents with Role-Aware Reasoning cites this paper.

Thinking in Character: Advancing Role-Playing Agents with Role-Aware Reasoning Benchmarking and Improving Large Vision-Language Models for Fundamental Visual Graph Understanding and Reasoning

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:40:58.093263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:40:53.230718Z digest=sha256:03538ad5553ab5b6b4aaebf60972d80736e2eeff4cdbd9e4ca56b89159da1a11