Pith. sign in

Paper Citation Record · LEDGER

Analyze-Prompt-Reason: A Collaborative Agent-Based Framework for Multi-Image Vision-Language Reasoning

As of 8 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 0 inbound Pith citation observations for arXiv:2508.00356.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.00356 v1

Coverage vector

measured 24 of 24 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T10:14:26.501485Z

measured 24 of 24 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

24 of 24 outbound references displayed

  • verified exact7
  • verified fuzzy1
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 57c4ef17-6644-495a-a376-98a2eb7fe313 · outbound

This paper cites an unresolved cited work.

Analyze-Prompt-Reason: A Collaborative Agent-Based Framework for Multi-Image Vision-Language Reasoning Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-06T10:14:27.083225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:14:26.383981Z digest=sha256:8510c2a1f62342659ce5091d20b4947be0c76c360f730095e3d9c56c9b942e7f

Observation 7c949231-5f48-44ea-9a23-e4db810e6bbd · outbound

This paper cites an unresolved cited work.

Analyze-Prompt-Reason: A Collaborative Agent-Based Framework for Multi-Image Vision-Language Reasoning Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-06T10:14:27.067450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:14:26.395786Z digest=sha256:0040a8c890a7d70fcc518393e7d851a635b37f31fc5ec7718aaea9c58513d74f

Observation ec7a6b27-faeb-4549-9bc1-9262a518c81c · outbound

This paper cites iEdit: Localised Text-guided Image Editing with Weak Supervision.

Analyze-Prompt-Reason: A Collaborative Agent-Based Framework for Multi-Image Vision-Language Reasoning iEdit: Localised Text-guided Image Editing with Weak Supervision

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:14:26.955157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:14:26.400829Z digest=sha256:05d3db7ad2305b2b58bf70e01bfa72f9f7a20408f4ff2a0f0bc11808d2e971ca

Observation d3bdebc5-3888-494c-b572-a38bb4c572ba · outbound

This paper cites nuScenes: A multimodal dataset for autonomous driving.

Analyze-Prompt-Reason: A Collaborative Agent-Based Framework for Multi-Image Vision-Language Reasoning nuScenes: A multimodal dataset for autonomous driving

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T10:14:26.406143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:14:26.406143Z digest=sha256:b22a6905366baec094bd157b49be720f1748084b2221e3586e3e7d25062261c5

Observation 8fc48e60-3e2c-482b-a6c0-e0198b28d2de · outbound

This paper cites WebQA: Multihop and Multimodal QA.

Analyze-Prompt-Reason: A Collaborative Agent-Based Framework for Multi-Image Vision-Language Reasoning WebQA: Multihop and Multimodal QA

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T10:14:26.411573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:14:26.411573Z digest=sha256:d24ddb02b5c019d885ffec3e33057a67f92e5fdce064f6cfba8c4b24d9aa16c3

Observation 718fb223-8954-4c7c-a9ed-14aa57031b23 · outbound

This paper cites an unresolved cited work.

Analyze-Prompt-Reason: A Collaborative Agent-Based Framework for Multi-Image Vision-Language Reasoning Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-06T10:14:27.051742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:14:26.417685Z digest=sha256:f4b2b7de64d9e48c9719a30880a7ed76a734567e95d3ac910f767eda01c19fe9

Observation 6a4d9f21-e11b-46b3-b594-55fede6f18aa · outbound

This paper cites Neural Naturalist: Generating Fine-Grained Image Comparisons.

Analyze-Prompt-Reason: A Collaborative Agent-Based Framework for Multi-Image Vision-Language Reasoning Neural Naturalist: Generating Fine-Grained Image Comparisons

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:14:26.900213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:14:26.422530Z digest=sha256:f7e4a7a2331b9c5e7c216fc53b38640c28bd234534d67f5bcb612cdfe658c66d

Observation bdc9d5da-fed6-44d1-9a24-02c2d089559a · outbound

This paper cites an unresolved cited work.

Analyze-Prompt-Reason: A Collaborative Agent-Based Framework for Multi-Image Vision-Language Reasoning Unresolved cited work

Reference 8

Resolution
verified exact
doi, observed 2026-08-06T10:14:26.555105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:14:26.428676Z digest=sha256:6c9d76c5f54dedc26f055b5173a98d6e9f51d91f020965155e393a887bfb0ca8

Observation f8835270-7cc0-44f3-b1a0-514a488701e8 · outbound

This paper cites VizWiz Grand Challenge: Answering Visual Questions from Blind People.

Analyze-Prompt-Reason: A Collaborative Agent-Based Framework for Multi-Image Vision-Language Reasoning VizWiz Grand Challenge: Answering Visual Questions from Blind People

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T10:14:26.433842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:14:26.433842Z digest=sha256:7113b5a39a890dcd0ff808d54917a8207da48f7dd3b619ae2a4cfb0ebe4b19c5

Observation 7afc94be-cd73-4ec8-bf06-3aedf9ce1de8 · outbound

This paper cites Automatic Spatially-aware Fashion Concept Discovery.

Analyze-Prompt-Reason: A Collaborative Agent-Based Framework for Multi-Image Vision-Language Reasoning Automatic Spatially-aware Fashion Concept Discovery

Reference 10

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T10:14:26.859611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:14:26.439127Z digest=sha256:94a1a4e33857c2701e1fee12d95c1a19edafb2d118aec2ec35da54c3723cfedd

Observation 03f9712d-06cc-4790-b42b-3eb47d297715 · outbound

This paper cites an unresolved cited work.

Analyze-Prompt-Reason: A Collaborative Agent-Based Framework for Multi-Image Vision-Language Reasoning Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-06T10:14:27.036541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:14:26.443906Z digest=sha256:fd8a320c5ad8f37d59ecc6c657efe1271c03b91d5e44cf299479d339e9c46b26

Observation 61f18b4e-39f1-4953-8d1b-a49da1445d0f · outbound

This paper cites Lim, and Edward H.

Analyze-Prompt-Reason: A Collaborative Agent-Based Framework for Multi-Image Vision-Language Reasoning Lim, and Edward H

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:14:27.020400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:14:26.448708Z digest=sha256:f6f65b961cfbd4ff904468d6dc8e50a0420f35a05d616d1a2701cc75a3cfa5af

Observation eec49654-8b85-42c3-8c4d-56a1da1665ee · outbound

This paper cites an unresolved cited work.

Analyze-Prompt-Reason: A Collaborative Agent-Based Framework for Multi-Image Vision-Language Reasoning Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-06T10:14:27.004968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:14:26.453408Z digest=sha256:ccf2c89d20e578cc3edd1e9f7eb5148acc1c24c6282f357556af073768eca8b4

Observation 8038a4a9-d55b-42fd-9584-5c218f8e5302 · outbound

This paper cites Textbook Question Answering with Multi-modal Context Graph Understanding and Self-supervised Open-set Comprehension.

Analyze-Prompt-Reason: A Collaborative Agent-Based Framework for Multi-Image Vision-Language Reasoning Textbook Question Answering with Multi-modal Context Graph Understanding and Self-supervised Open-set Comprehension

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T10:14:26.457867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:14:26.457867Z digest=sha256:a8e2c4280610686ce37281a0e354e9e54292b008e573448c3d09deff8d1fbb79

Observation 0e01d4e9-522d-4e50-a6b0-fba0710f9d9e · outbound

This paper cites an unresolved cited work.

Analyze-Prompt-Reason: A Collaborative Agent-Based Framework for Multi-Image Vision-Language Reasoning Unresolved cited work

Reference 15

Resolution
verified exact
doi, observed 2026-08-06T10:14:26.539249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:14:26.462613Z digest=sha256:ab2be17edce049c7ba591f4afcada1acd6a01773a088921b76b21836e36b1829

Observation 43534105-71b5-4352-9c37-10cd104a6e2f · outbound

This paper cites an unresolved cited work.

Analyze-Prompt-Reason: A Collaborative Agent-Based Framework for Multi-Image Vision-Language Reasoning Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-06T10:14:26.988233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:14:26.467296Z digest=sha256:707f7e23eb5ca8f0af84715de80b33d54341ab9de7e79c9f7c53aac01c596074

Observation b8386a33-20c2-4723-803b-f2ef52d0b13a · outbound

This paper cites DocVQA: A Dataset for VQA on Document Images.

Analyze-Prompt-Reason: A Collaborative Agent-Based Framework for Multi-Image Vision-Language Reasoning DocVQA: A Dataset for VQA on Document Images

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T10:14:26.472521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:14:26.472521Z digest=sha256:7e8a6ca9f4373f0fb918b7d6aa0d3a9df06ce54b6a01990c9724c92eedccfe09

Observation 4412677d-676f-48ff-98b2-2ff2f9f680b0 · outbound

This paper cites an unresolved cited work.

Analyze-Prompt-Reason: A Collaborative Agent-Based Framework for Multi-Image Vision-Language Reasoning Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T10:14:26.478099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:14:26.478099Z digest=sha256:c2634cedc6f4b535e6175d3466eeaaa79738c245eb1fd26ebe78a8da3ca0f018

Observation 22fbebb8-7b81-47b5-92da-e90b3b048edc · outbound

This paper cites Robust Change Captioning.

Analyze-Prompt-Reason: A Collaborative Agent-Based Framework for Multi-Image Vision-Language Reasoning Robust Change Captioning

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:14:26.721542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:14:26.482824Z digest=sha256:408d665744a741a11526ee4bc61e26b3ee22a41048f3c5ac5bf0ac2c3bbc7ce9

Observation d74dbec7-1fe0-4d7c-a424-92774c0d32e7 · outbound

This paper cites an unresolved cited work.

Analyze-Prompt-Reason: A Collaborative Agent-Based Framework for Multi-Image Vision-Language Reasoning Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T10:14:26.487693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:14:26.487693Z digest=sha256:39954db8862884aeda76f841690ee7b181ac33f20daeaa9b8b9a0128f4de69aa

Observation f6f0a749-f489-4708-9d49-0fffae55ef97 · outbound

This paper cites Totally Looks Like - How Humans Compare, Compared to Machines.

Analyze-Prompt-Reason: A Collaborative Agent-Based Framework for Multi-Image Vision-Language Reasoning Totally Looks Like - How Humans Compare, Compared to Machines

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:14:26.617394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:14:26.492057Z digest=sha256:4bc77e325a94afa59285bf1aa041f62a62391f8665de1740744726f1a698175a

Observation eca3f955-72b0-4d9c-875b-545a2c56042b · outbound

This paper cites ALFRED: A Benchmark for Interpreting Grounded Instructions for Everyday Tasks.

Analyze-Prompt-Reason: A Collaborative Agent-Based Framework for Multi-Image Vision-Language Reasoning ALFRED: A Benchmark for Interpreting Grounded Instructions for Everyday Tasks

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T10:14:26.496776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:14:26.496776Z digest=sha256:16b06898c882e225d1cc99d08a4d3daea97fc42639acb1516f8be72118644d4f

Observation 64e5c2e6-8ee0-4e94-9298-2f298e1ed724 · outbound

This paper cites NLVR2 Visual Bias Analysis.

Analyze-Prompt-Reason: A Collaborative Agent-Based Framework for Multi-Image Vision-Language Reasoning NLVR2 Visual Bias Analysis

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:14:26.577064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:14:26.501485Z digest=sha256:554ba09508b686d0fb27eec158a7e69c4f5a4b597e660847117fba2aab051e5b

Observation 2c619c72-0063-459e-a0b1-8cc98d7fccc2 · outbound

This paper cites VISION Datasets: A Benchmark for Vision-based InduStrial InspectiON.

Analyze-Prompt-Reason: A Collaborative Agent-Based Framework for Multi-Image Vision-Language Reasoning VISION Datasets: A Benchmark for Vision-based InduStrial InspectiON

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T10:14:26.389640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:14:26.389640Z digest=sha256:b8d77a3880afc538d3e5d8683f314a320ca01da345311accf2c4ecfe8e368c2d

Pith citing papers

No inbound Pith citation observations are available.