Pith. sign in

Paper Citation Record · LEDGER

Visual Haystacks: A Vision-Centric Needle-In-A-Haystack Benchmark

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2407.13766.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.13766 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:49:53.475228Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T06:36:52.559941Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 03f03589-0a7e-44de-a1b4-023e511b63e0 · inbound

MAGNET: A Multi-agent Framework for Finding Audio-Visual Needles by Reasoning over Multi-Video Haystacks cites this paper.

MAGNET: A Multi-agent Framework for Finding Audio-Visual Needles by Reasoning over Multi-Video Haystacks Visual Haystacks: A Vision-Centric Needle-In-A-Haystack Benchmark

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-07T05:49:53.475228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:49:53.475228Z digest=sha256:a9264a3db6592d072aebb903e79215a772ef1283b81a8b23c953e65df7e4106d

Observation 1d64bf2a-1c1b-4886-bc56-f5b34b402675 · inbound

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation cites this paper.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Visual Haystacks: A Vision-Centric Needle-In-A-Haystack Benchmark

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:48.044179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:48.044179Z digest=sha256:33ba2aa3664ab457a8e670f39677c20c65c19f09e3f9bf3dd45b00f275a2ea50

Observation 9d3bcf66-c07c-44de-9637-691bed3f7b76 · inbound

Temporal Chain of Thought: Long-Video Understanding by Thinking in Frames cites this paper.

Temporal Chain of Thought: Long-Video Understanding by Thinking in Frames Visual Haystacks: A Vision-Centric Needle-In-A-Haystack Benchmark

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T21:06:26.795465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:06:26.795465Z digest=sha256:a64f8839ec0c76cccdb0243e2322baa83031254fafafb54b0b1a1118e22ad250

Observation e8e2e635-b3b7-4dc6-b0a6-020f38e29882 · inbound

Decoding the Pulse of Reasoning VLMs in Multi-Image Understanding Tasks cites this paper.

Decoding the Pulse of Reasoning VLMs in Multi-Image Understanding Tasks Visual Haystacks: A Vision-Centric Needle-In-A-Haystack Benchmark

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-15T16:06:14.482917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T16:04:58.851378Z digest=sha256:17e1978d0fda332abff8bccde9286086a5c991c2ccba9d60e3cb6c5169afc61c

Observation cb694d30-1942-4dcb-969d-6b8ee29c11e1 · inbound

Utility-Oriented Visual Evidence Selection for Multimodal Retrieval-Augmented Generation cites this paper.

Utility-Oriented Visual Evidence Selection for Multimodal Retrieval-Augmented Generation Visual Haystacks: A Vision-Centric Needle-In-A-Haystack Benchmark

Reference 40

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T19:09:22.942744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-14T19:09:18.975682Z digest=sha256:3151b5ea6b5629ac1a06d31a7bba65db8171aff342ea6dbe3e437d5331d2dbd3

Observation 2259e90d-95dd-4db2-83a6-3189f1a20789 · inbound

Training Long-Context Vision-Language Models Effectively with Generalization Beyond 128K Context cites this paper.

Training Long-Context Vision-Language Models Effectively with Generalization Beyond 128K Context Visual Haystacks: A Vision-Centric Needle-In-A-Haystack Benchmark

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-14T19:17:50.089420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T19:16:07.851098Z digest=sha256:318315c9cc04e7656928b6b5f2a7b59a92ceb68415d4e5dc42f1ec59db2a4aa6

Observation fbce73d7-06dd-444f-a247-b711ec15e776 · inbound

Personal Visual Memory from Explicit and Implicit Evidence cites this paper.

Personal Visual Memory from Explicit and Implicit Evidence Visual Haystacks: A Vision-Centric Needle-In-A-Haystack Benchmark

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-06-29T12:43:25.439201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T12:40:34.740804Z digest=sha256:2dd66a60e1fce139900409a3cec285c8ca5367ee0d68397233e71abaea88e633

Observation 71aab449-6c2f-4f8a-aedc-21992cb0b003 · inbound

Invoice Haystack: Benchmarking Document Retrieval and Visual Question Answering Under Strong Visual Homogeneity cites this paper.

Invoice Haystack: Benchmarking Document Retrieval and Visual Question Answering Under Strong Visual Homogeneity Visual Haystacks: A Vision-Centric Needle-In-A-Haystack Benchmark

Reference 70

Resolution
verified exact
arxiv_id, observed 2026-07-04T19:30:07.775820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-25T21:15:12.242955Z digest=sha256:b0310796f66900633ffea04bf6333633e3f969659ed24522933180f3f205bf6a

Observation c0c2a662-ca93-4faa-b0f8-66391d037bad · inbound

Invoice Haystack: Benchmarking Document Retrieval and Visual Question Answering Under Strong Visual Homogeneity cites this paper.

Invoice Haystack: Benchmarking Document Retrieval and Visual Question Answering Under Strong Visual Homogeneity Visual Haystacks: A Vision-Centric Needle-In-A-Haystack Benchmark

Reference 70

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:23:51.256632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T05:07:28.537679Z digest=sha256:3173b23e3493aab5503ca796b5ea0ebc00dea7932cbe90c9fdd7a28200dffccd

Observation 339be329-fadd-4f7f-a564-1a93d51be974 · inbound

C3-Bench: A Context-Aware Change Captioning Benchmark cites this paper.

C3-Bench: A Context-Aware Change Captioning Benchmark Visual Haystacks: A Vision-Centric Needle-In-A-Haystack Benchmark

Reference 93

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T19:50:10.146177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-25T21:02:52.529391Z digest=sha256:44753a17c5ce9ecae220d216066ac58652c53ed80ab501239c6274a51af5e90e

Observation 7fdc33c5-da4a-458a-b748-ea7865372227 · inbound

Cognitive-structured Multimodal Agent for Multimodal Understanding, Generation, and Editing cites this paper.

Cognitive-structured Multimodal Agent for Multimodal Understanding, Generation, and Editing Visual Haystacks: A Vision-Centric Needle-In-A-Haystack Benchmark

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-07-10T06:36:52.561097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-10T06:30:28.462917Z digest=sha256:7dcb4489ef44382c1ff7da224bd2719e28c98cc8875f974267f57eb947dd039c

Observation ebfad781-b4d7-45e9-82e1-51941c47bea5 · inbound

ReToken: One Token to Improve Vision-Language Models for Visual Retrieval cites this paper.

ReToken: One Token to Improve Vision-Language Models for Visual Retrieval Visual Haystacks: A Vision-Centric Needle-In-A-Haystack Benchmark

Reference 45

Resolution
unresolved
no resolver link, observed 2026-07-31T01:43:24.234557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T01:43:24.234557Z digest=sha256:8de997db5a3110ba7794fb1158c779f4161f387d9b251350ae18ac46e85ec89b