Pith. sign in

Paper Citation Record · LEDGER

Vision Language Models See What You Want but not What You See

As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2410.00324.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.00324 v6

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T17:26:36.012121Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T22:29:09.850744Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c7333a23-3e30-407f-bacc-5dbf01ef70d6 · inbound

Explainability for Vision Foundation Models: A Survey cites this paper.

Explainability for Vision Foundation Models: A Survey Vision Language Models See What You Want but not What You See

Reference 259

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:36.012121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:36.012121Z digest=sha256:54d831d602f0530411a210329e43fbc3f0a441a18aeb4eb2c1049cb58308d04d

Observation c2d92f78-91cf-41c2-82ed-5306adaac3bc · inbound

Towards Embodied Cognition in Robots via Spatially Grounded Synthetic Worlds cites this paper.

Towards Embodied Cognition in Robots via Spatially Grounded Synthetic Worlds Vision Language Models See What You Want but not What You See

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T15:37:57.959908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:37:57.959908Z digest=sha256:c2debc62350cc3af7c38d43b445e505a75953c3d22d99fb10faf4b5de2012a21

Observation 21c28773-bd82-47f5-9e58-896f91791d1d · inbound

Egocentric Bias in Vision-Language Models cites this paper.

Egocentric Bias in Vision-Language Models Vision Language Models See What You Want but not What You See

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T03:04:43.096093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:04:43.096093Z digest=sha256:c2a7c560a255c827af9ac77e776303553bbc05452d28db5022cd01ffb08d1551

Observation 25bda350-029f-460d-959f-f0394c36f4b5 · inbound

Beyond Localization: A Comprehensive Diagnosis of Perspective-Conditioned Spatial Reasoning in MLLMs from Omnidirectional Images cites this paper.

Beyond Localization: A Comprehensive Diagnosis of Perspective-Conditioned Spatial Reasoning in MLLMs from Omnidirectional Images Vision Language Models See What You Want but not What You See

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:57:27.411345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-13T06:56:37.792662Z digest=sha256:0a5f5bd448d6b388ba53cb8fadba290d1e9db13cd5a295028e633666de3fe40c

Observation 0d9e3f85-ea0b-40c9-90b0-a5cc0b01eb35 · inbound

Beyond Localization: A Comprehensive Diagnosis of Perspective-Conditioned Spatial Reasoning in MLLMs from Omnidirectional Images cites this paper.

Beyond Localization: A Comprehensive Diagnosis of Perspective-Conditioned Spatial Reasoning in MLLMs from Omnidirectional Images Vision Language Models See What You Want but not What You See

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:39:29.056071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-14T21:38:09.573431Z digest=sha256:8b10a8e52970a651731b572f308df2f68a33e02a3c85f4cd32e4662d92069546

Observation fe7ae660-acf8-45df-bc89-221707fadb7a · inbound

Beyond Localization: A Comprehensive Diagnosis of Perspective-Conditioned Spatial Reasoning in MLLMs from Omnidirectional Images cites this paper.

Beyond Localization: A Comprehensive Diagnosis of Perspective-Conditioned Spatial Reasoning in MLLMs from Omnidirectional Images Vision Language Models See What You Want but not What You See

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-20T22:29:09.853037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-20T22:24:38.232462Z digest=sha256:3586b3513b268fde42b6d5f5ce4c69f1c17b44e525f0eee134b615188af3c088