Pith. sign in

Paper Citation Record · LEDGER

Retrieval-Augmented Perception: High-Resolution Image Perception Meets Visual RAG

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2503.01222.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.01222 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:52:42.077508Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T07:49:38.567142Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 38f47b3e-ace7-4142-987b-d54b6bf574b0 · inbound

Wireless Agentic AI with Retrieval-Augmented Multimodal Semantic Perception cites this paper.

Wireless Agentic AI with Retrieval-Augmented Multimodal Semantic Perception Retrieval-Augmented Perception: High-Resolution Image Perception Meets Visual RAG

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:42.077508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:42.077508Z digest=sha256:7982a525352ad2652cae3d886a76206f5e6f2a5b1b32a09cae4b672c2551aa1b

Observation 92ed591e-d0b9-442d-9796-31b2a290a10c · inbound

AMIA: Automatic Masking and Joint Intention Analysis Makes LVLMs Robust Jailbreak Defenders cites this paper.

AMIA: Automatic Masking and Joint Intention Analysis Makes LVLMs Robust Jailbreak Defenders Retrieval-Augmented Perception: High-Resolution Image Perception Meets Visual RAG

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:28.937202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:28.937202Z digest=sha256:921b5468eb1596ee49d0b930a6ac478d02ebe5d56e9b27a46c89b644d0ccd4f3

Observation 58ea860f-86fc-42d2-a043-baf6309f49f7 · inbound

BioMol-MQA: A Multi-Modal Question Answering Dataset For LLM Reasoning Over Bio-Molecular Interactions cites this paper.

BioMol-MQA: A Multi-Modal Question Answering Dataset For LLM Reasoning Over Bio-Molecular Interactions Retrieval-Augmented Perception: High-Resolution Image Perception Meets Visual RAG

Reference 128

Resolution
unresolved
no resolver link, observed 2026-08-07T10:19:24.680869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:19:24.680869Z digest=sha256:45103ac67cc46b642dafb3ac6947f98af44187c23adebdca2de94059d63e343e

Observation 87d51fb1-ac2f-4cbf-955d-bdb607f566d6 · inbound

Toward Edge General Intelligence with Multiple-Large Language Model (Multi-LLM): Architecture, Trust, and Orchestration cites this paper.

Toward Edge General Intelligence with Multiple-Large Language Model (Multi-LLM): Architecture, Trust, and Orchestration Retrieval-Augmented Perception: High-Resolution Image Perception Meets Visual RAG

Reference 172

Resolution
unresolved
no resolver link, observed 2026-08-06T21:16:15.005988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:16:15.005988Z digest=sha256:8aa3f3eb7e9c0b27ae780f1d4eac28fb4dbb5d84738fdd3275f6ecfb7f776601

Observation f0a70e9b-e492-45d8-908f-5118e1e39024 · inbound

Look Before You Zoom: Adaptive Routing for the Resolution-Context Trade-off in Visual RAG cites this paper.

Look Before You Zoom: Adaptive Routing for the Resolution-Context Trade-off in Visual RAG Retrieval-Augmented Perception: High-Resolution Image Perception Meets Visual RAG

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-04T07:49:38.569409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T12:54:06.319322Z digest=sha256:bfcd9828f6f200bc014750cf5e1799ca5fa050897f25e3b715e89aa671f94ea1

Observation c32f5145-e9a5-4c9d-a49b-0d5e6679c1bc · inbound

BVS: Bayesian Visual Search with Multimodal Large Language Model for Fine-grained Perception cites this paper.

BVS: Bayesian Visual Search with Multimodal Large Language Model for Fine-grained Perception Retrieval-Augmented Perception: High-Resolution Image Perception Meets Visual RAG

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-12T04:17:40.198357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-12T04:17:40.198357Z digest=sha256:b49e007d1fa52adfd58a28b64b00cbbc60d4df4bd8bab7d47d3463d7b7c9b9dd