Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T19:27:24.411116Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 2 inbound Pith citation observations for arXiv:2507.05673.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T19:27:24.411116Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-19T14:24:48.938948Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-19T14:27:24.320785Z
38 of 38 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation b11696a9-f229-41bf-9796-3b06c51ca008 · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea599908-8907-4dca-8505-f3669f4b8432 · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cfb5238d-3181-438e-bf97-f4684832be4a · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dda9be68-2083-447e-882c-f602fa44f548 · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 735c04a1-b202-45d7-bef8-c182ce1024cc · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation feb118dc-d6c3-4f77-8c56-8f42b185b65d · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48cedb75-8d1a-4cb2-a7da-4f1a98d7c6f9 · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 10b439b8-16b2-47e4-830d-d7e89eae4eb5 · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding ASSISTGUI: Task-Oriented Desktop Graphical User Interface Automation
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35d112e1-37a3-49af-8195-51200545f2d4 · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75cbf0a1-6094-4e83-887d-ecf61b755617 · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Fast R-CNN
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63c08db3-3a34-4587-9f8c-193de7897f46 · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 276b7c5f-f2da-4034-bed4-f0e9e97c2d14 · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 3f9e9c18-4adf-4f68-af3d-cfebe1dce810 · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6ff302a5-700c-4ddd-a218-2f5bde6350a2 · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding WebVoyager: Building an End-to-End Web Agent with Large Multimodal Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2b2a860-7b18-45cb-926e-31bc46cada4a · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c0be1938-6bc9-4572-b06e-cf80fb2b0fc6 · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a947842b-a14b-4ed0-8147-35ba94124eb5 · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding OmniACT: A Dataset and Benchmark for Enabling Multimodal Generalist Autonomous Agents for Desktop and Web
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b219ece9-c2af-4596-adde-93cbe84df7f8 · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17a058b4-165a-412c-9f3a-041b385c2bde · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation fbaa090f-817a-49f3-8b7b-b5c1508350b3 · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Decoupled Weight Decay Regularization
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c71ca54-07d1-4ae9-9208-bf4ed2964856 · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding SimCon Loss with Multiple Views for Text Supervised Semantic Segmentation
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 25748ec8-d0ed-4ec8-90d6-f76e2f8e7a44 · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e2c4afd2-b22f-4f25-8d06-8bbefc3e5711 · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 3a5e9143-37bf-4bbc-9f4e-7179113eb47a · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9d0aa4d6-f452-413e-a072-f397d1aca5bf · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c20fc2dd-7c26-493e-86bf-47179cab3c62 · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding ZoomEye: Enhancing Multimodal LLMs with Human-Like Zooming Capabilities through Tree-Based Image Exploration
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c14735c-94a5-4c62-8db0-a0738b40d152 · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4eb04b6c-2332-4a3f-b2cf-ab92b7ac183d · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation dc832ccf-2fa6-44b3-b3ed-8db6b69a42d4 · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 83061e98-a732-44e5-8b0d-5b14b60cb24a · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding You Only Look at Screens: Multimodal Chain-of-Action Agents
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dcb16118-a6f4-40b0-97c6-cbdcddd5e6d0 · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding GPT-4V(ision) is a Generalist Web Agent, if Grounded
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 442eedfe-ccec-4258-b0b9-1b30644a7616 · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 49cf1bcf-c013-4949-9e04-c522406f976c · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding AgentStudio: A Toolkit for Building General Virtual Agents
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1dd36eb-7289-4d88-a64b-1f410fae661a · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 984713c1-71fd-4899-8621-ea70afe44b03 · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d02d0340-1b12-4b81-aded-69b90d337191 · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 987e571c-f5c3-4a20-9dbd-d040f008a2cb · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding online" 'onlinestring :=
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09dce4fa-b1d5-4fb6-b51d-ba1a9cf5c44d · outbound
R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding write newline
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53b33f13-a7e0-4af3-b7ba-6800a7abb538 · inbound
BAMI: Training-Free Bias Mitigation in GUI Grounding R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 22f79aca-1484-4f11-9baf-260f110ea114 · inbound
DRS-GUI: Dynamic Region Search for Training-Free GUI Grounding R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.