Pith. sign in

Paper Citation Record · LEDGER

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding

As of 7 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 2 inbound Pith citation observations for arXiv:2507.05673.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.05673 v1

Coverage vector

measured 38 of 38 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:27:24.411116Z

measured 40 of 40 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-19T14:24:48.938948Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-19T14:27:24.320785Z

Reference resolution

38 of 38 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved37
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b11696a9-f229-41bf-9796-3b06c51ca008 · outbound

This paper cites GPT-4 Technical Report.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T19:27:24.212735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:27:24.212735Z digest=sha256:2936b1fc38b09f9a6a8ce45d400e528be631dcc51a50428ca135bf4ddac0f0aa

Observation ea599908-8907-4dca-8505-f3669f4b8432 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T19:27:24.218897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:27:24.218897Z digest=sha256:8647eb67f74aabcec08a5c9a035530e5990709151e7dc43f6e0febb0e1de392f

Observation cfb5238d-3181-438e-bf97-f4684832be4a · outbound

This paper cites an unresolved cited work.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T19:27:24.224150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:27:24.224150Z digest=sha256:3f69220ac87ad7b58fd25c4b938d4d98b11b39b991d1e2164bf8accd77a9c108

Observation dda9be68-2083-447e-882c-f602fa44f548 · outbound

This paper cites an unresolved cited work.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:27:25.051971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T19:27:24.229639Z digest=sha256:afb62bf6b751f5c75d93a49f6acb0fd7f7092b4087b315ef05fd3ef24b82c809

Observation 735c04a1-b202-45d7-bef8-c182ce1024cc · outbound

This paper cites an unresolved cited work.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:27:25.034436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T19:27:24.235829Z digest=sha256:a0b0c8e18e1d231e28332d286c9069cc460e0741ff1b31b17165cd092fe7f246

Observation feb118dc-d6c3-4f77-8c56-8f42b185b65d · outbound

This paper cites an unresolved cited work.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T19:27:24.241420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:27:24.241420Z digest=sha256:34160a195dce7097017f9068633125e1cb47b6e2997223e7b9a6070ea2b0c2b7

Observation 48cedb75-8d1a-4cb2-a7da-4f1a98d7c6f9 · outbound

This paper cites an unresolved cited work.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:27:25.003225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T19:27:24.247498Z digest=sha256:f33c624c5b191c0f9b0cbda19ff70aebc956b106094d3ed3a53c2db24f156b0a

Observation 10b439b8-16b2-47e4-830d-d7e89eae4eb5 · outbound

This paper cites ASSISTGUI: Task-Oriented Desktop Graphical User Interface Automation.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding ASSISTGUI: Task-Oriented Desktop Graphical User Interface Automation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T19:27:24.252097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:27:24.252097Z digest=sha256:b7e46a7b84ec8ca684bc1d9f6e406d188752296c0aed82cbd6645dcefdc8aaa6

Observation 35d112e1-37a3-49af-8195-51200545f2d4 · outbound

This paper cites an unresolved cited work.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T19:27:24.257232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:27:24.257232Z digest=sha256:5166f59f80c319b9de704773eeb0db62f77dce491c7fa15aa22f9a891ccb1059

Observation 75cbf0a1-6094-4e83-887d-ecf61b755617 · outbound

This paper cites Fast R-CNN.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Fast R-CNN

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T19:27:24.262142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:27:24.262142Z digest=sha256:1ecae53d3bc280b08d28ffb2065a68cfcda6cc75d5bc7e4cc71c6b06e60dfec0

Observation 63c08db3-3a34-4587-9f8c-193de7897f46 · outbound

This paper cites an unresolved cited work.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:27:24.973013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T19:27:24.267199Z digest=sha256:2d3d1529ff199081cc2dd4eb25f18d513eb3f387155326d8c53999b3f10a7112

Observation 276b7c5f-f2da-4034-bed4-f0e9e97c2d14 · outbound

This paper cites an unresolved cited work.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:27:24.956555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T19:27:24.272772Z digest=sha256:32521d04e03688357c37be33df22555a84c1fcc79c0cecf63d458c4023fd8696

Observation 3f9e9c18-4adf-4f68-af3d-cfebe1dce810 · outbound

This paper cites an unresolved cited work.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:27:24.940575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T19:27:24.277902Z digest=sha256:75f92d9ac69a262a10d030b253c47b1d8b202878c76bb94f0fb991537ffc0761

Observation 6ff302a5-700c-4ddd-a218-2f5bde6350a2 · outbound

This paper cites WebVoyager: Building an End-to-End Web Agent with Large Multimodal Models.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding WebVoyager: Building an End-to-End Web Agent with Large Multimodal Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T19:27:24.283142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:27:24.283142Z digest=sha256:ffee98cbdada7dfd3b5c7ef165e8e5aa3a65af5ea78d2967ce5e49d9bd8a1ce4

Observation e2b2a860-7b18-45cb-926e-31bc46cada4a · outbound

This paper cites an unresolved cited work.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:27:24.923855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T19:27:24.288554Z digest=sha256:460f5e5fbb2f846524add56b054a275653876ee505d85678ceac2d3daa197b6a

Observation c0be1938-6bc9-4572-b06e-cf80fb2b0fc6 · outbound

This paper cites an unresolved cited work.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T19:27:24.293943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:27:24.293943Z digest=sha256:10340b723fdf184b3060fb9a8065a4350687f4ea69daf3ac440fa88c19f7e334

Observation a947842b-a14b-4ed0-8147-35ba94124eb5 · outbound

This paper cites OmniACT: A Dataset and Benchmark for Enabling Multimodal Generalist Autonomous Agents for Desktop and Web.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding OmniACT: A Dataset and Benchmark for Enabling Multimodal Generalist Autonomous Agents for Desktop and Web

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T19:27:24.299572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:27:24.299572Z digest=sha256:1f9ff60db39e622b75de58028e20f122dca099d12a6957d4076d2bf0a40cced3

Observation b219ece9-c2af-4596-adde-93cbe84df7f8 · outbound

This paper cites an unresolved cited work.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T19:27:24.304727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:27:24.304727Z digest=sha256:d45987f6c49dc3f6b02d2a389333532d86256b0b31448c1944e3b0be432a523a

Observation 17a058b4-165a-412c-9f3a-041b385c2bde · outbound

This paper cites an unresolved cited work.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:27:24.883379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T19:27:24.309870Z digest=sha256:b8a62babaf3c024f96978191c073a0c7b1dbc45af627d9d73aa3a43b8e94026c

Observation fbaa090f-817a-49f3-8b7b-b5c1508350b3 · outbound

This paper cites Decoupled Weight Decay Regularization.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Decoupled Weight Decay Regularization

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T19:27:24.315672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:27:24.315672Z digest=sha256:c581f33955dc2f933dd32c50d938b425f5bad8d18c96cd4e10bb8f93cf5daf97

Observation 0c71ca54-07d1-4ae9-9208-bf4ed2964856 · outbound

This paper cites SimCon Loss with Multiple Views for Text Supervised Semantic Segmentation.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding SimCon Loss with Multiple Views for Text Supervised Semantic Segmentation

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-08-06T19:27:24.532162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T19:27:24.320886Z digest=sha256:dad9af5e884a73376ab42e0d8ab3b1cd257b8deca941ce36efb626f14539ce9f

Observation 25748ec8-d0ed-4ec8-90d6-f76e2f8e7a44 · outbound

This paper cites an unresolved cited work.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:27:24.865131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T19:27:24.326696Z digest=sha256:e96c60bfd949972ec8cfb122169326b1bceb90034d5e131976f93f1f291475f4

Observation e2c4afd2-b22f-4f25-8d06-8bbefc3e5711 · outbound

This paper cites an unresolved cited work.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:27:24.846905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T19:27:24.332396Z digest=sha256:a122a4024ac48c38d71aa1b47d729ce1ef341a5306368c9a901a05d4483fd307

Observation 3a5e9143-37bf-4bbc-9f4e-7179113eb47a · outbound

This paper cites an unresolved cited work.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:27:24.828189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T19:27:24.337540Z digest=sha256:ed534179e75711f90ec05870200c73da6ca2e7781bc31df0c3d39b64e3e36da9

Observation 9d0aa4d6-f452-413e-a072-f397d1aca5bf · outbound

This paper cites an unresolved cited work.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:27:24.810918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T19:27:24.343146Z digest=sha256:78d6d21a5985b983f765ac764798f1ec853f98cc4a370d4813600d370832b6cb

Observation c20fc2dd-7c26-493e-86bf-47179cab3c62 · outbound

This paper cites ZoomEye: Enhancing Multimodal LLMs with Human-Like Zooming Capabilities through Tree-Based Image Exploration.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding ZoomEye: Enhancing Multimodal LLMs with Human-Like Zooming Capabilities through Tree-Based Image Exploration

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T19:27:24.348097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:27:24.348097Z digest=sha256:52d91323581bd58b090f3a668e306fd58accb3ebdf36bec4ebae2df560928dd5

Observation 9c14735c-94a5-4c62-8db0-a0738b40d152 · outbound

This paper cites an unresolved cited work.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T19:27:24.353074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:27:24.353074Z digest=sha256:86522ecec66e7119932aaf98fed2ded963f43cd69da2d35548eafc3745e3d8e4

Observation 4eb04b6c-2332-4a3f-b2cf-ab92b7ac183d · outbound

This paper cites an unresolved cited work.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:27:24.781317Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T19:27:24.358476Z digest=sha256:9e7214e953a55dba9970bfc14d8044f20d718a69f20baaf4041b5ebf42e5ded8

Observation dc832ccf-2fa6-44b3-b3ed-8db6b69a42d4 · outbound

This paper cites an unresolved cited work.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:27:24.761739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T19:27:24.364462Z digest=sha256:b92922d3154479583444d5073d5ff211ecb290f243bcea7ce7d4e00e446f6a58

Observation 83061e98-a732-44e5-8b0d-5b14b60cb24a · outbound

This paper cites You Only Look at Screens: Multimodal Chain-of-Action Agents.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding You Only Look at Screens: Multimodal Chain-of-Action Agents

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T19:27:24.370000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:27:24.370000Z digest=sha256:6044731e5fe0952c1779816d7d186527e8182a878c0fa63c558937fa24521722

Observation dcb16118-a6f4-40b0-97c6-cbdcddd5e6d0 · outbound

This paper cites GPT-4V(ision) is a Generalist Web Agent, if Grounded.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding GPT-4V(ision) is a Generalist Web Agent, if Grounded

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T19:27:24.375110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:27:24.375110Z digest=sha256:3bdd599121e7f174651786a4adb7911e545c82fc4d94f98a7a6876a9f56e1450

Observation 442eedfe-ccec-4258-b0b9-1b30644a7616 · outbound

This paper cites an unresolved cited work.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:27:24.743586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T19:27:24.379963Z digest=sha256:c2b8acad744e40e3daa8e1ce6f74e3174c612898910849efc5754214bc4026c7

Observation 49cf1bcf-c013-4949-9e04-c522406f976c · outbound

This paper cites AgentStudio: A Toolkit for Building General Virtual Agents.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding AgentStudio: A Toolkit for Building General Virtual Agents

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T19:27:24.385204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:27:24.385204Z digest=sha256:43241d1b116e9882ef2bff9aa6b100e0de8f4c7df3745506aeca41e302d5f225

Observation f1dd36eb-7289-4d88-a64b-1f410fae661a · outbound

This paper cites an unresolved cited work.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:27:24.725332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T19:27:24.390134Z digest=sha256:7569ebba80e03864c09fbcb799b9fb67543fb11e5e817d7804f7b2324caef338

Observation 984713c1-71fd-4899-8621-ea70afe44b03 · outbound

This paper cites an unresolved cited work.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:27:24.708939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T19:27:24.394991Z digest=sha256:8ef4ca5dbdf654791779fc3d29738a702f4264ef00005fd614636cc2a6dea4de

Observation d02d0340-1b12-4b81-aded-69b90d337191 · outbound

This paper cites an unresolved cited work.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:27:24.691713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T19:27:24.399973Z digest=sha256:d5420dfd6f192fa530b53264f9cff7be2f7e8c41452604a35fa02d83b2b15bf3

Observation 987e571c-f5c3-4a20-9dbd-d040f008a2cb · outbound

This paper cites online" 'onlinestring :=.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding online" 'onlinestring :=

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T19:27:24.405903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:27:24.405903Z digest=sha256:5337a923c4e9f4cec32feb5a0d47dee1c4c137ee5a85e161c9622ab9b054b470

Observation 09dce4fa-b1d5-4fb6-b51d-ba1a9cf5c44d · outbound

This paper cites write newline.

R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding write newline

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T19:27:24.411116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:27:24.411116Z digest=sha256:28642f5e90ecac83ccfb5e634ba0d071cf1fa739e9e32b7b3724c29184a139ba

Pith citing papers

Observation 53b33f13-a7e0-4af3-b7ba-6800a7abb538 · inbound

BAMI: Training-Free Bias Mitigation in GUI Grounding cites this paper.

BAMI: Training-Free Bias Mitigation in GUI Grounding R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:16:10.135986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:20:42.718277Z digest=sha256:0150ac490eb2a662bb6d4aba7dfde08a67dc26304d19ebfce10fdc89123f4704

Observation 22f79aca-1484-4f11-9baf-260f110ea114 · inbound

DRS-GUI: Dynamic Region Search for Training-Free GUI Grounding cites this paper.

DRS-GUI: Dynamic Region Search for Training-Free GUI Grounding R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-19T14:27:24.323399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-19T14:24:48.938948Z digest=sha256:0468ad2b686fc983d65473d0edcc2bdc9a83f90e80d0d17ffc34c78f0b475929