Pith. sign in

Paper Citation Record · LEDGER

UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

As of 5 August 2026, this Paper Citation Record lists 18 of 18 outbound references and 44 inbound Pith citation observations for arXiv:2503.21620.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.21620 v5

Coverage vector

measured 18 of 18 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-16T11:02:41.335059Z

measured 62 of 62 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 44 of 44 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T07:02:42.105225Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

18 of 18 outbound references displayed

  • verified exact11
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch7

External citation measurements

1
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 1a32c312-b75e-4471-b2cd-b264538e1092 · outbound

This paper cites L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning.

UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-18T00:19:22.325854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T11:02:41.335059Z digest=sha256:736d8fabe22d19f91d832d867e0e4947c64e3876a8e155dced668d5469f7545a

Observation d53601c4-0c83-4dbc-b025-6e78ee9808b5 · outbound

This paper cites AMEX: Android Multi-annotation Expo Dataset for Mobile GUI Agents.

UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning AMEX: Android Multi-annotation Expo Dataset for Mobile GUI Agents

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T11:02:41.359560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T11:02:41.335059Z digest=sha256:d4ffefc1adcb23d88216aaddefc8648ae355477c13d8e9b36a674acf7b240f7a

Observation 4b1ebc86-ba6f-4939-980c-8c88ca5f7ee0 · outbound

This paper cites VisRL: Intention-Driven Visual Perception via Reinforced Reasoning.

UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning VisRL: Intention-Driven Visual Perception via Reinforced Reasoning

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T11:02:41.363985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T11:02:41.335059Z digest=sha256:533ea7cb466f00a675f6fdcbd8b90c555661e80e4adca178c4c61eea85636c46

Observation 59f969bf-94ad-4151-b5eb-f5be48228c2b · outbound

This paper cites Navigating the Digital World as Humans Do: Universal Visual Grounding for GUI Agents.

UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning Navigating the Digital World as Humans Do: Universal Visual Grounding for GUI Agents

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-16T21:09:01.446022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T11:02:41.335059Z digest=sha256:51b6399dad790e59d7cfa4e073ca65984a18c3bf79e46ae70ccd095e93996caa

Observation 738d9a85-dec2-4ef6-b95b-21871b661bb4 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-05-16T11:02:41.372022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T11:02:41.335059Z digest=sha256:8c044ff195d701d776fc3f00430b6e9de96667ef399b2f92516b21d745da1023

Observation a073bb6f-e00f-4805-aca9-43bc52576cd7 · outbound

This paper cites Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models.

UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-16T11:02:41.376523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T11:02:41.335059Z digest=sha256:4472a4bc74e0d864055747fcbce9454218baa8365e4085710943835baaee2371

Observation 9b356932-dbdb-4723-9f25-0d936f79de3c · outbound

This paper cites Think or not think: A study of ex- plicit thinking in rule-based visual reinforcement fine-tuning.

UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning Think or not think: A study of ex- plicit thinking in rule-based visual reinforcement fine-tuning

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T11:02:41.380780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T11:02:41.335059Z digest=sha256:bb848bfcfbc3b6b32087c2f717ebd123e070605dfa718fc73608ac3bf61a6f1e

Observation e80f9667-1b61-4f9f-a389-1ecc50b2b97f · outbound

This paper cites MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning.

UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-05-16T11:02:41.384566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T11:02:41.335059Z digest=sha256:e77e7d3ae231585abf992becf6544fa8b47b94f64f1f0d1ecdec2e09c505f55e

Observation ae09ade4-3711-4f58-b527-375c7db3f535 · outbound

This paper cites s1: Simple test-time scaling.

UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning s1: Simple test-time scaling

Reference 9

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T11:02:41.388399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T11:02:41.335059Z digest=sha256:14a39da4a9c6279d6f243e031523dbf10864339e595265be5ffc7d4bc3566e6f

Observation 0d8dac02-a753-44a4-9178-be2768b64533 · outbound

This paper cites UI-TARS: Pioneering Automated GUI Interaction with Native Agents.

UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning UI-TARS: Pioneering Automated GUI Interaction with Native Agents

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-05-16T11:02:41.392244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T11:02:41.335059Z digest=sha256:a2a96d60c701cbd17dc65922539cb1cf140cc43d85f93641666f24a6fdf968a7

Observation 7cc8e6ef-47fb-4c39-a084-7a5aadda903e · outbound

This paper cites Xiaoye Qu, Yafu Li, Zhaochen Su, Weigao Sun, Jianhao Yan, Dongrui Liu, Ganqu Cui, Daizong Liu, Shuxian Liang, Junxian He, and 1 others.

UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning Xiaoye Qu, Yafu Li, Zhaochen Su, Weigao Sun, Jianhao Yan, Dongrui Liu, Ganqu Cui, Daizong Liu, Shuxian Liang, Junxian He, and 1 others

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T11:02:41.396961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T11:02:41.335059Z digest=sha256:4782a20b88a4d557e54eb612002040d5e297249d2c7eb256b2ea67fe315c2123

Observation 14c204a1-4c1c-4817-8dd8-abd8c54b4fea · outbound

This paper cites Proximal Policy Optimization Algorithms.

UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning Proximal Policy Optimization Algorithms

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-16T11:02:41.401423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T11:02:41.335059Z digest=sha256:3593a4307bb468f09eeaf3d3ff78af6a7e3e3ea0f5f8258c89454fd349476a7b

Observation 30fcb531-eb22-45f2-844c-32d0ba019c9a · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-05-16T11:02:41.406052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T11:02:41.335059Z digest=sha256:8fc93bcdf4e2455d1f01cf06a44cb2880841bf0c8e4ed4f8734e5b1d20517f2d

Observation 175fc7dc-9e7b-4f2f-8e59-0433fef1f682 · outbound

This paper cites Dast: Difficulty-adaptive slow-thinking for large reasoning models.

UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning Dast: Difficulty-adaptive slow-thinking for large reasoning models

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:02:41.410687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T11:02:41.335059Z digest=sha256:fc01f5231bfdf697f83b05bd712ae4a45affa873be3457775541331f38d8d2a3

Observation 48dd0879-dda4-47d4-8612-8c40df87691e · outbound

This paper cites Smoothing Grounding and Reasoning for MLLM-Powered GUI Agents with Query-Oriented Pivot Tasks.

UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning Smoothing Grounding and Reasoning for MLLM-Powered GUI Agents with Query-Oriented Pivot Tasks

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T11:02:41.414608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T11:02:41.335059Z digest=sha256:39818b0ab48066268572bd6bea3af09926ae9161f31250d63667c1471bc339fa

Observation 642040bf-3f7d-48fc-aa3d-2925f678b824 · outbound

This paper cites Aguvis: Unified Pure Vision Agents for Autonomous GUI Interaction.

UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning Aguvis: Unified Pure Vision Agents for Autonomous GUI Interaction

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-18T04:09:42.002663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T11:02:41.335059Z digest=sha256:4598df4d82578087f24ba02b4c189c282066137904b9258379c3aef5f81f6ade

Observation f402604d-a719-47e6-9e8c-b9ab0ae39c0d · outbound

This paper cites AppAgent: Multimodal Agents as Smartphone Users.

UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning AppAgent: Multimodal Agents as Smartphone Users

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T10:16:44.196372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T11:02:41.335059Z digest=sha256:84b743199baa0f2f7d98573beddb1ae836c2fee2ec478cc6618b35b88711f985

Observation 465e5b14-2ac5-4d34-8ddc-9666c3be29e9 · outbound

This paper cites R1-Zero's "Aha Moment" in Visual Reasoning on a 2B Non-SFT Model.

UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning R1-Zero's "Aha Moment" in Visual Reasoning on a 2B Non-SFT Model

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-19T07:13:47.672124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T11:02:41.335059Z digest=sha256:642a02ff80ef164546f1866716a45cdf2ef362ef188cf060f9d91a71f445028d

Pith citing papers

Observation c122ab89-0bff-42b7-905b-c83b6917c094 · inbound

Video-R1: Reinforcing Video Reasoning in MLLMs cites this paper.

Video-R1: Reinforcing Video Reasoning in MLLMs UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:02:41.429566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-12T09:43:00.208065Z digest=sha256:298486ef623a3d30fe67219dd9fb504979e86943d30b72af374453076310e456

Observation 7f483ced-fdb1-42e2-a2d3-ece5c17903b3 · inbound

GUI-R1 : A Generalist R1-Style Vision-Language Action Model For GUI Agents cites this paper.

GUI-R1 : A Generalist R1-Style Vision-Language Action Model For GUI Agents UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:02:41.429566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-15T02:10:57.976448Z digest=sha256:fb448eb7493ab47c07ae18d50aee6a4ad8cb7deb5d980950115f89517cdd6a97

Observation 6b1f4060-ec7d-4940-9144-795639026de5 · inbound

InfiGUI-R1: Advancing Multimodal GUI Agents from Reactive Actors to Deliberative Reasoners cites this paper.

InfiGUI-R1: Advancing Multimodal GUI Agents from Reactive Actors to Deliberative Reasoners UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:54:44.089259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T13:54:44.011048Z digest=sha256:1b3824c8ab4ac6079c580f57355f6e31457d500d21998e7e35a40732c639a1c2

Observation 99380165-b6d4-40c7-ae33-0818969f5ce7 · inbound

Grounded Reinforcement Learning for Visual Reasoning cites this paper.

Grounded Reinforcement Learning for Visual Reasoning UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-05-22T01:05:52.038819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T01:05:18.801388Z digest=sha256:c25a58b931bfb7903cee0385b985fd8e731677d5694f6365417821bb2dea6a90

Observation f9da5ac2-c610-4785-b339-7f93a929d3ff · inbound

LPO: Towards Accurate GUI Agent Interaction via Location Preference Optimization cites this paper.

LPO: Towards Accurate GUI Agent Interaction via Location Preference Optimization UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-05-19T09:47:13.930425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T09:45:55.674155Z digest=sha256:6909cbc43e34290aabbab73b2861520c925bc092fdafb66b9ffba0d20fcd09da

Observation 2e241387-0bb2-4aad-96c0-fd5697de2790 · inbound

GTA1: GUI Test-time Scaling Agent cites this paper.

GTA1: GUI Test-time Scaling Agent UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-05-17T13:55:00.027972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-17T13:54:59.938216Z digest=sha256:8d102d1cbc518f58d9e19e33bd910693aa7a81f2df3ca2c3b54197edd1cb8246

Observation e709465c-ebf3-413a-9194-c98ac1178d4b · inbound

MAS-Bench: A Unified Benchmark for Shortcut-Augmented Hybrid Mobile GUI Agents cites this paper.

MAS-Bench: A Unified Benchmark for Shortcut-Augmented Hybrid Mobile GUI Agents UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-05-18T18:42:48.094392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-18T18:42:29.744124Z digest=sha256:257a3c1cb12ab41aa95472014895d72783e3dbe69b07b8e4a55637ed99764538

Observation 87c649d1-8fda-495c-90ac-9843f98a9100 · inbound

VeriOS: Query-Driven Proactive Human-Agent-GUI Interaction for Trustworthy OS Agents cites this paper.

VeriOS: Query-Driven Proactive Human-Agent-GUI Interaction for Trustworthy OS Agents UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-05-18T18:06:42.814671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T18:06:12.349285Z digest=sha256:35d37143a72af8d1b61e4e5cc0d3be69ef6df28924b7bf36193e2fb9a455efe5

Observation e3a6615d-e424-4527-b5db-24ed3520c2c5 · inbound

Enhancing Trustworthy GUI Grounding via Self-Critiqued Reinforcement Learning cites this paper.

Enhancing Trustworthy GUI Grounding via Self-Critiqued Reinforcement Learning UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 2015

Resolution
unresolved
no resolver link, observed 2026-08-04T07:02:42.105225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T07:02:42.105225Z digest=sha256:7c361aa0981d8c037532e2f517f37ae5f53b693a52c26ad0763591cf7d9f7188

Observation 3903365b-e94f-48ea-8b1b-8fdcf5098803 · inbound

Grounding Computer Use Agents on Human Demonstrations cites this paper.

Grounding Computer Use Agents on Human Demonstrations UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-03T23:06:05.031828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T23:06:05.031828Z digest=sha256:57e77f82560b9ddae714b5e3fd1ab6e13df0aae9b889468b14ed4fa19f3c6633

Observation 21967e0c-e5f6-42a9-a2a2-d93f2d683900 · inbound

Pushing Forward Pareto Frontiers of Proactive Agents with Behavioral Agentic Optimization cites this paper.

Pushing Forward Pareto Frontiers of Proactive Agents with Behavioral Agentic Optimization UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-03T00:13:27.015676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T00:13:27.015676Z digest=sha256:98a0701144124ecd1e8f9d162425f8d9769483130a60192ad4a6a74832348cf3

Observation 1f0fa10b-14d6-4802-a1ec-4caba6053144 · inbound

ProactiveMobile: A Comprehensive Benchmark for Boosting Proactive Intelligence on Mobile Devices cites this paper.

ProactiveMobile: A Comprehensive Benchmark for Boosting Proactive Intelligence on Mobile Devices UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:02:41.429566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-15T19:43:30.394715Z digest=sha256:921074b15208411e4b6d208dbe91f29dc8f884ea8ea88349e3bd6db421831a49

Observation 2ec95887-07b3-4f19-a857-2d25c2f5ad04 · inbound

GUI-Libra: Training Native GUI Agents to Reason and Act with Action-aware Supervision and Partially Verifiable RL cites this paper.

GUI-Libra: Training Native GUI Agents to Reason and Act with Action-aware Supervision and Partially Verifiable RL UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-02T20:51:42.979477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T20:51:42.979477Z digest=sha256:abb98055cf7f322953d8b048e8822e9c83b13e582a493700f6f386fa8bcfafe9

Observation 686823a1-c2c5-43f8-b652-c119a8004622 · inbound

From Pixels to Digital Agents: An Empirical Study on the Taxonomy and Technological Trends of Reinforcement Learning Environments cites this paper.

From Pixels to Digital Agents: An Empirical Study on the Taxonomy and Technological Trends of Reinforcement Learning Environments UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 212

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:02:41.429566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-15T01:20:03.181903Z digest=sha256:40c043fd1642183568d12516538286193b6c04639640dc364769022d40f163c8

Observation 7d9eb86a-a20e-4b22-a261-186518a55ea0 · inbound

UI-Zoomer: Uncertainty-Driven Adaptive Zoom-In for GUI Grounding cites this paper.

UI-Zoomer: Uncertainty-Driven Adaptive Zoom-In for GUI Grounding UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:02:41.429566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T14:06:55.472857Z digest=sha256:8f56970699c05ddad7be0ba2e6267238fe65b3f7669a58e3f228cc43c82811ff

Observation 591a083e-513b-4aff-8289-d9832d2795a4 · inbound

HalluClear: Diagnosing, Evaluating and Mitigating Hallucinations in GUI Agents cites this paper.

HalluClear: Diagnosing, Evaluating and Mitigating Hallucinations in GUI Agents UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:02:41.429566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T06:27:19.717895Z digest=sha256:a8037f7299d3464072f193e31553f02d55a40fb59612d24396addfc49e0ee868

Observation a1c5c2ad-218b-4c45-aedb-08ad9a1b0dcd · inbound

Measure Twice, Click Once: Co-evolving Proposer and Visual Critic via Reinforcement Learning for GUI Grounding cites this paper.

Measure Twice, Click Once: Co-evolving Proposer and Visual Critic via Reinforcement Learning for GUI Grounding UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 86

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T11:02:41.429566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-09T23:05:05.251150Z digest=sha256:e87275b9f0d61f61abde3da6ed39fea6ccb751a7e3d67ee4249ae156c254eaa2

Observation 8f0fc96a-22db-4d80-bb7c-6cf6e01a166b · inbound

OS-SPEAR: A Toolkit for the Safety, Performance,Efficiency, and Robustness Analysis of OS Agents cites this paper.

OS-SPEAR: A Toolkit for the Safety, Performance,Efficiency, and Robustness Analysis of OS Agents UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:02:41.429566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T03:51:54.310805Z digest=sha256:d4e57300ed86b5178cb73f654f944350630ea1722937f1130cdd5e24a064fd7e

Observation bb7783f5-fd71-4cda-842a-bb724068d3d9 · inbound

AutoGUI-v2: A Comprehensive Multi-Modal GUI Functionality Understanding Benchmark cites this paper.

AutoGUI-v2: A Comprehensive Multi-Modal GUI Functionality Understanding Benchmark UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:02:41.429566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T04:32:06.963243Z digest=sha256:b87bf2b43cf8404938cf7e8cb90072a510a5bb164c1aa4a6e55498c0312408d9

Observation effac9bc-4c14-49f3-8f77-4dcd44162e1a · inbound

Rethinking Agentic Reinforcement Learning In Large Language Models cites this paper.

Rethinking Agentic Reinforcement Learning In Large Language Models UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:02:41.429566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T06:30:09.945371Z digest=sha256:a44e8d8f11e59ab99c9116aa84fe730c460368c5e32d1fa6d534cb81d430f4f5

Observation 616a6e86-a5de-4642-a429-be711cf22ce0 · inbound

Rethinking Agentic Reinforcement Learning In Large Language Models cites this paper.

Rethinking Agentic Reinforcement Learning In Large Language Models UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:02:41.429566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T03:12:19.414358Z digest=sha256:2f3e602733286a2a723b6bbed0230f25eb45aea1d4d34d2f0a734a8be75f6f32

Observation bec50ab5-1854-464b-8a7c-2431ddcd3426 · inbound

Rethinking Agentic Reinforcement Learning In Large Language Models cites this paper.

Rethinking Agentic Reinforcement Learning In Large Language Models UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 55

Resolution
verified exact
local_arxiv, observed 2026-05-19T17:02:40.826157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T16:58:41.558250Z digest=sha256:40fbea941a00a7a486229ca157c9556660bf814f7fe81eca2df091cb7c40c823

Observation 660ae962-70cb-4a97-a0b2-2ca0082eaadd · inbound

On Training Large Language Models for Long-Horizon Tasks: An Empirical Study of Horizon Length cites this paper.

On Training Large Language Models for Long-Horizon Tasks: An Empirical Study of Horizon Length UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 65

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T11:02:41.429566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:13:25.735085Z digest=sha256:fafacf5a4aef893d9ac97e8ee66950d94b2748096ff3d5297997e6bc8ec79229

Observation 1b9ee0ab-0c41-439e-a4fd-70be11898d41 · inbound

ROSE: Rollout On Serving GPUs via Cooperative Elasticity for Agentic RL cites this paper.

ROSE: Rollout On Serving GPUs via Cooperative Elasticity for Agentic RL UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:02:41.429566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T05:14:14.168753Z digest=sha256:b4d8d9924658ae28e54f66f70a2489d57a5285e252a8177ee2330c369e6dc2c7

Observation 2a90d8cb-c65c-44d8-840e-40caa402f528 · inbound

ROSE: Rollout On Serving GPUs via Cooperative Elasticity for Agentic RL cites this paper.

ROSE: Rollout On Serving GPUs via Cooperative Elasticity for Agentic RL UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-05-21T08:39:53.245323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T08:39:31.911497Z digest=sha256:b2f294949050d6722bbac973ef583a3fed933ee292251482f317efbb66460143

Observation a7e78695-4457-4568-ba9a-05e757b73124 · inbound

BAMI: Training-Free Bias Mitigation in GUI Grounding cites this paper.

BAMI: Training-Free Bias Mitigation in GUI Grounding UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:02:41.429566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T12:20:42.718277Z digest=sha256:56783085af3b35b52ef184836e752f3cec8948043a225d8605df44eee85e0461

Observation 9857d44d-8310-44fa-926f-ac7b3030055e · inbound

Securing Computer-Use Agents: A Unified Architecture-Lifecycle Framework for Deployment-Grounded Reliability cites this paper.

Securing Computer-Use Agents: A Unified Architecture-Lifecycle Framework for Deployment-Grounded Reliability UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 112

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:02:41.429566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-11T01:15:19.239355Z digest=sha256:8ccf6b1b5351de39ecf69acfe048c19587c0911f0b42e2f754f5333debee4ac9

Observation cce6f51c-cb89-44c7-a540-b2eac298ac42 · inbound

LiteGUI: Distilling Compact GUI Agents with Reinforcement Learning cites this paper.

LiteGUI: Distilling Compact GUI Agents with Reinforcement Learning UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T11:02:41.429566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-11T02:18:57.917353Z digest=sha256:9eaa3e5a1fdd10766a822a8e23984f4e6c0a383a04ae0773156d20b83549f79c

Observation 0c915db7-0642-46c4-86a1-7e05d7f0dcda · inbound

ToolCUA: Towards Optimal GUI-Tool Path Orchestration for Computer Use Agents cites this paper.

ToolCUA: Towards Optimal GUI-Tool Path Orchestration for Computer Use Agents UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:02:41.429566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T03:51:54.401200Z digest=sha256:ebf81cf22e2b161116ffb3c2b7d4ddeff886c2d751932bb01ed16fc582eedeab

Observation fbf1e0b4-2e36-41ef-8a70-dc132db55ef2 · inbound

Beyond Binary: Reframing GUI Critique as Continuous Semantic Alignment cites this paper.

Beyond Binary: Reframing GUI Critique as Continuous Semantic Alignment UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 96

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:02:41.429566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-15T01:58:19.295247Z digest=sha256:3bc78e5a0960c47aed8180734e7ec3b6f3057cb78a5eaf3041e1f168ed8d3de9

Observation d5eff564-5ead-4032-8b15-3f4baa049594 · inbound

Beyond Binary: Reframing GUI Critique as Continuous Semantic Alignment cites this paper.

Beyond Binary: Reframing GUI Critique as Continuous Semantic Alignment UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 96

Resolution
verified exact
local_arxiv, observed 2026-05-19T16:47:40.374623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-19T16:45:16.963802Z digest=sha256:81a1c5cdcf10e16b189f9e43f7b7cebf8278c3b72457e6137c3bb179fb0b37f5

Observation 8798ea97-3dc5-435b-a8f9-d19559b28045 · inbound

PAGER: Bridging the Semantic-Execution Gap in Point-Precise Geometric GUI Control cites this paper.

PAGER: Bridging the Semantic-Execution Gap in Point-Precise Geometric GUI Control UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-05-20T18:48:53.397890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-20T18:46:20.417039Z digest=sha256:63ddd51b54c52604a0de5351093f8b8a2a2fc73989b637798693b70d8a584b03

Observation 7517af85-c9a7-48b7-8bd4-a1e5f857610e · inbound

SE-GA: Memory-Augmented Self-Evolution for GUI Agents cites this paper.

SE-GA: Memory-Augmented Self-Evolution for GUI Agents UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-05-19T20:42:46.418731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-19T20:38:37.264654Z digest=sha256:5e3baa2ffcb3202845eafb72560246ea8a8d699f6c622d42d4d24002b069f8ef

Observation 3eb1075e-7679-4380-8079-9f7f170807b3 · inbound

CaptchaMind: Training CAPTCHA Solvers via Reinforcement Learning with Explicit Reasoning Supervision cites this paper.

CaptchaMind: Training CAPTCHA Solvers via Reinforcement Learning with Explicit Reasoning Supervision UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 11

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T06:18:05.557028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-20T06:14:43.288597Z digest=sha256:5761d37c031ab5401db48d5f02b94590ae3e6ef40d65d325c0b28c29618ff370

Observation 9b1700a5-4a0d-4bf7-80d0-a0f55d351c39 · inbound

Mobile-Aptus: Confidence-Driven Proactive and Robust Interaction in MLLM-based Mobile-Using Agents cites this paper.

Mobile-Aptus: Confidence-Driven Proactive and Robust Interaction in MLLM-based Mobile-Using Agents UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-06-29T12:53:26.744924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-29T12:47:01.474220Z digest=sha256:254cdd75b00129f2e6a262f4cc5dac342d657947d66027431fe5d8c52c3b7fd9

Observation cbb2a9b2-312e-40a6-87b6-d45471c7e7d8 · inbound

GUI-C$^2$: Coarse-to-Fine GUI Grounding via Difficulty-Aware Reinforcement Learning cites this paper.

GUI-C$^2$: Coarse-to-Fine GUI Grounding via Difficulty-Aware Reinforcement Learning UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-06-28T22:52:44.704188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-28T22:52:04.755523Z digest=sha256:f1d77b0578a5ab4bf0d1fa7f21f3b1de2c4cc1a6defae08b2e28a24ea9ee0a82

Observation 7f0b0893-5f6b-423f-865e-9668c098a7e5 · inbound

Characterize Then Distill: Mechanistic Reasoning in Large Output Spaces cites this paper.

Characterize Then Distill: Mechanistic Reasoning in Large Output Spaces UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 113

Resolution
metadata mismatch
local_arxiv, observed 2026-06-27T22:31:21.483121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-27T22:22:52.690010Z digest=sha256:7f302411b2f35841895fe6567a65ecee1ee83ff1dba17ca8df122210f41e7c0f

Observation 9a01a07e-6ca5-49ae-8d2d-0ef0b0fa5b0f · inbound

GUI-AC: Enhancing Continual Learning in GUI Agents cites this paper.

GUI-AC: Enhancing Continual Learning in GUI Agents UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-07-03T04:27:36.607989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-27T13:56:09.049753Z digest=sha256:afcb1df6ad33166392942958d5fac07af5874f864051ebc41d858713aafbe7bb

Observation ba8aec33-f626-427b-9998-65c95332b402 · inbound

GUI-AC: Enhancing Continual Learning in GUI Agents cites this paper.

GUI-AC: Enhancing Continual Learning in GUI Agents UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-07-12T14:27:05.589465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T14:27:05.589465Z digest=sha256:44558eb28db45ba32b96e3d2ba20156db67699cec6863f09ab2478958502812d

Observation 4a70f267-80d1-463a-94ab-31ec97eb2408 · inbound

MobileForge: Annotation-Free Adaptation for Mobile GUI Agents with Hierarchical Feedback-Guided Policy Optimization cites this paper.

MobileForge: Annotation-Free Adaptation for Mobile GUI Agents with Hierarchical Feedback-Guided Policy Optimization UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-07-04T05:29:35.709543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-26T15:58:59.248141Z digest=sha256:e90ee31cb7151303fcc2b717883facad391415b124dfee1d8383efa93ad9bceb

Observation 150a535e-b9b9-452d-9b01-e5d7563b4fe2 · inbound

Empowering GUI Agents via Autonomous Experience Exploration and Hindsight Experience Utilization for Task Planning cites this paper.

Empowering GUI Agents via Autonomous Experience Exploration and Hindsight Experience Utilization for Task Planning UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 8

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T14:29:53.358674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-26T03:51:51.827622Z digest=sha256:7da746bb4c69c172ea76c40e3928d43e326a46d99be90b2dafa3d6889c93f8a9

Observation c35d5605-e388-4ae3-a272-a53c2b7861f2 · inbound

BashCoder-R1: Towards Robust and Explainable Bash Code Generation with Robustness-Aware Group Relative Policy Optimization cites this paper.

BashCoder-R1: Towards Robust and Explainable Bash Code Generation with Robustness-Aware Group Relative Policy Optimization UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-07-01T17:05:50.837929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-29T04:16:04.477464Z digest=sha256:36ac1a50f74335fd22378bc5657007f5056d928282154e25475e9b6f232edc56

Observation 6225b758-634e-44c6-af39-a1e7d021d8c8 · inbound

GUICrafter: Weakly-Supervised GUI Agent Leveraging Massive Unannotated Screenshots cites this paper.

GUICrafter: Weakly-Supervised GUI Agent Leveraging Massive Unannotated Screenshots UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-06-30T06:44:19.217044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T06:39:12.591090Z digest=sha256:70ed6c97fc960f05310f4f3a9100a9ba97e768ed78048c14cf8e59d74ba45d44

Observation 07201f30-c3f5-4cad-95b5-0168b9ddfc88 · inbound

Qwen-UI-Agent Technical Report: Toward Next-Generation Real-World Centric Foundation GUI Agents cites this paper.

Qwen-UI-Agent Technical Report: Toward Next-Generation Real-World Centric Foundation GUI Agents UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 118

Resolution
unresolved
no resolver link, observed 2026-07-31T14:04:45.815734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T14:04:45.815734Z digest=sha256:392e39218e34fd87283ef015e9ff728d617ad15ba1a33c9e162cc55eaf84a0b0