Pith. sign in

Paper Citation Record · LEDGER

EvoCUA-1.5: Online Reinforcement Learning for Multi-turn Computer-Use Agents

As of 6 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 1 inbound Pith citation observation for arXiv:2607.09773.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.09773 v1

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-14T16:00:52.709643Z

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-31T14:04:46.670034Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

17 of 17 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved17
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ed431483-db9b-4367-ac6f-d2adc7b33572 · outbound

This paper cites Qwen3-VL Technical Report.

EvoCUA-1.5: Online Reinforcement Learning for Multi-turn Computer-Use Agents Qwen3-VL Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-14T16:00:52.709643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T16:00:52.709643Z digest=sha256:88a4c9ae32ccc40856943d16da239bfb5d08c67753bbe372a136f5be47a368c8

Observation 30da3f35-9b80-4151-a2d1-1b1e0cb8d40a · outbound

This paper cites Xinyuan Wang, Bowen Wang, Dunjie Lu, Junlin Yang, Tianbao Xie, Junli Wang, Jiaqi Deng, Xiaole Guo, Yiheng Xu, Chen Henry Wu, et al.

EvoCUA-1.5: Online Reinforcement Learning for Multi-turn Computer-Use Agents Xinyuan Wang, Bowen Wang, Dunjie Lu, Junlin Yang, Tianbao Xie, Junli Wang, Jiaqi Deng, Xiaole Guo, Yiheng Xu, Chen Henry Wu, et al

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-14T16:00:52.709643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T16:00:52.709643Z digest=sha256:85433df0e3dca9a44a60a86bd9b1284b35379b1e33ea9db47814029387213b25

Observation d5ee4706-7dc8-41a3-9389-79ce4b75d5cb · outbound

This paper cites Zhihong Shao, Peiyi Wang, Qihao Zhu, Runxin Xu, Junxiao Song, Xiao Bi, Haowei Zhang, Mingchuan Zhang, YK Li, Yang Wu, et al.

EvoCUA-1.5: Online Reinforcement Learning for Multi-turn Computer-Use Agents Zhihong Shao, Peiyi Wang, Qihao Zhu, Runxin Xu, Junxiao Song, Xiao Bi, Haowei Zhang, Mingchuan Zhang, YK Li, Yang Wu, et al

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-14T16:00:52.709643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T16:00:52.709643Z digest=sha256:3cdbf77c990ad63fc0654194f4b6a361c33f75494eaf162df9b2f07eea905549

Observation ad25127f-d659-4ace-92ec-b86f0013b9be · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

EvoCUA-1.5: Online Reinforcement Learning for Multi-turn Computer-Use Agents DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-14T16:00:52.709643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T16:00:52.709643Z digest=sha256:b7c04418b13155a49e2c959f90f1f5f8e6ae258c506f92613659f5be9a7a1ecf

Observation ba6f4292-b5d8-476e-9aca-8823c6ea59a7 · outbound

This paper cites Proximal Policy Optimization Algorithms.

EvoCUA-1.5: Online Reinforcement Learning for Multi-turn Computer-Use Agents Proximal Policy Optimization Algorithms

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-14T16:00:52.709643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T16:00:52.709643Z digest=sha256:df26f11279cc54cf945a1a3cd4e8ef54912468e298e71050233e30a4963e1d5c

Observation 7c85903f-3c22-46ad-8e7f-98ae6e507e37 · outbound

This paper cites Haolong Yan, Jia Wang, Xin Huang, Yeqing Shen, Ziyang Meng, Zhimin Fan, Kaijun Tan, Jin Gao, Lieyu Shi, Mi Yang, et al.

EvoCUA-1.5: Online Reinforcement Learning for Multi-turn Computer-Use Agents Haolong Yan, Jia Wang, Xin Huang, Yeqing Shen, Ziyang Meng, Zhimin Fan, Kaijun Tan, Jin Gao, Lieyu Shi, Mi Yang, et al

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-14T16:00:52.709643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T16:00:52.709643Z digest=sha256:129530a22f9486e3c9efefd5e6ff7dee64ad63880133cb5fd866f2d5c952e925

Observation 26583e83-970d-4f7d-b981-62d9bd62333b · outbound

This paper cites Qwen Team.

EvoCUA-1.5: Online Reinforcement Learning for Multi-turn Computer-Use Agents Qwen Team

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-14T16:00:52.709643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T16:00:52.709643Z digest=sha256:8e229f51c34897e716e85cae40da422462176c72d857ed732e640ac8b0c718bd

Observation 23396511-308e-46b2-bc01-3b095fb38738 · outbound

This paper cites UI-TARS: Pioneering Automated GUI Interaction with Native Agents.

EvoCUA-1.5: Online Reinforcement Learning for Multi-turn Computer-Use Agents UI-TARS: Pioneering Automated GUI Interaction with Native Agents

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-14T16:00:52.709643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T16:00:52.709643Z digest=sha256:65af67c8740b9cd0cd8f46b25a977e04a117a288531fe843c2c3fe6d293dacf8

Observation 120c9b0d-907d-450a-88f3-61e14e84bc60 · outbound

This paper cites Mobile-Agent-v3: Fundamental Agents for GUI Automation.

EvoCUA-1.5: Online Reinforcement Learning for Multi-turn Computer-Use Agents Mobile-Agent-v3: Fundamental Agents for GUI Automation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-14T16:00:52.709643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T16:00:52.709643Z digest=sha256:6577883ac74ea3cec309ccd1d059ad6d5dc581cf04dcd4208f30ce62155653a3

Observation 502f1b07-d573-4f22-acc8-33313231f7e7 · outbound

This paper cites CUA-Gym: Scaling Verifiable Training Environments and Tasks for Computer-Use Agents.

EvoCUA-1.5: Online Reinforcement Learning for Multi-turn Computer-Use Agents CUA-Gym: Scaling Verifiable Training Environments and Tasks for Computer-Use Agents

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-14T16:00:52.709643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T16:00:52.709643Z digest=sha256:155d0709b5d85b21468d91d619976bcf1e859f17d12cdadd717a66f5a21c156d

Observation 9692b6c4-f382-46db-a7cc-936503549022 · outbound

This paper cites Windows Agent Arena: Evaluating Multi-Modal OS Agents at Scale.

EvoCUA-1.5: Online Reinforcement Learning for Multi-turn Computer-Use Agents Windows Agent Arena: Evaluating Multi-Modal OS Agents at Scale

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-14T16:00:52.709643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T16:00:52.709643Z digest=sha256:031d7548ee032f0e61e2a1e7b2c2fc05d90b0845c81ea7356c9ffd5423079847

Observation 9c60e594-4b8e-47be-976d-e07e74f2696a · outbound

This paper cites Digi-Q: Learning Q-Value Functions for Training Device-Control Agents.

EvoCUA-1.5: Online Reinforcement Learning for Multi-turn Computer-Use Agents Digi-Q: Learning Q-Value Functions for Training Device-Control Agents

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-14T16:00:52.709643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T16:00:52.709643Z digest=sha256:625cfa0f57f75c78eb8f27623965846dbba78efe9c960910e0a8c2f202ac61d8

Observation 4c40db38-e88f-4f15-944b-6de5e52afe14 · outbound

This paper cites Webarena: A realistic web environment for building autonomous agents.

EvoCUA-1.5: Online Reinforcement Learning for Multi-turn Computer-Use Agents Webarena: A realistic web environment for building autonomous agents

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-14T16:00:52.709643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T16:00:52.709643Z digest=sha256:30868ce21c416ae48292910b4d2236dd733486d26a483faa55d0e3a549c206c0

Observation 5680e1c6-393c-4d1b-a1c6-f7c00d71f8d6 · outbound

This paper cites WebFactory: Automated Compression of Foundational Language Intelligence into Grounded Web Agents.

EvoCUA-1.5: Online Reinforcement Learning for Multi-turn Computer-Use Agents WebFactory: Automated Compression of Foundational Language Intelligence into Grounded Web Agents

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-14T16:00:52.709643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T16:00:52.709643Z digest=sha256:628240a1c57195540e6b38d61e165adef0224fc507f8ae6296c65503f7c75b31

Observation 84f9d961-b2f2-45cc-baff-2608e971f185 · outbound

This paper cites Webagent-r1: Training web agents via end-to-end multi-turn reinforcement learning.

EvoCUA-1.5: Online Reinforcement Learning for Multi-turn Computer-Use Agents Webagent-r1: Training web agents via end-to-end multi-turn reinforcement learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-14T16:00:52.709643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T16:00:52.709643Z digest=sha256:478405119f112508b78f4b10a05949d2f53370012e9b84c63f2b6a449a7d42dd

Observation b006e2c5-2e0d-400e-a10f-3f7c569761ca · outbound

This paper cites Ui-s1: Advancing gui automation via semi-online reinforcement learning.arXiv preprint arXiv:2509.11543,.

EvoCUA-1.5: Online Reinforcement Learning for Multi-turn Computer-Use Agents Ui-s1: Advancing gui automation via semi-online reinforcement learning.arXiv preprint arXiv:2509.11543,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-14T16:00:52.709643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T16:00:52.709643Z digest=sha256:9d27b6e346827b4545da85f91276956b340b7b0db9cbc8b1403a8d34b5164c4f

Observation 8a8e46ba-7e5e-4102-972c-beca49e2ad4d · outbound

This paper cites Efficient multi-turn rl for gui agents via decoupled training and adaptive data curation.arXiv preprint arXiv:2509.23866,.

EvoCUA-1.5: Online Reinforcement Learning for Multi-turn Computer-Use Agents Efficient multi-turn rl for gui agents via decoupled training and adaptive data curation.arXiv preprint arXiv:2509.23866,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-14T16:00:52.709643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T16:00:52.709643Z digest=sha256:aba14f155e8ecd1e33848fab3f4c8b1b5bb8bbf2b406d0658cb5d861ff48e06d

Pith citing papers

Observation 27aad222-a72a-41cf-9a1b-e6a3de33970a · inbound

Qwen-UI-Agent Technical Report: Toward Next-Generation Real-World Centric Foundation GUI Agents cites this paper.

Qwen-UI-Agent Technical Report: Toward Next-Generation Real-World Centric Foundation GUI Agents EvoCUA-1.5: Online Reinforcement Learning for Multi-turn Computer-Use Agents

Reference 197

Resolution
unresolved
no resolver link, observed 2026-07-31T14:04:46.670034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T14:04:46.670034Z digest=sha256:66d97faf3d74bafcdd1f191f43437fa4b326281f9c77b336b62066640e78e04b