Pith. sign in

Paper Citation Record · LEDGER

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL

As of 8 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 0 inbound Pith citation observations for arXiv:2607.11185.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.11185 v1

Coverage vector

measured 31 of 31 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-14T06:25:23.264527Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

31 of 31 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved31
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0b86aae1-f8bf-4270-860a-14b421dadb7a · outbound

This paper cites Agent S2: A Compositional Generalist-Specialist Framework for Computer Use Agents.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Agent S2: A Compositional Generalist-Specialist Framework for Computer Use Agents

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:1eaa14c08f299cf2669d12e3abcd635654d7758f0a66b6249d54c4ba666b600d

Observation 4db3b6de-c2c8-4d0c-93ae-b18c8da5366f · outbound

This paper cites Qwen3-VL Technical Report.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Qwen3-VL Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:cfd392ee26f89cef2f5b3307b8cb6b189e48d469352c2cce863bdcafabdc8513

Observation 64a4e43f-3499-4e10-bb6b-12f54f590a3c · outbound

This paper cites Windows Agent Arena: Evaluating Multi-Modal OS Agents at Scale.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Windows Agent Arena: Evaluating Multi-Modal OS Agents at Scale

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:c064017600445e253e5d35e5074b1acc06982a35015bab449c68bce559ddab1e

Observation 4f1dacf1-6acf-45ba-96f4-b2408baecb7b · outbound

This paper cites Gui-genesis: Automated synthesis of efficient environments with verifiable rewards for gui agent post-training.arXiv preprint arXiv:2602.14093,.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Gui-genesis: Automated synthesis of efficient environments with verifiable rewards for gui agent post-training.arXiv preprint arXiv:2602.14093,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:6e1241889e8ad7ff86d2e8117e99a3f35f4f6d72c6854bf822701d75648e202b

Observation a7fe5b11-3441-4c07-b2bc-5d76ee54b478 · outbound

This paper cites Agentic Reinforced Policy Optimization.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Agentic Reinforced Policy Optimization

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:cc20502a7103420e584edfe7aea6fff5836ed880be810fd1f031bc5b9fa5a965

Observation 33f473d4-e0e8-4825-8b6e-65aeb4b337d0 · outbound

This paper cites Navigating the Digital World as Humans Do: Universal Visual Grounding for GUI Agents.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Navigating the Digital World as Humans Do: Universal Visual Grounding for GUI Agents

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:f36794ad9c5924c8353423939431846d0a6a19e2fb5ed919c0db7967c3ec7f78

Observation 1a3f5382-04c4-449d-98eb-82326fcd579e · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:7e926a300fd9d337e70e6678813fcf6b5ef1c359231003edb98e1fb4d62df2f5

Observation 74475d5f-119d-4f3f-92b8-a00338a67808 · outbound

This paper cites GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:8d656e17e3be090ef6eb8a3d1cf26e11be549e34c7498aa0c5f981b75bff545e

Observation 239cc74f-50ec-42de-85ec-e2362093f294 · outbound

This paper cites CoAct: A Global-Local Hierarchy for Autonomous Agent Collaboration.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL CoAct: A Global-Local Hierarchy for Autonomous Agent Collaboration

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:b5e0e2b5ca4403ed435ea91bb7e18ea6f44c447fd85dfff1451f9d649da37d3f

Observation 36f00fbb-47fe-489b-8ba3-734bb9eac714 · outbound

This paper cites Androidgen: Building an android language agent under data scarcity.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Androidgen: Building an android language agent under data scarcity

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:7e3bc542b4df7d74a394568c92c04f6e2a212879ac5311d5ad2515e767ff4e27

Observation 01ac9fef-c562-4224-a789-9452760bec0e · outbound

This paper cites VisualAgentBench: Towards Large Multimodal Models as Visual Foundation Agents.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL VisualAgentBench: Towards Large Multimodal Models as Visual Foundation Agents

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:0a7495e75abbd05928271c70d4a22f15c634996e96faeee014c37c61b1f0104d

Observation 3ea66aef-dc4e-4898-8121-a7bf36a89eb6 · outbound

This paper cites WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:dd422eeeb6ca6b71f75da8262eca73f968d3f70d837417451e9be4a0f9cd4428

Observation fb084d7f-f1d3-4193-812e-0e02e328cb22 · outbound

This paper cites UI-TARS: Pioneering Automated GUI Interaction with Native Agents.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL UI-TARS: Pioneering Automated GUI Interaction with Native Agents

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:5bef3fcd3e9336dd953356a7d0e3e5fcded0b75b9202cbf54c03516a7fecedbd

Observation 6fb7a1f6-f563-46ed-a0cc-59fa0b91d7f3 · outbound

This paper cites Seed1.8 Model Card: Towards Generalized Real-World Agency.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Seed1.8 Model Card: Towards Generalized Real-World Agency

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:98670a4f47b0db8745300dc2d75cb40c96f93421113959b43ea72e4890321c56

Observation 502dc265-67b8-467e-bd31-97eb0af9def3 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:25b22617b41748fa969387c4e6e8f08c2e31b318f350732508f99dcb07032e90

Observation f9893b93-b9d0-42f5-8f8a-aa23860c5100 · outbound

This paper cites MobileGUI-RL: Advancing Mobile GUI Agent through Reinforcement Learning in Online Environment.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL MobileGUI-RL: Advancing Mobile GUI Agent through Reinforcement Learning in Online Environment

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:db819596c5c1f751f0928223b07dd03f1eb38ceabcd04011a9956b3edd1456db

Observation 1ad2afee-0ccc-4c4c-b390-d2c729b00246 · outbound

This paper cites Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:09d3f148d78ec7bfe28cd6a816bd348a33695556c7be4fbe59f1762551db7e6d

Observation 4f051194-9f2c-4999-84e5-0e16fcbcb860 · outbound

This paper cites ScienceBoard: Evaluating Multimodal Autonomous Agents in Realistic Scientific Workflows.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL ScienceBoard: Evaluating Multimodal Autonomous Agents in Realistic Scientific Workflows

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:736625a59714088c3546fa1608e846f9bf6c90233f9f5c709a4d584f464f6392

Observation b74c6ba6-0d27-49b2-b84d-a9b00bd6fbf5 · outbound

This paper cites UI-TARS-2 Technical Report: Advancing GUI Agent with Multi-Turn Reinforcement Learning.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL UI-TARS-2 Technical Report: Advancing GUI Agent with Multi-Turn Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:fb27e3694fb03b00e9568a8e02b449f9046322c11de6474a6916bb0e01e0f18c

Observation 94ed22cc-330f-412d-bdaa-bb9e2358b852 · outbound

This paper cites Agentsynth: Scalable task generation for generalist computer-use agents.arXiv preprint arXiv:2506.14205, 2025a.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Agentsynth: Scalable task generation for generalist computer-use agents.arXiv preprint arXiv:2506.14205, 2025a

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:a855a22677c8f4dc78cd2d27c22f97b73da6711aff2109352af129403492944b

Observation 95e88f73-9f88-4eae-83ec-d3a228595f0d · outbound

This paper cites Scaling computer-use grounding via user interface decomposition and synthesis.arXiv preprint arXiv:2505.13227, 2025b.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Scaling computer-use grounding via user interface decomposition and synthesis.arXiv preprint arXiv:2505.13227, 2025b

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:71d69494906319c31a354ee226e33904f62ace086606644eac94e3fdde6c991d

Observation d5546386-87e0-42bb-accb-77f2d703c242 · outbound

This paper cites Mobilerl: Online agentic reinforcement learning for mobile gui agents.arXiv preprint arXiv:2509.18119,.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Mobilerl: Online agentic reinforcement learning for mobile gui agents.arXiv preprint arXiv:2509.18119,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:d06aae9ad8766456e29b1d0bdc87b904b670d32ed8d04622b8e7580ced1ad839

Observation 4e4b3f1c-7045-4b0a-bfa0-17a11abee2a9 · outbound

This paper cites AgentTrek: Agent Trajectory Synthesis via Guiding Replay with Web Tutorials.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL AgentTrek: Agent Trajectory Synthesis via Guiding Replay with Web Tutorials

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:4cc65b4d83571f1bd5ceb8273cfb264ce81a2c6b02d61d0421d61ddc245333d5

Observation 6202e2ba-0fee-4776-b1c7-df3dc0e56d20 · outbound

This paper cites Evocua: Evolving computer use agents via learning from scalable synthetic experience.arXiv preprint arXiv:2601.15876,.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Evocua: Evolving computer use agents via learning from scalable synthetic experience.arXiv preprint arXiv:2601.15876,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:cf65726a487839bb9edf6477eaccf27e3873e06ffba9a5117aefa1858336cbbf

Observation b4c8fc65-81ae-4a55-ba50-6873b7dba976 · outbound

This paper cites Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:b8f6c90bca979334b18fb2f53e1c40b5a5747591f41e603ff46096d934641af6

Observation b93b78de-be2a-44df-b0d3-6f5e972cddca · outbound

This paper cites Qwen3 Technical Report.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Qwen3 Technical Report

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:a0209529f86a1c90699ddec416d5ffee0dfc81721ebacdbf57f2ef1d0c0dc385

Observation 4a19d40a-4490-4ba9-934a-265c9c3a1bff · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:32c4db550a2d64389646e4bedf9242fbc8871d09d04be5de76a17eb761a5d02c

Observation 3ae1de40-ba82-4c3f-8ce2-69c056ecca52 · outbound

This paper cites Agentrl: Scaling agentic reinforcement learning with a multi-turn, multi-task framework.arXiv preprint arXiv:2510.04206,.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Agentrl: Scaling agentic reinforcement learning with a multi-turn, multi-task framework.arXiv preprint arXiv:2510.04206,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:2ba37293fc1e3ae3c42b21436cc8bf8501c439587d04095100c30f8cdfbfaea7

Observation 6695e059-b3a6-4f8a-9f5e-94fc0249b227 · outbound

This paper cites WebArena: A Realistic Web Environment for Building Autonomous Agents.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL WebArena: A Realistic Web Environment for Building Autonomous Agents

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:7594338eb310b1358b3cfdf9685168486c2fba83b62ac49a990988b14b4a493c

Observation 3282f93e-ae12-4ea7-9bc7-acbf88a40564 · outbound

This paper cites an unresolved cited work.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:456fb135521c1d5fe692a4080478d75329ea233c046faf789073703bcdc6c555

Observation b21034aa-9e33-4325-985d-262f88f9dfad · outbound

This paper cites Larger K values produce fewer but longer segments, requiring more rollout workers to keep the training engine fed.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Larger K values produce fewer but longer segments, requiring more rollout workers to keep the training engine fed

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:8e4d2c2eb657798871d799dbad6e8899b3b88427f4576e6b13c6daa48429eb75

Pith citing papers

No inbound Pith citation observations are available.