Pith. sign in

Paper Citation Record · LEDGER

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control

As of 22 August 2026, this Paper Citation Record lists 75 of 75 outbound references and 2 inbound Pith citation observations for arXiv:2508.20018.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.20018 v1

Coverage vector

measured 75 of 75 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T15:18:57.759016Z

measured 77 of 77 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T04:18:11.054622Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T20:56:13.527455Z

Reference resolution

75 of 75 outbound references displayed

  • verified exact3
  • verified fuzzy6
  • unresolved66
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 19a9428a-3b77-4944-a35d-8b37187c5f34 · outbound

This paper cites Chain-of-Thought Reasoning In The Wild Is Not Always Faithful.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Chain-of-Thought Reasoning In The Wild Is Not Always Faithful

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:50.382652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:50.382652Z digest=sha256:dcd3d3e99ae67307185606366b517e059a1c431a3df899398b59f322ab99f46e

Observation 9247b188-b405-46f1-9e63-c098dcd87294 · outbound

This paper cites Qwen2.5-VL Technical Report.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Qwen2.5-VL Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:50.510010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:50.510010Z digest=sha256:af53c177e2a569ad62fc80f200d72de48445d349649245b4508ce06fe33fe658

Observation 7b7b667c-f0b0-4f93-8021-8422d711e038 · outbound

This paper cites Distributed optimization and statistical learning via the alternating direction method of multipliers.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Distributed optimization and statistical learning via the alternating direction method of multipliers

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:50.622301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:50.622301Z digest=sha256:5cbe7048159d06b0c1061a50a09625c12df4ffd04ae9b2fd4783c480b92edd31

Observation 51f022b9-3a0f-4636-a5d4-c2f0a9355429 · outbound

This paper cites Multi-agent reinforcement learning: A review of challenges and applications.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Multi-agent reinforcement learning: A review of challenges and applications

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:50.743598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:50.743598Z digest=sha256:34c7ae12c1be78b59eff9a7840d2c5d966677e2fe8e2ef8536ca9ddca170ef3e

Observation e3fbfef5-8364-43bf-84be-fef9801b2dbc · outbound

This paper cites AMEX: Android Multi-annotation Expo Dataset for Mobile GUI Agents.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control AMEX: Android Multi-annotation Expo Dataset for Mobile GUI Agents

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:50.874256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:50.874256Z digest=sha256:6c42f2d9cc497d3f43a69f551d517bd067c1a121e3c8fc276e532e0caf74709f

Observation 7693ce73-db26-4be6-956f-e200c9cea4bc · outbound

This paper cites GUICourse: From General Vision Language Models to Versatile GUI Agents.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control GUICourse: From General Vision Language Models to Versatile GUI Agents

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:50.998350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:50.998350Z digest=sha256:656feb2b41ac3d3f3dacac5b35360eb1d7fada29329c5a25b3910a87d45f1339

Observation 4b37cf8a-1c2e-437e-95f0-0a5b2c3764b7 · outbound

This paper cites Multi-agent deep reinforcement learning for large-scale traffic signal control.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Multi-agent deep reinforcement learning for large-scale traffic signal control

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:59.714977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T15:18:51.107151Z digest=sha256:7c4791c5d29ef0cb94ef05fd017515692db237e1d3ca329ed2768dc745a7f6bc

Observation cee11735-16f4-47b2-8c2f-cc3bc3750374 · outbound

This paper cites SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:51.219851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:51.219851Z digest=sha256:ef7d8eede8139619b33a316e0a774ff3c65728ee7ba35a0aca753ade83769507

Observation c1ab69ff-a89c-409e-af21-f5939873460b · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Training Verifiers to Solve Math Word Problems

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:51.304192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:51.304192Z digest=sha256:87e9c47e02b0efa0ed924973f0d7ebe4b10c0739d87a5131964c3551e23918f9

Observation 31f0b094-f876-40d6-88b3-88b02d3ba3a4 · outbound

This paper cites Process Reinforcement through Implicit Rewards.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Process Reinforcement through Implicit Rewards

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:51.436257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:51.436257Z digest=sha256:fedf480564e32ae132d43dd1a8b94f8be3c32f18270598244a588ca1ed8afed5

Observation 0bc4c11f-ff7c-463c-8aba-3745abd36ad5 · outbound

This paper cites Is Independent Learning All You Need in the StarCraft Multi-Agent Challenge?.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Is Independent Learning All You Need in the StarCraft Multi-Agent Challenge?

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:51.513389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:51.513389Z digest=sha256:f1880a28699de8de00abbfc37a02cb784896f17a400624a909e20fb5cd4e1d87

Observation 9c28e3fa-83d7-488c-847c-7abee908cccf · outbound

This paper cites Improving factuality and reasoning in language models through multiagent debate.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Improving factuality and reasoning in language models through multiagent debate

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:51.621480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:51.621480Z digest=sha256:8f56b9cf751474b6b613d84c87bd1a730de7523705f2d58558e7dd6d25255054

Observation e5a78a3a-f459-42be-9329-c28463a4b888 · outbound

This paper cites Plan-and-Act: Improving Planning of Agents for Long-Horizon Tasks.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Plan-and-Act: Improving Planning of Agents for Long-Horizon Tasks

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:51.734773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:51.734773Z digest=sha256:fd3bc1c99fbb3b100037b39e90fbe3d9b532cb2f85165c786f481149cc7de2cc

Observation dcd2a098-f0f9-4964-882c-fd237304bff4 · outbound

This paper cites On alternating direction methods of multipliers: a historical perspective.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control On alternating direction methods of multipliers: a historical perspective

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:59.677196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T15:18:51.853425Z digest=sha256:601d7f3c28c4175b50e1ce005f3cfc0478995c3c8820f5a95b9cdb357d6e9603

Observation 0d5f1473-ec08-49cf-adb0-92ed47a6983b · outbound

This paper cites Towards Efficient Multi-Agent Learning Systems.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Towards Efficient Multi-Agent Learning Systems

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-05T15:18:59.149753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T15:18:51.963528Z digest=sha256:eb8acd1aeb9b42f0c3bdfe4eb2f3dc578c47f639f9774fd403e43ed6a7e3819f

Observation 6a21393a-2d07-49a2-841b-36ec9af8bfe8 · outbound

This paper cites Seed1.5-VL Technical Report.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Seed1.5-VL Technical Report

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:52.028713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:52.028713Z digest=sha256:9953ab90e0ab183489941d30911d1ff1cd341ef83d479b4939047ac5be4e4ced

Observation afb744fa-3099-443c-8f5c-25843fb7e772 · outbound

This paper cites LLM Multi-Agent Systems: Challenges and Open Problems.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control LLM Multi-Agent Systems: Challenges and Open Problems

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:52.148823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:52.148823Z digest=sha256:9e63741ab9ed0d7595b2265e1cd4cf518892e1eb305f81e7f1e64f8f6316d311

Observation 73c21b1c-42f9-43f4-9768-03b570521f72 · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Measuring Mathematical Problem Solving With the MATH Dataset

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:52.263559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:52.263559Z digest=sha256:70c160a09bf2831e0e0fe4f219bd89a8ba2b1237a2bad6db6a98d09898d5d90c

Observation 3e0bcf3f-125c-4b0a-83a4-b367c3b90271 · outbound

This paper cites OpenRLHF: An Easy-to-use, Scalable and High-performance RLHF Framework.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control OpenRLHF: An Easy-to-use, Scalable and High-performance RLHF Framework

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:52.384772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:52.384772Z digest=sha256:c14abd7f569969e25e60d5b84970bf99f37f5445a3bd9cea353376396213b452

Observation fd5fd225-1cce-4510-83b4-7298b75cac67 · outbound

This paper cites OWL: Optimized Workforce Learning for General Multi-Agent Assistance in Real-World Task Automation.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control OWL: Optimized Workforce Learning for General Multi-Agent Assistance in Real-World Task Automation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:52.500880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:52.500880Z digest=sha256:d024be24a5e59877baad4da88ce32db27da32b6ebd2e25262ce7e8e24b613e65

Observation 479c9983-0611-44a2-a6f6-92044d09bca0 · outbound

This paper cites Qwen2.5-Coder Technical Report.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Qwen2.5-Coder Technical Report

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:52.630612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:52.630612Z digest=sha256:82ca27e19de2d8a64531c1332ef135d9510129aac2d32a0ae62ddb61d2bb3b6b

Observation 53723fdb-ad20-4a58-8efc-0330967c96f2 · outbound

This paper cites Omniact: A dataset and benchmark for enabling multimodal generalist autonomous agents for desktop and web.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Omniact: A dataset and benchmark for enabling multimodal generalist autonomous agents for desktop and web

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:59.653591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T15:18:52.735866Z digest=sha256:c741653f047a403212b58017caef585ccc9cbd98d1b112de275ca622ff2d24e5

Observation 516ac072-e338-4278-8f3a-b64ce774873d · outbound

This paper cites Os-harm: A benchmark for measuring safety of computer use agents.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Os-harm: A benchmark for measuring safety of computer use agents

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:52.862777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:52.862777Z digest=sha256:d6261edbe48008234a298d2153d75546e794db84191ec0e5eecfb16adc08417e

Observation ebb81904-c3bd-4d6b-a063-2e0271315e49 · outbound

This paper cites Gonzalez, Hao Zhang, and Ion Stoica.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Gonzalez, Hao Zhang, and Ion Stoica

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:52.977243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:52.977243Z digest=sha256:50783a4675ce0411cac6aea239d011930e20f04bef548be93504c02910915b1d

Observation 5005a91c-eea3-4a3f-8c6b-4264ded03b52 · outbound

This paper cites ReGUIDE: Data Efficient GUI Grounding via Spatial Reasoning and Search.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control ReGUIDE: Data Efficient GUI Grounding via Spatial Reasoning and Search

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:53.098539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:53.098539Z digest=sha256:75e52fe35e0f9deffc8355d88053d970dba13bf1b5b008f206672b1577fa507b

Observation 9dd64cb2-b43d-4f66-ab36-7672dd6ce78c · outbound

This paper cites Towards Better Chain-of-Thought: A Reflection on Effectiveness and Faithfulness.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Towards Better Chain-of-Thought: A Reflection on Effectiveness and Faithfulness

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:53.252747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:53.252747Z digest=sha256:89dc81153a537102d10229afef84cbd0534ce5a0e6ba36f68e5b820b603abf2c

Observation e8bc6389-2215-4eb6-b20b-a0c28e13307d · outbound

This paper cites On the Effects of Data Scale on UI Control Agents.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control On the Effects of Data Scale on UI Control Agents

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:53.370344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:53.370344Z digest=sha256:e5b0bab5aac81f706a0cd1e522811794c17fb195b08fbda4ab647fd9b7d885ba

Observation 7284a6bd-a9f8-495d-a1aa-f45307bd4211 · outbound

This paper cites MARFT: Multi-Agent Reinforcement Fine-Tuning.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control MARFT: Multi-Agent Reinforcement Fine-Tuning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:53.497423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:53.497423Z digest=sha256:13d9720c5085cf0283e3b22df8e337647d27903cd95748400b72ef60bb164021

Observation 7b9a3d2f-7874-4a61-8926-71a1fd82fa76 · outbound

This paper cites InfiGUI-R1: Advancing Multimodal GUI Agents from Reactive Actors to Deliberative Reasoners.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control InfiGUI-R1: Advancing Multimodal GUI Agents from Reactive Actors to Deliberative Reasoners

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:53.536887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:53.536887Z digest=sha256:6d5510aee18f28eba4f7979ed2d5c932851855a97d02e46f0c3679104c69ffc0

Observation 458cafae-0407-4869-9f4c-0e3ea612a27f · outbound

This paper cites GUIOdyssey: A Comprehensive Dataset for Cross-App GUI Navigation on Mobile Devices.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control GUIOdyssey: A Comprehensive Dataset for Cross-App GUI Navigation on Mobile Devices

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:53.611757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:53.611757Z digest=sha256:e8353b48d1d78f719511d31eae130c3bb05f47d32e55ec71e0ba441ccb5325c5

Observation 4c2810cd-48c7-49a7-8fa4-e2f458c9bf5a · outbound

This paper cites UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:53.726458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:53.726458Z digest=sha256:c38b922f16cfe1c7560829e5cf855d4c66a78fc7f899fc4d688ef80f0ef66394

Observation 3a582d7b-166c-462e-a4f8-e4788bf0d221 · outbound

This paper cites GUI-R1 : A Generalist R1-Style Vision-Language Action Model For GUI Agents.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control GUI-R1 : A Generalist R1-Style Vision-Language Action Model For GUI Agents

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:53.855437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:53.855437Z digest=sha256:47e60bc728949d1d5e22dab52a8d5a437cfe089a4d45e6543310ed6700e964c8

Observation 49d3dfd9-8be8-4b2a-928c-4195f544d630 · outbound

This paper cites MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:53.972497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:53.972497Z digest=sha256:900341020ba5472adf40091fe0777f73c94c4cfa5ce2e1106677d029d4747ce7

Observation 09649300-7f31-4425-9b3b-92956ccfcfbc · outbound

This paper cites Building a Stable Planner: An Extended Finite State Machine Based Planning Module for Mobile GUI Agent.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Building a Stable Planner: An Extended Finite State Machine Based Planning Module for Mobile GUI Agent

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:54.078626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:54.078626Z digest=sha256:b245a2bcf752e7644c43664d4f8d6e5e29eab1415594b5fdec65516fc27c101e

Observation bdbf12c8-a705-442d-aefb-097151b4c606 · outbound

This paper cites ScreenAgent: A Vision Language Model-driven Computer Control Agent.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control ScreenAgent: A Vision Language Model-driven Computer Control Agent

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:54.236190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:54.236190Z digest=sha256:1234d2afe0c583d89e6e7a83d44e679f16d71365e3657d70ef29355e635271f9

Observation 0d78ccc9-e24d-4360-8109-b5a779693a66 · outbound

This paper cites Introducing gpt-5.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Introducing gpt-5

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:59.614982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T15:18:54.371511Z digest=sha256:0f11f372ccb54a7567a918ee75229e45612a1645d4f94785ad6712a6a2188520

Observation 3b7fe2e0-b28c-4a28-af38-7d38bba51389 · outbound

This paper cites UI-TARS: Pioneering Automated GUI Interaction with Native Agents.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control UI-TARS: Pioneering Automated GUI Interaction with Native Agents

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:54.490511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:54.490511Z digest=sha256:562d080fd344c76c7ca32269de7b887f0315a147c6a53d689c30ad157b9d4f53

Observation de128de1-a07d-49ec-a34e-c52f80ad201d · outbound

This paper cites Monotonic value function factorisation for deep multi-agent reinforcement learning.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Monotonic value function factorisation for deep multi-agent reinforcement learning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:54.649045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:54.649045Z digest=sha256:dad516b76a48d21bff6c2983d4ac19a4ede2bd24f011b95cdfa6176dfac6bb74

Observation 7b1898c5-e4b7-4cb8-a89d-fab0c23d349a · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control SAM 2: Segment Anything in Images and Videos

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:54.784911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:54.784911Z digest=sha256:32a8a94fea3e6129a21891c3cc65b63adf99a832bc985def4616ce4dc4cac0d0

Observation e8eb95c4-8a04-4a12-9b7e-00b83ed2644f · outbound

This paper cites Android in the Wild: A Large-Scale Dataset for Android Device Control.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Android in the Wild: A Large-Scale Dataset for Android Device Control

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:54.955390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:54.955390Z digest=sha256:67286efa6ab6306b8fd53af0ba802354ea0f36aeddb0e57d3f5591bfd21bf16e

Observation ee5652ef-c26f-4dd3-aada-6a145d69455b · outbound

This paper cites Trust region policy optimization.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Trust region policy optimization

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:55.056102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:55.056102Z digest=sha256:1d09ab9d0734c8a92b42d1f37406045b072cd8d8fa3d6d420c0f2c2dac375516

Observation 05916953-5ccd-4cb2-90e9-7a4c871e725b · outbound

This paper cites Proximal Policy Optimization Algorithms.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Proximal Policy Optimization Algorithms

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:55.169460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:55.169460Z digest=sha256:0e0077a1a6df5ee327cb4eefdac1a1313e5d99200a163fa227917b98420504f0

Observation 3e2203b9-d430-43a6-96c6-5973b923525a · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:55.330893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:55.330893Z digest=sha256:58568363294a04efc89c748757eb61fa032ba9b80ac6f2fb951b2f44dd49d50c

Observation cafeca55-bc45-4a98-9a89-3e4a2671c7c3 · outbound

This paper cites Hybridflow: A flexible and efficient rlhf framework.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Hybridflow: A flexible and efficient rlhf framework

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:55.493051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:55.493051Z digest=sha256:7ee8a5d2a64495a1699976643b02f90fde88cae3d793d718bc44b6378c581f68

Observation 8a1a5993-fd84-49b9-85bc-de438734ccf1 · outbound

This paper cites Towards trustworthy gui agents: A survey.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Towards trustworthy gui agents: A survey

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:55.630193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:55.630193Z digest=sha256:bf60c7c930f3cbddd85116ba91b81cb63a09b225e941ed8e3c5163ba558fe522

Observation f4e37ba6-4650-44be-9ec0-6ace697b82f9 · outbound

This paper cites Teaching Models to Balance Resisting and Accepting Persuasion.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Teaching Models to Balance Resisting and Accepting Persuasion

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:55.797459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:55.797459Z digest=sha256:ae96cb645d5f5214e7c15bbeaf7620b5e5e7c1b75713c2454444811aa9375778

Observation 42ad7824-693b-46b6-be49-870abb151ffb · outbound

This paper cites Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:55.902002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:55.902002Z digest=sha256:de2872e04be5c561386dda85df011ee9a269092974b90457e4ade9de32ebf43c

Observation 06b385d4-bb0a-4e4e-8317-be9f29a75b0d · outbound

This paper cites Value-Decomposition Networks For Cooperative Multi-Agent Learning.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:56.030027Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:56.030027Z digest=sha256:063cc11e5680da3cad4007a4148bd5f91cc54bd5b072438e2e1346b53b1f3fc7

Observation efee8467-2e75-4412-b4ee-714b52fea9a2 · outbound

This paper cites GUI-G$^2$: Gaussian Reward Modeling for GUI Grounding.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control GUI-G$^2$: Gaussian Reward Modeling for GUI Grounding

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:56.155953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:56.155953Z digest=sha256:98bb3f2d035b10ee22f17c17024d39f72312d3822e78865ccaa0272d247fad3a

Observation 3ea87f0f-c321-4b53-88f4-05c8e77552d2 · outbound

This paper cites Multi-Agent Collaboration Mechanisms: A Survey of LLMs.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Multi-Agent Collaboration Mechanisms: A Survey of LLMs

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:56.487424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:56.487424Z digest=sha256:a954dba34885c2f565cd2624b9fb48137731fed9d6cd6bab3f06d0f15c5aded6

Observation 15dfc07d-049f-45d0-b2ef-0db6b380bb91 · outbound

This paper cites Language models don't always say what they think: Unfaithful explanations in chain-of-thought prompting.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Language models don't always say what they think: Unfaithful explanations in chain-of-thought prompting

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:56.755000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:56.755000Z digest=sha256:d3c329ee952775eddc4bc6bc77d714ceda3b60f61772e694f23be68ce2a77c71

Observation 36e0ce03-7710-4111-bb7d-cc717cf9c8ed · outbound

This paper cites GUI Agents with Foundation Models: A Comprehensive Survey.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control GUI Agents with Foundation Models: A Comprehensive Survey

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:56.900795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:56.900795Z digest=sha256:d82fa4118b7c611f5db5302a9e66ea716ea3f8860108caaa0501a5505d955d3c

Observation a5853eb5-71da-4db5-a308-a125ecdc7a41 · outbound

This paper cites Model-based Multi-agent Reinforcement Learning: Recent Progress and Prospects.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Model-based Multi-agent Reinforcement Learning: Recent Progress and Prospects

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:57.030594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:57.030594Z digest=sha256:cc508a0986c346bde2bb78b04f80ddaf37c2dfab7b27278adafac88d52a73c61

Observation c4aa2014-ce7f-4a6d-bfa6-8d1e0746810c · outbound

This paper cites Order Matters: Agent-by-agent Policy Optimization.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Order Matters: Agent-by-agent Policy Optimization

Reference 54

Resolution
verified exact
local_arxiv, observed 2026-08-05T15:18:58.169804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T15:18:57.147045Z digest=sha256:722fd5a3f611050cba082d4821d3cd3360b6083d7fdffcb482ae90a0a0900968

Observation 4ac9fbff-9eb4-47ba-b500-9f9db7afe815 · outbound

This paper cites Mp-gui: Modality perception with mllms for gui understanding.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Mp-gui: Modality perception with mllms for gui understanding

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:59.533792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T15:18:57.243786Z digest=sha256:94b5c5af2f1f6856f3e48e75deb2c69af9189a72488f7ea593f6082941ab6ae0

Observation 9a0bd143-9fe1-4573-be92-c8a2cd182172 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Chain-of-thought prompting elicits reasoning in large language models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:57.352490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:57.352490Z digest=sha256:e00f700068a113fa5ddc078001bee0d70c842244a2201699e46640fdd3b7dc68

Observation c95e31b8-acbe-416a-93d9-b2944dca21e5 · outbound

This paper cites CMATH: Can Your Language Model Pass Chinese Elementary School Math Test?.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control CMATH: Can Your Language Model Pass Chinese Elementary School Math Test?

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:57.423607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:57.423607Z digest=sha256:584c49ed7f59dfc78916f1669e4d10f6b6d8b63ebcba7bf952f0a37cbc1723b4

Observation bbbc11b9-4f2a-471b-9fcc-94d29a09fc57 · outbound

This paper cites Autogen: Enabling next-gen llm applications via multi-agent conversations.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Autogen: Enabling next-gen llm applications via multi-agent conversations

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:57.545356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:57.545356Z digest=sha256:b1072784d7dba3eb43daf43986d811969b50f9d869188500b5c3585fead288e8

Observation 7648e41d-0bf0-4f75-911c-fcebbbcfce67 · outbound

This paper cites OS-ATLAS: A Foundation Action Model for Generalist GUI Agents.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control OS-ATLAS: A Foundation Action Model for Generalist GUI Agents

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:57.651708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:57.651708Z digest=sha256:111c2f1eef615fe4dc83c07ffe74af442741e3bc90776ce5030ec38e44525616

Observation 12716f24-d199-44e6-a838-6115cd11bfe8 · outbound

This paper cites Towards System 2 Reasoning in LLMs: Learning How to Think With Meta Chain-of-Thought.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Towards System 2 Reasoning in LLMs: Learning How to Think With Meta Chain-of-Thought

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:57.675232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:57.675232Z digest=sha256:4ba606a81481609ce5dba45456d898bc37135f18c95d64ab228ecc6fdda28d0c

Observation 8517c31a-4e46-4c87-a23f-9a9cacc65512 · outbound

This paper cites TradingAgents: Multi-Agents LLM Financial Trading Framework.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control TradingAgents: Multi-Agents LLM Financial Trading Framework

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:57.680409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:57.680409Z digest=sha256:46647a26739437d68ccb7b0e197c4893e032a8343b0a2ce150ad335398dbb0fe

Observation 3a10d33a-8219-4629-be6d-f6b3f46e4223 · outbound

This paper cites Aguvis: Unified Pure Vision Agents for Autonomous GUI Interaction.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Aguvis: Unified Pure Vision Agents for Autonomous GUI Interaction

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:57.686718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:57.686718Z digest=sha256:8977e06b5232fdf83ff00bd2b292bddc7362ae2aaf50971014ec15ef3ff7405e

Observation 62017697-9004-444f-bc61-ed919caba547 · outbound

This paper cites A Survey of ADMM Variants for Distributed Optimization: Problems, Algorithms and Features.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control A Survey of ADMM Variants for Distributed Optimization: Problems, Algorithms and Features

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:57.691983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:57.691983Z digest=sha256:f7d34d9a8520c939e747aaa31f0ccce1c6f26886edbda7ccda9c3f61f383e11e

Observation acbf7106-55c1-414b-9fcf-700eac3af2cc · outbound

This paper cites Enhancing Visual Grounding for GUI Agents via Self-Evolutionary Reinforcement Learning.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Enhancing Visual Grounding for GUI Agents via Self-Evolutionary Reinforcement Learning

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:57.697193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:57.697193Z digest=sha256:befa06edd9dd752d1d6123d003c61adffebec7b0f477fbe21996e4ebbe56e9b5

Observation 1d4c849e-b94b-4919-854d-6ae0ccbc72bf · outbound

This paper cites Android in the Zoo: Chain-of-Action-Thought for GUI Agents.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Android in the Zoo: Chain-of-Action-Thought for GUI Agents

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:57.701929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:57.701929Z digest=sha256:a24698cd7998a34a91f8bf08390c2a152aac9424b214722abf3aeb32081fdbc7

Observation 54090172-8602-4cd3-a23d-2e31ad3adf6e · outbound

This paper cites Does Chain-of-Thought Reasoning Help Mobile GUI Agent? An Empirical Study.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Does Chain-of-Thought Reasoning Help Mobile GUI Agent? An Empirical Study

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:57.707774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:57.707774Z digest=sha256:351bd011bafc53b6f39752588af0e3882bf5b9bbefe8f18252bcee1ed8f968c6

Observation 68065708-a7ee-4821-a03d-6e3de2cb1aee · outbound

This paper cites LongAgent: Scaling Language Models to 128k Context through Multi-Agent Collaboration.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control LongAgent: Scaling Language Models to 128k Context through Multi-Agent Collaboration

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:57.713495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:57.713495Z digest=sha256:00d461df992fd11e032f385bdb22645d68130c9bf2bffb02c51fbd92d5b5a2b2

Observation c1f48cc3-e73f-4732-98c0-d28292401d05 · outbound

This paper cites SiriuS: Self-improving Multi-agent Systems via Bootstrapped Reasoning.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control SiriuS: Self-improving Multi-agent Systems via Bootstrapped Reasoning

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:57.718882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:57.718882Z digest=sha256:2854a4b68344742c67935bceddcb2c63c214707880d06c53f08ac710302be43a

Observation ae57306c-22d6-4f8f-adc4-9a549141dbf7 · outbound

This paper cites An Electoral Approach to Diversify LLM-based Multi-Agent Collective Decision-Making.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control An Electoral Approach to Diversify LLM-based Multi-Agent Collective Decision-Making

Reference 69

Resolution
verified exact
local_arxiv, observed 2026-08-05T15:18:57.856062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T15:18:57.723824Z digest=sha256:7d87825e44aead86a6d6e5f843c8cfd17b8089d55a7291b094ee159af0b9a01b

Observation 255449bd-01f8-40f9-952c-d365d48d31fc · outbound

This paper cites Heterogeneous-agent reinforcement learning.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Heterogeneous-agent reinforcement learning

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:59.481072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T15:18:57.728971Z digest=sha256:55bba6a0bb9c2fb402bcb1e7113ad3077099dcd3dd24fe4b93d8e4d2cf44306f

Observation 94bccebf-59c1-42b4-8859-a20ac12ff831 · outbound

This paper cites GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:57.734021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:57.734021Z digest=sha256:db29155f91d45ad7d6d39408530eaa311711dc24e9169b6ea649fa34b26a9694

Observation 58f85316-a1b2-4b03-8e32-8a89c58b8ed1 · outbound

This paper cites InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:57.739096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:57.739096Z digest=sha256:64392526b870b8c7f57c64a1d40b7011f3d79ce3bc36e84480efaf791133547b

Observation 15fab545-d472-4144-b1a3-d03e31c4904f · outbound

This paper cites @esa (Ref.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control @esa (Ref

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:57.745032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:57.745032Z digest=sha256:44fc5397f44a9b5c8d22c3d7318155270cfd0dae4321e255b38a6bca5678d414

Observation 965654c6-fcef-47ac-a38e-df6f1c5d873b · outbound

This paper cites an unresolved cited work.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Unresolved cited work

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:57.753369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:57.753369Z digest=sha256:8c4b5f2036e55148899a2b5c5c53735b81a4e1183b36f307a446855fbb19358b

Observation c4935e77-8b93-42f9-9a55-92011d974d5f · outbound

This paper cites an unresolved cited work.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Unresolved cited work

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:57.759016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:57.759016Z digest=sha256:6eb1ea9199559bb121c11754d5e5cedceb0cfbe416d519c99a6afa4b72ce1ea7

Pith citing papers

Observation 9d05a0cc-0145-4168-b782-a48fe1b9e486 · inbound

Trust Region On-Policy Distillation cites this paper.

Trust Region On-Policy Distillation SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-07-01T20:56:13.528809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-28T17:38:50.313305Z digest=sha256:4143f7b5a135822cdc0adda5a2bd0736cba558188f3e86f5b5e229ab6f2536e3

Observation d7be8218-0e2a-4828-bc27-b9e00949bb21 · inbound

When the Judge Should Not Decide: Evidence-Locked, Non-Compensatory Selection Bounds LLM-Judge Failure in Reasoning Pipelines cites this paper.

When the Judge Should Not Decide: Evidence-Locked, Non-Compensatory Selection Bounds LLM-Judge Failure in Reasoning Pipelines SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-11T04:18:11.054622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:18:11.054622Z digest=sha256:3018cd53079a41560b0f8ba1b38ecfd14a675beb58889dedb652be18e16028fe