Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T15:18:57.759016Z
Paper Citation Record · LEDGER
As of 22 August 2026, this Paper Citation Record lists 75 of 75 outbound references and 2 inbound Pith citation observations for arXiv:2508.20018.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T15:18:57.759016Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-11T04:18:11.054622Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-01T20:56:13.527455Z
75 of 75 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 19a9428a-3b77-4944-a35d-8b37187c5f34 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Chain-of-Thought Reasoning In The Wild Is Not Always Faithful
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9247b188-b405-46f1-9e63-c098dcd87294 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Qwen2.5-VL Technical Report
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b7b667c-f0b0-4f93-8021-8422d711e038 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Distributed optimization and statistical learning via the alternating direction method of multipliers
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51f022b9-3a0f-4636-a5d4-c2f0a9355429 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Multi-agent reinforcement learning: A review of challenges and applications
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3fbfef5-8364-43bf-84be-fef9801b2dbc · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control AMEX: Android Multi-annotation Expo Dataset for Mobile GUI Agents
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7693ce73-db26-4be6-956f-e200c9cea4bc · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control GUICourse: From General Vision Language Models to Versatile GUI Agents
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b37cf8a-1c2e-437e-95f0-0a5b2c3764b7 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Multi-agent deep reinforcement learning for large-scale traffic signal control
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation cee11735-16f4-47b2-8c2f-cc3bc3750374 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1ab69ff-a89c-409e-af21-f5939873460b · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Training Verifiers to Solve Math Word Problems
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31f0b094-f876-40d6-88b3-88b02d3ba3a4 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Process Reinforcement through Implicit Rewards
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0bc4c11f-ff7c-463c-8aba-3745abd36ad5 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Is Independent Learning All You Need in the StarCraft Multi-Agent Challenge?
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c28e3fa-83d7-488c-847c-7abee908cccf · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Improving factuality and reasoning in language models through multiagent debate
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5a78a3a-f459-42be-9329-c28463a4b888 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Plan-and-Act: Improving Planning of Agents for Long-Horizon Tasks
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dcd2a098-f0f9-4964-882c-fd237304bff4 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control On alternating direction methods of multipliers: a historical perspective
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 0d5f1473-ec08-49cf-adb0-92ed47a6983b · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Towards Efficient Multi-Agent Learning Systems
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 6a21393a-2d07-49a2-841b-36ec9af8bfe8 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Seed1.5-VL Technical Report
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afb744fa-3099-443c-8f5c-25843fb7e772 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control LLM Multi-Agent Systems: Challenges and Open Problems
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73c21b1c-42f9-43f4-9768-03b570521f72 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Measuring Mathematical Problem Solving With the MATH Dataset
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e0bcf3f-125c-4b0a-83a4-b367c3b90271 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control OpenRLHF: An Easy-to-use, Scalable and High-performance RLHF Framework
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd5fd225-1cce-4510-83b4-7298b75cac67 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control OWL: Optimized Workforce Learning for General Multi-Agent Assistance in Real-World Task Automation
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 479c9983-0611-44a2-a6f6-92044d09bca0 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Qwen2.5-Coder Technical Report
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53723fdb-ad20-4a58-8efc-0330967c96f2 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Omniact: A dataset and benchmark for enabling multimodal generalist autonomous agents for desktop and web
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 516ac072-e338-4278-8f3a-b64ce774873d · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Os-harm: A benchmark for measuring safety of computer use agents
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ebb81904-c3bd-4d6b-a063-2e0271315e49 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Gonzalez, Hao Zhang, and Ion Stoica
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5005a91c-eea3-4a3f-8c6b-4264ded03b52 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control ReGUIDE: Data Efficient GUI Grounding via Spatial Reasoning and Search
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9dd64cb2-b43d-4f66-ab36-7672dd6ce78c · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Towards Better Chain-of-Thought: A Reflection on Effectiveness and Faithfulness
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8bc6389-2215-4eb6-b20b-a0c28e13307d · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control On the Effects of Data Scale on UI Control Agents
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7284a6bd-a9f8-495d-a1aa-f45307bd4211 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control MARFT: Multi-Agent Reinforcement Fine-Tuning
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b9a3d2f-7874-4a61-8926-71a1fd82fa76 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control InfiGUI-R1: Advancing Multimodal GUI Agents from Reactive Actors to Deliberative Reasoners
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 458cafae-0407-4869-9f4c-0e3ea612a27f · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control GUIOdyssey: A Comprehensive Dataset for Cross-App GUI Navigation on Mobile Devices
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c2810cd-48c7-49a7-8fa4-e2f458c9bf5a · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a582d7b-166c-462e-a4f8-e4788bf0d221 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control GUI-R1 : A Generalist R1-Style Vision-Language Action Model For GUI Agents
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49d3dfd9-8be8-4b2a-928c-4195f544d630 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09649300-7f31-4425-9b3b-92956ccfcfbc · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Building a Stable Planner: An Extended Finite State Machine Based Planning Module for Mobile GUI Agent
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bdbf12c8-a705-442d-aefb-097151b4c606 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control ScreenAgent: A Vision Language Model-driven Computer Control Agent
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d78ccc9-e24d-4360-8109-b5a779693a66 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Introducing gpt-5
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 3b7fe2e0-b28c-4a28-af38-7d38bba51389 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control UI-TARS: Pioneering Automated GUI Interaction with Native Agents
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de128de1-a07d-49ec-a34e-c52f80ad201d · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Monotonic value function factorisation for deep multi-agent reinforcement learning
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b1898c5-e4b7-4cb8-a89d-fab0c23d349a · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control SAM 2: Segment Anything in Images and Videos
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8eb95c4-8a04-4a12-9b7e-00b83ed2644f · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Android in the Wild: A Large-Scale Dataset for Android Device Control
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee5652ef-c26f-4dd3-aada-6a145d69455b · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Trust region policy optimization
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05916953-5ccd-4cb2-90e9-7a4c871e725b · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Proximal Policy Optimization Algorithms
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e2203b9-d430-43a6-96c6-5973b923525a · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cafeca55-bc45-4a98-9a89-3e4a2671c7c3 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Hybridflow: A flexible and efficient rlhf framework
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a1a5993-fd84-49b9-85bc-de438734ccf1 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Towards trustworthy gui agents: A survey
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4e37ba6-4650-44be-9ec0-6ace697b82f9 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Teaching Models to Balance Resisting and Accepting Persuasion
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42ad7824-693b-46b6-be49-870abb151ffb · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06b385d4-bb0a-4e4e-8317-be9f29a75b0d · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Value-Decomposition Networks For Cooperative Multi-Agent Learning
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation efee8467-2e75-4412-b4ee-714b52fea9a2 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control GUI-G$^2$: Gaussian Reward Modeling for GUI Grounding
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ea87f0f-c321-4b53-88f4-05c8e77552d2 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Multi-Agent Collaboration Mechanisms: A Survey of LLMs
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15dfc07d-049f-45d0-b2ef-0db6b380bb91 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Language models don't always say what they think: Unfaithful explanations in chain-of-thought prompting
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36e0ce03-7710-4111-bb7d-cc717cf9c8ed · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control GUI Agents with Foundation Models: A Comprehensive Survey
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5853eb5-71da-4db5-a308-a125ecdc7a41 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Model-based Multi-agent Reinforcement Learning: Recent Progress and Prospects
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4aa2014-ce7f-4a6d-bfa6-8d1e0746810c · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Order Matters: Agent-by-agent Policy Optimization
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 4ac9fbff-9eb4-47ba-b500-9f9db7afe815 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Mp-gui: Modality perception with mllms for gui understanding
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 9a0bd143-9fe1-4573-be92-c8a2cd182172 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Chain-of-thought prompting elicits reasoning in large language models
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c95e31b8-acbe-416a-93d9-b2944dca21e5 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control CMATH: Can Your Language Model Pass Chinese Elementary School Math Test?
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bbbc11b9-4f2a-471b-9fcc-94d29a09fc57 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Autogen: Enabling next-gen llm applications via multi-agent conversations
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7648e41d-0bf0-4f75-911c-fcebbbcfce67 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control OS-ATLAS: A Foundation Action Model for Generalist GUI Agents
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12716f24-d199-44e6-a838-6115cd11bfe8 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Towards System 2 Reasoning in LLMs: Learning How to Think With Meta Chain-of-Thought
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8517c31a-4e46-4c87-a23f-9a9cacc65512 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control TradingAgents: Multi-Agents LLM Financial Trading Framework
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a10d33a-8219-4629-be6d-f6b3f46e4223 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Aguvis: Unified Pure Vision Agents for Autonomous GUI Interaction
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62017697-9004-444f-bc61-ed919caba547 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control A Survey of ADMM Variants for Distributed Optimization: Problems, Algorithms and Features
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation acbf7106-55c1-414b-9fcf-700eac3af2cc · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Enhancing Visual Grounding for GUI Agents via Self-Evolutionary Reinforcement Learning
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d4c849e-b94b-4919-854d-6ae0ccbc72bf · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Android in the Zoo: Chain-of-Action-Thought for GUI Agents
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54090172-8602-4cd3-a23d-2e31ad3adf6e · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Does Chain-of-Thought Reasoning Help Mobile GUI Agent? An Empirical Study
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68065708-a7ee-4821-a03d-6e3de2cb1aee · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control LongAgent: Scaling Language Models to 128k Context through Multi-Agent Collaboration
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1f48cc3-e73f-4732-98c0-d28292401d05 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control SiriuS: Self-improving Multi-agent Systems via Bootstrapped Reasoning
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae57306c-22d6-4f8f-adc4-9a549141dbf7 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control An Electoral Approach to Diversify LLM-based Multi-Agent Collective Decision-Making
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 255449bd-01f8-40f9-952c-d365d48d31fc · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Heterogeneous-agent reinforcement learning
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 94bccebf-59c1-42b4-8859-a20ac12ff831 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58f85316-a1b2-4b03-8e32-8a89c58b8ed1 · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15fab545-d472-4144-b1a3-d03e31c4904f · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control @esa (Ref
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 965654c6-fcef-47ac-a38e-df6f1c5d873b · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Unresolved cited work
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4935e77-8b93-42f9-9a55-92011d974d5f · outbound
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Unresolved cited work
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d05a0cc-0145-4168-b782-a48fe1b9e486 · inbound
Trust Region On-Policy Distillation SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation d7be8218-0e2a-4828-bc27-b9e00949bb21 · inbound
When the Judge Should Not Decide: Evidence-Locked, Non-Compensatory Selection Bounds LLM-Judge Failure in Reasoning Pipelines SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.