Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-28T15:46:50.587684Z
Paper Citation Record · LEDGER
As of 12 August 2026, this Paper Citation Record lists 72 of 72 outbound references and 2 inbound Pith citation observations for arXiv:2606.02031.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-28T15:46:50.587684Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-04T00:45:32.126811Z
A source-named dated measurement, never combined with another source.
Source: cited_works
72 of 72 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 0ee32fac-a6fb-4f7d-ae52-b509651b9d76 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 7f02ee38-cd25-4965-8ed5-5784fb0fc937 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Surfer-H Meets Holo1: Cost-Efficient Web Agent Powered by Open Weights
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 3f5eb26f-bd39-4b89-bd5e-8cf024ff023a · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Fara-7b: An efficient agentic model for computer use
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 34702829-b451-433f-852d-0ae36c114e07 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents WebGym: Scaling Training Environments for Visual Web Agents with Realistic Tasks
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 694b5164-defb-4340-bb0d-7e54549e8f21 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents DigiRL: Training in-the-wild device-control agents with autonomous reinforcement learning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c55bcdab-1331-4eda-9752-5468ffed1a99 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Qwen3-VL Technical Report
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 2a4a2e64-87bf-4942-8f14-a23fbbc68db7 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Web agents with world models: Learning and leveraging environment dynamics in web navigation
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e54dd26e-4112-426c-a1b9-798a51d52b86 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents arXiv preprint arXiv:2510.12693 , year=
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 087543f3-3c2a-4b47-b709-a579650171d3 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents CAPTCHA Solving for Native GUI Agents: Automated Reasoning-Action Data Generation and Self-Corrective Training
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a41d36ae-09a0-4df7-8a3b-016c80e6aca7 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d254b99f-7e51-4a8b-90f2-e15b2dd88ad3 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Seeclick: Harnessing gui grounding for advanced visual gui agents
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7cca339-63da-499a-bbba-4ae98d91ec8c · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Mind2web: Towards a generalist agent for the web.Advances in Neural Information Processing Systems, 36:28091–28114, 2023
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b911a0f9-9836-4ff1-ae24-c93f367885bf · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Scaling laws for reward model overoptimization
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e74c58f-b2d7-4521-bedb-9ceb6e623bc7 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Navigating the digital world as humans do: Universal visual grounding for gui agents
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93b96fca-c486-4927-a8f5-2f21c814485a · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Is Your LLM Secretly a World Model of the Internet? Model-Based Planning for Web Agents
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation df3c7960-c60d-48c0-af59-127219d61ad9 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Deepseek-r1 incentivizes reasoning in llms through reinforcement learning.Nature, 645(8081):633–638, 2025
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f00ee4d4-aa91-423a-8866-6e5378792912 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents MolmoWeb: Open Visual Web Agent and Open Data for the Open Web
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 4c8c2e25-0f07-471a-b41a-568e3cb8d7e9 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Webvoyager: Building an end-to-end web agent with large multimodal models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22c49a81-ea69-43cf-8c15-5d79817fceae · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Openwebvoyager: Building multimodal web agents via iterative real-world ex- ploration, feedback and optimization
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7fe0526-818a-439f-9f56-0fc4732152e7 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Scalable data synthesis for computer use agents with step-level filtering,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 46d9c196-30b5-4289-b1dc-d5b5f0a6d50c · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Glm-4.1 v-thinking: Towards versatile multimodal reasoning with scalable reinforcement learning.arXiv e-prints, pages arXiv–2507, 2025
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1e13c10-827a-4c66-ba1f-01450b12a6df · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Embodied web agents: Bridging physical-digital realms for integrated agent intelligence.Advances in Neural Information Processing Systems, 38, 2026
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85403a82-da14-4642-8df6-f0a5815ed867 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Rethinking memory mechanisms of foundation agents in the second half
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 62901e59-52a5-4b31-916a-8d4ba98fbe3f · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 6aaa6ab2-3b78-4245-b8fc-559b8841929e · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Scalecua: Scaling open-source computer use agents with cross-platform data
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 501acd53-cba0-4fb0-be15-00a0630bcd25 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Visual-rft: Visual reinforcement fine-tuning
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34605548-2744-40eb-8380-9f91743c5182 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Agentrewardbench: Evaluating automatic evaluations of web agent trajectories
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 57443a4a-4acc-4129-af17-4b13b2443f17 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Ui-r1: Enhancing efficient action prediction of gui agents by reinforcement learning
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05b64c83-ada2-48db-9b2d-7b1014d58eec · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents GUI-R1 : A Generalist R1-Style Vision-Language Action Model For GUI Agents
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 7383006c-253c-4889-a0b0-b6b62341d313 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents DeepShop: A Benchmark for Deep Research Shopping Agents
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 1004efd3-c3e3-4718-b65d-039422044496 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Inform: Mitigating reward hacking in rlhf via information-theoretic reward modeling.Advances in Neural Information Processing Systems, 37:134387–134429, 2024
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 171f012a-c40f-4943-8958-199f6e341e8d · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents WebCanvas: Benchmarking Web Agents in Online Environments
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 223f4fdb-b8b4-4a80-b10a-9b68f958701c · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Orchard: An Open-Source Agentic Modeling Framework
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 7aedf6dd-e6a6-40da-a114-d2169af94b97 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Webrl: Training llm web agents via self-evolving online curriculum reinforcement learning
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91c4051b-f433-483e-8fe9-23c279cdabf8 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents UI-TARS: Pioneering Automated GUI Interaction with Native Agents
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation ead8682a-9df4-416b-abf4-ca0693d3d6e5 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation ba77acba-516f-4347-93cd-c284777f47c7 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 4cdf6d52-2b97-4470-91ac-6eef6690b316 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents OpenAI GPT-5 System Card
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 5afc4c37-dc9e-4fd3-a0d5-8f3c9e695a88 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Kimi-VL Technical Report
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 262c0d22-9ae3-4eaf-a79f-acb7ffbdded2 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents InSTA: Towards Internet-Scale Training For Agents
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a813d14d-4c60-4e1e-8b10-d9d48d09dc68 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents UI-TARS-2 Technical Report: Advancing GUI Agent with Multi-Turn Reinforcement Learning
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f9bf9220-b665-486a-afa8-42907aa94c0f · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Vl-rethinker: Incentivizing self-reflection of vision-language models with reinforcement learning.Advances in Neural Information Processing Systems, 38:30865–30891, 2026
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25022457-6040-4bca-9e95-16451e02d2c0 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Vagen: Reinforcing world model reasoning for multi-turn vlm agents.Advances in Neural Information Processing Systems, 38:172871–172933, 2026
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0239183f-d0ff-41a0-b994-aab23e15d362 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents WebXSkill: Skill Learning for Autonomous Web Agents
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 30281e7c-4687-4ea5-a483-c266c1f14a32 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents RAGEN: Understanding Self-Evolution in LLM Agents via Multi-Turn Reinforcement Learning
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation e05f5330-39e5-4994-b363-a04c2f89e91b · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Webagent-r1: Training web agents via end-to-end multi-turn reinforcement learning
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 262006fb-c5ba-4f1d-ba15-bbf7ccc52454 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents GUI-Actor: Coordinate-Free Visual Grounding for GUI Agents
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 7bdd7622-76bf-485a-b8ee-b50e7857d5a4 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Os-atlas: Foundation action model for generalist gui agents
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 045f0561-ef05-4383-b820-5ea605ff435f · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Aguvis: Unified Pure Vision Agents for Autonomous GUI Interaction
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 395a6145-0ae6-41b8-92cf-52a7c1c11030 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents An illusion of progress? assessing the current state of web agents
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5dbee843-c802-4ab4-9ade-b1966d29cae5 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Magma: A foundation model for multimodal ai agents
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f343030-6a3f-4511-a06e-38f37c337a2d · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Agentoccam: A simple yet strong baseline for llm-based web agents
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4c219b1-1b7e-43f8-8449-93a2d4a6f311 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 77bb24c1-5f61-48b7-9fd6-f36fbd2c4bce · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Regularizing hidden states enables learning generalizable reward model for llms.Advances in Neural Information Processing Systems, 37:62279–62309, 2024
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 721830e8-399d-43cd-ac98-29ab900578a2 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents GUI-Libra: Training native GUI agents to reason and act with action-aware supervision and partially verifiable RL, 2026
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 144a5e71-d1b0-4778-a7f3-9e7c715f3039 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents ReAct: Synergizing reasoning and acting in language models
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1999d737-fa6e-4c4e-8131-db2c737aa7ec · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents How do visual attributes influence web agents? a comprehensive evaluation of user interface design factors.arXiv preprint arXiv:2601.21961, 2026
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c9a3d777-1c56-423e-9faa-d8a393413545 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents DAPO: An Open-Source LLM Reinforcement Learning System at Scale
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 86abf6e1-d644-4487-bc7a-d7d169ca0424 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Fine-tuning large vision-language models as decision-making agents via reinforcement learning, 2024
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f68a811-a732-41e8-bbe1-f60d883382ad · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Beat: Visual backdoor attacks on vlm-based embodied agents via contrastive trigger learning, 2026
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41386195-9587-4283-87d4-fa3bd473a13d · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents AgentRL: Scaling agentic reinforcement learning with a multi-turn, multi-task framework, 2025
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 097481d2-63c9-4c0e-9c76-4b92003ea4a7 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents LlamaFactory: Unified efficient fine-tuning of 100+ language models
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97d1fb3d-bc28-4b9f-bc4f-35339278784d · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Deepresearcher: Scaling deep research via reinforcement learning in real-world environments
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5597776e-93f1-49ca-8b20-550bdda05c6c · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Proposer-agent-evaluator (PAE): Autonomous skill discovery for foundation model internet agents
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2cfaf7ae-1f8c-4241-914a-80b01801d842 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Workforceagent-r1: Incentivizing reasoning capability in llm-based web agents via reinforcement learning
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03c70b78-9ba4-4902-9a4d-679cdbe79ddf · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Alpine Ridge
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 3ab2f7fc-a598-4a61-9fcd-00b76f1d7a0a · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents name": ...,
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea5985b0-4562-4d13-8cf6-60e5c7c07c82 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Unresolved cited work
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f39a3e6f-bd2f-4a7a-a9be-a410e08f89fc · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents Use it to understand what the agent tried to do, but do not treat it as ground truth if it conflicts with the screenshots
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 780bd1dd-e6ba-4857-932f-4a5bf3017257 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents point 2d
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 89e4c9a6-38fa-49c0-865c-ac2224c9e55c · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents SHOP MEN'S
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32b51fec-4d25-440b-981b-527bb85dd4c0 · outbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents name": "hover
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11a59d82-1620-4a73-aa12-e8302c4c0355 · inbound
SeekJudge: A Practical Reward Framework for Reinforcement Learning in Computer-Use Agents OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0217b2c-d4d3-4b49-82a3-c2b19124985e · inbound
RMSWeb: Reflection, Failure-Mode Mining, and Salvage-DS for Web Agent Reinforcement Learning OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.