Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:01:25.268734Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 1 inbound Pith citation observation for arXiv:2505.22942.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:01:25.268734Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-16T09:33:30.444057Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-16T09:37:41.944603Z
53 of 53 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation d6d3f758-f93f-4b9f-b296-2cfa105bc81d · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning online" 'onlinestring :=
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 555dd279-badd-41eb-8dc7-0608c3a6e1a1 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning write newline
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c05b1af-4175-41d1-8cdf-042686b4f6e9 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Nemotron-4 340B Technical Report
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24c6ddbe-1b80-431a-a7c6-a1457d379561 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning SeeClick: Harnessing GUI Grounding for Advanced Visual GUI Agents
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 073aff89-ce94-4cf9-a649-f8929e728b63 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning The BrowserGym Ecosystem for Web Agent Research
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d04ee0b-94c0-4b75-9ba7-8a62058bd972 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4f3f1d3d-9b9b-4994-9aa0-26ef3b260ce3 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Laradji, Manuel Del Verme, Tom Marty, David Vazquez, Nicolas Chapados, and Alexandre Lacoste
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0339ea42-1057-4d10-864d-f1d8e9ec2727 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning The Llama 3 Herd of Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d477bd3-3331-4e95-9d3d-2fbb01500ea4 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 24b2fc55-e79a-4c59-9716-f6d51fe25a70 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83b894f4-da58-41bd-b9fd-b69d32ac0f09 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68a133eb-d901-42e3-9210-9afdc67ce733 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning A Real-World WebAgent with Planning, Long Context Understanding, and Program Synthesis
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b17aeedb-adc1-4c61-b128-beb80f898367 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning WebVoyager: Building an End-to-End Web Agent with Large Multimodal Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60e58c5d-6d0d-4c86-bada-cbe80b64a804 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning OpenWebVoyager: Building Multimodal Web Agents via Iterative Real-World Exploration, Feedback and Optimization
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 337f8532-e1f3-4e5b-86e5-2c897efd69c0 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning GPT-4o System Card
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 233fcc62-c834-4818-b132-b3642fa84a08 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning VideoWebArena: Evaluating Long Context Multimodal Agents with Video Understanding Web Tasks
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18091e6e-8d27-4e9c-8127-a67f3eceb49c · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bddcc80b-d7c3-4d49-969a-4b5a4bd78c04 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ee25ebe5-b132-47da-b113-31e06bc30f82 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning ST-WebAgentBench: A Benchmark for Evaluating Safety and Trustworthiness in Web Agents
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6658e12c-bddd-42d5-8bbc-633daa5d0b52 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c90a064-4a20-49b5-83b6-be24acdb7b0b · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning ReflecTool: Towards Reflection-Aware Tool-Augmented Clinical Agents
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5e43736-43ad-4bde-a342-009c3178061d · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning AgentBench: Evaluating LLMs as Agents
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0fa1ecbc-8652-4784-99a3-ed2dbcd2078e · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning LASER: LLM Agent with State-Space Exploration for Web Navigation
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67007de0-b117-4a67-a764-1da794453828 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b518536-05fb-4fb4-aae4-9fa037557982 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning s1: Simple test-time scaling
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83a9e7f2-1d7f-4b33-96f3-481373a19201 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning WebGPT: Browser-assisted question-answering with human feedback
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4825587-3a7f-4bc0-976e-06fd526f37c9 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 12d7c553-52f3-46d0-81d4-e8ea9dafa983 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 945135ee-9b28-4163-86b4-0b92a3c8fe41 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0d76ac06-0f6b-468c-9f6c-137e27ee6005 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Autonomous Evaluation and Refinement of Digital Agents
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63dfd185-4092-4c15-997a-74ee35241cf1 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning WebCanvas: Benchmarking Web Agents in Online Environments
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da49da50-7f77-48e0-8e68-bb3df4780cf2 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation dbc72f20-ca55-4d15-af67-e673b2f1ff4f · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning ToolRL: Reward is All Tool Learning Needs
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72453d07-1e30-4f0e-8792-46d2b5fe212c · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c73efc34-93f3-4204-b6af-ed3bbdee65c8 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46d2388d-2d3d-4ce4-bd7b-cc6325a520cb · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning HybridFlow: A Flexible and Efficient RLHF Framework
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02dc13a4-91d7-46d5-aa79-4997c9059013 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6cf7d8cc-a877-4e04-95a5-beb079905529 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f9d5973-3f37-4027-99f3-391fb6525e4c · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3afd9b5f-3988-417f-bb44-1cd1a5b200a8 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8e3605c6-8824-497f-b3e2-91711b2ce922 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1973b9e9-7deb-4ffe-9ccf-5fb05d9dcecd · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning RAGEN: Understanding Self-Evolution in LLM Agents via Multi-Turn Reinforcement Learning
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47af6178-9d8f-4070-8a8d-64b8fe54ba03 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning TheAgentCompany: Benchmarking LLM Agents on Consequential Real World Tasks
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 148e3680-84c5-44c1-aba5-a5839a598083 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Qwen2.5 Technical Report
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70230196-de45-4987-ba2e-0feb2d3e5680 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cbafc01e-53ce-4ff1-b7c3-a0d2912c1ecd · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c435f826-8d81-4429-85f1-58a81941880e · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76fe97b0-2dcd-4329-a6e2-899f652b19cf · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning AssistantBench: Can Web Agents Solve Realistic and Time-Consuming Tasks?
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9d98500-5871-425b-8269-0adb0b2176f0 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning GPT-4V(ision) is a Generalist Web Agent, if Grounded
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1f95386-dd9e-4eb1-a2a0-bc0118ab9296 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning WebArena: A Realistic Web Environment for Building Autonomous Agents
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7b712e4-e754-4120-836e-84c908bc680b · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Rossi, Somdeb Sarkhel, and Chao Zhang
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 65f17232-ffc7-4b7c-8295-3501c40d7130 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 99e6c23d-c187-416f-b266-476e79738164 · outbound
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Hephaestus: Improving Fundamental Agent Capabilities of Large Language Models through Continual Pre-Training
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c5f77f0-f1c3-468c-9463-1310ddc94d86 · inbound
DynaWeb: Model-Based Reinforcement Learning of Web Agents WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.