Pith. sign in

Paper Citation Record · LEDGER

WebDancer: Towards Autonomous Information Seeking Agency

As of 7 August 2026, this Paper Citation Record lists 73 of 73 outbound references and 41 inbound Pith citation observations for arXiv:2505.22648.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.22648 v3

Coverage vector

measured 73 of 73 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:09:05.825645Z

measured 114 of 114 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 41 of 41 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:26:56.960098Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

73 of 73 outbound references displayed

  • verified exact0
  • verified fuzzy20
  • unresolved53
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

2
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation f8d568ce-ae46-4c0a-aed5-e66671f677bb · outbound

This paper cites write newline.

WebDancer: Towards Autonomous Information Seeking Agency write newline

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:12.981441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:08:55.870617Z digest=sha256:da60db9940164ee24d329ec29cf8cb74de2873f85f2d4e7051c0ab5027820a83

Observation 382d04ec-3381-4951-87e8-320abfee75a2 · outbound

This paper cites Deep research system card, 2025 a.

WebDancer: Towards Autonomous Information Seeking Agency Deep research system card, 2025 a

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:12.655889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:08:55.997151Z digest=sha256:24a5e044a401a24bd9e2b6fcc26b198e63fda4b336f20db1f069362d96e8d909

Observation 37fe22e5-f4f5-4f78-9765-340954078ce1 · outbound

This paper cites Grok 3 beta — the age of reasoning agents, 2025.

WebDancer: Towards Autonomous Information Seeking Agency Grok 3 beta — the age of reasoning agents, 2025

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:12.454202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:08:56.143758Z digest=sha256:9a1d5e4584fb3c8bd44a27b3c78b89df72b04689f3a7e5aa636d2113f948c43d

Observation 9a078472-a785-4ce2-8e51-1ce3ae9a89af · outbound

This paper cites WebWalker: Benchmarking LLMs in Web Traversal.

WebDancer: Towards Autonomous Information Seeking Agency WebWalker: Benchmarking LLMs in Web Traversal

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:56.317112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:08:56.317112Z digest=sha256:119ad2a7b6cb62156cf0026e33d56e06cf606576bbf590a5d913b1ee1a17ab9d

Observation 9915ce69-54d7-4c96-9b37-27685d8b43c2 · outbound

This paper cites Webthinker: Empowering large reasoning models with deep research capability, 2025 a.

WebDancer: Towards Autonomous Information Seeking Agency Webthinker: Empowering large reasoning models with deep research capability, 2025 a

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:12.162675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:08:56.462849Z digest=sha256:8d693bc6f7614e8f9e4da8ea6397e8108d4b723d29ce78160950b106929b4921

Observation ff2bac49-de2a-4f32-8cf9-5ebc19b35708 · outbound

This paper cites Search-o1: Agentic Search-Enhanced Large Reasoning Models.

WebDancer: Towards Autonomous Information Seeking Agency Search-o1: Agentic Search-Enhanced Large Reasoning Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:56.595252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:08:56.595252Z digest=sha256:7b9df1c229369b576088469ff0a2eca914399611e4047bb0219e03f1ea48d1bf

Observation e367822e-dc8d-4eee-9ebd-df55e66d0c9a · outbound

This paper cites ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning.

WebDancer: Towards Autonomous Information Seeking Agency ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:56.703085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:08:56.703085Z digest=sha256:c0964c91a1e658a2a0272bf7617078017b40ab641f1ab6c7c82f776f45eba8a3

Observation 98ee34ba-55e2-4fb3-837c-2eda5145e58f · outbound

This paper cites R1-Searcher: Incentivizing the Search Capability in LLMs via Reinforcement Learning.

WebDancer: Towards Autonomous Information Seeking Agency R1-Searcher: Incentivizing the Search Capability in LLMs via Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:56.844666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:08:56.844666Z digest=sha256:3db57675225e5dc6db2f4661a3b8d6fdf3ccccc5bc042cf166a911a8a89b6c44

Observation b056d949-544a-4fce-a1c3-f6bc1f39c29e · outbound

This paper cites Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning.

WebDancer: Towards Autonomous Information Seeking Agency Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:56.968758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:08:56.968758Z digest=sha256:6a56b2606614c85fe73dd37513a74d48f085831bac909c7520a6ae0061b4468f

Observation 50258864-6235-44e6-8af3-0c6108a023b5 · outbound

This paper cites Simpledeepsearcher: Deep information seeking via web-powered reasoning trajectory synthesis.

WebDancer: Towards Autonomous Information Seeking Agency Simpledeepsearcher: Deep information seeking via web-powered reasoning trajectory synthesis

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:11.920975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:08:57.093646Z digest=sha256:016a36259bd92e14e332cb68134ec7599a233df589a3564dcc3bf528cf613ae7

Observation ff022502-11f1-4f3d-b46b-120bf26a6d41 · outbound

This paper cites DeepResearcher: Scaling Deep Research via Reinforcement Learning in Real-world Environments.

WebDancer: Towards Autonomous Information Seeking Agency DeepResearcher: Scaling Deep Research via Reinforcement Learning in Real-world Environments

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:57.252692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:08:57.252692Z digest=sha256:47bdd51b7f0159e3879021f0aaa76db799fac5d0f0b20f5d20ca4a3e9f73f3e8

Observation 390db571-2ad7-4dba-92de-d2c624053e6e · outbound

This paper cites React: Synergizing reasoning and acting in language models.

WebDancer: Towards Autonomous Information Seeking Agency React: Synergizing reasoning and acting in language models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:57.406564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:08:57.406564Z digest=sha256:4bddccf4f67b9fe4bc070eba16089a111fc81b123414f78bc2903211a0c51bcc

Observation 62aa6e68-5b27-4220-9eb3-07fcd164048e · outbound

This paper cites Gaia: a benchmark for general ai assistants.

WebDancer: Towards Autonomous Information Seeking Agency Gaia: a benchmark for general ai assistants

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:57.507064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:08:57.507064Z digest=sha256:ba4901d682004a9ddc794cdfcaa52262c89a108177f35591cfee5d148c81512b

Observation 3effb8c8-dcb4-4586-897e-25ba119dd69f · outbound

This paper cites Browsecomp: A simple yet challenging benchmark for browsing agents.

WebDancer: Towards Autonomous Information Seeking Agency Browsecomp: A simple yet challenging benchmark for browsing agents

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:11.604810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:08:57.596483Z digest=sha256:07e4b3d7f51345d89896e3d0f4a7151235abc2a4b45c765e887014f775ca2c9e

Observation a000de90-8595-4cf3-86c4-e1e11cf422e2 · outbound

This paper cites SynWorld: Virtual Scenario Synthesis for Agentic Action Knowledge Refinement.

WebDancer: Towards Autonomous Information Seeking Agency SynWorld: Virtual Scenario Synthesis for Agentic Action Knowledge Refinement

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:57.735461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:08:57.735461Z digest=sha256:8410390122bdfad489b1d1a4039a93c7e9d56e063d363b3f79240186c09ef370

Observation 3fe53b86-e25a-46c0-b98a-077ef4f289a5 · outbound

This paper cites TTRL: Test-Time Reinforcement Learning.

WebDancer: Towards Autonomous Information Seeking Agency TTRL: Test-Time Reinforcement Learning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:57.869227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:08:57.869227Z digest=sha256:33759d5137abbc52390396888affee0ae47736149f10521ee11d57e34510c77e

Observation 820e3921-aa3e-4902-b365-a3e0e076264f · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

WebDancer: Towards Autonomous Information Seeking Agency DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:58.025183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:08:58.025183Z digest=sha256:f29fbbd74f7e2358b8b8cd38edaf17a82579da39d5026f0bc814b31ef96979aa

Observation 33bf8a31-5ff5-4f51-9060-55b8ed4f19f8 · outbound

This paper cites Mintaka: A complex, natural, and multilingual dataset for end-to-end question answering.

WebDancer: Towards Autonomous Information Seeking Agency Mintaka: A complex, natural, and multilingual dataset for end-to-end question answering

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:11.290100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:08:58.164101Z digest=sha256:d26c87bd342caa9a3ea279e5ee56b2edf465ab8c11e8bf3a2ca468a016e77163

Observation ec90153f-738e-4ec9-8c15-9841d6b18fa4 · outbound

This paper cites Language models are few-shot learners.

WebDancer: Towards Autonomous Information Seeking Agency Language models are few-shot learners

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:58.280810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:08:58.280810Z digest=sha256:78b460b177ed23f2b42f87121600d356f45334313d71f0fbc49bb6a9c92c5059

Observation ed1ae12a-c7ab-41a4-b55b-9423f10e836b · outbound

This paper cites BrowseComp-ZH: Benchmarking Web Browsing Ability of Large Language Models in Chinese.

WebDancer: Towards Autonomous Information Seeking Agency BrowseComp-ZH: Benchmarking Web Browsing Ability of Large Language Models in Chinese

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:58.430964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:08:58.430964Z digest=sha256:6830fb623bbe878accb92cedb44429c37a611d2c9eb0d97cdcb8818bb3eaf6bf

Observation 7dfd91b4-a2b4-4671-908d-77be1033f587 · outbound

This paper cites Introducing simpleqa, 2025 b.

WebDancer: Towards Autonomous Information Seeking Agency Introducing simpleqa, 2025 b

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:11.051803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:08:58.524027Z digest=sha256:998fe364a6b0d6ac0ba92c30a5e4f891539b6ccf093356da2b509d78aa62872d

Observation b07e172d-ff0a-4afd-aa50-b57fd317ed6d · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

WebDancer: Towards Autonomous Information Seeking Agency Chain-of-thought prompting elicits reasoning in large language models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:58.644916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:08:58.644916Z digest=sha256:265b86f13d1685c029a91b8eef39213277937879c87b61d33f4010e958298e6d

Observation 34004d64-1731-4d15-8987-9d9d92763634 · outbound

This paper cites Agent models: Internalizing Chain-of-Action Generation into Reasoning models.

WebDancer: Towards Autonomous Information Seeking Agency Agent models: Internalizing Chain-of-Action Generation into Reasoning models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:58.804825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:08:58.804825Z digest=sha256:56109459c13ca06b674dcc4a9d0494f27479f4dbda5c7f4d80091c5bd374a439

Observation 7387e269-f360-4561-a7fc-aeedf26504ac · outbound

This paper cites Agent rl scaling law: Agent rl with spontaneous code execution for mathematical problem solving, 2025.

WebDancer: Towards Autonomous Information Seeking Agency Agent rl scaling law: Agent rl with spontaneous code execution for mathematical problem solving, 2025

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:10.724692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:08:58.923935Z digest=sha256:78b5711d06db247c0b5412d7f01029a9e510faf04406020d4d31f194d25037fa

Observation 1cd6c6b2-2f42-45dd-b4b8-418b0e0895dd · outbound

This paper cites Qwq-32b: Embracing the power of reinforcement learning, 2025 b.

WebDancer: Towards Autonomous Information Seeking Agency Qwq-32b: Embracing the power of reinforcement learning, 2025 b

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:10.527248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:08:59.080172Z digest=sha256:36f70b27d23f25a37299eaae24204ac3d68cf13ed46dff173b023cb844d29a21

Observation f2488c8d-a296-4f29-9554-1f41881f20e9 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

WebDancer: Towards Autonomous Information Seeking Agency DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:59.222336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:08:59.222336Z digest=sha256:9dfa637cabfb2a90b289e4e852d0d260f768db3ec4dd2e39d0ffcbea8432ee53

Observation 3de72f0a-1b06-41df-99f4-0dd164f7c077 · outbound

This paper cites A Comparative Study on Reasoning Patterns of OpenAI's o1 Model.

WebDancer: Towards Autonomous Information Seeking Agency A Comparative Study on Reasoning Patterns of OpenAI's o1 Model

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:59.356861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:08:59.356861Z digest=sha256:9c7b52a37be967d9633b949ec2cc9486b1de0473475d9ee6ace024c34a4622e3

Observation f0489ece-1f19-4a99-9546-1abd63dc5cb5 · outbound

This paper cites Humanity's Last Exam.

WebDancer: Towards Autonomous Information Seeking Agency Humanity's Last Exam

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:59.530903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:08:59.530903Z digest=sha256:7bdacfa3fcff262ee7b74a254c021b1aeba35700f95ce314c02b37ecd94cddc4

Observation 05a885bf-085a-4741-89e4-742cb11aae9e · outbound

This paper cites FireAct: Toward Language Agent Fine-tuning.

WebDancer: Towards Autonomous Information Seeking Agency FireAct: Toward Language Agent Fine-tuning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:59.724999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:08:59.724999Z digest=sha256:fd6b06e26cf6f71ceb7392afa4888d7d179176ffd21bd50684f46cdf761bb0c4

Observation 426d76a5-1199-4889-a3fb-0e9c2a936848 · outbound

This paper cites 100 Days After DeepSeek-R1: A Survey on Replication Studies and More Directions for Reasoning Language Models.

WebDancer: Towards Autonomous Information Seeking Agency 100 Days After DeepSeek-R1: A Survey on Replication Studies and More Directions for Reasoning Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:59.831868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:08:59.831868Z digest=sha256:719080ea5ab9569b4ac5d1994d8f83935c64dfa27c622f82f03cc1936352a2ed

Observation b1404a08-550c-4bef-88ba-dc02c5b91478 · outbound

This paper cites LLM Post-Training: A Deep Dive into Reasoning Large Language Models.

WebDancer: Towards Autonomous Information Seeking Agency LLM Post-Training: A Deep Dive into Reasoning Large Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:00.013818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:00.013818Z digest=sha256:0908dd5232228c634c1b6b981e67472c2872c4920224c2c749a8451a8f124e1b

Observation c2f54ddd-04a7-44a1-acca-2e0b343ffaa0 · outbound

This paper cites Reasoning beyond limits: Advances and open problems for llms.

WebDancer: Towards Autonomous Information Seeking Agency Reasoning beyond limits: Advances and open problems for llms

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:00.163556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:00.163556Z digest=sha256:f33bd422d6737376f3c331c30c7c718077f695648f93d08a2beb0b34bcd5b8be

Observation db12f1b6-a332-4c1f-ab4a-6aad545cd67b · outbound

This paper cites Judging llm-as-a-judge with mt-bench and chatbot arena.

WebDancer: Towards Autonomous Information Seeking Agency Judging llm-as-a-judge with mt-bench and chatbot arena

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:00.298633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:00.298633Z digest=sha256:9865b60959efc9542722463679298a63a9327e3a21af29df17f319f39f76e9ae

Observation 9befb21f-e1f7-4b72-b452-5f9c2a0ffa13 · outbound

This paper cites Seed-thinking-v1.

WebDancer: Towards Autonomous Information Seeking Agency Seed-thinking-v1

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:00.450391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:00.450391Z digest=sha256:05e15aff06a3b697e5e4de8422abc8bf347c0bc7ca5fc31ea4aea17d6c024bb9

Observation 8e894e89-aaa8-40d4-a55e-a065552fb817 · outbound

This paper cites A Unified Pairwise Framework for RLHF: Bridging Generative Reward Modeling and Policy Optimization.

WebDancer: Towards Autonomous Information Seeking Agency A Unified Pairwise Framework for RLHF: Bridging Generative Reward Modeling and Policy Optimization

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:00.584149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:00.584149Z digest=sha256:771fa4d02dda80fe1d885ca47bbf6a9fc3f95ffd3b93a82ac7145780079f525d

Observation 683ce7b7-1ec2-4dcf-9f87-7f1c31841609 · outbound

This paper cites Inference-time scaling for generalist reward modeling.

WebDancer: Towards Autonomous Information Seeking Agency Inference-time scaling for generalist reward modeling

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:00.744891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:00.744891Z digest=sha256:5368ea47463380db363ae91a6a8e5db98056d2f66082ede4d594b30fc376de62

Observation f1d4bc83-aa1d-4ef2-92da-3a6765b83014 · outbound

This paper cites Does Reinforcement Learning Really Incentivize Reasoning Capacity in LLMs Beyond the Base Model?.

WebDancer: Towards Autonomous Information Seeking Agency Does Reinforcement Learning Really Incentivize Reasoning Capacity in LLMs Beyond the Base Model?

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:00.873822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:00.873822Z digest=sha256:5c0d3a5d767769abb7e3c95f683eadec482ccda9d59220763fba8b9883d88689

Observation 8d1d3468-3a01-46f2-ab4b-598ff3905225 · outbound

This paper cites Andrew Bagnell.

WebDancer: Towards Autonomous Information Seeking Agency Andrew Bagnell

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:01.024864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:01.024864Z digest=sha256:8a933a0e9583ba75a90af06f47c13039bd27c7de5ad320e7fdc3c09b66d8352e

Observation 29d3063f-7785-4bad-bc40-d4c2b208edc6 · outbound

This paper cites Group-in-Group Policy Optimization for LLM Agent Training.

WebDancer: Towards Autonomous Information Seeking Agency Group-in-Group Policy Optimization for LLM Agent Training

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:01.160586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:01.160586Z digest=sha256:f4cb02357a9e29000101f58ad39577b7d17db782899efe0a31c0cce65d056f9d

Observation 09598273-5280-42df-a3df-46f108c56d45 · outbound

This paper cites Webagent-r1: Training web agents via end-to-end multi-turn reinforcement learning, 2025 a.

WebDancer: Towards Autonomous Information Seeking Agency Webagent-r1: Training web agents via end-to-end multi-turn reinforcement learning, 2025 a

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:01.327140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:01.327140Z digest=sha256:9670d4edf12c77b7be00f9a9bcbbbffcad7a5488db01f91c8a89e1ce70dc70ec

Observation fb7ec172-cb7b-4de7-b7d7-f1ff79948c2d · outbound

This paper cites ToRA: A Tool-Integrated Reasoning Agent for Mathematical Problem Solving.

WebDancer: Towards Autonomous Information Seeking Agency ToRA: A Tool-Integrated Reasoning Agent for Mathematical Problem Solving

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:01.456783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:01.456783Z digest=sha256:8934ec46b5de882c0862b47e4e933c4c12bdcfa7223dd8888c9077ff8764de99

Observation b223d237-71f5-46d8-9604-c6f325bd517f · outbound

This paper cites Small models struggle to learn from strong reasoners.

WebDancer: Towards Autonomous Information Seeking Agency Small models struggle to learn from strong reasoners

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:01.570298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:01.570298Z digest=sha256:951a0e553ae8efb42b63f525046ebea5f5c3578639677050d03c57be225f9512

Observation 9c394266-cf96-4a0d-b97c-0b3f09c76de3 · outbound

This paper cites Towards widening the distillation bottleneck for reasoning models, 2025.

WebDancer: Towards Autonomous Information Seeking Agency Towards widening the distillation bottleneck for reasoning models, 2025

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:10.267151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:01.692105Z digest=sha256:15ad2a9ab63cfdbc315b59e52ea960d66e86a831cf0ad5ec8adc6d09c203aafe

Observation cfea8b20-49ac-4482-9fa8-f0a24e71a682 · outbound

This paper cites Omnithink: Expanding knowledge boundaries in machine writing through thinking.

WebDancer: Towards Autonomous Information Seeking Agency Omnithink: Expanding knowledge boundaries in machine writing through thinking

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:01.840994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:01.840994Z digest=sha256:60635b8965cdc222cbcbe02429bb60cac21f8abec41a214927fd965d7bb61ac3

Observation d5922067-e435-4076-8065-c1f09325d1ee · outbound

This paper cites Advances and Challenges in Foundation Agents: From Brain-Inspired Intelligence to Evolutionary, Collaborative, and Safe Systems.

WebDancer: Towards Autonomous Information Seeking Agency Advances and Challenges in Foundation Agents: From Brain-Inspired Intelligence to Evolutionary, Collaborative, and Safe Systems

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:01.959758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:01.959758Z digest=sha256:1c6bdd9b31d62bb71a80aec3e1a511c1b630257f3c51f50eab9c6ae99c675bab

Observation 36d9c9fb-a405-4c4c-acf1-63ea8512166f · outbound

This paper cites Symbolic Learning Enables Self-Evolving Agents.

WebDancer: Towards Autonomous Information Seeking Agency Symbolic Learning Enables Self-Evolving Agents

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:02.084175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:02.084175Z digest=sha256:9cfee45d4e13715afa939f411a7a07a2e0041e5ce6f56df56e347d00822cf2db

Observation b5845360-9ec6-41c0-b41d-5d4a77b4bd63 · outbound

This paper cites Agents: An Open-source Framework for Autonomous Language Agents.

WebDancer: Towards Autonomous Information Seeking Agency Agents: An Open-source Framework for Autonomous Language Agents

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:02.260119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:02.260119Z digest=sha256:a8ea8a66a3de2a7d284e28350fb3b435731a16a6bd5d25452219e9fd57138ea0

Observation 86d78663-1a76-4cbe-ac17-628076eab81d · outbound

This paper cites Autoact: Automatic agent learning from scratch for qa via self-planning.

WebDancer: Towards Autonomous Information Seeking Agency Autoact: Automatic agent learning from scratch for qa via self-planning

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:09.964042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:02.384475Z digest=sha256:ae735570d41256e214293c211f63755e3c18952c0212c45a0f9a0e58f7f59fd7

Observation 7dcd458e-1732-42ea-a34e-4465467074d6 · outbound

This paper cites Agenttuning: Enabling generalized agent abilities for llms.

WebDancer: Towards Autonomous Information Seeking Agency Agenttuning: Enabling generalized agent abilities for llms

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:09.661536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:02.546363Z digest=sha256:50b115910809b373b2d1c3875eefed39a00de22824bc4c7462da0664e82ca8b9

Observation 7b6a82c6-ebb2-4d94-a76b-63f352695a06 · outbound

This paper cites Agent-flan: Designing data and methods of effective agent tuning for large language models.

WebDancer: Towards Autonomous Information Seeking Agency Agent-flan: Designing data and methods of effective agent tuning for large language models

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:09.395912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:02.700803Z digest=sha256:ee048d174408e7c28bd56a5e921452febac614497e7b23dd8c3e2d7c05a57e1c

Observation 6fc74151-5eff-4081-a180-57d2aad2b823 · outbound

This paper cites Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning.

WebDancer: Towards Autonomous Information Seeking Agency Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:02.846595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:02.846595Z digest=sha256:60deaf0b72e969a95815ce5b872799d21525a649fcb453614ad18bf0f31f92b4

Observation ae2efee1-4910-42b4-8ed5-d4c646fb20da · outbound

This paper cites ToolRL: Reward is All Tool Learning Needs.

WebDancer: Towards Autonomous Information Seeking Agency ToolRL: Reward is All Tool Learning Needs

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:02.987452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:02.987452Z digest=sha256:fabcba3469f683edf3e7b9b215d4971b98af9a56495d64bfc54d901f271b14b2

Observation 488f43eb-1c00-40da-812c-07e0c836561d · outbound

This paper cites StepTool: Enhancing Multi-Step Tool Usage in LLMs via Step-Grained Reinforcement Learning.

WebDancer: Towards Autonomous Information Seeking Agency StepTool: Enhancing Multi-Step Tool Usage in LLMs via Step-Grained Reinforcement Learning

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:03.114131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:03.114131Z digest=sha256:843766d707d864bae7b41681b12a693fb63c3db5801a409668c1bf18120a63df

Observation 4ca596bf-03d8-4243-a056-c1d6e952c666 · outbound

This paper cites an unresolved cited work.

WebDancer: Towards Autonomous Information Seeking Agency Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:09:09.085566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:03.244034Z digest=sha256:3d763494dc551458f8edafbec5f0ac42d11acdb130905636c695d100593f7818

Observation f75f80cf-6d2d-4f43-9e8b-77dd7c609ce3 · outbound

This paper cites Gonzalez, and Ion Stoica.

WebDancer: Towards Autonomous Information Seeking Agency Gonzalez, and Ion Stoica

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:03.380818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:03.380818Z digest=sha256:7e0907144076a8aa02a8e0b3627ffea00e6339502bfc6509449c533164e946cd

Observation 7b8a5994-d3b0-4637-b79b-efdd04762b0b · outbound

This paper cites API Agents vs. GUI Agents: Divergence and Convergence.

WebDancer: Towards Autonomous Information Seeking Agency API Agents vs. GUI Agents: Divergence and Convergence

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:03.542075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:03.542075Z digest=sha256:a766753042dc1e4b58cf0717a923048c9370612a74b72014c4f0b850d5d9f506

Observation 9fe3b9e6-96ed-4f07-827e-87dc5219e628 · outbound

This paper cites Assisting in writing wikipedia-like articles from scratch with large language models.

WebDancer: Towards Autonomous Information Seeking Agency Assisting in writing wikipedia-like articles from scratch with large language models

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:08.776098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:03.626317Z digest=sha256:59db79fd888b9305fb1f3ac737aa959421f5570a2453b6e52cdc6bf09219f39e

Observation e657de81-cdc6-4e85-b9a9-744b767f0e82 · outbound

This paper cites Qwen3 Technical Report.

WebDancer: Towards Autonomous Information Seeking Agency Qwen3 Technical Report

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:03.733534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:03.733534Z digest=sha256:27521b4c8f70348f69a04de5f73df7ca1ead70ba3eedb537c3a74aff6fc415c7

Observation a3dce6d3-837e-4e92-895e-f72c5fee07e4 · outbound

This paper cites Direct preference optimization: Your language model is secretly a reward model.

WebDancer: Towards Autonomous Information Seeking Agency Direct preference optimization: Your language model is secretly a reward model

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:03.871535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:03.871535Z digest=sha256:36fc269c3a5296be1da16c01a6fed280d91f096dd7de04341c24744e7ba5b2bb

Observation d65c987b-4266-4fa4-b131-8a0f52d078fe · outbound

This paper cites Owl: Optimized workforce learning for general multi-agent assistance in real-world task automation, 2025.

WebDancer: Towards Autonomous Information Seeking Agency Owl: Optimized workforce learning for general multi-agent assistance in real-world task automation, 2025

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:08.495719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:04.008095Z digest=sha256:747ea9392e935e121f88ffad8a52b3fbaed812329184b9f97fea7adbaec2da8b

Observation 41447d7c-6d75-4563-b41e-ee3fe3490545 · outbound

This paper cites Camel: Communicative agents for "mind" exploration of large language model society.

WebDancer: Towards Autonomous Information Seeking Agency Camel: Communicative agents for "mind" exploration of large language model society

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:04.126923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:04.126923Z digest=sha256:728fe83941a1bd86c00e170a2c797e90be80d4ebd9a9eaa3530b1ae70a3798df

Observation 57949fce-af3c-4042-8191-881772021c7e · outbound

This paper cites Openmanus: An open-source framework for building general ai agents, 2025.

WebDancer: Towards Autonomous Information Seeking Agency Openmanus: An open-source framework for building general ai agents, 2025

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:04.211048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:04.211048Z digest=sha256:d33d8adc6ca5b6ee89a1d3262deefbd5805a98ce0d7ac34510eda15f1ca51276

Observation 72e23d31-bb1c-4ee0-a77b-155ff4e9f133 · outbound

This paper cites Meet claude, 2025.

WebDancer: Towards Autonomous Information Seeking Agency Meet claude, 2025

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:08.236609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:04.375426Z digest=sha256:e6075204ff01d289ea656ab8b1873718eb24630a8d2c5d5127e7066552c8fe6d

Observation 22873573-f1d4-4fda-98fc-42acf0b86c05 · outbound

This paper cites MuSiQue: Multihop Questions via Single-hop Question Composition.

WebDancer: Towards Autonomous Information Seeking Agency MuSiQue: Multihop Questions via Single-hop Question Composition

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:04.496311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:04.496311Z digest=sha256:1beca7790f46956d349f1adb928e908dc6d155e1ff588ce1b113b55aa22ca618

Observation b64a7ff7-47dd-4fad-ad69-b541b6b37901 · outbound

This paper cites Smith, and Mike Lewis.

WebDancer: Towards Autonomous Information Seeking Agency Smith, and Mike Lewis

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:07.929847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:04.645992Z digest=sha256:723babed1f3784108ebc218365b8a2335afae2d9b341a1ded8a2365b1122b081

Observation 7a005871-ded2-4c7e-9bbc-5f927f7bd5f3 · outbound

This paper cites When not to trust language models: Investigating effectiveness of parametric and non-parametric memories, 2022.

WebDancer: Towards Autonomous Information Seeking Agency When not to trust language models: Investigating effectiveness of parametric and non-parametric memories, 2022

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:07.688637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:04.772045Z digest=sha256:e0039adc3f4a9b441aba393c24e77162e94d0c12909ce6b70afde57f2d137db5

Observation 34e9275a-2a5e-4599-a125-4aa3579187e3 · outbound

This paper cites Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps.

WebDancer: Towards Autonomous Information Seeking Agency Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:04.990968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:04.990968Z digest=sha256:a7e50cb1742c89ca0a44d7182e6bb2a8c1e928d88fde42c315c7fe7aa60841ec

Observation 907afced-e943-4859-8c16-27c3a6c77250 · outbound

This paper cites HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering.

WebDancer: Towards Autonomous Information Seeking Agency HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:05.113581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:05.113581Z digest=sha256:a2089845be5c9468cec83e022d86ae267d40d04bc397b22040b8122500efb78d

Observation 8199fc44-0643-4a1c-a0e0-4447c3abf24f · outbound

This paper cites Qwen2.5 Technical Report.

WebDancer: Towards Autonomous Information Seeking Agency Qwen2.5 Technical Report

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:05.276721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:05.276721Z digest=sha256:863c5898aaa009c7cd2809b42f0ebdc9a71a003a24ae1992e30ee4050d79db8c

Observation 151a4caa-0971-4d23-8136-f2beb2ff95ba · outbound

This paper cites Gpt-4 system card, 2022.

WebDancer: Towards Autonomous Information Seeking Agency Gpt-4 system card, 2022

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:07.392081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:05.434943Z digest=sha256:a159f82597e5bf02ae6a516c65f94edc027838e0b88cbde928cbce88d3892e1e

Observation feb5be03-6777-4e4b-8235-e0dd31cef140 · outbound

This paper cites HybridFlow: A Flexible and Efficient RLHF Framework.

WebDancer: Towards Autonomous Information Seeking Agency HybridFlow: A Flexible and Efficient RLHF Framework

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:05.548927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:05.548927Z digest=sha256:933a5c1f286390cdab1d8dfdb22b90b5693b9381e9cc36cb0d62b70637be52fd

Observation e55a45b3-fb37-4e59-b735-24fe7ea07497 · outbound

This paper cites Gonzalez, Hao Zhang, and Ion Stoica.

WebDancer: Towards Autonomous Information Seeking Agency Gonzalez, Hao Zhang, and Ion Stoica

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:05.713933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:05.713933Z digest=sha256:6b0b293e8104161276fdcfaa2f39dc689cda17a54ef86ae9fc22de4aa7778607

Observation 61070950-d5a9-4f66-b3e1-1cc4163fe0a0 · outbound

This paper cites write newline.

WebDancer: Towards Autonomous Information Seeking Agency write newline

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:05.825645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:05.825645Z digest=sha256:7720dbc05c5c2de391723a60440fe8bf607a8e8db6f792daf2ddaeec643cad1b

Pith citing papers

Observation 60bac806-d70e-4bf6-973e-e1018d9d2ba5 · inbound

Deep Research Agents: A Systematic Examination And Roadmap cites this paper.

Deep Research Agents: A Systematic Examination And Roadmap WebDancer: Towards Autonomous Information Seeking Agency

Reference 123

Resolution
unresolved
no resolver link, observed 2026-08-06T23:26:56.960098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:26:56.960098Z digest=sha256:2f1fe4a28d5fd97860125695a646a8a27179ecf7d24a4218f463d53c95466271

Observation d41b6370-029a-4e2c-9337-21c672dca542 · inbound

WebSailor: Navigating Super-human Reasoning for Web Agent cites this paper.

WebSailor: Navigating Super-human Reasoning for Web Agent WebDancer: Towards Autonomous Information Seeking Agency

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:37:09.737148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T15:37:09.572241Z digest=sha256:02cb2c3fb91672d77cadafee1ef6588939e39afb394dfcefbd7ac653ec54132b

Observation db155de4-5c38-43a4-973a-81b2d1342f6f · inbound

SciMaster: Towards General-Purpose Scientific AI Agents, Part I. X-Master as Foundation: Can We Lead on Humanity's Last Exam? cites this paper.

SciMaster: Towards General-Purpose Scientific AI Agents, Part I. X-Master as Foundation: Can We Lead on Humanity's Last Exam? WebDancer: Towards Autonomous Information Seeking Agency

Reference 1995

Resolution
unresolved
no resolver link, observed 2026-08-06T19:38:02.151745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:38:02.151745Z digest=sha256:e5f2df4956565b7c1930b5a036825babff0e83d475d3af4f7a7007acbe797be9

Observation 56369d0e-e182-47b1-837b-429fe542c79e · inbound

WebShaper: Agentically Data Synthesizing via Information-Seeking Formalization cites this paper.

WebShaper: Agentically Data Synthesizing via Information-Seeking Formalization WebDancer: Towards Autonomous Information Seeking Agency

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T15:46:06.469971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:46:06.469971Z digest=sha256:12bc06afd2d955bfe0394182e31f193649f5c04a938bad706ca97f92d919c4df

Observation 3e47c703-93d9-4b71-b4dd-36032489f028 · inbound

MetaAgent: Toward Self-Evolving Agent via Tool Meta-Learning cites this paper.

MetaAgent: Toward Self-Evolving Agent via Tool Meta-Learning WebDancer: Towards Autonomous Information Seeking Agency

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T10:17:13.686579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:17:13.686579Z digest=sha256:a6af0f6f7d9c8073e13641b02e1b26cb9b50a27cd42d7ade76e1d17877e46c2c

Observation e5141165-c921-4a66-8436-6ff4afbefcc0 · inbound

Chain-of-Agents: End-to-End Agent Foundation Models via Multi-Agent Distillation and Agentic RL cites this paper.

Chain-of-Agents: End-to-End Agent Foundation Models via Multi-Agent Distillation and Agentic RL WebDancer: Towards Autonomous Information Seeking Agency

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-05T23:57:35.940452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:57:35.940452Z digest=sha256:54e966aa3739aaad319203ba575a12388d95d1d0873aae065b3bad9a9447c94a

Observation cd686d36-58bd-44a3-a292-26692d62add7 · inbound

Encouraging Good Processes Without the Need for Good Answers: Reinforcement Learning for LLM Agent Planning cites this paper.

Encouraging Good Processes Without the Need for Good Answers: Reinforcement Learning for LLM Agent Planning WebDancer: Towards Autonomous Information Seeking Agency

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T15:44:14.916888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:44:14.916888Z digest=sha256:3dc07820bec9792dfdfbaa12323563def0927e3a91b3e4221dd311ea1df805bd

Observation c72e1ce8-6990-4c5b-a52c-768cbeea65cd · inbound

The Landscape of Agentic Reinforcement Learning for LLMs: A Survey cites this paper.

The Landscape of Agentic Reinforcement Learning for LLMs: A Survey WebDancer: Towards Autonomous Information Seeking Agency

Reference 114

Resolution
verified exact
arxiv_id, observed 2026-05-18T19:21:48.577572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-18T19:19:36.427337Z digest=sha256:efe66875041c2426fefb26336f737623aa0182eab8c9d0f29096d9ee8b5065e0

Observation 5cbeea76-2848-4078-94a0-17038199ad9b · inbound

MemSearcher: Training LLMs to Reason, Search and Manage Memory via End-to-End Reinforcement Learning cites this paper.

MemSearcher: Training LLMs to Reason, Search and Manage Memory via End-to-End Reinforcement Learning WebDancer: Towards Autonomous Information Seeking Agency

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-18T01:00:34.599718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-18T00:57:25.902674Z digest=sha256:d9ce9516fa022f24fe8e0fdb38a36829d352ed83a054a056917ba5091d55b633

Observation fce44f7a-14b2-42e2-aad8-36f767442317 · inbound

MiroThinker: Pushing the Performance Boundaries of Open-Source Research Agents via Model, Context, and Interactive Scaling cites this paper.

MiroThinker: Pushing the Performance Boundaries of Open-Source Research Agents via Model, Context, and Interactive Scaling WebDancer: Towards Autonomous Information Seeking Agency

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-17T21:45:17.924421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T21:44:18.744201Z digest=sha256:e61b16616dbce2be226c0e585f0c39278eee8b6833c33f2d5e0ef1cb133a3511

Observation 8552e89a-b1d7-4757-9c9e-8c6bce7f1b41 · inbound

MemSifter: Offloading LLM Memory Retrieval via Outcome-Driven Proxy Reasoning cites this paper.

MemSifter: Offloading LLM Memory Retrieval via Outcome-Driven Proxy Reasoning WebDancer: Towards Autonomous Information Seeking Agency

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-02T19:23:18.139598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T19:23:18.139598Z digest=sha256:0be6b09537c7652cfb1d8e680246b0173f3317bf4b38425613d58db2580de608

Observation ba891453-dd17-4b87-9c4a-b9366080617b · inbound

Evaluating the Search Agent in a Parallel World cites this paper.

Evaluating the Search Agent in a Parallel World WebDancer: Towards Autonomous Information Seeking Agency

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:06:19.324095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T17:02:10.963839Z digest=sha256:2b866874f6ec6d6ec11eedd9b247d98c6ba39a9ee2ffec00ba124283d739b276

Observation 7a0e1a78-836f-4800-99ae-b94a92b4af3a · inbound

LightThinker++: From Reasoning Compression to Memory Management cites this paper.

LightThinker++: From Reasoning Compression to Memory Management WebDancer: Towards Autonomous Information Seeking Agency

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:28:02.480259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T17:25:28.432170Z digest=sha256:68839bfd6c285d4dd9231c40ca2fe76cc19c1c3733c710193e8a23f4b46edce2

Observation 4b7562a8-8625-4289-a592-fbb5ebac7fec · inbound

GeoBrowse: A Geolocation Benchmark for Agentic Tool Use with Expert-Annotated Reasoning Traces cites this paper.

GeoBrowse: A Geolocation Benchmark for Agentic Tool Use with Expert-Annotated Reasoning Traces WebDancer: Towards Autonomous Information Seeking Agency

Reference 52

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T17:33:02.332097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T17:31:08.575993Z digest=sha256:455c427d22316a0e5d39c7a414c71a26bed95d2dc15c6fd07ed4d92f443565da

Observation 9c16834f-bdd2-4aee-95eb-8a7c45b030c2 · inbound

Mind DeepResearch Technical Report cites this paper.

Mind DeepResearch Technical Report WebDancer: Towards Autonomous Information Seeking Agency

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:50:20.538123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T11:46:49.178896Z digest=sha256:8fafd613be8bac4ba34ee7464f28a9d1aea8cc1fb6c6cc68dfe91bc2d41bcd39

Observation 1f5db6aa-dd7c-4772-84be-775b963e3483 · inbound

LiteResearcher: A Scalable Agentic RL Training Framework for Deep Research Agent cites this paper.

LiteResearcher: A Scalable Agentic RL Training Framework for Deep Research Agent WebDancer: Towards Autonomous Information Seeking Agency

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-07-05T15:11:10.887948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-07-05T15:03:50.420072Z digest=sha256:6811de6558df44b00a4842ea2063ac28d1aa3d884e6968af1a51b684ad36df9b

Observation d99191de-949b-4535-b09f-c41a927587fc · inbound

Agent-World: Scaling Real-World Environment Synthesis for Evolving General Agent Intelligence cites this paper.

Agent-World: Scaling Real-World Environment Synthesis for Evolving General Agent Intelligence WebDancer: Towards Autonomous Information Seeking Agency

Reference 105

Resolution
verified exact
arxiv_id, observed 2026-05-10T05:25:54.258876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T05:24:00.503836Z digest=sha256:24d675e06b25f982674501b6c9b9df59af173e460b951863efe5bb1779bf0115

Observation b4a2eb24-7917-42f4-94c3-1c854ecb4f8f · inbound

SiriusHelper: An LLM Agent-Based Operations Assistant for Big Data Platforms cites this paper.

SiriusHelper: An LLM Agent-Based Operations Assistant for Big Data Platforms WebDancer: Towards Autonomous Information Seeking Agency

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:46:24.924422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-09T21:07:01.601808Z digest=sha256:c752b17a3785c7beea1bbc12f799c8702cfc593b2944de055bbbeb21999d4fb2

Observation 0766f0c3-a8d3-4b55-8136-ea340fae67e0 · inbound

SciResearcher: Scaling Deep Research Agents for Frontier Scientific Reasoning cites this paper.

SciResearcher: Scaling Deep Research Agents for Frontier Scientific Reasoning WebDancer: Towards Autonomous Information Seeking Agency

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:01:08.288363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-09T14:18:14.048230Z digest=sha256:dd24e81bb823a5acee97e7607b282f9a8837c999434ea31253373128b606ddc6

Observation 17535dfe-d035-4a7e-a772-ad6615b86701 · inbound

PiCA: Pivot-Based Credit Assignment for Search Agentic Reinforcement Learning cites this paper.

PiCA: Pivot-Based Credit Assignment for Search Agentic Reinforcement Learning WebDancer: Towards Autonomous Information Seeking Agency

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:06:26.668662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T04:35:21.469086Z digest=sha256:6ac8224653d4fb648ea8b4f4b8cee9e2291ab445e03f009ed477fa82abca254e

Observation 45f8dd8f-338c-41e7-8655-3ace8ad76eb9 · inbound

PiCA: Pivot-Based Credit Assignment for Search Agentic Reinforcement Learning cites this paper.

PiCA: Pivot-Based Credit Assignment for Search Agentic Reinforcement Learning WebDancer: Towards Autonomous Information Seeking Agency

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:37:29.685084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:34:32.166480Z digest=sha256:5bd8df182cb1cca3b24d6d304e5c905738e53abe376d90c078c00241b4cba74d

Observation 72e0943b-fbdf-42bd-8f57-c924b1fdf245 · inbound

ViDR: Grounding Multimodal Deep Research Reports in Source Visual Evidence cites this paper.

ViDR: Grounding Multimodal Deep Research Reports in Source Visual Evidence WebDancer: Towards Autonomous Information Seeking Agency

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-14T19:19:23.778081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-14T19:18:52.801531Z digest=sha256:b8ccb989d7c913a32764744638ca03d30e6171cd596f203434f5aee1f02daca0

Observation c134a30e-0bce-4577-ada8-d90664d85262 · inbound

WARD: Adversarially Robust Defense of Web Agents Against Prompt Injections cites this paper.

WARD: Adversarially Robust Defense of Web Agents Against Prompt Injections WebDancer: Towards Autonomous Information Seeking Agency

Reference 70

Resolution
verified exact
arxiv_id, observed 2026-07-01T14:45:49.576978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T20:16:13.413064Z digest=sha256:018222bc3833127aa8a548aa6341a52e5b961d6e387225991a7fc605149b3a83

Observation 209b8218-c2bc-457c-8c60-b23ffcebee8d · inbound

Efficient Agentic Reasoning Through Self-Regulated Simulative Planning cites this paper.

Efficient Agentic Reasoning Through Self-Regulated Simulative Planning WebDancer: Towards Autonomous Information Seeking Agency

Reference 104

Resolution
verified exact
arxiv_id, observed 2026-05-22T06:34:40.874464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T06:33:36.846345Z digest=sha256:d2d7501f1598fafa1c6d9cab0ae1537ab2514c80ee4be8d65ea37594f9f8c22f

Observation 6ea571ab-11e9-4aef-ba31-66773e687e26 · inbound

AgentFugue: Agent Scaling for Long-Horizon Tasks through Collective Reasoning cites this paper.

AgentFugue: Agent Scaling for Long-Horizon Tasks through Collective Reasoning WebDancer: Towards Autonomous Information Seeking Agency

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-06-30T13:34:40.417297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T13:29:36.710152Z digest=sha256:061bb469a38ae3abc71a97e3ee853dfbb5535ad1c45cc5d8ffa676fa8a5238e3

Observation bc87a14a-d656-4f65-9f93-c601cb9a3d82 · inbound

Beyond Trajectory Rewards: Step-level Credit Assignment for Agentic Search via Graph Modeling cites this paper.

Beyond Trajectory Rewards: Step-level Credit Assignment for Agentic Search via Graph Modeling WebDancer: Towards Autonomous Information Seeking Agency

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:03:14.247506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-29T07:58:55.548382Z digest=sha256:c49e018ce4569582686859b20bfafa232d786465bc52617ac15cf28f5396ea42

Observation 7b02dcca-b0a2-4f0a-8ed2-012566e937b5 · inbound

SlimSearcher: Training Efficiency-Aware Web Agents via Adaptive Reward Gating cites this paper.

SlimSearcher: Training Efficiency-Aware Web Agents via Adaptive Reward Gating WebDancer: Towards Autonomous Information Seeking Agency

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T16:17:08.729780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-27T22:54:44.329613Z digest=sha256:9290ec490fc63866f3586399f248609311af524725f66f88276fe01704340bb1

Observation 4fd56797-b452-4c03-91ac-c692b3ba82d1 · inbound

Struct-Searcher: Agentic Structural Thinking Advances Multimodal Deep Information Seeking cites this paper.

Struct-Searcher: Agentic Structural Thinking Advances Multimodal Deep Information Seeking WebDancer: Towards Autonomous Information Seeking Agency

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T16:47:10.197485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-27T22:20:53.187362Z digest=sha256:f07098b0cc9e0456272854df72ca1245e35d5e3e380cb65ee4b08a1ff911c8ce

Observation c19194ca-6c64-4408-b54a-7a0fb423b165 · inbound

Toward Generalist Autonomous Research via Hypothesis-Tree Refinement cites this paper.

Toward Generalist Autonomous Research via Hypothesis-Tree Refinement WebDancer: Towards Autonomous Information Seeking Agency

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-06-27T09:40:47.133729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-27T09:34:41.800309Z digest=sha256:2e59311bbd7a6219bc9e81313ad7ce19e5d32ce1830351d99eb94209295341ed

Observation 4d045f32-4ff9-4b3f-86cc-1072348a3461 · inbound

GraphPO: Graph-based Policy Optimization for Reasoning Models cites this paper.

GraphPO: Graph-based Policy Optimization for Reasoning Models WebDancer: Towards Autonomous Information Seeking Agency

Reference 48

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T00:59:21.087774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T20:45:34.358581Z digest=sha256:1ae2a6b7466654ef635e1a81260de505a9d52a7c6df824035ed84789283e31ee

Observation 121804f5-a788-4537-89a1-3af1b118b299 · inbound

Reinforcement Learning without Ground-Truth Solutions can Improve LLMs cites this paper.

Reinforcement Learning without Ground-Truth Solutions can Improve LLMs WebDancer: Towards Autonomous Information Seeking Agency

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T13:59:51.925971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-26T04:47:47.691913Z digest=sha256:d874be834b8dd9f578b63f7741af8bdc89a5c50e7efc8958013639ad20aff2b9

Observation 71505682-a504-446f-b65b-9e32eac8a431 · inbound

SimpleSearch-VL: A Simple Recipe for Multimodal Agentic Deep Search cites this paper.

SimpleSearch-VL: A Simple Recipe for Multimodal Agentic Deep Search WebDancer: Towards Autonomous Information Seeking Agency

Reference 43

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T09:55:41.032622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-07-01T06:02:48.532478Z digest=sha256:0cf4388ffb6a4f815ddca00bcf39ff918cc16943e5ea303a8cec485f70dfb0a6

Observation a5f4314a-1061-4986-b8d9-c46a4256b682 · inbound

When RAG Meets Query Planning: Logical Query Trees for Resolving Exploratory Reasoning Problems cites this paper.

When RAG Meets Query Planning: Logical Query Trees for Resolving Exploratory Reasoning Problems WebDancer: Towards Autonomous Information Seeking Agency

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-07-02T06:56:43.880693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-02T06:47:57.679984Z digest=sha256:6c897ea7d02a113b63e01a815a15443f1875d048f95022d5c84cc47544f7053c

Observation 89af5c0f-afed-44dd-9301-55c5a4261637 · inbound

When RAG Meets Query Planning: Logical Query Trees for Resolving Exploratory Reasoning Problems cites this paper.

When RAG Meets Query Planning: Logical Query Trees for Resolving Exploratory Reasoning Problems WebDancer: Towards Autonomous Information Seeking Agency

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-07-03T19:08:49.349018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-03T19:05:04.593481Z digest=sha256:40c036281e17e053bda79a3817c52c196b535c8863a81215b980c50407078a21

Observation 5105af75-7e87-46d1-9b06-3f71130605b1 · inbound

SearchEyes: Towards Frontier Multimodal Deep Search Intelligence via Search World Simulation cites this paper.

SearchEyes: Towards Frontier Multimodal Deep Search Intelligence via Search World Simulation WebDancer: Towards Autonomous Information Seeking Agency

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-07-08T20:05:34.016144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-08T20:04:30.942032Z digest=sha256:347ab22c62b3a600028eb981da9b24499bcf6c5f1e8ce4615717b619e295fa7e

Observation f876bd42-1726-401d-b7e4-1736d07dd297 · inbound

Entropy Pacing Policy Optimization for Multi-Task Agentic Reinforcement Learning cites this paper.

Entropy Pacing Policy Optimization for Multi-Task Agentic Reinforcement Learning WebDancer: Towards Autonomous Information Seeking Agency

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-07-09T18:26:26.297654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-09T18:19:12.180499Z digest=sha256:016574829f3e347282ce1091dd456ee1c50109704ec6c3456ed414700b718de6

Observation 6e9d84fb-bc79-49c7-8494-a4803e013db4 · inbound

ToolAnchor: Anchoring Counterfactual Context to Boost Agentic Tool-use Capability cites this paper.

ToolAnchor: Anchoring Counterfactual Context to Boost Agentic Tool-use Capability WebDancer: Towards Autonomous Information Seeking Agency

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T06:36:48.233756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:36:48.233756Z digest=sha256:a65bba5d72de549ebc1cdf56ec73948938c1e879a9727bfcf3aa7a688554696a

Observation 6c0496fa-01bb-4c21-8b8d-35eb7d750d84 · inbound

TAPO: Transition-Aware Policy Optimization for LLM Agents cites this paper.

TAPO: Transition-Aware Policy Optimization for LLM Agents WebDancer: Towards Autonomous Information Seeking Agency

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-31T21:44:39.651881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T21:44:39.651881Z digest=sha256:c5c5c0ef560fb96c2353258788d176e725f51d9ab646acbfe806e47130ba20d7

Observation 39b823c0-e97b-4377-89f2-e2b6287f8d13 · inbound

Fetch-then-Explore: Decoupling Selection from Extraction over a Persistent Workspace for Search Agents cites this paper.

Fetch-then-Explore: Decoupling Selection from Extraction over a Persistent Workspace for Search Agents WebDancer: Towards Autonomous Information Seeking Agency

Reference 118

Resolution
unresolved
no resolver link, observed 2026-08-04T15:12:55.814757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T15:12:55.814757Z digest=sha256:9f6a223ec65ce4793b1633ed984a0486845b5fdd261de779b2d0dc65f8ba47fe

Observation 10978142-90b9-479d-b748-93a4baabba7c · inbound

From Simple QA to Deep Research: A Verifiable Benchmark Constructed through Iterative Task Evolution cites this paper.

From Simple QA to Deep Research: A Verifiable Benchmark Constructed through Iterative Task Evolution WebDancer: Towards Autonomous Information Seeking Agency

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-04T13:31:24.114972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:31:24.114972Z digest=sha256:ddd0dbe8c9a273472bc6e96b9d6791899150179ce61c396b645ad197748844ea

Observation 4b2f2681-3270-418e-bf12-105ca9a541a8 · inbound

Video-DeepResearch: Towards the Next-Generation Multimodal Deepresearch Agent cites this paper.

Video-DeepResearch: Towards the Next-Generation Multimodal Deepresearch Agent WebDancer: Towards Autonomous Information Seeking Agency

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T04:44:29.293979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T04:44:29.293979Z digest=sha256:ab2c6d54758c89aba3f17e8d8d8b3ff17c2d33780880e0df821d326bb00bf097