Pith. sign in

Paper Citation Record · LEDGER

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent

As of 21 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 6 inbound Pith citation observations for arXiv:2505.07596.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.07596 v1

Coverage vector

measured 63 of 63 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T22:17:44.975973Z

measured 69 of 69 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T18:47:14.522738Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T10:58:02.960267Z

Reference resolution

63 of 63 outbound references displayed

  • verified exact0
  • verified fuzzy35
  • unresolved28
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5e794911-03a6-4605-b3d7-5c29c44d8580 · outbound

This paper cites Back to basics: Revisiting reinforce style optimization for learning from human feedback in llms, 2024.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Back to basics: Revisiting reinforce style optimization for learning from human feedback in llms, 2024

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T22:17:44.688087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:17:44.688087Z digest=sha256:e63c2b33b20ef0bb4db84219d80873aa11b75c9f4f781104bdcf13a25578ada7

Observation 318dafe7-3ebc-47dd-9902-df54f3288c98 · outbound

This paper cites Teaching large language models to express knowledge boundary from their own signals, 2024.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Teaching large language models to express knowledge boundary from their own signals, 2024

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.828738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.692977Z digest=sha256:e0e2e13da8f0a5faf350ab78e7bc61f8626c63a55982d582ecf7c22eb4bd2bb8

Observation 2b7284b1-ceab-44d9-a1ef-d55edabb38d2 · outbound

This paper cites Pan, Wen Zhang, Huajun Chen, Fan Yang, Zenan Zhou, and Weipeng Chen.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Pan, Wen Zhang, Huajun Chen, Fan Yang, Zenan Zhou, and Weipeng Chen

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T22:17:44.697218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:17:44.697218Z digest=sha256:b2945950a56d36f89c1d84430d308c5721002207a64a38a6cd5b53667e4b229f

Observation 13ce135d-5b57-4b3e-9781-2601b2511989 · outbound

This paper cites Understand- ing the interplay between parametric and contextual knowledge for large language models, 2024.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Understand- ing the interplay between parametric and contextual knowledge for large language models, 2024

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.808355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.701420Z digest=sha256:2bc8434e40a1186317e5ee96b420e5d517665e5a7ebcad7b9f8b0bc40e67cd88

Observation 9c7b1947-4aca-4a23-bb79-74f5b435e698 · outbound

This paper cites an unresolved cited work.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T22:17:44.705762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:17:44.705762Z digest=sha256:8f17197e079fe2daa2e306a8992f69399c1cf2aeac8680c15440af9f3374f644

Observation 43796db7-39f9-4256-b2eb-791887f9ed1f · outbound

This paper cites Understand what llm needs: Dual preference alignment for retrieval-augmented generation.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Understand what llm needs: Dual preference alignment for retrieval-augmented generation

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.787222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.711046Z digest=sha256:012623ea8c9226da530e39314c438baa3ee20db9ed1eba123c4e72899699fa64

Observation 2777ac2a-e60a-418e-922d-35ae70ae1bd1 · outbound

This paper cites Enhancing noise robustness of retrieval-augmented language models with adaptive adversarial training.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Enhancing noise robustness of retrieval-augmented language models with adaptive adversarial training

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.775146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.716520Z digest=sha256:3293ea482a309a3421d8f43b9671f6a524a91f414bdbb82f11f20694824dde94

Observation fe2c52e9-639c-4b90-9350-1a5c97c9ea2a · outbound

This paper cites Retrieval-augmented generation for large language models: A survey, 2024.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Retrieval-augmented generation for large language models: A survey, 2024

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T22:17:44.720709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:17:44.720709Z digest=sha256:cb4dd5c385df0ea3b52d4ade9195454862bd03b879d783e8e3f514cbb6bd219a

Observation f6530895-3770-481a-8ad9-ace7656f1255 · outbound

This paper cites an unresolved cited work.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T22:17:44.724341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:17:44.724341Z digest=sha256:b8396d07679b0a52d229338ded59dbf1c7b03ecdd02af70fbced7fb0b5beef65

Observation 9419ac91-5013-4b9b-897f-b7b87994f2f7 · outbound

This paper cites Deeprag: Thinking to retrieval step by step for large language models, 2025.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Deeprag: Thinking to retrieval step by step for large language models, 2025

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.745503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.728148Z digest=sha256:6651fddaa1cdec9be888ae94d897a9735ae48355862afda34b699e27d21e8ff7

Observation 41215034-545a-4e6e-9a06-721d89abd0f5 · outbound

This paper cites Language models as knowledge bases: On entity representations, storage capacity, and paraphrased queries.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Language models as knowledge bases: On entity representations, storage capacity, and paraphrased queries

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.732094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.731691Z digest=sha256:7a6960f8cf1788dfd11aa702ccc6dc7fe7deb499a872e257b9551aa4289892f1

Observation 922e2833-3f30-4651-82ed-ca0b217b3d2f · outbound

This paper cites Constructing a multi-hop QA dataset for comprehensive evaluation of reasoning steps.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Constructing a multi-hop QA dataset for comprehensive evaluation of reasoning steps

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T22:17:44.739982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:17:44.739982Z digest=sha256:17d84e0e7dc3a490df6a81c2388e8e1df3eca957fde14f4462925a9933b93d03

Observation df4a3e8e-6f0e-426f-9540-ee28a32d5768 · outbound

This paper cites A survey on hallucination in large language models: Principles, taxonomy, challenges, and open questions.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent A survey on hallucination in large language models: Principles, taxonomy, challenges, and open questions

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.703278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.744001Z digest=sha256:018bfbb0ae1774024a9e9753dc0bf083030a79ad43fdbc637c210ec29276be01

Observation bbb45a1e-4976-4f9f-908a-4fae6aacf43e · outbound

This paper cites an unresolved cited work.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T22:17:44.747748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:17:44.747748Z digest=sha256:168ef67340410b23218c1e3a9614aa21a163bc1881abb131e84b3678747600e3

Observation aaef2a86-4e41-4443-a041-da67ec82a735 · outbound

This paper cites Xu, Luyu Gao, Zhiqing Sun, Qian Liu, Jane Dwivedi-Yu, Yiming Yang, Jamie Callan, and Graham Neubig.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Xu, Luyu Gao, Zhiqing Sun, Qian Liu, Jane Dwivedi-Yu, Yiming Yang, Jamie Callan, and Graham Neubig

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T22:17:44.751592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:17:44.751592Z digest=sha256:94e00d9a3f378bf6eb50ceb0bb8cc58eb25db2becea48cb226ce395aaafa928f

Observation a8f17d09-bb63-46ef-b0c6-b6f198d68f9b · outbound

This paper cites Search-r1: Training llms to reason and leverage search engines with reinforcement learning, 2025.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Search-r1: Training llms to reason and leverage search engines with reinforcement learning, 2025

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T22:17:44.755593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:17:44.755593Z digest=sha256:e4e8dcfcb2af27628f9361213df827197ca16407bd42cf706c3e6c864d20f0ca

Observation 857f1108-d1e8-4e50-a254-2c9d76c6609c · outbound

This paper cites FlashRAG: A Modular Toolkit for Efficient Retrieval-Augmented Generation Research.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent FlashRAG: A Modular Toolkit for Efficient Retrieval-Augmented Generation Research

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T22:17:44.759328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:17:44.759328Z digest=sha256:644a2b91fe00253d8e4b5c7c3be010e9f1326b338683625225db086da01f4a2d

Observation 772cee32-3e15-4245-a91e-6d057343b5b4 · outbound

This paper cites Cutting off the head ends the conflict: A mechanism for interpreting and mitigating knowledge conflicts in language models.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Cutting off the head ends the conflict: A mechanism for interpreting and mitigating knowledge conflicts in language models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.667276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.763847Z digest=sha256:1051ed3e78e279cdb4fb7ca8a4097fa35cba21c0a679a1371dfefae17889c4a3

Observation 71ebfbd6-6f6d-4dd0-95ed-d87c291b4b12 · outbound

This paper cites Dai, Jakob Uszkoreit, Quoc Le, and Slav Petrov.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Dai, Jakob Uszkoreit, Quoc Le, and Slav Petrov

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T22:17:44.768235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:17:44.768235Z digest=sha256:08ecf3e5d829c698f7ee7138e18edd53eb110f72855079ea06081a877bc7371c

Observation c412cae4-ee66-4832-91ed-17fe614aaafa · outbound

This paper cites Knowledge boundary of large language models: A survey, 2024.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Knowledge boundary of large language models: A survey, 2024

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T22:17:44.773187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:17:44.773187Z digest=sha256:3da1da67e854b07dad71fcbdd838e5383db8eb3e10ce35b3e1b315c82ed5cc70

Observation efdeb2da-6b3b-4cd1-883c-41cb27ba86f6 · outbound

This paper cites Remax: A simple, effective, and efficient reinforcement learning method for aligning large language models, 2024.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Remax: A simple, effective, and efficient reinforcement learning method for aligning large language models, 2024

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.637885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.780084Z digest=sha256:cd6234ca3c4dad0eb7c9800dac2894d636dd375b78f8589a401ecac7d1087dfa

Observation c46f7dcc-d60f-4d19-9371-5239fa369114 · outbound

This paper cites When not to trust language models: Investigating effectiveness of parametric and non-parametric memories.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent When not to trust language models: Investigating effectiveness of parametric and non-parametric memories

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.623618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.784624Z digest=sha256:d26428182af22207c9099997d8ce9700b31b726831db16304a8fc795f9cec465

Observation 9b3f5e4f-59f2-4c95-9bf9-cb1f0e9f8e23 · outbound

This paper cites Generation-augmented retrieval for open-domain question answering.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Generation-augmented retrieval for open-domain question answering

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.610941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.790070Z digest=sha256:ea357d6902eb04bf67f6f02d9463c435948a74f1f32ba006b7c4397870099375

Observation f2cb25be-590f-4e4d-a71d-757b2550e2b3 · outbound

This paper cites an unresolved cited work.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T22:17:44.798457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:17:44.798457Z digest=sha256:85018a7e4c8b796537f25112a14c996e961e517ce9c5be5c4762c0a0b9329edf

Observation b9d1343e-f22f-41d4-90bc-4d5a7b94f168 · outbound

This paper cites Direct preference optimization: Your language model is secretly a reward model.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Direct preference optimization: Your language model is secretly a reward model

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T22:17:44.812631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:17:44.812631Z digest=sha256:b2b48980b4667f2a0eb16ed4b5001f62947725d384035b27ed57e0843e18e209

Observation 086f09aa-5615-4edb-9e83-bb6ae020fcd2 · outbound

This paper cites an unresolved cited work.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T22:17:44.803493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:17:44.803493Z digest=sha256:0dbbe3e1b530e603b5d66cfa1dbd33efe92dafcaa55b95a263dd5252932d3373

Observation 2fe924eb-b74b-4547-a675-ceccb5f1fb42 · outbound

This paper cites High- dimensional continuous control using generalized advantage estimation, 2018.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent High- dimensional continuous control using generalized advantage estimation, 2018

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T22:17:44.823984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:17:44.823984Z digest=sha256:8b92f450f9e5da63cdb8aa1b1f1b51f6a26fa5c4ca537e4c3dd95121b25f1ffa

Observation d6d01a27-3856-4acf-81d8-74e820648ee3 · outbound

This paper cites Proximal policy optimization algorithms, 2017.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Proximal policy optimization algorithms, 2017

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T22:17:44.828149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:17:44.828149Z digest=sha256:062911d8a7b38c1734483e0196b8366861afcebb54e81b2a9da341af25ddacc7

Observation 6f5d9948-b9b6-4145-b554-cb736a012d50 · outbound

This paper cites Investigating the factual knowledge boundary of large language models with retrieval augmentation, 2024.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Investigating the factual knowledge boundary of large language models with retrieval augmentation, 2024

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.574185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.817903Z digest=sha256:14406e762b71a779487a523ce0157c8702f04d5d24c3704e1ac53a79a8381986

Observation d23a71f4-b674-4824-918d-13b61d5deb48 · outbound

This paper cites an unresolved cited work.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T22:17:44.836437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:17:44.836437Z digest=sha256:f78e6e109e727e661d2ec3fed5b1620d1fa4b6b2c087e135b556013d347f9f4b

Observation dcac03fc-5e50-49c1-adf1-bd9c030af051 · outbound

This paper cites Hybridflow: A flexible and efficient rlhf framework.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Hybridflow: A flexible and efficient rlhf framework

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.517812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.840607Z digest=sha256:52882928568458226eb4b51bbe2b076d7a0ffa7db981e32cb0c939d2d5beb48a

Observation 17b3ff7f-fce1-4bd3-ae9d-85cc6a7e6199 · outbound

This paper cites Enhancing retrieval-augmented large language models with iterative retrieval-generation synergy.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Enhancing retrieval-augmented large language models with iterative retrieval-generation synergy

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.539619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.832389Z digest=sha256:a7b04bb92281d00653724722ec2204b844139fcd5f5b14a0f0ba14b6e2802ff3

Observation 4c8fffb8-b695-4616-b6c6-2491fac9a78c · outbound

This paper cites Crossing the reward bridge: Expanding rl with verifiable rewards across diverse domains, 2025.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Crossing the reward bridge: Expanding rl with verifiable rewards across diverse domains, 2025

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T22:17:44.852213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:17:44.852213Z digest=sha256:91e87c03d0e7cf9880d54eab9f4943cf5c25ba76e05b04afc883e1b049a3ecf4

Observation 3f6e5b6c-c54f-45b2-8daa-d79aa0828ff7 · outbound

This paper cites Interleaving retrieval with chain-of-thought reasoning for knowledge-intensive multi-step questions.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Interleaving retrieval with chain-of-thought reasoning for knowledge-intensive multi-step questions

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.487933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.856679Z digest=sha256:a2e33499d641ef8f3dfd620161b478fc4282e4dba5342140527db12f0b93383f

Observation 2ba688b0-4a5a-4064-a62c-4363e65f8b14 · outbound

This paper cites R1-searcher: Incentivizing the search capability in llms via reinforcement learning, 2025.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent R1-searcher: Incentivizing the search capability in llms via reinforcement learning, 2025

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T22:17:44.844990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:17:44.844990Z digest=sha256:fdc89bc74a2f2fd4d599b996c9b3f5da15fe020b0010a31f3d68db8e3d6b07ed

Observation 30bab5bd-21a5-4877-91c4-9931280865f2 · outbound

This paper cites Text Embeddings by Weakly-Supervised Contrastive Pre-training.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Text Embeddings by Weakly-Supervised Contrastive Pre-training

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T22:17:44.865331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:17:44.865331Z digest=sha256:975fe72831ef2cccbf4ae17370c6ddfe142d8d95e6df29123299d137134d2d99

Observation 74d586ae-8701-4e13-b5a8-156af591f4ef · outbound

This paper cites Reinforcement learning enhanced llms: A survey, 2025.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Reinforcement learning enhanced llms: A survey, 2025

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T22:17:44.869175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:17:44.869175Z digest=sha256:dbff5234b44d9a1ce1c5debac6fab8ea718a1137e2acb816c49e1ccbcd5c52bd

Observation 96a9e5fa-1246-4121-b334-d3f1285263d8 · outbound

This paper cites Chain- of-retrieval augmented generation, 2025.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Chain- of-retrieval augmented generation, 2025

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.472290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.861168Z digest=sha256:f332e8c554be3d2816a6d2ca280fd74ec7e539b2161f2b089f650bbafd00eb12

Observation 77230c82-0c39-4663-803c-5ac0a9420b4d · outbound

This paper cites Rejection improves reliability: Training LLMs to refuse unknown questions using RL from knowledge feedback.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Rejection improves reliability: Training LLMs to refuse unknown questions using RL from knowledge feedback

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.425750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.878461Z digest=sha256:bc846c8dde544b46f04e019508fc6fb1fc4f1c31971157c79f27f218c54f5bcb

Observation da7a99f2-ec62-4ae1-99b1-cc529751a94d · outbound

This paper cites Knowledge conflicts for LLMs: A survey.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Knowledge conflicts for LLMs: A survey

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T22:17:44.882853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:17:44.882853Z digest=sha256:cdab48ebd5c5fdf868161eeb35161e960e2841d7244d23cc49971f185911796a

Observation 75e5eeec-447f-4cff-9442-cea832c7ff33 · outbound

This paper cites Perception of knowledge boundary for large language models through semi- open-ended question answering.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Perception of knowledge boundary for large language models through semi- open-ended question answering

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.444766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.873827Z digest=sha256:8fd2e79b595095e387794fa2779e7fc02311ec36443cd229b129a146dfc3e926

Observation 60f10727-0342-406c-908e-ddcaf9499d7f · outbound

This paper cites Auto-rag: Autonomous retrieval-augmented generation for large language models, 2024.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Auto-rag: Autonomous retrieval-augmented generation for large language models, 2024

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.366096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.892338Z digest=sha256:105fcfdf0ff17515fd95d8578716855b6155174d1dd377227e8473a735caf10e

Observation 26c51263-97d3-448a-b91f-8595f0d6bb00 · outbound

This paper cites Steering knowledge selection behaviours in LLMs via SAE-based representation engineering.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Steering knowledge selection behaviours in LLMs via SAE-based representation engineering

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.338919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.896900Z digest=sha256:7b92b9121a3eef64ccfdfa22d14c878f34a8779262751e26ea632ed6f0fa41cd

Observation ba462f99-1274-4893-87c2-f9ca5f4170e0 · outbound

This paper cites an unresolved cited work.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-15T22:17:45.395909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.887408Z digest=sha256:0a76073f00bd086c14ee94a5abfffb513f3eb5f299181387568caeb6d5380d09

Observation 7f90f6be-2aa0-4b5f-af83-f6cc378f5790 · outbound

This paper cites [Yes] " is generally preferable to.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent [Yes] " is generally preferable to

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.306002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.905563Z digest=sha256:234b5ffc82a8b52bb898346f82a46198bffc00b9a4afbed37a0d5cc2a36575e9

Observation 314ea1e6-f99b-4920-af3a-ad8763fe6e40 · outbound

This paper cites Knowing what llms do not know: A simple yet effective self-detection method, 2024.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Knowing what llms do not know: A simple yet effective self-detection method, 2024

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.320656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.901426Z digest=sha256:1950a8816426d2c189bc4f0d5d6ed3f0db09507ea192b585967c3ef6944234b3

Observation ae418808-71f6-41ca-83de-b3307fa6acf5 · outbound

This paper cites Guidelines: • The answer NA means that the abstract and introduction do not include the claims made in the paper.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Guidelines: • The answer NA means that the abstract and introduction do not include the claims made in the paper

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.292613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.910508Z digest=sha256:8ba90e8fccf8297822a8c4f75210822a72b6d9389fd6d465a054c5aacf5d8263

Observation 43abe185-5442-4de1-b99a-3a14c7acdb82 · outbound

This paper cites Limitations.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Limitations

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.276743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.915263Z digest=sha256:dfcf52f9d732dffd93a465c759d09272b93be71ef62536048e101218efc61af7

Observation 015264fc-8e40-44a1-85d0-09be1a65d6d9 · outbound

This paper cites Guidelines: • The answer NA means that the paper does not include theoretical results.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Guidelines: • The answer NA means that the paper does not include theoretical results

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-15T22:17:44.919153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:17:44.919153Z digest=sha256:be5f45da6bc43fe6448390f4ca3045579425c74dda2707d1c940cf0ab2645312

Observation c82e9049-4adb-4775-ba1b-e9278e1b7d87 · outbound

This paper cites Guidelines: • The answer NA means that the paper does not include experiments.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Guidelines: • The answer NA means that the paper does not include experiments

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.250491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.923537Z digest=sha256:167202cc8b60ab0fb1354e088c44e083a744646dbb47e4adc59bca3c48a55485

Observation 6eb8525a-0da1-40f0-85ea-12427b048107 · outbound

This paper cites Guidelines: • The answer NA means that paper does not include experiments requiring code.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Guidelines: • The answer NA means that paper does not include experiments requiring code

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.233202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.928151Z digest=sha256:5e9d44a70b53b6d8c360f8dde258ae7bb7781d8a0b4f3fef12df32e699a3520e

Observation dc9e8303-086e-41b0-8d2a-2830bd2854fb · outbound

This paper cites Guidelines: • The answer NA means that the paper does not include experiments.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Guidelines: • The answer NA means that the paper does not include experiments

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.220426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.932060Z digest=sha256:fbbb5b173936f12c7ef3ecd38034cb4883d1b75230d03026e19d1e724ad78a0f

Observation 3ebc1464-3fb6-420e-a1c7-00aebabddb1c · outbound

This paper cites Guidelines: • The answer NA means that the paper does not include experiments.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Guidelines: • The answer NA means that the paper does not include experiments

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.203947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.935976Z digest=sha256:72ffb0c95d25bf87bab934cc6b81171ee917ee4c3fc87e1d602ce83b8edb7072

Observation 1fb50bfe-9b39-4b55-a22f-32fdd9b645b5 · outbound

This paper cites Guidelines: • The answer NA means that the paper does not include experiments.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Guidelines: • The answer NA means that the paper does not include experiments

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.186859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.939856Z digest=sha256:015e25c24109695349229d6e9d9788ec7653ef56d8dab6353aff2b948ecfad9e

Observation d1953279-0ad4-49e3-8b2c-4fca4f91fc0c · outbound

This paper cites Guidelines: • The answer NA means that the authors have not reviewed the NeurIPS Code of Ethics.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Guidelines: • The answer NA means that the authors have not reviewed the NeurIPS Code of Ethics

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.172091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.944152Z digest=sha256:5b1449905dbf1e3151d8e23624d9a959a1e24e94b2c37e8b9600505a40543159

Observation 80cbd0e7-99cd-4504-93f4-503a8fcb387f · outbound

This paper cites Guidelines: • The answer NA means that there is no societal impact of the work performed.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Guidelines: • The answer NA means that there is no societal impact of the work performed

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.150108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.947820Z digest=sha256:9a33fa898d358f35774de4539931cb975a173835cdc6aa532567384a7d30d7e4

Observation 76998411-1def-40ee-b409-0faa1d25d772 · outbound

This paper cites Guidelines: • The answer NA means that the paper poses no such risks.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Guidelines: • The answer NA means that the paper poses no such risks

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.131609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.951572Z digest=sha256:5bb8bd4489b40795e9b9c953d6b355c5e19f814f3cb64f7dededed2d5b85145b

Observation 8c0f0f8c-158c-4d25-8ba8-d2483c14dcad · outbound

This paper cites Guidelines: • The answer NA means that the paper does not use existing assets.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Guidelines: • The answer NA means that the paper does not use existing assets

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.115850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.955713Z digest=sha256:006912ad31bed723b88241e5994cde137189ea0eadf6401a461b50b5850e7adf

Observation af5200a3-1b2a-47fb-9dcd-fb33b3bbdfe5 · outbound

This paper cites Guidelines: • The answer NA means that the paper does not release new assets.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Guidelines: • The answer NA means that the paper does not release new assets

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.098633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.960054Z digest=sha256:1e9d6e908c4e84847a37c4218504c6a238c990312ad90c4bcc512bdf8131592b

Observation 315b1df0-89d4-45a1-a901-4140da27109c · outbound

This paper cites 21 Guidelines: • The answer NA means that the paper does not involve crowdsourcing nor research with human subjects.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent 21 Guidelines: • The answer NA means that the paper does not involve crowdsourcing nor research with human subjects

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:17:45.083354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T22:17:44.964419Z digest=sha256:58a30c2873c5c9db8535dcd120a41d3d554b73ed7ec22463ed6110f9f3655991

Observation fbe7b101-6428-4cf8-a290-ef039aae77f7 · outbound

This paper cites Guidelines: • The answer NA means that the paper does not involve crowdsourcing nor research with human subjects.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Guidelines: • The answer NA means that the paper does not involve crowdsourcing nor research with human subjects

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-15T22:17:44.971158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:17:44.971158Z digest=sha256:59d99e7b4e09f7e1913763cef1b217db0ba8f5c0803b3eeaa26035ce957dcede

Observation 2d16432a-e796-48d3-be40-742e392c5015 · outbound

This paper cites Answer: [NA] Justification: The core method development in this research does not involve LLMs as any important, original, or non-standard components.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Answer: [NA] Justification: The core method development in this research does not involve LLMs as any important, original, or non-standard components

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-15T22:17:44.975973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:17:44.975973Z digest=sha256:b366c50926f2e34f4536bc7c505b069947bf8fd711aaf8bacf71057b4594902b

Observation cfe2f0be-f2a0-4a8f-adad-99000d3e84f3 · outbound

This paper cites an unresolved cited work.

Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent Unresolved cited work

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-15T22:17:44.807539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:17:44.807539Z digest=sha256:2dedff96580cac3791665c087df6f836c5960f2ab411b98268f0888fb79b4eda

Pith citing papers

Observation 5dfb4f5f-69a8-4cba-aecf-721da6f12517 · inbound

From Web Search towards Agentic Deep Research: Incentivizing Search with Reasoning Agents cites this paper.

From Web Search towards Agentic Deep Research: Incentivizing Search with Reasoning Agents Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T18:47:14.522738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:47:14.522738Z digest=sha256:eeb5190c05921bfc75775434449b3ff178be08ac4a083e635849c0852ab6f996

Observation 62e38eaa-d5a7-4027-b14f-a8ba8a6d8d69 · inbound

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning cites this paper.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.132290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.132290Z digest=sha256:c53cc2a8cbb3a1ab4820d1f10920403f0c515a97efcd82d59c5b0fab178e357a

Observation 6079eaaa-96a9-41c0-bb31-bd7465ac2de2 · inbound

Toward Efficient Agents: Memory, Tool learning, and Planning cites this paper.

Toward Efficient Agents: Memory, Tool learning, and Planning Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-03T09:21:36.356795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T09:21:36.356795Z digest=sha256:55785be8aff3fc8108fe1d1681cd5953790f77b3c850a531846b6f8d463f75cf

Observation 950293bd-e821-4423-ac2a-6214b052613d · inbound

Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application cites this paper.

Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:58:02.961702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-27T09:46:30.702256Z digest=sha256:0e0b1e1a84aee9262e6c1c1fe2ac592d296b06bee056a8161c013b5b91dd4457

Observation 3d71ee6f-d6f0-4204-bf38-4d3da85e4aab · inbound

DocArena: Turning Raw Documents into Controllable Training Environments for Document Search Agents cites this paper.

DocArena: Turning Raw Documents into Controllable Training Environments for Document Search Agents Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T12:53:26.525307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-29T12:50:16.625077Z digest=sha256:79fcf5422adb1b7c390ab46e291ea1ed07e5d377c82f4318cbb16317085af582

Observation 8ea84bdf-1eed-499a-b34c-00cfb0b051c6 · inbound

KbSD: Knowledge Boundary aware Self-Distillation for Behavioral Calibration in Agentic Search cites this paper.

KbSD: Knowledge Boundary aware Self-Distillation for Behavioral Calibration in Agentic Search Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T06:04:21.201378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T06:02:12.866129Z digest=sha256:31453b18363cff2275deaafc64436951ab7981b7f32acb81e9c0cdeccd6f612f