Pith. sign in

Paper Citation Record · LEDGER

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

As of 19 August 2026, this Paper Citation Record lists 71 of 71 outbound references and 15 inbound Pith citation observations for arXiv:2507.21836.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.21836 v1

Coverage vector

measured 71 of 71 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T12:23:55.366230Z

measured 86 of 86 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T23:57:35.839876Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

71 of 71 outbound references displayed

  • verified exact2
  • verified fuzzy5
  • unresolved64
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 1649d66e-f102-48f1-8b9c-95e736f8a63a · outbound

This paper cites H.; Meade, N.; and Reddy, S.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning H.; Meade, N.; and Reddy, S

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:23:56.496044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.032507Z digest=sha256:7fed78a35a77cb5489ec094bda2157ae5bb687d71f653f858ff428d3d1f6bdc0

Observation d971dcd4-8e1e-40ad-a97b-16364f1c9a1f · outbound

This paper cites ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.037983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.037983Z digest=sha256:38454eb8ae44278a1415d0343bf5e1c9832e772264a50c2042847ebaf1e9fb54

Observation a5bf6c4d-a3ea-43e3-aaaf-358b2689ce40 · outbound

This paper cites SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.043194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.043194Z digest=sha256:1553b06e5392fefce3b35265921c33037ba77ccde58829e949cce800d138d98b

Observation 004e29cd-6747-4537-9b98-15ed8151193a · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Training Verifiers to Solve Math Word Problems

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.048401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.048401Z digest=sha256:c0b34dd87c763c7e1d08d3238d89a5cd3d443635d2be359d3e455fccadce1af3

Observation 667090f5-b2be-4c70-92c3-c4c6f39f9513 · outbound

This paper cites Process Reinforcement through Implicit Rewards.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Process Reinforcement through Implicit Rewards

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.054695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.054695Z digest=sha256:af73345980f39ca8400a615c7c83007efcfe34227c89188675033ed1a0313ce6

Observation 723ab1c7-9245-4c8b-bad1-c1a073751a10 · outbound

This paper cites MATHSENSEI: A Tool-Augmented Large Language Model for Mathematical Reasoning.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning MATHSENSEI: A Tool-Augmented Large Language Model for Mathematical Reasoning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.060523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.060523Z digest=sha256:f0aa9c13a76bed5d0f8034434286cf24f132e49271b0274f60acb89f49db72c3

Observation 61e64d85-0ef9-4e7b-8d37-94c49eb4807c · outbound

This paper cites Tool-Star: Empowering LLM-Brained Multi-Tool Reasoner via Reinforcement Learning.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Tool-Star: Empowering LLM-Brained Multi-Tool Reasoner via Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.066336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.066336Z digest=sha256:708ec126fbc76d4a00de4fb9fe46330a3464a7ef78cb29d50b6fdfbb4d25727d

Observation 722b34d0-45fd-4132-b46f-6e2fefe6d266 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.481262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.071475Z digest=sha256:798ffc778db66901ec14d81e23e1c6965bcbefd78068d096c695f4b8d905c36b

Observation cd292c23-6eaa-4846-ae2f-6f785941d014 · outbound

This paper cites ReTool: Reinforcement Learning for Strategic Tool Use in LLMs.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning ReTool: Reinforcement Learning for Strategic Tool Use in LLMs

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.076366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.076366Z digest=sha256:c173c66d3f61cf0ac4fd3f81b0f07b5a92c26e5eb119f75d0712a33d5602c1f6

Observation 506720b2-0c4a-4ad0-8c3c-17e67987324b · outbound

This paper cites MT-R1-Zero: Advancing LLM-based Machine Translation via R1-Zero-like Reinforcement Learning.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning MT-R1-Zero: Advancing LLM-based Machine Translation via R1-Zero-like Reinforcement Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.081531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.081531Z digest=sha256:152930d7728abe9d25e5da685b5584325e96978052de5a56a7337d32c4e0e5aa

Observation f4143950-2e76-456e-a840-66278d750e9e · outbound

This paper cites Scaling Reasoning, Losing Control: Evaluating Instruction Following in Large Reasoning Models.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Scaling Reasoning, Losing Control: Evaluating Instruction Following in Large Reasoning Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.086609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.086609Z digest=sha256:3bf38386c1951fab86cacf4bdaf37137097cf81803f5d2336b0838ca77674c96

Observation c66ce643-c9e4-4bf4-b988-43e83aea6c5d · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.466437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.091473Z digest=sha256:7619221eb7b6f3c49519ae3f19c2883786efca4a99b4df2a5f5772545dd89889

Observation 1777e045-0a18-4569-bc3d-11b104450a00 · outbound

This paper cites ToRA: A Tool-Integrated Reasoning Agent for Mathematical Problem Solving.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning ToRA: A Tool-Integrated Reasoning Agent for Mathematical Problem Solving

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.096057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.096057Z digest=sha256:1d64c46a0e4955d4c8e11425bff57b589909f6f327faa410a46914fbb92a2586

Observation 051459f2-a29c-46cb-b473-9df3b54f74bd · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.101180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.101180Z digest=sha256:0d045daf599f328f988ca7bfd5aec7c2542ca5d47e50af73bab3dfc18d9a1be0

Observation 0cb9e14f-a257-455c-a47d-2bd7111d30a0 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.451211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.106630Z digest=sha256:72645f2f9cb4f4e48a9917053b82aea9cc90bcddd3d5e247a83a757db4d53a87

Observation 153d75da-1955-4df7-92ee-b8e74148f85e · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.435572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.112461Z digest=sha256:375a6e23d69a9c7125a7d16bdf24d0cfde84be3babd2ce47771e5bcc6e1e27f4

Observation b003dfd8-33e6-4e16-b24e-abcd817e0a4f · outbound

This paper cites D.; Sugawara, S.; and Aizawa, A.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning D.; Sugawara, S.; and Aizawa, A

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:23:56.419486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.117576Z digest=sha256:89c376635f322ef22e9031abf18ccc71e7467d01b22cb029f672d11421368f86

Observation 6c5ed4f8-2840-48dd-ad42-ea2489c879a8 · outbound

This paper cites D.; Phan, D.; Dohan, D.; Douglas, S.; Le, T.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning D.; Phan, D.; Dohan, D.; Douglas, S.; Le, T

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:23:56.401669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.122505Z digest=sha256:18ea193ae6c542de331c6bc9867e3ee2844dce3e8367c8af7cdb4cfbab14eecb

Observation 6c731bc9-f4b8-4805-8e1e-c245c8c2aa3c · outbound

This paper cites Xolver: Multi-Agent Reasoning with Holistic Experience Learning Just Like an Olympiad Team.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Xolver: Multi-Agent Reasoning with Holistic Experience Learning Just Like an Olympiad Team

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.127390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.127390Z digest=sha256:a7de781359a9bfa798fa062b24a67925fbb6792c3ce643e016361956d5ed2cda

Observation 62e38eaa-d5a7-4027-b14f-a8ba8a6d8d69 · outbound

This paper cites Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.132290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.132290Z digest=sha256:c53cc2a8cbb3a1ab4820d1f10920403f0c515a97efcd82d59c5b0fab178e357a

Observation 0d4d994f-ad08-48cc-8053-f3617a23e119 · outbound

This paper cites Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.137421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.137421Z digest=sha256:baed5f43f06e8fa92895a3d16a1ff2f805fe32b511aafb74017e070536aa06e3

Observation 5447aed0-7fdc-43fa-81b1-6cd23b228ad6 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.383410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.142335Z digest=sha256:185932300742ade2fd1a43b1910d42e0a3967e95ad2db01da3548711fe84128f

Observation 05d35a0f-afb0-4399-b1e3-d4cab0135aa1 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.367171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.147454Z digest=sha256:b091c4b233dd3c3493f1c5b87efed3b4ffc7a8710bd6a665fde9477c2d331f3a

Observation 38425f2b-6272-4f3d-a5e3-28f7dd2d9507 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.350634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.152237Z digest=sha256:91e045d046f07f497aff7be711089f12669708ccb7b427425b129e3bf10b7f39

Observation 6a646038-8ee4-413d-a456-803f07f20f48 · outbound

This paper cites Tulu 3: Pushing Frontiers in Open Language Model Post-Training.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Tulu 3: Pushing Frontiers in Open Language Model Post-Training

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.157210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.157210Z digest=sha256:009baedaa09e7703bd2fe3b3f08428b123c38569549b3b4f6ed42069746748eb

Observation 4074e7f4-6c65-4f28-9f79-343206e695be · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.333903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.162520Z digest=sha256:bdb96983866fc40d941f9cf420fcd8a239519ca81bc5c22edcb4313fb78817bd

Observation 9a5097c2-5f06-4905-9075-95d6fb402a0a · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.316621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.166741Z digest=sha256:d54c98c4a08cbf69a293b3e522699e097e012a73b221894426100fd45cc53bd2

Observation 74af3947-339f-4281-8dfb-911079c03537 · outbound

This paper cites When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.171218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.171218Z digest=sha256:abb5fe2f42cce4ed079537186616d5f79131c1fc447b7c4ac9b03847ea318fdc

Observation 4c924188-6d0e-4605-9afb-89eb632f15d7 · outbound

This paper cites ToRL: Scaling Tool-Integrated RL.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning ToRL: Scaling Tool-Integrated RL

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.175839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.175839Z digest=sha256:030b96612676f99a91cfc073b6cbcbcd8c97ea4c3716146ee3c9ece89f06c0f3

Observation 919ad49d-462a-47a8-a599-44f0fcfc8d33 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.300788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.180432Z digest=sha256:0c739bddfb81abe2a9e06b63548028423fac248387fe3b3c403c00ade9b0f653

Observation 0528640e-ea62-4b24-9fe0-e14742311a38 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.285005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.184726Z digest=sha256:1ca9c4f307c886a1869a7f1db9fba254724ac874f9e5d01863b95dcb8f9a299b

Observation a71d6dd3-a094-4582-b908-a31ccb4c8fee · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.267602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.189197Z digest=sha256:4200e26379cdbf7c610af7c05af60d69651942e1fae15b035bc520e15d116a19

Observation 1d714475-8228-49c5-acb7-c2b3f21ecbed · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.250233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.193451Z digest=sha256:46cdcda1fc52a82e8c22ea672ac25feda6f9248c19c7469fcc852a17babee19d

Observation 180f9b04-4a4c-4050-8278-f640fa91051b · outbound

This paper cites Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.198073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.198073Z digest=sha256:02d2e3555ca33472f6f5522f0bd9e8585fe91d1f5404a1c5da283cd3b7f872a4

Observation 6f59724b-f410-45ca-86f8-84102e88672e · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.202714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.202714Z digest=sha256:2e094e7f6854fe6fa91fbdf85fd1d25e181ec44282ea017f3a272e4c4fb256e7

Observation 7a388184-643a-4935-a613-f8e0d224f6b4 · outbound

This paper cites Large Language Models: A Survey.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Large Language Models: A Survey

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.207383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.207383Z digest=sha256:ee8dd2acf93684b5d89b86d2d49ed2edd9f7e2ae5a93fd2f64259c2697213ca1

Observation d2a83e75-6cef-4e21-8055-4f24690cc144 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.219104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.212112Z digest=sha256:79c66cd118ad860d8aa73871ba1e3658a07242a4130fbdf79f8eab41e9cd9dcf

Observation a4ae1eb0-31bc-45d5-b200-cc97b6efef60 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.216656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.216656Z digest=sha256:60b4c9ce8bb9a9510e2b669993866e899b63009c678551f5d8f2778c793a5135

Observation dcec9936-1e97-4b00-aca9-930e2feb9e69 · outbound

This paper cites A.; and Lewis, M.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning A.; and Lewis, M

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:23:56.191248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.221344Z digest=sha256:2f645d1b0304b06f13999b49bc1d09283d8c52b9133af2b7f45ad53af84c24c8

Observation ded6630f-1569-4a74-8d97-7b33ed584986 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.174745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.225783Z digest=sha256:730695f4f42fe95440412a4942f2b8a90a3c297a49966b5d997ea359125c2aba

Observation 9f4d5364-552d-49dd-8964-6fa7027d9c82 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.158300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.230434Z digest=sha256:2f7c98771c075e837547c9e6419f351a85996103513c1055c4945b49c7f61105

Observation db7c884a-9ba1-46c0-8ddc-4171e6432d93 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.143058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.235001Z digest=sha256:b9e6427fe8f77da37adea8c64c5b5494dcf4fc19f93c38a911243bd6d1def475

Observation 14dee0d9-eb93-4f7a-ba8f-46497b18493d · outbound

This paper cites D.; Ermon, S.; and Finn, C.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning D.; Ermon, S.; and Finn, C

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.239596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.239596Z digest=sha256:55a2725334ae57d89cf444e432acdce23ae9c61c4cec6b497d7802375542637b

Observation 4ab36d37-786e-4545-9b60-4a0dd497893a · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.244037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.244037Z digest=sha256:5695fc432b7c9dbf8034e82ec562dccb5b84b9454567e8e6998165eb82a42b25

Observation 91440656-8a7c-4351-8cb4-5444c2e6743f · outbound

This paper cites Proximal Policy Optimization Algorithms.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Proximal Policy Optimization Algorithms

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.248574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.248574Z digest=sha256:03dcb24a0cf86794dbb834e4363404feec40df1f0be0f86a3c42126bcc852d47

Observation ec5e9d4d-52de-4882-a99e-3a50f9b73404 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.106655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.253394Z digest=sha256:3a8b7bde34645b6670e54bfa7ff2a84a8dfcef353faded7ce1ff43244a30870c

Observation 75f9ce17-0646-4097-81cc-c9f357540fff · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.257750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.257750Z digest=sha256:9f83f6cacaf5c11a5cd2598d9030cb52f2802486126edc6074838651515b3888

Observation 5dacf8b6-d784-4531-8e85-b9387d26ae5f · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.262322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.262322Z digest=sha256:be554bf882400a7eb3f98d5aeb29dc0ab52af9a73cf449aecda9fe1172f26784

Observation 0fb7dae3-9ebc-459a-bbb2-487fe04d0a30 · outbound

This paper cites R1-Searcher: Incentivizing the Search Capability in LLMs via Reinforcement Learning.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning R1-Searcher: Incentivizing the Search Capability in LLMs via Reinforcement Learning

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.266683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.266683Z digest=sha256:944f70bca998f97e558b83441ae1f2d8c8a1ef883c4d0f9fb5c249846cfcb2bf

Observation 7a829d1d-6953-4abc-9344-af55a7addeb6 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.081150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.271585Z digest=sha256:0b25ae1c5a65ebd042c076a14c08dd133615a7ce19e2bd94f45d0d4d3c33b788

Observation 930de274-7d92-4bba-afa9-7a90c589e8c4 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.065674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.275706Z digest=sha256:eb8fe4089f88bf7e7973ca7a4797ca5dde6f2ddad29598b2e72749f39ce83212

Observation a227ce7d-db50-46a7-8b24-353e9e43bc3a · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.279697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.279697Z digest=sha256:7a01278bd7d72016edda2160a5f6fb5f46f4ebffda5b395502ace79e20aabd70

Observation 3977e03d-20a8-4cd6-8f54-dbbbe0ecf6d9 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.035175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.283718Z digest=sha256:19434b1f3a6a2d856659bcc7928d94d7ede507b855448a337e2bf845d2e7701e

Observation 87884ee7-f47c-4fa4-820c-7f8c3f3aabc9 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.016815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.288398Z digest=sha256:4f88f872351f068790ffb8f328ad69b046371b9d496b465ecb9b30cb275f7512

Observation 0fd5e305-96e1-4806-8b5f-7d0512d9a7ac · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:55.999918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.292695Z digest=sha256:5f554aa145e2c13ddd459e8982f95bb0fa389c876f8feb0e501b3e65fc196c6d

Observation 6dd7b948-b9de-43da-ab64-09920117af32 · outbound

This paper cites Text Embeddings by Weakly-Supervised Contrastive Pre-training.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Text Embeddings by Weakly-Supervised Contrastive Pre-training

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.296956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.296956Z digest=sha256:64f058d372305a917e41242edb322db29e63232cbc28b333de32433c29633e36

Observation 2169dfa0-4492-4921-b001-ead51af07151 · outbound

This paper cites MenatQA: A New Dataset for Testing the Temporal Comprehension and Reasoning Abilities of Large Language Models.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning MenatQA: A New Dataset for Testing the Temporal Comprehension and Reasoning Abilities of Large Language Models

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-08-06T12:23:55.500019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.301646Z digest=sha256:e2c2cbf11c58dac36aeb4c7968f93e0eca199bd395c9e69566f59ee0113a66d1

Observation deb46779-f24f-42fa-8baa-c1861eaf09da · outbound

This paper cites Structural Entropy Guided Agent for Detecting and Repairing Knowledge Deficiencies in LLMs.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Structural Entropy Guided Agent for Detecting and Repairing Knowledge Deficiencies in LLMs

Reference 58

Resolution
verified exact
local_arxiv, observed 2026-08-06T12:23:55.477185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.307048Z digest=sha256:797247e29ae373fbac6b303c61245d11512fd49abefdec57d19a49032e0f1f05

Observation 9c056e49-0b23-4e46-84e5-0772bcc0b99f · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:55.983520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.311727Z digest=sha256:111432f01634e0d3bb32863cf013d7949d24b585f015185f653ca46e4c7dcafd

Observation b6dce6f5-72b0-4a74-9bc9-a36e24b926d2 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:55.968767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.316038Z digest=sha256:9efeab74f8d0155e2eaad645dc97a2705d6d596b9bd53f05ec9156b271e18eb0

Observation 79dfff32-e5a3-4db3-b580-4d25e33b316a · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 61

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:55.952044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.320345Z digest=sha256:abbd96817850e1d087cf64170cefddc9ec9c076f60be389370a5ade6daca7bd9

Observation a2a71f0e-8640-4bf5-9bea-2f4fa965e8aa · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.324339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.324339Z digest=sha256:10a6f8d10b8bc5bd5f6511559acccb11b8a68321ba5c84ff5e303bee9fc3d6a4

Observation bfd1390e-235f-4271-b985-235d5a08d6d6 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:55.926226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.328625Z digest=sha256:6eb0008d54d230e6c37d04a22625e3ae80dc24654d28619e714302931a6bb48f

Observation 512c3165-da91-4b50-ae51-162ed3e576f1 · outbound

This paper cites C.; Kaddar, Y.; Blunsom, P.; Staton, S.; and Gal, Y.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning C.; Kaddar, Y.; Blunsom, P.; Staton, S.; and Gal, Y

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:23:55.910593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.332803Z digest=sha256:b7ac99eea8a4230eeda3930baa874e13e923506c1e3555298c6d4c4fa354e8a6

Observation 681164e1-4561-4ded-87fc-f8956956aa80 · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.337141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.337141Z digest=sha256:44f5b2b902516390d7dc184e91449de48e82ab1d2abc0668281bd0c9e5c3bbb4

Observation 9af55f1e-16aa-4fac-8920-52e7930facdb · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.342076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.342076Z digest=sha256:19f1e1afbd7f0e64db909a7a71769c22a4a191a7b02e8d13210ad589be751c78

Observation c36f7776-b63e-49e8-9354-a906a54ca132 · outbound

This paper cites SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.346347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.346347Z digest=sha256:6e49640e1338c22b08a4c7ce29087201634d5495e40bca21b523e61c7eddf820

Observation 7c79cf1b-5bcc-4831-a59e-0453df71363e · outbound

This paper cites Instruction-Following Evaluation for Large Language Models.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Instruction-Following Evaluation for Large Language Models

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.351954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.351954Z digest=sha256:a4df63a32bb87faffb7e5998ce22201c3a52fa9f9a355774edc38cab13551126

Observation 9d1213f1-8b9f-4d62-8751-4accbfd81558 · outbound

This paper cites MEM1: Learning to Synergize Memory and Reasoning for Efficient Long-Horizon Agents.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning MEM1: Learning to Synergize Memory and Reasoning for Efficient Long-Horizon Agents

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.356601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.356601Z digest=sha256:7f9be46c3dc618dd3d2fe77177fa8e5cb47fc808aff3e44e51c2f39800a3c796

Observation 81019d0f-7131-47a8-a919-75900c09714b · outbound

This paper cites , " * write output.state after.block = add.period write newline.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning , " * write output.state after.block = add.period write newline

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.361388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.361388Z digest=sha256:91d9011fc0083e70d22390a320a568d21b2d017b83715c56904eccbc3251a8d5

Observation dab0f0fa-d2db-4448-9f87-dabbe20ebe69 · outbound

This paper cites write newline.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning write newline

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.366230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.366230Z digest=sha256:389603c44ebb35848cb0100b3eeaaadbd30f22e32551d22915688dcc6f8c7664

Pith citing papers

Observation be2f730c-502f-49ba-afd5-8c9677e23336 · inbound

Chain-of-Agents: End-to-End Agent Foundation Models via Multi-Agent Distillation and Agentic RL cites this paper.

Chain-of-Agents: End-to-End Agent Foundation Models via Multi-Agent Distillation and Agentic RL AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-05T23:57:35.839876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:57:35.839876Z digest=sha256:54da19906f7a84f023db003493ca30a778ff980d29b1e0967b0de4a74c8fdeaa

Observation 1ab86500-5dbb-4b41-98a7-8edf3269a992 · inbound

Learning How to Use Tools, Not Just When: Pattern-Aware Tool-Integrated Reasoning cites this paper.

Learning How to Use Tools, Not Just When: Pattern-Aware Tool-Integrated Reasoning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T14:49:51.602433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:49:51.602433Z digest=sha256:fba4418dc3f3913144d6a57a17e36261d6b87a866116e907077a7209b59c5131

Observation 9bce88ff-83e7-4b1d-ac9e-3ded845d3cf5 · inbound

SIGMA: Search-Augmented On-Demand Knowledge Integration for Agentic Mathematical Reasoning cites this paper.

SIGMA: Search-Augmented On-Demand Knowledge Integration for Agentic Mathematical Reasoning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T06:57:16.012421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:57:16.012421Z digest=sha256:4b3a156f39b658ac2bc737f570cef762b079c37909b1a5eb68477270a6b096b6

Observation db6d5a77-182a-41f6-9d2e-2679ff59999e · inbound

Lost in Execution: On the Multilingual Robustness of Tool Calling in Large Language Models cites this paper.

Lost in Execution: On the Multilingual Robustness of Tool Calling in Large Language Models AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T11:44:07.353535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T11:44:07.353535Z digest=sha256:c39310ad55107f367f6d8e316793cdbb840b571c83134431af2056a0cf65cc63

Observation 728c2d82-7c0c-4661-ae86-f1a6a9442a46 · inbound

Toward Efficient Agents: Memory, Tool learning, and Planning cites this paper.

Toward Efficient Agents: Memory, Tool learning, and Planning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 142

Resolution
unresolved
no resolver link, observed 2026-08-03T09:21:45.096788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T09:21:45.096788Z digest=sha256:4bc37467a76e809767e205e100783174875a45803540711d763a88479fe9af06

Observation ac93b3b1-bcd6-4ace-9141-1a7860485d3f · inbound

Reasoning and Tool-use Compete in Agentic RL:From Quantifying Interference to Disentangled Tuning cites this paper.

Reasoning and Tool-use Compete in Agentic RL:From Quantifying Interference to Disentangled Tuning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-03T05:56:57.193714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:56:57.193714Z digest=sha256:48a27332f6a5cd7708ee521eb6c001585b917e254f66a8f93d8ca73265f04e13

Observation 1f3e87f2-3491-4609-a0ec-8b177e8d9f8f · inbound

AnomalyAgent: Agentic Industrial Anomaly Synthesis via Tool-Augmented Reinforcement Learning cites this paper.

AnomalyAgent: Agentic Industrial Anomaly Synthesis via Tool-Augmented Reinforcement Learning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-10T18:35:42.825362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T18:32:31.955785Z digest=sha256:0030d60d710586bc0641f94c044899d43d122a7e8e580c04b53ac7dd8d93d0cd

Observation d0287446-3114-4176-936c-605bc0cd6bad · inbound

See Further, Think Deeper: Advancing VLM's Reasoning Ability with Low-level Visual Cues and Reflection cites this paper.

See Further, Think Deeper: Advancing VLM's Reasoning Ability with Low-level Visual Cues and Reflection AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:36:17.711402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-08T04:46:16.497585Z digest=sha256:a8afc0fb635f080c6bd8647aacc5f99adc828af0a3650db99a1f57e307077b8d

Observation e63939b4-76ff-4201-b75d-125c52cf8ac7 · inbound

Tools as Continuous Flow for Evolving Agentic Reasoning cites this paper.

Tools as Continuous Flow for Evolving Agentic Reasoning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-11T01:45:51.740826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-11T01:30:19.859374Z digest=sha256:aed634427c74bd8aa05f1e3eb71936ad45cb5091b511c72f8412a75219b1eb5f

Observation 9f34793a-d609-49d1-806b-6a4190184c37 · inbound

SkillMaster: Toward Autonomous Skill Mastery in LLM Agents cites this paper.

SkillMaster: Toward Autonomous Skill Mastery in LLM Agents AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:36:24.078934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-12T00:59:03.361037Z digest=sha256:98ca685d78ac40903eba630120f47fea238bfd5bc05565bb15cb6bb9413ec90c

Observation a3d8a207-1068-45b0-b0d7-6ec3032c8f74 · inbound

SkillMaster: Toward Autonomous Skill Mastery in LLM Agents cites this paper.

SkillMaster: Toward Autonomous Skill Mastery in LLM Agents AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:32:29.862504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T07:30:09.985448Z digest=sha256:004588870457d962e1376e97917adac9a163fb871a69191a208ce27a82f1e767

Observation feb34b55-7224-4cf2-8cbb-144ff2b23b32 · inbound

TIDE-Bench: Task-Aware and Diagnostic Evaluation of Tool-Integrated Reasoning cites this paper.

TIDE-Bench: Task-Aware and Diagnostic Evaluation of Tool-Integrated Reasoning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:31:23.948998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-12T02:47:31.830410Z digest=sha256:0bc28ba63aee68a222ca9970ecbd605163fae753b8d3a004001e9ca51ee9808d

Observation d0e8bf0e-7946-4e2d-aca4-9c72cdbb2022 · inbound

Learning When Not to Act: Mitigating Tool Abuse in Agentic Reinforcement Learning cites this paper.

Learning When Not to Act: Mitigating Tool Abuse in Agentic Reinforcement Learning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:46:20.257307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-28T14:59:57.983338Z digest=sha256:fddebe5b8b5bcdc1eadf85111c53d34ecc8951deca213883cd4f658d068f7112

Observation 60831c2a-f04d-4715-80d4-9105dcd716b8 · inbound

Modularized Reinforcement Learning on LLMs: From MDP Creation to Exploration and Learning cites this paper.

Modularized Reinforcement Learning on LLMs: From MDP Creation to Exploration and Learning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 223

Resolution
verified exact
arxiv_id, observed 2026-07-04T08:09:40.721699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-26T12:15:08.304150Z digest=sha256:3ed29e5f7072e26ace57cdd597dd5dafd69fd8f681bb3a0f761c5df1c5442f8a

Observation 998ceb97-64bb-4584-b4fa-38fd5bc6ced9 · inbound

TurnSight: Turn-Level Hindsight Self-Distillation for Tool-Integrated Reasoning cites this paper.

TurnSight: Turn-Level Hindsight Self-Distillation for Tool-Integrated Reasoning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T04:18:15.288921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T04:18:15.288921Z digest=sha256:38eb449b23ec1dc74a121fbcd05b04d810374bdd5113238b39fa4ac5be086d4c