Pith. sign in

Paper Citation Record · LEDGER

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

As of 8 August 2026, this Paper Citation Record lists 71 of 71 outbound references and 15 inbound Pith citation observations for arXiv:2507.21836.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.21836 v1

Coverage vector

measured 71 of 71 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T12:23:55.366230Z

measured 86 of 86 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T23:57:35.839876Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

71 of 71 outbound references displayed

  • verified exact2
  • verified fuzzy5
  • unresolved64
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 1649d66e-f102-48f1-8b9c-95e736f8a63a · outbound

This paper cites H.; Meade, N.; and Reddy, S.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning H.; Meade, N.; and Reddy, S

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:23:56.496044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.032507Z digest=sha256:d46730cc475db77dc3ef7fb99d7981850c7e5535b53ce15d73e2d0da090beec1

Observation d971dcd4-8e1e-40ad-a97b-16364f1c9a1f · outbound

This paper cites ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.037983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.037983Z digest=sha256:50b9719c3418a26946c6d8f136ce7dbe1b04629fd1fdbec8e1a60d59b729273a

Observation a5bf6c4d-a3ea-43e3-aaaf-358b2689ce40 · outbound

This paper cites SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.043194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.043194Z digest=sha256:99957faed1f4dcd688e85c2f42117a501d652392a3c9bb97a644a0e8265df0dd

Observation 004e29cd-6747-4537-9b98-15ed8151193a · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Training Verifiers to Solve Math Word Problems

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.048401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.048401Z digest=sha256:9db51fbbcd8c066f46be2c0eb7aaa57cfd08d0918a9a1de2681c76bd2d5420e3

Observation 667090f5-b2be-4c70-92c3-c4c6f39f9513 · outbound

This paper cites Process Reinforcement through Implicit Rewards.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Process Reinforcement through Implicit Rewards

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.054695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.054695Z digest=sha256:10db8303e5de1a2d42fb9b8f63bfd8083828dab5fa3c814b6e53d6a71d664081

Observation 723ab1c7-9245-4c8b-bad1-c1a073751a10 · outbound

This paper cites MATHSENSEI: A Tool-Augmented Large Language Model for Mathematical Reasoning.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning MATHSENSEI: A Tool-Augmented Large Language Model for Mathematical Reasoning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.060523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.060523Z digest=sha256:4a9bf301f8196bc87bee5f20a16568097ead57984d24f0cf6f612094d351505e

Observation 61e64d85-0ef9-4e7b-8d37-94c49eb4807c · outbound

This paper cites Tool-Star: Empowering LLM-Brained Multi-Tool Reasoner via Reinforcement Learning.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Tool-Star: Empowering LLM-Brained Multi-Tool Reasoner via Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.066336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.066336Z digest=sha256:c0cd0a80a5a1efae97322ef69408e90de86cea1004026c8f047a1c745c4af904

Observation 722b34d0-45fd-4132-b46f-6e2fefe6d266 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.481262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.071475Z digest=sha256:645d22f49a1bfe2a5759df9682fa0a4cd08c0dd761c6a40cd6623c4c5686e89f

Observation cd292c23-6eaa-4846-ae2f-6f785941d014 · outbound

This paper cites ReTool: Reinforcement Learning for Strategic Tool Use in LLMs.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning ReTool: Reinforcement Learning for Strategic Tool Use in LLMs

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.076366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.076366Z digest=sha256:0833493bd9274a0f0e10123fac1a8a1bf517ae45cdcaafb161ed4ef11c21a27a

Observation 506720b2-0c4a-4ad0-8c3c-17e67987324b · outbound

This paper cites MT-R1-Zero: Advancing LLM-based Machine Translation via R1-Zero-like Reinforcement Learning.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning MT-R1-Zero: Advancing LLM-based Machine Translation via R1-Zero-like Reinforcement Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.081531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.081531Z digest=sha256:e57440d5185024fa23d9a36c413ce812b12aa606f8c212989a0029fe922fde38

Observation f4143950-2e76-456e-a840-66278d750e9e · outbound

This paper cites Scaling Reasoning, Losing Control: Evaluating Instruction Following in Large Reasoning Models.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Scaling Reasoning, Losing Control: Evaluating Instruction Following in Large Reasoning Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.086609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.086609Z digest=sha256:91d402c449c98f36b7ca51188e6f66a999724cbc399e670ee20bd487cb2726e8

Observation c66ce643-c9e4-4bf4-b988-43e83aea6c5d · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.466437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.091473Z digest=sha256:3246e593fd52ffd1abe55829cc0dce479d4c7f2869a1268bff7f49198f7ea0e3

Observation 1777e045-0a18-4569-bc3d-11b104450a00 · outbound

This paper cites ToRA: A Tool-Integrated Reasoning Agent for Mathematical Problem Solving.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning ToRA: A Tool-Integrated Reasoning Agent for Mathematical Problem Solving

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.096057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.096057Z digest=sha256:a1897e69fe23c16e81628e9ad3492597836f3dd1ba034429964b52813e281c42

Observation 051459f2-a29c-46cb-b473-9df3b54f74bd · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.101180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.101180Z digest=sha256:c3c34cf7ddd901981bc109ea512a4ed8761d4c83bb7a1e5060ab43d48b8f72a6

Observation 0cb9e14f-a257-455c-a47d-2bd7111d30a0 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.451211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.106630Z digest=sha256:d2c6d50c370c2013bb32b8ba1acd06c123a74386761dbce44c53ae5c1cc7025e

Observation 153d75da-1955-4df7-92ee-b8e74148f85e · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.435572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.112461Z digest=sha256:572c104d56d0296b69478a9427b630cf1f348f13489f26409a4090927b1ef3c1

Observation b003dfd8-33e6-4e16-b24e-abcd817e0a4f · outbound

This paper cites D.; Sugawara, S.; and Aizawa, A.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning D.; Sugawara, S.; and Aizawa, A

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:23:56.419486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.117576Z digest=sha256:6681da5cebf4517876bb1214879dee0c293f5b8d7f0a5aa3ea9e028875b69284

Observation 6c5ed4f8-2840-48dd-ad42-ea2489c879a8 · outbound

This paper cites D.; Phan, D.; Dohan, D.; Douglas, S.; Le, T.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning D.; Phan, D.; Dohan, D.; Douglas, S.; Le, T

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:23:56.401669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.122505Z digest=sha256:eebc1d61e76768e1b2b79c8f1b4f13ea890245c159bdefd33d2df2069f717391

Observation 6c731bc9-f4b8-4805-8e1e-c245c8c2aa3c · outbound

This paper cites Xolver: Multi-Agent Reasoning with Holistic Experience Learning Just Like an Olympiad Team.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Xolver: Multi-Agent Reasoning with Holistic Experience Learning Just Like an Olympiad Team

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.127390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.127390Z digest=sha256:e61fb0076c2a831a31d992025c1de96d20bd2222c12038941152419f00b0af70

Observation 62e38eaa-d5a7-4027-b14f-a8ba8a6d8d69 · outbound

This paper cites Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.132290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.132290Z digest=sha256:c0cc676fd373a8b120aac2c06ff6a0b52691794382d47b20bbd98e9102527d6a

Observation 0d4d994f-ad08-48cc-8053-f3617a23e119 · outbound

This paper cites Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.137421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.137421Z digest=sha256:5550a223414fc3bb8d6906fa1279cafc1d7491127320a2c65881c1781653f648

Observation 5447aed0-7fdc-43fa-81b1-6cd23b228ad6 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.383410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.142335Z digest=sha256:f6b9fa95b74d35e79893d5898885ccce3003b2ba0ac817674769f2eb15c4e87f

Observation 05d35a0f-afb0-4399-b1e3-d4cab0135aa1 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.367171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.147454Z digest=sha256:e82ad51773802474b97ae96574fc95798d577801d1e656ea695169808d000aeb

Observation 38425f2b-6272-4f3d-a5e3-28f7dd2d9507 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.350634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.152237Z digest=sha256:8ace9d6e15ab398ae33ab47ebe416fc1e3f374f503a078eea035f77c4d14d133

Observation 6a646038-8ee4-413d-a456-803f07f20f48 · outbound

This paper cites Tulu 3: Pushing Frontiers in Open Language Model Post-Training.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Tulu 3: Pushing Frontiers in Open Language Model Post-Training

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.157210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.157210Z digest=sha256:6188b61b522d524c5c8f99de4d2beed90b006bcf6705c7d8818a1ccc7a362ccd

Observation 4074e7f4-6c65-4f28-9f79-343206e695be · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.333903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.162520Z digest=sha256:f89c9e7f0c67e61b4e3b15641737d5ee6571df2282248a5f0323c17b2bf511e7

Observation 9a5097c2-5f06-4905-9075-95d6fb402a0a · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.316621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.166741Z digest=sha256:b6cabdde689900300eb975ec0bed402cf6f9e2d1181d0577ccfbefc7b3775f11

Observation 74af3947-339f-4281-8dfb-911079c03537 · outbound

This paper cites When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.171218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.171218Z digest=sha256:6d6623d06a7cbf1be47cca3344dab28c11609dcf46f9d829c8074374d55821a1

Observation 4c924188-6d0e-4605-9afb-89eb632f15d7 · outbound

This paper cites ToRL: Scaling Tool-Integrated RL.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning ToRL: Scaling Tool-Integrated RL

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.175839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.175839Z digest=sha256:349dfc2a4b4e83b97e6d9bb62506d157b778c82d994143ced5df314d080ae94a

Observation 919ad49d-462a-47a8-a599-44f0fcfc8d33 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.300788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.180432Z digest=sha256:b6d474274ba402f53250254738f5787f81cc00a9b393c5aaeb7244ea2fcb301b

Observation 0528640e-ea62-4b24-9fe0-e14742311a38 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.285005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.184726Z digest=sha256:edd65d3f446942edbc6108674efacdc5a0a4395dc099ea0038f82e319c476c0b

Observation a71d6dd3-a094-4582-b908-a31ccb4c8fee · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.267602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.189197Z digest=sha256:269c911988fb58a160671c9f9a1df640bd1e1c05cf9d5aab0b12f9937ddfc06e

Observation 1d714475-8228-49c5-acb7-c2b3f21ecbed · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.250233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.193451Z digest=sha256:76bb0078706dbe8729c14b2e3deb9f34ddb133282eb12a731ad81d9b62cc92e3

Observation 180f9b04-4a4c-4050-8278-f640fa91051b · outbound

This paper cites Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.198073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.198073Z digest=sha256:8d667a853a520f74607e813833655863081e5adc57694c4a7c361885bb3e6354

Observation 6f59724b-f410-45ca-86f8-84102e88672e · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.202714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.202714Z digest=sha256:32dc875baf2900a600471e1bab4b6aa73170605774151f29fddb3d86c1d81374

Observation 7a388184-643a-4935-a613-f8e0d224f6b4 · outbound

This paper cites Large Language Models: A Survey.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Large Language Models: A Survey

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.207383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.207383Z digest=sha256:ac08002dbee3520669271b17bc2da0a8600534d344c772b9359bc0aa011e0f83

Observation d2a83e75-6cef-4e21-8055-4f24690cc144 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.219104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.212112Z digest=sha256:422c74c4cd76ea1b5fa3419ab53e4725eb7b1b7a2d81b42bdb551b75c30603bf

Observation a4ae1eb0-31bc-45d5-b200-cc97b6efef60 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.216656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.216656Z digest=sha256:d7553bfb3efffc594c70fdde1c2e1289dcb71f98d492c9aeacca9b3d78fd79d1

Observation dcec9936-1e97-4b00-aca9-930e2feb9e69 · outbound

This paper cites A.; and Lewis, M.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning A.; and Lewis, M

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:23:56.191248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.221344Z digest=sha256:8b1aba5ca7c0e1dc06f39bf5b17db15cbb264fee10687140f71c4846490c0f75

Observation ded6630f-1569-4a74-8d97-7b33ed584986 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.174745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.225783Z digest=sha256:84ff13b6b72444890d693324323563cae1a37bd080ccb4a1ac75bf8b2ba02fdd

Observation 9f4d5364-552d-49dd-8964-6fa7027d9c82 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.158300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.230434Z digest=sha256:19a0397461144458eb1af03b3daf0cfec44bc1fcc23e201db24e9508fcc09243

Observation db7c884a-9ba1-46c0-8ddc-4171e6432d93 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.143058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.235001Z digest=sha256:6cf50f035c91a6303bcb0bc57643467d2c5a7d64e5c599f629f0ffc8a28701d5

Observation 14dee0d9-eb93-4f7a-ba8f-46497b18493d · outbound

This paper cites D.; Ermon, S.; and Finn, C.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning D.; Ermon, S.; and Finn, C

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.239596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.239596Z digest=sha256:7dc80f8a8d5d0a8f89ef6f674f90988574d26781e48099024ea2e04dfeea75a2

Observation 4ab36d37-786e-4545-9b60-4a0dd497893a · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.244037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.244037Z digest=sha256:dea14ccce7862e17692b672cf7972dedb232586e6d53017f85b1ab7a17d05a39

Observation 91440656-8a7c-4351-8cb4-5444c2e6743f · outbound

This paper cites Proximal Policy Optimization Algorithms.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Proximal Policy Optimization Algorithms

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.248574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.248574Z digest=sha256:854c1046d45c20c3ab6e587b9ac701036bcacbd8a5ba0ce4273cd399cba6a771

Observation ec5e9d4d-52de-4882-a99e-3a50f9b73404 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.106655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.253394Z digest=sha256:41ffaa95542a398b09de4e1c534a1b6df04853467ab72cf96236675e5fdbbd9e

Observation 75f9ce17-0646-4097-81cc-c9f357540fff · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.257750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.257750Z digest=sha256:a0f6be40719db107417726f30a193b8c121b19f748b044892646d2280a6c6932

Observation 5dacf8b6-d784-4531-8e85-b9387d26ae5f · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.262322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.262322Z digest=sha256:7cfd8475a92b9c0bfb2ca32516894cea77c1fdcc14248a4075fb945b0c51e21f

Observation 0fb7dae3-9ebc-459a-bbb2-487fe04d0a30 · outbound

This paper cites R1-Searcher: Incentivizing the Search Capability in LLMs via Reinforcement Learning.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning R1-Searcher: Incentivizing the Search Capability in LLMs via Reinforcement Learning

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.266683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.266683Z digest=sha256:832b8b59b4d80597fc1740c3e6b96aff801741109d3af9e15381c5d324dbe838

Observation 7a829d1d-6953-4abc-9344-af55a7addeb6 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.081150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.271585Z digest=sha256:8800f3d942e504c1d95e01c1e510116340f2af971b782161abd669b27b78bd5d

Observation 930de274-7d92-4bba-afa9-7a90c589e8c4 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.065674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.275706Z digest=sha256:6554f8c505242ec1ac476f5114ed0c8977e57830730b0b35c9f15ec73f0ca11e

Observation a227ce7d-db50-46a7-8b24-353e9e43bc3a · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.279697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.279697Z digest=sha256:fe7165f012f668effc8d7ceceddafc716da352e6708b7a761600a53e4c3ddcb8

Observation 3977e03d-20a8-4cd6-8f54-dbbbe0ecf6d9 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.035175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.283718Z digest=sha256:90ed0cffc9bc60fbfb35f8d732e7a9e5cb0b1f2b208e33dfa381420f50af581d

Observation 87884ee7-f47c-4fa4-820c-7f8c3f3aabc9 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.016815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.288398Z digest=sha256:ca4bf885da60f1f874c7e44b14c34b4df6515d3f0f8adcdd7aca79c25d9c62e3

Observation 0fd5e305-96e1-4806-8b5f-7d0512d9a7ac · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:55.999918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.292695Z digest=sha256:869ad9456b780796c8f7c18b2be360e5d4362ddf04d838db677051986610a788

Observation 6dd7b948-b9de-43da-ab64-09920117af32 · outbound

This paper cites Text Embeddings by Weakly-Supervised Contrastive Pre-training.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Text Embeddings by Weakly-Supervised Contrastive Pre-training

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.296956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.296956Z digest=sha256:1d84624f969a24d32c621949c8e3e98b0a79e59e2ded7725d068186009e31d11

Observation 2169dfa0-4492-4921-b001-ead51af07151 · outbound

This paper cites MenatQA: A New Dataset for Testing the Temporal Comprehension and Reasoning Abilities of Large Language Models.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning MenatQA: A New Dataset for Testing the Temporal Comprehension and Reasoning Abilities of Large Language Models

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-08-06T12:23:55.500019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.301646Z digest=sha256:c4b4ab31084604050c547fbbf29ee214461e4fb8af910c69c9e5b1493071c0dc

Observation deb46779-f24f-42fa-8baa-c1861eaf09da · outbound

This paper cites Structural Entropy Guided Agent for Detecting and Repairing Knowledge Deficiencies in LLMs.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Structural Entropy Guided Agent for Detecting and Repairing Knowledge Deficiencies in LLMs

Reference 58

Resolution
verified exact
local_arxiv, observed 2026-08-06T12:23:55.477185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.307048Z digest=sha256:8e6e0e94e72dba87954d367ae1e568ba2d76ab195e51ff867410329b431c3fe1

Observation 9c056e49-0b23-4e46-84e5-0772bcc0b99f · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:55.983520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.311727Z digest=sha256:d8137534baaa00cd19d97b9c8b857f579077277bc6b4b626220a95ec3463b1cb

Observation b6dce6f5-72b0-4a74-9bc9-a36e24b926d2 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:55.968767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.316038Z digest=sha256:aa1d1c32b7e9367e4a5c6d5524283e1d9399ce649ac70a230e9ee9fd7873c0a5

Observation 79dfff32-e5a3-4db3-b580-4d25e33b316a · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 61

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:55.952044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.320345Z digest=sha256:cec24af858bda91e887b5b5e2d67aa24c2925fd59d6d505b0689853dda5233bd

Observation a2a71f0e-8640-4bf5-9bea-2f4fa965e8aa · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.324339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.324339Z digest=sha256:5c17e71cac23ac04c997bf2555f53f1f6d7c4215fcea7aff4f7500d864605b07

Observation bfd1390e-235f-4271-b985-235d5a08d6d6 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:55.926226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.328625Z digest=sha256:e5fbdd749fda7277ab654ca6b17882357974d8fef3a85800ceb35648ccb8a1b6

Observation 512c3165-da91-4b50-ae51-162ed3e576f1 · outbound

This paper cites C.; Kaddar, Y.; Blunsom, P.; Staton, S.; and Gal, Y.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning C.; Kaddar, Y.; Blunsom, P.; Staton, S.; and Gal, Y

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:23:55.910593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.332803Z digest=sha256:27e63c2db2a9654d016c0efcb263daa5b6d0862d745a08c7b40d9890c5ec5cd6

Observation 681164e1-4561-4ded-87fc-f8956956aa80 · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.337141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.337141Z digest=sha256:8857e35781eeb55abb6a596a5a4aeb4b7359b584fccd912d1f4aa4f9f08a6d45

Observation 9af55f1e-16aa-4fac-8920-52e7930facdb · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.342076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.342076Z digest=sha256:f951fcfd8e57d4c99ee2458fedebf3ff36761e52e6ad513d652a6e0d0dfadcd7

Observation c36f7776-b63e-49e8-9354-a906a54ca132 · outbound

This paper cites SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.346347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.346347Z digest=sha256:0ac5a33a2b14227c229ddf2e63da50bded3d95f3bf1cd9f8dadafc2fe9e72483

Observation 7c79cf1b-5bcc-4831-a59e-0453df71363e · outbound

This paper cites Instruction-Following Evaluation for Large Language Models.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Instruction-Following Evaluation for Large Language Models

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.351954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.351954Z digest=sha256:4667c1fd927ef167c63bcf0b5e84b0c8d0f17603c1fb854c6e3810e5fef01a46

Observation 9d1213f1-8b9f-4d62-8751-4accbfd81558 · outbound

This paper cites MEM1: Learning to Synergize Memory and Reasoning for Efficient Long-Horizon Agents.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning MEM1: Learning to Synergize Memory and Reasoning for Efficient Long-Horizon Agents

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.356601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.356601Z digest=sha256:a9a76e512005f40055752734aac8debe35262c78d817fa305b781cbf3e5e8286

Observation 81019d0f-7131-47a8-a919-75900c09714b · outbound

This paper cites , " * write output.state after.block = add.period write newline.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning , " * write output.state after.block = add.period write newline

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.361388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.361388Z digest=sha256:e42f6926d33856a53b4fc28312f165c51487562304ce521073946829b38f5dfa

Observation dab0f0fa-d2db-4448-9f87-dabbe20ebe69 · outbound

This paper cites write newline.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning write newline

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.366230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.366230Z digest=sha256:235a55f46feefc514a313ee1feb759aa3ea8e9b2a850eeabe27f78d41dd03b92

Pith citing papers

Observation be2f730c-502f-49ba-afd5-8c9677e23336 · inbound

Chain-of-Agents: End-to-End Agent Foundation Models via Multi-Agent Distillation and Agentic RL cites this paper.

Chain-of-Agents: End-to-End Agent Foundation Models via Multi-Agent Distillation and Agentic RL AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-05T23:57:35.839876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:57:35.839876Z digest=sha256:4ed959764bcc66cc983b76385b591583f9a2d178455fd7b68798ac7ce15e7690

Observation 1ab86500-5dbb-4b41-98a7-8edf3269a992 · inbound

Learning How to Use Tools, Not Just When: Pattern-Aware Tool-Integrated Reasoning cites this paper.

Learning How to Use Tools, Not Just When: Pattern-Aware Tool-Integrated Reasoning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T14:49:51.602433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:49:51.602433Z digest=sha256:7291e3c8cecbb3649278f7874bb4dd3d484942f2af5f321309a8f64663a5d276

Observation 9bce88ff-83e7-4b1d-ac9e-3ded845d3cf5 · inbound

SIGMA: Search-Augmented On-Demand Knowledge Integration for Agentic Mathematical Reasoning cites this paper.

SIGMA: Search-Augmented On-Demand Knowledge Integration for Agentic Mathematical Reasoning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T06:57:16.012421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:57:16.012421Z digest=sha256:d6932f65dd8de808578aaa5f45e1bd653c60a6424778fa5c166e5efc73df3aee

Observation db6d5a77-182a-41f6-9d2e-2679ff59999e · inbound

Lost in Execution: On the Multilingual Robustness of Tool Calling in Large Language Models cites this paper.

Lost in Execution: On the Multilingual Robustness of Tool Calling in Large Language Models AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T11:44:07.353535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T11:44:07.353535Z digest=sha256:fba825345242641bac6d1fa328a5ec411eb15b337cb2bfe01ed3979552b7d710

Observation 728c2d82-7c0c-4661-ae86-f1a6a9442a46 · inbound

Toward Efficient Agents: Memory, Tool learning, and Planning cites this paper.

Toward Efficient Agents: Memory, Tool learning, and Planning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 142

Resolution
unresolved
no resolver link, observed 2026-08-03T09:21:45.096788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T09:21:45.096788Z digest=sha256:9eabf975a937c5d93d2fc27c61685d50e6081f7269f7d25990f30fbe833e0c56

Observation ac93b3b1-bcd6-4ace-9141-1a7860485d3f · inbound

Reasoning and Tool-use Compete in Agentic RL:From Quantifying Interference to Disentangled Tuning cites this paper.

Reasoning and Tool-use Compete in Agentic RL:From Quantifying Interference to Disentangled Tuning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-03T05:56:57.193714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:56:57.193714Z digest=sha256:3c1173d5fb599c945f0d1eadde44f565b040927a0dfdca73988ab52daaaf962b

Observation 1f3e87f2-3491-4609-a0ec-8b177e8d9f8f · inbound

AnomalyAgent: Agentic Industrial Anomaly Synthesis via Tool-Augmented Reinforcement Learning cites this paper.

AnomalyAgent: Agentic Industrial Anomaly Synthesis via Tool-Augmented Reinforcement Learning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-10T18:35:42.825362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T18:32:31.955785Z digest=sha256:935b1308a3138f2850cb109b5f9a1354fb194c4d7895a05c7d1d7d09c38467ab

Observation d0287446-3114-4176-936c-605bc0cd6bad · inbound

See Further, Think Deeper: Advancing VLM's Reasoning Ability with Low-level Visual Cues and Reflection cites this paper.

See Further, Think Deeper: Advancing VLM's Reasoning Ability with Low-level Visual Cues and Reflection AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:36:17.711402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T04:46:16.497585Z digest=sha256:f9013cea343c5241d635f541fc498acf7adead2a674fdc3ed3aff672a7c371a4

Observation e63939b4-76ff-4201-b75d-125c52cf8ac7 · inbound

Tools as Continuous Flow for Evolving Agentic Reasoning cites this paper.

Tools as Continuous Flow for Evolving Agentic Reasoning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-11T01:45:51.740826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-11T01:30:19.859374Z digest=sha256:069057c13afcce9e8a536ffb8d398ed61dde0b478d525b66ccaf7cd9287bc4bf

Observation 9f34793a-d609-49d1-806b-6a4190184c37 · inbound

SkillMaster: Toward Autonomous Skill Mastery in LLM Agents cites this paper.

SkillMaster: Toward Autonomous Skill Mastery in LLM Agents AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:36:24.078934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T00:59:03.361037Z digest=sha256:b9f466696f59a8dec2eef94afb7cf44ff1a28ffb4e008694d0e9d7507f9a8858

Observation a3d8a207-1068-45b0-b0d7-6ec3032c8f74 · inbound

SkillMaster: Toward Autonomous Skill Mastery in LLM Agents cites this paper.

SkillMaster: Toward Autonomous Skill Mastery in LLM Agents AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:32:29.862504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:09.985448Z digest=sha256:5d822316e7d96cd212c46630075b33cf5233a20e367f216b5d9ad50c7db85e94

Observation feb34b55-7224-4cf2-8cbb-144ff2b23b32 · inbound

TIDE-Bench: Task-Aware and Diagnostic Evaluation of Tool-Integrated Reasoning cites this paper.

TIDE-Bench: Task-Aware and Diagnostic Evaluation of Tool-Integrated Reasoning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:31:23.948998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-12T02:47:31.830410Z digest=sha256:8af635df780d206dbb7bad0dd3942550b327d8497fce88ac8f81c9867dd713a9

Observation d0e8bf0e-7946-4e2d-aca4-9c72cdbb2022 · inbound

Learning When Not to Act: Mitigating Tool Abuse in Agentic Reinforcement Learning cites this paper.

Learning When Not to Act: Mitigating Tool Abuse in Agentic Reinforcement Learning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:46:20.257307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-28T14:59:57.983338Z digest=sha256:0fb1b6aaf95c8e8faafccccc56e5ca66d344335a16e596073f003b3a0f3f04d4

Observation 60831c2a-f04d-4715-80d4-9105dcd716b8 · inbound

Modularized Reinforcement Learning on LLMs: From MDP Creation to Exploration and Learning cites this paper.

Modularized Reinforcement Learning on LLMs: From MDP Creation to Exploration and Learning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 223

Resolution
verified exact
arxiv_id, observed 2026-07-04T08:09:40.721699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T12:15:08.304150Z digest=sha256:022e50dfce3ee59a7c23b6f44063aeb7cefd30be4e0f3809a0dc6a236000216b

Observation 998ceb97-64bb-4584-b4fa-38fd5bc6ced9 · inbound

TurnSight: Turn-Level Hindsight Self-Distillation for Tool-Integrated Reasoning cites this paper.

TurnSight: Turn-Level Hindsight Self-Distillation for Tool-Integrated Reasoning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T04:18:15.288921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T04:18:15.288921Z digest=sha256:9f91b69d60f9050d38716b1287587427143aff318e36a92d806c72a579e27e30