Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 31 inbound Pith citation observations for arXiv:2505.00024.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:44:36.354899Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation f32edaa0-bd51-4cb7-81a7-533956c446c1 · inbound
The Hallucination Tax of Reinforcement Finetuning Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f62d10de-d64c-42de-b2d7-77681a765934 · inbound
Visual Agentic Reinforcement Fine-Tuning Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6fc74151-5eff-4081-a180-57d2aad2b823 · inbound
WebDancer: Towards Autonomous Information Seeking Agency Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7932a883-2bfe-4888-803c-20a7cea739a8 · inbound
Open CaptchaWorld: A Comprehensive Web-based Platform for Testing and Benchmarking Multimodal LLM Agents Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d3b4cfe-bc69-4358-96c9-0d690ce2e958 · inbound
StepFun-Prover Preview: Let's Think and Verify Step by Step Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 159b4990-e5c7-4400-b315-9e5dbff73feb · inbound
MUA-RL: Multi-turn User-interacting Agent Reinforcement Learning for agentic tool use Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2a6fe6a-5e9d-4f0b-9f1d-f754a154fbe8 · inbound
How Can Input Reformulation Improve Tool Usage Accuracy in a Complex Dynamic Environment? A Study on $\tau$-bench Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d24ae44-7827-42b1-bf90-4ea7efc56487 · inbound
LLaVA-Critic-R1: Your Critic Model is Secretly a Strong Policy Model Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d928ccf2-7034-4c4f-a90f-1b64febeffe8 · inbound
Reinforced Visual Perception with Tools Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a2fa7a9-9b77-4d60-865b-50ae03e35d12 · inbound
Webscale-RL: Automated Data Pipeline for Scaling RL Data to Pretraining Levels Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bb5db8fd-4415-4d2c-bb58-9cc652e33de5 · inbound
MURPHY: Feedback-Aware GRPO with Retrospective Credit Assignment for Multi-Turn Code Generation Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9a615701-ee8e-4f43-88ba-0606415905e3 · inbound
Entropy-Preserving Supervised Fine-Tuning via Adaptive Self-Distillation for Large Reasoning Models Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e642e121-706e-41e9-a0b6-4f77905bd761 · inbound
LAST: Leveraging Tools as Hints to Enhance Spatial Reasoning for Multimodal Large Language Models Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f7167414-ff93-4c04-9fb4-de6f59864fb9 · inbound
Controllable and Verifiable Tool-Use Data Synthesis for Agentic Reinforcement Learning Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 142182ba-3d01-4ab8-8105-7d86ff9c2db3 · inbound
Democratizing Tool Learning with Environments Fully Simulated by a Free 8B Language Model Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 97593b0d-1f44-478e-8c0e-abbda9ea3e8c · inbound
R2IF: Aligning Reasoning with Decisions via Composite Rewards for Interpretable LLM Function Calling Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5287c186-a7de-47a5-a21d-5b2c0e598478 · inbound
R2IF: Aligning Reasoning with Decisions via Composite Rewards for Interpretable LLM Function Calling Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 50d2d743-3da8-4f94-95d3-0edc9010f587 · inbound
CuraView: A Multi-Agent Framework for Medical Hallucination Detection with GraphRAG-Enhanced Knowledge Verification Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 60d09fb1-7a28-4c2c-9a49-5cb1d35291c4 · inbound
RubricRefine: Improving Tool-Use Agent Reliability with Training-Free Pre-Execution Refinement Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1691925d-e988-412a-b22d-2a65c095f354 · inbound
RubricRefine: Improving Tool-Use Agent Reliability with Training-Free Pre-Execution Refinement Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e24e6cda-8f49-44bc-8708-0ee6cce84a0f · inbound
RubricRefine: Improving Tool-Use Agent Reliability with Training-Free Pre-Execution Refinement Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b610b465-0527-41fb-bae3-78b36d5cee19 · inbound
Entropy Polarity in Reinforcement Fine-Tuning: Direction, Asymmetry, and Control Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 43e4fe24-08dc-4499-af74-b9b8f36dbb38 · inbound
Entropy Polarity in Reinforcement Fine-Tuning: Direction, Asymmetry, and Control Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e15b3f73-7bf1-4501-ab0a-b930c5a74237 · inbound
Reinforcement Learning for Tool-Calling Agents in Fast Healthcare Interoperability Resources (FHIR) Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 131cda4a-c012-41cc-bbb6-824cbd1c1fa7 · inbound
Entropy-KL Divergence-based Token Masking: A Novel Approach for Selective Fine-tuning of Large Language Models Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 623b7556-ad61-4315-ad77-6ff632f35973 · inbound
On Effectiveness and Efficiency of Agentic Tool-calling and RL Training Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fa3b5909-2615-42e9-8579-37aa52a39b5e · inbound
Synthesize and Reward -- Reinforcement Learning for Multi-Step Tool Use in Live Environments Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 53b32c7b-1cbe-468d-aa76-2e2eed45add2 · inbound
Capability-Aligned Hierarchical Learning for Tool-Augmented LLMs Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 96f730d3-8d75-45fa-b278-cbec30e2f639 · inbound
Pushing the Limits of LLM Tool Calling via Experiential Knowledge Integration and Activation Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8ded1493-716c-4061-a89a-18136fec6c84 · inbound
TCPO: Turn-Level Credit Policy Optimization Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a641ada2-d208-4a43-987a-a284cdd29e67 · inbound
TurnSight: Turn-Level Hindsight Self-Distillation for Tool-Integrated Reasoning Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.