Pith. sign in

Paper Citation Record · LEDGER

AgenticRL: Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation

As of 7 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 0 inbound Pith citation observations for arXiv:2606.03963.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.03963 v3

Coverage vector

measured 31 of 31 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-28T10:03:54.185019Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

31 of 31 outbound references displayed

  • verified exact8
  • verified fuzzy0
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 871cf118-69e7-4651-9e2e-4a75d0078055 · outbound

This paper cites Sim-to-Real Deep Reinforcement Learning based Obstacle Avoidance for UAVs under Measurement Uncertainty.

AgenticRL: Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation Sim-to-Real Deep Reinforcement Learning based Obstacle Avoidance for UAVs under Measurement Uncertainty

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-02T03:26:29.346422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T10:03:54.185019Z digest=sha256:a5363df5f8db93302d4e55b826d65683485c70787a0a3db382d92ae1edab06cb

Observation ef190c82-393b-489b-a046-3df6bb59ac1f · outbound

This paper cites W ANG, X.

AgenticRL: Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation W ANG, X

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-28T10:03:54.185019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:03:54.185019Z digest=sha256:5d3aaadc14b3ec1f843eea6a29027ba500d02ab6a4643f508091167100088f69

Observation 20e938b5-48c6-489e-b8ef-3ffbd64dd179 · outbound

This paper cites an unresolved cited work.

AgenticRL: Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-06-28T10:03:54.185019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:03:54.185019Z digest=sha256:d6e8a00df21e7293d33ccfe6138c51e301a4ba73b7599231fa979bec3ebecf75

Observation 47021b31-b007-4f96-8a3c-9ab43765867d · outbound

This paper cites an unresolved cited work.

AgenticRL: Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-06-28T10:03:54.185019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:03:54.185019Z digest=sha256:d43ec1fe555970e34311b596ce9bfad4b1876ca04f8307a93985345c79352b40

Observation 5686bb34-5bc1-418c-b2bc-712acb4a53f8 · outbound

This paper cites Sadigh, A.

AgenticRL: Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation Sadigh, A

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-28T10:03:54.185019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:03:54.185019Z digest=sha256:816b6065e75498733757cbdadf8619dd848fb83003d4c09b4087d6f6147f79ef

Observation 6be3d70a-4a1c-4414-b309-4e77ed25849f · outbound

This paper cites Bıyık, N.

AgenticRL: Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation Bıyık, N

Reference 6

Resolution
unresolved
no resolver link, observed 2026-06-28T10:03:54.185019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:03:54.185019Z digest=sha256:a5eb2c2f55029729def79e70a5cd58a311b2dd30b6aeb89f5381aa0747f84929

Observation 111eb76f-6768-4b1a-afcc-078ea7e02fc6 · outbound

This paper cites Liang, W.

AgenticRL: Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation Liang, W

Reference 7

Resolution
unresolved
no resolver link, observed 2026-06-28T10:03:54.185019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:03:54.185019Z digest=sha256:9145d86b1977d25a2a01d525fe5fa0de29b434aa645df2d05f037856d76e10c1

Observation 5c6d4373-d173-4728-a85b-8a5754bcd145 · outbound

This paper cites an unresolved cited work.

AgenticRL: Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-06-28T10:03:54.185019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:03:54.185019Z digest=sha256:3343d5eaaf4c4ef47ab986900e742d9211ee412e7015b34e6147ab69b45e9ab9

Observation 0cbc2f94-14f0-413f-9e3b-cfe1e132cb71 · outbound

This paper cites an unresolved cited work.

AgenticRL: Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-28T10:03:54.185019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:03:54.185019Z digest=sha256:856dc673c646dc0d664d6fd5857c9315ca1252998f27bdf923b7038420855182

Observation 7044cc48-c7c8-457e-ab7c-04b93b2d7c7a · outbound

This paper cites Mill ´an-Arias, R.

AgenticRL: Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation Mill ´an-Arias, R

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-28T10:03:54.185019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:03:54.185019Z digest=sha256:b4a2f71bee5753f2c1c85bfbc2ceda79502386ce5673ed38e3cc701bc3d42982

Observation 942a4be4-c320-4265-9062-d5d24690aa7e · outbound

This paper cites an unresolved cited work.

AgenticRL: Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-06-28T10:03:54.185019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:03:54.185019Z digest=sha256:173a1406861dbce433898d6fc4ec03e88aea59718820079f857002b091b927d4

Observation c4176a5f-bcb8-41c7-8e42-3dba430a69e7 · outbound

This paper cites Devidze, G.

AgenticRL: Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation Devidze, G

Reference 12

Resolution
unresolved
no resolver link, observed 2026-06-28T10:03:54.185019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:03:54.185019Z digest=sha256:81f35550f0db9ef4aa35819e1388fd42b9e3c1c18028586299136f71db654020

Observation ba4bb9db-bdcf-46bf-ab43-a72b432051fe · outbound

This paper cites an unresolved cited work.

AgenticRL: Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-06-28T10:03:54.185019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:03:54.185019Z digest=sha256:0c9e4d7620072e1e3b7f5592e07999cad01b5f3e11aa6c457023740b2fb204c9

Observation e96c6407-5986-4661-8ef0-6b48b78dbb18 · outbound

This paper cites an unresolved cited work.

AgenticRL: Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-06-28T10:03:54.185019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:03:54.185019Z digest=sha256:0895e69bccbb8bdacb4786fbd6ffbe0e2d55a6e688e5b08db303de0dee3ba32e

Observation 5e171a60-d64d-45e7-9cbe-5b8f5c4a0bc8 · outbound

This paper cites Zhang, Y.

AgenticRL: Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation Zhang, Y

Reference 15

Resolution
unresolved
no resolver link, observed 2026-06-28T10:03:54.185019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:03:54.185019Z digest=sha256:687713fc73bce42892103be53547ebba260d52d705bff4ba48ac0769c90b80f0

Observation 12d88875-2be1-49f8-89de-add60ecc7900 · outbound

This paper cites Venuto, S.

AgenticRL: Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation Venuto, S

Reference 16

Resolution
unresolved
no resolver link, observed 2026-06-28T10:03:54.185019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:03:54.185019Z digest=sha256:94f3e12a928897887faf478b9dd9444d245f085a2d06abed2c0eed762cd30406

Observation fd6eece2-cc46-4388-a619-b89680570d10 · outbound

This paper cites RoboReward: General-purpose vision- language reward models for robotics.

AgenticRL: Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation RoboReward: General-purpose vision- language reward models for robotics

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-02T03:26:29.342667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T10:03:54.185019Z digest=sha256:8cf5e9849107044909cfef01a70db32aa30623d670a7eb2b0387e49a1ec34b84

Observation add91cbe-8919-44dd-bdab-93e8c759c185 · outbound

This paper cites Zhang, H.

AgenticRL: Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation Zhang, H

Reference 18

Resolution
unresolved
no resolver link, observed 2026-06-28T10:03:54.185019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:03:54.185019Z digest=sha256:194f65d2a0a5e24f9849da531651489ddbbd6b66254d333a53c357cef66404c0

Observation e30ed320-ac69-48f2-b193-5e82d5564596 · outbound

This paper cites an unresolved cited work.

AgenticRL: Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-06-28T10:03:54.185019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:03:54.185019Z digest=sha256:3c72e4acf316ceaba8143957a320a7062c175c5e3f5c3f9e0dee5c1a3a58b401

Observation b523f900-d17c-41b2-ac40-5cc185fbd853 · outbound

This paper cites Agentrl: Scaling agentic reinforcement learning with a multi-turn, multi-task framework.

AgenticRL: Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation Agentrl: Scaling agentic reinforcement learning with a multi-turn, multi-task framework

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-02T03:26:29.327123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T10:03:54.185019Z digest=sha256:0914b0a4be817164668d220476c9a02c88cb207972689a1a453ef5bb3d4c5f46

Observation b55f1ba7-cf7d-4b55-b1b0-60b03a905c3b · outbound

This paper cites Agent-RLVR: Training Software Engineering Agents via Guidance and Environment Rewards.

AgenticRL: Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation Agent-RLVR: Training Software Engineering Agents via Guidance and Environment Rewards

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-07-02T03:26:29.320110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T10:03:54.185019Z digest=sha256:24e2805c3eb770f0e20d404ed93b3cbfa57d6e1fb896e52dbf6e40b8daf9321e

Observation f7724ba0-2e11-44cf-a3a8-c3087742c8a5 · outbound

This paper cites Exploring Reasoning Reward Model for Agents.

AgenticRL: Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation Exploring Reasoning Reward Model for Agents

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-07-02T03:26:29.327537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T10:03:54.185019Z digest=sha256:33cc1b1a493ab4c9c5ffe66f334fb729a47a2a2b962970fcd4b2aa66dc078200

Observation 2fa5c8cd-2e1d-4666-a6ab-e196f6207750 · outbound

This paper cites OpenReward: Learning to Reward Long-form Agentic Tasks via Reinforcement Learning.

AgenticRL: Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation OpenReward: Learning to Reward Long-form Agentic Tasks via Reinforcement Learning

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-07-02T03:26:29.330841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T10:03:54.185019Z digest=sha256:aa5d3adfb8722f737d4cf7b85b92df3e46c26f645041bf0e9656e39231f598e0

Observation 2687e264-44eb-4af6-a54a-cd201819304c · outbound

This paper cites LiteResearcher: A Scalable Agentic RL Training Framework for Deep Research Agent.

AgenticRL: Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation LiteResearcher: A Scalable Agentic RL Training Framework for Deep Research Agent

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-07-02T03:26:29.341523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T10:03:54.185019Z digest=sha256:1dacd8a2f1897ae82f4c9ff29c8468422ae790c6dc592752f24ec888941ef454

Observation bde6fc1e-dbf8-448b-b715-2f10a959f5a0 · outbound

This paper cites an unresolved cited work.

AgenticRL: Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-06-28T10:03:54.185019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:03:54.185019Z digest=sha256:c186f9bfcd64bedb1c8aaed34fdb9aa4fc507c4ed2f050c8eb62dcb79d5821df

Observation 3e6aaadb-c3ae-4f80-8428-88469ee2ecb6 · outbound

This paper cites Panerati, H.

AgenticRL: Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation Panerati, H

Reference 26

Resolution
unresolved
no resolver link, observed 2026-06-28T10:03:54.185019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:03:54.185019Z digest=sha256:44632ff2affb58f25165634ef94384e185e315e95e223aca9c758556c85d07eb

Observation 15a70218-5469-4832-8cb2-c6c09f17222f · outbound

This paper cites Proximal Policy Optimization Algorithms.

AgenticRL: Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation Proximal Policy Optimization Algorithms

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-07-02T03:26:29.339578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T10:03:54.185019Z digest=sha256:d9d75db60c805e4e9f5f487f8410fcd2b3e5d2a54e69db2682e6205ff8a41245

Observation 8f8a7688-7ec4-447d-ae25-028fde444e0a · outbound

This paper cites Kaufmann, L.

AgenticRL: Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation Kaufmann, L

Reference 28

Resolution
unresolved
no resolver link, observed 2026-06-28T10:03:54.185019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:03:54.185019Z digest=sha256:4f95015e17f1283ff172cc02b56848737822a57795254ea7cca47ed35cf889eb

Observation c4e61c3d-acde-4712-9f2a-8c49b1e4e07c · outbound

This paper cites an unresolved cited work.

AgenticRL: Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-06-28T10:03:54.185019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:03:54.185019Z digest=sha256:fbb41a1ff091493158aeda167fa10fba5784a12f45f9b40f7598f4d1c8df04a3

Observation 550abb45-52a2-46de-8df5-c1691622811f · outbound

This paper cites Ahmed, N.

AgenticRL: Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation Ahmed, N

Reference 30

Resolution
unresolved
no resolver link, observed 2026-06-28T10:03:54.185019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:03:54.185019Z digest=sha256:c82ff97e5ad8d15b3a99b46f3cbc18a36494ec86ce8bd2c1f113cac882d5ca55

Observation 9b2f8ec9-8149-4b2f-a7fe-f40994a5b9d6 · outbound

This paper cites Eysenbach and S.

AgenticRL: Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation Eysenbach and S

Reference 31

Resolution
unresolved
no resolver link, observed 2026-06-28T10:03:54.185019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:03:54.185019Z digest=sha256:d6050b35fe9650f6ea5bcef085df74d7adfdca4e59a714644b5fbbef000e6706

Pith citing papers

No inbound Pith citation observations are available.