Pith. sign in

Paper Citation Record · LEDGER

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control

As of 18 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 1 inbound Pith citation observation for arXiv:2502.02265.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.02265 v1

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T12:49:59.170317Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-27T15:27:51.585099Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

30 of 30 outbound references displayed

  • verified exact0
  • verified fuzzy18
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 005660fd-f11d-40c9-bf38-a89d949fdacc · outbound

This paper cites Hindsight experience replay.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Hindsight experience replay

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.607414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.035484Z digest=sha256:be37ffd32051cdb86198f436204544dbe976e99a0cd95143cb9278569fdc6f87

Observation 6e179587-d8f0-4ba9-86c4-665a4ad47c42 · outbound

This paper cites P., Maghade, D., Sondkar, S., and Pawar, S.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control P., Maghade, D., Sondkar, S., and Pawar, S

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.592147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.040935Z digest=sha256:bd138a7020638055dcd410356f8f3bccca16ad33e233e9a12406d95a3b38dead

Observation abd3d449-c43e-4b46-a973-dc9bf510f0a0 · outbound

This paper cites Exploration by random network distillation.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Exploration by random network distillation

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.576396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.045753Z digest=sha256:4779594bc4bb712deb8969abdcb8d027a23c29a3a75efdba474f19d4b61b1314

Observation 03dd188a-2241-4d44-9027-1269f5a20e47 · outbound

This paper cites Reinforcement learning for control: Performance, stability, and deep approximators.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Reinforcement learning for control: Performance, stability, and deep approximators

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.561296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.050519Z digest=sha256:65cf4ec0609a7f661c5c4d5aa8f2a46bc648b5cd79edd6b7cdbc0fad69960794

Observation 008b4ff0-7dca-470a-9775-df1beccedce0 · outbound

This paper cites Learning sim-to-real dense object descriptors for robotic manipulation.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Learning sim-to-real dense object descriptors for robotic manipulation

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.546171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.055719Z digest=sha256:6e83d4b1203e7d835f6fba3c0405ca1398a2595ca248b4db3bcbe559a8c89800

Observation a60f5adf-8c81-4a66-85e9-48918bc402b3 · outbound

This paper cites Neural-network-based nonlinear optimal terminal guidance with impact angle constraints.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Neural-network-based nonlinear optimal terminal guidance with impact angle constraints

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.531047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.060598Z digest=sha256:77ce69f4d77842c8d28979f783ed67b375c190ac7626548d4ac2108e8dfcf436

Observation bb575392-bd9b-47db-9728-84df08a3f9bf · outbound

This paper cites an unresolved cited work.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-09T12:49:59.515527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.065842Z digest=sha256:4ec52bcd7329cf9cb5de1702f7edcc6b01854566d32ed8295dbc782ab2a79928

Observation 86175193-adb9-4711-9cb5-0aa029de2293 · outbound

This paper cites Addressing function approximation error in actor-critic methods.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Addressing function approximation error in actor-critic methods

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T12:49:59.070277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:49:59.070277Z digest=sha256:bf78db7a3647416c3871b9a6da338aabb0ab3e270a1e883e2aa995f546fd40fe

Observation e0797492-a89d-44a9-9f02-1cc9270fa322 · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-09T12:49:59.075006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:49:59.075006Z digest=sha256:f91faab970abe5b71ab3c732e6aa251a68a68e12e5923b585b429fe97ed712f2

Observation faf70b9f-6882-46ae-84e0-f2c305af9901 · outbound

This paper cites Learning to utilize shaping rewards: A new approach of reward shaping.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Learning to utilize shaping rewards: A new approach of reward shaping

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.481764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.079365Z digest=sha256:129f350c1e0bd777d22ff571d8b90d89665a2184a2c2c39ad07cc46d1b274d4f

Observation 48605718-3d47-4a0b-a625-fdb1949b5017 · outbound

This paper cites D., Tsounis, V., Hwangbo, J., Bodie, K., Fankhauser, P., Bloesch, M., et al.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control D., Tsounis, V., Hwangbo, J., Bodie, K., Fankhauser, P., Bloesch, M., et al

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.466447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.083704Z digest=sha256:3ef4ca6111d05aa66741487abb5fa9c65bccf55787fe0f6e77a6b3cb2e064367

Observation 948f4ce5-1522-4428-b26e-620c297cfa90 · outbound

This paper cites R., Sobh, I., Talpaert, V., Mannion, P., Al Sallab, A.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control R., Sobh, I., Talpaert, V., Mannion, P., Al Sallab, A

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-09T12:49:59.088067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:49:59.088067Z digest=sha256:459d2273f12ada4869c009a4e279b30d41b9f5ad72d688423a4ee0ac0a632b9b

Observation 8d6f2deb-e3db-46b5-b1b5-7b9f5a4a4dfd · outbound

This paper cites H., and Chong, G.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control H., and Chong, G

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.442341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.092820Z digest=sha256:6b70923904c4165b34b54331862888e8c076c0553c1ce5806ae19ff622c269f8

Observation 7b3ee027-c958-4823-8cad-d6c96086fd1f · outbound

This paper cites Continuous control with deep reinforcement learning.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Continuous control with deep reinforcement learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-09T12:49:59.097406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:49:59.097406Z digest=sha256:8f556526c5d265a7a3446fa495c7669b70780dec94bb5311a8463267fb138ee9

Observation 19f774e0-a4fd-48fb-8f70-91d5a83473ba · outbound

This paper cites A review of industrial mimo decoupling control.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control A review of industrial mimo decoupling control

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.427187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.102440Z digest=sha256:078e805d8b6fdd3808e30868c962e5cfcef4c44330ade73832f4236e6878fd23

Observation d5e18518-2e86-4153-bfd2-06607d1a89fd · outbound

This paper cites Goal-Conditioned Reinforcement Learning: Problems and Solutions.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-09T12:49:59.106851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:49:59.106851Z digest=sha256:5e14ba251a9f6ab68af8e22d35daa5621be79ab5c1d073dfb189bdfee8863709

Observation 88376a46-728d-4a23-aacb-6bdf8e95806a · outbound

This paper cites A., Veness, J., Bellemare, M.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control A., Veness, J., Bellemare, M

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-09T12:49:59.111561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:49:59.111561Z digest=sha256:32d75468647e3d5ea4c3416aa5ccfef80e4531f51f0591209c3e192fd5d5cfba

Observation e285dbe8-e8e8-46a8-a728-d4ce6efa5bc4 · outbound

This paper cites Goal-directed planning via hindsight experience replay.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Goal-directed planning via hindsight experience replay

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.402026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.115909Z digest=sha256:781c92ea6ea43f96f6dfa9e15c99ec037b6a7bf26de446cdc4c3923944344fba

Observation cb9b5b77-5a19-4227-a0b5-cdb3e0cbddc8 · outbound

This paper cites A., and Darrell, T.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control A., and Darrell, T

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-09T12:49:59.120185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:49:59.120185Z digest=sha256:79d9a12f004def96a6f0f4effb82a387fc9551767a933b5c69f0229d971006fd

Observation 448fbb73-893e-477b-9963-880893f87a36 · outbound

This paper cites Dynamic-matching adaptive sliding mode control for hypersonic vehicles.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Dynamic-matching adaptive sliding mode control for hypersonic vehicles

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.377494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.124777Z digest=sha256:26ee40239eb73f329371fc454331ae8ef520e3bcdc686cd005310754c0530b2e

Observation 06734728-c786-440a-97ac-89c7bdb52c5b · outbound

This paper cites Discovering blind spots in reinforcement learning.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Discovering blind spots in reinforcement learning

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.362366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.129734Z digest=sha256:1e48c08b2c686699827dd40b375aacc137f37937e2f14ed7b6878d85e50bee07

Observation c6bec956-1b08-41e5-97e5-6f9dba0e8240 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Proximal Policy Optimization Algorithms

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-09T12:49:59.134118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:49:59.134118Z digest=sha256:56e845e6f80b450b57092e2d80745c86a2e180d5aeb82f0fbeae2a8247a47582

Observation 2cdffbd7-5206-4799-8fe2-ff334b6439bd · outbound

This paper cites Deterministic policy gradient algorithms.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Deterministic policy gradient algorithms

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-09T12:49:59.138954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:49:59.138954Z digest=sha256:b9e9ab458f93973322d316f9126fad22b357381ebeb14a6937806bea58d824e7

Observation f353c4f6-4994-4597-959e-d7c5f1028dee · outbound

This paper cites Action robust reinforcement learning and applications in continuous control.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Action robust reinforcement learning and applications in continuous control

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-09T12:49:59.143546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:49:59.143546Z digest=sha256:570da10b9e882da15910702d64431e23c9925deaecac3d03618d78083f4ae951

Observation 407097ac-ed4c-47a2-925a-55e19f8e6c23 · outbound

This paper cites D., Michi, A., Chervonyi, Y., Davies, I., Paduraru, C., Lazic, N., Felici, F., Ewalds, T., Donner, C., Galperti, C., et al.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control D., Michi, A., Chervonyi, Y., Davies, I., Paduraru, C., Lazic, N., Felici, F., Ewalds, T., Donner, C., Galperti, C., et al

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.325736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.148057Z digest=sha256:066ae92af8d0a8b85bf1b796b71271cba276bc466f7000d8e74c49dd6abf463d

Observation 69cefe13-0d49-40f6-9cef-8c89416365ee · outbound

This paper cites A brief survey on nonlinear control using adaptive dynamic programming under engineering-oriented complexities.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control A brief survey on nonlinear control using adaptive dynamic programming under engineering-oriented complexities

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.310143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.152626Z digest=sha256:45e5dc9ff9371560f3a9252901bfff1f03fae76e01d7d16a27dd54fc6725b974

Observation 8a103e4a-59e7-478a-9bd3-c47d20057cfc · outbound

This paper cites P., Qingqing, L., and Westerlund, T.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control P., Qingqing, L., and Westerlund, T

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.294548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.156933Z digest=sha256:57b4fdb34dbba0bd525621dc501c8874e406cd396c34cc89e25664d14e48600d

Observation e0d97cee-bc2e-4aa1-8468-00c75548ee32 · outbound

This paper cites P., and Westerlund, T.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control P., and Westerlund, T

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.279619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.161470Z digest=sha256:7a78f127a930227a6d169883da87db881aaf45d3822fa35d8cb54f72d5a6bc9e

Observation b34a8f89-1956-4252-bf22-d8918cfa874a · outbound

This paper cites A comprehensive survey on transfer learning.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control A comprehensive survey on transfer learning

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.262839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.165901Z digest=sha256:ea8019d0d1bd0fb7491dcc472d08ddb23cd46bed39519fa7fef8c0067c6b6429

Observation 82e864a6-f2f0-4395-a749-cc4b93895863 · outbound

This paper cites write newline.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control write newline

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-09T12:49:59.170317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:49:59.170317Z digest=sha256:93bcb23dfd6ee6834ccebc75f976718ef38c9a6b2e9f0c6dceac6bd0d24102f9

Pith citing papers

Observation ff0e69d0-98c0-4cf9-9d63-2b1a0dafcd24 · inbound

Tracking the Effective Surface Area of Non-Convex Satellites cites this paper.

Tracking the Effective Surface Area of Non-Convex Satellites Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-06-27T15:31:00.509587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-27T15:27:51.585099Z digest=sha256:bc742486f14a3e2a7cb421ec3ee3eebdf5894d6c1de0326ffeb1ce1bc720934e