Pith. sign in

Paper Citation Record · LEDGER

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control

As of 18 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 1 inbound Pith citation observation for arXiv:2502.02265.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.02265 v1

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T12:49:59.170317Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-27T15:27:51.585099Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

30 of 30 outbound references displayed

  • verified exact0
  • verified fuzzy18
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 005660fd-f11d-40c9-bf38-a89d949fdacc · outbound

This paper cites Hindsight experience replay.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Hindsight experience replay

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.607414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.035484Z digest=sha256:430f8205196fcacbe088376f3802e539ac52977b22d1037de5c5d125dcbdef10

Observation 6e179587-d8f0-4ba9-86c4-665a4ad47c42 · outbound

This paper cites P., Maghade, D., Sondkar, S., and Pawar, S.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control P., Maghade, D., Sondkar, S., and Pawar, S

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.592147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.040935Z digest=sha256:e41d60dcd3d415ff9a03075e9c72168428c2fe8876144166f64bb2bbe26462cf

Observation abd3d449-c43e-4b46-a973-dc9bf510f0a0 · outbound

This paper cites Exploration by random network distillation.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Exploration by random network distillation

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.576396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.045753Z digest=sha256:f0005a5f197853b863123a984752ab0d35c78bd8c9ca0f37d3afd89373fbfad9

Observation 03dd188a-2241-4d44-9027-1269f5a20e47 · outbound

This paper cites Reinforcement learning for control: Performance, stability, and deep approximators.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Reinforcement learning for control: Performance, stability, and deep approximators

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.561296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.050519Z digest=sha256:aa7caf97da5ba90adc057d180202fd65d09b831d256df0fab09728b900b14086

Observation 008b4ff0-7dca-470a-9775-df1beccedce0 · outbound

This paper cites Learning sim-to-real dense object descriptors for robotic manipulation.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Learning sim-to-real dense object descriptors for robotic manipulation

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.546171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.055719Z digest=sha256:b1d17526162c4ef4386a9afcdec4feb1a86e57959a2ae4fc8ddebf9265f9d012

Observation a60f5adf-8c81-4a66-85e9-48918bc402b3 · outbound

This paper cites Neural-network-based nonlinear optimal terminal guidance with impact angle constraints.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Neural-network-based nonlinear optimal terminal guidance with impact angle constraints

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.531047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.060598Z digest=sha256:c610298b29d141d72ddcad1471d76138765000ebbf9bbe34c7d10d5d2e084b21

Observation bb575392-bd9b-47db-9728-84df08a3f9bf · outbound

This paper cites an unresolved cited work.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-09T12:49:59.515527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.065842Z digest=sha256:fdad431743a2df75cd019aefb2cfc3c8e4995d4a91a7e67f1a8a3009168a7ce2

Observation 86175193-adb9-4711-9cb5-0aa029de2293 · outbound

This paper cites Addressing function approximation error in actor-critic methods.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Addressing function approximation error in actor-critic methods

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T12:49:59.070277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:49:59.070277Z digest=sha256:bf78db7a3647416c3871b9a6da338aabb0ab3e270a1e883e2aa995f546fd40fe

Observation e0797492-a89d-44a9-9f02-1cc9270fa322 · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-09T12:49:59.075006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:49:59.075006Z digest=sha256:f91faab970abe5b71ab3c732e6aa251a68a68e12e5923b585b429fe97ed712f2

Observation faf70b9f-6882-46ae-84e0-f2c305af9901 · outbound

This paper cites Learning to utilize shaping rewards: A new approach of reward shaping.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Learning to utilize shaping rewards: A new approach of reward shaping

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.481764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.079365Z digest=sha256:40e2dede84b52cb19b7ad670e2b08a3abba8d30734956410131a34837b51b24a

Observation 48605718-3d47-4a0b-a625-fdb1949b5017 · outbound

This paper cites D., Tsounis, V., Hwangbo, J., Bodie, K., Fankhauser, P., Bloesch, M., et al.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control D., Tsounis, V., Hwangbo, J., Bodie, K., Fankhauser, P., Bloesch, M., et al

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.466447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.083704Z digest=sha256:0276cb6d8d8042469dab56014dbf274b59e5421f6714bb566821ec28a4c41a45

Observation 948f4ce5-1522-4428-b26e-620c297cfa90 · outbound

This paper cites R., Sobh, I., Talpaert, V., Mannion, P., Al Sallab, A.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control R., Sobh, I., Talpaert, V., Mannion, P., Al Sallab, A

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-09T12:49:59.088067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:49:59.088067Z digest=sha256:459d2273f12ada4869c009a4e279b30d41b9f5ad72d688423a4ee0ac0a632b9b

Observation 8d6f2deb-e3db-46b5-b1b5-7b9f5a4a4dfd · outbound

This paper cites H., and Chong, G.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control H., and Chong, G

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.442341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.092820Z digest=sha256:46de8c07407caf8dd340c0bffc8c47a2e1d836733db77f7bbddda0208bafd9e4

Observation 7b3ee027-c958-4823-8cad-d6c96086fd1f · outbound

This paper cites Continuous control with deep reinforcement learning.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Continuous control with deep reinforcement learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-09T12:49:59.097406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:49:59.097406Z digest=sha256:8f556526c5d265a7a3446fa495c7669b70780dec94bb5311a8463267fb138ee9

Observation 19f774e0-a4fd-48fb-8f70-91d5a83473ba · outbound

This paper cites A review of industrial mimo decoupling control.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control A review of industrial mimo decoupling control

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.427187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.102440Z digest=sha256:59fd2f7b04192a13ed092d2fc68ca67cefdfa6a2d610aa62df9657b5be2d7988

Observation d5e18518-2e86-4153-bfd2-06607d1a89fd · outbound

This paper cites Goal-Conditioned Reinforcement Learning: Problems and Solutions.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-09T12:49:59.106851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:49:59.106851Z digest=sha256:5e14ba251a9f6ab68af8e22d35daa5621be79ab5c1d073dfb189bdfee8863709

Observation 88376a46-728d-4a23-aacb-6bdf8e95806a · outbound

This paper cites A., Veness, J., Bellemare, M.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control A., Veness, J., Bellemare, M

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-09T12:49:59.111561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:49:59.111561Z digest=sha256:32d75468647e3d5ea4c3416aa5ccfef80e4531f51f0591209c3e192fd5d5cfba

Observation e285dbe8-e8e8-46a8-a728-d4ce6efa5bc4 · outbound

This paper cites Goal-directed planning via hindsight experience replay.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Goal-directed planning via hindsight experience replay

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.402026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.115909Z digest=sha256:3e7df03adc09adc54d586bb96414fddad101c56006a0bb6b8cd0a2ae4484e19e

Observation cb9b5b77-5a19-4227-a0b5-cdb3e0cbddc8 · outbound

This paper cites A., and Darrell, T.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control A., and Darrell, T

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-09T12:49:59.120185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:49:59.120185Z digest=sha256:79d9a12f004def96a6f0f4effb82a387fc9551767a933b5c69f0229d971006fd

Observation 448fbb73-893e-477b-9963-880893f87a36 · outbound

This paper cites Dynamic-matching adaptive sliding mode control for hypersonic vehicles.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Dynamic-matching adaptive sliding mode control for hypersonic vehicles

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.377494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.124777Z digest=sha256:6afc998a0b2abfe197c017dc01d8fb7ab3f6c7dc760b20528a061dd3d164d8d5

Observation 06734728-c786-440a-97ac-89c7bdb52c5b · outbound

This paper cites Discovering blind spots in reinforcement learning.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Discovering blind spots in reinforcement learning

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.362366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.129734Z digest=sha256:4e692e522ae2ad300649c766267f5db1c73fa4ba4d872c2d2c1c86359b48a4a8

Observation c6bec956-1b08-41e5-97e5-6f9dba0e8240 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Proximal Policy Optimization Algorithms

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-09T12:49:59.134118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:49:59.134118Z digest=sha256:56e845e6f80b450b57092e2d80745c86a2e180d5aeb82f0fbeae2a8247a47582

Observation 2cdffbd7-5206-4799-8fe2-ff334b6439bd · outbound

This paper cites Deterministic policy gradient algorithms.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Deterministic policy gradient algorithms

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-09T12:49:59.138954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:49:59.138954Z digest=sha256:b9e9ab458f93973322d316f9126fad22b357381ebeb14a6937806bea58d824e7

Observation f353c4f6-4994-4597-959e-d7c5f1028dee · outbound

This paper cites Action robust reinforcement learning and applications in continuous control.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Action robust reinforcement learning and applications in continuous control

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-09T12:49:59.143546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:49:59.143546Z digest=sha256:570da10b9e882da15910702d64431e23c9925deaecac3d03618d78083f4ae951

Observation 407097ac-ed4c-47a2-925a-55e19f8e6c23 · outbound

This paper cites D., Michi, A., Chervonyi, Y., Davies, I., Paduraru, C., Lazic, N., Felici, F., Ewalds, T., Donner, C., Galperti, C., et al.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control D., Michi, A., Chervonyi, Y., Davies, I., Paduraru, C., Lazic, N., Felici, F., Ewalds, T., Donner, C., Galperti, C., et al

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.325736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.148057Z digest=sha256:d1e6330d8e6f26e63f7bd6d7223e46fec701f017580bcd36aea23b096f611df1

Observation 69cefe13-0d49-40f6-9cef-8c89416365ee · outbound

This paper cites A brief survey on nonlinear control using adaptive dynamic programming under engineering-oriented complexities.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control A brief survey on nonlinear control using adaptive dynamic programming under engineering-oriented complexities

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.310143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.152626Z digest=sha256:9408094eb658b3b90ecb0c9f9375ddbdf007243c45cc728ae8ce3c738ab1e72e

Observation 8a103e4a-59e7-478a-9bd3-c47d20057cfc · outbound

This paper cites P., Qingqing, L., and Westerlund, T.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control P., Qingqing, L., and Westerlund, T

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.294548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.156933Z digest=sha256:2bffe5f514cb30244d964445ec610554e75ea4534f5b3e9c55e591282cc85fc2

Observation e0d97cee-bc2e-4aa1-8468-00c75548ee32 · outbound

This paper cites P., and Westerlund, T.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control P., and Westerlund, T

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.279619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.161470Z digest=sha256:ad4b2ebbb6457879507aa880ef7fe2c1bec3c90bfa45c21d96bdacae9ad4952e

Observation b34a8f89-1956-4252-bf22-d8918cfa874a · outbound

This paper cites A comprehensive survey on transfer learning.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control A comprehensive survey on transfer learning

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.262839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.165901Z digest=sha256:b0dcd2b4e90a69f7895f1767bd6b5e1782d33fdff8ef7e134a245c3bccec92e1

Observation 82e864a6-f2f0-4395-a749-cc4b93895863 · outbound

This paper cites write newline.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control write newline

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-09T12:49:59.170317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:49:59.170317Z digest=sha256:93bcb23dfd6ee6834ccebc75f976718ef38c9a6b2e9f0c6dceac6bd0d24102f9

Pith citing papers

Observation ff0e69d0-98c0-4cf9-9d63-2b1a0dafcd24 · inbound

Tracking the Effective Surface Area of Non-Convex Satellites cites this paper.

Tracking the Effective Surface Area of Non-Convex Satellites Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-06-27T15:31:00.509587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-06-27T15:27:51.585099Z digest=sha256:4f2dec86bd2c08e58c50950baefed9c37d7c799331647983dca782c22f8e7360