Pith. sign in

Paper Citation Record · LEDGER

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control

As of 22 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 1 inbound Pith citation observation for arXiv:2502.02265.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.02265 v1

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T12:49:59.170317Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-27T15:27:51.585099Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

30 of 30 outbound references displayed

  • verified exact0
  • verified fuzzy18
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 005660fd-f11d-40c9-bf38-a89d949fdacc · outbound

This paper cites Hindsight experience replay.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Hindsight experience replay

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.607414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.035484Z digest=sha256:7e6b5c73077121fe131772b2f63be9d813d809dbc15ad98da8362571d97b3bc0

Observation 6e179587-d8f0-4ba9-86c4-665a4ad47c42 · outbound

This paper cites P., Maghade, D., Sondkar, S., and Pawar, S.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control P., Maghade, D., Sondkar, S., and Pawar, S

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.592147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.040935Z digest=sha256:5e376eb0b2ff1e2867484c25e8bbfc8a9fc19f716120170fc71cedc5b01ad124

Observation abd3d449-c43e-4b46-a973-dc9bf510f0a0 · outbound

This paper cites Exploration by random network distillation.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Exploration by random network distillation

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.576396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.045753Z digest=sha256:54eb241e5a67a5e68bd2d2d067f5d225a992a21684ed888b07edf5a2f198f480

Observation 03dd188a-2241-4d44-9027-1269f5a20e47 · outbound

This paper cites Reinforcement learning for control: Performance, stability, and deep approximators.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Reinforcement learning for control: Performance, stability, and deep approximators

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.561296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.050519Z digest=sha256:6b1d88c5ceb73cb73d89586fb41cc50093b7ac059b0612cc0c7ec8512cbff05c

Observation 008b4ff0-7dca-470a-9775-df1beccedce0 · outbound

This paper cites Learning sim-to-real dense object descriptors for robotic manipulation.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Learning sim-to-real dense object descriptors for robotic manipulation

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.546171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.055719Z digest=sha256:a29680fe434daacb0970fbe5201784d9ce1553cebca4a827506ca341c8269f08

Observation a60f5adf-8c81-4a66-85e9-48918bc402b3 · outbound

This paper cites Neural-network-based nonlinear optimal terminal guidance with impact angle constraints.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Neural-network-based nonlinear optimal terminal guidance with impact angle constraints

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.531047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.060598Z digest=sha256:0793d57eb8e93c82767c4ae1e16a667b5635af559fe264bf1c109ee076bca23e

Observation bb575392-bd9b-47db-9728-84df08a3f9bf · outbound

This paper cites an unresolved cited work.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-09T12:49:59.515527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.065842Z digest=sha256:d19067647527c89a8a985e211dfd2fa74cfeee9e6fea4d9d3fdf4d158472a1ed

Observation 86175193-adb9-4711-9cb5-0aa029de2293 · outbound

This paper cites Addressing function approximation error in actor-critic methods.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Addressing function approximation error in actor-critic methods

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T12:49:59.070277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:49:59.070277Z digest=sha256:bf78db7a3647416c3871b9a6da338aabb0ab3e270a1e883e2aa995f546fd40fe

Observation e0797492-a89d-44a9-9f02-1cc9270fa322 · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-09T12:49:59.075006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:49:59.075006Z digest=sha256:f91faab970abe5b71ab3c732e6aa251a68a68e12e5923b585b429fe97ed712f2

Observation faf70b9f-6882-46ae-84e0-f2c305af9901 · outbound

This paper cites Learning to utilize shaping rewards: A new approach of reward shaping.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Learning to utilize shaping rewards: A new approach of reward shaping

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.481764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.079365Z digest=sha256:37bd9597c7c3a52baee6629bdbced9555d23d59a8e5e0d24442c0466d831351b

Observation 48605718-3d47-4a0b-a625-fdb1949b5017 · outbound

This paper cites D., Tsounis, V., Hwangbo, J., Bodie, K., Fankhauser, P., Bloesch, M., et al.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control D., Tsounis, V., Hwangbo, J., Bodie, K., Fankhauser, P., Bloesch, M., et al

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.466447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.083704Z digest=sha256:67c65f36ab182a3a383220e0983cf26fe9f892d81b87055ec4329df29c4478d2

Observation 948f4ce5-1522-4428-b26e-620c297cfa90 · outbound

This paper cites R., Sobh, I., Talpaert, V., Mannion, P., Al Sallab, A.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control R., Sobh, I., Talpaert, V., Mannion, P., Al Sallab, A

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-09T12:49:59.088067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:49:59.088067Z digest=sha256:459d2273f12ada4869c009a4e279b30d41b9f5ad72d688423a4ee0ac0a632b9b

Observation 8d6f2deb-e3db-46b5-b1b5-7b9f5a4a4dfd · outbound

This paper cites H., and Chong, G.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control H., and Chong, G

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.442341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.092820Z digest=sha256:6048fc04708b2e77b170ed25acec823bba0d28a54294ca4ab56c19b008dc4945

Observation 7b3ee027-c958-4823-8cad-d6c96086fd1f · outbound

This paper cites Continuous control with deep reinforcement learning.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Continuous control with deep reinforcement learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-09T12:49:59.097406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:49:59.097406Z digest=sha256:8f556526c5d265a7a3446fa495c7669b70780dec94bb5311a8463267fb138ee9

Observation 19f774e0-a4fd-48fb-8f70-91d5a83473ba · outbound

This paper cites A review of industrial mimo decoupling control.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control A review of industrial mimo decoupling control

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.427187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.102440Z digest=sha256:bf053015e8ff33713e2cb4aa5f3aa365e3a30ba3a0fd0f29044b6e93b478b8ce

Observation d5e18518-2e86-4153-bfd2-06607d1a89fd · outbound

This paper cites Goal-Conditioned Reinforcement Learning: Problems and Solutions.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-09T12:49:59.106851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:49:59.106851Z digest=sha256:5e14ba251a9f6ab68af8e22d35daa5621be79ab5c1d073dfb189bdfee8863709

Observation 88376a46-728d-4a23-aacb-6bdf8e95806a · outbound

This paper cites A., Veness, J., Bellemare, M.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control A., Veness, J., Bellemare, M

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-09T12:49:59.111561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:49:59.111561Z digest=sha256:32d75468647e3d5ea4c3416aa5ccfef80e4531f51f0591209c3e192fd5d5cfba

Observation e285dbe8-e8e8-46a8-a728-d4ce6efa5bc4 · outbound

This paper cites Goal-directed planning via hindsight experience replay.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Goal-directed planning via hindsight experience replay

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.402026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.115909Z digest=sha256:ec45f453ce1874a89c290e55b9ba8bfde76f5b67479c561bf9cc74127608da66

Observation cb9b5b77-5a19-4227-a0b5-cdb3e0cbddc8 · outbound

This paper cites A., and Darrell, T.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control A., and Darrell, T

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-09T12:49:59.120185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:49:59.120185Z digest=sha256:79d9a12f004def96a6f0f4effb82a387fc9551767a933b5c69f0229d971006fd

Observation 448fbb73-893e-477b-9963-880893f87a36 · outbound

This paper cites Dynamic-matching adaptive sliding mode control for hypersonic vehicles.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Dynamic-matching adaptive sliding mode control for hypersonic vehicles

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.377494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.124777Z digest=sha256:a137df49bce551084ae36f204200c905e3d1e79575146ebc9df6d491a91c1e96

Observation 06734728-c786-440a-97ac-89c7bdb52c5b · outbound

This paper cites Discovering blind spots in reinforcement learning.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Discovering blind spots in reinforcement learning

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.362366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.129734Z digest=sha256:34e6560f78aba153f7600ca2e7ef3689d5a96330243d636d13d32795bdcc27a9

Observation c6bec956-1b08-41e5-97e5-6f9dba0e8240 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Proximal Policy Optimization Algorithms

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-09T12:49:59.134118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:49:59.134118Z digest=sha256:2c05c98d859d1497d47cfb600eab1a9be55fe7d18ce3aaf3980f5538c5019557

Observation 2cdffbd7-5206-4799-8fe2-ff334b6439bd · outbound

This paper cites Deterministic policy gradient algorithms.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Deterministic policy gradient algorithms

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-09T12:49:59.138954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:49:59.138954Z digest=sha256:b9e9ab458f93973322d316f9126fad22b357381ebeb14a6937806bea58d824e7

Observation f353c4f6-4994-4597-959e-d7c5f1028dee · outbound

This paper cites Action robust reinforcement learning and applications in continuous control.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control Action robust reinforcement learning and applications in continuous control

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-09T12:49:59.143546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:49:59.143546Z digest=sha256:570da10b9e882da15910702d64431e23c9925deaecac3d03618d78083f4ae951

Observation 407097ac-ed4c-47a2-925a-55e19f8e6c23 · outbound

This paper cites D., Michi, A., Chervonyi, Y., Davies, I., Paduraru, C., Lazic, N., Felici, F., Ewalds, T., Donner, C., Galperti, C., et al.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control D., Michi, A., Chervonyi, Y., Davies, I., Paduraru, C., Lazic, N., Felici, F., Ewalds, T., Donner, C., Galperti, C., et al

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.325736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.148057Z digest=sha256:76c8268f4873e7978beb5407db00562c60c35cbaf267b1e451261e540341d8d6

Observation 69cefe13-0d49-40f6-9cef-8c89416365ee · outbound

This paper cites A brief survey on nonlinear control using adaptive dynamic programming under engineering-oriented complexities.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control A brief survey on nonlinear control using adaptive dynamic programming under engineering-oriented complexities

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.310143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.152626Z digest=sha256:ec50856092a9cfb2e22ee0b4d393ad0215d4f7a0fe56378504879f2bd77746de

Observation 8a103e4a-59e7-478a-9bd3-c47d20057cfc · outbound

This paper cites P., Qingqing, L., and Westerlund, T.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control P., Qingqing, L., and Westerlund, T

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.294548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.156933Z digest=sha256:f5fcce49ef1448dbe91187f0f5152f20b02f783dada22f7796b11b598af944ae

Observation e0d97cee-bc2e-4aa1-8468-00c75548ee32 · outbound

This paper cites P., and Westerlund, T.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control P., and Westerlund, T

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.279619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.161470Z digest=sha256:19330b03bdcba444bce98dcdadc9318bf6d2df0e5140658e68ec90623dd609af

Observation b34a8f89-1956-4252-bf22-d8918cfa874a · outbound

This paper cites A comprehensive survey on transfer learning.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control A comprehensive survey on transfer learning

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:49:59.262839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-09T12:49:59.165901Z digest=sha256:4c8d78d3c9ff06809573435b660e8e94a3976dd29452f9e69037043c97cada91

Observation 82e864a6-f2f0-4395-a749-cc4b93895863 · outbound

This paper cites write newline.

Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control write newline

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-09T12:49:59.170317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:49:59.170317Z digest=sha256:93bcb23dfd6ee6834ccebc75f976718ef38c9a6b2e9f0c6dceac6bd0d24102f9

Pith citing papers

Observation ff0e69d0-98c0-4cf9-9d63-2b1a0dafcd24 · inbound

Tracking the Effective Surface Area of Non-Convex Satellites cites this paper.

Tracking the Effective Surface Area of Non-Convex Satellites Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-06-27T15:31:00.509587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-06-27T15:27:51.585099Z digest=sha256:766544519046fc659f2123b584bdf32c8dba4e5c52e3673bd197a87475cccf90