Pith. sign in

Paper Citation Record · LEDGER

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change

As of 19 August 2026, this Paper Citation Record lists 100 of 246 outbound references and 0 inbound Pith citation observations for arXiv:2505.10330.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.10330 v1

Coverage vector

measured 100 of 246 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:16:00.018724Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 246 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved99
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8bb213dc-9450-46f7-b3fd-2b6099c64b29 · outbound

This paper cites Mastering the game of go without human knowledge,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Mastering the game of go without human knowledge,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.069279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.069279Z digest=sha256:7eef9ae03a06cfebe77777ad29f93afdc6a45861050088aa39b4b77610d98185

Observation f1d598b9-b60d-44cb-bcea-26b056cab1bf · outbound

This paper cites Mastering atari, go, chess and shogi by planning with a learned model,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Mastering atari, go, chess and shogi by planning with a learned model,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.077876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.077876Z digest=sha256:238f4c8c0ff78dc99594761d3aab73cc1b969c23801262111c687a75d47c6468

Observation c49256e1-822a-4cd1-9957-b794514bb4ab · outbound

This paper cites Grandmaster level in starcraft ii using multi-agent reinforcement learning,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Grandmaster level in starcraft ii using multi-agent reinforcement learning,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.084850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.084850Z digest=sha256:19b50b31180d30e414790e753d0b25890ff3db9abcb98247d0a7f0ac8e2c19f0

Observation c5f77ab1-08bf-4a91-b0e4-b52429619df8 · outbound

This paper cites Dota 2 with Large Scale Deep Reinforcement Learning.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Dota 2 with Large Scale Deep Reinforcement Learning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.091559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.091559Z digest=sha256:52fde7554fdbb3ffe2cac32683bfb07192277df89d94511c5244c36529839a13

Observation 8549a4f1-5811-4511-98c9-6885b84f8aee · outbound

This paper cites Agent57: Outperforming the human atari benchmark,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Agent57: Outperforming the human atari benchmark,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.099808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.099808Z digest=sha256:5c3b4ba69b85ca0da03ad41e6cf7690866bcb2f19f76b5ecffb4c9faa78763dc

Observation a9ad6ea3-71c7-494c-9814-82a1b315d9e0 · outbound

This paper cites Reinforcement learning based recommender systems: A survey,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Reinforcement learning based recommender systems: A survey,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.113283Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.113283Z digest=sha256:c75679b90fe0bad6399859bb884db594ca08901091096837741ec3aee95df049

Observation 8787a233-f3d0-4d8f-9132-a9af6cf52f1d · outbound

This paper cites Deepmind ai reduces google data centre cooling bill by 40%,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Deepmind ai reduces google data centre cooling bill by 40%,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.126194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.126194Z digest=sha256:303de7c2fa97ff1ad70d5a09382aa6503ad9749384cfdd5ee8f5f209458eda76

Observation 646eedef-cc8e-4f49-b633-505c07d9fab7 · outbound

This paper cites Gnu-rl: A precocial reinforcement learning so- lution for building hvac control using a differentiable mpc policy,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Gnu-rl: A precocial reinforcement learning so- lution for building hvac control using a differentiable mpc policy,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.133657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.133657Z digest=sha256:721d4fa2cb7e219415c4ff96e774cf4d49eb25f06b338331dfdd8b3c542c1083

Observation b0a39276-513c-4fb5-9374-9fa80e2814a4 · outbound

This paper cites Magnetic control of tokamak plasmas through deep reinforce- ment learning,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Magnetic control of tokamak plasmas through deep reinforce- ment learning,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.145334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.145334Z digest=sha256:072cd6e6827c90765d4c6413a1d5a1454fa8dde7f6a29cd8b5fd65535dab1097

Observation 7edde538-4863-4b5a-83d5-8bfb448839a6 · outbound

This paper cites Adversarial policies beat superhuman go ais,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Adversarial policies beat superhuman go ais,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.155739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.155739Z digest=sha256:858c720babd1972d002623d6f4118250928d5d11cf6af02e9ed17f51e5aa0310

Observation af2d9cac-75a0-4ae1-b2b1-61604bdac3ad · outbound

This paper cites Adaptation in constant utility non-stationary environments.,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Adaptation in constant utility non-stationary environments.,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.161373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.161373Z digest=sha256:9a152ae04b47428c2bfa395d880e4aea4e715a2c00a6c46dc13527e7ceeb4617

Observation 074e49ff-b87b-447a-a19c-54a6f7d03619 · outbound

This paper cites an unresolved cited work.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.175120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.175120Z digest=sha256:a2dabe18e87d25324f40630d268540a4b95cfd0d11f2a5c40e06791464d6e24c

Observation a2369159-0aec-464f-976d-cbb0314a3961 · outbound

This paper cites Impact of timing in post-warning prepositioning decisions on performance measures of disaster management: A 163 real-life application,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Impact of timing in post-warning prepositioning decisions on performance measures of disaster management: A 163 real-life application,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.181875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.181875Z digest=sha256:7e51f0ab076da28caea9582b8a6a1c69d2d7611d0695fdc17bcbd82e53c8db27

Observation bee051dd-7a95-4a27-a96b-1dbfbd3ba6f1 · outbound

This paper cites Domain randomization for transferring deep neural networks from simulation to the real world,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Domain randomization for transferring deep neural networks from simulation to the real world,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.190382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.190382Z digest=sha256:75e541042ae80f5b2ddc0c3f6e4534efd7a6e1fe17f441f2ae1247f3694b4ef5

Observation f6f5d207-465d-4eeb-ac17-ec1b171158f6 · outbound

This paper cites Learning optimal adap- tation strategies in unpredictable motor tasks,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Learning optimal adap- tation strategies in unpredictable motor tasks,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.198548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.198548Z digest=sha256:cac6bf4fb256da7e7538a522f680ca1f40b3eeb40d3de680233a1ab2539203b7

Observation f8612988-d5f0-4eef-9623-d2bd9a89fe36 · outbound

This paper cites Reward learning: Reinforcement, incentives, and expectations,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Reward learning: Reinforcement, incentives, and expectations,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.210439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.210439Z digest=sha256:92d9e056676aad29f5632f0f430691a8aa9b9fae9af328acc2a1f10f40f60551

Observation d62bc700-fd5f-499a-90a3-9ed6e906c8ec · outbound

This paper cites Towards precision holography.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Towards precision holography

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.221123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.221123Z digest=sha256:c6ae27afeef3760e952f5bbbecb2408402b91d665c5607214eddcbad1167fe37

Observation c23d9383-47bb-428d-a8c1-3c1eb5e75a59 · outbound

This paper cites Deep neural networks for youtube rec- ommendations,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Deep neural networks for youtube rec- ommendations,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.227562Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.227562Z digest=sha256:a42ff1e70d7b7aaefb3b070fcf114e14bd77030271417ac3865943c9fe764489

Observation 1accb39b-d7ba-4401-9899-aa4f85be4955 · outbound

This paper cites Scheduling on a budget: Avoiding stale recommendations with timely updates,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Scheduling on a budget: Avoiding stale recommendations with timely updates,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.236715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.236715Z digest=sha256:a388e868e70a89d5d2c2cf0abc6e3db39130142911c38bba5ccfb0da674c252a

Observation ab83c578-080a-4238-abbc-def669a488b2 · outbound

This paper cites Catastrophic interference in connectionist net- works: The sequential learning problem,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Catastrophic interference in connectionist net- works: The sequential learning problem,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.249385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.249385Z digest=sha256:c97b03fd41cbecb0614ac87c1aaeb9b6d3f2918c054ede5fb682b9baa4ee9431

Observation 3b8e47fc-3c03-43bf-86b5-4a19971b3148 · outbound

This paper cites Learning to predict by the methods of temporal differences,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Learning to predict by the methods of temporal differences,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.257424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.257424Z digest=sha256:1cf2a103ef624f0d554e2f9e54adcc2a651f5fe733280c2f808e9eb7acdc8b54

Observation a86c677b-3431-4fd4-b80d-033953f4dd8e · outbound

This paper cites Analysis of temporal-diffference learning with func- tion approximation,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Analysis of temporal-diffference learning with func- tion approximation,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.267179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.267179Z digest=sha256:ddbcf96aa2cd68261d2cd2ec1ac3cf0a27af5fc060ef2ab69fe24fb680490801

Observation 3c6ccfc8-6e62-4ea6-9593-08ac31f9e694 · outbound

This paper cites Human-level control through deep reinforcement learning,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Human-level control through deep reinforcement learning,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.277183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.277183Z digest=sha256:f2ecb67636555a6be34cf8525144175fab3f81e443416cb646a9aa363184b302

Observation 0c6c19b7-8dff-4885-a1c6-a42f99cfcad4 · outbound

This paper cites Deep Reinforcement Learning and the Deadly Triad.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Deep Reinforcement Learning and the Deadly Triad

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.289096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.289096Z digest=sha256:273d9f0cab69e30ce2047e2186e51d46d81434757e4e3d42db05517cfce5efc3

Observation f7e9ba2d-1b8c-498f-95b0-416f42011241 · outbound

This paper cites Simple statistical gradient-following algorithms for connectionist reinforcement learning,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Simple statistical gradient-following algorithms for connectionist reinforcement learning,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.299736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.299736Z digest=sha256:17da7e60d62cbf87f03277934fc4ff451a2b807c33c6fdff0d75f78617da9dfd

Observation 35b57e3e-b9be-41c7-ba46-c22dc1ce44b0 · outbound

This paper cites Asynchronous methods for deep reinforcement learning,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Asynchronous methods for deep reinforcement learning,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.309292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.309292Z digest=sha256:6ff797fe42689388edd60d11428f7c1815b644f58fb0a849dc2d715a453d7efa

Observation a41617e8-6cbd-47b7-b593-472866fd3b01 · outbound

This paper cites Impala: Scalable distributed deep-rl with importance weighted actor-learner architectures,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Impala: Scalable distributed deep-rl with importance weighted actor-learner architectures,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.316184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.316184Z digest=sha256:33c1b846381f226c507b9cf42cf094b2604d5baff250490fd4b8549a11f5f089

Observation 16062bad-540c-40c8-8a77-2a1b55b195bc · outbound

This paper cites Continuous control with deep reinforcement learning,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Continuous control with deep reinforcement learning,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.324781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.324781Z digest=sha256:105c671248f6804923c08d134c083c9f7438c3af680e2d8532321778b2cae625

Observation 945f7532-0f0b-4353-ae67-1ee4d13088d4 · outbound

This paper cites Distributed distributional deterministic policy gradients,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Distributed distributional deterministic policy gradients,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.335060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.335060Z digest=sha256:f349b659b7bdcab3b813610fb83bdc352b012751a0ab0a620851c3e20c86ea3d

Observation 62c1d1da-3bcd-4d85-8c46-9edcb8be888e · outbound

This paper cites Scalable trust-region method for deep reinforcement learning using kronecker-factored approximation,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Scalable trust-region method for deep reinforcement learning using kronecker-factored approximation,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.346698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.346698Z digest=sha256:91fcb0acd78c478380a972d7df55fe17e0d79711d26cba281987fe58772898a9

Observation 02d67d68-b000-47bb-a1dc-d36558a3c55f · outbound

This paper cites Soft actor-critic: Off-policy maxi- mum entropy deep reinforcement learning with a stochastic actor,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Soft actor-critic: Off-policy maxi- mum entropy deep reinforcement learning with a stochastic actor,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.354303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.354303Z digest=sha256:b656c6bf1a959c2bf4d8a71007837933e089e96cafefcf7636f21de3aab3ff57

Observation 58588fe6-dc46-4e88-aba3-8b8681a924ef · outbound

This paper cites Trust region pol- icy optimization,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Trust region pol- icy optimization,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.369754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.369754Z digest=sha256:fac90eca9881345da34784aef0e39c3dd7b34cf5ed9403f71750f568523a97a0

Observation 88e42701-3dea-44a9-99d1-2e07398935cd · outbound

This paper cites Proximal Policy Optimization Algorithms.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Proximal Policy Optimization Algorithms

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.377816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.377816Z digest=sha256:0bc43f1ba70bc05e345b23511ee046404a8121c39985425ed49b240855780aae

Observation dd4d2e19-2898-4e64-bdbc-e559d47f153a · outbound

This paper cites Dyna, an integrated architecture for learning, planning, and reacting,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Dyna, an integrated architecture for learning, planning, and reacting,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.384898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.384898Z digest=sha256:646b0ffb0d38d6791fa188190255e0d0ba99f4d3c96d5f70977fddea158574cc

Observation 6bb3b444-52e1-4069-b7d3-3ef783a3f54f · outbound

This paper cites First return, then explore,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change First return, then explore,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.395462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.395462Z digest=sha256:23f5865f21c338f8839a67f282c85ddf074ba55a8982c79d8c4eb18c148af8f9

Observation 4a0ed717-3cb0-47e0-9674-b2ce1d30610f · outbound

This paper cites Learning latent dynamics for planning from pixels,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Learning latent dynamics for planning from pixels,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.406216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.406216Z digest=sha256:54486701b0ab2daf8d0301b1c4f77c4d7481db1e8dc6ad8cae339b8d20d1850a

Observation 6ec25d40-c564-4cfc-9958-99194229e867 · outbound

This paper cites Curious model-building control systems,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Curious model-building control systems,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.418445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.418445Z digest=sha256:dd8c5b29a63b44ef5822e4472ea4c58ecd1d7ad4aa23f29969e1665cb72c1e9a

Observation f56d2ea2-50f9-4060-9e51-c18d150b418c · outbound

This paper cites On Learning to Think: Algorithmic Information Theory for Novel Combinations of Reinforcement Learning Controllers and Recurrent Neural World Models.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change On Learning to Think: Algorithmic Information Theory for Novel Combinations of Reinforcement Learning Controllers and Recurrent Neural World Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.428205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.428205Z digest=sha256:47a910fcb219eb4088352671149e852c00b3a73b2323adc4ca2ff5a6011ec7e0

Observation 18c87a44-82aa-4a6c-a073-bbf03641872c · outbound

This paper cites Recurrent world models facilitate policy evolution,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Recurrent world models facilitate policy evolution,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.437614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.437614Z digest=sha256:05028ff17533be3db5036820acf988e703e126029a2b0b16e310a46ac8ca86af

Observation ed22c21e-7630-4db0-bb89-f39ec53129d3 · outbound

This paper cites Dream to control: Learning behav- iors by latent imagination,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Dream to control: Learning behav- iors by latent imagination,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.448190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.448190Z digest=sha256:634b606588faf5d119471162178cd7cb6623ce79eed0dd699b42e3a44b511476

Observation 7fc91ebd-9bd0-4bde-b830-847e45b5a062 · outbound

This paper cites Mastering atari with discrete world models,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Mastering atari with discrete world models,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.464751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.464751Z digest=sha256:ff8f03683a6968948d89f3e23530dc317af129301c7eac61c9382cf63f594f40

Observation f04e09b9-1be3-458e-aa98-d378d4291cd3 · outbound

This paper cites Mastering Diverse Domains through World Models.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Mastering Diverse Domains through World Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.472534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.472534Z digest=sha256:4f2f770ad93cb6e62e9697218f3bf8c71188af65d893c5ab26685d43795fe382

Observation 5b345e16-a794-4728-a06a-8084bf0f0700 · outbound

This paper cites On the Properties of Neural Machine Translation: Encoder-Decoder Approaches.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change On the Properties of Neural Machine Translation: Encoder-Decoder Approaches

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.485835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.485835Z digest=sha256:a13da214e2b37661c843b65663aa68206282d3f08c417dfd6a45c11801789173

Observation 67b5deec-6cfc-4b56-b67c-f299ef73d8bd · outbound

This paper cites Convolutional networks for images, speech, and time series,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Convolutional networks for images, speech, and time series,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.492201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.492201Z digest=sha256:7dc7a1c2c331cdf63d2300c24c20132ddf3bff85f05ce902568a3a862b40328c

Observation c7d134be-e586-4825-9cd3-1e8f6610bc10 · outbound

This paper cites Estimating or Propagating Gradients Through Stochastic Neurons for Conditional Computation.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Estimating or Propagating Gradients Through Stochastic Neurons for Conditional Computation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.498195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.498195Z digest=sha256:f63c3e0148f4b8a6786458e66534310e444557b0a9ee7b252185519f0bc96597

Observation 6f2691ef-5777-46ef-b275-9dfce39284dc · outbound

This paper cites Dm control: Software and tasks for continuous con- trol,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Dm control: Software and tasks for continuous con- trol,

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.507948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.507948Z digest=sha256:93283fae29074cf2c45d40f1fe6c67864feea302b2bdc5bdf918e653c1fcd4e8

Observation 0e902dc0-ae37-426c-976e-cd0e7fdcf3c7 · outbound

This paper cites The arcade learning environment: An evaluation platform for general agents,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change The arcade learning environment: An evaluation platform for general agents,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.515230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.515230Z digest=sha256:d08943720846db66082f7a8bd99f35df353d24bf934cc6c89241f4ec2aed19b5

Observation 373b9b0a-ee95-4973-9969-b4f351d66902 · outbound

This paper cites Policy invariance under reward transforma- tions: Theory and application to reward shaping,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Policy invariance under reward transforma- tions: Theory and application to reward shaping,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.521131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.521131Z digest=sha256:b7f3f07a4b8e92bc23d226f36bd6864e1e5b3f9e75560e29c4ef2f04b41d743d

Observation d46664b3-7be0-4861-9f84-5e1cc92ec8ce · outbound

This paper cites Self-improving reactive agents based on reinforcement learning, plan- ning and teaching,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Self-improving reactive agents based on reinforcement learning, plan- ning and teaching,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.530020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.530020Z digest=sha256:08f324b65264fa8128ffee7f926f228c638fd5d9a18b1aea4a771f4004ea639a

Observation badbb7e8-9f76-4d1f-946d-d40e6b9ec770 · outbound

This paper cites Sample efficient actor-critic with experience replay,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Sample efficient actor-critic with experience replay,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.537162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.537162Z digest=sha256:dcad0ce3c382294c6b8b80679d26212733bf85d611129b4324fcbcc54d969a74

Observation 63dc1484-2984-4929-8633-a13d189143e0 · outbound

This paper cites Rainbow: Combining improvements in deep reinforcement learn- ing,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Rainbow: Combining improvements in deep reinforcement learn- ing,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.545797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.545797Z digest=sha256:f04cf96266ee7e9260c1a7c467464b9609116a59183222cb9547e8795da073e6

Observation 2c3f1207-1be5-45d3-911d-b2b49a2c82c4 · outbound

This paper cites A Deeper Look at Experience Replay.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change A Deeper Look at Experience Replay

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.552889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.552889Z digest=sha256:b43fd02985199504df37e39d234a05969604c742596a220af41d8590d20480b1

Observation a84b1911-47e9-4e67-8861-60e084505eb1 · outbound

This paper cites Revisiting fundamentals of experience replay,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Revisiting fundamentals of experience replay,

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.560650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.560650Z digest=sha256:1e7b262da0f69d6f0678e840ec35dbaebab2000cc0bc9b9571138eeb6de5a65a

Observation 2be1cbd3-8fe3-4fd6-b634-012b09fd6ff4 · outbound

This paper cites Prioritized experience replay,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Prioritized experience replay,

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.571827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.571827Z digest=sha256:47099689fbb9598b3cfdb894d5d901ed8a3c63985360079e77538759695a7883

Observation 887dd01e-658c-4e6b-b6a4-58075e686d7d · outbound

This paper cites Prioritized experience replay method based on experience reward,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Prioritized experience replay method based on experience reward,

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.579543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.579543Z digest=sha256:1c6eec17ecc3c7186af7cd28ca0ed6b3de7724623dee5bef7655223370ab36f9

Observation bf2d5765-a429-4875-9021-55001a1a28f7 · outbound

This paper cites Model-augmented prioritized experience replay,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Model-augmented prioritized experience replay,

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.587135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.587135Z digest=sha256:8c9dc64b2171ca9b5cf22e3c7e9729885b76f8f0a7d499c8fdf7c25f04334ae0

Observation 21827c24-b782-4c12-b54a-9cdff7bc22da · outbound

This paper cites Prioritized experience replay based on dynamics priority,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Prioritized experience replay based on dynamics priority,

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.593754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.593754Z digest=sha256:e940cd01f04960a1e6035a0999930b962ee23e057777d9876a0fee0a42d31c16

Observation 28880909-9ef5-4f2b-95e2-716081aa983e · outbound

This paper cites Image augmentation is all you need: Reg- ularizing deep reinforcement learning from pixels,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Image augmentation is all you need: Reg- ularizing deep reinforcement learning from pixels,

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.600732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.600732Z digest=sha256:0e50d1bc9157cee7f49244ac9c08abf1c67eff60e9639c0fd4b6664e83066a9d

Observation bb3261a3-2354-4fab-9fa7-3b8d165774e3 · outbound

This paper cites Temporal difference learning for model pre- dictive control,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Temporal difference learning for model pre- dictive control,

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.608644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.608644Z digest=sha256:fc7da93ebc86fda05dff56ddf3f8458bd3b44e590c61885fc264c922bfb2d463

Observation 951eaa83-08f4-44f7-875d-a76e668421e0 · outbound

This paper cites Transfer Learning in Deep Reinforcement Learning: A Survey.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Transfer Learning in Deep Reinforcement Learning: A Survey

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.617648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.617648Z digest=sha256:2af20ef1332d901cf1e0390a2c3d7b66b310fbc4a6d5e2623f5c2c74e384de10

Observation ac750a00-b921-4bb4-b3a7-bbb561a2299d · outbound

This paper cites Distilling the Knowledge in a Neural Network.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Distilling the Knowledge in a Neural Network

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.629259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.629259Z digest=sha256:1379de2e56839ab8695f16764373a94c122cc2c7ed716fc6c6530e71b7e813a0

Observation bbfd7cca-92d8-457e-8cdd-354b96510e35 · outbound

This paper cites Knowledge distillation: A survey,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Knowledge distillation: A survey,

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.642310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.642310Z digest=sha256:13e5ce0f49156e4a1c2dba89c75b4ca6e20e288d23e7a2239b27322e71abe498

Observation f5b1d7c9-131f-4cb5-a799-62ab7283ce40 · outbound

This paper cites Teaching on a budget: Agents advising agents in rein- forcement learning,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Teaching on a budget: Agents advising agents in rein- forcement learning,

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.650908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.650908Z digest=sha256:52aa34be18b3e7b8c05e3074c756153a348f4ded572ffae4850a4b061f24f045

Observation 65a7e75e-8a59-4218-8d61-f9c10140c608 · outbound

This paper cites Online transfer learning in reinforcement learning do- mains,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Online transfer learning in reinforcement learning do- mains,

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.661472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.661472Z digest=sha256:1683b56b0b3c35e93877e3e850c57895d68b3cd70cfb885c7a82ea89e48847b5

Observation 3fec6310-43f4-4baf-8400-abe9f017ca53 · outbound

This paper cites A survey on transfer learning,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change A survey on transfer learning,

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.673376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.673376Z digest=sha256:fe19396fe8c0502f31531788e5eb93b543e7d86aebfa1679b7c7634c55d542c5

Observation f6b7be33-429f-4471-aff3-7e626e0f478c · outbound

This paper cites Transfer learning for reinforcement learning domains: A survey.,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Transfer learning for reinforcement learning domains: A survey.,

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.682328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.682328Z digest=sha256:5c35ee5d49c5cc90a0b0b8879c434e02f11aa4fca398d45e55f634f0a026df14

Observation e75e4662-eb63-4180-9447-11933bcccff1 · outbound

This paper cites A Comprehensive Survey on Test-Time Adaptation under Distribution Shifts.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change A Comprehensive Survey on Test-Time Adaptation under Distribution Shifts

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.692146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.692146Z digest=sha256:5d31fff7bc8f798ca3d0ae35e20b263aec6923e05ba6895a444b95734ce68951

Observation 7d2f22a2-1761-4126-a62a-7f75083599cc · outbound

This paper cites A review of novelty detection,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change A review of novelty detection,

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.701112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.701112Z digest=sha256:93781c4846235948b22fe62bbecac12b2e559609542f2956417848a90148c65a

Observation 558d6a72-f865-4a75-8344-4d1000c6f88c · outbound

This paper cites Towards a unifying framework for formal theories of novelty,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Towards a unifying framework for formal theories of novelty,

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.709383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.709383Z digest=sha256:2f05616c6feff0329efd5eb6cb4daf4d22f5eaf70a97b0cbfbc9ba8573311035

Observation d4e83569-863d-413b-aa9a-e99ca2d53c82 · outbound

This paper cites Open-world learning for radically autonomous agents,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Open-world learning for radically autonomous agents,

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.717946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.717946Z digest=sha256:ef1e53b9c70e4fde1810b8f463f6dcfd62af19ac13bba3369f561baec4ba8b74

Observation 13fd4986-b491-4ed7-a7f1-c23a1214b0f3 · outbound

This paper cites Mixtbn: A fully test-time adaptation method for visual reinforce- ment learning on robotic manipulation,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Mixtbn: A fully test-time adaptation method for visual reinforce- ment learning on robotic manipulation,

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.730032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.730032Z digest=sha256:e86c939c2695328e7f9512a27512b509a00feedd64d36deb9b68c54636e7a48c

Observation 506e06de-0f6c-4e4c-967b-c2cfd0644945 · outbound

This paper cites Active test-time adaptation: Theoretical analyses and an algorithm,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Active test-time adaptation: Theoretical analyses and an algorithm,

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.745099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.745099Z digest=sha256:1c12338180cecfa74466a8914b68221258216458e7d11a5f0d353ac221c20cc1

Observation d69bb3f9-5363-4488-9c46-4db95f5d3cff · outbound

This paper cites Unknown sample discovery for source free open set domain adaptation,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Unknown sample discovery for source free open set domain adaptation,

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.754014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.754014Z digest=sha256:131bef2248b6b307f56586df5669b3c711a2c1efd33d8b0aa31a9c556abd1eb9

Observation 86a5621a-af8d-4542-bd5d-eff244394314 · outbound

This paper cites Hidden-mode markov decision processes for nonstationary sequential decision making,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Hidden-mode markov decision processes for nonstationary sequential decision making,

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.775138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.775138Z digest=sha256:fc07bb3458bfd713ebdfb4afb8b6fcdabbdf57775150f0324b1ed22781b4a1b2

Observation 55f29354-0515-4d56-89dd-d7eb3a83b818 · outbound

This paper cites Choosing search heuristics by non-stationary reinforcement learn- ing,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Choosing search heuristics by non-stationary reinforcement learn- ing,

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.786510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.786510Z digest=sha256:8c698b775a7ea37766f8aa137e5b35b29c2863c719d6daafcb9e697ab4a574ea

Observation 15c88ad6-90e1-4269-8d00-b9cf1aa2b763 · outbound

This paper cites Non-stationary reinforcement learning without prior knowledge: An optimal black-box approach,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Non-stationary reinforcement learning without prior knowledge: An optimal black-box approach,

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.799270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.799270Z digest=sha256:c8ae425278df8bbee7761aaf3dea5254b01becd4fbb3cdafb8f2cfcdb9cb6750

Observation 6474dc5c-ee3b-41bb-88f1-bbf9ddcfd52e · outbound

This paper cites Near-optimal model- free reinforcement learning in non-stationary episodic mdps,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Near-optimal model- free reinforcement learning in non-stationary episodic mdps,

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.811735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.811735Z digest=sha256:e86e98d02c97e17a69ee77a771899c52b253020146bbcb2aaf2c9b4d23b1d3d0

Observation 68f75129-2aab-4295-b69a-6b5d0005fe42 · outbound

This paper cites Non-stationary reinforcement learning under general function approximation,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Non-stationary reinforcement learning under general function approximation,

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.819317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.819317Z digest=sha256:9f1a1b09991e46bc70e66cecc0631dcf56eee1d9a2d70a61eb4a68ed9f48cf32

Observation f5ad77ed-1116-4b98-8967-2c6100a10b5d · outbound

This paper cites Addressing environment non-stationarity by repeat- ing q-learning updates,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Addressing environment non-stationarity by repeat- ing q-learning updates,

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.826698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.826698Z digest=sha256:7c3e9044acb3205adc11de4154ca31c663ca36649c5d44ae31fd37ba552ba23b

Observation 4fd715e3-9bdb-4ef1-b845-316ab26f6ef8 · outbound

This paper cites Non-stationary markov decision processes, a worst-case approach using model-based reinforcement learning,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Non-stationary markov decision processes, a worst-case approach using model-based reinforcement learning,

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.839173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.839173Z digest=sha256:b7a10b09e75a4666b130f358fef3c59ed165cd7f3a9506efbc43c48b9ebde3ce

Observation 9de3e21d-e44e-4e85-a7c4-4eb6ac2ec82a · outbound

This paper cites Reinforcement learning algorithm for non-stationary environments,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Reinforcement learning algorithm for non-stationary environments,

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.847130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.847130Z digest=sha256:20b15503bfda81ccb1b8560bbb4585b2b3823aa1ef5d3ba6347dafec58aa21fc

Observation 575e6f73-7b01-44d9-a2c4-484e39a41db0 · outbound

This paper cites Reactive exploration to cope with non-stationarity in life- long reinforcement learning,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Reactive exploration to cope with non-stationarity in life- long reinforcement learning,

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.858236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.858236Z digest=sha256:7d43d5fed9cb2fa2be0c153a5a4f8c99c29777da7e1f1e0942129a7fd97030e9

Observation 3644160b-304e-427c-b30b-62cfa6749b00 · outbound

This paper cites Transfer in reinforcement learning: A framework and a survey,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Transfer in reinforcement learning: A framework and a survey,

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.872781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.872781Z digest=sha256:f8b4a12559e6ea222c61375da3d4290cd401516ef57cbd9aa9403e0eaf01b382

Observation 9389818f-b3cd-40fd-98a9-52b142a18d76 · outbound

This paper cites Cross-modal domain adaptation for cost-efficient visual reinforcement learning,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Cross-modal domain adaptation for cost-efficient visual reinforcement learning,

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.881669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.881669Z digest=sha256:9ee9c19473037f6061319d8945c526c0ccf68d46d36d911d80034fd1028e7440

Observation 7879d77e-0e4d-4cb7-9b88-f95b63e0c972 · outbound

This paper cites Deep reinforcement learning amidst lifelong non- stationarity,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Deep reinforcement learning amidst lifelong non- stationarity,

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.889965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.889965Z digest=sha256:210415192c6e21b4ff7c3948709caa9c40265484b13ab3c49312da0b2810902f

Observation 56d56f0c-6e62-4b3d-b297-60328daf7490 · outbound

This paper cites Model-based nov- elty adaptation for open-world ai,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Model-based nov- elty adaptation for open-world ai,

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.900392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.900392Z digest=sha256:e5ab00eb58eddc08676510cfbb5143a93b8a703ef5046c5213c1ac8d13cb32ca

Observation 139fd434-fe0d-47ba-a6de-b7b3fabcfcf2 · outbound

This paper cites Detecting and adapting to novelty in games,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Detecting and adapting to novelty in games,

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.906124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.906124Z digest=sha256:8de65593617c024a2a081f4a6f284adfb76f318ddaa78ce9b7ef4247e84ee0c9

Observation 05e3ef52-8c99-4259-a8d1-4afd3b6bea1f · outbound

This paper cites Spotter: Extending symbolic planning operators through targeted reinforcement learning,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Spotter: Extending symbolic planning operators through targeted reinforcement learning,

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.915271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.915271Z digest=sha256:71dd610651949dde1a0ed3a1ec9370df078c7c98e20e7f6c502b34ff3ccd6886

Observation d3c11078-a8b5-47ed-a070-7c91546c2ea9 · outbound

This paper cites An integrated architecture for online adaptation to novelty in open worlds using probabilistic programming and novelty-aware planning,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change An integrated architecture for online adaptation to novelty in open worlds using probabilistic programming and novelty-aware planning,

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.929097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.929097Z digest=sha256:92b1ffe9b057f3b64ebb8a8c81d4259ecb70eef2eab14678474254d4bfc4a81d

Observation 03e8f72d-2128-4d16-828f-768e541c3d4e · outbound

This paper cites Lifelong machine learning systems: Beyond learning algorithms,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Lifelong machine learning systems: Beyond learning algorithms,

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.938416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.938416Z digest=sha256:b5e25807fb20979a7db16e9652cddf6e0ff96c6e11efc7c9aaa89ab71e6fc5f0

Observation b634438a-fd1c-4970-9325-a0c36b931914 · outbound

This paper cites Online learning and online convex optimization,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Online learning and online convex optimization,

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.947252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.947252Z digest=sha256:a597d933cee08b13c8617ade92fbbeaa60f713fc405226ded02965e9577e2cea

Observation a3890c45-8d6e-47ea-b98f-efa00d81004f · outbound

This paper cites Introduction to online convex optimization,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Introduction to online convex optimization,

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.955337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.955337Z digest=sha256:cc69cf78e14c461acdf92c2b1207e75f2f141238c00fb3e57276aad6cdde802d

Observation c74011aa-0093-46f0-b1c4-b3b7358044e2 · outbound

This paper cites Measuring catas- trophic forgetting in neural networks,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Measuring catas- trophic forgetting in neural networks,

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.963420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.963420Z digest=sha256:823435ea747f9eeb6f35449153593011faf765374dbe6b9552294bab4c5ac3aa

Observation af57c857-b637-430a-9a20-f7357cfd3fd8 · outbound

This paper cites Memory efficient experience replay for streaming learning,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Memory efficient experience replay for streaming learning,

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.971825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.971825Z digest=sha256:1f5a9672c08da614dc9d33413419547695ed21a3cc22c349598d3a218e6e0010

Observation 333e7cbe-c8be-4d13-8436-cdb184485355 · outbound

This paper cites Memory-Efficient Semi-Supervised Continual Learning: The World is its Own Replay Buffer.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Memory-Efficient Semi-Supervised Continual Learning: The World is its Own Replay Buffer

Reference 95

Resolution
verified exact
local_arxiv, observed 2026-08-15T21:16:02.803170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:15:59.979001Z digest=sha256:b8fcf970e9e4cb4e90715edff0d25c8acce4f51d22bb0886500478eedce759fa

Observation 0a5f28b3-0f63-4859-a0c3-c6671931149a · outbound

This paper cites Reinforcement learning with gaussian pro- cesses,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Reinforcement learning with gaussian pro- cesses,

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.987683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.987683Z digest=sha256:7f8bd7b3ac22689e81301d20703cea3b2ce85d9d40c0eec10342f5eac860136a

Observation acd0dfca-ade8-4287-975d-c5ddbb3c3399 · outbound

This paper cites Chevalier-Boisvert, L.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Chevalier-Boisvert, L

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.994086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.994086Z digest=sha256:cad998aa725d746d358dfa024b350f925171f97bd54ec76ff8cd4d807e5e2d27

Observation 399c1633-f32f-4532-a424-8730b4830a13 · outbound

This paper cites A multi-agent simulator for generating novelty in monopoly,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change A multi-agent simulator for generating novelty in monopoly,

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-15T21:16:00.004212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:16:00.004212Z digest=sha256:84407d71dd79726bbb1b94ccaba4be6a48f99da14f9e842505f8a1b4a76568a0

Observation f026a157-e36f-4327-9ee7-0c6efc7721b6 · outbound

This paper cites Novelty gen- eration framework for ai agents in angry birds style physics games,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Novelty gen- eration framework for ai agents in angry birds style physics games,

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-15T21:16:00.012629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:16:00.012629Z digest=sha256:35bd8abfa21ce0e1a7e3964c203360a2edebebfb6c32cdbd4ef9fb177499cc96

Observation 405165c0-6884-4373-b9d5-c509a3d9a614 · outbound

This paper cites Schmidhuber, A possibility for implementing curiosity and boredom in model- building neural controllers, 1991.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Schmidhuber, A possibility for implementing curiosity and boredom in model- building neural controllers, 1991

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-15T21:16:00.018724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:16:00.018724Z digest=sha256:d83c573aeda7a19f7e9d79529a379daaf3d694ca48d8ee6bec78ec03ce2dca29

Pith citing papers

No inbound Pith citation observations are available.