Pith. sign in

Paper Citation Record · LEDGER

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change

As of 17 August 2026, this Paper Citation Record lists 100 of 246 outbound references and 0 inbound Pith citation observations for arXiv:2505.10330.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.10330 v1

Coverage vector

measured 100 of 246 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:16:00.018724Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 246 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved99
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8bb213dc-9450-46f7-b3fd-2b6099c64b29 · outbound

This paper cites Mastering the game of go without human knowledge,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Mastering the game of go without human knowledge,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.069279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.069279Z digest=sha256:797a9727e30e8d1ef72735a49b73c04f65e8fe19521693bab415bbc9f3bc93c6

Observation f1d598b9-b60d-44cb-bcea-26b056cab1bf · outbound

This paper cites Mastering atari, go, chess and shogi by planning with a learned model,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Mastering atari, go, chess and shogi by planning with a learned model,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.077876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.077876Z digest=sha256:6ebb42b291d1715a8c594b17ba63200d7560b62ff9e808aa3efb4232c04c6166

Observation c49256e1-822a-4cd1-9957-b794514bb4ab · outbound

This paper cites Grandmaster level in starcraft ii using multi-agent reinforcement learning,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Grandmaster level in starcraft ii using multi-agent reinforcement learning,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.084850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.084850Z digest=sha256:1241cf5d141b826e34f242db8c04d810c8cd5cc8126f2b055c7a1d566a564840

Observation c5f77ab1-08bf-4a91-b0e4-b52429619df8 · outbound

This paper cites Dota 2 with Large Scale Deep Reinforcement Learning.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Dota 2 with Large Scale Deep Reinforcement Learning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.091559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.091559Z digest=sha256:0206c5f89df2b4799e9031f7859328de60bfab1e54fb56d7d342dbe7a029c9dc

Observation 8549a4f1-5811-4511-98c9-6885b84f8aee · outbound

This paper cites Agent57: Outperforming the human atari benchmark,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Agent57: Outperforming the human atari benchmark,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.099808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.099808Z digest=sha256:aa79751752c855e8a034b4107667b58a01b3f8d3545e2e30642e05d0b60166cb

Observation a9ad6ea3-71c7-494c-9814-82a1b315d9e0 · outbound

This paper cites Reinforcement learning based recommender systems: A survey,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Reinforcement learning based recommender systems: A survey,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.113283Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.113283Z digest=sha256:4124ce54bbf559070f020081024d94d38a9f9a438b407ee53716f119e0384101

Observation 8787a233-f3d0-4d8f-9132-a9af6cf52f1d · outbound

This paper cites Deepmind ai reduces google data centre cooling bill by 40%,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Deepmind ai reduces google data centre cooling bill by 40%,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.126194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.126194Z digest=sha256:17dde82ef1293c55900e576b0a1a330282be7b82684fd4ea38c8da2be7b2bd12

Observation 646eedef-cc8e-4f49-b633-505c07d9fab7 · outbound

This paper cites Gnu-rl: A precocial reinforcement learning so- lution for building hvac control using a differentiable mpc policy,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Gnu-rl: A precocial reinforcement learning so- lution for building hvac control using a differentiable mpc policy,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.133657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.133657Z digest=sha256:1432562c5730739855bb264a0dca895b36d1029aa5376a1d5ba2c0e6cfff3ac1

Observation b0a39276-513c-4fb5-9374-9fa80e2814a4 · outbound

This paper cites Magnetic control of tokamak plasmas through deep reinforce- ment learning,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Magnetic control of tokamak plasmas through deep reinforce- ment learning,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.145334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.145334Z digest=sha256:2385ee4de87118c0b27e9391cf0de3302a9673cd64c28273186f8586bcba2956

Observation 7edde538-4863-4b5a-83d5-8bfb448839a6 · outbound

This paper cites Adversarial policies beat superhuman go ais,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Adversarial policies beat superhuman go ais,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.155739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.155739Z digest=sha256:5af6b859a0a30eed50b912dc71bb47d75769b23d96a81244401c7404becbff0a

Observation af2d9cac-75a0-4ae1-b2b1-61604bdac3ad · outbound

This paper cites Adaptation in constant utility non-stationary environments.,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Adaptation in constant utility non-stationary environments.,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.161373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.161373Z digest=sha256:d9647220f3e3fd4f1a01b9a0ab45c0ad6aa23182b829e6912fd56b464cb6946a

Observation 074e49ff-b87b-447a-a19c-54a6f7d03619 · outbound

This paper cites an unresolved cited work.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.175120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.175120Z digest=sha256:3c46213cd94f15541dc7ff622ddb51bfebbc056244928ca3e6acd41e434d1817

Observation a2369159-0aec-464f-976d-cbb0314a3961 · outbound

This paper cites Impact of timing in post-warning prepositioning decisions on performance measures of disaster management: A 163 real-life application,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Impact of timing in post-warning prepositioning decisions on performance measures of disaster management: A 163 real-life application,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.181875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.181875Z digest=sha256:0b619038d47b08a9959b989007bca34b0aed0e4b53591a3afddebad35d2db7ac

Observation bee051dd-7a95-4a27-a96b-1dbfbd3ba6f1 · outbound

This paper cites Domain randomization for transferring deep neural networks from simulation to the real world,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Domain randomization for transferring deep neural networks from simulation to the real world,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.190382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.190382Z digest=sha256:9a1dcf12173d498972101f6e1e19dfaa8d6352df336d47ef3779144e5421f77c

Observation f6f5d207-465d-4eeb-ac17-ec1b171158f6 · outbound

This paper cites Learning optimal adap- tation strategies in unpredictable motor tasks,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Learning optimal adap- tation strategies in unpredictable motor tasks,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.198548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.198548Z digest=sha256:1c6ca89ea955af6a921abe0506b2be73041d40ff23d1003441cb7aa92b486be4

Observation f8612988-d5f0-4eef-9623-d2bd9a89fe36 · outbound

This paper cites Reward learning: Reinforcement, incentives, and expectations,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Reward learning: Reinforcement, incentives, and expectations,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.210439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.210439Z digest=sha256:7f0e48559f23767a6a9bafbe59c3245cbd6e224d6eb04ac0670ad7251754832f

Observation d62bc700-fd5f-499a-90a3-9ed6e906c8ec · outbound

This paper cites Towards precision holography.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Towards precision holography

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.221123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.221123Z digest=sha256:841c30311ad9b5bec05216b7600915afc77e427e5e2d0f88ef88f9985f201322

Observation c23d9383-47bb-428d-a8c1-3c1eb5e75a59 · outbound

This paper cites Deep neural networks for youtube rec- ommendations,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Deep neural networks for youtube rec- ommendations,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.227562Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.227562Z digest=sha256:8a853e1d279ab0bc76a459fda6219be935a110e31fce79ea8ccbf2a4d6f515a9

Observation 1accb39b-d7ba-4401-9899-aa4f85be4955 · outbound

This paper cites Scheduling on a budget: Avoiding stale recommendations with timely updates,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Scheduling on a budget: Avoiding stale recommendations with timely updates,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.236715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.236715Z digest=sha256:c6e61b8da1dd87c5b857ee43bb85a0c433433d31409948986962a121b138f827

Observation ab83c578-080a-4238-abbc-def669a488b2 · outbound

This paper cites Catastrophic interference in connectionist net- works: The sequential learning problem,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Catastrophic interference in connectionist net- works: The sequential learning problem,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.249385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.249385Z digest=sha256:77d5d029a29dcc0dd026e9814df176fbaccc33f03cf2c8667c59bb532ee69f68

Observation 3b8e47fc-3c03-43bf-86b5-4a19971b3148 · outbound

This paper cites Learning to predict by the methods of temporal differences,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Learning to predict by the methods of temporal differences,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.257424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.257424Z digest=sha256:96c2313cd0a681aba98992c6dec542cf174483240dec35872efea8f4e3bfb59c

Observation a86c677b-3431-4fd4-b80d-033953f4dd8e · outbound

This paper cites Analysis of temporal-diffference learning with func- tion approximation,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Analysis of temporal-diffference learning with func- tion approximation,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.267179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.267179Z digest=sha256:759a58b3e06dc659aa3a9a077715c375f1e36fc4a9db77376e0a5076933418cb

Observation 3c6ccfc8-6e62-4ea6-9593-08ac31f9e694 · outbound

This paper cites Human-level control through deep reinforcement learning,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Human-level control through deep reinforcement learning,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.277183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.277183Z digest=sha256:136870301d5bd5e3184431c1f9e33ab28619561722a0b1ead0cabc599ab64209

Observation 0c6c19b7-8dff-4885-a1c6-a42f99cfcad4 · outbound

This paper cites Deep Reinforcement Learning and the Deadly Triad.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Deep Reinforcement Learning and the Deadly Triad

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.289096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.289096Z digest=sha256:a5c2d0c197a502733505f2f7eaa28c6489429b6a53230611b234265ca531ff3d

Observation f7e9ba2d-1b8c-498f-95b0-416f42011241 · outbound

This paper cites Simple statistical gradient-following algorithms for connectionist reinforcement learning,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Simple statistical gradient-following algorithms for connectionist reinforcement learning,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.299736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.299736Z digest=sha256:f6f9b0dbc7018e14a27a0019ba97b3371772ddc66ab722c17ba766a13be9459b

Observation 35b57e3e-b9be-41c7-ba46-c22dc1ce44b0 · outbound

This paper cites Asynchronous methods for deep reinforcement learning,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Asynchronous methods for deep reinforcement learning,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.309292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.309292Z digest=sha256:8824816afd1acf6b5ed53638cb6069bef05f5c88abc8b655f778d257cc19c58f

Observation a41617e8-6cbd-47b7-b593-472866fd3b01 · outbound

This paper cites Impala: Scalable distributed deep-rl with importance weighted actor-learner architectures,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Impala: Scalable distributed deep-rl with importance weighted actor-learner architectures,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.316184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.316184Z digest=sha256:e10499cbbc5d763e8ae4f4b501ce18d28ac98141bca246f729d4a1d1125b96f9

Observation 16062bad-540c-40c8-8a77-2a1b55b195bc · outbound

This paper cites Continuous control with deep reinforcement learning,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Continuous control with deep reinforcement learning,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.324781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.324781Z digest=sha256:c5ebcd2bde9d1238467a675c6ec8a0128457d0bac9a6a22418c22ef39df0aadc

Observation 945f7532-0f0b-4353-ae67-1ee4d13088d4 · outbound

This paper cites Distributed distributional deterministic policy gradients,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Distributed distributional deterministic policy gradients,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.335060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.335060Z digest=sha256:a8b7f728763f5a7df0c4e8a6a9227381c6ad22befde9ffecf86ea6509c79794a

Observation 62c1d1da-3bcd-4d85-8c46-9edcb8be888e · outbound

This paper cites Scalable trust-region method for deep reinforcement learning using kronecker-factored approximation,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Scalable trust-region method for deep reinforcement learning using kronecker-factored approximation,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.346698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.346698Z digest=sha256:4cce2a6122a169bd2d7dc063081cd61c7b85885f0decb4ba97c46cbad19b074e

Observation 02d67d68-b000-47bb-a1dc-d36558a3c55f · outbound

This paper cites Soft actor-critic: Off-policy maxi- mum entropy deep reinforcement learning with a stochastic actor,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Soft actor-critic: Off-policy maxi- mum entropy deep reinforcement learning with a stochastic actor,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.354303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.354303Z digest=sha256:35bfb1e33718e0246ad65ceb63e23afc2b3ef0322ad9928d3661c72695a934f5

Observation 58588fe6-dc46-4e88-aba3-8b8681a924ef · outbound

This paper cites Trust region pol- icy optimization,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Trust region pol- icy optimization,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.369754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.369754Z digest=sha256:95cb0b5cbfda4fed56178ab218cd96d5529ac5b3202faca04d94c5790051cdb1

Observation 88e42701-3dea-44a9-99d1-2e07398935cd · outbound

This paper cites Proximal Policy Optimization Algorithms.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Proximal Policy Optimization Algorithms

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.377816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.377816Z digest=sha256:4cafb3416615155545adf05e8a45324ab0717ceb194583d0f587af9c8703e22a

Observation dd4d2e19-2898-4e64-bdbc-e559d47f153a · outbound

This paper cites Dyna, an integrated architecture for learning, planning, and reacting,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Dyna, an integrated architecture for learning, planning, and reacting,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.384898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.384898Z digest=sha256:79fd8cda3c2d35df0c306efcbf38de4858ad22b3827f05f546ee8067ac387e7a

Observation 6bb3b444-52e1-4069-b7d3-3ef783a3f54f · outbound

This paper cites First return, then explore,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change First return, then explore,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.395462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.395462Z digest=sha256:e456656984730270153e01d675201babd27c29bf9d44473791dec20cd48fb48f

Observation 4a0ed717-3cb0-47e0-9674-b2ce1d30610f · outbound

This paper cites Learning latent dynamics for planning from pixels,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Learning latent dynamics for planning from pixels,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.406216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.406216Z digest=sha256:93fb9ec37d06c33e24d094ca031497cf3f22940c4e9f75b6c4999074fcee7368

Observation 6ec25d40-c564-4cfc-9958-99194229e867 · outbound

This paper cites Curious model-building control systems,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Curious model-building control systems,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.418445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.418445Z digest=sha256:ca6903ef29ceec727ad6d86f15402d7b60057e1f62eb853061c891d57187f067

Observation f56d2ea2-50f9-4060-9e51-c18d150b418c · outbound

This paper cites On Learning to Think: Algorithmic Information Theory for Novel Combinations of Reinforcement Learning Controllers and Recurrent Neural World Models.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change On Learning to Think: Algorithmic Information Theory for Novel Combinations of Reinforcement Learning Controllers and Recurrent Neural World Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.428205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.428205Z digest=sha256:f649edac3befa335d6bc73fed4db9a76f6f464811dc9ea23a050c30a70f8ae40

Observation 18c87a44-82aa-4a6c-a073-bbf03641872c · outbound

This paper cites Recurrent world models facilitate policy evolution,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Recurrent world models facilitate policy evolution,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.437614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.437614Z digest=sha256:f7cb650142e0397636ced26bb118a0892796d3d5a1ebf8b9fc44b61f4f7dfaec

Observation ed22c21e-7630-4db0-bb89-f39ec53129d3 · outbound

This paper cites Dream to control: Learning behav- iors by latent imagination,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Dream to control: Learning behav- iors by latent imagination,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.448190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.448190Z digest=sha256:1bc37f279195245afb8b0ab67dc405683405c740051167f98cca3e854a207dbc

Observation 7fc91ebd-9bd0-4bde-b830-847e45b5a062 · outbound

This paper cites Mastering atari with discrete world models,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Mastering atari with discrete world models,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.464751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.464751Z digest=sha256:00eb28b84cbe4232ea94627184c01f09fa5344f7369a66e851c6d407779bdf96

Observation f04e09b9-1be3-458e-aa98-d378d4291cd3 · outbound

This paper cites Mastering Diverse Domains through World Models.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Mastering Diverse Domains through World Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.472534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.472534Z digest=sha256:3a93e8e3e4d7b736388396adc7b4002d165d1af4897b31b55a9b571bbf0fd8e5

Observation 5b345e16-a794-4728-a06a-8084bf0f0700 · outbound

This paper cites On the Properties of Neural Machine Translation: Encoder-Decoder Approaches.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change On the Properties of Neural Machine Translation: Encoder-Decoder Approaches

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.485835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.485835Z digest=sha256:fe3f042b93bd2624df8032cb47cb7a7b7ec6a03ccd5f52d9f8d8ef5b007b6926

Observation 67b5deec-6cfc-4b56-b67c-f299ef73d8bd · outbound

This paper cites Convolutional networks for images, speech, and time series,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Convolutional networks for images, speech, and time series,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.492201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.492201Z digest=sha256:4b26e66591ed4bf393636f4d6667442146561e1635e8ad3d729df23a0f628088

Observation c7d134be-e586-4825-9cd3-1e8f6610bc10 · outbound

This paper cites Estimating or Propagating Gradients Through Stochastic Neurons for Conditional Computation.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Estimating or Propagating Gradients Through Stochastic Neurons for Conditional Computation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.498195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.498195Z digest=sha256:e1c886436ce63ab0bad283d26ccaba28d394c83c7f6c5da116d37b85bdd286a1

Observation 6f2691ef-5777-46ef-b275-9dfce39284dc · outbound

This paper cites Dm control: Software and tasks for continuous con- trol,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Dm control: Software and tasks for continuous con- trol,

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.507948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.507948Z digest=sha256:d4e0f4d3ea7254aa0ff9ce767deaf6cf6ba2a075992cd674574642ffa4e632f7

Observation 0e902dc0-ae37-426c-976e-cd0e7fdcf3c7 · outbound

This paper cites The arcade learning environment: An evaluation platform for general agents,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change The arcade learning environment: An evaluation platform for general agents,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.515230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.515230Z digest=sha256:e281c7a2a119d49e90561c01f11aa35db779317c93ca6a7bbfce968d57a727d4

Observation 373b9b0a-ee95-4973-9969-b4f351d66902 · outbound

This paper cites Policy invariance under reward transforma- tions: Theory and application to reward shaping,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Policy invariance under reward transforma- tions: Theory and application to reward shaping,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.521131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.521131Z digest=sha256:a736af704e740552da248fe74ab3f7e2b8a7b8804f4f4b436f0ae8d027b60a48

Observation d46664b3-7be0-4861-9f84-5e1cc92ec8ce · outbound

This paper cites Self-improving reactive agents based on reinforcement learning, plan- ning and teaching,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Self-improving reactive agents based on reinforcement learning, plan- ning and teaching,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.530020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.530020Z digest=sha256:e34a96370157993e212987a6def5c8d26af3a50ce66ae783b8ac1f3fb3d2fa82

Observation badbb7e8-9f76-4d1f-946d-d40e6b9ec770 · outbound

This paper cites Sample efficient actor-critic with experience replay,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Sample efficient actor-critic with experience replay,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.537162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.537162Z digest=sha256:6f76871ab694a2f5df212c143c079da433c4adcbf618c60f304fa0d777e33282

Observation 63dc1484-2984-4929-8633-a13d189143e0 · outbound

This paper cites Rainbow: Combining improvements in deep reinforcement learn- ing,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Rainbow: Combining improvements in deep reinforcement learn- ing,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.545797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.545797Z digest=sha256:d1006eb9b227ec0a920397faef7ecd05ce5d27fbb080e002984ba1c8f17e4aaf

Observation 2c3f1207-1be5-45d3-911d-b2b49a2c82c4 · outbound

This paper cites A Deeper Look at Experience Replay.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change A Deeper Look at Experience Replay

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.552889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.552889Z digest=sha256:99fa9a6d15e90b868d94700aac315778dcd4eece377aa1d0e180b5b44018e5fe

Observation a84b1911-47e9-4e67-8861-60e084505eb1 · outbound

This paper cites Revisiting fundamentals of experience replay,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Revisiting fundamentals of experience replay,

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.560650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.560650Z digest=sha256:c99d6de62bde6110825a15b389e8503f9a2759b8660ab01de96b1f74cfe8a643

Observation 2be1cbd3-8fe3-4fd6-b634-012b09fd6ff4 · outbound

This paper cites Prioritized experience replay,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Prioritized experience replay,

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.571827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.571827Z digest=sha256:d571787f74c03677b72363ffbc8a99ef601cf8c2963cb0172359920c9ebf9503

Observation 887dd01e-658c-4e6b-b6a4-58075e686d7d · outbound

This paper cites Prioritized experience replay method based on experience reward,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Prioritized experience replay method based on experience reward,

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.579543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.579543Z digest=sha256:574a526aeec6e0dcb29ecfcc7ab407b2fd56a41930738d5c842a09784fd514f6

Observation bf2d5765-a429-4875-9021-55001a1a28f7 · outbound

This paper cites Model-augmented prioritized experience replay,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Model-augmented prioritized experience replay,

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.587135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.587135Z digest=sha256:d8b812ca1e42957504d6d2384a35a1dc6f3b50eaaf48265fb06c1a9c535d6982

Observation 21827c24-b782-4c12-b54a-9cdff7bc22da · outbound

This paper cites Prioritized experience replay based on dynamics priority,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Prioritized experience replay based on dynamics priority,

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.593754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.593754Z digest=sha256:38bcb54af6fb7c04b85538a22e064e379d2fa54e0e209a1e238c8fc8c7317109

Observation 28880909-9ef5-4f2b-95e2-716081aa983e · outbound

This paper cites Image augmentation is all you need: Reg- ularizing deep reinforcement learning from pixels,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Image augmentation is all you need: Reg- ularizing deep reinforcement learning from pixels,

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.600732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.600732Z digest=sha256:61a27f40b880f13db20b077fa790c9ce871c9340b611bbd3b7bf91f505096529

Observation bb3261a3-2354-4fab-9fa7-3b8d165774e3 · outbound

This paper cites Temporal difference learning for model pre- dictive control,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Temporal difference learning for model pre- dictive control,

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.608644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.608644Z digest=sha256:26fa2ec807fe002d48c8e4613d962c965665d1ba3eac96878ee50253e0c1b28a

Observation 951eaa83-08f4-44f7-875d-a76e668421e0 · outbound

This paper cites Transfer Learning in Deep Reinforcement Learning: A Survey.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Transfer Learning in Deep Reinforcement Learning: A Survey

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.617648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.617648Z digest=sha256:add12427512bbdddcf2c0953071ecb611457d3617b8e5ce08041c272ebc0b7d5

Observation ac750a00-b921-4bb4-b3a7-bbb561a2299d · outbound

This paper cites Distilling the Knowledge in a Neural Network.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Distilling the Knowledge in a Neural Network

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.629259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.629259Z digest=sha256:58f5af213325945b823dc5c29881f756ad35999a82cc0deb688ecf0590cb9aec

Observation bbfd7cca-92d8-457e-8cdd-354b96510e35 · outbound

This paper cites Knowledge distillation: A survey,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Knowledge distillation: A survey,

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.642310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.642310Z digest=sha256:fb1c2471d9009a897618c57a1643ee94a850f654a22b0f9d11b6508eaa535168

Observation f5b1d7c9-131f-4cb5-a799-62ab7283ce40 · outbound

This paper cites Teaching on a budget: Agents advising agents in rein- forcement learning,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Teaching on a budget: Agents advising agents in rein- forcement learning,

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.650908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.650908Z digest=sha256:0c937f36dddc1fb70256c49d02c709f3d941800407a6e540c64a92cdc0974dc8

Observation 65a7e75e-8a59-4218-8d61-f9c10140c608 · outbound

This paper cites Online transfer learning in reinforcement learning do- mains,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Online transfer learning in reinforcement learning do- mains,

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.661472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.661472Z digest=sha256:db79b810764a22d4be81a82c1a280311f93fef93ddc690d31371a744d44de900

Observation 3fec6310-43f4-4baf-8400-abe9f017ca53 · outbound

This paper cites A survey on transfer learning,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change A survey on transfer learning,

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.673376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.673376Z digest=sha256:159529ce90e03b8d4e11439405ab35127d56ff0260ea06b5880e99bb7a173da3

Observation f6b7be33-429f-4471-aff3-7e626e0f478c · outbound

This paper cites Transfer learning for reinforcement learning domains: A survey.,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Transfer learning for reinforcement learning domains: A survey.,

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.682328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.682328Z digest=sha256:f9a99edfe3b10f18853637d895bca58b75edb4e76b41e70f12867148a8faa693

Observation e75e4662-eb63-4180-9447-11933bcccff1 · outbound

This paper cites A Comprehensive Survey on Test-Time Adaptation under Distribution Shifts.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change A Comprehensive Survey on Test-Time Adaptation under Distribution Shifts

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.692146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.692146Z digest=sha256:ad1623543f4c95e037458cb119b0f9a9806616401dc6f61cc41d3ebd9b054c17

Observation 7d2f22a2-1761-4126-a62a-7f75083599cc · outbound

This paper cites A review of novelty detection,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change A review of novelty detection,

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.701112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.701112Z digest=sha256:a20aeaded2e610da300303085145dcb016dd63e7c689d0ac84577f619f0c6b01

Observation 558d6a72-f865-4a75-8344-4d1000c6f88c · outbound

This paper cites Towards a unifying framework for formal theories of novelty,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Towards a unifying framework for formal theories of novelty,

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.709383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.709383Z digest=sha256:6c22215938e6bc0df2f01d4aa352b0768efc35da4c918a9db2d9f2433808a96d

Observation d4e83569-863d-413b-aa9a-e99ca2d53c82 · outbound

This paper cites Open-world learning for radically autonomous agents,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Open-world learning for radically autonomous agents,

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.717946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.717946Z digest=sha256:363eecf151b3d6c4b65048eba407e9a6ed9f5a861f647c86a0da9cc6b0f73b8b

Observation 13fd4986-b491-4ed7-a7f1-c23a1214b0f3 · outbound

This paper cites Mixtbn: A fully test-time adaptation method for visual reinforce- ment learning on robotic manipulation,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Mixtbn: A fully test-time adaptation method for visual reinforce- ment learning on robotic manipulation,

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.730032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.730032Z digest=sha256:411f740454c20e350264c14de86b3b0de5be7128b14d55fa313e3f33fdeffb42

Observation 506e06de-0f6c-4e4c-967b-c2cfd0644945 · outbound

This paper cites Active test-time adaptation: Theoretical analyses and an algorithm,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Active test-time adaptation: Theoretical analyses and an algorithm,

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.745099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.745099Z digest=sha256:08f159622c4d0760768733a34afbd30cc1d249b2037532606ed6c94f92b8af0d

Observation d69bb3f9-5363-4488-9c46-4db95f5d3cff · outbound

This paper cites Unknown sample discovery for source free open set domain adaptation,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Unknown sample discovery for source free open set domain adaptation,

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.754014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.754014Z digest=sha256:bab5d817716f7c12caa3a72c39ea40f785630e42ffe74759123b4fb2e491ba44

Observation 86a5621a-af8d-4542-bd5d-eff244394314 · outbound

This paper cites Hidden-mode markov decision processes for nonstationary sequential decision making,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Hidden-mode markov decision processes for nonstationary sequential decision making,

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.775138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.775138Z digest=sha256:9f75cfcde2ff3c5819712bda44d15f5d45733c8b27f354441a331efd470a662f

Observation 55f29354-0515-4d56-89dd-d7eb3a83b818 · outbound

This paper cites Choosing search heuristics by non-stationary reinforcement learn- ing,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Choosing search heuristics by non-stationary reinforcement learn- ing,

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.786510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.786510Z digest=sha256:af2989159d4a3b1b107825bffe709c34de88340bfe20af63c51e9119037d78ab

Observation 15c88ad6-90e1-4269-8d00-b9cf1aa2b763 · outbound

This paper cites Non-stationary reinforcement learning without prior knowledge: An optimal black-box approach,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Non-stationary reinforcement learning without prior knowledge: An optimal black-box approach,

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.799270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.799270Z digest=sha256:2026d9ccba52d7b7ff77e7f0cdf8cc7800d4884424b4d77399db4d432f02a187

Observation 6474dc5c-ee3b-41bb-88f1-bbf9ddcfd52e · outbound

This paper cites Near-optimal model- free reinforcement learning in non-stationary episodic mdps,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Near-optimal model- free reinforcement learning in non-stationary episodic mdps,

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.811735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.811735Z digest=sha256:6c67d6a47f907d04f69b2e9c191ec1f559314c56e2fbce62355d94bddf1afaee

Observation 68f75129-2aab-4295-b69a-6b5d0005fe42 · outbound

This paper cites Non-stationary reinforcement learning under general function approximation,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Non-stationary reinforcement learning under general function approximation,

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.819317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.819317Z digest=sha256:3c1adf6044c5014114b28eca4a2093cea4317af0181077a0481edccc091cd610

Observation f5ad77ed-1116-4b98-8967-2c6100a10b5d · outbound

This paper cites Addressing environment non-stationarity by repeat- ing q-learning updates,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Addressing environment non-stationarity by repeat- ing q-learning updates,

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.826698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.826698Z digest=sha256:f47329b0a03191adb7e03add8503316e3a051864192066a218e2249125a03677

Observation 4fd715e3-9bdb-4ef1-b845-316ab26f6ef8 · outbound

This paper cites Non-stationary markov decision processes, a worst-case approach using model-based reinforcement learning,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Non-stationary markov decision processes, a worst-case approach using model-based reinforcement learning,

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.839173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.839173Z digest=sha256:9fe30daeb3ff29fffc98014fcf5a678bf9018bf52d23b853c05bc2f597bc5692

Observation 9de3e21d-e44e-4e85-a7c4-4eb6ac2ec82a · outbound

This paper cites Reinforcement learning algorithm for non-stationary environments,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Reinforcement learning algorithm for non-stationary environments,

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.847130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.847130Z digest=sha256:03568e2bd9d706d1e263d3207a5ae8030d2fc5cc966acf7f6988ae3a8d86816f

Observation 575e6f73-7b01-44d9-a2c4-484e39a41db0 · outbound

This paper cites Reactive exploration to cope with non-stationarity in life- long reinforcement learning,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Reactive exploration to cope with non-stationarity in life- long reinforcement learning,

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.858236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.858236Z digest=sha256:dde37a52fae4a055ca2c25122c8f81f50da22b15a09a7dd2606225f6f9cb16c1

Observation 3644160b-304e-427c-b30b-62cfa6749b00 · outbound

This paper cites Transfer in reinforcement learning: A framework and a survey,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Transfer in reinforcement learning: A framework and a survey,

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.872781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.872781Z digest=sha256:164bc05dbb71a766033cfe79f2f9cdc8eedc2193537abb65d70806dbfa019af6

Observation 9389818f-b3cd-40fd-98a9-52b142a18d76 · outbound

This paper cites Cross-modal domain adaptation for cost-efficient visual reinforcement learning,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Cross-modal domain adaptation for cost-efficient visual reinforcement learning,

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.881669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.881669Z digest=sha256:c37dee6e76e2cfae6be61bea4ed2b8c0be2a5a71f235cc55b80e202d8b36739c

Observation 7879d77e-0e4d-4cb7-9b88-f95b63e0c972 · outbound

This paper cites Deep reinforcement learning amidst lifelong non- stationarity,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Deep reinforcement learning amidst lifelong non- stationarity,

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.889965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.889965Z digest=sha256:1ce8875fec69d834db931f4d117c559dde7c6e9ceb0861f7a225460b255c23c4

Observation 56d56f0c-6e62-4b3d-b297-60328daf7490 · outbound

This paper cites Model-based nov- elty adaptation for open-world ai,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Model-based nov- elty adaptation for open-world ai,

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.900392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.900392Z digest=sha256:66d8f059e49fdd2d1b76ab6ef0829eff9348f88cefd5d41db8b863bf68a9bf8b

Observation 139fd434-fe0d-47ba-a6de-b7b3fabcfcf2 · outbound

This paper cites Detecting and adapting to novelty in games,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Detecting and adapting to novelty in games,

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.906124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.906124Z digest=sha256:30a1d411c32b459904ad1e3aba4d4485d408cad6c8eef000e955d527dc0e2f78

Observation 05e3ef52-8c99-4259-a8d1-4afd3b6bea1f · outbound

This paper cites Spotter: Extending symbolic planning operators through targeted reinforcement learning,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Spotter: Extending symbolic planning operators through targeted reinforcement learning,

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.915271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.915271Z digest=sha256:f6ad5bab1cff218374982014fba18f3c0b7011746e2f5f6805c74de8a88d66b6

Observation d3c11078-a8b5-47ed-a070-7c91546c2ea9 · outbound

This paper cites An integrated architecture for online adaptation to novelty in open worlds using probabilistic programming and novelty-aware planning,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change An integrated architecture for online adaptation to novelty in open worlds using probabilistic programming and novelty-aware planning,

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.929097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.929097Z digest=sha256:af20d917dbb107a4e735d62ed0de52ffaa21c52c39167dab09d75a5f5ae214a6

Observation 03e8f72d-2128-4d16-828f-768e541c3d4e · outbound

This paper cites Lifelong machine learning systems: Beyond learning algorithms,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Lifelong machine learning systems: Beyond learning algorithms,

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.938416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.938416Z digest=sha256:6a0b5e860ccf9164bcadcbe7d6afe4efa1daf4d337ccf3555421958042565cb2

Observation b634438a-fd1c-4970-9325-a0c36b931914 · outbound

This paper cites Online learning and online convex optimization,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Online learning and online convex optimization,

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.947252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.947252Z digest=sha256:5f56ba1e86520735caa78d4be3cacd85c48af8cd918c99ade00d7bed57c3dcaf

Observation a3890c45-8d6e-47ea-b98f-efa00d81004f · outbound

This paper cites Introduction to online convex optimization,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Introduction to online convex optimization,

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.955337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.955337Z digest=sha256:a7e7ade62934286269e77a0797da6125b388878f0a95ea25a35b01c309f36a5f

Observation c74011aa-0093-46f0-b1c4-b3b7358044e2 · outbound

This paper cites Measuring catas- trophic forgetting in neural networks,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Measuring catas- trophic forgetting in neural networks,

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.963420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.963420Z digest=sha256:9259bb481355419c52d6e8ef93d56f8337976375627b5b4f607893625503d3fe

Observation af57c857-b637-430a-9a20-f7357cfd3fd8 · outbound

This paper cites Memory efficient experience replay for streaming learning,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Memory efficient experience replay for streaming learning,

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.971825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.971825Z digest=sha256:f9d7b03a1f344c8c81feb6825a18939663f9f1c98bbc53409456e806481a1d29

Observation 333e7cbe-c8be-4d13-8436-cdb184485355 · outbound

This paper cites Memory-Efficient Semi-Supervised Continual Learning: The World is its Own Replay Buffer.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Memory-Efficient Semi-Supervised Continual Learning: The World is its Own Replay Buffer

Reference 95

Resolution
verified exact
local_arxiv, observed 2026-08-15T21:16:02.803170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T21:15:59.979001Z digest=sha256:a124425314f74e703d8cc9c69a6679f1515d64d0c2b69de112ef01e9838a4943

Observation 0a5f28b3-0f63-4859-a0c3-c6671931149a · outbound

This paper cites Reinforcement learning with gaussian pro- cesses,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Reinforcement learning with gaussian pro- cesses,

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.987683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.987683Z digest=sha256:90f246d3840a93433e6c117e76852fb56144249ac0d9dc92b345bd45268fa3f9

Observation acd0dfca-ade8-4287-975d-c5ddbb3c3399 · outbound

This paper cites Chevalier-Boisvert, L.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Chevalier-Boisvert, L

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-15T21:15:59.994086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:15:59.994086Z digest=sha256:5c6fe5e01baea211e4e7ad07d9017180ee595faf293d64e68a994133e5ee72fd

Observation 399c1633-f32f-4532-a424-8730b4830a13 · outbound

This paper cites A multi-agent simulator for generating novelty in monopoly,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change A multi-agent simulator for generating novelty in monopoly,

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-15T21:16:00.004212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:16:00.004212Z digest=sha256:f6cb51f1d8c0ce98d1ab3ba083e5eb03b7acc469381dc5960b58069c88d14299

Observation f026a157-e36f-4327-9ee7-0c6efc7721b6 · outbound

This paper cites Novelty gen- eration framework for ai agents in angry birds style physics games,.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Novelty gen- eration framework for ai agents in angry birds style physics games,

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-15T21:16:00.012629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:16:00.012629Z digest=sha256:86646f6bbd112137258497e336f33ee652653eb909fc55c47f861d1ac1b965ae

Observation 405165c0-6884-4373-b9d5-c509a3d9a614 · outbound

This paper cites Schmidhuber, A possibility for implementing curiosity and boredom in model- building neural controllers, 1991.

Efficient Adaptation of Reinforcement Learning Agents to Sudden Environmental Change Schmidhuber, A possibility for implementing curiosity and boredom in model- building neural controllers, 1991

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-15T21:16:00.018724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:16:00.018724Z digest=sha256:5fe05427caca07b330e5c74b60af2baf4a12d058cb4e816a3f2e4df71e44a21e

Pith citing papers

No inbound Pith citation observations are available.