Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T19:07:25.450652Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 69 of 69 outbound references and 1 inbound Pith citation observation for arXiv:2502.10200.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T19:07:25.450652Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-19T06:36:56.956656Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
69 of 69 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation 1ec14da0-79da-4785-90b6-29f5601b08a4 · outbound
Dynamic Reinforcement Learning for Actors Can I say, now machines can think?
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 28349bd4-8bc5-4941-8b46-6ebec2dfccc2 · outbound
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c0abd69-1693-432d-b8c0-603607caf7f3 · outbound
Dynamic Reinforcement Learning for Actors Any Target Function Exists in a Neighborhood of Any Sufficiently Wide Random Network: A Geometrical Perspective
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1c7bb315-e3da-4197-bf53-2ca238a407ad · outbound
Dynamic Reinforcement Learning for Actors Andrychowicz, F
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 259cb42c-a97a-4417-a27d-fcfc03b285b6 · outbound
Dynamic Reinforcement Learning for Actors Azizi and G
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 377bd31a-6354-49d7-811a-6b6bbe73a8d4 · outbound
Dynamic Reinforcement Learning for Actors Berlyne and W
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c853da98-0f1a-4364-b8fe-df3bc71d3fb0 · outbound
Dynamic Reinforcement Learning for Actors Active Divergence with Generative Deep Learning -- A Survey and Taxonomy
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c484f11f-3c07-4f53-bd01-8ebdf8385ffb · outbound
Dynamic Reinforcement Learning for Actors Statement on AI risk, 2025 a
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ead8142e-04d5-40a2-b5b3-3867c44848f2 · outbound
Dynamic Reinforcement Learning for Actors An overview of catastrophic AI risks, 2025 b
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c52952a1-8c40-4e3e-a4ef-03244c2af132 · outbound
Dynamic Reinforcement Learning for Actors Art or Artifice? Large Language Models and the False Promise of Creativity
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26a6625f-4522-48ef-b3ec-5a134256a63d · outbound
Dynamic Reinforcement Learning for Actors The alternative uses test, 2018
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4ddcb73b-9a6c-4eb9-8b9a-b7c24c5ae6b5 · outbound
Dynamic Reinforcement Learning for Actors Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9aa920ac-df9a-4ec0-a580-31af5c6ce784 · outbound
Dynamic Reinforcement Learning for Actors Unresolved cited work
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e6541dd2-6962-439b-b162-da5c76e9c24d · outbound
Dynamic Reinforcement Learning for Actors Creative Beam Search: LLM-as-a-Judge For Improving Response Generation
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c6b79a1-bce9-445a-b509-10bc37ffacc5 · outbound
Dynamic Reinforcement Learning for Actors Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 04004886-f679-42fd-b756-62e850f00150 · outbound
Dynamic Reinforcement Learning for Actors Fujimoto, H
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 95473fc4-335b-4e89-9cc1-bfb984da374f · outbound
Dynamic Reinforcement Learning for Actors Research priorities for robust and beneficial artificial intelligence: An open letter, 2015
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d300d54c-f16b-44d4-9b4f-e2e93dc2e476 · outbound
Dynamic Reinforcement Learning for Actors Large Language Models Are Not Strong Abstract Reasoners
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a2574c0-1756-4d9d-ae57-7a3cc3da2e9d · outbound
Dynamic Reinforcement Learning for Actors Goto and K
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ccef354b-ffe4-41ca-990e-653835134b3e · outbound
Dynamic Reinforcement Learning for Actors Unresolved cited work
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1bbccc8e-a87c-4b59-bdf6-0ef6844955aa · outbound
Dynamic Reinforcement Learning for Actors Unresolved cited work
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac096356-3161-4daf-8dfa-c9693e4791c9 · outbound
Dynamic Reinforcement Learning for Actors Haarnoja, A
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7dbe6518-1db2-4038-85dc-f67eeb01d6bc · outbound
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0912888d-f5fa-42a4-a746-fd74fd9a6fba · outbound
Dynamic Reinforcement Learning for Actors Creativity in AI: Progresses and Challenges
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 438312f1-2ed0-43a6-87ba-195065a4c6a8 · outbound
Dynamic Reinforcement Learning for Actors BRAINTEASER: Lateral Thinking Puzzles for Large Language Models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47709292-6a33-488f-a45b-f0f35bd9f59c · outbound
Dynamic Reinforcement Learning for Actors Creativity
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 47f466da-ef4c-4232-baa8-1c231d480aff · outbound
Dynamic Reinforcement Learning for Actors Khachaturyan, S
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b5361bcc-dd35-4e15-ab15-ab4093a9582c · outbound
Dynamic Reinforcement Learning for Actors Koivisto and S
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4f34e5fa-38a8-4f73-aece-c427366422b1 · outbound
Dynamic Reinforcement Learning for Actors Unresolved cited work
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e94891ad-91c4-41bc-9e3d-34f99356c730 · outbound
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation baa3fa8b-e7ae-4a56-b9f9-4bfb0d0b1347 · outbound
Dynamic Reinforcement Learning for Actors AI as Humanity's Salieri: Quantifying Linguistic Creativity of Language Models via Systematic Attribution of Machine Text against Web Text
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64867c5a-43a9-479a-8ab3-647acc2032cd · outbound
Dynamic Reinforcement Learning for Actors Matsuki and K
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation dcc27c0c-fdb8-4a56-bff2-43d80cff4c55 · outbound
Dynamic Reinforcement Learning for Actors Matsuki, Y
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d62fe747-6738-4245-bdb7-464459f28e64 · outbound
Dynamic Reinforcement Learning for Actors McCulloch and W
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a623cb31-b245-4796-abf5-cd563d500e77 · outbound
Dynamic Reinforcement Learning for Actors Comparing Humans, GPT-4, and GPT-4V On Abstraction and Reasoning Tasks
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b7c4dc0-c179-4bb3-b4ed-4e10f2f5aba2 · outbound
Dynamic Reinforcement Learning for Actors Unresolved cited work
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4c17a51-bb00-45b6-ad81-8d7f197e4ec4 · outbound
Dynamic Reinforcement Learning for Actors Asynchronous Methods for Deep Reinforcement Learning
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d1bdc33-c100-4c6d-9621-38acb2cb02f0 · outbound
Dynamic Reinforcement Learning for Actors Characterising the Creative Process in Humans and Large Language Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e3cc862-1886-4615-8d71-feb6f9f05d46 · outbound
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ba9a4523-4bae-4a4c-995c-0f04e6588a3c · outbound
Dynamic Reinforcement Learning for Actors OpenAI Five , 2019
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7d817cc5-7d64-4846-acb4-c42cad585af0 · outbound
Dynamic Reinforcement Learning for Actors GPT-4 , 2023
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 45d1ec93-e47c-4a30-bd16-4aa64f3ad1b2 · outbound
Dynamic Reinforcement Learning for Actors GPT-4 Technical Report
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6f3f176-c7af-4d0e-9bfd-c363115b7966 · outbound
Dynamic Reinforcement Learning for Actors Is Temperature the Creativity Parameter of Large Language Models?
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd14ff61-dea1-47f7-8e25-9b68f6f0f1df · outbound
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bbf03383-89cf-482e-a3d6-80f36187af94 · outbound
Dynamic Reinforcement Learning for Actors Unresolved cited work
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f5e8aa09-3ccf-4ba8-b3ae-a5403abb21f3 · outbound
Dynamic Reinforcement Learning for Actors Unresolved cited work
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b260205c-741e-4f7d-b7e2-88b0f7dcb6c9 · outbound
Dynamic Reinforcement Learning for Actors Sawatsubashi, M
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0b462b86-7a7d-4100-9c3a-a52baeb23fc6 · outbound
Dynamic Reinforcement Learning for Actors Prioritized Experience Replay
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e8545b7-1c80-47e6-b156-7e3e5479e483 · outbound
Dynamic Reinforcement Learning for Actors Schrittwieser, I
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f6d76df-97fd-4a7e-9849-4024a80b709c · outbound
Dynamic Reinforcement Learning for Actors Schulman, F
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4968dc46-e457-406f-a229-667915bec91f · outbound
Dynamic Reinforcement Learning for Actors Unresolved cited work
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 52e13e57-b625-484f-8339-6315628f90fc · outbound
Dynamic Reinforcement Learning for Actors Unresolved cited work
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 54b98b14-b07b-4a97-9015-c9bd09a26fa0 · outbound
Dynamic Reinforcement Learning for Actors Communications that Emerge through Reinforcement Learning Using a (Recurrent) Neural Network
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 33424638-049a-4dd4-87a1-231258f290b2 · outbound
Dynamic Reinforcement Learning for Actors Functions that Emerge through End-to-End Reinforcement Learning - The Direction for Artificial General Intelligence -
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6f2d4dfe-5aba-4b7e-a397-15e3d7adb2ef · outbound
Dynamic Reinforcement Learning for Actors Shibata and K
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 44d3b80a-8a8e-4fdd-b614-855afaf68761 · outbound
Dynamic Reinforcement Learning for Actors Shibata and Y
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5ec49e9a-0d84-426a-b665-2f8f5000210e · outbound
Dynamic Reinforcement Learning for Actors Shibata and Y
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5bb07383-ff04-4a80-947c-f193d8a4c461 · outbound
Dynamic Reinforcement Learning for Actors Shibata, T
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ba4ae6c7-14e9-4da0-b737-dc48e0b74834 · outbound
Dynamic Reinforcement Learning for Actors Unresolved cited work
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 87713aa5-2dcf-4f16-8bca-6267426a7050 · outbound
Dynamic Reinforcement Learning for Actors Unresolved cited work
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb6d35b8-cef8-4e61-82b7-2004467e4f7f · outbound
Dynamic Reinforcement Learning for Actors Evaluating the Factual Consistency of Large Language Models Through News Summarization
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc849c82-3175-4ab0-913b-26b928d99c6c · outbound
Dynamic Reinforcement Learning for Actors Unresolved cited work
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 86ff78fb-4582-4079-b070-c53680d4c48e · outbound
Dynamic Reinforcement Learning for Actors Unresolved cited work
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ba79934-6a43-44e3-927a-48126c95bca5 · outbound
Dynamic Reinforcement Learning for Actors Attention Is All You Need
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ee40a16-25a4-4651-bd07-51b1109295b4 · outbound
Dynamic Reinforcement Learning for Actors Vinyals, I
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 218db37f-57fb-4ea6-bc5c-381fbf37392b · outbound
Dynamic Reinforcement Learning for Actors Unresolved cited work
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8f1ca0b8-3041-4d76-a353-93973d99a0a4 · outbound
Dynamic Reinforcement Learning for Actors Yamashita and J
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac7a4ff4-5507-4e75-8529-9f068298d984 · outbound
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c005c9ce-927a-41a9-a49f-f7ab3b8cf065 · outbound
Dynamic Reinforcement Learning for Actors Assessing and Understanding Creativity in Large Language Models
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7be5454f-03bf-4c06-9b4d-673946b1e8ed · inbound
Temporally smoothed incremental model-based heuristic dynamic programming for command-filtered cascaded online learning flight control Dynamic Reinforcement Learning for Actors
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.