Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:30:30.809999Z
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 82 of 82 outbound references and 4 inbound Pith citation observations for arXiv:2505.21731.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:30:30.809999Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-08T03:13:07.963343Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-08T03:14:31.657309Z
82 of 82 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 134c8065-0c24-4436-a59c-2deac18b0316 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence write newline
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 579804ee-f6ef-4f18-b9aa-e86cb09b64bb · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence S., Courville, A
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9240cbd6-8e12-459f-9fb3-d2e7dbcd135a · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence R., Smith, K
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 120536c2-1522-4d03-bdf1-a65d55a067c8 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence The option-critic architecture
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 74c003d1-e5d9-44f1-9a6a-9f5a283bd392 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence P., Piot, B., Kapturowski, S., Sprechmann, P., Vitvitskyi, A., Guo, Z
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 1f3b2b39-9776-401d-9db0-14d438427cd6 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence P., Sprechmann, P., Vitvitskyi, A., Guo, Z
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 3e94862e-7013-40fa-af7f-cde9bfa1abab · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence L., Saxe, R., and Tenenbaum, J
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 676bf4e1-ca45-4d46-9246-d7fcfec8388d · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence H., Wang, R., and Manchester, I
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 93433e11-e145-468c-88f0-837ad138f22c · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Verifiable reinforcement learning via policy extraction
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 06f56c74-38ea-435e-87ac-e127784cdfcc · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence W., Hamrick, J
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 6e85f31b-bc9e-4260-949f-471045e0b002 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence G., Naddaf, Y., Veness, J., and Bowling, M
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3246711-7ea7-4c6b-895d-d87bc759b5b8 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence G., Dabney, W., and Munos, R
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e238dab9-de2e-4ac4-ad4d-371827d9f625 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Multi-objective causal bayesian optimization
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 794a1c90-393d-4399-9eb5-66c28a46ef36 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Deep reinforcement learning via object-centric attention
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 5d18e6fa-2960-40d7-aecc-abc79588d168 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 8ff71175-8dfa-4725-bdc8-ad8f8a12984e · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Galois: boosting deep reinforcement learning via generalizable logic synthesis
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e3853768-dce4-4b87-b112-96e7cb60e8ae · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 2789924c-4570-4ff5-8eb3-f36a3fc34d0a · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Quantifying generalization in reinforcement learning
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d32eb536-fd92-4c5c-8127-380f4323e35f · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Leveraging procedural generation to benchmark reinforcement learning
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 1875bd8c-6f3a-4eec-827f-aceaac255c8f · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence and Lake, B
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 2d2d5b48-bc6a-4b93-8c68-f22df1a3d950 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence D., Gershman, S
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 7e1e23fd-16b9-49fe-9537-00a9ab938a20 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence S., and Kersting, K
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 192e05ac-606d-4493-90f7-fd5428b00371 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Boosting object representation learning via motion and object continuity
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ee4bc06d-c081-4e95-8911-71830373944b · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence OCAtari : O bject-centric Atari 2600 reinforcement learning environments
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 272d3c3e-472b-44ce-af8b-119f3d5d9b35 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Interpretable concept bottlenecks to align reinforcement learning agents
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 538e8492-d9ea-45d7-95de-6fb302af25e2 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Unresolved cited work
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 5dcd3f36-6599-4907-8580-4ad6d661b94e · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence L., Koch, J., Sharkey, L
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation a5b9b7e6-bde0-4aa0-bb9b-febde3c9e8b9 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence P., and Kersting, K
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 1913a38d-58be-46bb-a9ac-f74471f60e8b · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence and Rothkopf, C
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 72705a8d-f14d-4c78-8a7f-40cc153c6b8d · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Unresolved cited work
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation a1873249-222c-4f9c-b53c-03303b19f054 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence IMPALA: scalable distributed deep-rl with importance weighted actor-learner architectures
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d9b21243-9d79-402b-9e35-bea9e6f6dfb7 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence C., and Bowling, M
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 56c903a7-f8ff-4d4a-881a-61bf55a5e9df · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Unresolved cited work
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 6dae4b7e-fa38-46c8-81f0-6a578433c42e · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence S., Brendel, W., Bethge, M., and Wichmann, F
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 3faca16e-e516-4a4a-948d-fd86a99f5976 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Atari agents, 2022
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation fb6793f3-a24a-450f-9b6e-40fa850c6464 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Visualizing and understanding atari agents, 2018
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 125c1fff-cf81-47e1-b668-68c8000dccd9 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence E., Pechenizkiy, M., and Mocanu, D
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 73d9257c-f359-48fd-8b32-23d3db41e002 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence H., Codel, C., Hofmann, K., Houghton, B., Kuno, N., Milani, S., Mohanty, S., Liebana, D
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 9a7df026-9a1b-4b49-a424-f7216e3cff77 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Benchmarking the spectrum of agent capabilities
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3c179ad-7de0-4550-8f6e-d5f6a1068edb · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence P., Ba, J., and Norouzi, M
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 3721e626-eb9e-4732-9136-39e1c5ee2159 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence T., Wang, Z., Heess, N., and Riedmiller, M
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 43a621b7-75ef-42e1-8da8-92d65c9bd741 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence S., and Kersting, K
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 2b74f479-a4cf-4d98-9bf0-f2d7c1359a53 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Rainbow: Combining improvements in deep reinforcement learning
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation fb38c080-68fd-458f-a903-44f960647f4e · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Unresolved cited work
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ace4704a-3670-496e-a8fc-0b06c0e4dfc1 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Adversarial examples are not bugs, they are features
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d3edb5b0-824d-4d0d-9ff6-85cb10855569 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence and Luo, S
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c78e215e-f521-4030-a423-853442247cc7 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence u ml, J., W \
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 5fa309cf-6886-46a4-9340-698ea206b8d3 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Unresolved cited work
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 34622877-14ce-42c2-bf3b-178fba0114c7 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Objective robustness in deep reinforcement learning
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 771a805f-c648-46b6-b2de-da1f8133fceb · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Interpretable and editable programmatic tree policies for reinforcement learning
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 652f6741-6d1d-4d1e-908e-582af4b86c39 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence R., Holyoak, K
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ac3a0fcf-ceda-46c7-a5a6-bda62020c866 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence M., Ullman, T
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 42422dfc-9f94-4c5c-b909-9aee2fe4f6ee · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Sub-policy adaptation for hierarchical reinforcement learning
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 598fa7ff-51f2-45ac-a73d-90793e8269d8 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence V., Sun, W., Singh, G., Deng, F., Jiang, J., and Ahn, S
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ceb82f8c-577a-4884-9a0a-99e86ccf5c7e · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Hierarchical programmatic reinforcement learning via learning to compose programs
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 50aa84f4-1e8c-4128-8a12-7f0b80986f43 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Insight: End-to-end neuro-symbolic visual reinforcement learning with language explanations
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 386998f7-4044-41c3-bbe8-9e66f93a7bf9 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence C., Bellemare, M
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e4e9ae23-f78e-4359-b390-7e22ea1572c6 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Sympol: Symbolic tree-based on-policy reinforcement learning
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 7d34ea38-5ea1-44ef-8262-bd700c754bdb · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence The perception of causality
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 77127910-8ab2-488c-8786-10902b1bfd82 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence H., Mohanty, S
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 2c7a38ea-21a3-4d72-926f-d05e7791c194 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence A., Veness, J., Bellemare, M
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 018f80ff-5a84-45c3-96e3-90b647baadb7 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Robust reinforcement learning: A review of foundations and recent advances
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ec94d5bb-fd7d-4491-8498-c7a512511dda · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Unresolved cited work
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 158dd0aa-e211-4d10-a186-61d7b3457e33 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Robust adversarial reinforcement learning
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 2d042f70-164f-4593-b547-6cd503594f5b · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Proximal policy optimization algorithms
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f62cab48-44ef-4e73-8924-5be07c177370 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Unresolved cited work
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation b6a6d657-8ed5-46e8-8baa-6591eb1263a1 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Reinforcement learning in strategy-based and atari games: A review of google deepminds innovations
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation bc146152-e5cc-4cf7-8cc7-2a845f72e286 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence S., and Kersting, K
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 2af951df-e7fb-45db-8a30-46dff3455326 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Unresolved cited work
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 2782b6a3-9823-4960-8ea0-16a72533fb0c · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Right for the right concept: Revising neuro-symbolic concepts by interacting with their explanations
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 19b5f7dd-0862-41ac-822c-551180500649 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Unresolved cited work
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c10350ad-2fc0-4750-8a28-538e0e119e4c · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Scaling up robust mdps using function approximation
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 708b08cd-e187-4ad1-a846-7433ecc88d27 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Unresolved cited work
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c367caee-a864-4c4e-97a4-095f854d039f · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence B., Kemp, C., Griffiths, T
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 9653dbf1-59c0-470e-aa83-7587194e3c4e · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Domain randomization for transferring deep neural networks from simulation to the real world
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c63006ba-c43e-43cf-a3ae-9d99a38d283d · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Munchausen reinforcement learning
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation b8709bae-7a9d-4bd2-8a38-4d87fa0cd5ae · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Iterated q -network: Beyond one-step bellman updates in deep reinforcement learning
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d59baca6-7d38-4c74-9488-745f91582963 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence S., Rothkopf, C
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation b1696a84-ee83-4107-b976-6c316e03bc37 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Towards generalizable reinforcement learning via causality-guided self-adaptive representations
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation fe302e13-807a-468f-8082-8bba517d5958 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence F., Raposo, D., Santoro, A., Bapst, V., Li, Y., Babuschkin, I., Tuyls, K., Reichert, D
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 73f0200f-3136-4fa3-8639-6d9f4bf4ded3 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence Natural environment benchmarks for reinforcement learning
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 2129f278-c5a9-4432-95c8-9ab95c88d773 · outbound
Deep Reinforcement Learning Agents are not even close to Human Intelligence N., et al
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 5c9da058-ebd4-4926-989a-b3dc64a38a1e · inbound
SLR: Automated Synthesis for Scalable Logical Reasoning Deep Reinforcement Learning Agents are not even close to Human Intelligence
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f075b78c-d9a8-42c9-a05a-53de452c851b · inbound
ActivationReasoning: Logical Reasoning in Latent Activation Spaces Deep Reinforcement Learning Agents are not even close to Human Intelligence
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 2155e889-86ac-4ca7-ab09-c222f376b1ea · inbound
GRAIL: Autonomous Concept Grounding for Neuro-Symbolic Reinforcement Learning Deep Reinforcement Learning Agents are not even close to Human Intelligence
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 048c7cc3-3145-45d9-b093-68c9ca79fa90 · inbound
Bridging Physical Reasoning and Task Generalization via Visual Action Outcome Reasoning Alignment Deep Reinforcement Learning Agents are not even close to Human Intelligence
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.