Pith. sign in

Paper Citation Record · LEDGER

Deep Reinforcement Learning Agents are not even close to Human Intelligence

As of 11 August 2026, this Paper Citation Record lists 82 of 82 outbound references and 4 inbound Pith citation observations for arXiv:2505.21731.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.21731 v1

Coverage vector

measured 82 of 82 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:30:30.809999Z

measured 86 of 86 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-08T03:13:07.963343Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-08T03:14:31.657309Z

Reference resolution

82 of 82 outbound references displayed

  • verified exact0
  • verified fuzzy66
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 134c8065-0c24-4436-a59c-2deac18b0316 · outbound

This paper cites write newline.

Deep Reinforcement Learning Agents are not even close to Human Intelligence write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:22.452178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:22.452178Z digest=sha256:78e91bdce534d931cf92cc42a9ab8e60bd68ac28aaa90568e7f14a842a3d1bdc

Observation 579804ee-f6ef-4f18-b9aa-e86cb09b64bb · outbound

This paper cites S., Courville, A.

Deep Reinforcement Learning Agents are not even close to Human Intelligence S., Courville, A

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:22.556997Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:22.556997Z digest=sha256:705aa69cb5dccccfbc31bfc1874eff7cf0a2f9b867313977246f257e264332a6

Observation 9240cbd6-8e12-459f-9fb3-d2e7dbcd135a · outbound

This paper cites R., Smith, K.

Deep Reinforcement Learning Agents are not even close to Human Intelligence R., Smith, K

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:46.919899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:22.674956Z digest=sha256:6ae8f7c91e4719c43d7d2609b00d533563171b8c5c8e59c82bf6ff477523cbe1

Observation 120536c2-1522-4d03-bdf1-a65d55a067c8 · outbound

This paper cites The option-critic architecture.

Deep Reinforcement Learning Agents are not even close to Human Intelligence The option-critic architecture

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:46.687208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:22.737487Z digest=sha256:305109f017ccbd51c20e297effb3cb852fbef8397f17bec8681897658e01c0f7

Observation 74c003d1-e5d9-44f1-9a6a-9f5a283bd392 · outbound

This paper cites P., Piot, B., Kapturowski, S., Sprechmann, P., Vitvitskyi, A., Guo, Z.

Deep Reinforcement Learning Agents are not even close to Human Intelligence P., Piot, B., Kapturowski, S., Sprechmann, P., Vitvitskyi, A., Guo, Z

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:46.435022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:22.820055Z digest=sha256:95213196c0fb09a74141eb6ce1232936da37432614fc26f199949beebebe39ad

Observation 1f3b2b39-9776-401d-9db0-14d438427cd6 · outbound

This paper cites P., Sprechmann, P., Vitvitskyi, A., Guo, Z.

Deep Reinforcement Learning Agents are not even close to Human Intelligence P., Sprechmann, P., Vitvitskyi, A., Guo, Z

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:46.153627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:22.893975Z digest=sha256:6fa3716ade9043fa4e87101ef4d49be2ca481e8afcca950f6cda7cad7b383813

Observation 3e94862e-7013-40fa-af7f-cde9bfa1abab · outbound

This paper cites L., Saxe, R., and Tenenbaum, J.

Deep Reinforcement Learning Agents are not even close to Human Intelligence L., Saxe, R., and Tenenbaum, J

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:45.954955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.026307Z digest=sha256:0f97ecf925b13e30308158a68596bc1104859192fa00bc25a86c6d2138511b5a

Observation 676bf4e1-ca45-4d46-9246-d7fcfec8388d · outbound

This paper cites H., Wang, R., and Manchester, I.

Deep Reinforcement Learning Agents are not even close to Human Intelligence H., Wang, R., and Manchester, I

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:45.757560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.121760Z digest=sha256:093bfda0efa69b011b11f8870d612e014f1d5a2521e33642bf63d1c05dfa8f4b

Observation 93433e11-e145-468c-88f0-837ad138f22c · outbound

This paper cites Verifiable reinforcement learning via policy extraction.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Verifiable reinforcement learning via policy extraction

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:45.550446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.225533Z digest=sha256:029e0895462ddec25d9167377cdfdf1546e62706061bafb8b331ffbffa78829b

Observation 06f56c74-38ea-435e-87ac-e127784cdfcc · outbound

This paper cites W., Hamrick, J.

Deep Reinforcement Learning Agents are not even close to Human Intelligence W., Hamrick, J

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:45.383378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.305350Z digest=sha256:e18c7902d26d29daeb9bd0ade0c2a9eaeffb3b43bb5fb031c70b7889f81763f7

Observation 6e85f31b-bc9e-4260-949f-471045e0b002 · outbound

This paper cites G., Naddaf, Y., Veness, J., and Bowling, M.

Deep Reinforcement Learning Agents are not even close to Human Intelligence G., Naddaf, Y., Veness, J., and Bowling, M

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:23.367460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:23.367460Z digest=sha256:06fc3fa95ac689787b7dc4627a1c1e15ed18272c6a56f3fab5e4fd31f3689977

Observation b3246711-7ea7-4c6b-895d-d87bc759b5b8 · outbound

This paper cites G., Dabney, W., and Munos, R.

Deep Reinforcement Learning Agents are not even close to Human Intelligence G., Dabney, W., and Munos, R

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:45.228526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.454742Z digest=sha256:08976969cd4b785ad203f004f29218edfb7483b24e11c14486ae0963e2e5b90c

Observation e238dab9-de2e-4ac4-ad4d-371827d9f625 · outbound

This paper cites Multi-objective causal bayesian optimization.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Multi-objective causal bayesian optimization

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:45.014623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.584889Z digest=sha256:a8abf85c9822e3c84f4248cb477fec4ea90ae930141cc02c261b36f3b0f3ecec

Observation 794a1c90-393d-4399-9eb5-66c28a46ef36 · outbound

This paper cites Deep reinforcement learning via object-centric attention.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Deep reinforcement learning via object-centric attention

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:44.883287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.672931Z digest=sha256:904a8dc521822b1325a662dab2d2c2e383e7578626d913bb1ad5677a820bccfd

Observation 5d18e6fa-2960-40d7-aecc-abc79588d168 · outbound

This paper cites an unresolved cited work.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:30:44.722421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.745902Z digest=sha256:4613f1d66243fefb49709e014f40d6a02319b2ae88c9536d380769115d2c6f35

Observation 8ff71175-8dfa-4725-bdc8-ad8f8a12984e · outbound

This paper cites Galois: boosting deep reinforcement learning via generalizable logic synthesis.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Galois: boosting deep reinforcement learning via generalizable logic synthesis

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:44.505967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.806792Z digest=sha256:1136eff9d3694a061a03e73199f21a7688a5239a7b5e3885b63aad1da32b9a8c

Observation e3853768-dce4-4b87-b112-96e7cb60e8ae · outbound

This paper cites an unresolved cited work.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:30:44.345921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.874994Z digest=sha256:52ffdba680adf6d74d290e103241e80734783b0d723b21192e65e624b32abc25

Observation 2789924c-4570-4ff5-8eb3-f36a3fc34d0a · outbound

This paper cites Quantifying generalization in reinforcement learning.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Quantifying generalization in reinforcement learning

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:44.164387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.981980Z digest=sha256:202741df376cbf124116892e6165a18d54e441ef0dc219a45a35e2d10c5504e0

Observation d32eb536-fd92-4c5c-8127-380f4323e35f · outbound

This paper cites Leveraging procedural generation to benchmark reinforcement learning.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Leveraging procedural generation to benchmark reinforcement learning

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:43.981249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.057439Z digest=sha256:2c0679bc80c21cba9b213bcbdd862670efea93c63c71bf93c68c1a5bef7ec779

Observation 1875bd8c-6f3a-4eec-827f-aceaac255c8f · outbound

This paper cites and Lake, B.

Deep Reinforcement Learning Agents are not even close to Human Intelligence and Lake, B

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:43.774743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.117765Z digest=sha256:be3050c1c50b7817f236ffa640db8f53cc9be4f524b6f68e15393cb9cf86b03a

Observation 2d2d5b48-bc6a-4b93-8c68-f22df1a3d950 · outbound

This paper cites D., Gershman, S.

Deep Reinforcement Learning Agents are not even close to Human Intelligence D., Gershman, S

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:43.483560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.181658Z digest=sha256:48daf12ead3f16b4d9c7cc43adefd3e78b3737849249d9ff722a0afd800c4e10

Observation 7e1e23fd-16b9-49fe-9537-00a9ab938a20 · outbound

This paper cites S., and Kersting, K.

Deep Reinforcement Learning Agents are not even close to Human Intelligence S., and Kersting, K

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:43.259217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.237690Z digest=sha256:c074c69000ca19eb741f973f2ef6d85e4d1546e6ee81d41bff3792478731f1cc

Observation 192e05ac-606d-4493-90f7-fd5428b00371 · outbound

This paper cites Boosting object representation learning via motion and object continuity.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Boosting object representation learning via motion and object continuity

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:43.011284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.294299Z digest=sha256:07b0e7e143a44f2e0e17703a0f5770fcefa708dae84b2a0cf1f1d54f4b10beb4

Observation ee4bc06d-c081-4e95-8911-71830373944b · outbound

This paper cites OCAtari : O bject-centric Atari 2600 reinforcement learning environments.

Deep Reinforcement Learning Agents are not even close to Human Intelligence OCAtari : O bject-centric Atari 2600 reinforcement learning environments

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:42.782945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.378097Z digest=sha256:3e98ddedd1db7c25cbc67465e2b487c8cf838d556e09ecb4f8324db448d31118

Observation 272d3c3e-472b-44ce-af8b-119f3d5d9b35 · outbound

This paper cites Interpretable concept bottlenecks to align reinforcement learning agents.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Interpretable concept bottlenecks to align reinforcement learning agents

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:42.579471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.604945Z digest=sha256:ad98fc3161c6cad03dc6d836ef000b873b11a4894640e144b8ec8fd67034cdea

Observation 538e8492-d9ea-45d7-95de-6fb302af25e2 · outbound

This paper cites an unresolved cited work.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:30:42.334222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.734745Z digest=sha256:7eb826d65ff09e1a4451b71bc5ad299ce54bc3f956eca95df3406909d752b6a0

Observation 5dcd3f36-6599-4907-8580-4ad6d661b94e · outbound

This paper cites L., Koch, J., Sharkey, L.

Deep Reinforcement Learning Agents are not even close to Human Intelligence L., Koch, J., Sharkey, L

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:42.088428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.797068Z digest=sha256:adee98f8b38176ed56528dffb568729eeb9599f78acad1064bc37472729816c7

Observation a5b9b7e6-bde0-4aa0-bb9b-febde3c9e8b9 · outbound

This paper cites P., and Kersting, K.

Deep Reinforcement Learning Agents are not even close to Human Intelligence P., and Kersting, K

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:41.859081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.864805Z digest=sha256:e95b8bce001166edf772b30a407973503706632c890e271c071cc2c6a74dfcc5

Observation 1913a38d-58be-46bb-a9ac-f74471f60e8b · outbound

This paper cites and Rothkopf, C.

Deep Reinforcement Learning Agents are not even close to Human Intelligence and Rothkopf, C

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:41.623553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.984867Z digest=sha256:d9dcf86818c2748935f16ef46deb9c9d9dbedac5756514d1a450ff6500881767

Observation 72705a8d-f14d-4c78-8a7f-40cc153c6b8d · outbound

This paper cites an unresolved cited work.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:30:41.364656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:25.094899Z digest=sha256:749e70e709e259e3c44bc8d282e6bcedeb2044c5e1bf6a7b1723e628efbb2f1d

Observation a1873249-222c-4f9c-b53c-03303b19f054 · outbound

This paper cites IMPALA: scalable distributed deep-rl with importance weighted actor-learner architectures.

Deep Reinforcement Learning Agents are not even close to Human Intelligence IMPALA: scalable distributed deep-rl with importance weighted actor-learner architectures

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:41.125147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:25.305027Z digest=sha256:b6393cde7acf919e131fc74f6cacef1f92ae40a74dddfddfa9d983c6dcf5c18d

Observation d9b21243-9d79-402b-9e35-bea9e6f6dfb7 · outbound

This paper cites C., and Bowling, M.

Deep Reinforcement Learning Agents are not even close to Human Intelligence C., and Bowling, M

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:40.860028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:25.475266Z digest=sha256:ba8600b9727db1bb8c5451ae526f43cfbc2b45cf472c5caee9ad84fb15198835

Observation 56c903a7-f8ff-4d4a-881a-61bf55a5e9df · outbound

This paper cites an unresolved cited work.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:30:40.665376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:25.675365Z digest=sha256:d8404c11f093b7996ac9beae32e581a0807e4116391dc6fe69bd85e338c9369c

Observation 6dae4b7e-fa38-46c8-81f0-6a578433c42e · outbound

This paper cites S., Brendel, W., Bethge, M., and Wichmann, F.

Deep Reinforcement Learning Agents are not even close to Human Intelligence S., Brendel, W., Bethge, M., and Wichmann, F

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:40.441644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:25.814825Z digest=sha256:4f905325ac4ee9c3f1773a798cdd4798e25498d8b3325ab4520bed895608d007

Observation 3faca16e-e516-4a4a-948d-fd86a99f5976 · outbound

This paper cites Atari agents, 2022.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Atari agents, 2022

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:40.241239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:25.934736Z digest=sha256:7a5ae013fa2a07c9a5ece02bc8992b5d73e637573084a3b59a5cc5971429eb90

Observation fb6793f3-a24a-450f-9b6e-40fa850c6464 · outbound

This paper cites Visualizing and understanding atari agents, 2018.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Visualizing and understanding atari agents, 2018

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:40.033921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:26.012180Z digest=sha256:e2cfd58b4bccbfc1d0150821269069e00af8df2e3b1df1cc6fbdff6df2bc3f02

Observation 125c1fff-cf81-47e1-b668-68c8000dccd9 · outbound

This paper cites E., Pechenizkiy, M., and Mocanu, D.

Deep Reinforcement Learning Agents are not even close to Human Intelligence E., Pechenizkiy, M., and Mocanu, D

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:39.842610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:26.125076Z digest=sha256:3df73b808cc02d97b8b15674a3e67c22e499859010070250d3db3ff216a984ea

Observation 73d9257c-f359-48fd-8b32-23d3db41e002 · outbound

This paper cites H., Codel, C., Hofmann, K., Houghton, B., Kuno, N., Milani, S., Mohanty, S., Liebana, D.

Deep Reinforcement Learning Agents are not even close to Human Intelligence H., Codel, C., Hofmann, K., Houghton, B., Kuno, N., Milani, S., Mohanty, S., Liebana, D

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:39.699258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:26.256710Z digest=sha256:3f210a73eb35f26a411e426ce8c759527091d5da2b4dbfc7e9464c97d05f2f17

Observation 9a7df026-9a1b-4b49-a424-f7216e3cff77 · outbound

This paper cites Benchmarking the spectrum of agent capabilities.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Benchmarking the spectrum of agent capabilities

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:26.384892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:26.384892Z digest=sha256:aa2e618eb010658e77cb825faecd1853509a95c003ee7222eb04ce10531d75e6

Observation b3c179ad-7de0-4550-8f6e-d5f6a1068edb · outbound

This paper cites P., Ba, J., and Norouzi, M.

Deep Reinforcement Learning Agents are not even close to Human Intelligence P., Ba, J., and Norouzi, M

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:39.571980Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:26.458693Z digest=sha256:def6967474030bb23028bf0e6f33acdfab9dde43cf2674ffd6b06d6dc58a1328

Observation 3721e626-eb9e-4732-9136-39e1c5ee2159 · outbound

This paper cites T., Wang, Z., Heess, N., and Riedmiller, M.

Deep Reinforcement Learning Agents are not even close to Human Intelligence T., Wang, Z., Heess, N., and Riedmiller, M

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:39.420596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:26.572422Z digest=sha256:483483383ddc4d41a14451d3d041bdae161b4105ad6a436f35f872e475bac6cd

Observation 43a621b7-75ef-42e1-8da8-92d65c9bd741 · outbound

This paper cites S., and Kersting, K.

Deep Reinforcement Learning Agents are not even close to Human Intelligence S., and Kersting, K

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:39.230738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:26.897578Z digest=sha256:7440b7214c554a1b7fcd116ba109d13f5b7e5972f9b421f9454d30b635c7f4ad

Observation 2b74f479-a4cf-4d98-9bf0-f2d7c1359a53 · outbound

This paper cites Rainbow: Combining improvements in deep reinforcement learning.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Rainbow: Combining improvements in deep reinforcement learning

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:39.073626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:26.997295Z digest=sha256:fa778d2bf608c01f25ea6c26b408092bd7c401c3080adb3c5f3a3edf2b1bf8e6

Observation fb38c080-68fd-458f-a903-44f960647f4e · outbound

This paper cites an unresolved cited work.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:30:38.820110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:27.126459Z digest=sha256:72efa372c8557f3cf0ea1822049fd026b96869eab94c7db98b72fa61613b5470

Observation ace4704a-3670-496e-a8fc-0b06c0e4dfc1 · outbound

This paper cites Adversarial examples are not bugs, they are features.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Adversarial examples are not bugs, they are features

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:38.594220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:27.264745Z digest=sha256:d6d0a2b7bdc44eeeb63534d007fbd915507dca3e89f0a818d9c8e7033ee6c456

Observation d3edb5b0-824d-4d0d-9ff6-85cb10855569 · outbound

This paper cites and Luo, S.

Deep Reinforcement Learning Agents are not even close to Human Intelligence and Luo, S

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:38.383136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:27.445351Z digest=sha256:9df4905ed6c473befdbd0c6d2f18e009f1c756da3fa5165b36e8026d7a93578d

Observation c78e215e-f521-4030-a423-853442247cc7 · outbound

This paper cites u ml, J., W \.

Deep Reinforcement Learning Agents are not even close to Human Intelligence u ml, J., W \

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:38.193348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:27.574379Z digest=sha256:fbc48182f559faadc07238e222854a1a5e55034b4386eb8d34a6c56f9bd9431d

Observation 5fa309cf-6886-46a4-9340-698ea206b8d3 · outbound

This paper cites an unresolved cited work.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:30:37.958212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:27.676237Z digest=sha256:5cec77f631ce1eb39bf396bbff359703b965ee5c23e84501fcfe4176b99647fb

Observation 34622877-14ce-42c2-bf3b-178fba0114c7 · outbound

This paper cites Objective robustness in deep reinforcement learning.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Objective robustness in deep reinforcement learning

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:37.735890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:27.864819Z digest=sha256:a54bda191431afead11a8465cc1a70945c286c1a83b9e86e7172c6a207db8c66

Observation 771a805f-c648-46b6-b2de-da1f8133fceb · outbound

This paper cites Interpretable and editable programmatic tree policies for reinforcement learning.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Interpretable and editable programmatic tree policies for reinforcement learning

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:37.509789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:27.954897Z digest=sha256:b0f33239a754d32897c52417cf9ff3756a6488aef2ae06a7e6eed5abc508c1fa

Observation 652f6741-6d1d-4d1e-908e-582af4b86c39 · outbound

This paper cites R., Holyoak, K.

Deep Reinforcement Learning Agents are not even close to Human Intelligence R., Holyoak, K

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:37.340172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.045609Z digest=sha256:af5cfec1e62f5afe040c8b0879a5eb4baddd48d6613362828c45b8e01153ffdf

Observation ac3a0fcf-ceda-46c7-a5a6-bda62020c866 · outbound

This paper cites M., Ullman, T.

Deep Reinforcement Learning Agents are not even close to Human Intelligence M., Ullman, T

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:37.056247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.234741Z digest=sha256:9cc581b302a7dbc5cfb5ba62e96d700ffdc83ee20b8563d460a12764dd3da9cf

Observation 42422dfc-9f94-4c5c-b909-9aee2fe4f6ee · outbound

This paper cites Sub-policy adaptation for hierarchical reinforcement learning.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Sub-policy adaptation for hierarchical reinforcement learning

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:36.831752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.504315Z digest=sha256:68d3e37455a9feb282af24f61aa843663a37ae47f1278b16e4643be80de30a68

Observation 598fa7ff-51f2-45ac-a73d-90793e8269d8 · outbound

This paper cites V., Sun, W., Singh, G., Deng, F., Jiang, J., and Ahn, S.

Deep Reinforcement Learning Agents are not even close to Human Intelligence V., Sun, W., Singh, G., Deng, F., Jiang, J., and Ahn, S

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:36.669772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.556179Z digest=sha256:52f5e9d4342902e5ee4d1432ea439f5833b646618782f60f93927dd06e62bbd6

Observation ceb82f8c-577a-4884-9a0a-99e86ccf5c7e · outbound

This paper cites Hierarchical programmatic reinforcement learning via learning to compose programs.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Hierarchical programmatic reinforcement learning via learning to compose programs

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:36.480713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.633760Z digest=sha256:f530cec569c9afc036861a2b120c8ae1a1236a9615a677b8e0b0002cdd198ef9

Observation 50aa84f4-1e8c-4128-8a12-7f0b80986f43 · outbound

This paper cites Insight: End-to-end neuro-symbolic visual reinforcement learning with language explanations.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Insight: End-to-end neuro-symbolic visual reinforcement learning with language explanations

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:36.280885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.691611Z digest=sha256:487a239e363bfb7160306c34015c2889b6d6d50164083339bf63c7eeb2eb65af

Observation 386998f7-4044-41c3-bbe8-9e66f93a7bf9 · outbound

This paper cites C., Bellemare, M.

Deep Reinforcement Learning Agents are not even close to Human Intelligence C., Bellemare, M

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:36.044746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.735806Z digest=sha256:06751188f87bb2c94a6c34f0a1f403189a584c430171d1dd4ba4ac2fba22d6d4

Observation e4e9ae23-f78e-4359-b390-7e22ea1572c6 · outbound

This paper cites Sympol: Symbolic tree-based on-policy reinforcement learning.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Sympol: Symbolic tree-based on-policy reinforcement learning

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:35.888727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.802987Z digest=sha256:16477b60e15e7aecc50798c60b74daae1c3517d67426b3f2559e6c69d37a67fb

Observation 7d34ea38-5ea1-44ef-8262-bd700c754bdb · outbound

This paper cites The perception of causality.

Deep Reinforcement Learning Agents are not even close to Human Intelligence The perception of causality

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:35.718227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.843655Z digest=sha256:041c111a4f3f519064c430d688efdd85fe1817809badec39cdcc8e8dfd711426

Observation 77127910-8ab2-488c-8786-10902b1bfd82 · outbound

This paper cites H., Mohanty, S.

Deep Reinforcement Learning Agents are not even close to Human Intelligence H., Mohanty, S

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:35.582517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.908966Z digest=sha256:1f6507049388c5f732a343ff6a64a982bc0e0219d6e9554dd376edddffed2b23

Observation 2c7a38ea-21a3-4d72-926f-d05e7791c194 · outbound

This paper cites A., Veness, J., Bellemare, M.

Deep Reinforcement Learning Agents are not even close to Human Intelligence A., Veness, J., Bellemare, M

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:35.329174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.937695Z digest=sha256:0bcd78e92914e001f11324351fb3889a0caf8e7947063ca93e920212704e3ca2

Observation 018f80ff-5a84-45c3-96e3-90b647baadb7 · outbound

This paper cites Robust reinforcement learning: A review of foundations and recent advances.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Robust reinforcement learning: A review of foundations and recent advances

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:35.141234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:29.064298Z digest=sha256:c6f773cf7c29cc03de24b61bb29790b2cefa55d885005ae12c0b188a041c2673

Observation ec94d5bb-fd7d-4491-8498-c7a512511dda · outbound

This paper cites an unresolved cited work.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:30:34.956408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:29.173116Z digest=sha256:fb70a8920d2aba2409d41e5c328b7288cacd405bb42e4ca3c47abc3b0aeb1527

Observation 158dd0aa-e211-4d10-a186-61d7b3457e33 · outbound

This paper cites Robust adversarial reinforcement learning.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Robust adversarial reinforcement learning

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:34.763140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:29.286940Z digest=sha256:2a880c590a5bbd1b5c5af055a43ef902afd5210335e353e121e0d65a78c3cf20

Observation 2d042f70-164f-4593-b547-6cd503594f5b · outbound

This paper cites Proximal policy optimization algorithms.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Proximal policy optimization algorithms

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:34.536879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:29.350658Z digest=sha256:5c5665fb7eca2249544bcca4f5012e92aacebb1597c20ad799eff5cdbc85f469

Observation f62cab48-44ef-4e73-8924-5be07c177370 · outbound

This paper cites an unresolved cited work.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Unresolved cited work

Reference 66

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:30:34.369761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:29.457219Z digest=sha256:0e25578e5d07d5fa083e13b7a133437b6bfc33a7e55702120ac4133a7d1a99d0

Observation b6a6d657-8ed5-46e8-8baa-6591eb1263a1 · outbound

This paper cites Reinforcement learning in strategy-based and atari games: A review of google deepminds innovations.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Reinforcement learning in strategy-based and atari games: A review of google deepminds innovations

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:34.128710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:29.569395Z digest=sha256:03765b7826361bff13c62ddc90358b4784b52893b1429948e2dabf5e91293dbd

Observation bc146152-e5cc-4cf7-8cc7-2a845f72e286 · outbound

This paper cites S., and Kersting, K.

Deep Reinforcement Learning Agents are not even close to Human Intelligence S., and Kersting, K

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:33.938188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:29.630331Z digest=sha256:23f5b2f3b20850792ff3a929ea933d20f4df2f3b0916729ffa95d7a21dafe9db

Observation 2af951df-e7fb-45db-8a30-46dff3455326 · outbound

This paper cites an unresolved cited work.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Unresolved cited work

Reference 69

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:30:33.797320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:29.709457Z digest=sha256:7d211a6bce85baf0645ed6479ac2ca29089ca72c6721599f804b413dc0e1d061

Observation 2782b6a3-9823-4960-8ea0-16a72533fb0c · outbound

This paper cites Right for the right concept: Revising neuro-symbolic concepts by interacting with their explanations.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Right for the right concept: Revising neuro-symbolic concepts by interacting with their explanations

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:33.618260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:29.785574Z digest=sha256:7e9c560a9f555275b5fa2ffdaeabdf221b0129da08d1c7d81792f2d969c43163

Observation 19b5f7dd-0862-41ac-822c-551180500649 · outbound

This paper cites an unresolved cited work.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Unresolved cited work

Reference 71

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:30:33.411410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:29.900069Z digest=sha256:7744fc8a1f510a12797a060850b4bac59adf1a5f7ac3a36101037490e62d6704

Observation c10350ad-2fc0-4750-8a28-538e0e119e4c · outbound

This paper cites Scaling up robust mdps using function approximation.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Scaling up robust mdps using function approximation

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:33.215847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:29.943136Z digest=sha256:ee51193913ebd71f0079defd6e85b2d59ca463db2948df250a0878d97ba81dae

Observation 708b08cd-e187-4ad1-a846-7433ecc88d27 · outbound

This paper cites an unresolved cited work.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Unresolved cited work

Reference 73

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:30:32.968146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:30.039935Z digest=sha256:d9752f3cbe0dde1175c88a66a505641762a5ccb3726da4fc66bf44f7e9d51cd9

Observation c367caee-a864-4c4e-97a4-095f854d039f · outbound

This paper cites B., Kemp, C., Griffiths, T.

Deep Reinforcement Learning Agents are not even close to Human Intelligence B., Kemp, C., Griffiths, T

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:32.742695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:30.172366Z digest=sha256:4b70ba8e60678ae9ec0c679bdc78f8ae88c4b61d1fe9f1d070de55a038b9a189

Observation 9653dbf1-59c0-470e-aa83-7587194e3c4e · outbound

This paper cites Domain randomization for transferring deep neural networks from simulation to the real world.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Domain randomization for transferring deep neural networks from simulation to the real world

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:32.543806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:30.251354Z digest=sha256:8f2ffd5862599cda56be996a7cc62d8f13cea43ed5e49f8da3a634b4a91e5fc0

Observation c63006ba-c43e-43cf-a3ae-9d99a38d283d · outbound

This paper cites Munchausen reinforcement learning.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Munchausen reinforcement learning

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:32.305476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:30.293466Z digest=sha256:e5e3168f0bd0bb2b90c496197cace6dd58878d3bd01fe488108ac43a6d35b683

Observation b8709bae-7a9d-4bd2-8a38-4d87fa0cd5ae · outbound

This paper cites Iterated q -network: Beyond one-step bellman updates in deep reinforcement learning.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Iterated q -network: Beyond one-step bellman updates in deep reinforcement learning

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:32.116826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:30.390115Z digest=sha256:64e1a793d59d38147a8eaadd3371d0f46a3567b146548d3043cb4141dd46b256

Observation d59baca6-7d38-4c74-9488-745f91582963 · outbound

This paper cites S., Rothkopf, C.

Deep Reinforcement Learning Agents are not even close to Human Intelligence S., Rothkopf, C

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:31.868781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:30.468977Z digest=sha256:a1d537b379090c2c9c9eaf3bfa3f5181b17384a44e402f2f59edd0846a8d6734

Observation b1696a84-ee83-4107-b976-6c316e03bc37 · outbound

This paper cites Towards generalizable reinforcement learning via causality-guided self-adaptive representations.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Towards generalizable reinforcement learning via causality-guided self-adaptive representations

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:31.586099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:30.538591Z digest=sha256:708c9aa9f83db6fc852bbfee498d0f34df7904439e53d401fb7d6e0a95d89010

Observation fe302e13-807a-468f-8082-8bba517d5958 · outbound

This paper cites F., Raposo, D., Santoro, A., Bapst, V., Li, Y., Babuschkin, I., Tuyls, K., Reichert, D.

Deep Reinforcement Learning Agents are not even close to Human Intelligence F., Raposo, D., Santoro, A., Bapst, V., Li, Y., Babuschkin, I., Tuyls, K., Reichert, D

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:31.382070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:30.653772Z digest=sha256:1383450fe8111456731fea4681e27d72ced5c9aafd9601f371feb1faf30dece8

Observation 73f0200f-3136-4fa3-8639-6d9f4bf4ded3 · outbound

This paper cites Natural environment benchmarks for reinforcement learning.

Deep Reinforcement Learning Agents are not even close to Human Intelligence Natural environment benchmarks for reinforcement learning

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:31.221306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:30.724744Z digest=sha256:5254aea380f13d2aea17c5ce1320df17df3f548a196dcdb8c43edb4175918f2e

Observation 2129f278-c5a9-4432-95c8-9ab95c88d773 · outbound

This paper cites N., et al.

Deep Reinforcement Learning Agents are not even close to Human Intelligence N., et al

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:31.003953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T13:30:30.809999Z digest=sha256:debc3aeccb1c35915ee07ec365b17c6aebb541d44d33205e189979db660e456c

Pith citing papers

Observation 5c9da058-ebd4-4926-989a-b3dc64a38a1e · inbound

SLR: Automated Synthesis for Scalable Logical Reasoning cites this paper.

SLR: Automated Synthesis for Scalable Logical Reasoning Deep Reinforcement Learning Agents are not even close to Human Intelligence

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-19T08:52:13.803998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-19T08:47:17.008113Z digest=sha256:9cc97cfa78f8b71a9fadf71147322b3fead7cbbf3642495c73495660ab90c65d

Observation f075b78c-d9a8-42c9-a05a-53de452c851b · inbound

ActivationReasoning: Logical Reasoning in Latent Activation Spaces cites this paper.

ActivationReasoning: Logical Reasoning in Latent Activation Spaces Deep Reinforcement Learning Agents are not even close to Human Intelligence

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T05:45:56.222552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-18T05:43:45.863209Z digest=sha256:158c3aec54063bc260da963bd9abde6fe186cbf316cbf60f1a2d7da95a49d83b

Observation 2155e889-86ac-4ca7-ab09-c222f376b1ea · inbound

GRAIL: Autonomous Concept Grounding for Neuro-Symbolic Reinforcement Learning cites this paper.

GRAIL: Autonomous Concept Grounding for Neuro-Symbolic Reinforcement Learning Deep Reinforcement Learning Agents are not even close to Human Intelligence

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T07:21:55.178067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-10T07:20:32.282630Z digest=sha256:86f42622fd41c8a343a7050317295d6d65ebc20a06be53b3b7f8ebb964479aab

Observation 048c7cc3-3145-45d9-b093-68c9ca79fa90 · inbound

Bridging Physical Reasoning and Task Generalization via Visual Action Outcome Reasoning Alignment cites this paper.

Bridging Physical Reasoning and Task Generalization via Visual Action Outcome Reasoning Alignment Deep Reinforcement Learning Agents are not even close to Human Intelligence

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-07-08T03:14:31.658780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-07-08T03:13:07.963343Z digest=sha256:7b14f74ee3f25bfb69b1e622e1e3d258cc598258aed0e15a0d9fbab9ca968e29