Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T19:09:54.006504Z
Paper Citation Record · LEDGER
As of 12 August 2026, this Paper Citation Record lists 100 of 300 outbound references and 0 inbound Pith citation observations for arXiv:2412.07177.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T19:09:54.006504Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
100 of 300 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation f670229a-7884-4a60-a7c8-b7924ffe79e2 · outbound
Effective Reward Specification in Deep Reinforcement Learning write newline
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ce6d39f-ebc7-4668-9d47-185b987f835d · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf6479eb-f08c-4c86-8f4c-3976865bd0b1 · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9582fd8-9708-4b24-893a-6f74f4a45b4f · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e41861d2-2621-4006-a9d9-57e5d91e27e3 · outbound
Effective Reward Specification in Deep Reinforcement Learning and Ng, A
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 089b4a0c-6696-494f-853d-b41283a0127d · outbound
Effective Reward Specification in Deep Reinforcement Learning and Ng, A
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9720b40-68cf-48ee-a7c9-0ee8239e1cf2 · outbound
Effective Reward Specification in Deep Reinforcement Learning K., Littman, M., Precup, D., and Singh, S
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation adfd7db3-5749-44a6-8eba-ed931304f984 · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be446dbb-fbc9-4be0-b634-f49cbfb80154 · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a8a3d86c-0c6a-499e-8b9d-b6123b1aacb4 · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e22357e-ddb0-4685-bbcf-7166c49ae5ea · outbound
Effective Reward Specification in Deep Reinforcement Learning Feudal Multi-Agent Hierarchies for Cooperative Reinforcement Learning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b3ede61-5ab0-4aaf-a747-cec72cf31152 · outbound
Effective Reward Specification in Deep Reinforcement Learning Do As I Can, Not As I Say: Grounding Language in Robotic Affordances
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7324c04-d8f0-4e08-b7ba-719ab1f929be · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4da7c10a-aab0-40f4-896d-aaaea15babd1 · outbound
Effective Reward Specification in Deep Reinforcement Learning Solving Rubik's Cube with a Robot Hand
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation add28bfa-6651-4b5e-a342-db45149cb1a9 · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91c3085d-16f1-4244-97b7-213c21c114fe · outbound
Effective Reward Specification in Deep Reinforcement Learning Deep Reinforcement Learning for Navigation in AAA Video Games
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a974e33a-8fd6-4da6-96d7-8583e287f451 · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 595fc39b-f931-4b4d-abbb-d9a20d1d694f · outbound
Effective Reward Specification in Deep Reinforcement Learning A Survey of Exploration Methods in Reinforcement Learning
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd3de39f-95d6-46ae-8a40-2e46c3625b5b · outbound
Effective Reward Specification in Deep Reinforcement Learning Concrete Problems in AI Safety
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d1d635b-38b9-48a4-aaa5-f5256f9fa57f · outbound
Effective Reward Specification in Deep Reinforcement Learning Explaining Reinforcement Learning to Mere Mortals: An Empirical Study
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e341f36f-1766-48f1-9580-b6bfd63d85c2 · outbound
Effective Reward Specification in Deep Reinforcement Learning P., and Zaremba, W
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d7f753a-6bbe-4747-acbc-fe25c17ebb63 · outbound
Effective Reward Specification in Deep Reinforcement Learning M., Baker, B., Chociej, M., Jozefowicz, R., McGrew, B., Pachocki, J., Petron, A., Plappert, M., Powell, G., Ray, A., et al
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0a6d152-ee32-46e6-aa7b-c3f586a92ed5 · outbound
Effective Reward Specification in Deep Reinforcement Learning and Doshi, P
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 628072b7-53e6-49e4-a569-91408f889bcb · outbound
Effective Reward Specification in Deep Reinforcement Learning Accurately and Efficiently Interpreting Human-Robot Instructions of Varying Granularities
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53131749-ec93-4df1-bf40-04dda5322d80 · outbound
Effective Reward Specification in Deep Reinforcement Learning A General Language Assistant as a Laboratory for Alignment
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13785aa4-9151-4ed9-91d5-cfab30392e51 · outbound
Effective Reward Specification in Deep Reinforcement Learning DynGFN: Towards Bayesian Inference of Gene Regulatory Networks with GFlowNets
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bbb50bcb-4d33-4ddd-8651-a3f00a12f6de · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5fd203ce-372e-46af-ab1d-4e7b8a4e940e · outbound
Effective Reward Specification in Deep Reinforcement Learning Layer Normalization
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de6390a0-20a3-407b-91f7-316ce7f5691b · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b51f118-118b-4bce-9f9c-9d0bd1048d01 · outbound
Effective Reward Specification in Deep Reinforcement Learning Learning to Understand Goal Specifications by Modelling Reward
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1972b0e1-6630-4a0f-bd70-08c3b8ddf945 · outbound
Effective Reward Specification in Deep Reinforcement Learning Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f65dbb71-f6f0-4ef5-91f6-5a2a5fc7896c · outbound
Effective Reward Specification in Deep Reinforcement Learning P., O’malley, M
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b4979f7-dcff-4dff-9d1f-43fb575db73d · outbound
Effective Reward Specification in Deep Reinforcement Learning and Narayanan, S
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa328edd-1c67-4763-b170-851b136f5d51 · outbound
Effective Reward Specification in Deep Reinforcement Learning L., Waytowich, N
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9f9c7dc-196f-4ed8-bad2-3f9bb884c6e4 · outbound
Effective Reward Specification in Deep Reinforcement Learning Graph augmented Deep Reinforcement Learning in the GameRLand3D environment
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 6c883da5-c171-4883-aee1-8d3402a55021 · outbound
Effective Reward Specification in Deep Reinforcement Learning G., Candido, S., Castro, P
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 889c3a95-3e21-460b-834e-2bfb943f18c6 · outbound
Effective Reward Specification in Deep Reinforcement Learning A Distributional Perspective on Reinforcement Learning
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ccca836-b239-4653-986a-01eb009f15c3 · outbound
Effective Reward Specification in Deep Reinforcement Learning G., Naddaf, Y., Veness, J., and Bowling, M
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ee61b4a-6e58-4f43-953c-c0b5a7c81cbf · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d72c8e5e-efd3-4fef-82ba-911ded7c3725 · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c2d210c-b3ce-4eec-b8d6-079a4182122d · outbound
Effective Reward Specification in Deep Reinforcement Learning J., Tiwari, M., and Bengio, E
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04f0d674-0eb1-49b8-9c6b-a0e03efa5b1b · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c08540c-3345-488b-8f3e-b0a45aeda7af · outbound
Effective Reward Specification in Deep Reinforcement Learning Dota 2 with Large Scale Deep Reinforcement Learning
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ddfd02f1-1dd0-47ed-8614-3cccdf70e470 · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46888275-f4a3-4d5b-aac9-e4832d064882 · outbound
Effective Reward Specification in Deep Reinforcement Learning R., Paolini, G
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81cc074e-07d3-4d19-a596-10041e5454f3 · outbound
Effective Reward Specification in Deep Reinforcement Learning Value constrained model-free continuous control
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 9eca8b83-2bda-490f-884b-869f29da442b · outbound
Effective Reward Specification in Deep Reinforcement Learning B., Shah, J., Niekum, S., Stone, P., and Allievi, A
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d16b75b8-bc84-4b85-853c-c0933d4b2282 · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4eb1e295-03c8-4341-8886-743a9bd84f65 · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 992bda25-33e8-4904-a021-b0c9ff3f0139 · outbound
Effective Reward Specification in Deep Reinforcement Learning D., Abel, D., and Dabney, W
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7de053a7-3e23-4bf6-9bf8-26736579e514 · outbound
Effective Reward Specification in Deep Reinforcement Learning OpenAI Gym
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14aa8807-f81d-43a9-be22-00705510bf67 · outbound
Effective Reward Specification in Deep Reinforcement Learning Extrapolating Beyond Suboptimal Demonstrations via Inverse Reinforcement Learning from Observations
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 239d6426-e858-4a33-b371-605d06227559 · outbound
Effective Reward Specification in Deep Reinforcement Learning H., and Vaucher, A
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19bf22df-ddb1-4b32-8e56-758bfff2917a · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b02448d-cd8d-4a4a-a6f3-040bad8b795c · outbound
Effective Reward Specification in Deep Reinforcement Learning Balancing Constraints and Rewards with Meta-Gradient D4PG
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 02ff1b52-1fc7-4ddb-8da8-65787e12fc85 · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ee31b36-4e34-46c4-8b35-7772d11a96eb · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26c6f9e7-cf8c-48d0-825a-abac7f9c29f7 · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0a9aee7-44c5-43f6-b7f5-5097d89ab056 · outbound
Effective Reward Specification in Deep Reinforcement Learning u rnkranz, J., H \
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b879b8c-1080-45df-8598-7506052ce900 · outbound
Effective Reward Specification in Deep Reinforcement Learning G., and Singh, S
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea2188bb-ae4b-4ce4-ad80-d572e4f02603 · outbound
Effective Reward Specification in Deep Reinforcement Learning BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94e48b24-8898-4d09-8247-f5de408aea82 · outbound
Effective Reward Specification in Deep Reinforcement Learning and Kim, K.-E
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7edc8b98-99dc-4e0f-af39-9b949900490a · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a65f529c-6fd3-4e2f-9358-2d9686dcad04 · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation caead2f4-d565-4a16-b601-04a375a763eb · outbound
Effective Reward Specification in Deep Reinforcement Learning Lyapunov-based Safe Policy Optimization for Continuous Control
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44b13125-9c77-41dd-8aa1-5684c4b07b6a · outbound
Effective Reward Specification in Deep Reinforcement Learning F., Leike, J., Brown, T., Martic, M., Legg, S., and Amodei, D
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2758b84a-e19c-47c0-ba0c-8d85921589f9 · outbound
Effective Reward Specification in Deep Reinforcement Learning and Amodei, D
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d750245-79b7-4c5b-83e7-4fb08c060cd9 · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5d4711a-99d6-4167-a660-35c939c1dfed · outbound
Effective Reward Specification in Deep Reinforcement Learning and Niekum, S
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85619297-2c24-467d-9076-c0fadba86258 · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3780a273-9c48-4032-a1fb-c49de9d14ff4 · outbound
Effective Reward Specification in Deep Reinforcement Learning G., and Silver, D
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16dcd7c3-7925-4efb-9112-b146b1a77c9f · outbound
Effective Reward Specification in Deep Reinforcement Learning G., and Munos, R
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9562dfe-e4ee-4a2e-8b28-cb5b2815f8c0 · outbound
Effective Reward Specification in Deep Reinforcement Learning Safe Exploration in Continuous Action Spaces
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fac7ee7e-2ea4-4aea-85ff-8e0a3cfa6ad1 · outbound
Effective Reward Specification in Deep Reinforcement Learning Robotic Table Tennis: A Case Study into a High Speed Learning System
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15297a9f-e93b-4ee4-9b89-b5847006731b · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 001f70bd-cbd3-49e0-916e-f6f2aeda4092 · outbound
Effective Reward Specification in Deep Reinforcement Learning and Dennis, J
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27da0fd3-5053-4011-a8fc-d1c524a851e9 · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3e78f71-fb45-4da2-8549-60d605283617 · outbound
Effective Reward Specification in Deep Reinforcement Learning and Hinton, G
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6481a611-15a4-4696-8bf7-50a41fd79e94 · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2c90d26-5996-498a-b768-75edae709482 · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 840cd487-d96a-4fdf-b43f-492de65b5755 · outbound
Effective Reward Specification in Deep Reinforcement Learning Off-Policy Actor-Critic
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7d33566-2f04-4e2a-bc4b-e8f91f0f4e92 · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f24761a7-93a9-4391-bcbd-ec185d154813 · outbound
Effective Reward Specification in Deep Reinforcement Learning Navigation Turing Test (NTT): Learning to Evaluate Human-Like Navigation
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 9ef83963-391c-4eb3-9626-493e391a1fad · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 119dc798-9155-4d1a-bdd8-c78ebf838a5c · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da1f14d4-daf7-49d3-8e4f-5fa224e7e0f5 · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 86
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d48b8830-6454-4c79-a88c-9a5f82028692 · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b3be87f-bb72-460c-a6e4-7105fdf137d7 · outbound
Effective Reward Specification in Deep Reinforcement Learning Adapting Auxiliary Losses Using Gradient Similarity
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c18a82c-181b-4524-abf1-92d04fec2c01 · outbound
Effective Reward Specification in Deep Reinforcement Learning J., Li, J., Paduraru, C., Gowal, S., and Hester, T
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2f5c98f-554d-4115-8e66-ef28c1c9709f · outbound
Effective Reward Specification in Deep Reinforcement Learning Challenges of Real-World Reinforcement Learning
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c92bf16c-a7b9-4c30-b04e-397ca3f3ed6f · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 91
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99be24b8-21b7-4ab8-ad38-9033571aafd8 · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9c15b42-3b19-488e-b03c-07cbbd1e7e9c · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation faab6a52-19b8-4e91-bb55-212383f004b9 · outbound
Effective Reward Specification in Deep Reinforcement Learning and Schuffenhauer, A
Reference 94
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c2c311c-7bfd-4269-bdc8-f213738306d2 · outbound
Effective Reward Specification in Deep Reinforcement Learning and Gao, J
Reference 95
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9bb42d6f-36d9-44ee-80e2-991f392c92c2 · outbound
Effective Reward Specification in Deep Reinforcement Learning Hyperbolic Discounting and Learning over Multiple Horizons
Reference 96
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e84dc15-c216-4e34-9b02-16f2c88580e5 · outbound
Effective Reward Specification in Deep Reinforcement Learning Bridging the Gap: A Survey on Integrating (Human) Feedback for Natural Language Generation
Reference 97
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3164b18-2f78-43eb-bf76-2278eabe7854 · outbound
Effective Reward Specification in Deep Reinforcement Learning A Connection between Generative Adversarial Networks, Inverse Reinforcement Learning, and Energy-Based Models
Reference 98
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6382932-a9f2-460d-b85a-2396b92285ca · outbound
Effective Reward Specification in Deep Reinforcement Learning Unresolved cited work
Reference 99
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ecfacdc-344f-4fdf-b7a1-031a4d44fca1 · outbound
Effective Reward Specification in Deep Reinforcement Learning A., de Freitas, N., and Whiteson, S
Reference 100
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.