Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2310.00166.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T11:36:11.387527Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T20:48:55.285309Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation b96f90a6-1417-4d05-a595-cd4b91c98aa0 · inbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Motif: Intrinsic Motivation from Artificial Intelligence Feedback
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed1a5933-e60f-4642-9cad-c860934eda93 · inbound
BALROG: Benchmarking Agentic LLM and VLM Reasoning On Games Motif: Intrinsic Motivation from Artificial Intelligence Feedback
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6629c4af-c145-4cfa-a771-f0e5e1be95dc · inbound
Probing for Consciousness in Machines Motif: Intrinsic Motivation from Artificial Intelligence Feedback
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85820567-6bd6-4198-a2bc-470379680d37 · inbound
Effective Reward Specification in Deep Reinforcement Learning Motif: Intrinsic Motivation from Artificial Intelligence Feedback
Reference 165
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce1ce659-408c-478b-b279-083d16f8f96f · inbound
A Self-Improving Coding Agent Motif: Intrinsic Motivation from Artificial Intelligence Feedback
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b73f263-94f0-46dd-a90f-7947a64dbdd7 · inbound
LLMs are Greedy Agents: Effects of RL Fine-tuning on Decision-Making Abilities Motif: Intrinsic Motivation from Artificial Intelligence Feedback
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e49f440-9299-4c4c-aa45-57d9ed67b91b · inbound
PRISM: Projection-based Reward Integration for Scene-Aware Real-to-Sim-to-Real Transfer with Few Demonstrations Motif: Intrinsic Motivation from Artificial Intelligence Feedback
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d09e3dd3-6de9-48af-818e-5448436fe3e6 · inbound
TREND: Tri-teaching for Robust Preference-based Reinforcement Learning with Demonstrations Motif: Intrinsic Motivation from Artificial Intelligence Feedback
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46f564a4-a23e-4fbd-a944-409edba38e42 · inbound
Enhancing Rating-Based Reinforcement Learning to Effectively Leverage Feedback from Large Vision-Language Models Motif: Intrinsic Motivation from Artificial Intelligence Feedback
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1be62d4a-6672-4734-b03a-985e95e42ca8 · inbound
Reward Models in Deep Reinforcement Learning: A Survey Motif: Intrinsic Motivation from Artificial Intelligence Feedback
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4eeb02b1-76de-49a5-af09-5122d2f86184 · inbound
FOUNDER: Grounding Foundation Models in World Models for Open-Ended Embodied Decision Making Motif: Intrinsic Motivation from Artificial Intelligence Feedback
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 395ebe7e-bb63-45b4-bdfe-8a918f69e06b · inbound
LLM Economist: Large Population Models and Mechanism Design in Multi-Agent Generative Simulacra Motif: Intrinsic Motivation from Artificial Intelligence Feedback
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2632fedb-e8d5-4038-9601-fb61ee64f41b · inbound
Timing the Message: Language-Based Notifications for Time-Critical Assistive Settings Motif: Intrinsic Motivation from Artificial Intelligence Feedback
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99071007-bad2-430b-9938-23976410309c · inbound
Hierarchical Behaviour Spaces Motif: Intrinsic Motivation from Artificial Intelligence Feedback
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 7c38ead9-20fb-4471-8c24-b690bc4e6f19 · inbound
Agentick: A Unified Benchmark for General Sequential Decision-Making Agents Motif: Intrinsic Motivation from Artificial Intelligence Feedback
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 2e722bf5-783c-4481-9561-1108d5eb2fa6 · inbound
Goal-Conditioned Agents that Learn Everything All at Once Motif: Intrinsic Motivation from Artificial Intelligence Feedback
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e4a6465b-583c-466b-be0e-1c2f9121b2f2 · inbound
VLM-AR3L: Vision-Language Models for Absolute and Relative Rewards in Reinforcement Learning Motif: Intrinsic Motivation from Artificial Intelligence Feedback
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 471e7edd-6813-4514-9ae8-69eaff48942b · inbound
VLM-AR3L: Vision-Language Models for Absolute and Relative Rewards in Reinforcement Learning Motif: Intrinsic Motivation from Artificial Intelligence Feedback
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 7d1ad87f-dcd3-44f3-9305-a9b5f4ae0224 · inbound
Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills Motif: Intrinsic Motivation from Artificial Intelligence Feedback
Reference 121
Source-reported events for the cited work
Unavailable: canonical work link unavailable.