Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-25T23:35:03.577967Z
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 0 inbound Pith citation observations for arXiv:2606.24622.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-25T23:35:03.577967Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
50 of 50 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation e0f1e3e4-9853-43c5-9dd2-10607f5a05d1 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback A systematic study on reinforcement learning based applications,
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 734e6cfc-0ea3-40dd-8d04-f959d0b0cf03 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Deep reinforcement learning for autonomous driving: A survey,
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9acd538f-4867-48f8-bea5-f643018b3c42 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Deep reinforcement learning for robotics: A survey of real-world successes,
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d628a25-175e-49c9-a257-e17f252f9ec2 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback A review on reinforcement learning: Introduction and applications in industrial process control,
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6efe7846-c89a-47c8-8aca-83c854f0ce70 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Training language models to follow instructions with human feedback
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 31770af0-6afb-4646-bfd5-c94705e7c396 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Mastering the game of go with deep neural networks and tree search,
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4952f2d4-39bd-44ed-8b84-6a581194ea21 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Defining and characterizing reward gaming,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 280248a8-afff-4f6e-9910-820c6c6079fa · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Reward learning from human preferences and demonstrations in atari,
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92295d92-e493-4c23-ac81-aeccc9907af6 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Deep reinforcement learning from human preferences,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d9726e2-5ace-487e-b7cb-c481e0309617 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Human-in-the-Loop Deep Reinforcement Learning with Application to Autonomous Driving
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 6cacb8aa-fded-42e3-86bd-3385714149d0 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback The utility of explainable ai in ad hoc human-machine teaming,
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2818fa3-3989-4636-9789-3e64e4039060 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Open Problems and Fundamental Limitations of Reinforcement Learning from Human Feedback
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 03251450-8d6d-4743-aa7b-840cacc514ce · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Unresolved cited work
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0adaf2d5-9fa4-408c-85c5-0a469256b7f1 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback An overview of the action space for deep reinforcement learning,
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f690b0a8-1cbb-4ea6-93f7-f207bb2ba447 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Human-level control through deep reinforcement learning,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e084eeb8-84d0-440e-a3f7-582b3bc72455 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Apprenticeship learning via inverse rein- forcement learning,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2cc9e3d-f69f-4c05-827e-a9aaead82ca6 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback DQN-TAMER: Human-in-the-Loop Reinforcement Learning with Intractable Feedback
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 7127ee49-baaf-4be9-99f8-d999492dfd71 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback PEBBLE: Feedback-Efficient Interactive Reinforcement Learning via Relabeling Experience and Unsupervised Pre-training
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation a642df2a-df81-4add-97a5-b3b176611927 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback SURF: Semi-supervised Reward Learning with Data Augmentation for Feedback-efficient Preference-based Reinforcement Learning
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation efd56d09-8f33-48f8-8e89-583583adf205 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Weak Human Preference Supervision For Deep Reinforcement Learning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 93605ef0-3991-445f-936e-20a5ce3b29ad · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback A survey on interactive reinforcement learning: Design principles and open challenges,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation ec1b64f0-6d15-443e-9788-450974ae89f8 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Leveraging human guidance for deep reinforcement learning tasks,
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8b4feff-9a10-4b08-80c2-3c5e5217bf8d · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Knowledge-based causal attribution: The abnormal conditions focus model
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ccbeded-bafc-436c-9e0d-70cddb4a0302 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Collective ex- plainable ai: Explaining cooperative strategies and agent contribution in multiagent reinforcement learning with shapley values,
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe01a410-39aa-4112-826c-709aa91356df · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Visualizing and understanding atari agents,
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cce16177-4309-4f79-b2b7-f9a3b6c38933 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback A Survey on Explainable Reinforcement Learning: Concepts, Algorithms, Challenges
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 2811f7cd-e373-4b7e-a6c9-26ea84c42aa3 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Explainable deep reinforcement learning: state of the art and challenges,
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6fab2a8d-de65-438c-a29a-25aefb881173 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback B-pref: Benchmarking preference-based reinforcement learning,
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dbc08dbc-28a6-4977-8501-26b3482094de · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Hydra - a framework for elegantly configuring complex applications,
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d707d3f9-0055-4f6e-9058-f40c3ef0f269 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Kazuma Tsuji, Ken’ichiro Tanaka, and Sebastian Pokutta
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 8cc002af-4a8a-4e1d-96da-5b2b677bff84 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 6f805bff-4e57-42bb-9448-567f431b30cd · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 2db9a9f8-8bc1-4eb4-bab4-75c3dacec16a · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor,
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4645bb36-af0f-4a25-adc7-772fe0386020 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Soft actor-critic for discrete action settings,
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b2466e7-060b-4d5b-bf60-7bc7bc4210bc · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Rank analysis of incomplete block designs: I. the method of paired comparisons,
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad534517-0e90-4143-b268-8e1079188333 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Maximizing the efficiency of human feedback in AI alignment: a comparative analysis
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation dea1463d-b4fc-4386-b456-349acd775110 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Local and global explanations of agent behavior: Integrating strategy summaries with saliency maps,
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba013a27-0995-4bf7-b32c-7346e362e45b · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Captum: A unified and generic model inter- pretability library for pytorch,
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation acb14a84-dde6-47f6-8ed8-0f38f1842cf2 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Axiomatic attribution for deep networks,
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad34cda8-f031-49ff-b032-95f9dd109e22 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback A unified approach to interpreting model predictions,
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 186fdd71-0186-4bef-a834-71b5b27b9368 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Estimating training data influence by tracing gradient descent,
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9dae72b6-d693-4986-9033-68d4bf6187fa · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Mongodb: The developer data platform,
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31a52f26-b387-4667-8225-b8d75c936d7f · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Next.js: The react framework,
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63355d77-2d82-46f6-89cb-97087e1caa72 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback The arcade learning environment: An evaluation platform for general agents,
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19fb7d8c-1937-4834-9ed0-1bd446b4c1bb · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Loadster: A load testing & website stress testing tool,
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b336d9d-fd91-44f3-8ad7-21708d016557 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Available: https://loadster.app/
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df27150e-5b92-42eb-9c7a-1abab6b7f803 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Mujoco: A physics engine for model-based control,
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e2cb0ca-a83e-4731-846b-dbfca2b34261 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Trust Region Policy Optimization
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation bf0c0586-61db-445b-80db-e3ebf182fe08 · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Asynchronous Methods for Deep Reinforcement Learning
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 21ac579f-dc02-4006-a8ad-b142f77816aa · outbound
Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Widening the pipeline in human-guided reinforcement learning with explanation and context-aware data augmentation,
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.