Pith. sign in

Paper Citation Record · LEDGER

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback

As of 21 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 0 inbound Pith citation observations for arXiv:2606.24622.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.24622 v1

Coverage vector

measured 50 of 50 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-25T23:35:03.577967Z

measured 50 of 50 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

50 of 50 outbound references displayed

  • verified exact14
  • verified fuzzy0
  • unresolved35
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e0f1e3e4-9853-43c5-9dd2-10607f5a05d1 · outbound

This paper cites A systematic study on reinforcement learning based applications,.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback A systematic study on reinforcement learning based applications,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:c668dc2a497d7f720cc231189a96b798b3cf6a5ca9877ca09211cec7eeccd00b

Observation 734e6cfc-0ea3-40dd-8d04-f959d0b0cf03 · outbound

This paper cites Deep reinforcement learning for autonomous driving: A survey,.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Deep reinforcement learning for autonomous driving: A survey,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:b8e66f446cec293a97456f119dfc360f4b42a8221019eb94fd75be4dce66feb1

Observation 9acd538f-4867-48f8-bea5-f643018b3c42 · outbound

This paper cites Deep reinforcement learning for robotics: A survey of real-world successes,.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Deep reinforcement learning for robotics: A survey of real-world successes,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:48ed6d7c1e971de18f3b05e22ff0880b8d04bd58b7d74ea27a751336028ce98e

Observation 2d628a25-175e-49c9-a257-e17f252f9ec2 · outbound

This paper cites A review on reinforcement learning: Introduction and applications in industrial process control,.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback A review on reinforcement learning: Introduction and applications in industrial process control,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:2c6c36bfb2d003102e402ebb30954277f2c309699b884aa70cbd76c6e6ee75cd

Observation 6efe7846-c89a-47c8-8aca-83c854f0ce70 · outbound

This paper cites Training language models to follow instructions with human feedback.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Training language models to follow instructions with human feedback

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-07-04T17:40:00.101740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:8eb1e81a4c055f683eceda89603de32f498fc73587cdc6e1e4548cb2f01f6a66

Observation 31770af0-6afb-4646-bfd5-c94705e7c396 · outbound

This paper cites Mastering the game of go with deep neural networks and tree search,.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Mastering the game of go with deep neural networks and tree search,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:e0347a329b90999aeaafe40552fbdf18f8de3ce441bd4dac14677ae85982a8ae

Observation 4952f2d4-39bd-44ed-8b84-6a581194ea21 · outbound

This paper cites Defining and characterizing reward gaming,.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Defining and characterizing reward gaming,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:97778a668a7bf194facdc1c74b9eb358df9c0fd5d1ba09840bcbf860bd46d7c9

Observation 280248a8-afff-4f6e-9910-820c6c6079fa · outbound

This paper cites Reward learning from human preferences and demonstrations in atari,.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Reward learning from human preferences and demonstrations in atari,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:71a0dde7cc5f0c52a31851e2a7e45a56f461fe9a1a1126792aac442513d29b1a

Observation 92295d92-e493-4c23-ac81-aeccc9907af6 · outbound

This paper cites Deep reinforcement learning from human preferences,.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Deep reinforcement learning from human preferences,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:b32577af276decdca423c3c663e76ac6b1f1c13f807e697eba8a3a3c5a76b1ea

Observation 0d9726e2-5ace-487e-b7cb-c481e0309617 · outbound

This paper cites Human-in-the-Loop Deep Reinforcement Learning with Application to Autonomous Driving.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Human-in-the-Loop Deep Reinforcement Learning with Application to Autonomous Driving

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-07-04T17:40:00.119855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:9b1526c2903209a5203ec8bb21abf5cd512007d6e8a90a1fb4fdb4e8339c54a8

Observation 6cacb8aa-fded-42e3-86bd-3385714149d0 · outbound

This paper cites The utility of explainable ai in ad hoc human-machine teaming,.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback The utility of explainable ai in ad hoc human-machine teaming,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:dd11d2593a0abb87ea43c087fa557f07c188cdf25a52e5fd0247c303b4626960

Observation c2818fa3-3989-4636-9789-3e64e4039060 · outbound

This paper cites Open Problems and Fundamental Limitations of Reinforcement Learning from Human Feedback.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Open Problems and Fundamental Limitations of Reinforcement Learning from Human Feedback

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-07-04T17:40:00.117386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:9a94a6fb3fbb8e2d9e485b892cdcfa35ab1f778ae3b467a2314028a8a6a4f208

Observation 03251450-8d6d-4743-aa7b-840cacc514ce · outbound

This paper cites an unresolved cited work.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:724bc7018e3a46080953c9e508da4e8d93c203c171a05e960e719979f97d8409

Observation 0adaf2d5-9fa4-408c-85c5-0a469256b7f1 · outbound

This paper cites An overview of the action space for deep reinforcement learning,.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback An overview of the action space for deep reinforcement learning,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:1f049066469deb3d002b7665fe517be95ac4e4c08d5b3df3ddbeb697f01f2c3a

Observation f690b0a8-1cbb-4ea6-93f7-f207bb2ba447 · outbound

This paper cites Human-level control through deep reinforcement learning,.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Human-level control through deep reinforcement learning,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:33631d07a28692a78b49dca980df4f7fa1fe50d5fcc76ae2302558c262da46fc

Observation e084eeb8-84d0-440e-a3f7-582b3bc72455 · outbound

This paper cites Apprenticeship learning via inverse rein- forcement learning,.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Apprenticeship learning via inverse rein- forcement learning,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:2bc17f79f980c21079397b66d184946538fe64acc8e5a0d67e18722115571464

Observation f2cc9e3d-f69f-4c05-827e-a9aaead82ca6 · outbound

This paper cites DQN-TAMER: Human-in-the-Loop Reinforcement Learning with Intractable Feedback.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback DQN-TAMER: Human-in-the-Loop Reinforcement Learning with Intractable Feedback

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-07-04T17:40:00.114755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:4a5df6c15a184a142974d436733de0e69a11845f190b868bac97d33c4ba646e8

Observation 7127ee49-baaf-4be9-99f8-d999492dfd71 · outbound

This paper cites PEBBLE: Feedback-Efficient Interactive Reinforcement Learning via Relabeling Experience and Unsupervised Pre-training.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback PEBBLE: Feedback-Efficient Interactive Reinforcement Learning via Relabeling Experience and Unsupervised Pre-training

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-07-04T17:40:00.090762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:4cee3e300a692aa5161ddea2c41b8461c594ec36b45418e644296a56104a5946

Observation a642df2a-df81-4add-97a5-b3b176611927 · outbound

This paper cites SURF: Semi-supervised Reward Learning with Data Augmentation for Feedback-efficient Preference-based Reinforcement Learning.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback SURF: Semi-supervised Reward Learning with Data Augmentation for Feedback-efficient Preference-based Reinforcement Learning

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-07-04T17:40:00.087977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:ee51bf05934a1acd184124ea1f4edcf32a427f0c5c23e6fdadeb058d46e3e165

Observation efd56d09-8f33-48f8-8e89-583583adf205 · outbound

This paper cites Weak Human Preference Supervision For Deep Reinforcement Learning.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Weak Human Preference Supervision For Deep Reinforcement Learning

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-04T17:40:00.113955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:2e1c58713b92312e2d2fd7d32fac7b465d82b245b9c4abbdc46df8d2c79d5026

Observation 93605ef0-3991-445f-936e-20a5ce3b29ad · outbound

This paper cites A survey on interactive reinforcement learning: Design principles and open challenges,.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback A survey on interactive reinforcement learning: Design principles and open challenges,

Reference 21

Resolution
verified exact
doi, observed 2026-06-25T23:38:41.377751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:c66b6a57366190dae8499bd77d18787038d00d043630848697ffb5807491cd56

Observation ec1b64f0-6d15-443e-9788-450974ae89f8 · outbound

This paper cites Leveraging human guidance for deep reinforcement learning tasks,.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Leveraging human guidance for deep reinforcement learning tasks,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:a439fc484165ff1821451dc33b353ef175fa0f24bb39cd50ce628c238c0980f1

Observation f8b4feff-9a10-4b08-80c2-3c5e5217bf8d · outbound

This paper cites Knowledge-based causal attribution: The abnormal conditions focus model.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Knowledge-based causal attribution: The abnormal conditions focus model

Reference 23

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:be0e78abe335b269576ff66a4b27dba2dfec0643d80042a2a57c075551d2e589

Observation 4ccbeded-bafc-436c-9e0d-70cddb4a0302 · outbound

This paper cites Collective ex- plainable ai: Explaining cooperative strategies and agent contribution in multiagent reinforcement learning with shapley values,.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Collective ex- plainable ai: Explaining cooperative strategies and agent contribution in multiagent reinforcement learning with shapley values,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:fbb6cc7ed950dd5623ff0c279281f5b217aadfb49dbdd5c8920d024ee167496c

Observation fe01a410-39aa-4112-826c-709aa91356df · outbound

This paper cites Visualizing and understanding atari agents,.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Visualizing and understanding atari agents,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:c056efd46c7754e3884c279a83032e2e73900b1c2919ca006fa5028005255fff

Observation cce16177-4309-4f79-b2b7-f9a3b6c38933 · outbound

This paper cites A Survey on Explainable Reinforcement Learning: Concepts, Algorithms, Challenges.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback A Survey on Explainable Reinforcement Learning: Concepts, Algorithms, Challenges

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-07-04T17:40:00.100803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:98c07453f553c279de2b7ac645e83bcd75e421bf57058d9312fe1f71ee20a8a3

Observation 2811f7cd-e373-4b7e-a6c9-26ea84c42aa3 · outbound

This paper cites Explainable deep reinforcement learning: state of the art and challenges,.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Explainable deep reinforcement learning: state of the art and challenges,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:aa95e59fa881a748ad81d9ffc5efa0bf443a6a33a26608483afee062f5da3b55

Observation 6fab2a8d-de65-438c-a29a-25aefb881173 · outbound

This paper cites B-pref: Benchmarking preference-based reinforcement learning,.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback B-pref: Benchmarking preference-based reinforcement learning,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:fb26bb830f1924aeb423057b9fa976572f9643b5291daaa9a5378cdea6956945

Observation dbc08dbc-28a6-4977-8501-26b3482094de · outbound

This paper cites Hydra - a framework for elegantly configuring complex applications,.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Hydra - a framework for elegantly configuring complex applications,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:913d3b722bb308094039f265e09b5b3d7d5ec417c8cc7501fac6c8ba31083f90

Observation d707d3f9-0055-4f6e-9058-f40c3ef0f269 · outbound

This paper cites Kazuma Tsuji, Ken’ichiro Tanaka, and Sebastian Pokutta.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Kazuma Tsuji, Ken’ichiro Tanaka, and Sebastian Pokutta

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-07-04T17:40:00.104551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:2780da382fe7be9696183f58963f4a03310543c4cef7293732c3891e633aecbf

Observation 8cc002af-4a8a-4e1d-96da-5b2b677bff84 · outbound

This paper cites Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-07-04T17:40:00.106038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:697c052cdcf44853de2b25e7a8c4fb214249670e05799ef415d59dc0a45641b8

Observation 6f805bff-4e57-42bb-9448-567f431b30cd · outbound

This paper cites BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T17:40:00.109677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:39030a24441e75d9d44f2e5faa410a0939e9b2ee95bf6dedd0633fdd1e853482

Observation 2db9a9f8-8bc1-4eb4-bab4-75c3dacec16a · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor,.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:083df3b2b80d4ba167c6cf637092bf2d8f4f7d4eec10f9bb6c27ff324a561dc4

Observation 4645bb36-af0f-4a25-adc7-772fe0386020 · outbound

This paper cites Soft actor-critic for discrete action settings,.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Soft actor-critic for discrete action settings,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:da5c6b4306b632e754dcc0b637b16511bcce43c38e1dcc9691c9e6e3207e3893

Observation 3b2466e7-060b-4d5b-bf60-7bc7bc4210bc · outbound

This paper cites Rank analysis of incomplete block designs: I. the method of paired comparisons,.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Rank analysis of incomplete block designs: I. the method of paired comparisons,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:8fdd4386a538dfea6b397f553c3d79d266568d6b60765b361782ca36f3316676

Observation ad534517-0e90-4143-b268-8e1079188333 · outbound

This paper cites Maximizing the efficiency of human feedback in AI alignment: a comparative analysis.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Maximizing the efficiency of human feedback in AI alignment: a comparative analysis

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-07-04T17:40:00.112152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:a310d37c6f3b9409bab9ec391156801dbdbf4acdbb4ffc07581b413da6ba1f59

Observation dea1463d-b4fc-4386-b456-349acd775110 · outbound

This paper cites Local and global explanations of agent behavior: Integrating strategy summaries with saliency maps,.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Local and global explanations of agent behavior: Integrating strategy summaries with saliency maps,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:9dff37c97ae304ef43a913097263b97b619723e5a644bfe8d10c91a1b8c541e0

Observation ba013a27-0995-4bf7-b32c-7346e362e45b · outbound

This paper cites Captum: A unified and generic model inter- pretability library for pytorch,.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Captum: A unified and generic model inter- pretability library for pytorch,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:db9731f49d8faa862b09e590710d45fe60a910c09a8bda76ef81d24277f1ca68

Observation acb14a84-dde6-47f6-8ed8-0f38f1842cf2 · outbound

This paper cites Axiomatic attribution for deep networks,.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Axiomatic attribution for deep networks,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:9f47597d5a94728e7e9df854af585b12c9ec510a47161398013254bc794f9b71

Observation ad34cda8-f031-49ff-b032-95f9dd109e22 · outbound

This paper cites A unified approach to interpreting model predictions,.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback A unified approach to interpreting model predictions,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:74e2c5b758339416714c3aa74227637f5f4fee99871e06c13baefe33c8208fbf

Observation 186fdd71-0186-4bef-a834-71b5b27b9368 · outbound

This paper cites Estimating training data influence by tracing gradient descent,.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Estimating training data influence by tracing gradient descent,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:5526f18422ea0dee500cc4804cbfef74e9e4447d93918fd2c4b47aab80453bd6

Observation 9dae72b6-d693-4986-9033-68d4bf6187fa · outbound

This paper cites Mongodb: The developer data platform,.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Mongodb: The developer data platform,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:9bbdbeeed2d5878de726881c87f52a7cc8ebb44620f3b8f763e4eb4e96fef776

Observation 31a52f26-b387-4667-8225-b8d75c936d7f · outbound

This paper cites Next.js: The react framework,.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Next.js: The react framework,

Reference 43

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:e2812f23189fa05aa917cf317ece3349591223a5432849d155db8b33108a9b0f

Observation 63355d77-2d82-46f6-89cb-97087e1caa72 · outbound

This paper cites The arcade learning environment: An evaluation platform for general agents,.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback The arcade learning environment: An evaluation platform for general agents,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:cbf96636881c2c88b015c77f3a1a51e132c1fcee6b7d09ef8a3eb7f667afccf5

Observation 19fb7d8c-1937-4834-9ed0-1bd446b4c1bb · outbound

This paper cites Loadster: A load testing & website stress testing tool,.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Loadster: A load testing & website stress testing tool,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:989c1ac836a3dca385a362d7fed945f962cc874ebc1bd945d6c597bf1b9e4cb4

Observation 0b336d9d-fd91-44f3-8ad7-21708d016557 · outbound

This paper cites Available: https://loadster.app/.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Available: https://loadster.app/

Reference 46

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:3f6582c5cd7acabaa7f23f210a92d5b7d9394b919df3886f5ce5f168da425881

Observation df27150e-5b92-42eb-9c7a-1abab6b7f803 · outbound

This paper cites Mujoco: A physics engine for model-based control,.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Mujoco: A physics engine for model-based control,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:fca29d2282b08e54b056d86fc0fbcfe6b33233526478ae79baac35d5ca5ec00a

Observation 7e2cb0ca-a83e-4731-846b-dbfca2b34261 · outbound

This paper cites Trust Region Policy Optimization.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Trust Region Policy Optimization

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-07-04T17:40:00.096607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:1479fd6b32cfa1ce9a92cf91d6196c0a02732793af8e73cd112829b6aa9dd6ec

Observation bf0c0586-61db-445b-80db-e3ebf182fe08 · outbound

This paper cites Asynchronous Methods for Deep Reinforcement Learning.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Asynchronous Methods for Deep Reinforcement Learning

Reference 49

Resolution
verified exact
local_arxiv, observed 2026-07-04T17:40:00.108311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:375d9aa8f93c30a6fa8e3283ecd13004fe41a12d3f1265be885eba98f732f4e6

Observation 21ac579f-dc02-4006-a8ad-b142f77816aa · outbound

This paper cites Widening the pipeline in human-guided reinforcement learning with explanation and context-aware data augmentation,.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Widening the pipeline in human-guided reinforcement learning with explanation and context-aware data augmentation,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-06-25T23:35:03.577967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:db21cc3e49bf1e1fa25c87e95a0cc6fefcd292330d2453f336d641ddb8acf19b

Pith citing papers

No inbound Pith citation observations are available.