Pith. sign in

Paper Citation Record · LEDGER

Imitation Learning via Focused Satisficing

As of 19 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 0 inbound Pith citation observations for arXiv:2505.14820.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.14820 v2

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:35:36.141307Z

measured 35 of 35 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

35 of 35 outbound references displayed

  • verified exact2
  • verified fuzzy26
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 89718f58-42b5-4b34-94c5-a3ae893b2ca9 · outbound

This paper cites an unresolved cited work.

Imitation Learning via Focused Satisficing Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:35:39.568926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:35:33.586848Z digest=sha256:8fe21fafbcb3685e37cd9c90bc19a24a76f6538b9efe3aa92275f2404c1ae298

Observation b8529d1b-38a4-473c-b962-bf873ebb59cd · outbound

This paper cites (11) where α(j) k = 1 fk(ξ)−fk( ˜ξ(j)) is the hinge slope that makes demonstration ˜ξ(j) exactly where the subdominance becomes zero.

Imitation Learning via Focused Satisficing (11) where α(j) k = 1 fk(ξ)−fk( ˜ξ(j)) is the hinge slope that makes demonstration ˜ξ(j) exactly where the subdominance becomes zero

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:35:37.000313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:35:35.994262Z digest=sha256:924142a55cab3f821f654fabff6414e6a3dbc0a4ff57d6d641deefdd5d7afdaf

Observation 567eb4dc-46ef-4eda-a3ae-c704e96e1521 · outbound

This paper cites Extrapolating beyond sub- optimal demonstrations via inverse reinforcement learning from observations.

Imitation Learning via Focused Satisficing Extrapolating beyond sub- optimal demonstrations via inverse reinforcement learning from observations

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:35:39.529854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:35:34.069702Z digest=sha256:4b943cfebaf363b5e2a58d75cab7034f4097965697ec1452afcdba57784c0a7a

Observation 80d3df44-a186-442c-b4ca-f83d0de777e4 · outbound

This paper cites an unresolved cited work.

Imitation Learning via Focused Satisficing Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:35:36.837938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:35:36.042227Z digest=sha256:e2ae855c244b59acf84a74dd9533c244d259df8fb861d0c00a8fcfcd44353821

Observation c42fc5e4-aa07-442b-b18b-e99483566b8a · outbound

This paper cites Learning from Suboptimal Demonstration via Self-Supervised Reward Regression.

Imitation Learning via Focused Satisficing Learning from Suboptimal Demonstration via Self-Supervised Reward Regression

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:35:36.506393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:35:34.323405Z digest=sha256:a29090aaabe853abebabe1f03078cba98dcb475db325f813056dc57e88bd346b

Observation 0aefb1ac-c840-4de0-b9f1-c74e2a912cbd · outbound

This paper cites Deep reinforcement learning from human preferences.Ad- vances in Neural Information Processing Systems , 30,.

Imitation Learning via Focused Satisficing Deep reinforcement learning from human preferences.Ad- vances in Neural Information Processing Systems , 30,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:35:39.463498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:35:34.392724Z digest=sha256:32cd6a366f187688846b5614ad1392000e6db7a8bf3faea74c75b48aa99c45d7

Observation b1e1749e-267b-46a1-9646-f8560bd66cff · outbound

This paper cites an unresolved cited work.

Imitation Learning via Focused Satisficing Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:35:39.273548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:35:34.575158Z digest=sha256:93ea980e12489d8848f369caacb6007948af45fe635093a2493f51389f522f0e

Observation 0949aa94-09fb-4491-9b86-1b12fb6d3502 · outbound

This paper cites Superhuman fairness.

Imitation Learning via Focused Satisficing Superhuman fairness

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:35:39.006266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:35:34.887251Z digest=sha256:4108ef5d9ed1d9af576bf71aa2d66ca5a5ddc48a0ea1867a4caa5309c180e643

Observation 4916cc7e-3ec7-457e-83f1-55342638685c · outbound

This paper cites The magical number seven, plus or minus two: Some limits on our capacity for pro- cessing information.

Imitation Learning via Focused Satisficing The magical number seven, plus or minus two: Some limits on our capacity for pro- cessing information

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:35:38.852336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:35:34.922402Z digest=sha256:52bf102a97061ee0f139faa950612d712b123e1473c7ff82d44c4660b3b7656e

Observation 18318c54-1983-4868-8572-15e048cf9704 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Imitation Learning via Focused Satisficing Proximal Policy Optimization Algorithms

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:35.319508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:35:35.319508Z digest=sha256:0aaf86aba70845277f748b220257839c07478fba5aa6e2d472f4a0c73012d9e9

Observation 9e24915e-2ea6-4a85-8970-1c1f9b214ae1 · outbound

This paper cites Rational choice and the structure of the environment.

Imitation Learning via Focused Satisficing Rational choice and the structure of the environment

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:35:38.154925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:35:35.385964Z digest=sha256:0b59e0b1641931c2b1f133bbb9ab4dfd9d9cc412edd779a1521b7939ce8844e6

Observation f1efaa88-2a2c-449d-9fc9-c4dbe255a07a · outbound

This paper cites Policy gradient methods for reinforcement learning with function approx- imation.

Imitation Learning via Focused Satisficing Policy gradient methods for reinforcement learning with function approx- imation

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:35:38.061670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:35:35.430377Z digest=sha256:47d6aaec8072ea88b6cab39933643c2d4c7bf507e2bdc5d678947fba486db6ca

Observation 1672cabb-1ffb-469d-acea-42e35c33b066 · outbound

This paper cites Bounds on error expectation for support vector machines.

Imitation Learning via Focused Satisficing Bounds on error expectation for support vector machines

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:35:37.669355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:35:35.627625Z digest=sha256:c0f62b26e3570752e7de4113a2406a87ce0307758317f46913a37bcc29559327

Observation 10c490d2-fb9e-4d1d-9d7e-5145c5c12264 · outbound

This paper cites Effective analysis of reac- tion time data.

Imitation Learning via Focused Satisficing Effective analysis of reac- tion time data

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:35:37.512584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:35:35.689680Z digest=sha256:0a77e4ac59f755dfe73a4f6df9883e49cf6a3428dfd7998b9b5a48210d1bb13e

Observation 16d40326-967d-4cf0-b49c-4b2b2396107c · outbound

This paper cites Imitation learning from imperfect demonstra- tion.

Imitation Learning via Focused Satisficing Imitation learning from imperfect demonstra- tion

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:35:37.275972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:35:35.792257Z digest=sha256:6c62f0943c3688876254f1230a7cc7d4402c0585afa05581d8d4bc5c2e3db3e8

Observation ce191812-b140-4123-873b-26bfd2ca04b7 · outbound

This paper cites Confidence-Aware Imitation Learning from Demonstrations with Varying Optimality.

Imitation Learning via Focused Satisficing Confidence-Aware Imitation Learning from Demonstrations with Varying Optimality

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:35:36.315940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:35:35.864821Z digest=sha256:43e973574d056697b26c91c5a8ecae82d4f34408d42cec594cb83878b3e3b67f

Observation 6f06fde6-2ef1-489f-ac1e-37ac0d3c89a2 · outbound

This paper cites Ziebart, Sanjiban Choudhury, Xinyan Yan, and Paul Vernaza.

Imitation Learning via Focused Satisficing Ziebart, Sanjiban Choudhury, Xinyan Yan, and Paul Vernaza

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:35:37.111038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:35:35.955448Z digest=sha256:461f1ee3c565640d3ef7d0392cb22f44140cd81c0029e059cc1d6debee8256d3

Observation 30bbfac7-c60f-4310-9077-6ee7ed4c0217 · outbound

This paper cites This toy examples gives us a peek into the source of this degeneracy.

Imitation Learning via Focused Satisficing This toy examples gives us a peek into the source of this degeneracy

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:35:36.623491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:35:36.141307Z digest=sha256:7f94978f01fd8fd070fe04475447f846d6df1c35742eb72aa323b10a29779e4e

Observation 642a20e8-45f4-4375-83c5-dfceb3bae3c2 · outbound

This paper cites Openai gym,.

Imitation Learning via Focused Satisficing Openai gym,

Reference 1952

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:35:39.539257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:35:33.995624Z digest=sha256:f081c1d8730915c8e4e8e07ff462826a4776e0b7438e0e012cb4d41990f719d6

Observation 0d4d7fbd-d284-4517-9baf-3c8f4726b1a2 · outbound

This paper cites Algorithms for inverse reinforcement learning.

Imitation Learning via Focused Satisficing Algorithms for inverse reinforcement learning

Reference 1956

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:35:38.701008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:35:34.961395Z digest=sha256:2f3f7d721d1dc967990123a444154cc9f28f02a1023e5fc42d48285ebe0e6245

Observation 671cb537-3c3d-4e11-a207-10562b30f609 · outbound

This paper cites Advantages of robotic assistance over a manual approach in simulated subretinal injections and its relevance for gene therapy.

Imitation Learning via Focused Satisficing Advantages of robotic assistance over a manual approach in simulated subretinal injections and its relevance for gene therapy

Reference 1964

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:35:39.172429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:35:34.681688Z digest=sha256:c41308a00947b2b526fab5e8ab0965f1ae7b31faec5a5495872b189cd59613a8

Observation 0ab965ba-710a-47e6-b237-3b72cf175061 · outbound

This paper cites Direct preference optimization: Your lan- guage model is secretly a reward model.Advances in Neu- ral Information Processing Systems, 36,.

Imitation Learning via Focused Satisficing Direct preference optimization: Your lan- guage model is secretly a reward model.Advances in Neu- ral Information Processing Systems, 36,

Reference 1991

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:35:38.455155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:35:35.178015Z digest=sha256:0ee050831836e563c87b033d86eeae6a343d41a59c8e064b02ee709ed2538e3a

Observation 6282de34-0edc-4f95-9407-fa405ce5b3e9 · outbound

This paper cites A game-theoretic approach to apprenticeship learning.

Imitation Learning via Focused Satisficing A game-theoretic approach to apprenticeship learning

Reference 1999

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:35:37.901702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:35:35.463615Z digest=sha256:da108935be89e9df79acbe9e33bc91d7b2fbc14e1965fa323e5ccc35880abdfa

Observation d305f85c-9281-4f78-bc05-001d56c2d282 · outbound

This paper cites An Algorithmic Perspective on Imitation Learning.

Imitation Learning via Focused Satisficing An Algorithmic Perspective on Imitation Learning

Reference 2000

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:35.024041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:35:35.024041Z digest=sha256:58464e78134a3dccf86dda3fb5275c67cfe29ad75e8d4c31d6a901a7c5258280

Observation 762903c9-13cc-46c7-b187-8de77b6eea95 · outbound

This paper cites Concrete Problems in AI Safety.

Imitation Learning via Focused Satisficing Concrete Problems in AI Safety

Reference 2004

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:33.685232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:35:33.685232Z digest=sha256:e1d8622357096f3aa486997b0a83db78d42a3d0ede16d6c783fbfacf1ee1fd85

Observation c34949cc-d14b-4b3d-a549-cebc43e89063 · outbound

This paper cites Robust imitation learning from noisy demonstrations.

Imitation Learning via Focused Satisficing Robust imitation learning from noisy demonstrations

Reference 2007

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:35:37.752139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:35:35.567714Z digest=sha256:c7edd3be1ec6043e4c927646a5789758cfe2268c11beefdefe8b3a740a88d0f7

Observation 8b910c09-371c-47ce-8d93-3f98899132e8 · outbound

This paper cites A survey of preference-based reinforcement learning methods.

Imitation Learning via Focused Satisficing A survey of preference-based reinforcement learning methods

Reference 2008

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:35:37.394464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:35:35.743587Z digest=sha256:ec6674f28d6e6020e6af9f5c64d434880662ea17f1ec8b73a5d904f6681c9f68

Observation c437e437-e3eb-4c32-9b1b-fd886c24c99d · outbound

This paper cites Pitfalls of learning a reward function online.

Imitation Learning via Focused Satisficing Pitfalls of learning a reward function online

Reference 2016

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:35:39.559053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:35:33.823094Z digest=sha256:8f6adcff7471a234c6a477556fbe12dc56aa3a98badc9b28e60e2e5e0f777ff9

Observation f4e82e5f-f619-4c75-978f-4a1cc2924837 · outbound

This paper cites Learning robust rewards with adverserial inverse rein- forcement learning.

Imitation Learning via Focused Satisficing Learning robust rewards with adverserial inverse rein- forcement learning

Reference 2017

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:35:39.372211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:35:34.478420Z digest=sha256:e96a4e6ab5b7bcc791ff56939b1afb394d36153b145b5aa23fcc97307452cb95

Observation ce59d8ef-055c-4a50-b2c8-db239f1e7664 · outbound

This paper cites Pomerleau.

Imitation Learning via Focused Satisficing Pomerleau

Reference 2018

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:35:38.560191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:35:35.079168Z digest=sha256:a59a8a34313efba94efef8609395525644caedaef74202bebd0bcfe3daaaf453

Observation d3cc2954-95e1-476c-b6ae-55442ce90f2f · outbound

This paper cites Brown, Wonjoon Goo, and Scott Niekum.

Imitation Learning via Focused Satisficing Brown, Wonjoon Goo, and Scott Niekum

Reference 2019

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:35:39.516949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:35:34.153521Z digest=sha256:a07006a603116d86d0719075469bbe13e8fb4df725dfc652ee6e73455fb63827

Observation 52eb1291-eb47-49ac-9cc2-b91efbc3935f · outbound

This paper cites Distance minimization for reward learn- ing from scored trajectories.Proceedings of the AAAI Con- ference on Artificial Intelligence, 30(1),.

Imitation Learning via Focused Satisficing Distance minimization for reward learn- ing from scored trajectories.Proceedings of the AAAI Con- ference on Artificial Intelligence, 30(1),

Reference 2020

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:35:39.508110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:35:34.247089Z digest=sha256:8355f015b3e10885f347184ae8084962a8ab5e310b6de8522b6f8f47f690f541

Observation fedb857e-e6b1-4e19-9f48-407a89167f90 · outbound

This paper cites Rank analysis of incomplete block designs: I.

Imitation Learning via Focused Satisficing Rank analysis of incomplete block designs: I

Reference 2021

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:35:39.549659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:35:33.935001Z digest=sha256:121e6681d55a4f6b7fff21bb4cc56d91efcca43722179677da0e8237e96806d5

Observation fffdea06-e94a-4121-a594-bed435bdfef7 · outbound

This paper cites Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems.

Imitation Learning via Focused Satisficing Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:34.756014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:35:34.756014Z digest=sha256:db66c1ea6a054b53f5e3f32dff716e65b4f9b8384a213c38b12ee596005bc4a4

Observation 74fe25e3-1786-45ec-8ad7-01aeff79e74c · outbound

This paper cites Stable-baselines3: Reliable reinforcement learning implementations.

Imitation Learning via Focused Satisficing Stable-baselines3: Reliable reinforcement learning implementations

Reference 2024

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:35:38.291195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T15:35:35.243155Z digest=sha256:649826b666380ce47228e5aa3ba0bb1f8d69fe79ad06c25873db7521b78222a1

Pith citing papers

No inbound Pith citation observations are available.