Pith. sign in

Paper Citation Record · LEDGER

Motif: Intrinsic Motivation from Artificial Intelligence Feedback

As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2310.00166.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2310.00166 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 19 of 19 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:36:11.387527Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T20:48:55.285309Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b96f90a6-1417-4d05-a595-cd4b91c98aa0 · inbound

A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback cites this paper.

A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Motif: Intrinsic Motivation from Artificial Intelligence Feedback

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-12T16:30:09.439343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:30:09.439343Z digest=sha256:adf5c7ee3364b19a96b5db6f044ad1c8b67d6a8e9171996f93f55e33bd0b644e

Observation ed1a5933-e60f-4642-9cad-c860934eda93 · inbound

BALROG: Benchmarking Agentic LLM and VLM Reasoning On Games cites this paper.

BALROG: Benchmarking Agentic LLM and VLM Reasoning On Games Motif: Intrinsic Motivation from Artificial Intelligence Feedback

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T16:21:08.226030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T16:21:08.226030Z digest=sha256:fad78afbf45f505ecace698e79486426d0f8dba237233e71e45b5e8079e68801

Observation 6629c4af-c145-4cfa-a771-f0e5e1be95dc · inbound

Probing for Consciousness in Machines cites this paper.

Probing for Consciousness in Machines Motif: Intrinsic Motivation from Artificial Intelligence Feedback

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T13:27:12.728339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:27:12.728339Z digest=sha256:ccfa9747c33a81d0f2bc2313281f0ae61779077dfadd8f20129beb715997f93e

Observation 85820567-6bd6-4198-a2bc-470379680d37 · inbound

Effective Reward Specification in Deep Reinforcement Learning cites this paper.

Effective Reward Specification in Deep Reinforcement Learning Motif: Intrinsic Motivation from Artificial Intelligence Feedback

Reference 165

Resolution
unresolved
no resolver link, observed 2026-08-11T19:09:54.562965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T19:09:54.562965Z digest=sha256:a9d7f07304950d8ef4d617d9ee03805537b4cce506c88cf819355b14adec697b

Observation ce1ce659-408c-478b-b279-083d16f8f96f · inbound

A Self-Improving Coding Agent cites this paper.

A Self-Improving Coding Agent Motif: Intrinsic Motivation from Artificial Intelligence Feedback

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T11:36:11.387527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:36:11.387527Z digest=sha256:eec2cca206165969b1a619ec8e587b4b30a4325337bdf0e0edb0532b957ae9d5

Observation 8b73f263-94f0-46dd-a90f-7947a64dbdd7 · inbound

LLMs are Greedy Agents: Effects of RL Fine-tuning on Decision-Making Abilities cites this paper.

LLMs are Greedy Agents: Effects of RL Fine-tuning on Decision-Making Abilities Motif: Intrinsic Motivation from Artificial Intelligence Feedback

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-16T11:16:29.681423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:16:29.681423Z digest=sha256:dfd7c94fae06a6b2d1605a34d0b0f9bb44b65c4f1f0d665292bd1e5d21d9d85e

Observation 5e49f440-9299-4c4c-aa45-57d9ed67b91b · inbound

PRISM: Projection-based Reward Integration for Scene-Aware Real-to-Sim-to-Real Transfer with Few Demonstrations cites this paper.

PRISM: Projection-based Reward Integration for Scene-Aware Real-to-Sim-to-Real Transfer with Few Demonstrations Motif: Intrinsic Motivation from Artificial Intelligence Feedback

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-16T05:31:55.924776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:31:55.924776Z digest=sha256:ca986b4795a7901edeaa3cf681922de5ddc24f2c1aaef2752e54bc51724aa173

Observation d09e3dd3-6de9-48af-818e-5448436fe3e6 · inbound

TREND: Tri-teaching for Robust Preference-based Reinforcement Learning with Demonstrations cites this paper.

TREND: Tri-teaching for Robust Preference-based Reinforcement Learning with Demonstrations Motif: Intrinsic Motivation from Artificial Intelligence Feedback

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-15T22:52:23.991919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:52:23.991919Z digest=sha256:7e0f17442e1051c6793a28c77c7291850b7623ec5dd793cdc1a68ddca2195347

Observation 46f564a4-a23e-4fbd-a944-409edba38e42 · inbound

Enhancing Rating-Based Reinforcement Learning to Effectively Leverage Feedback from Large Vision-Language Models cites this paper.

Enhancing Rating-Based Reinforcement Learning to Effectively Leverage Feedback from Large Vision-Language Models Motif: Intrinsic Motivation from Artificial Intelligence Feedback

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T20:11:50.010153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:11:50.010153Z digest=sha256:667c8955a105b9963b141f38d63462de213344e15259a5829c5b4ac54db8d3bb

Observation 1be62d4a-6672-4734-b03a-985e95e42ca8 · inbound

Reward Models in Deep Reinforcement Learning: A Survey cites this paper.

Reward Models in Deep Reinforcement Learning: A Survey Motif: Intrinsic Motivation from Artificial Intelligence Feedback

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-15T19:38:50.645727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:38:50.645727Z digest=sha256:670f48f2246eb89817ca0f1f1d1084a8abe11e603b65d4a6f5b5b7815da71f9c

Observation 4eeb02b1-76de-49a5-af09-5122d2f86184 · inbound

FOUNDER: Grounding Foundation Models in World Models for Open-Ended Embodied Decision Making cites this paper.

FOUNDER: Grounding Foundation Models in World Models for Open-Ended Embodied Decision Making Motif: Intrinsic Motivation from Artificial Intelligence Feedback

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T17:12:28.459946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:12:28.459946Z digest=sha256:c35dfa905f19ce6a0fdfe5f114f3608c348c7a6432042825cb820828e9ea54ca

Observation 395ebe7e-bb63-45b4-bdfe-8a918f69e06b · inbound

LLM Economist: Large Population Models and Mechanism Design in Multi-Agent Generative Simulacra cites this paper.

LLM Economist: Large Population Models and Mechanism Design in Multi-Agent Generative Simulacra Motif: Intrinsic Motivation from Artificial Intelligence Feedback

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T15:28:43.636971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:28:43.636971Z digest=sha256:ad077286fa50fd8ea6347adeeb7ae14d364255f2cafaea9c9f89d9693dfe41b3

Observation 2632fedb-e8d5-4038-9601-fb61ee64f41b · inbound

Timing the Message: Language-Based Notifications for Time-Critical Assistive Settings cites this paper.

Timing the Message: Language-Based Notifications for Time-Critical Assistive Settings Motif: Intrinsic Motivation from Artificial Intelligence Feedback

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T22:17:50.436207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:17:50.436207Z digest=sha256:7e435e3c25f17ae16ce2f976957504f22d6dd8ffadf803e50152a6bf1c9c6a9b

Observation 99071007-bad2-430b-9938-23976410309c · inbound

Hierarchical Behaviour Spaces cites this paper.

Hierarchical Behaviour Spaces Motif: Intrinsic Motivation from Artificial Intelligence Feedback

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:01:12.847279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-08T03:33:48.527621Z digest=sha256:0d8cf6ea8396b698aba5d67d81eda1d9e126acdb79b171f223fc25bc4fc91b69

Observation 7c38ead9-20fb-4471-8c24-b690bc4e6f19 · inbound

Agentick: A Unified Benchmark for General Sequential Decision-Making Agents cites this paper.

Agentick: A Unified Benchmark for General Sequential Decision-Making Agents Motif: Intrinsic Motivation from Artificial Intelligence Feedback

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:25:58.745765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-05-11T01:22:32.713175Z digest=sha256:a1fe0e0c676c09d633c4051d9d8c4f8dbff125fee6a92a9a52444b4c4e674874

Observation 2e722bf5-783c-4481-9561-1108d5eb2fa6 · inbound

Goal-Conditioned Agents that Learn Everything All at Once cites this paper.

Goal-Conditioned Agents that Learn Everything All at Once Motif: Intrinsic Motivation from Artificial Intelligence Feedback

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:00:21.663285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-05-25T04:59:48.867927Z digest=sha256:4af9605ea7998ca3427b5a853054c3de5e74a7033fa3fae0b8baddffacf1333b

Observation e4a6465b-583c-466b-be0e-1c2f9121b2f2 · inbound

VLM-AR3L: Vision-Language Models for Absolute and Relative Rewards in Reinforcement Learning cites this paper.

VLM-AR3L: Vision-Language Models for Absolute and Relative Rewards in Reinforcement Learning Motif: Intrinsic Motivation from Artificial Intelligence Feedback

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-07-02T11:56:54.894458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-07-02T11:52:25.035266Z digest=sha256:53545910803bac7cc07365f36d28f64327bf4ff95619a333c5237b5bebbc9a16

Observation 471e7edd-6813-4514-9ae8-69eaff48942b · inbound

VLM-AR3L: Vision-Language Models for Absolute and Relative Rewards in Reinforcement Learning cites this paper.

VLM-AR3L: Vision-Language Models for Absolute and Relative Rewards in Reinforcement Learning Motif: Intrinsic Motivation from Artificial Intelligence Feedback

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:48:55.287255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-07-03T20:43:34.222848Z digest=sha256:4a4b9617ea3911358290a67312c4f4a1b2c30a39f91e5d58517ab2f778cb66ca

Observation 7d1ad87f-dcd3-44f3-9305-a9b5f4ae0224 · inbound

Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills cites this paper.

Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills Motif: Intrinsic Motivation from Artificial Intelligence Feedback

Reference 121

Resolution
unresolved
no resolver link, observed 2026-08-04T19:45:34.833075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:45:34.833075Z digest=sha256:c8e638d7784371be0e0b34e9b04be0822f32c38470d36739e4cfdbe79554acf1