Pith. sign in

Paper Citation Record · LEDGER

Highly Efficient Self-Adaptive Reward Shaping for Reinforcement Learning

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:2408.03029.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2408.03029 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 4 of 4 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T16:35:01.478900Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T09:59:44.620366Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3e597636-2626-49ea-8caf-c9ffcbacd396 · inbound

Occupancy Reward Shaping: Improving Credit Assignment for Offline Goal-Conditioned Reinforcement Learning cites this paper.

Occupancy Reward Shaping: Improving Credit Assignment for Offline Goal-Conditioned Reinforcement Learning Highly Efficient Self-Adaptive Reward Shaping for Reinforcement Learning

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-10T00:24:47.176236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T00:19:48.053466Z digest=sha256:c3c2ef39f88d473f8b047ac99b865bb6fec1207d0d429ab8b049c3297824ec62

Observation 5d24a28d-6e8f-403d-8d68-7a6b57df9834 · inbound

Learning Process Rewards via Success Visitation Matching for Efficient RL cites this paper.

Learning Process Rewards via Success Visitation Matching for Efficient RL Highly Efficient Self-Adaptive Reward Shaping for Reinforcement Learning

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-07-04T09:59:44.621838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T09:20:35.062060Z digest=sha256:a4df367277fee3675bc889408ea546aca9c4a7e1c0e8f1818de24086dd55c001

Observation 6d0d0be8-2b84-40f4-a650-31a938adcc53 · inbound

Information-Based Exploration via Random Features for Reinforcement Learning cites this paper.

Information-Based Exploration via Random Features for Reinforcement Learning Highly Efficient Self-Adaptive Reward Shaping for Reinforcement Learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T16:35:01.478900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T16:35:01.478900Z digest=sha256:973ddf95b6003e94800480f405080f73cfc651d78bb62b91a555cee315dea36d

Observation 60712595-d350-4103-9b40-62289167d629 · inbound

Meta-Learned Reward Shaping for Reinforcement Learning from Human Feedback cites this paper.

Meta-Learned Reward Shaping for Reinforcement Learning from Human Feedback Highly Efficient Self-Adaptive Reward Shaping for Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T03:14:11.955077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:14:11.955077Z digest=sha256:c5b95c7d4984e6f02d841e675750a49ed38aaf436ef2db3d61bbc05921f3fb80