Pith. sign in

Paper Citation Record · LEDGER

PufferLib: Making Reinforcement Learning Libraries and Environments Play Nice

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2406.12905.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.12905 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:57:40.035699Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T01:29:22.376416Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 7f1aee20-8479-4610-9cc4-a576405ad418 · inbound

Gymnasium: A Standard Interface for Reinforcement Learning Environments cites this paper.

Gymnasium: A Standard Interface for Reinforcement Learning Environments PufferLib: Making Reinforcement Learning Libraries and Environments Play Nice

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:29:49.500000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-11T17:29:49.186565Z digest=sha256:208bb060bb4f1638c0ac817d0b40ad24a8e2890c40d6441b5f4d3df51c030ab6

Observation db7404eb-e5c0-4b34-a830-334337701ea0 · inbound

The challenge of hidden gifts in multi-agent reinforcement learning cites this paper.

The challenge of hidden gifts in multi-agent reinforcement learning PufferLib: Making Reinforcement Learning Libraries and Environments Play Nice

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:40.035699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:57:40.035699Z digest=sha256:ef66b46a9456629ea364c7af4d9bc8a8a3569d2d5f732187b686cdc7382a1f81

Observation 45f8d382-7bb5-44ef-a732-8d09c8df9a06 · inbound

Combining Pre-Trained Models for Enhanced Feature Representation in Reinforcement Learning cites this paper.

Combining Pre-Trained Models for Enhanced Feature Representation in Reinforcement Learning PufferLib: Making Reinforcement Learning Libraries and Environments Play Nice

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T18:51:58.571631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:51:58.571631Z digest=sha256:a6a68466b3c90458a2aaad368ba535a425aa9d405f1442f0a8b5e9b08d5f6751

Observation 90f46c78-a691-4335-bd15-fac083517cec · inbound

Scalable Option Learning in High-Throughput Environments cites this paper.

Scalable Option Learning in High-Throughput Environments PufferLib: Making Reinforcement Learning Libraries and Environments Play Nice

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-18T20:06:49.697339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-05-18T20:04:58.064472Z digest=sha256:3f39dc3d7a215c488c87f6ccfed325d64f42c69542876a2ace54997e941cc97c

Observation 5c6bc485-9575-4ff9-a1c2-f3c46bd62297 · inbound

Automatic Generation of High-Performance RL Environments cites this paper.

Automatic Generation of High-Performance RL Environments PufferLib: Making Reinforcement Learning Libraries and Environments Play Nice

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-21T11:05:01.485239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-21T11:04:11.718672Z digest=sha256:17b6ca0457972bd02f6bbf6aec4a69b447b72b08ec2161be61490656946d6bda

Observation b466de6e-450e-468d-a3ca-1e3a786ffa26 · inbound

A High-Throughput Compute-Efficient POMDP Hide-And-Seek-Engine (HASE) for Multi-Agent Operations cites this paper.

A High-Throughput Compute-Efficient POMDP Hide-And-Seek-Engine (HASE) for Multi-Agent Operations PufferLib: Making Reinforcement Learning Libraries and Environments Play Nice

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:36:26.191830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-07T10:27:55.337253Z digest=sha256:ba2bab960e9a1fbd0b1485283e8f4961d2c4980c53ea1180414f139fbdc0f721

Observation e27902df-9854-42bf-b392-d4d79a8775cb · inbound

Equivariant Reinforcement Learning for Clifford Quantum Circuit Synthesis cites this paper.

Equivariant Reinforcement Learning for Clifford Quantum Circuit Synthesis PufferLib: Making Reinforcement Learning Libraries and Environments Play Nice

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:46:43.941390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-12T03:58:55.389930Z digest=sha256:1eb5806733f2b0196ca564a78d736deadfd5142f11ddf64693b52fe688a50c9b

Observation 6d542ba8-78fd-4013-b0d6-b2849523dfb5 · inbound

CoPark: Learning Reactive Parking via Self-Play cites this paper.

CoPark: Learning Reactive Parking via Self-Play PufferLib: Making Reinforcement Learning Libraries and Environments Play Nice

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-07-02T03:36:29.819429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-28T09:52:45.407210Z digest=sha256:fe876d1f1d0d48f86631ba83ceef16e657a493c41074e2937a76fef7b97b8760

Observation dd39ac08-8dad-49d4-b895-5cff951c3f25 · inbound

TerraTransfer: Learning End-to-End Driving Policies Without Expert Demonstrations cites this paper.

TerraTransfer: Learning End-to-End Driving Policies Without Expert Demonstrations PufferLib: Making Reinforcement Learning Libraries and Environments Play Nice

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-07-03T18:38:49.989093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-27T02:27:51.946679Z digest=sha256:cc9450037d1c9fc1b65fe4f75cab4e92fee05aa3a1e89f9533b93a86e33cecd3

Observation 657f0e52-a255-41b6-acb8-7a6c76044372 · inbound

TerraTransfer: Learning End-to-End Driving Policies Without Expert Demonstrations cites this paper.

TerraTransfer: Learning End-to-End Driving Policies Without Expert Demonstrations PufferLib: Making Reinforcement Learning Libraries and Environments Play Nice

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T11:08:37.763307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:08:37.763307Z digest=sha256:2c80632315d71ac5f57117a42e198a654bce3708c59a94d089b1e84fdf8f3400

Observation 2ac989f6-7aa7-46f5-a8dc-7aae20cf8587 · inbound

Human-like autonomy emerges from self-play and a pinch of human data cites this paper.

Human-like autonomy emerges from self-play and a pinch of human data PufferLib: Making Reinforcement Learning Libraries and Environments Play Nice

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-03T14:28:32.004591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-27T07:01:12.737217Z digest=sha256:5a57362e12d3f70b9313a97c0def3ec8991254e39b3c060a72c8b10c070eb0aa

Observation dc75fd39-b52b-4269-83a7-da6589ea4b09 · inbound

Scaling Self-Play for End-to-End Driving cites this paper.

Scaling Self-Play for End-to-End Driving PufferLib: Making Reinforcement Learning Libraries and Environments Play Nice

Reference 100

Resolution
verified exact
arxiv_id, observed 2026-07-04T01:29:22.377860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-26T20:22:03.938518Z digest=sha256:aea21a1876d6ba3fd68fdb925b734d7e6b24280c7135f244560a0a92ad2a9823

Observation 6db98e29-7bfb-4f56-a6fb-1af29c203139 · inbound

Pictura: Perspective-View Self-Play at Scale for Driving cites this paper.

Pictura: Perspective-View Self-Play at Scale for Driving PufferLib: Making Reinforcement Learning Libraries and Environments Play Nice

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-01T00:58:58.420429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:58:58.420429Z digest=sha256:1b4d055b1c177111fb2d9c9017a609ce759234688a6b2f692df9da9cd22858b6