Pith. sign in

Paper Citation Record · LEDGER

POPGym: Benchmarking Partially Observable Reinforcement Learning

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2303.01859.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2303.01859 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:24:25.955282Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T20:27:22.562202Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 923dfd12-a7d8-45a1-ab01-8732785c259d · inbound

How Far Can LLMs Improve from Experience? Measuring Test-Time Learning Ability in LLMs with Human Comparison cites this paper.

How Far Can LLMs Improve from Experience? Measuring Test-Time Learning Ability in LLMs with Human Comparison POPGym: Benchmarking Partially Observable Reinforcement Learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T00:24:25.955282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:24:25.955282Z digest=sha256:d2485fbb56f9bf9d79f559789634932ab5d0626eea3a1b8de989337eb004393c

Observation 6adb8350-b470-4884-8cee-2a8e9ef67971 · inbound

Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks cites this paper.

Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks POPGym: Benchmarking Partially Observable Reinforcement Learning

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T11:03:27.386150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:03:27.386150Z digest=sha256:8616be00dc2547d0199e7d5ced15d44ba49a2638838175a4e9880dcf0ddc1c9e

Observation d5447535-73d6-4df3-b03d-c627294d3599 · inbound

Belief-State RWKV for Reinforcement Learning under Partial Observability cites this paper.

Belief-State RWKV for Reinforcement Learning under Partial Observability POPGym: Benchmarking Partially Observable Reinforcement Learning

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-13T21:58:19.950545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T21:56:09.642064Z digest=sha256:df03132f71a26fb91026b00c35e1494878a0a2317925f4bef704b6206fc2e781

Observation 134b0871-2313-4fb8-98ae-4796c0df24a0 · inbound

Why Linear Recurrent Memory Works in Partially Observable Reinforcement Learning cites this paper.

Why Linear Recurrent Memory Works in Partially Observable Reinforcement Learning POPGym: Benchmarking Partially Observable Reinforcement Learning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-06-28T23:32:47.443609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T23:27:39.024454Z digest=sha256:26b4a49202bc825f310f963536b616f2f5202384c52284c8d262869d57e48b74

Observation 43866d0e-11d0-4c14-8d8f-dec804b4bcde · inbound

Belief-Aware Scheduling for Predictive Wildfire Hazard Mapping under Sparse-Window Telemetry cites this paper.

Belief-Aware Scheduling for Predictive Wildfire Hazard Mapping under Sparse-Window Telemetry POPGym: Benchmarking Partially Observable Reinforcement Learning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:27:22.563742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T20:24:01.961913Z digest=sha256:3c786dced27aa9c3fc98dd044777a0c8b9fb3ac0bb0a2f37858ac748d91df5cd

Observation 4ec736aa-66df-4009-9630-4030d0f80ee8 · inbound

ASK in the Dark: Uncertainty-Gated LLM Assistance under Partial Observability cites this paper.

ASK in the Dark: Uncertainty-Gated LLM Assistance under Partial Observability POPGym: Benchmarking Partially Observable Reinforcement Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-12T07:43:59.534909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T07:43:59.534909Z digest=sha256:f3697a171e63b15f76518edc0af75b34ea3391636f7247b50613046959615943