Pith. sign in

Paper Citation Record · LEDGER

FutureWorld: A Live Reinforcement Learning Environment for Predictive Agents with Real-World Outcome Rewards

As of 4 August 2026, this Paper Citation Record lists 3 of 3 outbound references and 0 inbound Pith citation observations for arXiv:2604.26733.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2604.26733 v4

Coverage vector

measured 3 of 3 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-11T11:50:26.030339Z

measured 3 of 3 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

3 of 3 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 25088d7d-5bec-46f7-bc20-57dbb390f5d3 · outbound

This paper cites Reinforcement Learning for Long-Horizon Interactive LLM Agents.

FutureWorld: A Live Reinforcement Learning Environment for Predictive Agents with Real-World Outcome Rewards Reinforcement Learning for Long-Horizon Interactive LLM Agents

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-19T17:22:42.017929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T17:20:07.747499Z digest=sha256:63281d77ece419b3d352db33c52f59f7e20dd31cf997e2d0ea85f207eaa84b14

Observation 27bde80e-076b-48e8-9e59-49b15b015e0a · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

FutureWorld: A Live Reinforcement Learning Environment for Predictive Agents with Real-World Outcome Rewards DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-05-19T17:22:41.691326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-11T11:50:26.030339Z digest=sha256:656a407277db9a1ebf62e80e19191b9ddf19138cddbfd17c2961afc7881a0806

Observation 8457dff8-85e5-45cb-b037-9d27563d8ebc · outbound

This paper cites VisualWebArena: Evaluating multimodal agents on realistic visual web tasks.

FutureWorld: A Live Reinforcement Learning Environment for Predictive Agents with Real-World Outcome Rewards VisualWebArena: Evaluating multimodal agents on realistic visual web tasks

Reference 3

Resolution
metadata mismatch
doi, observed 2026-05-19T17:22:41.693769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T17:20:07.747499Z digest=sha256:c1ccaa35d702252b20e4b81338d9e9935e4219c832f0a83ab1f88474b262618a

Pith citing papers

No inbound Pith citation observations are available.