Pith. sign in

Paper Citation Record · LEDGER

RVI-SAC: Average Reward Off-Policy Deep Reinforcement Learning

As of 16 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 2 inbound Pith citation observations for arXiv:2408.01972.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2408.01972 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 2 of 2 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T20:56:51.074552Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-10T20:13:58.440169Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3577ed12-96a2-4ede-afe8-fbe2a3f4b295 · inbound

An Empirical Study of Deep Reinforcement Learning in Continuing Tasks cites this paper.

An Empirical Study of Deep Reinforcement Learning in Continuing Tasks RVI-SAC: Average Reward Off-Policy Deep Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T20:56:51.074552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T20:56:51.074552Z digest=sha256:11a23a54a16be3e48f82ca4c292a9485e81b6525d1cecf24102edab443f0b5b3

Observation b7e82671-4a27-4b1f-b938-ed127239bce1 · inbound

Average-Reward Soft Actor-Critic cites this paper.

Average-Reward Soft Actor-Critic RVI-SAC: Average Reward Off-Policy Deep Reinforcement Learning

Reference 2018

Resolution
verified exact
local_arxiv, observed 2026-08-10T20:13:58.447879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T20:13:57.513390Z digest=sha256:a634b3d8cbb43145abb0be7821bbb8d234617f2b62603987d4c02a1b11dce75f