Pith. sign in

Paper Citation Record · LEDGER

Uncertainty Weighted Actor-Critic for Offline Reinforcement Learning

As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2105.08140.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2105.08140 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T22:17:41.939227Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T08:43:15.128113Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c7714f3f-054e-41d6-b19f-15301083ebf4 · inbound

Mitigating Relative Over-Generalization in Multi-Agent Reinforcement Learning cites this paper.

Mitigating Relative Over-Generalization in Multi-Agent Reinforcement Learning Uncertainty Weighted Actor-Critic for Offline Reinforcement Learning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-12T19:02:50.926448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T19:02:50.926448Z digest=sha256:8a9adba177f34d479b4ca05f79ccbcf97e29126762311a9676eb5380684fb184

Observation 9e394043-acc7-4829-a3c8-2dce28da6763 · inbound

Combining Bayesian Inference and Reinforcement Learning for Agent Decision Making: A Review cites this paper.

Combining Bayesian Inference and Reinforcement Learning for Agent Decision Making: A Review Uncertainty Weighted Actor-Critic for Offline Reinforcement Learning

Reference 192

Resolution
unresolved
no resolver link, observed 2026-08-15T22:17:41.939227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:17:41.939227Z digest=sha256:546e06617b3f6b65626c560ca637bc60da6ddefa39c5bb80460b69c368ec2c6c

Observation cd89333e-74bc-4e73-9bc8-3c96b01c2613 · inbound

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning cites this paper.

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning Uncertainty Weighted Actor-Critic for Offline Reinforcement Learning

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:56:00.734820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-10T16:25:25.739019Z digest=sha256:e53cd5a2d77370840c21a64b57b3f336c5025cf2e43290b18fcde15a50d3e86b

Observation 9c4d26f6-1093-42de-b6f9-f4cb7be3b953 · inbound

Beyond Penalization: Diffusion-based Out-of-Distribution Detection and Selective Regularization in Offline Reinforcement Learning cites this paper.

Beyond Penalization: Diffusion-based Out-of-Distribution Detection and Selective Regularization in Offline Reinforcement Learning Uncertainty Weighted Actor-Critic for Offline Reinforcement Learning

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:41:50.547178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-12T02:17:25.783688Z digest=sha256:74570a1ec98b698b516f89d84958b4b2e706f78ef6281ab4a88cfbf3c08f7651

Observation 1273e49c-7c95-4b3f-84d3-f000ba58e8fb · inbound

UNIQ: Conformal Calibration for Adaptive Conservatism in Offline Reinforcement Learning cites this paper.

UNIQ: Conformal Calibration for Adaptive Conservatism in Offline Reinforcement Learning Uncertainty Weighted Actor-Critic for Offline Reinforcement Learning

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:43:15.129743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-06-29T08:39:34.884726Z digest=sha256:d33580168caedf77745ee2f2729a932417dee0aa3e100ba329048d44867702a5

Observation b02f5f93-5f63-4162-94f2-c034f54f4e04 · inbound

Multi-Turn On-Policy Distillation with Prefix Replay cites this paper.

Multi-Turn On-Policy Distillation with Prefix Replay Uncertainty Weighted Actor-Critic for Offline Reinforcement Learning

Reference 176

Resolution
unresolved
no resolver link, observed 2026-07-11T13:53:36.775836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T13:53:36.775836Z digest=sha256:7b920e9ba9b32fd97615c190abc13161e67f9a4e2c6ac5c784be8aee91ccb13b

Observation 3e3ed705-f1f9-4f08-9ccb-495cb2c68033 · inbound

Multi-Turn On-Policy Distillation with Prefix Replay cites this paper.

Multi-Turn On-Policy Distillation with Prefix Replay Uncertainty Weighted Actor-Critic for Offline Reinforcement Learning

Reference 177

Resolution
unresolved
no resolver link, observed 2026-08-02T08:40:52.852816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T08:40:52.852816Z digest=sha256:1b9bad1693bc06baf1cde1d1eb7f68f161b795e35d371835c62f8085232049c7

Observation a074f1fd-fa65-4776-a518-e65e3bfced2f · inbound

Collaborative Weighting with Pessimistic Critic for Mitigating Overestimation in Off-Policy Reinforcement Learning cites this paper.

Collaborative Weighting with Pessimistic Critic for Mitigating Overestimation in Off-Policy Reinforcement Learning Uncertainty Weighted Actor-Critic for Offline Reinforcement Learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T14:19:58.343808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:19:58.343808Z digest=sha256:ab9459bee026cc2df908039ecb5aa56a1c629a5c76584e47f6817bd724057d23

Observation 9b78a919-2db4-40d0-8f21-70201db89ad9 · inbound

Uncertainty-Guided LLM Semantic Augmentation for Heterogeneous Treatment Effect Estimation cites this paper.

Uncertainty-Guided LLM Semantic Augmentation for Heterogeneous Treatment Effect Estimation Uncertainty Weighted Actor-Critic for Offline Reinforcement Learning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-01T12:41:27.456257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:41:27.456257Z digest=sha256:13e3cb798e71b6dc31480d6a6243775e4577e1705c49409df09cc7c7c49b7c45

Observation 685f7411-20ce-4c0a-968f-f571d87c969c · inbound

Convex-Hull-Neighborhood Smooth Dual Generalization: Controlling Local Correction Propagation in Offline RL cites this paper.

Convex-Hull-Neighborhood Smooth Dual Generalization: Controlling Local Correction Propagation in Offline RL Uncertainty Weighted Actor-Critic for Offline Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T14:57:30.244330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:57:30.244330Z digest=sha256:0d721dcece36c88ba50dc49d6629a64a7f4b51d28ea7787cc7b9f2462862e980