Pith. sign in

Paper Citation Record · LEDGER

Behavior-Regularized Diffusion Policy Optimization for Offline Reinforcement Learning

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:2502.04778.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.04778 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 4 of 4 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:20:37.619351Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-25T07:35:29.370333Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 2dff5418-ebe0-492d-be11-b3819b4cd23c · inbound

Decision Flow Policy Optimization cites this paper.

Decision Flow Policy Optimization Behavior-Regularized Diffusion Policy Optimization for Offline Reinforcement Learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T14:20:37.619351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:20:37.619351Z digest=sha256:aa640b712477c670d27856d161f16e516c6a55dd8c42225ba3ff458777f54e78

Observation 64bb61ce-4e78-461e-abf5-ab87c39ae074 · inbound

D2 Actor Critic: Diffusion Actor Meets Distributional Critic cites this paper.

D2 Actor Critic: Diffusion Actor Meets Distributional Critic Behavior-Regularized Diffusion Policy Optimization for Offline Reinforcement Learning

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-25T07:35:29.373696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-25T07:31:27.330951Z digest=sha256:09f8b87d4588bef6ba528c60db7eab7e488496349732e80fa08887a453f74954

Observation 0223a517-9a63-446c-864c-802982f50c09 · inbound

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning cites this paper.

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning Behavior-Regularized Diffusion Policy Optimization for Offline Reinforcement Learning

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:56:00.695725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T16:25:25.739019Z digest=sha256:b7063eea6f5e1675d0b3dae833889fe5d8787c133aed4d2bf027a6ae19345273

Observation aeb489c3-3cea-402c-9aba-b3e92b38eb51 · inbound

ReBRAC-v2: The Return of the King cites this paper.

ReBRAC-v2: The Return of the King Behavior-Regularized Diffusion Policy Optimization for Offline Reinforcement Learning

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-06T00:32:46.502554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:32:46.502554Z digest=sha256:943bbd10397f8c2e4576211b43eb930f2b6d1482ca7631873e3707cb972d3337