Pith. sign in

Paper Citation Record · LEDGER

Learning from Random Demonstrations: Offline Reinforcement Learning with Importance-Sampled Diffusion Models

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2405.19878.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.19878 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:17:38.372928Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-25T07:00:27.058047Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 4dcc9a95-1838-4aca-ab63-8e42fe3210c1 · inbound

Residual Reward Models for Preference-based Reinforcement Learning cites this paper.

Residual Reward Models for Preference-based Reinforcement Learning Learning from Random Demonstrations: Offline Reinforcement Learning with Importance-Sampled Diffusion Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T21:17:38.372928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:17:38.372928Z digest=sha256:9dfe54bbe7b288eabbf7f060666a8b3e0bdeb17e6747b04789950e1d9aa247b1

Observation 779ff215-163e-4f9d-99aa-21a7e530b946 · inbound

Tail-Risk-Safe Monte Carlo Tree Search under PAC-Level Guarantees cites this paper.

Tail-Risk-Safe Monte Carlo Tree Search under PAC-Level Guarantees Learning from Random Demonstrations: Offline Reinforcement Learning with Importance-Sampled Diffusion Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T23:22:29.353384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:22:29.353384Z digest=sha256:65345b5a6d1aaa5a9700948354f9372c3cd9b87937f037c6e223617606b4e608

Observation 0359c6b4-2d84-487d-b532-483db052a787 · inbound

Perception Graph for Cognitive Attack Reasoning in Augmented Reality cites this paper.

Perception Graph for Cognitive Attack Reasoning in Augmented Reality Learning from Random Demonstrations: Offline Reinforcement Learning with Importance-Sampled Diffusion Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T13:32:58.153412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:32:58.153412Z digest=sha256:8359092f6301c64b737be43f0eb6c26af8af3cd3ee27b2ba90a94892209a5980

Observation 0d703c39-7987-4691-bf5f-81e74f239498 · inbound

MINT: Minimal Information Neuro-Symbolic Tree for Objective-Driven Knowledge-Gap Reasoning and Active Elicitation cites this paper.

MINT: Minimal Information Neuro-Symbolic Tree for Objective-Driven Knowledge-Gap Reasoning and Active Elicitation Learning from Random Demonstrations: Offline Reinforcement Learning with Importance-Sampled Diffusion Models

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-16T07:17:30.684934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-16T07:13:42.705619Z digest=sha256:dfa032a9f2231ad583b86febf85172858fe6a1763f9119f1e9ee41f75eba6367

Observation ed0f8d50-468a-4a83-948e-4b7e5014a300 · inbound

IntentScore: Intent-Conditioned Action Evaluation for Computer-Use Agents cites this paper.

IntentScore: Intent-Conditioned Action Evaluation for Computer-Use Agents Learning from Random Demonstrations: Offline Reinforcement Learning with Importance-Sampled Diffusion Models

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:25:54.181045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T19:08:21.527478Z digest=sha256:bedc20221722138617855f9f68f34fc9ac120e2eeb2df9d03f99e8105c5bc5d1

Observation a260149c-f192-4266-940a-148941706656 · inbound

IntentScore: Intent-Conditioned Action Evaluation for Computer-Use Agents cites this paper.

IntentScore: Intent-Conditioned Action Evaluation for Computer-Use Agents Learning from Random Demonstrations: Offline Reinforcement Learning with Importance-Sampled Diffusion Models

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-25T07:00:27.062676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-25T06:56:56.794206Z digest=sha256:858356be0646d5038dd68a51188bc5c18ff2535ed3ee37c16094946c308d652a