Pith. sign in

Paper Citation Record · LEDGER

Diffusion Models for Reinforcement Learning: A Survey

As of 24 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 23 inbound Pith citation observations for arXiv:2311.01223.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2311.01223 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 23 of 23 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T00:07:53.191949Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T04:27:36.344632Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 990409e2-9687-4e34-b126-f2b05034b4d9 · inbound

Diffusion Policy Policy Optimization cites this paper.

Diffusion Policy Policy Optimization Diffusion Models for Reinforcement Learning: A Survey

Reference 108

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:48:15.046748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-16T08:48:14.776754Z digest=sha256:0d2134a6a5a542bd519d7b78d5a5c91ef12a3f6e5ac7bb98ff9b384ca2cee964

Observation a33c2e8f-595f-448c-ba50-fecddc257fad · inbound

Exploratory Diffusion Model for Unsupervised Reinforcement Learning cites this paper.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Diffusion Models for Reinforcement Learning: A Survey

Reference 70

Resolution
malformed identifier
no resolver link, observed 2026-08-08T13:21:18.474336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.474336Z digest=sha256:8ec1869131bfd41e73d1c70b5c4b47791545b070765e8919a637f51c0a9f8f23

Observation fd65aef8-d687-4e42-a755-0c3f22f182b3 · inbound

Latent Theory of Mind: A Decentralized Diffusion Architecture for Cooperative Manipulation cites this paper.

Latent Theory of Mind: A Decentralized Diffusion Architecture for Cooperative Manipulation Diffusion Models for Reinforcement Learning: A Survey

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:01.759305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:01.759305Z digest=sha256:8f2935d0b0c63eebc496ab9f3ee4e1acbfc1ba4515d5ceec148cb92771594fa7

Observation b0a145b2-c1a7-45a0-b940-986bb199780b · inbound

Improved Sample Complexity For Diffusion Model Training Without Empirical Risk Minimizer Access cites this paper.

Improved Sample Complexity For Diffusion Model Training Without Empirical Risk Minimizer Access Diffusion Models for Reinforcement Learning: A Survey

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-19T12:37:17.434588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-19T12:36:56.621522Z digest=sha256:84e358624e44d3b854d373f146fdd494eef02886a7824488e283c83d962318bc

Observation d5fe7f67-3312-425e-b01b-c10578630d15 · inbound

STITCH-OPE: Trajectory Stitching with Guided Diffusion for Off-Policy Evaluation cites this paper.

STITCH-OPE: Trajectory Stitching with Guided Diffusion for Off-Policy Evaluation Diffusion Models for Reinforcement Learning: A Survey

Reference 63

Resolution
malformed identifier
no resolver link, observed 2026-08-07T13:54:07.158408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:07.158408Z digest=sha256:592d729876bf907296008e3d390c2c941bf29926c6833b02d6efc8e7ca594c15

Observation e9268e5d-894a-41da-860e-3e09c4e33196 · inbound

BadReward: Clean-Label Poisoning of Reward Models in Text-to-Image RLHF cites this paper.

BadReward: Clean-Label Poisoning of Reward Models in Text-to-Image RLHF Diffusion Models for Reinforcement Learning: A Survey

Reference 38

Resolution
malformed identifier
no resolver link, observed 2026-08-07T11:18:05.968907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:18:05.968907Z digest=sha256:4ffe2fa4955eb97c5f1be25d73567de0c7c8ede009c44cdc390f87554316a245

Observation 8f474f92-de88-4c47-8ede-ea7f051e9588 · inbound

An Optimization-Augmented Control Framework for Single and Coordinated Multi-Arm Robotic Manipulation cites this paper.

An Optimization-Augmented Control Framework for Single and Coordinated Multi-Arm Robotic Manipulation Diffusion Models for Reinforcement Learning: A Survey

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T23:40:38.939757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:40:38.939757Z digest=sha256:abfcceb291356762447bcd3de281e8180b401013e22d596c7493c1495ff03c33

Observation 80d00b20-8190-40f6-87d5-e785056e23db · inbound

Object-centric Denoising Diffusion Models for Physical Reasoning cites this paper.

Object-centric Denoising Diffusion Models for Physical Reasoning Diffusion Models for Reinforcement Learning: A Survey

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:28.621893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:41:28.621893Z digest=sha256:e91ccfd826e43087892976be331a822a8def38d07aed381dec7a587b9a3e88ea

Observation 2d786222-d7f8-4638-bd94-1b854de13bc4 · inbound

Diffusion-RL Based Air Traffic Conflict Detection and Resolution Method cites this paper.

Diffusion-RL Based Air Traffic Conflict Detection and Resolution Method Diffusion Models for Reinforcement Learning: A Survey

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-05T11:23:27.972831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:23:27.972831Z digest=sha256:fe098f3935fbdd1b72a683078bbe6929e602a7571fa79c1ab22b9cb44cac4924

Observation 22d8c262-be43-4a70-bc7a-13ea3c60a729 · inbound

Rectified Schr\"odinger Bridge Matching for Few-Step Visual Navigation cites this paper.

Rectified Schr\"odinger Bridge Matching for Few-Step Visual Navigation Diffusion Models for Reinforcement Learning: A Survey

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T23:00:48.553791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-10T19:25:36.704959Z digest=sha256:47521f246d50b18b79876ba9884d903eabbf9b8b0b8375ade7c3dd03045d188e

Observation d400361a-b222-4095-9e17-8903a6d3197f · inbound

TacticGen: Grounding Adaptable and Scalable Generation of Football Tactics cites this paper.

TacticGen: Grounding Adaptable and Scalable Generation of Football Tactics Diffusion Models for Reinforcement Learning: A Survey

Reference 47

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T12:06:03.147765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-10T04:11:45.093169Z digest=sha256:8055b333af82bd4c19a0f5d826d7ba4bffcb4f9529439a60afec6c39167902a7

Observation b6368822-6ee4-4c2e-8e3b-14ce852c4fff · inbound

Muninn: Your Trajectory Diffusion Model But Faster cites this paper.

Muninn: Your Trajectory Diffusion Model But Faster Diffusion Models for Reinforcement Learning: A Survey

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:56:27.715605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-12T03:49:48.492957Z digest=sha256:75fda1c4adffbedf5b05d8b4f2591f92217ede5a39b2af46f97f96866f1586df

Observation b05d732e-d771-4571-ad4f-0c7d7018ff48 · inbound

Aligning Flow Map Policies with Optimal Q-Guidance cites this paper.

Aligning Flow Map Policies with Optimal Q-Guidance Diffusion Models for Reinforcement Learning: A Survey

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:07:17.435900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-13T05:04:22.603786Z digest=sha256:7907c70a6ed7741c7e5c2a55145f92c7e591120045bf90a54c8eb71bbab088a2

Observation 10ca82c1-d5ee-4da0-8ffc-7833d1d47d47 · inbound

Ada-Diffuser: Latent-Aware Adaptive Diffusion for Decision-Making cites this paper.

Ada-Diffuser: Latent-Aware Adaptive Diffusion for Decision-Making Diffusion Models for Reinforcement Learning: A Survey

Reference 84

Resolution
verified exact
arxiv_id, observed 2026-05-20T20:59:02.029551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-05-20T20:54:31.025488Z digest=sha256:491b03d35569483e64a52467cf000db4caadf435708a3850f5f4a9d629316c03

Observation 61fbfce4-ed00-472c-b982-9255bf511eac · inbound

Global Convergence of Sampling-Based Nonconvex Optimization through Diffusion-Style Smoothing cites this paper.

Global Convergence of Sampling-Based Nonconvex Optimization through Diffusion-Style Smoothing Diffusion Models for Reinforcement Learning: A Survey

Reference 239

Resolution
verified exact
arxiv_id, observed 2026-05-20T20:18:59.794591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-05-20T20:15:44.030714Z digest=sha256:110eee4bb5ddd6bb3533e396018923e9fd9b7d44e7420eaa1ecfbe124af4a4f6

Observation a8eb4fef-b0c8-4868-b016-98a21583bc01 · inbound

From Denoising to Decision Making: A Survey on Diffusion Model-Enabled Deep Reinforcement Learning for Wireless Networks cites this paper.

From Denoising to Decision Making: A Survey on Diffusion Model-Enabled Deep Reinforcement Learning for Wireless Networks Diffusion Models for Reinforcement Learning: A Survey

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T20:53:58.376467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-29T20:49:07.030872Z digest=sha256:1f65c2d16bfc08ee67889e7bc69a263d0e3bb7018b546c668ef7f9cc2a3da356

Observation f084e249-8118-4c72-9fb9-e11ca282f95a · inbound

From Noise to Control: Parameterized Diffusion Policies cites this paper.

From Noise to Control: Parameterized Diffusion Policies Diffusion Models for Reinforcement Learning: A Survey

Reference 67

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T19:46:10.592151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-06-28T22:06:53.667826Z digest=sha256:57bd60a3fb4d5f179eb40f5fa430b92c35df0361873c32369a9296dd9ea31535

Observation 1a95e9e3-abe0-456e-b1d9-b66dec6042aa · inbound

Conditional Graph Diffusion for Negotiation Support: Overcoming Discrete Infeasibility and Preference Elicitation Gaps cites this paper.

Conditional Graph Diffusion for Negotiation Support: Overcoming Discrete Infeasibility and Preference Elicitation Gaps Diffusion Models for Reinforcement Learning: A Survey

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-07-02T01:16:24.728046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T12:18:58.904218Z digest=sha256:0f23107b468247104be90195ddcfc59ef0d5ccd45fc093856bac1f2dcbb20aca

Observation b17c9bfa-a6ca-42d9-82fd-c87f924ea5c7 · inbound

Fast and Highly Expressive Policy Learning for Offline Reinforcement Learning via Bootstrapped Flow Q-Learning cites this paper.

Fast and Highly Expressive Policy Learning for Offline Reinforcement Learning via Bootstrapped Flow Q-Learning Diffusion Models for Reinforcement Learning: A Survey

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-07-03T04:27:36.346168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-27T13:57:31.180928Z digest=sha256:674afcee67d8282e76862bc2bf54c336ae41188c56970881a8c0d5ef19edd832

Observation 2aa5115a-ddad-480b-b1cb-8c5a465161e8 · inbound

RS-Diffuser: Risk-Sensitive Diffusion Planning with Distributional Value Guidance cites this paper.

RS-Diffuser: Risk-Sensitive Diffusion Planning with Distributional Value Guidance Diffusion Models for Reinforcement Learning: A Survey

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:23:51.463309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-29T05:05:39.828659Z digest=sha256:17dfda092eeaf471d601605e134376c5f515d696cd025175b58f4ea596c71677

Observation 9ee80221-eceb-4d0f-ba47-ce763e985432 · inbound

VINE: Taming Generative Control Policies for Reinforcement Learning cites this paper.

VINE: Taming Generative Control Policies for Reinforcement Learning Diffusion Models for Reinforcement Learning: A Survey

Reference 49

Resolution
unresolved
no resolver link, observed 2026-07-14T12:17:04.321971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:17:04.321971Z digest=sha256:767e437482928c83bbcc840d00b863b01caf14ca3910fa07bef29de23e2e4914

Observation 7fda30b7-b65e-4380-9d25-efe1ee9adb08 · inbound

JAGG: Jacobian-Aggregated Group Gradient for Efficient GRPO Training of Diffusion Models cites this paper.

JAGG: Jacobian-Aggregated Group Gradient for Efficient GRPO Training of Diffusion Models Diffusion Models for Reinforcement Learning: A Survey

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-01T17:41:04.258735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T17:41:04.258735Z digest=sha256:c6fd3340cb05247886af964ac049fe5893f35a5b6daf157eccd9c20fc7bdcd64

Observation 4885a165-2807-4a09-b1c7-cb56c57db289 · inbound

Diffusion Models in Finance: A Survey cites this paper.

Diffusion Models in Finance: A Survey Diffusion Models for Reinforcement Learning: A Survey

Reference 102

Resolution
unresolved
no resolver link, observed 2026-08-16T00:07:53.191949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:07:53.191949Z digest=sha256:9c5eaa4ba35a4f81d46831b3c6b05a2585551c1a124998fe4a5cede0ec8048de