Pith. sign in

Paper Citation Record · LEDGER

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning

As of 22 August 2026, this Paper Citation Record lists 79 of 79 outbound references and 9 inbound Pith citation observations for arXiv:2502.02316.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.02316 v2

Coverage vector

measured 79 of 79 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T12:38:14.369887Z

measured 88 of 88 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T19:28:41.716079Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T08:59:43.293825Z

Reference resolution

79 of 79 outbound references displayed

  • verified exact0
  • verified fuzzy32
  • unresolved45
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 203afd33-f713-4722-bd17-58eab248641f · outbound

This paper cites write newline.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.087464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.087464Z digest=sha256:08f27a4bc539e2d2a01a631d6b91b7c6c8ba06330acad58e28d9678984d574a5

Observation 0b88acfd-b643-49d0-81f6-9cd71d0fb879 · outbound

This paper cites an unresolved cited work.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-09T12:38:15.182823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.093314Z digest=sha256:4703f9177d1207d687e7c1343368f300ba59dc5fbc35a8d9249f1d1dbee31f0f

Observation ee4f3bbf-ae68-44f5-bb53-dff5915592a1 · outbound

This paper cites S., Courville, A.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning S., Courville, A

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.097793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.097793Z digest=sha256:88d76828cf62999f6d9468095924430db6d4d2cdf898229a7793797bdb7add57

Observation 3a48d042-edd4-4ae0-9f7f-5076055408ed · outbound

This paper cites Iterated Denoising Energy Matching for Sampling from Boltzmann Densities.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Iterated Denoising Energy Matching for Sampling from Boltzmann Densities

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.101529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.101529Z digest=sha256:13ac81be4c4583af38be1fe2f9f1cf2ddd8e2336bfe42486748fef6d841e1a6d

Observation 1ec41a2f-6a9f-44b1-9967-682e6019bcb6 · outbound

This paper cites an unresolved cited work.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.105807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.105807Z digest=sha256:542a9997702b1859c9c1942d4d20b8d30ad3c78ab58e07e2aa7af57cd9e40024

Observation acf952e3-931d-467a-b8c5-3b34a8ce8f42 · outbound

This paper cites Efficient gradient-free variational inference using policy search.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Efficient gradient-free variational inference using policy search

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:38:15.154632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.109558Z digest=sha256:4b484f94eb49f837896b6754ce7eba1fab5043b69460d58f895315b6c7624eaf

Observation 0a0eb642-e926-44f4-b258-70818517bb84 · outbound

This paper cites G., Dabney, W., and Munos, R.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning G., Dabney, W., and Munos, R

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.113362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.113362Z digest=sha256:79b66ee96fa65b85b43e5c5d1b53183e3e63317cbb7e326ccc48afb7e53b2055

Observation 1e46971f-fe93-4dee-908a-e3e3943f2acc · outbound

This paper cites An optimal control perspective on diffusion-based generative modeling.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning An optimal control perspective on diffusion-based generative modeling

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:38:15.134988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.117266Z digest=sha256:a83ec18055df5bd539b7f6924e5cc408579210750275d554c428eec09c65f93f

Observation c8944175-5d6f-43dd-9944-97d74e1ce9ff · outbound

This paper cites Crossq: Batch normalization in deep reinforcement learning for greater sample efficiency and simplicity.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Crossq: Batch normalization in deep reinforcement learning for greater sample efficiency and simplicity

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:38:15.123699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.120792Z digest=sha256:512d0b6bb914228d34cf97b272e6b1d0e2da2438029b829e76ce611733ccf82b

Observation 24adc66f-f8fd-4d60-9a19-b4e0607a3011 · outbound

This paper cites Underdamped diffusion bridges with applications to sampling.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Underdamped diffusion bridges with applications to sampling

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:38:15.111698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.124244Z digest=sha256:6b796a656af8070a646c9c695a8d8de077701d5e527e0f227f411f096a2ca779

Observation 33f623d8-5a90-49f1-b532-dcf9af660a1f · outbound

This paper cites End-to-end learning of gaussian mixture priors for diffusion sampler.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning End-to-end learning of gaussian mixture priors for diffusion sampler

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:38:15.100057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.128205Z digest=sha256:2c1a7a67a20d7ffc681b207bf34847f03463e8b697c73120b495a09e10b4f6ae

Observation 6df01a18-2c97-49ef-bbbd-a2ed5019a634 · outbound

This paper cites Openai gym, 2016.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Openai gym, 2016

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:38:15.088157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.131657Z digest=sha256:a4c5c3511d604567be35fa71c4a9a6c1dfdd0865ffb15001d77035f3c3274705

Observation ce76b231-caf9-4a2b-8ba3-cd28ad1c8ff1 · outbound

This paper cites Tightness without Counterexamples: A New Approach and New Results for Prophet Inequalities.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Tightness without Counterexamples: A New Approach and New Results for Prophet Inequalities

Reference 13

Resolution
metadata mismatch
local_arxiv, observed 2026-08-09T12:38:14.567153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.135117Z digest=sha256:3d13cb6da315556df5d5c4501d297068001013d00242cd6a0fd7b7058bbe4b37

Observation af2db694-a77b-4c8a-be46-091df0cbb3a6 · outbound

This paper cites Offline reinforcement learning via high-fidelity generative behavior modeling.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Offline reinforcement learning via high-fidelity generative behavior modeling

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:38:15.075686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.139164Z digest=sha256:958a9023ad3b9ffd28d96c2b1fe7e967486436d9a8fe4ae312f3de3509ad1252

Observation 463b306f-e696-4d5d-a00e-8141d85cd41a · outbound

This paper cites Sequential Controlled Langevin Diffusions.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Sequential Controlled Langevin Diffusions

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.142543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.142543Z digest=sha256:d7b5bc8a0cbd105bc63c9b352c621ad036fa4c159974c9016f5f1897ed3b61c2

Observation 656a7fcb-641d-49db-a591-b4ae7eeadf75 · outbound

This paper cites Sequential controlled langevin diffusions.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Sequential controlled langevin diffusions

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:38:15.064866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.145887Z digest=sha256:2710554a03846c3dae120e1d6097ff5e23e51b1286432be2dfe8ee366368f277

Observation af58e749-21bc-44d5-9cb4-7d66cd8c2d97 · outbound

This paper cites Diffusion policy: Visuomotor policy learning via action diffusion.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Diffusion policy: Visuomotor policy learning via action diffusion

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.148959Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.148959Z digest=sha256:2864ed34fa6cefa9c8cdbac4a71190136d5568f096ddbbaa3f817e0d51f5790a

Observation 8be1a765-d9d0-4d70-8b05-ad06f40602ae · outbound

This paper cites an unresolved cited work.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.151776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.151776Z digest=sha256:a4cdf7270dac1b3b32d685945e6dde2790c8a35cd1a05f8aef6dc36623329f9f

Observation a7279a3d-3df2-4438-bbbe-ad2886b1b3e9 · outbound

This paper cites A stochastic control approach to reciprocal diffusion processes.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning A stochastic control approach to reciprocal diffusion processes

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:38:15.038657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.154882Z digest=sha256:556966f7036b1aeee231633866a0f54557f6b2f6213b4c383dfd73c60bc55dd3

Observation e5e22f54-b8a5-40ec-8f10-5b7533763a68 · outbound

This paper cites Sequential monte carlo samplers.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Sequential monte carlo samplers

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.157903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.157903Z digest=sha256:e874e438d4babdc1c310b11dce7e46e870670ded702521efdfe8d2599feb875e

Observation 44b83fed-ed29-42bb-9776-08a3c386a602 · outbound

This paper cites Diffusion-based reinforcement learning via q-weighted variational policy optimization.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Diffusion-based reinforcement learning via q-weighted variational policy optimization

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:38:15.019019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.160824Z digest=sha256:b97122e4674420253ed111fddf4f10d16ef4b4f283f639b5901dc5da8f9d8749

Observation 38113b63-1b87-44a1-9ff4-3a91fb5f4048 · outbound

This paper cites and Jin, C.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning and Jin, C

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:38:15.000021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.163769Z digest=sha256:d9df370abca9b69c7125ee7853a3c458972fbe87f38fe7e01add86efa4a54b39

Observation 4e73b97f-d16e-44bc-8199-27bacc96820a · outbound

This paper cites G., and Strathmann, H.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning G., and Strathmann, H

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:38:14.988685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.167254Z digest=sha256:f6761b5d4c0466caa07a4887cb4e263644e3db0798486a621eacaf1045c53b11

Observation 7c606447-b7d7-4c41-a6fb-043a47e367a8 · outbound

This paper cites Diffusion actor-critic: Formulating constrained policy iteration as diffusion noise regression for offline reinforcement learning.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Diffusion actor-critic: Formulating constrained policy iteration as diffusion noise regression for offline reinforcement learning

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:38:14.977582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.170851Z digest=sha256:40e51e4c2175c7b1d40024d0fea34e243a928adb0c3a1ff7eb9c807eda8ac921

Observation 9513d163-9d7d-4f26-9adf-765f629775ab · outbound

This paper cites and Domke, J.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning and Domke, J

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.174440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.174440Z digest=sha256:4180978147fbd1ff31da166039cb532ad808190df5af41a33203cf5020801f5e

Observation f0559be9-0003-4eec-ac08-c22b2a65e4e7 · outbound

This paper cites Reinforcement learning with deep energy-based policies.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Reinforcement learning with deep energy-based policies

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.178080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.178080Z digest=sha256:0ba14525009ffae6f3a5d952778d5b2538fccef7ade93a17ab177d80d20ee6a1

Observation 27bf4433-5563-4170-8138-df713dab2298 · outbound

This paper cites Latent space policies for hierarchical reinforcement learning.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Latent space policies for hierarchical reinforcement learning

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:38:14.953089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.182904Z digest=sha256:eed67a2939ecc16e58e55d3ec90c1db32f72dd0bb30fd24219d31b1b201552a3

Observation 9d579252-3905-44b2-9baa-656604c177a0 · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:38:14.942335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.186408Z digest=sha256:ce29c48966ce3ea18171cca04eeb35512354afd30159ec4fabc41f95e8d65457

Observation e8e530f3-f95b-423b-9935-697a180111ae · outbound

This paper cites Soft Actor-Critic Algorithms and Applications.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Soft Actor-Critic Algorithms and Applications

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.189969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.189969Z digest=sha256:96a805d9fa0a72e789e68ce0087e292a74213026030d3a0cba45ad36e74bf920

Observation 99b2ad79-7c1e-41aa-bb7f-361477653df9 · outbound

This paper cites IDQL: Implicit Q-Learning as an Actor-Critic Method with Diffusion Policies.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning IDQL: Implicit Q-Learning as an Actor-Critic Method with Diffusion Policies

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.193834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.193834Z digest=sha256:aa5575578a7269e7b4d403f6a6cb433b2ad721520872c7c35c535f469a7f96d5

Observation c395e1eb-532b-4fd5-b9a9-98c9649bfa0b · outbound

This paper cites an unresolved cited work.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-09T12:38:14.930480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.197642Z digest=sha256:95137bceaf6e6c7d2fe2092c485d46c1213a5300822b69e6466c28d5feca4257

Observation b813972e-5a86-417c-8a3a-17c3f56d04a9 · outbound

This paper cites Denoising diffusion probabilistic models.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Denoising diffusion probabilistic models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.201237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.201237Z digest=sha256:5cf7e6d9a633856b79f0b42e2c52690209c66f36abbdcbef5aa966ba3e9d4fe3

Observation c4cd67f8-4626-4693-9bec-ce5961d57d34 · outbound

This paper cites Schr{\"o}dinger-F{\"o}llmer Sampler: Sampling without Ergodicity.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Schr{\"o}dinger-F{\"o}llmer Sampler: Sampling without Ergodicity

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.204618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.204618Z digest=sha256:26a3f1d17d78ab299783ddbf608e4b77dab430d92107cf8d2eca34a4829f3034

Observation cc8de289-3168-40d6-87da-1bbb62609350 · outbound

This paper cites and Dayan, P.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning and Dayan, P

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.208548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.208548Z digest=sha256:c9ab4af131730eaca2117056216537009799ddf23b5d7d0a226115b224ff5e68

Observation b30079e6-bd0e-4bdc-b0fd-08b46ee2c299 · outbound

This paper cites N., and Precup, D.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning N., and Precup, D

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:38:14.906381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.212244Z digest=sha256:098eca400fc789f5ea54299834b96357a7692e54b72be0816c4396f8aeef6865

Observation e7811df2-12a0-49fb-85f8-da4e426314a7 · outbound

This paper cites Planning with diffusion for flexible behavior synthesis.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Planning with diffusion for flexible behavior synthesis

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.215741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.215741Z digest=sha256:cc287693b7984e8d023f046ff4837b4abeb55cc5c8aa6715dadd3bf1c822f843

Observation 0db103f9-73cf-4dbe-b5a4-f6b31590b611 · outbound

This paper cites Efficient diffusion policies for offline reinforcement learning.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Efficient diffusion policies for offline reinforcement learning

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:38:14.888260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.219173Z digest=sha256:2fee4f4f6dd6f56fafbeb63374660d6e5ffd215fe04a50b8ae28965e1b12174d

Observation 4a3676ed-ea1c-434d-bd8f-eb1b1a52240f · outbound

This paper cites Elucidating the design space of diffusion-based generative models.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Elucidating the design space of diffusion-based generative models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.222831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.222831Z digest=sha256:1b13cfcdf091e32c3116946e432331eaec4593eef6e1c8c373b45409298f1ea7

Observation 619b962a-90db-43ce-af48-88db610c544a · outbound

This paper cites Auto-Encoding Variational Bayes.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Auto-Encoding Variational Bayes

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.226467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.226467Z digest=sha256:1190e3ef1ae310df4c1f7694e08568af874e08f43556a49b25d978c8b83ae7a6

Observation 7d12ff7c-cc09-4caf-9e15-e7eae1c7b510 · outbound

This paper cites Batch reinforcement learning.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Batch reinforcement learning

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:38:14.870021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.230327Z digest=sha256:4aef379d187cfdcf417031fe56112cda22a7c56177f97a169ea6b4cab824350d

Observation a4f75e3d-f77f-496d-b070-bb2927c87969 · outbound

This paper cites Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.234145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.234145Z digest=sha256:9d7fb36092b32757c4b9d20648cac64303f1a86ac325c2438d1b99c8ea789dc5

Observation f2638d14-02d0-4a9c-a77e-617b1d26d5a7 · outbound

This paper cites TOP - ERL : Transformer-based off-policy episodic reinforcement learning.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning TOP - ERL : Transformer-based off-policy episodic reinforcement learning

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:38:14.858848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.238193Z digest=sha256:c3caa1fe16974d8ecdff31b9d90b28e3eb3aa0fe8c01628386386a6259fd9028

Observation 4809b0cd-c429-4c4f-b974-d3063e183d79 · outbound

This paper cites Learning Multimodal Behaviors from Scratch with Diffusion Policy Gradient.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Learning Multimodal Behaviors from Scratch with Diffusion Policy Gradient

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.242058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.242058Z digest=sha256:a92b0b0a812d5222c5689c0312693561881a895158eef6ecf5c142d6df934218

Observation 87c1541d-f4e7-4906-affb-894f06ebfd19 · outbound

This paper cites Contrastive energy prediction for exact energy-guided diffusion sampling in offline reinforcement learning.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Contrastive energy prediction for exact energy-guided diffusion sampling in offline reinforcement learning

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.245151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.245151Z digest=sha256:28da7d89500ddd1be14587483e7a471807f11e606473fb25eefb8b98d9eeb132

Observation dadef05a-23ca-40c2-a96b-e75e2feae13f · outbound

This paper cites K., S nderby, S.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning K., S nderby, S

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:38:14.839088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.248289Z digest=sha256:ebf6198c3f8dd7420f9681ba72edd805b2cc6b8f99a209faae5443f44fac6ab0

Observation 3e71aaa3-867c-479a-87cb-e80a15e5436f · outbound

This paper cites Diffusion-dice: In-sample diffusion guidance for offline reinforcement learning.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Diffusion-dice: In-sample diffusion guidance for offline reinforcement learning

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:38:14.827504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.251712Z digest=sha256:ae6ec13a037f8c37297634f26e98bc9bbd5a73ea2c29ce8af41ee775d4ea4ea6

Observation ee81219c-e089-40aa-86c3-9eda621f0f89 · outbound

This paper cites S 2 ac: Energy-based reinforcement learning with stein soft actor critic.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning S 2 ac: Energy-based reinforcement learning with stein soft actor critic

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:38:14.817002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.254946Z digest=sha256:eeaf3631743c513d373c35e3c708b71685b8c0f0e6d3393037ec62fca3a035a1

Observation ec591d77-e5aa-410d-9103-9717ef832ab5 · outbound

This paper cites and Cygan, M.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning and Cygan, M

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.258104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.258104Z digest=sha256:28e7a1ebf77d84bea3bbfbd5c70ad172c7508b01605b77c02d025d97b43f3ad3

Observation 21ce8d20-963b-4570-b0b2-b1edcab3dbad · outbound

This paper cites Bigger, regularized, optimistic: scaling for compute and sample efficient continuous control.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Bigger, regularized, optimistic: scaling for compute and sample efficient continuous control

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:38:14.799394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.261037Z digest=sha256:f46679dccd65e293518dcb55511e39e57156c6707d47785fe09e8796e60cf241

Observation 0d0cf4b9-2ca5-4469-9217-ad12a71951db · outbound

This paper cites Dynamical theories of Brownian motion, volume 101.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Dynamical theories of Brownian motion, volume 101

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:38:14.788288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.264042Z digest=sha256:8c497a64f2be1144d9f54cd298e60a93c5570da661f2481030e95f789b5fb8e9

Observation 3e64df8a-1914-47a3-ae76-c34cd0f92849 · outbound

This paper cites A unified view of entropy-regularized Markov decision processes.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning A unified view of entropy-regularized Markov decision processes

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.267814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.267814Z digest=sha256:c04ca4bfa41a5bccf6e54f2138f06b0fafa411caf8272e2fbf3ee3b1d7b40b41

Observation faf27320-67d3-4d0b-9d64-e93315eb1116 · outbound

This paper cites The primacy bias in deep reinforcement learning.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning The primacy bias in deep reinforcement learning

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.271660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.271660Z digest=sha256:5b1388540a661f8ddaf2ad041284d8d2946a544da5e8328ec0ef969eb4014f73

Observation 36c8d1e6-630b-49e8-904c-4fd4f53c5b79 · outbound

This paper cites Learned Reference-based Diffusion Sampling for multi-modal distributions.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Learned Reference-based Diffusion Sampling for multi-modal distributions

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.275215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.275215Z digest=sha256:dd1b2d6592314d4bcaebfab7279a1805a5681efbd5d7a1d6d13666a87f025192

Observation 79dc8dc3-bef3-450e-8809-6ebb33be4dd1 · outbound

This paper cites Transport meets variational inference: Controlled monte carlo diffusions.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Transport meets variational inference: Controlled monte carlo diffusions

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.278880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.278880Z digest=sha256:0d267c43c30c79be27e0334412125e2fce2fd75c4437a6d54f08ae7b4e21b781

Observation 65b5d15c-9889-4cad-a67d-59fb089fc7c4 · outbound

This paper cites Learning a diffusion model policy from rewards via q-score matching.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Learning a diffusion model policy from rewards via q-score matching

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:38:14.761370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.282288Z digest=sha256:6597e77926e79f4a4f1296f7d057aa665a6effecbcd9636ae059b8016e389aa9

Observation a09c9356-1294-41c0-aced-32c735fcf081 · outbound

This paper cites Hierarchical variational models.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Hierarchical variational models

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:38:14.748420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.286001Z digest=sha256:886d4673067e6d636262c7b5244d24018b3d92e39a176097f6747c6658e5beb9

Observation 48824140-2a73-4abc-8ac3-5b49f513846a · outbound

This paper cites Goal conditioned imitation learning using score-based diffusion policies.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Goal conditioned imitation learning using score-based diffusion policies

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:38:14.736983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.289502Z digest=sha256:fc7e514c337eec8065d934fb20e81623e115f1c68f87076e3474ba9865993510

Observation bf533dbe-5035-4d60-95d8-7fc5f0b1272a · outbound

This paper cites and Berner, J.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning and Berner, J

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:38:14.725710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.293053Z digest=sha256:17481ef59313192c9bbbfab04ca0c4a56fb53745e272935bfb6a09d4f1f6a83b

Observation 8066ba5f-6ffb-4f13-aae4-5d5bfa5373fd · outbound

This paper cites and Solin, A.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning and Solin, A

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.296976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.296976Z digest=sha256:3fc7ddb1f0a0fe8d8237903bb1448fbe9be6af253080bf19a1ec4ae03d29ff98

Observation 49810f60-65f7-42f6-be15-4089eae599a3 · outbound

This paper cites Deep unsupervised learning using nonequilibrium thermodynamics.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Deep unsupervised learning using nonequilibrium thermodynamics

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.301800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.301800Z digest=sha256:c06a12bd8472e74f06552fdf26145851c09fff9fa4260a5b671983a08b65b19a

Observation 4a897596-edc0-4c6f-a83a-80b65c8d7af8 · outbound

This paper cites P., Kumar, A., Ermon, S., and Poole, B.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning P., Kumar, A., Ermon, S., and Poole, B

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.305552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.305552Z digest=sha256:5a2b03c6c3d358b11e3a56e0924a157e486a7cffefc1d46c00ea2880cef815d0

Observation 8a635f3c-89eb-4d7e-8911-58e4c8084279 · outbound

This paper cites an unresolved cited work.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Unresolved cited work

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.309171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.309171Z digest=sha256:c4c4361f3df67c4b7523cc823e195ba825c4c9ab74873a468046cce21955b702

Observation 185596b5-8064-407d-b48a-2920076db62b · outbound

This paper cites Robot trajectory optimization using approximate inference.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Robot trajectory optimization using approximate inference

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.312599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.312599Z digest=sha256:c042e16a6b1cc8aae087fb18efacde019bed3905addef0ab9a9b71a05ce9350f

Observation 9763709c-6b80-4d9c-8840-3b43a1535995 · outbound

This paper cites The Variational Gaussian Process.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning The Variational Gaussian Process

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.316160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.316160Z digest=sha256:5a3a381908eca6306febc0d95f9b3774fe76a7e2741efbb8d883e7907fc1b0bf

Observation ff43e272-9320-469a-8ee0-6c9e1396a6f4 · outbound

This paper cites dm\_control: Software and tasks for continuous control.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning dm\_control: Software and tasks for continuous control

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.320118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.320118Z digest=sha256:c8dc41a5e83c5e483281b6dadb1f43bdc6e286dbca9f979edc811730ca194372

Observation 69b67f08-7650-40ed-a1b9-5f1687fedf9b · outbound

This paper cites and Raginsky, M.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning and Raginsky, M

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.323715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.323715Z digest=sha256:ff6c98ae705e3f6b87d1764e9a1bf6853633da946b62c55db8a0b1bae20f154a

Observation e46bd896-c502-4a9e-abb2-6eaf23ffb790 · outbound

This paper cites S., and Doucet, A.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning S., and Doucet, A

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:38:14.665066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.327292Z digest=sha256:7c64843986a8d8fc808f95f905926fb306458b6d9980833a502b99acb069e351

Observation 1c194d97-1288-425b-aac3-d239516cb03e · outbound

This paper cites u sken, N. Bayesian learning via neural schr \.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning u sken, N. Bayesian learning via neural schr \

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:38:14.653971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.331007Z digest=sha256:3a2ae9921a56c7a53f53ff008b4a141c7a1fee2ab8ddc9070407213dd6da6f3b

Observation 3f4c3ba9-d9a9-40e2-a540-ad9028288e9b · outbound

This paper cites A connection between score matching and denoising autoencoders.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning A connection between score matching and denoising autoencoders

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.334714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.334714Z digest=sha256:cb563482b03ba1260f7162dfa6952669d7af423329485f360342c99a3e4b0990

Observation a017ff77-6f1f-454f-b3f9-96bd28d87566 · outbound

This paper cites Learning to Draw Samples: With Application to Amortized MLE for Generative Adversarial Learning.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Learning to Draw Samples: With Application to Amortized MLE for Generative Adversarial Learning

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.338402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.338402Z digest=sha256:cd27a19420aedb26ff2ef9221dbe246bef04331e6bd349cf0759a3bda47d6e87

Observation d88d61d6-91e0-4f62-bfc3-e70fa9569d23 · outbound

This paper cites an unresolved cited work.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Unresolved cited work

Reference 71

Resolution
unresolved
raw_fallback, observed 2026-08-09T12:38:14.634493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.342096Z digest=sha256:f31805ecea0f3ad1db6dafa6adab271023cd740dbe8887857e43434af78e05ff

Observation 9a98b497-f72f-4c08-a67e-62799cf7fc74 · outbound

This paper cites J., and Zhou, M.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning J., and Zhou, M

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:38:14.623216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.345551Z digest=sha256:b7073477d6550073ca85d78775bd5fa0c63ff5d4d5e23add49793644464f26f3

Observation 27df332f-43c7-4b99-b0d4-4a8734bb82af · outbound

This paper cites and Teh, Y.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning and Teh, Y

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.349023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.349023Z digest=sha256:0ff2a1cd3e43e1e253a17c52b47c763294478125b7d20c25f65d7cdb3107065d

Observation 7c3a774b-37f9-490d-853f-abe83da17fb7 · outbound

This paper cites Policy Representation via Diffusion Probability Model for Reinforcement Learning.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Policy Representation via Diffusion Probability Model for Reinforcement Learning

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.352435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.352435Z digest=sha256:d5fe7123d1fe26fa7cb5cd7e0b8c0ff11a5b29aca26d5a0bf262c3ec5b87ede6

Observation c586b889-f9f7-41f2-bae9-420dbd3113fb · outbound

This paper cites Diffusion Generative Flow Samplers: Improving learning signals through partial trajectory optimization.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Diffusion Generative Flow Samplers: Improving learning signals through partial trajectory optimization

Reference 75

Resolution
metadata mismatch
local_arxiv, observed 2026-08-09T12:38:14.419170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.356711Z digest=sha256:5b9eef52e6ae75ca95ae44efffcf8a01deec5d0eae246e92bc0400c16772760f

Observation d08cbb27-c32e-473a-908d-8bf8a07390c5 · outbound

This paper cites Path Integral Sampler: a stochastic control approach for sampling.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Path Integral Sampler: a stochastic control approach for sampling

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.360227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.360227Z digest=sha256:ad7ed63fb7302a274c3f855234271aba8da2725e4f38af787763b95c20c4cc33

Observation b879a459-e67a-4782-b1f2-ca061de9297b · outbound

This paper cites Variational distillation of diffusion policies into mixture of experts.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Variational distillation of diffusion policies into mixture of experts

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T12:38:14.604750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-09T12:38:14.363751Z digest=sha256:3b0881a1b606e5660740f3f98520ce526fa56e93ecea0ec2f889cecf9c7d1932

Observation 1f801739-e362-4d39-832d-213bf9037dcb · outbound

This paper cites an unresolved cited work.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning Unresolved cited work

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.366824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.366824Z digest=sha256:ce7add4c897f2516998401ea0f57c08c472c7cc7376bfc7f5297e6563f6b4957

Observation fb8be11c-3caf-4f86-8891-52385aa1d29e · outbound

This paper cites D., Maas, A.

DIME:Diffusion-Based Maximum Entropy Reinforcement Learning D., Maas, A

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-09T12:38:14.369887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:38:14.369887Z digest=sha256:3912bae96a48b3a33ca367173be09f259d775365bfbd8a8dbf7023f9f024b3c4

Pith citing papers

Observation 716b42f4-aae2-4756-8cff-99c54aa377b6 · inbound

Efficient Online Reinforcement Learning for Diffusion Policy cites this paper.

Efficient Online Reinforcement Learning for Diffusion Policy DIME:Diffusion-Based Maximum Entropy Reinforcement Learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-09T19:28:41.716079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:28:41.716079Z digest=sha256:872ad9bff3fe97346ff30468a36c0f49c9217bcff753d97bd4ed73bbbc47d6d5

Observation 4f41f3bf-524c-4d7d-b5ec-4ce13f26c927 · inbound

Exploratory Diffusion Model for Unsupervised Reinforcement Learning cites this paper.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning DIME:Diffusion-Based Maximum Entropy Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.155791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.155791Z digest=sha256:4b308a16bb3f5def35a5c74c5ab9bb25c2d264aad331dfef00811b4b0946c275

Observation 196e2f66-e015-4103-afdd-e02e0063ebff · inbound

Towards Adaptive External Communication in Autonomous Vehicles: A Conceptual Design Framework cites this paper.

Towards Adaptive External Communication in Autonomous Vehicles: A Conceptual Design Framework DIME:Diffusion-Based Maximum Entropy Reinforcement Learning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T19:28:32.623166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:28:32.623166Z digest=sha256:a842e2ffd206b9ff91bac4e611b3377101d2fab7f16bac47ddd9cb6d12b8a0e2

Observation 8d1c6be2-7ef3-4375-abe0-0e9490bfdf28 · inbound

Reinforcement Learning with Discrete Diffusion Policies for Combinatorial Action Spaces cites this paper.

Reinforcement Learning with Discrete Diffusion Policies for Combinatorial Action Spaces DIME:Diffusion-Based Maximum Entropy Reinforcement Learning

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-21T21:15:38.660604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-21T21:14:54.053177Z digest=sha256:8c71684eed4f8881a24dbaf321c88ba633ace205f59079607bb53fa389a172aa

Observation 05f10b2f-7194-469d-bc18-beb18aa47576 · inbound

Diffusion-Augmented Markov Decision Processes for Maximum Entropy Reinforcement Learning cites this paper.

Diffusion-Augmented Markov Decision Processes for Maximum Entropy Reinforcement Learning DIME:Diffusion-Based Maximum Entropy Reinforcement Learning

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-03T19:12:08.549060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T19:12:08.549060Z digest=sha256:27cbf98ca7b636e08deada980e207b9d324a20e605572925c11949552dc4acaf

Observation 381c37d3-3f20-4a73-ad9e-f78ea0c83536 · inbound

What Does Flow Matching Bring To TD Learning? cites this paper.

What Does Flow Matching Bring To TD Learning? DIME:Diffusion-Based Maximum Entropy Reinforcement Learning

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-15T16:36:17.860115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-15T16:32:29.432272Z digest=sha256:2e4e3114ded59cb8cb308f2bae6162e5ab667342ddeb94cb9e48a0a2c7f89932

Observation 7cac92bc-43fa-43a4-be54-8d13bcb37acb · inbound

GeMPO: Generalized Measure Matching for Online Diffusion Reinforcement Learning cites this paper.

GeMPO: Generalized Measure Matching for Online Diffusion Reinforcement Learning DIME:Diffusion-Based Maximum Entropy Reinforcement Learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-14T23:47:45.866615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T23:47:45.866615Z digest=sha256:c6a56055116b96708d4dcb9c1baebab3facc90a2055c253f6c378306618abbb0

Observation ced346c7-6915-4d10-92b2-1a7f1e1466dc · inbound

GenPO++: Generative Policy Optimization with Jacobian-free Likelihood Ratios cites this paper.

GenPO++: Generative Policy Optimization with Jacobian-free Likelihood Ratios DIME:Diffusion-Based Maximum Entropy Reinforcement Learning

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-02T16:27:08.518956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-27T22:47:10.062975Z digest=sha256:7f782028488fbce72f4c154395c62bbea9d54536d5f9b354e55ef5dc3d235139

Observation 7de4cdf8-7b7b-45db-a400-1f8dc992d368 · inbound

Scalable Maximum Entropy Reinforcement Learning for Diffusion Policies via Adjoint Matching cites this paper.

Scalable Maximum Entropy Reinforcement Learning for Diffusion Policies via Adjoint Matching DIME:Diffusion-Based Maximum Entropy Reinforcement Learning

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-04T08:59:43.295757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-26T10:38:05.185415Z digest=sha256:0d4fb2a439f85757497e01662d9554192673c37d6e415f3cc0056e0c823e588f