Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2411.01302.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:35:12.088553Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T17:18:43.049717Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 16902dd1-f9ca-46cb-999f-a6f9e1d98eef · inbound
Beyond separability: convergence rate of vanishing viscosity approximations to mean field games via FBSDE stability Regret of exploratory policy improvement and $q$-learning
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ee90eef-f7da-4936-9625-53dedca82274 · inbound
Upper and lower bounds for local Lipschitz stability of Bayesian posteriors Regret of exploratory policy improvement and $q$-learning
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8fcc4cb7-4208-41ba-83b1-a123ac4ebeb2 · inbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Regret of exploratory policy improvement and $q$-learning
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa758c4c-9bfa-4c24-a10a-f3f9ce6b2f26 · inbound
Continuous-time reinforcement learning for optimal switching over multiple regimes Regret of exploratory policy improvement and $q$-learning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f992bf67-3ff9-4351-92c9-233dce063556 · inbound
ART for Diffusion Sampling: A Reinforcement Learning Approach to Timestep Schedule Regret of exploratory policy improvement and $q$-learning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0769ca52-5e5d-4178-af31-555dd1da43d9 · inbound
Conditional Diffusion Guidance under Hard Constraint: A Stochastic Analysis Approach Regret of exploratory policy improvement and $q$-learning
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d49273f0-8ba8-43eb-bdab-8fa3f4fd8fbc · inbound
Amortized Guidance for Image Inpainting with Pretrained Diffusion Models Regret of exploratory policy improvement and $q$-learning
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d3f5a2e7-5589-424f-b2e4-a564f8992746 · inbound
ART for Diffusion Sampling: Continuous-Time Control and Actor-Critic Learning Regret of exploratory policy improvement and $q$-learning
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 74239336-d89e-431c-8d7b-b46b887269a7 · inbound
ART for Diffusion Sampling: Continuous-Time Control and Actor-Critic Learning Regret of exploratory policy improvement and $q$-learning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e94006f6-03e4-4e2e-963a-39107ecb0fa9 · inbound
A Continuous-Time Reinforcement Learning Framework for Fine-Tuning Discrete Diffusion Models Regret of exploratory policy improvement and $q$-learning
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c764ff7-65ef-41ed-ac5a-2d826cecf34a · inbound
Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Regret of exploratory policy improvement and $q$-learning
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.