Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T21:33:38.321579Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 5 inbound Pith citation observations for arXiv:2507.00358.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T21:33:38.321579Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-01T11:22:20.635669Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T17:18:43.032059Z
49 of 49 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 53c011f9-07ed-4843-baeb-28990310a5df · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Sublinear Regret for a Class of Continuous-Time Linear-Quadratic Reinforcement Learning Problems
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a416a5af-1070-4a81-87d0-f759c5104c62 · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Unresolved cited work
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3319b20d-c984-487a-8f7e-7b9980ff7326 · outbound
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a0260570-6ce0-4368-bfca-79e54bd1e914 · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Stochastic linear quadratic regulators with indefinite control weight costs,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9a56b895-bbdd-4b7b-9b39-ad595502bcb5 · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Well-posedness and attainability of indefinite stochastic linear quadratic control in infinite time horizon,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 56b4a220-ed46-4e8d-bf98-140e0303d25c · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Linear matrix inequalities, riccati equations, and indefinite stochastic linear quadratic controls,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6a310869-e25f-46c5-b6e2-e8ea4743ab12 · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Solvability and asymptotic behavior of generalized riccati equations arising in indefinite stochastic lq controls,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5304aed6-de98-4bb2-9c9f-dd2410cfa39f · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems A primal-dual semi-definite programming approach to linear quadratic control,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c4910c53-0ea6-4635-b660-f8dc3fdc3498 · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Stabilization control for itˆ o stochastic system with indefinite state and control weight costs,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0cc3dcd9-378d-4d3e-a63d-a3fe1fe2e24e · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Optimal regulators for a class of nonlinear stochastic systems,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 224c8670-cdc7-48ce-bd8a-0e48d8985ff7 · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Indefinite linear-quadratic optimal control of mean-field stochas- tic differential equation with jump diffusion: an equivalent cost functional method,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9ffde277-e768-4d2b-a6fa-a3f77730792b · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Stochastic linear quadratic optimal control problems with regime- switching jumps in infinite horizon,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c30d62f0-e968-4f7c-925b-97b18f70c6c6 · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems On estimating the expected return on the market: An exploratory investiga- tion,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 22553b20-340f-4dc7-840b-ee89e39fa168 · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Unresolved cited work
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7081615b-f0b1-426c-86b4-783149d67788 · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Rustem and M
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 37632502-7169-4d11-845d-733f8f20b593 · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Unresolved cited work
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c75ffc6c-70b2-4f62-ae92-8ac9ea4ad5bf · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems A survey on intrinsic motivation in reinforcement learning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ec4e6e4-16a0-4143-a74f-b77f89cebd07 · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Formal theory of creativity, fun, and intrinsic motivation (1990–2010),
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 86e125e1-8cb6-47dd-acbd-2a38acffa9ff · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Surprise-Based Intrinsic Motivation for Deep Reinforcement Learning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 347fd813-d80b-41d9-a067-a87eebf867ef · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Curiosity-driven exploration by self- supervised prediction,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 84763f26-e404-4299-83e2-4f9271c8ff3d · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Large-Scale Study of Curiosity-Driven Learning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10d21d05-a695-4b4e-9f32-1cfd6bd8fa23 · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Exploration by Random Network Distillation
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9708eae-0a06-4e27-8b0a-39fdd77d889f · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Randomized prior functions for deep reinforcement learning,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bc9ef8ca-a6c4-49c3-b588-e076d4af1322 · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Fast active learning for pure exploration in reinforcement learning,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e0d6309e-1719-4bf5-a7d2-7193a8dae6a2 · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems # exploration: A study of count-based exploration for deep reinforcement learning,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 16a5722b-ccca-4208-bb58-b7412a6a7e36 · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Go-Explore: a New Approach for Hard-Exploration Problems
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d28035f0-aadb-4915-83d3-90f2e7e7cc3d · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems First return, then explore,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b25f14d2-0ef2-4fc4-98fc-5b3db13d6108 · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Leveraging Demonstrations for Deep Reinforcement Learning on Robotics Problems with Sparse Rewards
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f31a89c-dfa3-46ad-a787-31beb912227c · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems A comprehensive survey on safe reinforcement learning,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4a86cb93-5458-470a-acfa-5d08381dd836 · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Trial without Error: Towards Safe Reinforcement Learning via Human Intervention
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e800d42-c001-4948-8579-cb69a9048ba5 · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Policy gradient in continuous time,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 12a82e21-f152-4421-bd38-9659339fcf81 · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Making deep Q-learning methods robust to time dis- cretization,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9bf35128-1d6c-40f9-8dff-3e1a08431cdc · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Time discretization-invariant safe action repetition for policy gradient methods,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6a3f0064-a959-44be-bc37-060185985191 · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Optimal scheduling of entropy regularizer for continuous-time linear-quadratic reinforcement learning,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a8021655-0c66-493c-bb20-04aee29c2885 · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Indefinite stochastic riccati equations,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7accc150-d4ad-4931-80bc-e33c5cef8d0a · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Existence of solutions to a class of indefinite stochastic riccati equations,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ccc0b981-160e-4ae4-9d31-cb9faf4049c9 · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Reinforcement learning in continuous time and space: A stochastic control approach,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8f088e2d-74bb-468c-8c19-ee4d41f8cfcc · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Exploration-exploitation trade-off for continuous-time episodic reinforcement learning with linear-convex models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd631aed-b86c-4aaf-a250-0ecdc1143c55 · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Policy gradient and actor-critic learning in continuous time and space: Theory and algorithms,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 99628674-cdfd-489c-8807-1efccfe2d68d · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Policy evaluation and temporal-difference learning in continuous time and space: A martingale approach,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d2f9f109-02ae-4cbe-a0d1-57cad82d45d9 · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems A stochastic approximation method,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bddc39d1-6174-4d91-ba31-6f58c63fe5d9 · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Stochastic approximation,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 52b5835a-44b2-4069-ac9a-03d5f5a4f485 · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Unresolved cited work
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ddca0fb6-7710-483a-9b99-e3eade5a629b · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems An overview of stochastic approximation,
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ac6aff8d-77e2-43be-a347-3e08a1c4bc21 · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems A stochastic approximation algorithm with varying bounds,
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a05ae417-2113-4264-88f5-45e8b3888bc7 · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems General bounds and finite-time improvement for the Kiefer-Wolfowitz stochastic approximation algorithm,
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3d5c3919-a8a0-4fce-80c2-6332c5d319a7 · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems A convergence theorem for non negative almost supermartin- gales and some applications,
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9b5aaec9-6d3e-4734-8818-9053c94d0707 · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Unresolved cited work
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8fcc4cb7-4208-41ba-83b1-a123ac4ebeb2 · outbound
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Regret of exploratory policy improvement and $q$-learning
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec665d63-6095-424a-ab99-5b351f53c8d3 · inbound
ART for Diffusion Sampling: A Reinforcement Learning Approach to Timestep Schedule Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9818ebf3-6f70-41c9-8d27-84dac68bed2a · inbound
Amortized Guidance for Image Inpainting with Pretrained Diffusion Models Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation dd9a3f69-c1e9-47d8-9199-dfe392315231 · inbound
ART for Diffusion Sampling: Continuous-Time Control and Actor-Critic Learning Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 10076fa3-b771-462d-89e9-b6081230cc75 · inbound
ART for Diffusion Sampling: Continuous-Time Control and Actor-Critic Learning Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a8504d5-6aa5-4220-9480-8210bd94ac96 · inbound
Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.