Pith. sign in

Paper Citation Record · LEDGER

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems

As of 8 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 5 inbound Pith citation observations for arXiv:2507.00358.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.00358 v2

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:33:38.321579Z

measured 54 of 54 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T11:22:20.635669Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T17:18:43.032059Z

Reference resolution

49 of 49 outbound references displayed

  • verified exact0
  • verified fuzzy34
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 53c011f9-07ed-4843-baeb-28990310a5df · outbound

This paper cites Sublinear Regret for a Class of Continuous-Time Linear-Quadratic Reinforcement Learning Problems.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Sublinear Regret for a Class of Continuous-Time Linear-Quadratic Reinforcement Learning Problems

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:32.491006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:32.491006Z digest=sha256:d5fa8d76fd09135a9198056345e8ff001f6c53b59d627e13fdc031b5d9a933c1

Observation a416a5af-1070-4a81-87d0-f759c5104c62 · outbound

This paper cites an unresolved cited work.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:33:50.572667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:32.598128Z digest=sha256:4419630e4de9c5e417a7e6f92130db2f61818f9c8f3ad6ee6ddac19377ece4f8

Observation 3319b20d-c984-487a-8f7e-7b9980ff7326 · outbound

This paper cites Yong and X.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Yong and X

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:50.346689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:32.682350Z digest=sha256:e46388b4d44e62a2d0bad9b5aa47f1b72597b335c9becb5ce260dd6984b930a0

Observation a0260570-6ce0-4368-bfca-79e54bd1e914 · outbound

This paper cites Stochastic linear quadratic regulators with indefinite control weight costs,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Stochastic linear quadratic regulators with indefinite control weight costs,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:50.100535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:32.785792Z digest=sha256:b3d1fffea4c2d8167c0be0e270c8ca3d86fc0759c56fb0721683103b13108749

Observation 9a56b895-bbdd-4b7b-9b39-ad595502bcb5 · outbound

This paper cites Well-posedness and attainability of indefinite stochastic linear quadratic control in infinite time horizon,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Well-posedness and attainability of indefinite stochastic linear quadratic control in infinite time horizon,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:49.849387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:32.875427Z digest=sha256:0aab7a3ae1f72fe631117a4ef1e484d78342e504163355449cc5fe15f7b6ea78

Observation 56b4a220-ed46-4e8d-bf98-140e0303d25c · outbound

This paper cites Linear matrix inequalities, riccati equations, and indefinite stochastic linear quadratic controls,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Linear matrix inequalities, riccati equations, and indefinite stochastic linear quadratic controls,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:49.628482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:32.970981Z digest=sha256:bd2da2980ad474ca24e3ffef97f07c9f781eff054e2f970ed6ec990fc4fdc35e

Observation 6a310869-e25f-46c5-b6e2-e8ea4743ab12 · outbound

This paper cites Solvability and asymptotic behavior of generalized riccati equations arising in indefinite stochastic lq controls,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Solvability and asymptotic behavior of generalized riccati equations arising in indefinite stochastic lq controls,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:49.359104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:33.097423Z digest=sha256:a89f0c28686d8f189ce308e14b38b896ebb4e4a8234a26548c98712c86fbba66

Observation 5304aed6-de98-4bb2-9c9f-dd2410cfa39f · outbound

This paper cites A primal-dual semi-definite programming approach to linear quadratic control,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems A primal-dual semi-definite programming approach to linear quadratic control,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:47.792609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:33.187398Z digest=sha256:ea7a4de6228e40336ee2ced2890040d10b2fe63d22cf0cc4c73bc1bf6ff10202

Observation c4910c53-0ea6-4635-b660-f8dc3fdc3498 · outbound

This paper cites Stabilization control for itˆ o stochastic system with indefinite state and control weight costs,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Stabilization control for itˆ o stochastic system with indefinite state and control weight costs,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:47.001170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:33.296546Z digest=sha256:bb3ef4cc9e8deddb686f6c84749d2d2bb24860fbb6555cf991706b60ce834745

Observation 0cc3dcd9-378d-4d3e-a63d-a3fe1fe2e24e · outbound

This paper cites Optimal regulators for a class of nonlinear stochastic systems,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Optimal regulators for a class of nonlinear stochastic systems,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:46.604066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:33.397836Z digest=sha256:891ce37ef3ca8b85a7897f66bad82d09f5e06db7139f25edfe7975a1d86957e6

Observation 224c8670-cdc7-48ce-bd8a-0e48d8985ff7 · outbound

This paper cites Indefinite linear-quadratic optimal control of mean-field stochas- tic differential equation with jump diffusion: an equivalent cost functional method,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Indefinite linear-quadratic optimal control of mean-field stochas- tic differential equation with jump diffusion: an equivalent cost functional method,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:46.262692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:33.501366Z digest=sha256:834cca65c75056e1c482c609df8a693603de23914574b16293c5161e8db50244

Observation 9ffde277-e768-4d2b-a6fa-a3f77730792b · outbound

This paper cites Stochastic linear quadratic optimal control problems with regime- switching jumps in infinite horizon,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Stochastic linear quadratic optimal control problems with regime- switching jumps in infinite horizon,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:45.935693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:33.595650Z digest=sha256:f76ae2ad49ad0f752c204016ea8fa40ab7692a5647a8f5291b00760542a31252

Observation c30d62f0-e968-4f7c-925b-97b18f70c6c6 · outbound

This paper cites On estimating the expected return on the market: An exploratory investiga- tion,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems On estimating the expected return on the market: An exploratory investiga- tion,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:45.607054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:33.693784Z digest=sha256:a495e40d142bb6efcd087a40a6819459310d0943b610a70b1ce581e93f1322ba

Observation 22553b20-340f-4dc7-840b-ee89e39fa168 · outbound

This paper cites an unresolved cited work.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:33:45.269404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:33.776096Z digest=sha256:2e24736937110759b9b94adc5b8e44b3b91dc2e7949fc2f7fb5fdf1bc498f480

Observation 7081615b-f0b1-426c-86b4-783149d67788 · outbound

This paper cites Rustem and M.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Rustem and M

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:45.008450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:33.870263Z digest=sha256:317a3f1061ab8fb3f60dc8b600082f9f6f8e42e7eb0a6e7eaa9d0d87ae65a24e

Observation 37632502-7169-4d11-845d-733f8f20b593 · outbound

This paper cites an unresolved cited work.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:33:44.729693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:33.944653Z digest=sha256:e9da4818618425b97c7f1f7bc0aa6c7b13789fc2f4bce434981ed132704e0dce

Observation c75ffc6c-70b2-4f62-ae92-8ac9ea4ad5bf · outbound

This paper cites A survey on intrinsic motivation in reinforcement learning.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems A survey on intrinsic motivation in reinforcement learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:34.035769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:34.035769Z digest=sha256:cdb8f9620bfde0617abfa59e63d186e1238a635143128d1b22147b0f78cda3d7

Observation 5ec4e6e4-16a0-4143-a74f-b77f89cebd07 · outbound

This paper cites Formal theory of creativity, fun, and intrinsic motivation (1990–2010),.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Formal theory of creativity, fun, and intrinsic motivation (1990–2010),

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:44.443994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:34.126427Z digest=sha256:8736a4270e45606254a8f4a5764f087667dfc21c867289ffe3fd9b33b6f574bd

Observation 86e125e1-8cb6-47dd-acbd-2a38acffa9ff · outbound

This paper cites Surprise-Based Intrinsic Motivation for Deep Reinforcement Learning.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Surprise-Based Intrinsic Motivation for Deep Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:34.257469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:34.257469Z digest=sha256:2709e546e783627b5405494ab9336df9014cd361505c46f5c6643dc885b8de3f

Observation 347fd813-d80b-41d9-a067-a87eebf867ef · outbound

This paper cites Curiosity-driven exploration by self- supervised prediction,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Curiosity-driven exploration by self- supervised prediction,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:44.157257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:34.353252Z digest=sha256:3892f2fa553fbcef7e59f09651b96f9f2e4587132de3d3caa258ebeab2df6ecb

Observation 84763f26-e404-4299-83e2-4f9271c8ff3d · outbound

This paper cites Large-Scale Study of Curiosity-Driven Learning.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Large-Scale Study of Curiosity-Driven Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:34.448173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:34.448173Z digest=sha256:dca9b24071f0c86be3fbd2b77bd9213d39169861ad572eece6df1fe0292eb7ad

Observation 10d21d05-a695-4b4e-9f32-1cfd6bd8fa23 · outbound

This paper cites Exploration by Random Network Distillation.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Exploration by Random Network Distillation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:34.568573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:34.568573Z digest=sha256:1d881cefa49df692940e7986698055bba16c189fa68625dcd875a3c1a6d0b7dc

Observation d9708eae-0a06-4e27-8b0a-39fdd77d889f · outbound

This paper cites Randomized prior functions for deep reinforcement learning,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Randomized prior functions for deep reinforcement learning,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:43.818381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:34.642461Z digest=sha256:109da91be5e911966faf9807c46b9f1be95edcee930e76a815a9085816e2df9b

Observation bc9ef8ca-a6c4-49c3-b588-e076d4af1322 · outbound

This paper cites Fast active learning for pure exploration in reinforcement learning,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Fast active learning for pure exploration in reinforcement learning,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:43.572233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:34.727222Z digest=sha256:405097dd84eea02c452216c3ac4a3b0e393216566613f332009a85fcb05c9f1e

Observation e0d6309e-1719-4bf5-a7d2-7193a8dae6a2 · outbound

This paper cites # exploration: A study of count-based exploration for deep reinforcement learning,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems # exploration: A study of count-based exploration for deep reinforcement learning,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:43.310875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:34.830998Z digest=sha256:201a3476402513c4608f554ca3d783a1e1ff12bdb968e1907f08c87845f18ebb

Observation 16a5722b-ccca-4208-bb58-b7412a6a7e36 · outbound

This paper cites Go-Explore: a New Approach for Hard-Exploration Problems.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Go-Explore: a New Approach for Hard-Exploration Problems

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:34.943251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:34.943251Z digest=sha256:7b928216842ce7f26c1302587a768c430c1da1deaba501ffa35c019eac850d4b

Observation d28035f0-aadb-4915-83d3-90f2e7e7cc3d · outbound

This paper cites First return, then explore,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems First return, then explore,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:43.039615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:35.051690Z digest=sha256:280d29c272d852ce9b849dc66f76080384fe6a8b4fd8ccbb7f94fca26b848723

Observation b25f14d2-0ef2-4fc4-98fc-5b3db13d6108 · outbound

This paper cites Leveraging Demonstrations for Deep Reinforcement Learning on Robotics Problems with Sparse Rewards.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Leveraging Demonstrations for Deep Reinforcement Learning on Robotics Problems with Sparse Rewards

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:35.132039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:35.132039Z digest=sha256:d2fa7d37c94d4077905ba3437d03c36d940f877e3c50c6737c738665e9e0ad95

Observation 3f31a89c-dfa3-46ad-a787-31beb912227c · outbound

This paper cites A comprehensive survey on safe reinforcement learning,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems A comprehensive survey on safe reinforcement learning,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:42.698606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:35.225247Z digest=sha256:60ebb79f9138f3f58a18dc9c38fec884ddf86ea9be25f440c97d36dd4c119c99

Observation 4a86cb93-5458-470a-acfa-5d08381dd836 · outbound

This paper cites Trial without Error: Towards Safe Reinforcement Learning via Human Intervention.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Trial without Error: Towards Safe Reinforcement Learning via Human Intervention

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:35.315253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:35.315253Z digest=sha256:246926ed78f0406beec584c887511888699cb10208f8e1429a6736b6e9e71564

Observation 2e800d42-c001-4948-8579-cb69a9048ba5 · outbound

This paper cites Policy gradient in continuous time,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Policy gradient in continuous time,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:42.482119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:35.459696Z digest=sha256:fd72b0bb8a7bc4af4d763dd23569d4fe2dc863d55bdf209d9e2d1e39731b2d50

Observation 12a82e21-f152-4421-bd38-9659339fcf81 · outbound

This paper cites Making deep Q-learning methods robust to time dis- cretization,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Making deep Q-learning methods robust to time dis- cretization,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:42.225469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:35.546997Z digest=sha256:53b6f43ec947416fd44a80fa32b19cd099a49d3e9692ef94366ad75a56960f71

Observation 9bf35128-1d6c-40f9-8dff-3e1a08431cdc · outbound

This paper cites Time discretization-invariant safe action repetition for policy gradient methods,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Time discretization-invariant safe action repetition for policy gradient methods,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:42.037479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:35.719240Z digest=sha256:8beaadbee37f5389a6c3cb41afd573b446c17af5042bedce2408ff2f7094bd08

Observation 6a3f0064-a959-44be-bc37-060185985191 · outbound

This paper cites Optimal scheduling of entropy regularizer for continuous-time linear-quadratic reinforcement learning,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Optimal scheduling of entropy regularizer for continuous-time linear-quadratic reinforcement learning,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:41.858896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:35.850763Z digest=sha256:dd167723b63104c07875723a75252ca5069a77ce38fbfa8b83d7aafcc22ebdeb

Observation a8021655-0c66-493c-bb20-04aee29c2885 · outbound

This paper cites Indefinite stochastic riccati equations,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Indefinite stochastic riccati equations,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:41.714543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:36.049474Z digest=sha256:f27dcfa1cb38f654a96b30427ad1a8fd78df4c0f65b5abdd17ef9d6bb44f1a97

Observation 7accc150-d4ad-4931-80bc-e33c5cef8d0a · outbound

This paper cites Existence of solutions to a class of indefinite stochastic riccati equations,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Existence of solutions to a class of indefinite stochastic riccati equations,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:41.550446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:36.115610Z digest=sha256:0b72529f61a56720655951e3cc2440775c346c1257eac453b26c824a93f4627f

Observation ccc0b981-160e-4ae4-9d31-cb9faf4049c9 · outbound

This paper cites Reinforcement learning in continuous time and space: A stochastic control approach,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Reinforcement learning in continuous time and space: A stochastic control approach,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:41.370950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:36.290103Z digest=sha256:c3261b602a1de0c3c3a08f89850ed21dda6ed2e64a45592bc7b5e623ed666111

Observation 8f088e2d-74bb-468c-8c19-ee4d41f8cfcc · outbound

This paper cites Exploration-exploitation trade-off for continuous-time episodic reinforcement learning with linear-convex models.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Exploration-exploitation trade-off for continuous-time episodic reinforcement learning with linear-convex models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:36.415315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:36.415315Z digest=sha256:1221c8308b1c82dda84ba91e3ebdf797d0a2ebd8f364cd0481c49f005e13ff8d

Observation dd631aed-b86c-4aaf-a250-0ecdc1143c55 · outbound

This paper cites Policy gradient and actor-critic learning in continuous time and space: Theory and algorithms,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Policy gradient and actor-critic learning in continuous time and space: Theory and algorithms,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:41.177508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:36.559039Z digest=sha256:c31a2651de416b0e39c50a92d207f38b6d7a1446cf519901111fa1b08861ee26

Observation 99628674-cdfd-489c-8807-1efccfe2d68d · outbound

This paper cites Policy evaluation and temporal-difference learning in continuous time and space: A martingale approach,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Policy evaluation and temporal-difference learning in continuous time and space: A martingale approach,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:41.048499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:36.759306Z digest=sha256:682ec704b2b060f97668bbf38f384f3950446e7d2cd6ded822c595687b8c5877

Observation d2f9f109-02ae-4cbe-a0d1-57cad82d45d9 · outbound

This paper cites A stochastic approximation method,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems A stochastic approximation method,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:40.850349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:36.893967Z digest=sha256:de85d781d188fea4fca16d6cf7f250d37289fe8ec6f261fa7ef8400f2dfb05d1

Observation bddc39d1-6174-4d91-ba31-6f58c63fe5d9 · outbound

This paper cites Stochastic approximation,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Stochastic approximation,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:40.619413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:37.102211Z digest=sha256:ae34af2a9dce6f6d6d99fe538b7d9dd011d2ac561ebe39348dc1f24cfb0fc0d5

Observation 52b5835a-44b2-4069-ac9a-03d5f5a4f485 · outbound

This paper cites an unresolved cited work.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:33:40.425098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:37.259548Z digest=sha256:1b413b51b415969ea69a07a95352ffc7cc6650330690ce35e6f30eaa29bd48d6

Observation ddca0fb6-7710-483a-9b99-e3eade5a629b · outbound

This paper cites An overview of stochastic approximation,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems An overview of stochastic approximation,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:40.198447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:37.457228Z digest=sha256:1d257697fb58cb08c857667197a770145d5f1600c1e17877c8d22a5e93ea2959

Observation ac6aff8d-77e2-43be-a347-3e08a1c4bc21 · outbound

This paper cites A stochastic approximation algorithm with varying bounds,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems A stochastic approximation algorithm with varying bounds,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:39.861209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:37.676106Z digest=sha256:f35385a680a1812c322453b3d9e541b75938b1f02a9ac417ce9a3f28eb72fc30

Observation a05ae417-2113-4264-88f5-45e8b3888bc7 · outbound

This paper cites General bounds and finite-time improvement for the Kiefer-Wolfowitz stochastic approximation algorithm,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems General bounds and finite-time improvement for the Kiefer-Wolfowitz stochastic approximation algorithm,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:39.528730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:37.874175Z digest=sha256:15931cf976f020b4015ba7e1c00863f848aed03806d01d84e73b6bab2df26e5b

Observation 3d5c3919-a8a0-4fce-80c2-6332c5d319a7 · outbound

This paper cites A convergence theorem for non negative almost supermartin- gales and some applications,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems A convergence theorem for non negative almost supermartin- gales and some applications,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:39.185660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:38.010679Z digest=sha256:495e962cd846df0ff9086b73393fce2303d499b513ccf80c7b25e0fceed80035

Observation 9b5aaec9-6d3e-4734-8818-9053c94d0707 · outbound

This paper cites an unresolved cited work.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:33:38.840537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:33:38.138033Z digest=sha256:a9831e73452aa25073d5e17ce434ac51dd66841c7e3005e2ec514709f5f04bf7

Observation 8fcc4cb7-4208-41ba-83b1-a123ac4ebeb2 · outbound

This paper cites Regret of exploratory policy improvement and $q$-learning.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Regret of exploratory policy improvement and $q$-learning

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:38.321579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:38.321579Z digest=sha256:0d13f6349ba248f825d4f776e7bb20e186227d4a44b2faa277ade58acabba81b

Pith citing papers

Observation ec665d63-6095-424a-ab99-5b351f53c8d3 · inbound

ART for Diffusion Sampling: A Reinforcement Learning Approach to Timestep Schedule cites this paper.

ART for Diffusion Sampling: A Reinforcement Learning Approach to Timestep Schedule Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-16T10:42:45.321033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T10:42:23.768313Z digest=sha256:289e8980ee360c90debc25ab1061c608c3d623c72de715629e8af665ce099cf2

Observation 9818ebf3-6f70-41c9-8d27-84dac68bed2a · inbound

Amortized Guidance for Image Inpainting with Pretrained Diffusion Models cites this paper.

Amortized Guidance for Image Inpainting with Pretrained Diffusion Models Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T19:37:52.603822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-14T19:33:27.641167Z digest=sha256:af3dcb63c38934315623a6e4f099dba90fbbd0a1e7a6510c3390b2cb60dd5378

Observation dd9a3f69-c1e9-47d8-9199-dfe392315231 · inbound

ART for Diffusion Sampling: Continuous-Time Control and Actor-Critic Learning cites this paper.

ART for Diffusion Sampling: Continuous-Time Control and Actor-Critic Learning Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-03T17:18:43.033802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-03T17:14:04.821073Z digest=sha256:6060b729ccc74dc3d4fdb1d4ff04051a09fd683a41a31038c411f2182c6ba5fb

Observation 10076fa3-b771-462d-89e9-b6081230cc75 · inbound

ART for Diffusion Sampling: Continuous-Time Control and Actor-Critic Learning cites this paper.

ART for Diffusion Sampling: Continuous-Time Control and Actor-Critic Learning Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-12T08:25:48.021715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T08:25:48.021715Z digest=sha256:063532d5250e13b955b11e988f7537fbed0f694da2a44d0ca9211fd5c33e368a

Observation 2a8504d5-6aa5-4220-9480-8210bd94ac96 · inbound

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies cites this paper.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.635669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.635669Z digest=sha256:926d98cc64da7a2a900447cd7e1556bf0542610a25219464cf5366835cfb9c8a