Pith. sign in

Paper Citation Record · LEDGER

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems

As of 8 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 5 inbound Pith citation observations for arXiv:2507.00358.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.00358 v2

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:33:38.321579Z

measured 54 of 54 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T11:22:20.635669Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T17:18:43.032059Z

Reference resolution

49 of 49 outbound references displayed

  • verified exact0
  • verified fuzzy34
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 53c011f9-07ed-4843-baeb-28990310a5df · outbound

This paper cites Sublinear Regret for a Class of Continuous-Time Linear-Quadratic Reinforcement Learning Problems.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Sublinear Regret for a Class of Continuous-Time Linear-Quadratic Reinforcement Learning Problems

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:32.491006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:32.491006Z digest=sha256:d5fa8d76fd09135a9198056345e8ff001f6c53b59d627e13fdc031b5d9a933c1

Observation a416a5af-1070-4a81-87d0-f759c5104c62 · outbound

This paper cites an unresolved cited work.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:33:50.572667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:32.598128Z digest=sha256:0d3b439efc214c6f0afe00e09376d178d3f9079c48275b3c35476589eb3360eb

Observation 3319b20d-c984-487a-8f7e-7b9980ff7326 · outbound

This paper cites Yong and X.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Yong and X

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:50.346689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:32.682350Z digest=sha256:4bcb04be170cc3d2a4a3a06650f02b10d38b6087ae73df35de5c26f1eda4cc0f

Observation a0260570-6ce0-4368-bfca-79e54bd1e914 · outbound

This paper cites Stochastic linear quadratic regulators with indefinite control weight costs,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Stochastic linear quadratic regulators with indefinite control weight costs,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:50.100535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:32.785792Z digest=sha256:e731fefebf9eaad2e37b8a0a4ceada1803a370dce66bf4b4fe7e708d593992f1

Observation 9a56b895-bbdd-4b7b-9b39-ad595502bcb5 · outbound

This paper cites Well-posedness and attainability of indefinite stochastic linear quadratic control in infinite time horizon,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Well-posedness and attainability of indefinite stochastic linear quadratic control in infinite time horizon,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:49.849387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:32.875427Z digest=sha256:76ea00cfeecccbcfb99942c0b59e8fb42cc110e5dee64ad14cdfb958b8972f52

Observation 56b4a220-ed46-4e8d-bf98-140e0303d25c · outbound

This paper cites Linear matrix inequalities, riccati equations, and indefinite stochastic linear quadratic controls,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Linear matrix inequalities, riccati equations, and indefinite stochastic linear quadratic controls,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:49.628482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:32.970981Z digest=sha256:11f08abc8548d102b6bce2e2725aed8f952dcd3777b6132a7fa2a5f3f9a5107c

Observation 6a310869-e25f-46c5-b6e2-e8ea4743ab12 · outbound

This paper cites Solvability and asymptotic behavior of generalized riccati equations arising in indefinite stochastic lq controls,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Solvability and asymptotic behavior of generalized riccati equations arising in indefinite stochastic lq controls,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:49.359104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:33.097423Z digest=sha256:07fa24658a3ea5a64bad80e42446bc94702654be1e591a9f5f6ff4b65de55b80

Observation 5304aed6-de98-4bb2-9c9f-dd2410cfa39f · outbound

This paper cites A primal-dual semi-definite programming approach to linear quadratic control,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems A primal-dual semi-definite programming approach to linear quadratic control,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:47.792609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:33.187398Z digest=sha256:9b442b1b211fb890f778510b339630526ecd4fa6a026b66129dc2d5db2da1805

Observation c4910c53-0ea6-4635-b660-f8dc3fdc3498 · outbound

This paper cites Stabilization control for itˆ o stochastic system with indefinite state and control weight costs,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Stabilization control for itˆ o stochastic system with indefinite state and control weight costs,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:47.001170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:33.296546Z digest=sha256:6413dcf30baf02e30ce8f8315b0d5f920e5e0d80484a1177bc95b9479f4178d5

Observation 0cc3dcd9-378d-4d3e-a63d-a3fe1fe2e24e · outbound

This paper cites Optimal regulators for a class of nonlinear stochastic systems,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Optimal regulators for a class of nonlinear stochastic systems,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:46.604066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:33.397836Z digest=sha256:cdb2def0bf1323b8c966138e29f253c5d046265d79f0c6f620e5e986c18a01b1

Observation 224c8670-cdc7-48ce-bd8a-0e48d8985ff7 · outbound

This paper cites Indefinite linear-quadratic optimal control of mean-field stochas- tic differential equation with jump diffusion: an equivalent cost functional method,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Indefinite linear-quadratic optimal control of mean-field stochas- tic differential equation with jump diffusion: an equivalent cost functional method,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:46.262692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:33.501366Z digest=sha256:d8586008b6f5a949ba73478cc9e8256084322202dd5a9aed77c66295421e1e4c

Observation 9ffde277-e768-4d2b-a6fa-a3f77730792b · outbound

This paper cites Stochastic linear quadratic optimal control problems with regime- switching jumps in infinite horizon,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Stochastic linear quadratic optimal control problems with regime- switching jumps in infinite horizon,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:45.935693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:33.595650Z digest=sha256:9ab080ea3942917e0d4bc8f3a78f54c1660259aa03ca7d81ae5bf5a3aff0f5e6

Observation c30d62f0-e968-4f7c-925b-97b18f70c6c6 · outbound

This paper cites On estimating the expected return on the market: An exploratory investiga- tion,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems On estimating the expected return on the market: An exploratory investiga- tion,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:45.607054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:33.693784Z digest=sha256:cd88c50848405fe717840de22f5c043acf797d9bc4fd39592608719e6341268c

Observation 22553b20-340f-4dc7-840b-ee89e39fa168 · outbound

This paper cites an unresolved cited work.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:33:45.269404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:33.776096Z digest=sha256:d1b921fc94d0d3460bcbccaa9fbffebefdfe57fa4596a6b78b681b00cb1ced49

Observation 7081615b-f0b1-426c-86b4-783149d67788 · outbound

This paper cites Rustem and M.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Rustem and M

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:45.008450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:33.870263Z digest=sha256:8b3043ce6f6e933ab862491e7a7b7272bcd7136415240488329c2981847f46e0

Observation 37632502-7169-4d11-845d-733f8f20b593 · outbound

This paper cites an unresolved cited work.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:33:44.729693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:33.944653Z digest=sha256:c451c27b687e4b14f2c9abfeab77907ff7176ffe96df77f63839d632049e97d0

Observation c75ffc6c-70b2-4f62-ae92-8ac9ea4ad5bf · outbound

This paper cites A survey on intrinsic motivation in reinforcement learning.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems A survey on intrinsic motivation in reinforcement learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:34.035769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:34.035769Z digest=sha256:cdb8f9620bfde0617abfa59e63d186e1238a635143128d1b22147b0f78cda3d7

Observation 5ec4e6e4-16a0-4143-a74f-b77f89cebd07 · outbound

This paper cites Formal theory of creativity, fun, and intrinsic motivation (1990–2010),.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Formal theory of creativity, fun, and intrinsic motivation (1990–2010),

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:44.443994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:34.126427Z digest=sha256:47889d0b48cf310bf1370cdaa3fd3ee9c494feaead3204bda471a06376e1ccc4

Observation 86e125e1-8cb6-47dd-acbd-2a38acffa9ff · outbound

This paper cites Surprise-Based Intrinsic Motivation for Deep Reinforcement Learning.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Surprise-Based Intrinsic Motivation for Deep Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:34.257469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:34.257469Z digest=sha256:2709e546e783627b5405494ab9336df9014cd361505c46f5c6643dc885b8de3f

Observation 347fd813-d80b-41d9-a067-a87eebf867ef · outbound

This paper cites Curiosity-driven exploration by self- supervised prediction,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Curiosity-driven exploration by self- supervised prediction,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:44.157257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:34.353252Z digest=sha256:ba0b4d98a5aff59327ce41d35cfdfd8d4c1a91e30cac0c2a1af240a083573e5a

Observation 84763f26-e404-4299-83e2-4f9271c8ff3d · outbound

This paper cites Large-Scale Study of Curiosity-Driven Learning.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Large-Scale Study of Curiosity-Driven Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:34.448173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:34.448173Z digest=sha256:dca9b24071f0c86be3fbd2b77bd9213d39169861ad572eece6df1fe0292eb7ad

Observation 10d21d05-a695-4b4e-9f32-1cfd6bd8fa23 · outbound

This paper cites Exploration by Random Network Distillation.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Exploration by Random Network Distillation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:34.568573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:34.568573Z digest=sha256:1d881cefa49df692940e7986698055bba16c189fa68625dcd875a3c1a6d0b7dc

Observation d9708eae-0a06-4e27-8b0a-39fdd77d889f · outbound

This paper cites Randomized prior functions for deep reinforcement learning,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Randomized prior functions for deep reinforcement learning,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:43.818381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:34.642461Z digest=sha256:31e3775a5c64c50feec93119eda7424a4a412b1b1350dd27c84487386a3510f6

Observation bc9ef8ca-a6c4-49c3-b588-e076d4af1322 · outbound

This paper cites Fast active learning for pure exploration in reinforcement learning,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Fast active learning for pure exploration in reinforcement learning,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:43.572233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:34.727222Z digest=sha256:a99ca2a4cb6ceb06edf67f7d74ffb77a51883459361c62c8c41d4aa42c46cc1b

Observation e0d6309e-1719-4bf5-a7d2-7193a8dae6a2 · outbound

This paper cites # exploration: A study of count-based exploration for deep reinforcement learning,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems # exploration: A study of count-based exploration for deep reinforcement learning,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:43.310875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:34.830998Z digest=sha256:8c02d63074a2a374e61a5b88fda514b6f2ec66662e0018cf07daf512e2010a6c

Observation 16a5722b-ccca-4208-bb58-b7412a6a7e36 · outbound

This paper cites Go-Explore: a New Approach for Hard-Exploration Problems.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Go-Explore: a New Approach for Hard-Exploration Problems

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:34.943251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:34.943251Z digest=sha256:7b928216842ce7f26c1302587a768c430c1da1deaba501ffa35c019eac850d4b

Observation d28035f0-aadb-4915-83d3-90f2e7e7cc3d · outbound

This paper cites First return, then explore,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems First return, then explore,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:43.039615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:35.051690Z digest=sha256:dc2330137b98e9ccdc0c22850274972422715ec04bf9dda5ec83a6eb7e879c34

Observation b25f14d2-0ef2-4fc4-98fc-5b3db13d6108 · outbound

This paper cites Leveraging Demonstrations for Deep Reinforcement Learning on Robotics Problems with Sparse Rewards.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Leveraging Demonstrations for Deep Reinforcement Learning on Robotics Problems with Sparse Rewards

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:35.132039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:35.132039Z digest=sha256:d2fa7d37c94d4077905ba3437d03c36d940f877e3c50c6737c738665e9e0ad95

Observation 3f31a89c-dfa3-46ad-a787-31beb912227c · outbound

This paper cites A comprehensive survey on safe reinforcement learning,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems A comprehensive survey on safe reinforcement learning,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:42.698606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:35.225247Z digest=sha256:995a7362eb3f9b282f12c0011d386c9e9eb66e3e93df29c4f563fe6f2f607ca7

Observation 4a86cb93-5458-470a-acfa-5d08381dd836 · outbound

This paper cites Trial without Error: Towards Safe Reinforcement Learning via Human Intervention.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Trial without Error: Towards Safe Reinforcement Learning via Human Intervention

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:35.315253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:35.315253Z digest=sha256:246926ed78f0406beec584c887511888699cb10208f8e1429a6736b6e9e71564

Observation 2e800d42-c001-4948-8579-cb69a9048ba5 · outbound

This paper cites Policy gradient in continuous time,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Policy gradient in continuous time,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:42.482119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:35.459696Z digest=sha256:1bef1ad2c7dc60378e323549201aaf9ff312982b171c36fd5101b78096970a91

Observation 12a82e21-f152-4421-bd38-9659339fcf81 · outbound

This paper cites Making deep Q-learning methods robust to time dis- cretization,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Making deep Q-learning methods robust to time dis- cretization,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:42.225469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:35.546997Z digest=sha256:7a57d6682b9b9be3309dba4219c3fc2e8f058dce84b0b02e2ebd9e33ba0d5749

Observation 9bf35128-1d6c-40f9-8dff-3e1a08431cdc · outbound

This paper cites Time discretization-invariant safe action repetition for policy gradient methods,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Time discretization-invariant safe action repetition for policy gradient methods,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:42.037479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:35.719240Z digest=sha256:0b3aff13dfed5dc1ae7c4b4bd03f77f0413d9a56c68b863fad051f5c7c60d1b3

Observation 6a3f0064-a959-44be-bc37-060185985191 · outbound

This paper cites Optimal scheduling of entropy regularizer for continuous-time linear-quadratic reinforcement learning,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Optimal scheduling of entropy regularizer for continuous-time linear-quadratic reinforcement learning,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:41.858896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:35.850763Z digest=sha256:8cbee79835d082df70884b6dcdd22decb46d1aa0ab74d1babb6d33379cd1627f

Observation a8021655-0c66-493c-bb20-04aee29c2885 · outbound

This paper cites Indefinite stochastic riccati equations,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Indefinite stochastic riccati equations,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:41.714543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:36.049474Z digest=sha256:c9b011a82e068bd55d098f477b3fd809353eac578c5fc018acfca9e9275f9c4d

Observation 7accc150-d4ad-4931-80bc-e33c5cef8d0a · outbound

This paper cites Existence of solutions to a class of indefinite stochastic riccati equations,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Existence of solutions to a class of indefinite stochastic riccati equations,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:41.550446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:36.115610Z digest=sha256:5e0c74b4333dd24a64a0a91fe96a848f276611bfa71040cce83d6077084f88c1

Observation ccc0b981-160e-4ae4-9d31-cb9faf4049c9 · outbound

This paper cites Reinforcement learning in continuous time and space: A stochastic control approach,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Reinforcement learning in continuous time and space: A stochastic control approach,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:41.370950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:36.290103Z digest=sha256:a807089f33d477ff1489339bf0d41f88a5f06bc4c07f404b81b024dd85e88b5b

Observation 8f088e2d-74bb-468c-8c19-ee4d41f8cfcc · outbound

This paper cites Exploration-exploitation trade-off for continuous-time episodic reinforcement learning with linear-convex models.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Exploration-exploitation trade-off for continuous-time episodic reinforcement learning with linear-convex models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:36.415315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:36.415315Z digest=sha256:1221c8308b1c82dda84ba91e3ebdf797d0a2ebd8f364cd0481c49f005e13ff8d

Observation dd631aed-b86c-4aaf-a250-0ecdc1143c55 · outbound

This paper cites Policy gradient and actor-critic learning in continuous time and space: Theory and algorithms,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Policy gradient and actor-critic learning in continuous time and space: Theory and algorithms,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:41.177508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:36.559039Z digest=sha256:1843d9b5928c70f4188bd2ac5bfbe10953f238043f1e74cef3ae5db6bbfef611

Observation 99628674-cdfd-489c-8807-1efccfe2d68d · outbound

This paper cites Policy evaluation and temporal-difference learning in continuous time and space: A martingale approach,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Policy evaluation and temporal-difference learning in continuous time and space: A martingale approach,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:41.048499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:36.759306Z digest=sha256:5eaa330b733ef38b8e5d40ef9e15de751acdb44f29c9b76233601927068fb8be

Observation d2f9f109-02ae-4cbe-a0d1-57cad82d45d9 · outbound

This paper cites A stochastic approximation method,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems A stochastic approximation method,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:40.850349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:36.893967Z digest=sha256:bd74c06e5af083d1ada5dc1709a5208d9526d3aa1f10f8d9bb6309e6d91cba57

Observation bddc39d1-6174-4d91-ba31-6f58c63fe5d9 · outbound

This paper cites Stochastic approximation,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Stochastic approximation,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:40.619413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:37.102211Z digest=sha256:2fd2f322bb2172fbda81d83603f463525e2ac342398fa8e3c7cd13ca55aa5f01

Observation 52b5835a-44b2-4069-ac9a-03d5f5a4f485 · outbound

This paper cites an unresolved cited work.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:33:40.425098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:37.259548Z digest=sha256:513439d6767e31332998dd022728230c5ab0965fb3c8209c9ba3c0efbab40520

Observation ddca0fb6-7710-483a-9b99-e3eade5a629b · outbound

This paper cites An overview of stochastic approximation,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems An overview of stochastic approximation,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:40.198447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:37.457228Z digest=sha256:f1f4950b22cdb3ca236fe88682944eb642c544a766a7711c9a8147f037111223

Observation ac6aff8d-77e2-43be-a347-3e08a1c4bc21 · outbound

This paper cites A stochastic approximation algorithm with varying bounds,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems A stochastic approximation algorithm with varying bounds,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:39.861209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:37.676106Z digest=sha256:283825701e5c1ae1bc9d158d3969beb5c834e0d117765723ace5f264e2ff4bc4

Observation a05ae417-2113-4264-88f5-45e8b3888bc7 · outbound

This paper cites General bounds and finite-time improvement for the Kiefer-Wolfowitz stochastic approximation algorithm,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems General bounds and finite-time improvement for the Kiefer-Wolfowitz stochastic approximation algorithm,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:39.528730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:37.874175Z digest=sha256:aa0cb91c2fab6b19922487bca122a44d64e4de6afd1e34d05409820789ff8ad9

Observation 3d5c3919-a8a0-4fce-80c2-6332c5d319a7 · outbound

This paper cites A convergence theorem for non negative almost supermartin- gales and some applications,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems A convergence theorem for non negative almost supermartin- gales and some applications,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:39.185660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:38.010679Z digest=sha256:2aea88058c21f196067ae8cff9de8301ab519579e54d3ee6d2a0e4c90a11b508

Observation 9b5aaec9-6d3e-4734-8818-9053c94d0707 · outbound

This paper cites an unresolved cited work.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:33:38.840537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:38.138033Z digest=sha256:4dda1a7a4ef88a74bebd113349e210cbb8f4baa9fe95f7881ec9435591a1286f

Observation 8fcc4cb7-4208-41ba-83b1-a123ac4ebeb2 · outbound

This paper cites Regret of exploratory policy improvement and $q$-learning.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Regret of exploratory policy improvement and $q$-learning

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:38.321579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:38.321579Z digest=sha256:0d13f6349ba248f825d4f776e7bb20e186227d4a44b2faa277ade58acabba81b

Pith citing papers

Observation ec665d63-6095-424a-ab99-5b351f53c8d3 · inbound

ART for Diffusion Sampling: A Reinforcement Learning Approach to Timestep Schedule cites this paper.

ART for Diffusion Sampling: A Reinforcement Learning Approach to Timestep Schedule Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-16T10:42:45.321033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T10:42:23.768313Z digest=sha256:daf7406eb4ad10793b0bdac3e77fcfa15aae9c05a1cb1f8544f74f67b5e47bc3

Observation 9818ebf3-6f70-41c9-8d27-84dac68bed2a · inbound

Amortized Guidance for Image Inpainting with Pretrained Diffusion Models cites this paper.

Amortized Guidance for Image Inpainting with Pretrained Diffusion Models Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T19:37:52.603822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-14T19:33:27.641167Z digest=sha256:c6a5c93bf954d448e36cc3e9d43a02922bd9aeb9ffa53d7da7f5df84ce18f47f

Observation dd9a3f69-c1e9-47d8-9199-dfe392315231 · inbound

ART for Diffusion Sampling: Continuous-Time Control and Actor-Critic Learning cites this paper.

ART for Diffusion Sampling: Continuous-Time Control and Actor-Critic Learning Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-03T17:18:43.033802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-03T17:14:04.821073Z digest=sha256:73ba636aa69c0e214e8a4124ab2f7ac26213389b48cd533c14bfad9c48b36923

Observation 10076fa3-b771-462d-89e9-b6081230cc75 · inbound

ART for Diffusion Sampling: Continuous-Time Control and Actor-Critic Learning cites this paper.

ART for Diffusion Sampling: Continuous-Time Control and Actor-Critic Learning Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-12T08:25:48.021715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T08:25:48.021715Z digest=sha256:063532d5250e13b955b11e988f7537fbed0f694da2a44d0ca9211fd5c33e368a

Observation 2a8504d5-6aa5-4220-9480-8210bd94ac96 · inbound

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies cites this paper.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.635669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.635669Z digest=sha256:926d98cc64da7a2a900447cd7e1556bf0542610a25219464cf5366835cfb9c8a