Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-24T11:25:12.822765Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 29 of 29 outbound references and 2 inbound Pith citation observations for arXiv:2209.07059.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-24T11:25:12.822765Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-11T15:52:37.070457Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-07T05:30:53.567417Z
29 of 29 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c21db4c5-ae56-4c38-b755-d15e2c0a2151 · outbound
Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Second order elliptic equations and elliptic systems, volume 174 of Trans- lations of Mathematical Monographs
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 9eb77077-ee9a-4560-ac47-0d357e567849 · outbound
Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Learning equilibrium mean-variance strategy
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d94b767f-a519-42a9-ace3-fc64780a318a · outbound
Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 394a5215-505d-431b-969b-a03390425b50 · outbound
Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Exploratory LQG mean field games with entropy regularization
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 98050124-d32b-48e2-acad-658b994783a8 · outbound
Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Taming the noise in reinforcement learning via soft updates
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 3c103ba0-2e30-4dfc-bf0c-e4a2d8d2ebb2 · outbound
Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Trudinger
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b0c7577d-06ce-4d79-9ff3-d08d2312b5e4 · outbound
Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Entropy regularization for mean field games with learning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 2531fded-2d8f-40ba-818a-776c6ef5cdfb · outbound
Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Reinforcement learning with deep energy-based policies
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 6d969779-70cf-402b-b603-9dc3e556422f · outbound
Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 7ee46a17-1f7d-4d20-8550-129021c20016 · outbound
Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Jacka and Aleksandar Mijatovi´ c
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation df8a86c2-e2bf-4f59-a05e-b4d533172a61 · outbound
Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 385fdfac-bede-4493-a528-e9c26c281144 · outbound
Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 4811f6d9-e3bd-4642-8d3c-0ad973a3c75b · outbound
Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems q-learning in continuous time
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 94e53293-c1d9-43b9-aaa9-d3ef331ce335 · outbound
Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Unresolved cited work
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c105a17b-02b9-49df-96a5-eacf875b0f3c · outbound
Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Kerimkulov, D
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation abb087c2-99ac-48ff-9021-700f50d94391 · outbound
Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Exponential convergence and stability of Howard’s policy improvement algorithm for controlled diffusions
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation cf0bef36-12f0-446c-a26d-b991c1564ab5 · outbound
Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Policy iterations for reinforcement learning problems in continuous time and space—fundamental theory and methods
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 184c3cd6-3214-45db-a1a2-7f4fa125436b · outbound
Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Value Iteration in Continuous Actions, States and Time
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f554acd8-3220-49de-8d23-195d447c70f1 · outbound
Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Higher chain formula proved by combinatorics
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 8144c54c-9df6-4d9c-9a4d-10050f192096 · outbound
Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Unresolved cited work
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 556a6338-a26e-4640-aa46-ebcbd6355afc · outbound
Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Regularity and stability of feedback relaxed controls
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 3b8c6086-367b-4fde-b741-f5e766181509 · outbound
Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Unresolved cited work
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 3c9b46e0-4dc8-42bc-96db-63ed5a8c38f4 · outbound
Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Policy iteration for the deterministic control problems -- a viscosity approach
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation db841bae-6f80-44e9-8320-57b46e065a59 · outbound
Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Exploratory hjb equations and their convergence
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 7282fa5a-4398-4db2-8cd0-e1b8769f71dd · outbound
Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Continuous-time reinforcement learning control: A review of theoretical results, insights on performance, and needs for new designs
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a6e86ab9-f662-4571-b749-8bf30913d064 · outbound
Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Reinforcement learning in continuous time and space: a stochastic control approach
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 07ab9f59-b436-4839-b75f-6bbf4dbcb1b4 · outbound
Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Continuous-time mean-variance portfolio selection: a reinforcement learning framework
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a9de8540-d1ad-4efa-b18c-664f5e3b8aa6 · outbound
Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Ziebart, J
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 9011900a-b31c-4c85-a773-828d553a3769 · outbound
Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Ziebart, Andrew L
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 515f29ec-a540-42de-864a-95f813e05948 · inbound
Continuous-time optimal investment with portfolio constraints: a reinforcement learning approach Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2e2057b-20b1-4630-8632-f97c71617ab8 · inbound
Continuous Policy and Value Iteration for Stochastic Control Problems and Its Convergence Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems
Reference 8510
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.