Pith. sign in

Paper Citation Record · LEDGER

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems

As of 13 August 2026, this Paper Citation Record lists 29 of 29 outbound references and 2 inbound Pith citation observations for arXiv:2209.07059.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2209.07059 v5

Coverage vector

measured 29 of 29 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-24T11:25:12.822765Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T15:52:37.070457Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T05:30:53.567417Z

Reference resolution

29 of 29 outbound references displayed

  • verified exact2
  • verified fuzzy21
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c21db4c5-ae56-4c38-b755-d15e2c0a2151 · outbound

This paper cites Second order elliptic equations and elliptic systems, volume 174 of Trans- lations of Mathematical Monographs.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Second order elliptic equations and elliptic systems, volume 174 of Trans- lations of Mathematical Monographs

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.217603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:a14daa235d1dd980e3e1fedf854bd27ec838209317a0bbba0d2cb9307990603c

Observation 9eb77077-ee9a-4560-ac47-0d357e567849 · outbound

This paper cites Learning equilibrium mean-variance strategy.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Learning equilibrium mean-variance strategy

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.153180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:eb7e123caa8de9de19f26231f4d87a120f3df5b24e858529194500d042cd0882

Observation d94b767f-a519-42a9-ace3-fc64780a318a · outbound

This paper cites an unresolved cited work.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-05-24T11:26:09.163985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:1a1305ce82453d5d6077dfb6949cce3abff313afb47149a0bedf7ba3e86e6d77

Observation 394a5215-505d-431b-969b-a03390425b50 · outbound

This paper cites Exploratory LQG mean field games with entropy regularization.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Exploratory LQG mean field games with entropy regularization

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.240206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:bd0d4c874d7d8b264eaef14e7d814661c0519db49a938f0e1011ddc53c44d4a6

Observation 98050124-d32b-48e2-acad-658b994783a8 · outbound

This paper cites Taming the noise in reinforcement learning via soft updates.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Taming the noise in reinforcement learning via soft updates

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.246705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:28a34c90c89350d84a11b14effb92e7c9e4eb806b404a5b8a53a73d56289f39d

Observation 3c103ba0-2e30-4dfc-bf0c-e4a2d8d2ebb2 · outbound

This paper cites Trudinger.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Trudinger

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.202064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:23a109b2467c42493752a392e75af4ea050ea218c7804599da293956980380cb

Observation b0c7577d-06ce-4d79-9ff3-d08d2312b5e4 · outbound

This paper cites Entropy regularization for mean field games with learning.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Entropy regularization for mean field games with learning

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.146960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:6204fcc2e3dc5e4a000f994f4ea7a4424d508f7b31efdff54bff2a5db204da95

Observation 2531fded-2d8f-40ba-818a-776c6ef5cdfb · outbound

This paper cites Reinforcement learning with deep energy-based policies.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Reinforcement learning with deep energy-based policies

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.232513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:65b13dba9935c646cc90d7e05059151644b22d7909b3d97c5e193f4f29a2d445

Observation 6d969779-70cf-402b-b603-9dc3e556422f · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.224900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:f1e37249f48ffa2bb1db85f6424be18308333a4790ff7219ce5ec2819a32de6f

Observation 7ee46a17-1f7d-4d20-8550-129021c20016 · outbound

This paper cites Jacka and Aleksandar Mijatovi´ c.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Jacka and Aleksandar Mijatovi´ c

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.167727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:7e77abdd8987da676035d199517e761b8aa0f621ea4fab42d86c5835b7caf1a6

Observation df8a86c2-e2bf-4f59-a05e-b4d533172a61 · outbound

This paper cites an unresolved cited work.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-05-24T11:26:09.236857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:daec8bd198a00da7cad114895c893c846824ba21e720010851495fcb4c474cf4

Observation 385fdfac-bede-4493-a528-e9c26c281144 · outbound

This paper cites an unresolved cited work.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-05-24T11:26:09.176502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:65f60cfa35bffa43226c0b501e0fc650a37e793ee096992eb2b262d3a4aeb540

Observation 4811f6d9-e3bd-4642-8d3c-0ad973a3c75b · outbound

This paper cites q-learning in continuous time.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems q-learning in continuous time

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.232890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:a3ed3e18ba4466e3cd9e9fcef7bbe41df1adc61504a86a560cd59860c916e05d

Observation 94e53293-c1d9-43b9-aaa9-d3ef331ce335 · outbound

This paper cites an unresolved cited work.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-05-24T11:26:09.242907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:8581e0e9b924a6247920680af66096af086e5f5f1b6556de8b6fb66293d5be3b

Observation c105a17b-02b9-49df-96a5-eacf875b0f3c · outbound

This paper cites Kerimkulov, D.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Kerimkulov, D

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.236215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:8bd03dd33be2ced72fd961d1c1585efe06d261b98ddd274b9cb39a59136d92e2

Observation abb087c2-99ac-48ff-9021-700f50d94391 · outbound

This paper cites Exponential convergence and stability of Howard’s policy improvement algorithm for controlled diffusions.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Exponential convergence and stability of Howard’s policy improvement algorithm for controlled diffusions

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.220497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:2b4f10ac6965ef7ba902504788e6751c218849d02338f48f4dcf4ccc19b9a88c

Observation cf0bef36-12f0-446c-a26d-b991c1564ab5 · outbound

This paper cites Policy iterations for reinforcement learning problems in continuous time and space—fundamental theory and methods.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Policy iterations for reinforcement learning problems in continuous time and space—fundamental theory and methods

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.186718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:299f5ffa83b7fd3507f7d3be999b239ebb38122f294ec18103d622382be48017

Observation 184c3cd6-3214-45db-a1a2-7f4fa125436b · outbound

This paper cites Value Iteration in Continuous Actions, States and Time.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Value Iteration in Continuous Actions, States and Time

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-24T11:26:09.005905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:58f28778db09145a4f6664b32d6cec5c9dc974b236b680e16d83f042f04b5e63

Observation f554acd8-3220-49de-8d23-195d447c70f1 · outbound

This paper cites Higher chain formula proved by combinatorics.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Higher chain formula proved by combinatorics

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.239670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:b26de8dd4b2dfce484fcbf0f6f4ffe2d93e14807058800006df01f94da48a724

Observation 8144c54c-9df6-4d9c-9a4d-10050f192096 · outbound

This paper cites an unresolved cited work.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-05-24T11:26:09.215950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:bc88dfd6cba1d76e541d842710c279918c0a896a5daf70cdf886533fbeb0011d

Observation 556a6338-a26e-4640-aa46-ebcbd6355afc · outbound

This paper cites Regularity and stability of feedback relaxed controls.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Regularity and stability of feedback relaxed controls

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.213788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:e291def112f5eb3b29ccd471b6d5fe555c3e2f65bdf0a2639041baab978954f3

Observation 3b8c6086-367b-4fde-b741-f5e766181509 · outbound

This paper cites an unresolved cited work.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-05-24T11:26:09.206427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:72948ab62e083219a788aaaaab645daea534d2374d72adef63be91cc1e27b917

Observation 3c9b46e0-4dc8-42bc-96db-63ed5a8c38f4 · outbound

This paper cites Policy iteration for the deterministic control problems -- a viscosity approach.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Policy iteration for the deterministic control problems -- a viscosity approach

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-24T11:26:08.991619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:4da8af6c8a2d1c1a5a06798c4bc75c8709b4b8597592e762b4d9b083f7dd5075

Observation db841bae-6f80-44e9-8320-57b46e065a59 · outbound

This paper cites Exploratory hjb equations and their convergence.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Exploratory hjb equations and their convergence

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.229034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:0563a52745bfd80813d8944888f3f9c967b6f03dbda07d6e3a20cfe3a7172651

Observation 7282fa5a-4398-4db2-8cd0-e1b8769f71dd · outbound

This paper cites Continuous-time reinforcement learning control: A review of theoretical results, insights on performance, and needs for new designs.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Continuous-time reinforcement learning control: A review of theoretical results, insights on performance, and needs for new designs

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.160372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:9543bbd97b6b9764e3a2d0bae26a37cfcead7d8fa574074a2766c40e4c5baea6

Observation a6e86ab9-f662-4571-b749-8bf30913d064 · outbound

This paper cites Reinforcement learning in continuous time and space: a stochastic control approach.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Reinforcement learning in continuous time and space: a stochastic control approach

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.243588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:09eefc1a3cf7b4f6436889f50ef1b81727559c2e046d3144afc3d2b3e76f8dac

Observation 07ab9f59-b436-4839-b75f-6bbf4dbcb1b4 · outbound

This paper cites Continuous-time mean-variance portfolio selection: a reinforcement learning framework.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Continuous-time mean-variance portfolio selection: a reinforcement learning framework

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.211783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:54d32b524791e6f3c00941ea9fc82946fddd4e52723ec71b28a448b46ebb8340

Observation a9de8540-d1ad-4efa-b18c-664f5e3b8aa6 · outbound

This paper cites Ziebart, J.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Ziebart, J

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.224593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:9cacde65e280d00794c5f376f7581b8602a1ac4e87d56f333c868176bc64f555

Observation 9011900a-b31c-4c85-a773-828d553a3769 · outbound

This paper cites Ziebart, Andrew L.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Ziebart, Andrew L

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.207878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:3ec5b238c4a38da8a58d1eabd6bba5b2ddca2875eba476f25a6e40253b69ec08

Pith citing papers

Observation 515f29ec-a540-42de-864a-95f813e05948 · inbound

Continuous-time optimal investment with portfolio constraints: a reinforcement learning approach cites this paper.

Continuous-time optimal investment with portfolio constraints: a reinforcement learning approach Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-11T15:52:37.070457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:52:37.070457Z digest=sha256:767f0ea10181175f6957df5b336487e10e12bd8ec811da3acf994c7efe7f38ba

Observation d2e2057b-20b1-4630-8632-f97c71617ab8 · inbound

Continuous Policy and Value Iteration for Stochastic Control Problems and Its Convergence cites this paper.

Continuous Policy and Value Iteration for Stochastic Control Problems and Its Convergence Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems

Reference 8510

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:30:53.573835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T05:30:53.529461Z digest=sha256:d294ca8a07bd9db72cbb9956f71c63672663f3e6dccf0c5db4644fef30da228a