Pith. sign in

Paper Citation Record · LEDGER

Entropic Regularization of Markov Decision Processes

As of 4 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 0 inbound Pith citation observations for arXiv:1907.04214.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1907.04214 v2

Coverage vector

measured 51 of 51 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-25T01:34:29.047426Z

measured 51 of 51 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

51 of 51 outbound references displayed

  • verified exact11
  • verified fuzzy40
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4ca5afae-4c7a-4482-9413-685f2684cfb7 · outbound

This paper cites Markov Decision Processes: Discrete Stochastic Dynamic Programming ; John Wiley & Sons: Hoboken, NJ, USA.

Entropic Regularization of Markov Decision Processes Markov Decision Processes: Discrete Stochastic Dynamic Programming ; John Wiley & Sons: Hoboken, NJ, USA

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.442857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:30c8cb06c900c998fc306e698d7d86b2dc8c9935285f759cad58d42ff204b35d

Observation f80f228d-e4c3-4244-8757-fe889684089d · outbound

This paper cites Reinforcement Learning: An Introduction ; MIT Press: Cambridge, MA, USA.

Entropic Regularization of Markov Decision Processes Reinforcement Learning: An Introduction ; MIT Press: Cambridge, MA, USA

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.446335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:20b9ba6016b353d23110f9aaf2082d2b505e51b60f050924279d007e5e9778ce

Observation 4d2aabc3-2dd7-47ea-a754-3ed7c198781e · outbound

This paper cites A survey on policy search for robotics.Found.

Entropic Regularization of Markov Decision Processes A survey on policy search for robotics.Found

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.525896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:84de4ce18356b4b47b0968bc5c440807f3d2fa279f6bd81659aca784a496bbb2

Observation cde12b75-023d-469f-ae36-31326970a7d2 · outbound

This paper cites Dynamic Programming.

Entropic Regularization of Markov Decision Processes Dynamic Programming

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.567568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:ed3c1d4fbedf761399b410ff21336772b04c44c60156953ace10eec910719ca5

Observation c11b2ef4-e891-4128-b155-467c1601c9fa · outbound

This paper cites A Natural Policy Gradient.

Entropic Regularization of Markov Decision Processes A Natural Policy Gradient

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.564176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:3e480c99bd085c156d3d65f2c3a7cbd39e4425dde1a6402b4279ecb0a06fa021

Observation 90d87262-2064-4d47-b2d0-ba4a8c1d6f92 · outbound

This paper cites Relative Entropy Policy Search.

Entropic Regularization of Markov Decision Processes Relative Entropy Policy Search

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.425028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:40344936c81cdba673555814203c5e6bc53e92c01de4c78aff79088d123c7585

Observation a43a6983-4833-4b9f-a93f-96e8b01054ac · outbound

This paper cites Trust Region Policy Optimization.

Entropic Regularization of Markov Decision Processes Trust Region Policy Optimization

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.581387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:b5a07df049d095ce3943fc08002cf6293811b8112be559879463863b32d392e0

Observation 4a6c47a1-f714-4e1e-973f-306249e10795 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Entropic Regularization of Markov Decision Processes Proximal Policy Optimization Algorithms

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-05-25T01:35:10.837484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:346078b5af120898d71a41f31d4cb3abe89dec1861fa9b75300070be0cb1d414

Observation 2415fe7d-8436-4e4c-809e-aa671659114c · outbound

This paper cites Improving predictive inference under covariate shift by weighting the log-likelihood function.

Entropic Regularization of Markov Decision Processes Improving predictive inference under covariate shift by weighting the log-likelihood function

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.554100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:ff331d07d26c868af447ac627da6c9469b05bbc7d6b86f98783b0f6ae1c550a1

Observation 2ea71631-65a5-44c2-9d08-e1de542bd9de · outbound

This paper cites A unified view of entropy-regularized Markov decision processes.

Entropic Regularization of Markov Decision Processes A unified view of entropy-regularized Markov decision processes

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-05-25T01:35:10.850379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:28597edecd446b4de5081abbd548277c381cbe6f37935e3f8a8e9cdca9814e38

Observation fb63fada-3a65-4d37-90eb-fa44b54fd1dc · outbound

This paper cites Proximal Algorithms.

Entropic Regularization of Markov Decision Processes Proximal Algorithms

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.588654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:af916597a68d3381082c85c134f1d2641f0f2aac9e3e082ebd4c2c284d0b9506

Observation a1ad0364-035e-409a-8421-f7723474f946 · outbound

This paper cites An elementary introduction to information geometry.

Entropic Regularization of Markov Decision Processes An elementary introduction to information geometry

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-25T01:35:10.807274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:8cd92016269c939b814f3d1c1f0048ebee9bf13b6fb53e5eb75f15686d29a6de

Observation dd44e36a-ffbb-43ee-ba07-31510f19f4dc · outbound

This paper cites Generative Adversarial Nets.

Entropic Regularization of Markov Decision Processes Generative Adversarial Nets

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.421812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:fad86db757986743351318284b35886339924e63a22baf82fcd188b5e13a5e29

Observation 05e5a00d-5032-4d40-af0d-fbb6869267ff · outbound

This paper cites Geometrical Insights for Implicit Generative Modeling.

Entropic Regularization of Markov Decision Processes Geometrical Insights for Implicit Generative Modeling

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.428681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:fa062f40f570178a6998781f5dde6375d54bc49f79252510571db1ff27e96e1c

Observation 487d6822-2a58-40dc-8042-d806d3031818 · outbound

This paper cites f-GAN: Training Generative Neural Samplers using Variational Divergence Minimization.

Entropic Regularization of Markov Decision Processes f-GAN: Training Generative Neural Samplers using Variational Divergence Minimization

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.432302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:4027c927fa97560eb7d867aad555c18619157231b40ae77fdadc52db20eaedb3

Observation 789a4559-43d4-4aad-a648-1eff5bfd2873 · outbound

This paper cites Entropic Proximal Mappings with Applications to Nonlinear Programming.

Entropic Regularization of Markov Decision Processes Entropic Proximal Mappings with Applications to Nonlinear Programming

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.578067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:e87d585edc014491a9660fe86bbaa26302619c4507ef630bba00c041a15cb17e

Observation 70394ce9-076b-44bd-833e-f342010eff36 · outbound

This paper cites Problem complexity and method efficiency in optimization.

Entropic Regularization of Markov Decision Processes Problem complexity and method efficiency in optimization

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.529250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:949fbd00c189ab45d936a2bae22e00a2d99d692d3b4c67340986ca719ad43f66

Observation 4f52db6c-710c-40e0-a4e8-6939914a366a · outbound

This paper cites Mirror descent and nonlinear projected subgradient methods for convex optimization.

Entropic Regularization of Markov Decision Processes Mirror descent and nonlinear projected subgradient methods for convex optimization

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.560795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:81190748ce3356ee7f20fbc861fb84c6430e0e83b9048c574a656ea6ad346ec6

Observation 5b185220-9c80-4744-a835-a5fa4ad98645 · outbound

This paper cites A measure of asymptotic efficiency for tests of a hypothesis based on the sum of observations.

Entropic Regularization of Markov Decision Processes A measure of asymptotic efficiency for tests of a hypothesis based on the sum of observations

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.585086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:439679ba4d6b8ce8798ba327a0576753d706fe662a5044786dd5900efbddd626

Observation dd4ad908-5d19-482c-b24a-1c2342b73d21 · outbound

This paper cites Differential-Geometrical Methods in Statistics ; Springer: New York, NY, USA.

Entropic Regularization of Markov Decision Processes Differential-Geometrical Methods in Statistics ; Springer: New York, NY, USA

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.453225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:efad144ea47d858cfe7405bc234d36ce7b7ed02d69ad905f6eaf24c6f34ed405

Observation 5e885707-4d5f-440b-a1c2-18b4e4388a74 · outbound

This paper cites Families of alpha- beta- and gamma- divergences: Flexible and robust measures of Similarities.

Entropic Regularization of Markov Decision Processes Families of alpha- beta- and gamma- divergences: Flexible and robust measures of Similarities

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.439500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:5974b284566190ca1c988247dbdd29d26f933a40ed28707e37b7f09942046be1

Observation a51c0b3d-fe93-40f5-b8d9-a86cb70b0863 · outbound

This paper cites A Notation for Markov Decision Processes.

Entropic Regularization of Markov Decision Processes A Notation for Markov Decision Processes

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-05-25T01:35:10.824053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:0cf3d337c525a0aa2a14499064685b04d56d8d1ebcb02787606ad2069a842fc8

Observation 0a7e1787-d63e-43ee-8215-ca5915e9604d · outbound

This paper cites Policy Gradient Methods for Reinforcement Learning with Function Approximation.

Entropic Regularization of Markov Decision Processes Policy Gradient Methods for Reinforcement Learning with Function Approximation

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.405133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:d5482cf97359979f4045642c7cdac39cf1b688d5ffe5b5bd015f7dfa49779301

Observation bf933ea6-ea26-4203-84e2-9092d5de6572 · outbound

This paper cites Natural Actor-Critic.

Entropic Regularization of Markov Decision Processes Natural Actor-Critic

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.436061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:649c37cbf7613ec5a0fcf652d31a89e656d11204dc1308be0f3cd7164787ae9f

Observation 79d45a85-9698-46b2-9445-717448228549 · outbound

This paper cites High-Dimensional Continuous Control Using Generalized Advantage Estimation.

Entropic Regularization of Markov Decision Processes High-Dimensional Continuous Control Using Generalized Advantage Estimation

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-05-25T01:35:10.856824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:f213b8fc585814661c4dd07008c1c85cb22d44a3931bd870a5a4da9c067151b3

Observation c37fcee8-ee75-41f7-b769-18a4bef77f5a · outbound

This paper cites Eine informationstheoretische Ungleichung und ihre Anwendung auf den Beweis der Ergodizität von Markoffschen Ketten.

Entropic Regularization of Markov Decision Processes Eine informationstheoretische Ungleichung und ihre Anwendung auf den Beweis der Ergodizität von Markoffschen Ketten

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.408522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:adaf832ecf627cce03fe7f55a66017d34dd9cc72541aa50cc889324b2f8058f8

Observation 7fc63687-8142-4f72-b0f8-176dd898db82 · outbound

This paper cites Information Geometric Measurements of Generalisation; Technical Report; Aston University: Birmingham, UK.

Entropic Regularization of Markov Decision Processes Information Geometric Measurements of Generalisation; Technical Report; Aston University: Birmingham, UK

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.571324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:4e48e70e62affbc26d3523b26607b3f410a65c0ea043ba891d970e483e79c30c

Observation 93b4aca0-472d-4c91-abf1-3f4959107eb7 · outbound

This paper cites Simple statistical gradient-following methods for connectionist reinforcement learning.

Entropic Regularization of Markov Decision Processes Simple statistical gradient-following methods for connectionist reinforcement learning

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.418515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:642e3cdf70e30e6b0098f0e8ee2c35673fe61d8c5e8a1df2729c4011239b15ef

Observation c49d21f0-fe76-4b3d-a6cf-55bb496aa490 · outbound

This paper cites Graphical Models, Exponential Families, and Variational Inference.

Entropic Regularization of Markov Decision Processes Graphical Models, Exponential Families, and Variational Inference

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.536527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:adf36640a610b3d23e95094de40806d964a5fc13f3e18fd1ab2ab4c3c4559abf

Observation e944a79c-e81d-42ca-8eaf-90096feb2ec8 · outbound

This paper cites Residual Algorithms: Reinforcement Learning with Function Approximation.

Entropic Regularization of Markov Decision Processes Residual Algorithms: Reinforcement Learning with Function Approximation

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.533042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:962d43e3365b0211089e8bf036a99c5dee0c569eeb40e15ac10dd015f28334d7

Observation 5ec633d3-528a-497c-8906-d8410ceae7b6 · outbound

This paper cites Policy Evaluation with Temporal Differences: A Survey and Comparison.

Entropic Regularization of Markov Decision Processes Policy Evaluation with Temporal Differences: A Survey and Comparison

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.449695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:748f1bbe2d98bbeaf943cec63230c9cd285f0d60cf76598da0adbdce7447b2f7

Observation fbc52316-854a-4cb2-a675-63c86c9b421f · outbound

This paper cites F-divergence inequalities.

Entropic Regularization of Markov Decision Processes F-divergence inequalities

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.539574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:d7970691390ce7d3fc1e3a21a9f0b70e0b0503a7eb598d668e98c5fbd26544be

Observation 11c4d318-20ed-4983-9986-727bce10d3f1 · outbound

This paper cites Regret Analysis of Stochastic and Nonstochastic Multi-armed Bandit Problems.

Entropic Regularization of Markov Decision Processes Regret Analysis of Stochastic and Nonstochastic Multi-armed Bandit Problems

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.543222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:42ac3fd7419bc5fdc7cf51238aacd20c1bded69e53a586b6edcca7274b649d7c

Observation c85c889a-5cb7-49f5-a53c-cf130dccc233 · outbound

This paper cites The Non-Stochastic Multi-Armed Bandit Problem.SIAM J.

Entropic Regularization of Markov Decision Processes The Non-Stochastic Multi-Armed Bandit Problem.SIAM J

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.550566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:cb43415e9a8e0850c6950d10849847856de2eef4bca452a5f82c83b2b0fecd26

Observation efc81a9f-2ef9-4f00-a143-ab818269fa93 · outbound

This paper cites Bayesian Reinforcement Learning: A Survey.

Entropic Regularization of Markov Decision Processes Bayesian Reinforcement Learning: A Survey

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.557486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:c604778394dc4b1f034cbc6f27633c6eb9b419e94cb58f32dbb53f383c43d774

Observation 0b1bbcc7-ceaa-4cae-a424-f05b07ec5cfa · outbound

This paper cites OpenAI Gym.

Entropic Regularization of Markov Decision Processes OpenAI Gym

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-05-25T01:35:10.813015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:f79aa78a2ef1eebe9d710025abc2ace224d91f37b8514fc9ca9e960d76a12b46

Observation 686dde6f-72c2-444a-abdd-c72c5f1cad12 · outbound

This paper cites Information theory of decisions and actions.

Entropic Regularization of Markov Decision Processes Information theory of decisions and actions

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.517777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:55dbfadf984ed6893bcf95dba0b08592f93deff00f7e871a1d61dc6f2af230f8

Observation 87b5f4df-6a56-4447-859f-7659e7c44911 · outbound

This paper cites Autonomy: An information theoretic perspective.

Entropic Regularization of Markov Decision Processes Autonomy: An information theoretic perspective

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.521369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:c10671fc2c92afb76eb0469468049f9949e2122ecb52f584c6f89ac9af00674e

Observation 3eb0224f-ce67-4363-86f7-4a4665a68885 · outbound

This paper cites An information-theoretic approach to curiosity-driven reinforcement learning.

Entropic Regularization of Markov Decision Processes An information-theoretic approach to curiosity-driven reinforcement learning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.456346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:10f565ab17062cadc9cfea14587510cfe0b8646b43ccb1ae3eb7023171cb1ce3

Observation 8aecab8a-e7b4-4bb9-8d89-9c5ce5924af5 · outbound

This paper cites Bounded rationality, abstraction, and hierarchical decision-making: An information-theoretic optimality principle.

Entropic Regularization of Markov Decision Processes Bounded rationality, abstraction, and hierarchical decision-making: An information-theoretic optimality principle

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.546982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:50a70a88884582d5812b448d9f18fc5afd873b04b232dd4a34b8b474bf163211

Observation 78027efc-413b-43e6-bac2-fc4e42585850 · outbound

This paper cites Information theory—the bridge connecting bounded rational game theory and statistical physics.

Entropic Regularization of Markov Decision Processes Information theory—the bridge connecting bounded rational game theory and statistical physics

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.514038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:9f01b6946dbaba81b75ce86818f0c7f074e3cce2420c094740de7bbe1c6f7001

Observation 259a522f-702a-4322-9bb2-186974b4928a · outbound

This paper cites A Theory of Regularized Markov Decision Processes.

Entropic Regularization of Markov Decision Processes A Theory of Regularized Markov Decision Processes

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-05-25T01:35:10.863131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:dd837fb9d205f24219ff55593aac85a03563c196c546f16b8e4f6b06370fd808

Observation f189041b-e566-41ff-a85b-262ebde19c7c · outbound

This paper cites A Regularized Approach to Sparse Optimal Policy in Reinforcement Learning.

Entropic Regularization of Markov Decision Processes A Regularized Approach to Sparse Optimal Policy in Reinforcement Learning

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-25T01:35:10.818545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:81fb4daf7f7697ff27d85f98e3f9725ae13185b32334f8dcc909db028e12985b

Observation 9af662a9-8bd8-4ba1-bf4f-71415aa378b3 · outbound

This paper cites Path Consistency Learning in Tsallis Entropy Regularized MDPs.

Entropic Regularization of Markov Decision Processes Path Consistency Learning in Tsallis Entropy Regularized MDPs

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-05-25T01:35:10.831638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:f3f0ccd52873f14845062ca0700e79a3d9baffa9bcc4bf549b293ebf35582ef0

Observation 193ea8ce-7727-4804-8cdf-9240efd3a30e · outbound

This paper cites Tsallis Reinforcement Learning: A Unified Framework for Maximum Entropy Reinforcement Learning.

Entropic Regularization of Markov Decision Processes Tsallis Reinforcement Learning: A Unified Framework for Maximum Entropy Reinforcement Learning

Reference 45

Resolution
verified exact
local_arxiv, observed 2026-05-25T01:35:10.843852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:0c41623d8ba34603cb8b307d8e7837df787cdf0043963a8c3a4eeb6eab2bf6f9

Observation 6387c191-33f1-48ee-8b54-82aef60790bb · outbound

This paper cites Sparse Markov decision processes with causal sparse Tsallis entropy regularization for reinforcement learning.

Entropic Regularization of Markov Decision Processes Sparse Markov decision processes with causal sparse Tsallis entropy regularization for reinforcement learning

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.509716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:644497c6269eec0ae5e4b076b49c932b5bee36655594537d9f1fdba645aafbe2

Observation 4807cc18-ae8a-4f03-8b6c-931b2b527904 · outbound

This paper cites Maximum Causal Tsallis Entropy Imitation Learning.

Entropic Regularization of Markov Decision Processes Maximum Causal Tsallis Entropy Imitation Learning

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.415216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:d7c9db153a324575029f5b164d5d77dc633c01734ddee83041451705355c67a9

Observation 2147e710-cf1a-4061-ab3b-21a137bcffa9 · outbound

This paper cites Proximal Reinforcement Learning: A New Theory of Sequential Decision Making in Primal-Dual Spaces.

Entropic Regularization of Markov Decision Processes Proximal Reinforcement Learning: A New Theory of Sequential Decision Making in Primal-Dual Spaces

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-05-25T01:35:10.800991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:cbf0a810661514fb2e538667311b4f5c4f3a0a00228a740c00d2e25757720ccd

Observation 63b70ca9-abbe-4542-858a-dd71d889e830 · outbound

This paper cites Markov processes and the H-theorem.

Entropic Regularization of Markov Decision Processes Markov processes and the H-theorem

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.574574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:07e0a0eae077a8f32cc3ad01de78486245e2036e1f9c569ee0c8c144119c5feb

Observation 4bab8e75-5611-40f5-a7b6-2cceec3e9131 · outbound

This paper cites A General Class of Coefficients of Divergence of One Distribution from Another.

Entropic Regularization of Markov Decision Processes A General Class of Coefficients of Divergence of One Distribution from Another

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.411930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:7db6e9370788cdec826ec10cf507ffc64cba9b87da3c2507a9b5e176fb3fec56

Observation d646aca9-c28e-40dd-8ff2-cb314217a33e · outbound

This paper cites Convex Optimization; Cambridge University Press: Cambridge, UK, 2004; 487p.

Entropic Regularization of Markov Decision Processes Convex Optimization; Cambridge University Press: Cambridge, UK, 2004; 487p

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T01:35:11.505146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T01:34:29.047426Z digest=sha256:5b29d7103968269e705c0e363ebef20252624be0a56a3248df9392b3e18058fb

Pith citing papers

No inbound Pith citation observations are available.