Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-24T02:16:04.596951Z
Paper Citation Record · LEDGER
As of 22 July 2026, this Paper Citation Record lists 60 of 60 outbound references and 2 inbound Pith citation observations for arXiv:2404.14442.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-24T02:16:04.596951Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-07-22T06:31:00.163083+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-10T18:45:48.543849Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-11T23:36:38.356022Z
60 of 60 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 28cc24c6-c65a-486a-85fa-197fe053a6ce · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 97aacc98-5917-43ab-8f17-ce5e51f28664 · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Human-level control through deep rein- forcement learning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 3cc5354f-bf5f-4605-92aa-a3b3884f803e · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Q-learning
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation af840059-9b62-4580-bd77-db791cf081e8 · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Convergence of st ochastic iterative dynamic programming algorithms
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation e79e4f8b-0f33-47c4-a390-8f4b462ea87f · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms The ODE method for convergence of stochastic approximation and reinforcement learning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 6f1771e3-8a63-461b-8875-d4a6458d8872 · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms A unified switching system perspective and con vergence analysis of Q- learning algorithms
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 940e2db8-f018-4452-a4d3-d18c1d92a81c · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Asynchronous stochastic approximation and Q- learning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 183a7283-d48a-4457-8366-773752a8d3e1 · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Conv ergence results for single-step on-policy reinforcement-learning algorithms
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation f7574daf-02c6-4a4f-b65f-4bfe96c4dde6 · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms The asymptotic convergence-rate of Q-lear ning
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation ec2fd1a0-1e56-461f-92f9-b5246cd4e3c8 · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Learning rates for Q-learning
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 143ef549-f556-4ede-bd22-6b3cb1548ee3 · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Error bounds for constant step- size Q-learning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation cab0cf58-eaa2-4bbc-8c10-2ea6b677ea25 · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Stochastic approximation with cone-contractive operators: Sharp $\ell_\infty$-bounds for $Q$-learning
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 09b93c36-6021-4eba-827f-4b394ae47756 · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Finite-Time Analysis of Asynchronous Stochastic Approximation and $Q$-Learning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 3480e319-29f3-457a-a428-e5f7bc10fc2b · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Sample Complexity of Asynchronous Q-Learning: Sharper Analysis and Variance Reduction
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 886d6936-c725-4e7f-8730-96008331033e · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms A Lyapunov Theory for Finite-Sample Guarantees of Asynchronous Q-Learning and TD-Learning Variants
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 476ec9ca-3f23-4922-9344-d2f89aa69f06 · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Finite-sample analysis of contractive stochastic approxim ation using smooth convex envelopes
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation abfdc2e7-f26b-4d39-870d-914537ff166b · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Final iteration convergence of Q-learning: Switching s ystem approach
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 5cabdf23-b006-4da3-9391-af2b85ac67a8 · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms A convergento(n) temporal-difference algorithm for off-policy learning with linear function approximation
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 7a3deb8f-347d-4773-bf42-85f8de6abaa5 · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Fast gradient-descent methods for temporal-difference learnin g with linear function approx- imation
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 693dfff2-9f50-4540-880a-b40f038321cf · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Gradient temporal- difference learning with regularized corrections
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 204f6c31-5ca5-4b2d-b4b6-72f465bb9dee · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms New versions of gradien t temporal difference learning
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 4aba900d-c1bd-4a37-9a71-398796c44c6e · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms An analysis of reinforc ement learning with function approximation
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 6d5195b9-accc-45cb-8d72-6cac39780f78 · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Bhatnagar, H
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation b73ef3ca-4540-4b9d-accf-a535f093520c · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Reinforceme nt learning with deep energy- based policies
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation caac22d1-c0e6-4d43-8d40-de4bfbb0653e · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Revisiting the softmax Bellman o perator: New benefits and new perspective
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 17a14c53-3214-4ba0-8949-8cb9d795d29f · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Reinforcemen t learning with dynamic Boltzmann softmax updates
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 1433e224-7f39-4771-875b-1805dec87a05 · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms An alternative softmax operator for reinforcement learning
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation c9cbe175-75fe-4ad1-9b5a-0807d9c9e361 · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Smoothed Q-learning
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation c3c3ef14-6946-4d41-a43a-5a3413f9668f · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms An analog scheme for fixed point computation. i. theory
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 4ccd7d8d-3fc9-4caf-8f24-dc6aee59cf7a · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Liberzon, Switching in systems and control
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation c4216a61-52e0-47cb-99a4-797ce582ce96 · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Unified finite-time error analysis of soft q -learning
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation cd0296a3-0135-48af-8b76-e08200c65e47 · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms SBEED: Convergent reinforcement learning with nonlinear function approximation
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation e7867490-3ee2-4be7-bba1-08f01f5863cf · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms On the Properties of the Softmax Function with Application in Game Theory and Reinforcement Learning
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation c7906621-92a0-4863-9e1e-4e29722253d2 · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Nonlinear systems
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 735754e0-fd8d-4e28-9998-1128d45754de · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Unresolved cited work
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 24dc0017-51c2-4ee4-b328-425d730c6cd4 · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Finite-time analysis of asynchronous q-lea rning under diminishing step-size from control-theoretic view
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 050e7beb-d5ec-420e-9067-70bae24944e8 · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Kushner and G
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 32c53f3e-5a1b-4ad3-b662-a2f586b033a8 · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms A stochastic approximation method
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation c172d941-e6d5-42d4-87ed-2fa7b06a56d8 · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Note on the derivatives with respect to a param eter of the solutions of a system of differential equations
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 95d2fe58-6c54-4e9e-a64a-e294ae96110a · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Unresolved cited work
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation cef9e683-cc38-4edf-8937-083424958d00 · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Unresolved cited work
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 57f7574a-3fbe-4ece-924f-b63212d35649 · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Unresolved cited work
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 0d6a051b-712a-4517-a989-a9d2cc61b94b · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms In addition, there exists a constant C0 < ∞ such that for any initial θ0 ∈ Rn, we have E[∥εk+1∥2 2|Gk] ≤ C0(1 + ∥θk∥2 2), ∀k ≥ 0
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 2a9b36a8-47c3-4aaa-950e-ec246c9e1213 · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms (17) Lemma 2 ( [5, Borkar and Meyn theorem])
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 17ef970e-76f9-4dba-9839-10f99c0c24c1 · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Unresolved cited work
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation bd502002-6b8c-4ffb-93ce-801bda678d54 · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms e., xt → H as t → ∞
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 953d76c0-123b-4329-9b10-d072fc4b4707 · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Unresolved cited work
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 847b5ecc-aada-4ca9-a4a9-3bfd396e8cfb · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms In addition, there exists a constant C0 < ∞ such that for any initial θ0 ∈ Rn, we have E[∥εk+1∥2 2|Gk] ≤ C0(1 + ∥θk∥2 2), ∀k ≥ 0 with probability one
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation fedb40c7-9b3e-4a52-a6c6-962792dfecda · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Lemma 3 ( [38, Robbins and Monro theorem])
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation bf4b9f89-163f-47a2-b39c-5d7cd9189a5d · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Unresolved cited work
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation bab10916-60a6-4e11-8ef4-d5546f33dcde · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Unresolved cited work
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 6ab92562-9ef0-4b49-a867-6ad80717bbd4 · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Unresolved cited work
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 81f93788-b4a2-4b42-bc07-c7a8f24104bc · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Unresolved cited work
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 957cb86b-a636-4770-9002-2d478f663253 · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Unresolved cited work
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation ba9ee95d-defb-4058-8cb7-b920533667ff · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Unresolved cited work
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 88507e15-fe43-4157-9df0-a16ac61a0deb · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Unresolved cited work
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation c3e5cecb-3e8e-476f-9c74-c433b45fc8f4 · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Unresolved cited work
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 1cebaad9-9ca5-4902-849a-8696c85bcc4a · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Unresolved cited work
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 0d384e2e-326c-4254-8c83-ed913515de1b · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Unresolved cited work
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation eebe205d-ebd2-4f70-bb88-8a39c3398e8c · outbound
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms Similar to the ODE case, the first three rows (max, L SE, and mellowmax) show that the trajectories converge toward their fixe d points as proved in Theorem 5
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation eeed0fd3-3546-4dcc-8add-dfbd564138a0 · inbound
Contraction-Aligned Analysis of Soft Bellman Residual Minimization with Weighted Lp-Norm for Markov Decision Problem Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.
Observation 624a9226-b504-4f17-9350-74df4165231b · inbound
Safe-Support Q-Learning: Learning without Unsafe Exploration Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.