Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:36:23.070508Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 2 inbound Pith citation observations for arXiv:2505.14564.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:36:23.070508Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-21T19:15:30.759880Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-21T19:20:31.176550Z
22 of 22 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 97971769-4e0c-44b1-aa2e-c90c255db885 · outbound
Bellman operator convergence enhancements in reinforcement learning algorithms Accessed on 13/03/2024
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 88a77697-2960-4bfc-a9da-32b850b2e55e · outbound
Bellman operator convergence enhancements in reinforcement learning algorithms An alternative softmax operator for reinforcement learning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22af8a4b-41b4-401c-b211-b638843799b6 · outbound
Bellman operator convergence enhancements in reinforcement learning algorithms Lipschitz Continuity in Model-based Reinforcement Learning
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30701fa9-700d-44a8-a382-4029391ac28f · outbound
Bellman operator convergence enhancements in reinforcement learning algorithms Speedy q-learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c70850de-13a6-4feb-9af8-4da18a2eeeb7 · outbound
Bellman operator convergence enhancements in reinforcement learning algorithms Neuronlike adaptive elements that can solve difficult learning control problems.IEEE transactions on systems, man, and cybernetics, (5):834–846, 1983
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 385bb917-34a3-4712-8395-c5d0f5b6d3d3 · outbound
Bellman operator convergence enhancements in reinforcement learning algorithms Increasing the action gap: New operators for reinforcement learning
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5dc4b47a-a18f-475e-a079-cbee99a440f0 · outbound
Bellman operator convergence enhancements in reinforcement learning algorithms Q-learning and enhanced policy iteration in discounted dynamic programming
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ee427491-64a4-4073-bbf6-1bbd0e756822 · outbound
Bellman operator convergence enhancements in reinforcement learning algorithms The Value Function Polytope in Reinforcement Learning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3396a85f-3cab-412a-ba59-edf573264ac2 · outbound
Bellman operator convergence enhancements in reinforcement learning algorithms Addison-Wesley Professional, 2019
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 077553d5-df00-491a-93ba-764231b43acd · outbound
Bellman operator convergence enhancements in reinforcement learning algorithms Topological Foundations of Reinforcement Learning
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 94e157b0-8bd1-4747-a679-eb98229684e7 · outbound
Bellman operator convergence enhancements in reinforcement learning algorithms Reinforcement learning essay (aims-cameroon)
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 12ddeb6d-a788-43bb-b167-c188ddef1d34 · outbound
Bellman operator convergence enhancements in reinforcement learning algorithms Metrics and continuity in reinforcement learning
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 47ff5210-c516-4a95-be5e-bd0fa24ae701 · outbound
Bellman operator convergence enhancements in reinforcement learning algorithms Markov decision processes and dynamic programming, 2013
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3e58590e-34a4-433b-8663-793648d0ee5d · outbound
Bellman operator convergence enhancements in reinforcement learning algorithms A General Family of Robust Stochastic Operators for Reinforcement Learning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 83833abb-6141-4a9d-ac0f-03679beb8b2e · outbound
Bellman operator convergence enhancements in reinforcement learning algorithms Efficient memory-based learning for robot control
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a2041a33-4389-44b0-b9f4-c833db4bdcbf · outbound
Bellman operator convergence enhancements in reinforcement learning algorithms John Wiley & Sons, 2013
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ef987062-9d92-4e6a-bf8b-2526c97377cb · outbound
Bellman operator convergence enhancements in reinforcement learning algorithms Efficient Model-free Reinforcement Learning in Metric Spaces
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 604ee969-6e6f-4f26-b1d7-2859d4bf060b · outbound
Bellman operator convergence enhancements in reinforcement learning algorithms Generalization in reinforcement learning: Successful examples using sparse coarse coding.Advances in neural information processing systems, 8, 1995
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5dc5b02-45a1-42f9-9128-f1da5c38c23c · outbound
Bellman operator convergence enhancements in reinforcement learning algorithms MIT press, 2018
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4633efd3-8bb1-47ae-95f3-7945135c3b9e · outbound
Bellman operator convergence enhancements in reinforcement learning algorithms Nova Science Publishers, New York, 2021
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 024b3833-47fe-415b-b7c1-5109dbabce50 · outbound
Bellman operator convergence enhancements in reinforcement learning algorithms Github : Basic Reinforcement Learning
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 31244681-40c7-4d4d-af3f-5c18478d4b53 · outbound
Bellman operator convergence enhancements in reinforcement learning algorithms Learning from delayed rewards
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5e064f49-fe9c-44bf-a386-8cd8c3911954 · inbound
Carbon-Aware Intrusion Detection: A Comparative Study of Supervised and Unsupervised DRL for Sustainable IoT Edge Gateways Bellman operator convergence enhancements in reinforcement learning algorithms
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 92b1a724-e897-400f-b5a6-3aec6e20797d · inbound
TabQL: In-Context Q-Learning with Tabular Foundation Models Bellman operator convergence enhancements in reinforcement learning algorithms
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.