Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-11T17:29:49.186565Z
Paper Citation Record · LEDGER
As of 4 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 100 inbound Pith citation observations for arXiv:2407.17032.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-11T17:29:49.186565Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-03T06:30:56.289259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-03T16:50:39.504717Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-10T06:15:00.866473Z
34 of 34 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 4cd2fe2d-4ae0-4f3b-bcb7-695472a7e997 · outbound
Gymnasium: A Standard Interface for Reinforcement Learning Environments Hindsight Experience Replay
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 7a2ab3fb-be5c-48af-a7a1-597576227b25 · outbound
Gymnasium: A Standard Interface for Reinforcement Learning Environments DeepMind Lab2D
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 47ba3158-f8f6-4f10-99d9-13016bb3d87d · outbound
Gymnasium: A Standard Interface for Reinforcement Learning Environments Dota 2 with Large Scale Deep Reinforcement Learning
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation c2b8b9d1-3d81-43d4-8c9f-0bc4ced60bf3 · outbound
Gymnasium: A Standard Interface for Reinforcement Learning Environments Jumanji: a Diverse Suite of Scalable Reinforcement Learning Environments in JAX
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation b1ccc781-d854-4b8a-8e83-e8c95781704c · outbound
Gymnasium: A Standard Interface for Reinforcement Learning Environments 10 Yann Bouteiller, Edouard GEZE, GobeX, Stefan Kuhn, and pius
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 2140442b-99ea-4ae9-a6a9-444fdf8c3e9b · outbound
Gymnasium: A Standard Interface for Reinforcement Learning Environments James Bradbury, Roy Frostig, Peter Hawkins, Matthew James Johnson, Chris Leary, Dougal Maclaurin, George Necula, Adam Paszke, Jake VanderPlas, Skye Wanderman-Milne, and Qiao Zhang
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 82dfe8da-00d1-4e51-af99-91d2ef66bc45 · outbound
Gymnasium: A Standard Interface for Reinforcement Learning Environments OpenAI Gym
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 29c766d6-d04f-443d-9164-287c68df9a9d · outbound
Gymnasium: A Standard Interface for Reinforcement Learning Environments Dopamine: A Research Framework for Deep Reinforcement Learning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation b9d3c1e1-02ec-42ad-b175-483b43ce798d · outbound
Gymnasium: A Standard Interface for Reinforcement Learning Environments ICU-Sepsis: A Benchmark MDP Built from Real Medical Data
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 43309982-9de4-4ee2-b6e5-20e8161f5ba2 · outbound
Gymnasium: A Standard Interface for Reinforcement Learning Environments Accelerating Reinforcement Learning through GPU Atari Emulation
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 45f8087c-9b2e-4864-869b-69f171fa9bac · outbound
Gymnasium: A Standard Interface for Reinforcement Learning Environments OCAtari: Object-Centric Atari 2600 Reinforcement Learning Environments
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e505c5a5-281e-4d75-9c32-a743585f3e73 · outbound
Gymnasium: A Standard Interface for Reinforcement Learning Environments Alegre, Ann Nowé, Ana L
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 6a9d7e02-f680-4442-9920-1cce1b8bb4b7 · outbound
Gymnasium: A Standard Interface for Reinforcement Learning Environments Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 5a4769fb-9a4c-44e1-934a-12ae09581d4b · outbound
Gymnasium: A Standard Interface for Reinforcement Learning Environments Acme: A Research Framework for Distributed Reinforcement Learning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 3986d04a-32dc-415c-8fcb-d5f74d588f80 · outbound
Gymnasium: A Standard Interface for Reinforcement Learning Environments Open RL Benchmark: Comprehensive Tracked Experiments for Reinforcement Learning
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation a56f97f1-baa0-4e74-b9cb-987d8424917e · outbound
Gymnasium: A Standard Interface for Reinforcement Learning Environments Littman, and Anthony R
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 8dfd740d-e136-4f10-a72a-2dcf61f958e5 · outbound
Gymnasium: A Standard Interface for Reinforcement Learning Environments GPUDrive: Data-driven, multi-agent driving simulation at 1 million FPS
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 28b4214c-4e7e-482a-84d7-be39b08f8ab4 · outbound
Gymnasium: A Standard Interface for Reinforcement Learning Environments ViZDoom: A Doom-based AI Research Platform for Visual Reinforcement Learning
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 9f0b7086-9d97-4fd3-9433-66dfd4d0b058 · outbound
Gymnasium: A Standard Interface for Reinforcement Learning Environments Reinforcement Learning on Web Interfaces Using Workflow-Guided Exploration
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 1d64acc1-f01a-4972-8322-7a90de2d153e · outbound
Gymnasium: A Standard Interface for Reinforcement Learning Environments Playing Atari with Deep Reinforcement Learning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e8131224-c63d-4fd7-a1fc-9185ed1cd0b9 · outbound
Gymnasium: A Standard Interface for Reinforcement Learning Environments Human-level control through deep reinforcement learning
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 951569ed-fd37-43a5-82d5-0f64b2f2ceb3 · outbound
Gymnasium: A Standard Interface for Reinforcement Learning Environments Behaviour Suite for Reinforcement Learning
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 05960e65-0149-4b05-a8f7-30612bab9e4e · outbound
Gymnasium: A Standard Interface for Reinforcement Learning Environments Proximal Policy Optimization Algorithms
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation ba409965-1b10-484d-936f-75b5a6402f34 · outbound
Gymnasium: A Standard Interface for Reinforcement Learning Environments Mask Atari for Deep Reinforcement Learning as POMDP Benchmarks
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 7f1aee20-8479-4610-9cc4-a576405ad418 · outbound
Gymnasium: A Standard Interface for Reinforcement Learning Environments PufferLib: Making Reinforcement Learning Libraries and Environments Play Nice
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 40010bf7-6710-44fa-8473-23467e9d535f · outbound
Gymnasium: A Standard Interface for Reinforcement Learning Environments PyFlyt -- UAV Simulation Environments for Reinforcement Learning Research
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation aaf865b1-736a-4343-adaf-663e70cd6bfc · outbound
Gymnasium: A Standard Interface for Reinforcement Learning Environments Multiplayer Support for the Arcade Learning Environment
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation b6698461-fc85-434a-867e-7f4fd4d33335 · outbound
Gymnasium: A Standard Interface for Reinforcement Learning Environments Mujoco: A physics en- gine for model-based control, in: 2012 IEEE/RSJ International Con- ference on Intelligent Robots and Systems, IEEE
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 5d71b5d0-159d-40d0-a35c-4a37b0eccf94 · outbound
Gymnasium: A Standard Interface for Reinforcement Learning Environments 2020 , issn =
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 22b1d74f-e0e4-4154-9188-acb09ad8e412 · outbound
Gymnasium: A Standard Interface for Reinforcement Learning Environments Tianshou: a Highly Modularized Deep Reinforcement Learning Library
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 69376ba7-5857-4018-81b4-35b60e4d6ff4 · outbound
Gymnasium: A Standard Interface for Reinforcement Learning Environments Kenny Young and Tian Tian
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation b918eef3-a7da-4d4f-b861-6ee4a0c7baba · outbound
Gymnasium: A Standard Interface for Reinforcement Learning Environments MinAtar: An Atari-Inspired Testbed for Thorough and Reproducible Reinforcement Learning Experiments
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e9ac05e1-c4d7-4699-b34f-433c7a1c65d8 · outbound
Gymnasium: A Standard Interface for Reinforcement Learning Environments Siyuan Zhang and Nan Jiang
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation dcf5e911-491b-4013-b9ad-76fa0aa02644 · outbound
Gymnasium: A Standard Interface for Reinforcement Learning Environments Meta-World: A Benchmark and Evaluation for Multi-Task and Meta Reinforcement Learning
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 1ad05e3f-1abb-400b-8ee8-3e6c7d5bc467 · inbound
Fast State Stabilization using Deep Reinforcement Learning for Measurement-based Quantum Feedback Control Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e3984839-5ff8-4ae8-a313-c5bbae50d858 · inbound
Simultaneous Multi-die Floorplanning and Technology Assignment Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation dc687816-0c83-40b4-9119-be571fae50fa · inbound
Scalable Multi-Task Learning through Spiking Neural Networks with Adaptive Task-Switching Policy for Intelligent Autonomous Agents Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 76ee0218-c645-4b50-a96e-061bf73025c9 · inbound
Parameter-Efficient Distributional RL via Normalizing Flows and a Geometry-Aware Cram\'er Surrogate Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 26fcbf04-c1e7-4db9-9f69-92a9ccdc7c54 · inbound
Leveraging Analytic Gradients in Provably Safe Reinforcement Learning Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation f45010df-8e02-4ec1-adec-3a1a9117f319 · inbound
DR-SAC: Distributionally Robust Soft Actor-Critic for Reinforcement Learning under Uncertainty Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation ddd202c7-ae37-4a30-8447-3b99b380b73b · inbound
A Survey of Continual Reinforcement Learning Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation a59e7710-90a6-467e-bbb0-bb781937730f · inbound
Deep Double Q-learning Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation f7f68170-0374-4a9d-acd1-92ff3a8dc84a · inbound
Adaptive Ensemble Aggregation for Actor-Critics Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 5080fb25-fd1f-480b-8f92-a5523977057c · inbound
Multimodal Remote Inference Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 702ff302-e68b-4952-a337-6e711e05ce63 · inbound
A Review On Safe Reinforcement Learning Using Lyapunov and Barrier Functions Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 134
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 07600789-4410-451c-9e67-1c81d76bf389 · inbound
Frictional Q-Learning Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 6887066f-0c03-4bd1-97f9-612a5687b85b · inbound
Activation Function Design Sustains Plasticity in Continual Learning Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation b4df839b-8ef7-49df-8dae-7dab23f409c4 · inbound
Sample-Efficient and Smooth Cross-Entropy Method Model Predictive Control Using Deterministic Samples Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 0e918685-e03b-490d-8bf8-5b012e6cfbc8 · inbound
Transformer-Guided Deep Reinforcement Learning for Optimal Takeoff Trajectory Design of an eVTOL Drone Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 2424d9ec-c41b-4199-a6d3-96a67b382a48 · inbound
Hybrid-AIRL: Enhancing Inverse Reinforcement Learning with Supervised Expert Guidance Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 6fa6eba9-beb3-414e-8c84-3f9f0a826e2d · inbound
Bench-Push: Benchmarking Pushing-based Navigation and Manipulation Tasks for Mobile Robots Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3c5a018-ffd2-4a77-aa52-d87da4bbf2b6 · inbound
Neural CDEs as Correctors for Learned Time Series Models Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation d9f81cab-8180-4215-9a49-47f63f87206c · inbound
The HydroGym Reinforcement Learning Platform for Fluid Dynamics Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be4c433f-3587-481d-9bfb-28318e12f44e · inbound
About Time: Model-free Reinforcement Learning with Timed Reward Machines Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e6f94848-e6ac-44c4-b7c4-2a52a9098c75 · inbound
Quantum Nonlinearity for Optical Neural Computing Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be9ed1c7-6ae6-4714-acfd-bb9e1c5669a7 · inbound
Plug-and-Play Benchmarking of Reinforcement Learning Algorithms for Large-Scale Flow Control Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6654283-b7bc-4612-bdc2-65b07539f193 · inbound
GraphAllocBench: A Flexible Benchmark for Preference-Conditioned Multi-Objective Policy Learning Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 1983
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ecb2fed-6451-453d-9547-b6a274ba7a3f · inbound
Stabilizing the Q-Gradient Field for Policy Smoothness in Actor-Critic Methods Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aeb34403-e2bb-4364-9ba7-e810f91b3b1f · inbound
Agile Reinforcement Learning through Separable Neural Architecture and Applications Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 938737e7-a3c8-49b3-8879-d20a0ff525ce · inbound
SUSD: Structured Unsupervised Skill Discovery through State Factorization Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 2012
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46381aa2-bfae-4a83-83a6-21bda85e38d4 · inbound
The hidden risks of temporal resampling in clinical reinforcement learning Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 8ce43951-02aa-4c32-9da1-fd317ea5175d · inbound
Pushing Forward Pareto Frontiers of Proactive Agents with Behavioral Agentic Optimization Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d07f17c4-bf89-49c1-82cd-606914e9b0c2 · inbound
MoralityGym: A Benchmark for Evaluating Hierarchical Moral Alignment in Sequential Decision-Making Agents Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 106
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 68e8a048-bded-4632-a9cf-b15a1136354f · inbound
RLGT: A reinforcement learning framework for extremal graph theory Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 412567f6-b7de-4094-85a3-813aadda6717 · inbound
Asymptotically Optimal Sequential Testing with Markovian Data Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e04756ea-2138-4fbc-9f8b-6d4201bd6a28 · inbound
Online World Modeling Enables Real-World Inverse Reinforcement Learning from Observation Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b95d71bc-a1f9-49af-906e-eeb623321e46 · inbound
HY-WU (Part I): An Extensible Functional Neural Memory Framework and An Instantiation in Text-Guided Image Editing Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f82c125-154d-45af-aa97-ec4c8389fe68 · inbound
A Survey of Reinforcement Learning For Economics Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0983aeb4-6f21-4a5f-bc14-1ddbf0418525 · inbound
Automatic Generation of High-Performance RL Environments Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 0d12ab06-5198-4034-8ee9-93d26aa0eb85 · inbound
WestWorld: A Knowledge-Encoded Scalable Trajectory World Model for Diverse Robotic Systems Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 0b8a735f-c51c-41e5-a642-9a0b52630854 · inbound
LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation eb33c335-3a86-411a-93e6-c0eea64e0500 · inbound
Research Novelty in Information Systems Journals After ChatGPT: Differences Across Institutional Language Contexts Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2da56bd9-ac95-45da-9a9c-979ccd5fb88c · inbound
Temporal Logic Control of Nonlinear Stochastic Systems with Online Performance Optimization Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 82c0e00a-6532-41c3-b085-c0ed6ff1c9fc · inbound
FlashSAC: Fast and Stable Off-Policy Reinforcement Learning for High-Dimensional Robot Control Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation ac930a4a-9087-4b3b-8972-4a2775899f8e · inbound
FlashSAC: Fast and Stable Off-Policy Reinforcement Learning for High-Dimensional Robot Control Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 35d27494-b206-40b0-96cb-47564bfeb305 · inbound
Gym-Anything: Turn any Software into an Agent Environment Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation f5c93907-05d9-4019-bc1f-a31a7a913153 · inbound
RAMP: Hybrid DRL for Online Learning of Numeric Action Models Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation b0edf386-0068-491b-a43a-595b228b0729 · inbound
SafeAdapt: Provably Safe Policy Updates in Deep Reinforcement Learning Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation c71f7138-7a7a-4d13-9d0b-2260efe82d3f · inbound
GPU-Accelerated Continuous-Time Successive Convexification for Contact-Implicit Legged Locomotion Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e3c9d446-9494-4470-9654-b63c66a024a7 · inbound
[COMP25] The Automated Negotiating Agents Competition (ANAC) 2025 Challenges and Results Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 4f003452-8588-4760-9c6d-b6d6b6097cda · inbound
Beyond Single-Model Optimization: Preserving Plasticity in Continual Reinforcement Learning Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f60a74c0-6484-4a8a-8930-4ad7b94ce600 · inbound
Efficient Federated RLHF via Zeroth-Order Policy Optimization Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 67f64fa7-21cb-408d-bbad-219f4bdd540e · inbound
Beyond Bellman: High-Order Generator Regression for Continuous-Time Policy Evaluation Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation c24c3f55-33f8-4999-b328-8d0c7dd47223 · inbound
RL-ABC: Reinforcement Learning for Accelerator Beamline Control Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation c9ae02b6-77aa-452e-8942-b0f8eb253b20 · inbound
Cyber Defense Benchmark: Agentic Threat Hunting Evaluation for LLMs in SecOps Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 843bb37b-6d6e-4a03-9df5-8633bf4045d2 · inbound
Benefits of Low-Cost Bio-Inspiration in the Age of Overparametrization Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 846606e3-da43-46a0-9da3-4d1f94647727 · inbound
Replay-buffer engineering for noise-robust quantum circuit optimization Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation eb4729b3-9039-45f5-ae73-17271a39e760 · inbound
SpecRLBench: A Benchmark for Generalization in Specification-Guided Reinforcement Learning Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation aee757c0-49d9-46de-98a6-b48d0997e3b4 · inbound
KinDER: A Physical Reasoning Benchmark for Robot Learning and Planning Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 6cef9daf-55b2-4021-848d-24eb401a2ca5 · inbound
A High-Throughput Compute-Efficient POMDP Hide-And-Seek-Engine (HASE) for Multi-Agent Operations Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 1a18a7ab-c157-464b-8c4b-646c84dd8d1c · inbound
Your Loss is My Gain: Low Stake Attacks on Liquid Staking Pools Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 2180fb82-a11c-4be6-b42b-be9544e69170 · inbound
Breaking the Computational Barrier: Provably Efficient Actor-Critic for Low-Rank MDPs Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 5045a22c-35ab-447d-9b57-45b74d31c37b · inbound
A Multi-View Media Profiling Suite: Resources, Evaluation, and Analysis Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation c36c3f00-1402-40bc-8a79-fd8900a8d0af · inbound
PACE: Parameter Change for Unsupervised Environment Design Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e74b8414-3dea-40a7-a85c-03f7683dced1 · inbound
Perturb and Correct: Post-Hoc Ensembles using Affine Redundancy Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation b12f7b7a-b0d2-478c-be31-5fe0b5064eb7 · inbound
Stable GFlowNets with Probabilistic Guarantees Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 1e8a6a7f-572f-4583-9f9a-37aa4dd0cb5b · inbound
Zero-Shot, Safe and Time-Efficient UAV Navigation via Potential-Based Reward Shaping, Control Lyapunov and Barrier Functions Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 207a6697-804f-4d84-951d-7d65cd4d25e1 · inbound
Training Non-Differentiable Networks via Optimal Transport Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 3fd50936-1b7d-423e-9448-324f3a2867ae · inbound
Bridging the Gap Between Average and Discounted TD Learning Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 158da49e-1e54-4dae-bdf4-f70888606603 · inbound
ANO: A Principled Approach to Robust Policy Optimization Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 5b13e0b7-7c9e-44a6-83b8-8998fc5256c2 · inbound
Enhancing RL Generalizability in Robotics through SHAP Analysis of Algorithms and Hyperparameters Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 162b1949-5a21-4f25-87a3-c75fae374c83 · inbound
Enhancing RL Generalizability in Robotics through SHAP Analysis of Algorithms and Hyperparameters Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 67c23977-be50-4503-ba65-2dd429e54f54 · inbound
Extending Differential Temporal Difference Methods for Episodic Problems Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 8d3081a2-af6b-42e8-b0f6-04f455eb6e29 · inbound
Manifold Steering Reveals the Shared Geometry of Neural Network Representation and Behavior Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 160
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 95470c55-0608-4c4b-8887-a4a54a6580d5 · inbound
SOPE: Stabilizing Off-Policy Evaluation for Online RL with Prior Data Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation acd0fd55-2ae8-4abe-a073-f7b53325d2ee · inbound
SOPE: Stabilizing Off-Policy Evaluation for Online RL with Prior Data Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 336dcce8-05a3-40ab-803c-2688c86a6353 · inbound
Causal Reinforcement Learning for Complex Card Games: A Magic The Gathering Benchmark Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation a78fd08b-e7c9-46aa-9f4c-5d22555b8d53 · inbound
AdaGamma: State-Dependent Discounting for Temporal Adaptation in Reinforcement Learning Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation fa132447-0738-46ef-af65-b86162cde18b · inbound
Beyond the Independence Assumption: Finite-Sample Guarantees for Deep Q-Learning under $\tau$-Mixing Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e7486c67-4604-407b-be41-54b9499ba9b0 · inbound
Agentick: A Unified Benchmark for General Sequential Decision-Making Agents Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation a62cd7d9-b02e-4c33-9fba-1f821390babc · inbound
Agentick: A Unified Benchmark for General Sequential Decision-Making Agents Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation ec253524-8436-41c9-a6dd-a2256f406f7f · inbound
Reflective Prompted Policy Optimization: Trajectory-Grounded Revision and Salience Bias Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 32d6f508-ee97-4b25-990f-7081447cdc9b · inbound
Learning When to Stop: Selective Imitation Learning Under Arbitrary Dynamics Shift Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 77669a96-ff19-4e39-8ecc-1cc8cf0ad116 · inbound
Learning When to Stop: Selective Imitation Learning Under Arbitrary Dynamics Shift Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e046d3de-47df-49ed-a95b-636afa8a1255 · inbound
Towards Generalist Game Players: An Investigation of Foundation Models in the Game Multiverse Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 163
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 6ace471f-92f0-40f3-be52-3920914b9610 · inbound
Towards Generalist Game Players: An Investigation of Foundation Models in the Game Multiverse Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 163
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 6ab60d41-7c0f-4098-90bb-ad112c1ef884 · inbound
Robust Probabilistic Shielding for Safe Offline Reinforcement Learning Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 86cd9959-e117-4691-9a7a-f24d9fef8575 · inbound
TuniQ: Autotuning Compilation Passes for Quantum Workloads at Scale for Effectiveness and Efficiency Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 6c8800c4-36c0-480a-920a-9039effada32 · inbound
Augmented Lagrangian Method for Last-Iterate Convergence for Constrained MDPs Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation c0607a29-0cae-4f0b-88b5-5e8e094bac8d · inbound
Learning When to Act: Communication-Efficient Reinforcement Learning via Run-Time Assurance Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 7f60a866-ded5-4b7c-ae9c-3ad0fe4700ac · inbound
Critic-Driven Voronoi-Quantization for Distilling Deep RL Policies to Explainable Models Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation c6832d17-d591-4e2c-9597-18ad1de8311e · inbound
Chrono-Gymnasium: An Open-Source, Gymnasium-Compatible Distributed Simulation Framework Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 236bb79c-c07c-40fe-b1b0-2de68588f6d4 · inbound
Learning Selective Merge Policies for Deadline-Constrained Coded Caching via Deep Reinforcement Learning Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation f5214580-3491-490f-a012-b1ecb8bd8cc5 · inbound
Learning Selective Merge Policies for Deadline-Constrained Coded Caching via Deep Reinforcement Learning Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 6f7ee100-2056-4600-a74e-5b8bebabeb3b · inbound
AtlasVA: Self-Evolving Visual Skill Memory for Teacher-Free VLM Agents Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation ec72341c-efe2-43c7-b204-e6b7fd44e9e6 · inbound
OSCToM: RL-Guided Adversarial Generation for High-Order Theory of Mind Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 434d659b-ceb6-4c72-aad6-a97ac7ce6634 · inbound
Hack-Verifiable Environments: Towards Evaluating Reward Hacking at Scale Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 0d9ceeef-ef5f-4ede-a589-277ab7314472 · inbound
stable-worldmodel: A Platform for Reproducible World Modeling Research and Evaluation Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 2fb51365-d222-4e43-98f5-7ba26eb10f64 · inbound
Score-Based One-step MeanFlow Policy Optimization Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation f47238c6-5d8b-4993-80bc-02e0e25620f5 · inbound
Quantum Frog: Emergent Cooperation and Difficulty Scaling in a Quantized-Time Cooperative Game Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 0dfdb80a-7112-49be-92c5-ba8fc8deea52 · inbound
Distilling Game Code World Model Generation into Lightweight Large Language Models Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 85933b69-edae-42c1-a1db-25f63d3a3a95 · inbound
Cost-Aware Adaptive Conformal Inference for Runtime Assurance in Dynamic Environments Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 636da051-2c89-4e39-b7b0-f02b3547825b · inbound
When Does Adaptive Guidance Help? Belief-Aware Privileged Distillation for Autonomous Driving Under Partial Observability Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 99af8fe6-e639-42de-a804-1f26e58a7b5d · inbound
Adversarial Dual On-Policy Distillation from Expressive Teacher Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.