Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-13T14:32:41.948156Z
Paper Citation Record · LEDGER
As of 5 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 100 inbound Pith citation observations for arXiv:1812.05905.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-13T14:32:41.948156Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T13:55:18.610364Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
13 of 13 outbound references displayed
External citation measurements
1955
pith, observed 2026-08-05T02:28:24.338817Z
Observation 120c6f10-402a-451f-a694-cd7837bfd158 · outbound
Soft Actor-Critic Algorithms and Applications Maximum a Posteriori Policy Optimisation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation db963ab7-95cf-4745-a90f-af682c73a1b5 · outbound
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3c434479-d9c5-428f-8cd0-956793b96077 · outbound
Soft Actor-Critic Algorithms and Applications Addressing Function Approximation Error in Actor-Critic Methods
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6214e5e7-2d98-4b8c-a768-3ca46054c422 · outbound
Soft Actor-Critic Algorithms and Applications The Reactor: A fast and sample-efficient Actor-Critic agent for Reinforcement Learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 9a3daa3f-8c80-4580-aaf7-70243aa818dc · outbound
Soft Actor-Critic Algorithms and Applications Q-Prop: Sample-Efficient Policy Gradient with An Off-Policy Critic
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4d1bfb09-22c5-4c4d-9d21-488274632bc0 · outbound
Soft Actor-Critic Algorithms and Applications Deep Reinforcement Learning that Matters
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c649c500-bee0-4e4b-b875-47f8f39a874f · outbound
Soft Actor-Critic Algorithms and Applications Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6b3b9993-beef-4bd2-8bc5-0ae548031259 · outbound
Soft Actor-Critic Algorithms and Applications Continuous control with deep reinforcement learning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a60cd17a-8fb2-439a-a9cb-89331ef8be1b · outbound
Soft Actor-Critic Algorithms and Applications Playing Atari with Deep Reinforcement Learning
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e9513ef5-76ce-46d4-aa1f-fe42ab84d8da · outbound
Soft Actor-Critic Algorithms and Applications Trust-PCL: An Off-Policy Trust Region Method for Continuous Control
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ed83ccbc-3954-48cc-bde3-f1919af2aca1 · outbound
Soft Actor-Critic Algorithms and Applications Equivalence Between Policy Gradients and Soft Q-Learning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c4c0db9e-1043-4521-a4a0-8772abbadb48 · outbound
Soft Actor-Critic Algorithms and Applications Dexterous Manipulation with Deep Reinforcement Learning: Efficient, General, and Low-Cost
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 37a23961-452a-4e81-8400-e5b869491c7f · outbound
Soft Actor-Critic Algorithms and Applications In that sense, discounted policy gradients typically do not optimize the true discounted objective
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 7da5b571-da5e-456f-877b-c72e364a954e · inbound
Solving Rubik's Cube with a Robot Hand Soft Actor-Critic Algorithms and Applications
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0b5d07ec-d09a-4f32-8783-5397d40a35de · inbound
robosuite: A Modular Simulation Framework and Benchmark for Robot Learning Soft Actor-Critic Algorithms and Applications
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1453e9c2-b39d-4c48-a384-d93b0f74a5a3 · inbound
FP-IRL: Fokker--Planck Inverse Reinforcement Learning -- A Physics-Constrained Approach to Markov Decision Processes Soft Actor-Critic Algorithms and Applications
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0387d6f7-4683-4d6b-abad-47135f081e56 · inbound
TD-MPC2: Scalable, Robust World Models for Continuous Control Soft Actor-Critic Algorithms and Applications
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e2cbd209-c7fb-4a05-b9ce-5170d3ce790d · inbound
CROP: Conservative Reward for Model-based Offline Policy Optimization Soft Actor-Critic Algorithms and Applications
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0b010a92-4cb6-4395-ae42-61d86fdc3713 · inbound
Koopman-Assisted Reinforcement Learning Soft Actor-Critic Algorithms and Applications
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e09dadb6-cde6-4d75-9b6d-d041c992053f · inbound
Optimal Gait Control for a Tendon-driven Soft Quadruped Robot by Model-based Reinforcement Learning Soft Actor-Critic Algorithms and Applications
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c2b80a4c-f453-45f2-9592-6320e793fbbe · inbound
FinTSB: A Comprehensive and Practical Benchmark for Financial Time Series Forecasting Soft Actor-Critic Algorithms and Applications
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b9a4d2fc-82d9-45f8-b5ea-511557d7374a · inbound
Kalman Filter Enhanced GRPO for Reinforcement Learning-Based Language Model Reasoning Soft Actor-Critic Algorithms and Applications
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 9760c728-ec29-4196-9e4b-5c7a7e81932b · inbound
Accelerated Learning with Linear Temporal Logic using Differentiable Simulation Soft Actor-Critic Algorithms and Applications
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e8e5e04d-8eb4-4849-94b5-d83b1ea8b6be · inbound
DR-SAC: Distributionally Robust Soft Actor-Critic for Reinforcement Learning under Uncertainty Soft Actor-Critic Algorithms and Applications
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 514b3a63-17a3-4436-9091-30655c9af17b · inbound
Deep Double Q-learning Soft Actor-Critic Algorithms and Applications
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 441c3157-77e6-4b7d-8c39-ea13c868c73b · inbound
EXPO: Stable Reinforcement Learning with Expressive Policies Soft Actor-Critic Algorithms and Applications
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1ecd2410-400d-4ea5-9193-50e0c50907ba · inbound
Joint Scheduling of Deferrable and Nondeferrable Demand with Colocated Stochastic Supply Soft Actor-Critic Algorithms and Applications
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0926d32f-33a8-4238-a6a1-9a2b6d55eefe · inbound
Relative Entropy Pathwise Policy Optimization Soft Actor-Critic Algorithms and Applications
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3c12a234-e4d3-4de2-b97d-61f99e7fb5d7 · inbound
Adaptive Ensemble Aggregation for Actor-Critics Soft Actor-Critic Algorithms and Applications
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 532e41a7-e47f-452f-afcb-f6236261a1d2 · inbound
Centralized Adaptive Sampling for Reliable Co-Training of Independent Multi-Agent Policies Soft Actor-Critic Algorithms and Applications
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation dacb7721-7cc3-412a-aec8-d89b4dd0d079 · inbound
First Order Model-Based RL through Decoupled Backpropagation Soft Actor-Critic Algorithms and Applications
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4ab26b5-cdc7-4628-ab95-277411eca475 · inbound
Pinching Antenna System (PASS) Enhanced Covert Communications: Against Warden via Sensing Soft Actor-Critic Algorithms and Applications
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c81b2da2-632d-453e-b65a-b60dc87ec669 · inbound
Attention and Risk-Aware Decision Framework for Safe Autonomous Driving Soft Actor-Critic Algorithms and Applications
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ecb03fdc-7934-4344-b447-05ecf8c4d8c3 · inbound
Dissecting Discrete Soft Actor-Critic: Limitations and Principled Alternatives Soft Actor-Critic Algorithms and Applications
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation cb6c0065-2f26-432c-a2fa-7dd6b8c9e369 · inbound
Real-time reinforcement learning for turbulent state-dependent control in a bluff-body wake Soft Actor-Critic Algorithms and Applications
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 05af16a5-ca92-4f9a-8214-d4dd341bcb93 · inbound
CORB-Planner: Corridor as Observations for RL Planning in High-Speed Flight Soft Actor-Critic Algorithms and Applications
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6996be6-81ad-4798-8b84-3b0317f39487 · inbound
High-Precision and High-Efficiency Trajectory Tracking for Excavators Based on Closed-Loop Dynamics Soft Actor-Critic Algorithms and Applications
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation bfb88eb3-2014-44bd-b51e-3d4650a46604 · inbound
Optimisation of Resource Allocation in Heterogeneous Wireless Networks Using Deep Reinforcement Learning Soft Actor-Critic Algorithms and Applications
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 358cb053-1672-455c-8516-d252540d03e6 · inbound
Optimisation of Resource Allocation in Heterogeneous Wireless Networks Using Deep Reinforcement Learning Soft Actor-Critic Algorithms and Applications
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f473e237-c902-4191-878e-2294cefac0a7 · inbound
Use the Online Network If You Can: Towards Fast and Stable Reinforcement Learning Soft Actor-Critic Algorithms and Applications
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 9e3c1d92-83e3-4ca5-88ba-e1ecd28a8324 · inbound
Optimal control of the future via prospective learning with control Soft Actor-Critic Algorithms and Applications
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 73d398b9-9885-420e-993b-52bf937ebf85 · inbound
Self-Supervised Multisensory Pretraining for Contact-Rich Robot Reinforcement Learning Soft Actor-Critic Algorithms and Applications
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f6c6dc51-1296-4c12-98e6-e1a0679284c3 · inbound
Self-Supervised Multisensory Pretraining for Contact-Rich Robot Reinforcement Learning Soft Actor-Critic Algorithms and Applications
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fcdc9dcd-3e93-43ce-bb7e-19946af7fa7c · inbound
FireScope: Wildfire Risk Raster Prediction with a Chain-of-Thought Oracle Soft Actor-Critic Algorithms and Applications
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a1643d85-eec6-468e-87a1-d17242d0dd14 · inbound
FireScope: Wildfire Risk Raster Prediction with a Chain-of-Thought Oracle Soft Actor-Critic Algorithms and Applications
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 95d99f5d-bda2-4a7c-b026-08f969a348f9 · inbound
R2PS: Worst-Case Robust Real-Time Pursuit Strategies under Partial Observability Soft Actor-Critic Algorithms and Applications
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b69198c5-b6ba-4006-961e-e3c6265b7ae8 · inbound
Prismatic World Model: Learning Compositional Dynamics for Planning in Hybrid Systems Soft Actor-Critic Algorithms and Applications
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 9e980c30-3165-4cd3-bca0-3ba4620548e6 · inbound
RL-AWB: Deep Reinforcement Learning for Auto White Balance Correction in Low-Light Night-time Scenes Soft Actor-Critic Algorithms and Applications
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f68c516d-3a0b-4499-8585-60b1fd0c7a06 · inbound
RL-AWB: Deep Reinforcement Learning for Auto White Balance Correction in Low-Light Night-time Scenes Soft Actor-Critic Algorithms and Applications
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74b6f97a-63fa-4c28-9916-1c04bc9487b0 · inbound
Coupling Smoothed Particle Hydrodynamics with Multi-Agent Deep Reinforcement Learning for Cooperative Control of Point Absorbers Soft Actor-Critic Algorithms and Applications
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17303f39-74c6-4d13-8187-ce9fbfcc348b · inbound
From Classical to Quantum Reinforcement Learning and Its Applications in Quantum Control: A Beginner's Tutorial Soft Actor-Critic Algorithms and Applications
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eefce765-f798-4859-8b44-f3855decb498 · inbound
Error Amplification Limits ANN-to-SNN Conversion in Continuous Control Soft Actor-Critic Algorithms and Applications
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b401733e-c3c8-42a7-8dbd-bdf11baa7076 · inbound
Agile Reinforcement Learning through Separable Neural Architecture and Applications Soft Actor-Critic Algorithms and Applications
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 246dd066-841b-41c7-a065-a2bd62a92113 · inbound
LC-SAC: Lyapunov-Constrained Soft Actor-Critic via Koopman Operator Theory for Trajectory Tracking and Stabilization Soft Actor-Critic Algorithms and Applications
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f273cae3-e53d-4d1c-8099-b10c5b7bf316 · inbound
Coupled Local and Global World Models for Efficient First Order RL Soft Actor-Critic Algorithms and Applications
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7789059b-9a07-42cb-a407-bdfedfc5dea1 · inbound
Robust SAC-Enabled UAV-RIS Assisted Secure MISO Systems With Untrusted EH Receivers Soft Actor-Critic Algorithms and Applications
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 612cc045-a79b-4a03-ad1e-82dd76e199ed · inbound
Maximin Robust Bayesian Experimental Design Soft Actor-Critic Algorithms and Applications
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 486c2dc2-3e05-461c-86b7-7b244acf5b49 · inbound
Simulation Distillation: Pretraining World Models in Simulation for Rapid Real-World Adaptation Soft Actor-Critic Algorithms and Applications
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 95851a73-591a-427b-802f-61c734bfc990 · inbound
DRL-Based Spectrum Sharing for RIS-Aided Local High-Quality Wireless Networks Soft Actor-Critic Algorithms and Applications
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f36c4ba-1005-4714-82f6-deffa8506f90 · inbound
Delayed homomorphic reinforcement learning for environments with delayed feedback Soft Actor-Critic Algorithms and Applications
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d4321f1b-bd07-41da-ab1b-6dc4d73471b0 · inbound
Incremental Residual Reinforcement Learning Toward Real-World Learning for Social Navigation Soft Actor-Critic Algorithms and Applications
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e3da9749-3322-4bf0-b126-e5faa8493991 · inbound
Incremental Residual Reinforcement Learning Toward Real-World Learning for Social Navigation Soft Actor-Critic Algorithms and Applications
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad244b88-38e9-479f-a4f9-8c82dd39c708 · inbound
SafeMind: A Risk-Aware Differentiable Control Framework for Adaptive and Safe Quadruped Locomotion Soft Actor-Critic Algorithms and Applications
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 161cb248-9b4a-4ba7-a6b3-3289aa1add26 · inbound
New Scheme Adaption Strategy for Hyperbolic Conservation Laws Soft Actor-Critic Algorithms and Applications
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a25bd6b-59ec-4436-a3dc-1a4cc8762a3f · inbound
Physics-Informed Reinforcement Learning of Spatial Density Velocity Potentials for Map-Free Racing Soft Actor-Critic Algorithms and Applications
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 7e004d75-46f9-47e9-9147-4cd2a2c46b10 · inbound
Simple but Stable, Fast and Safe: Achieve End-to-end Control by High-Fidelity Differentiable Simulation Soft Actor-Critic Algorithms and Applications
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation bcd90826-73e6-42d1-a4e3-db7bc3fc3c28 · inbound
Autonomous Diffractometry Enabled by Visual Reinforcement Learning Soft Actor-Critic Algorithms and Applications
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 246ae235-502a-4748-8ab8-5f186da7c2d8 · inbound
Flexible Empowerment at Reasoning with Extended Best-of-N Sampling Soft Actor-Critic Algorithms and Applications
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f429daf2-bae0-43f2-8041-b6cc0bf7ff14 · inbound
Self-Predictive Representation for Autonomous UAV Object-Goal Navigation Soft Actor-Critic Algorithms and Applications
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 08555fc4-88c0-4ca8-baca-aee0bf2a8c7d · inbound
KinDER: A Physical Reasoning Benchmark for Robot Learning and Planning Soft Actor-Critic Algorithms and Applications
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0dfdda28-ccf0-44f6-b255-fcc7cc77e079 · inbound
Atomic-Probe Governance for Skill Updates in Compositional Robot Policies Soft Actor-Critic Algorithms and Applications
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a8d20b2b-5a54-45d8-b892-a8e20b0a844b · inbound
Atomic-Probe Governance for Skill Updates in Compositional Robot Policies Soft Actor-Critic Algorithms and Applications
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 50b569e9-fbc3-4852-9b7e-873121f41dbf · inbound
Odysseus: Scaling VLMs to 100+ Turn Decision-Making in Games via Reinforcement Learning Soft Actor-Critic Algorithms and Applications
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c64ff3a5-6cab-4554-a6d5-fdbca793c2f5 · inbound
Counter-Dyna: Data-Efficient RL-Based HVAC Control using Counterfactual Building Models Soft Actor-Critic Algorithms and Applications
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3edbc64a-2ec4-4ecc-ac04-2cc95416e8ad · inbound
Offline Reinforcement Learning for Rotation Profile Control in Tokamaks Soft Actor-Critic Algorithms and Applications
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 46b446e5-9acb-40bb-bceb-fb78c2e2e1cf · inbound
Beyond Penalization: Diffusion-based Out-of-Distribution Detection and Selective Regularization in Offline Reinforcement Learning Soft Actor-Critic Algorithms and Applications
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 7f6539ee-7775-4f35-8318-14e0b1b27013 · inbound
REAP: Reinforcement-Learning End-to-End Autonomous Parking with Gaussian Splatting Simulator for Real2Sim2Real Transfer Soft Actor-Critic Algorithms and Applications
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f728f5d6-0d94-4980-8799-36e8b2609b69 · inbound
Generative Actor-Critic with Soft Bridge Policies Soft Actor-Critic Algorithms and Applications
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e11322e0-4c08-4519-9884-45916e93b863 · inbound
A Single Deep Preference-Conditioned Policy for Learning Pareto Coverage Sets Soft Actor-Critic Algorithms and Applications
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ab3fee38-09ea-4d14-8349-e644b5aee0e6 · inbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Soft Actor-Critic Algorithms and Applications
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 7e07ff32-fd7a-4938-90f1-87ac4ebf505a · inbound
TuniQ: Autotuning Compilation Passes for Quantum Workloads at Scale for Effectiveness and Efficiency Soft Actor-Critic Algorithms and Applications
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 90e442bd-16f2-4b8b-8e18-88ce7c9cc47c · inbound
Nautilus: From One Prompt to Plug-and-Play Robot Learning Soft Actor-Critic Algorithms and Applications
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c059de7c-b138-4742-8226-202d6bbb6492 · inbound
Nautilus: From One Prompt to Plug-and-Play Robot Learning Soft Actor-Critic Algorithms and Applications
Reference 99
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8dbbcbc8-35ec-49d4-ad5d-bdbd6a21e90f · inbound
Debiased Model-based Representations for Sample-efficient Continuous Control Soft Actor-Critic Algorithms and Applications
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 7dbaf19c-c29e-41af-a8bc-c7fbe3ff6ca9 · inbound
Self-Supervised On-Policy Reinforcement Learning via Contrastive Proximal Policy Optimisation Soft Actor-Critic Algorithms and Applications
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 60af9d34-8a1a-4a3a-b563-8b928eb23b68 · inbound
R2R2: Robust Representation for Intensive Experience Reuse via Redundancy Reduction in Self-Predictive Learning Soft Actor-Critic Algorithms and Applications
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e5f0bbdb-9bdf-46a2-825c-660402716ff9 · inbound
Policy Optimization in Hybrid Discrete-Continuous Action Spaces via Mixed Gradients Soft Actor-Critic Algorithms and Applications
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 9c5eccaa-60df-4a4e-a737-9e2a990171ff · inbound
Distributionally Robust Multi-Task Reinforcement Learning via Adaptive Task Sampling Soft Actor-Critic Algorithms and Applications
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation be7ff438-f973-4d72-b57e-75db7d7bf3c0 · inbound
EfficientTDMPC: Improved MPC Objectives for Sample-Efficient Continuous Control Soft Actor-Critic Algorithms and Applications
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f764aca8-ad6f-40ba-9bc5-752ad0dff51b · inbound
COOPO: Cyclic Offline-Online Policy Optimization Algorithm Soft Actor-Critic Algorithms and Applications
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 68510778-92a3-4b7a-95d0-7640fa2cc6e5 · inbound
Unleashing the Power of Tree-of-Thoughts for Edge-Enabled AIGC Service Provisioning Soft Actor-Critic Algorithms and Applications
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c1d18125-8e2e-43bd-8c0e-ead9a19d46fc · inbound
Partial Fusion of Neural Networks: Efficient Tradeoffs Between Ensembles and Weight Aggregation Soft Actor-Critic Algorithms and Applications
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b0840afa-5357-4aab-939a-6edec6a4c150 · inbound
Reflex: Reinforcement Learning with Reflection Symmetry Exploitation in State-Based Continuous Control Soft Actor-Critic Algorithms and Applications
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation eacdfee5-c726-48e8-9c5e-b414eb30ab37 · inbound
Goal-Conditioned Agents that Learn Everything All at Once Soft Actor-Critic Algorithms and Applications
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a7e7025e-882f-41bf-8b02-e1785d540b8e · inbound
Understanding Goal Generalisation in Sequential Reinforcement Learning Soft Actor-Critic Algorithms and Applications
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 346445e5-14e9-4adf-812b-2e032678a10c · inbound
Word Class Representations Spontaneously Emerge from Successor Representations Trained on Natural Language Soft Actor-Critic Algorithms and Applications
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a6b07761-01a5-4046-a14f-e24972cf40f7 · inbound
Bridging the Gap: Enabling Soft Actor Critic for High Performance Legged Locomotion Soft Actor-Critic Algorithms and Applications
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a3802ecb-525b-4e2a-b5d7-ad2cb187694e · inbound
ParkingWorld: End-to-End Autonomous Parking Reinforcement Learning from Corrective Experience in 3DGS Simulation Soft Actor-Critic Algorithms and Applications
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d809ab9a-ecb5-4227-8481-a457ed53e326 · inbound
EXPO-FT: Sample-Efficient Reinforcement Learning Finetuning for Vision-Language-Action Models Soft Actor-Critic Algorithms and Applications
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1af66703-cc0f-47dd-b29e-8b324fe69872 · inbound
Scaling World-Model Reinforcement Learning Through Diffusion Policy Optimization Soft Actor-Critic Algorithms and Applications
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2f4f78f1-c0fa-4fbf-8729-9d23ad3aec25 · inbound
Heterogeneous AAV Logistics Task Allocation: A Reinforcement Learning Enhanced Overlapping Coalition Formation Game Approach Soft Actor-Critic Algorithms and Applications
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f7fa21a0-6846-4e33-b6cc-81b67887081e · inbound
Efficient On-policy Visual-RL via Stochastic Decoupled Policy Gradient Soft Actor-Critic Algorithms and Applications
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 75d29eb4-16df-438c-b73d-d77b161a34f6 · inbound
Adaptive Reinforcement Learning for Robust Open Quantum System Control: A Multi-Task Framework with Temporal Optimization Soft Actor-Critic Algorithms and Applications
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 9642b65c-f81c-4540-b7d6-0277269d4224 · inbound
ZAPS-DA: Zero-Phase Action Policy Smoothing with Decoupled Actor for Continuous Control in Reinforcement Learning Soft Actor-Critic Algorithms and Applications
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5add10f4-062f-4cc2-a46e-fcc1e5c470c2 · inbound
ZAPS-DA: Zero-Phase Action Policy Smoothing with Decoupled Actor for Continuous Control in Reinforcement Learning Soft Actor-Critic Algorithms and Applications
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 908e6a21-0cb5-4c9d-93c0-24de49ec733a · inbound
Direct High-Magnetic-Field Coupling to Stripe Order in a Cuprate Superconductor Soft Actor-Critic Algorithms and Applications
Reference 137
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 852670b4-f74e-4d17-bb46-63804936b03f · inbound
Space-sampled Value Decay: Forgetting Mechanisms for Non-stationary Deep Reinforcement Learning Soft Actor-Critic Algorithms and Applications
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 709adedb-9564-40d2-9354-47072441e2ad · inbound
Deep Reinforcement Learning for Adaptive Power Allocation in ISAC Systems with Mobile Target Soft Actor-Critic Algorithms and Applications
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 96c91ba2-5833-4094-aa8d-80ab9e5d1b00 · inbound
Time-Slotted Multi-Cluster UAV AirComp with Energy-Awareness: A Pointer Network-Assisted Soft Actor-Critic Learning Framework Soft Actor-Critic Algorithms and Applications
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 199b1bbe-590a-4160-abb5-1ba3ecfa0379 · inbound
Augmenting Game AI with Deep Reinforcement Learning Soft Actor-Critic Algorithms and Applications
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 631f69be-bbcb-42b8-8a4f-ffa630c6ebc2 · inbound
ReLaTS: a Reinforcement Learning-based method for dynamically determining the coupling Time Step in multi-scale simulations of self-gravitating systems Soft Actor-Critic Algorithms and Applications
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 8fd919cc-c036-4be4-8388-ec4d777416f7 · inbound
Reward-free Pretraining for Reinforcement Learning via Occupancy Coverage Maximization Soft Actor-Critic Algorithms and Applications
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f78aa3de-6ca5-43b0-9713-86d047a9f5f5 · inbound
Modularized Reinforcement Learning on LLMs: From MDP Creation to Exploration and Learning Soft Actor-Critic Algorithms and Applications
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.