Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-12T04:33:08.990277Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 0 inbound Pith citation observations for arXiv:2605.09157.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-12T04:33:08.990277Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
53 of 53 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 06d884c3-8bf8-4ebd-ae18-fcea3c060862 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Reinforcement learning: Theory and algorithms
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f1c6a39a-bc3c-4ee0-bfeb-cab54e2c18a7 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Understanding the impact of entropy on policy optimization
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b32971d6-eae0-44ce-a01b-38de8aa42cb1 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Maximum Entropy Reinforcement Learning with Mixture Policies
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e2be3b7e-52b7-4277-91c9-591d28336bc4 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic On the sample complexity and metastability of heavy-tailed policy search in continuous control
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 05b4147b-b880-403a-80eb-33d724805114 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Estimating or Propagating Gradients Through Stochastic Neurons for Conditional Computation
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2fc38d5b-9d69-4c65-9198-9ee53fdcec2c · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic JAX : composable transformations of P ython+ N um P y programs
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 5f3b3339-1883-4c3c-bca2-1c56f4043eb4 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic OpenAI Gym
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7db74fb0-6519-4492-b292-b43139624b99 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic On upper and lower bounds for the variance of a function of a random variable
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 3ba1d749-aa97-4144-b0ec-7bc13c1d5770 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Myosuite: A contact-rich simulation suite for musculoskeletal motor control
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 118e43a9-53e5-4c8f-b83e-3e05fb59c4eb · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Specializing versatile skill libraries using local mixture of experts
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 909acd8f-cbe1-45ec-9780-dd1435c3acb8 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Improving stochastic policy gradients in continuous control with deep reinforcement learning using the beta distribution
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a66c1beb-ea0d-428b-b567-aa9ab1150f10 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Hierarchical relative entropy policy search
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b1c7b8da-de2e-4be9-bd8a-ffea9a47fbc9 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Model-free reinforcement learning with continuous action in practice
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b32533d6-f5cf-4a08-b6c8-c9c783185888 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Addressing function approximation error in actor-critic methods
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6cb18d3e-4abf-4b81-998f-bf71e8ebb0c2 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Uncertainty in deep learning
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 62e38cab-91c7-4fe7-bfc4-ec742dbdce88 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Acquiring diverse robot skills via maximum entropy deep reinforcement learning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ee7d3d50-1cf3-4368-8805-dc3180aa5f10 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Reinforcement learning with deep energy-based policies
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6651af1a-3c40-41d8-9399-0b80f53dfc78 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ec6809bb-1d1e-429c-b991-6a88a7570380 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ab3fee38-09ea-4d14-8349-e644b5aee0e6 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Soft Actor-Critic Algorithms and Applications
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0cb0d849-adb8-470f-9f09-e0b366946746 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Learning latent dynamics for planning from pixels
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 24c601ab-44b5-42d1-bb89-1caa69715d2a · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Learning continuous control policies by stochastic value gradients
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9bdc9302-e4c8-40e7-962e-0a641dce73a8 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Off-policy Maximum Entropy Reinforcement Learning : Soft Actor-Critic with Advantage Weighted Mixture Policy(SAC-AWMP)
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 05787a78-2e17-40e7-8e28-bac7c0d6a8cd · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Generalization in Dexterous Manipulation via Geometry-Aware Multi-Task Learning
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 5bd5f78e-1ac0-4fc7-9760-2cdbb4358992 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Categorical reparameterization with gumbel-softmax
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 8204c685-b8ff-4d7a-8205-18a6fe6554bb · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Adam: A Method for Stochastic Optimization
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation edc0f208-e1a9-40ad-885c-6cfe51ab770f · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Student-t policy in reinforcement learning to acquire global optimum of robot control
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d19de21d-04e0-434b-8da2-c5d9069155be · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Model-free policy learning with reward gradients
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation edbf07ce-4e8c-4b45-81e6-b909cb9ab6a8 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Stochastic latent actor-critic: Deep reinforcement learning with a latent variable model
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 401bcb7b-f6a2-438f-a555-c3c2390f38ee · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Continuous control with deep reinforcement learning
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 5efc6f9f-c731-4be9-bd12-3b99f976e4c2 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic The concrete distribution: A continuous relaxation of discrete random variables
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 778148b4-3481-47bf-a690-7047da57cfa5 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Leveraging exploration in off-policy algorithms via normalizing flows
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 230c3560-1cd5-42af-9c9f-aaf452f73c92 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic S\ 2\ AC : Energy-based reinforcement learning with stein soft actor critic
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 886fd1d4-3473-43bd-9540-cfb50801a3fa · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Reducing reparameterization gradient variance
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 8d6b7756-aefc-4310-af21-97711315d561 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Monte carlo gradient estimation in machine learning
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6e72d158-8db5-491b-8aad-b2dafa102802 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Robot skill adaptation via soft actor-critic gaussian mixture models
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a00a0810-6996-4cd5-a81f-d8d212df6b29 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Greedy actor-critic: A new conditional cross-entropy method for policy improvement
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ac3acd16-06f0-43fe-a8b1-58c906e84da7 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Pytorch: An imperative style, high-performance deep learning library
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 74b000b0-27a9-498e-bb44-2d6f490ea4e3 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Multi-Goal Reinforcement Learning: Challenging Robotics Environments and Request for Research
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 94c67878-6e7c-49af-a7fb-a6104e4d653d · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Probabilistic Mixture-of-Experts for Efficient Deep Reinforcement Learning
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 64d3372a-b612-4f17-b336-2fe671f8e3d3 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Proximal Policy Optimization Algorithms
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation db6aecc0-cc08-491f-989d-461bb9b13083 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Strength through diversity: Robust behavior learning via mixture policies
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 829b5396-cb1f-47cb-a712-8bb4668cee7a · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Reinforcement learning: An introduction
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d2a5742c-9dbb-4f2d-851a-0a58ca80a436 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Policy gradient methods for reinforcement learning with function approximation
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 4feeb983-2734-402d-8b34-b4e00a139782 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Implicit Policy for Reinforcement Learning
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e07a08fa-4270-4e26-89b5-a175398eb3a5 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic DeepMind Control Suite
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b285df28-723d-41dc-9c03-e28c7ae86042 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic SciPy 1.0: fundamental algorithms for scientific computing in Python
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation fb737d3b-6080-41b3-b898-2656305d3810 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Diffusion policies as an expressive policy class for offline reinforcement learning
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 47089ed9-a9d7-4e04-812d-60245b9a7c24 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Simple statistical gradient-following algorithms for connectionist reinforcement learning
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7ceb34f8-9a68-479a-8a09-19a580605d33 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Variance reduction properties of the reparameterization trick
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 8be5421a-2b10-4700-88c2-94ef8f603f4d · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Meta-world: A benchmark and evaluation for multi-task and meta reinforcement learning
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 4f2c4491-6930-4690-9a65-3fa6f9570eb8 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Latent state marginalization as a low-cost approach for improving exploration
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 570cccc1-1752-43e0-b5fa-78d2ef814852 · outbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Model-based reparameterization policy gradient methods: Theory and practical algorithms
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
No inbound Pith citation observations are available.