Pith. sign in

Paper Citation Record · LEDGER

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm

As of 18 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 1 inbound Pith citation observation for arXiv:2412.06139.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.06139 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T20:02:45.818390Z

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-10T16:53:48.187942Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T07:56:00.523665Z

Reference resolution

28 of 28 outbound references displayed

  • verified exact1
  • verified fuzzy19
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1dd89f82-0b1b-4c64-80bd-bc3fb0dd824a · outbound

This paper cites A survey on intrinsic motivation in reinforcement learning.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm A survey on intrinsic motivation in reinforcement learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T20:02:45.685985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:02:45.685985Z digest=sha256:4003c99acc3dac778c645a7e16634951211137752611d5fdcf8e674487a0a094

Observation 3613d47c-4572-450f-86a8-306a8719793e · outbound

This paper cites Learning and Policy Search in Stochastic Dynamical Systems with Bayesian Neural Networks.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm Learning and Policy Search in Stochastic Dynamical Systems with Bayesian Neural Networks

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T20:02:45.715371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:02:45.715371Z digest=sha256:4d36f0277e85127586b50c1b1b94a9d520870b5c15a4bf9498a033a2a921a4e4

Observation 7db7953c-11f5-493b-948d-6c466888f356 · outbound

This paper cites , 2017] Prafulla Dhariwal, Christopher Hesse, Oleg Klimov, Alex Nichol, Matthias Plappert, Alec Radford, John Schulman, Szymon Sidor, Yuhuai Wu, and Peter Zhokhov.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm , 2017] Prafulla Dhariwal, Christopher Hesse, Oleg Klimov, Alex Nichol, Matthias Plappert, Alec Radford, John Schulman, Szymon Sidor, Yuhuai Wu, and Peter Zhokhov

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.245200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:02:45.726904Z digest=sha256:5a61fab0f136e01249f74e928812abeb9c1bd6b4cf293a1f9caab86832215587

Observation f257143d-5db1-4930-84e8-20560d8a57b7 · outbound

This paper cites A comprehensive survey on safe reinforcement learning.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm A comprehensive survey on safe reinforcement learning

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.215485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:02:45.736303Z digest=sha256:d53d8d919749a646476b66f3ee422641896564cf0efb09877c72d7d873b95ff2

Observation ecb6fd95-aeb9-46fd-b05d-dd83104fcdd4 · outbound

This paper cites an unresolved cited work.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-11T20:02:46.194137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:02:45.745359Z digest=sha256:f5ced5b3f9395c87577a8059b188078d4d0424c40b8a4b74d7b9f320bde6c1a6

Observation 832226f7-308c-429b-bf3e-2faeed70392c · outbound

This paper cites , 2016] Rein Houthooft, Xi Chen, Yan Duan, John Schulman, Filip De Turck, and Pieter Abbeel.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm , 2016] Rein Houthooft, Xi Chen, Yan Duan, John Schulman, Filip De Turck, and Pieter Abbeel

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.171607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:02:45.750534Z digest=sha256:22a845b78eb982f31e6e9e6698105767272ffc3394f5b2f8829c6e1c781d16c7

Observation 104d555d-73c1-4c02-bd3f-a654af097d37 · outbound

This paper cites , 2019] Michael Janner, Justin Fu, Marvin Zhang, and Sergey Levine.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm , 2019] Michael Janner, Justin Fu, Marvin Zhang, and Sergey Levine

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.154522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:02:45.754158Z digest=sha256:09815adfb832cd4a131f713562e4fcd9b705fbc35a2ab80fe130235a35339b33

Observation b8658f0c-fd8e-45d1-aeb6-80ff90e71e7f · outbound

This paper cites What uncertainties do we need in bayesian deep learn- ing for computer vision? Advances in neural informa- tion processing systems, 30,.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm What uncertainties do we need in bayesian deep learn- ing for computer vision? Advances in neural informa- tion processing systems, 30,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.140403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:02:45.757808Z digest=sha256:a4a4ad8cb19448bfacb59165180e7f001b249d6a5c346c13995e812d6ae64db4

Observation 3dbbfe57-0df0-4f43-8f69-0378f032b92b · outbound

This paper cites , 2013] Jens Kober, J Andrew Bagnell, and Jan Peters.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm , 2013] Jens Kober, J Andrew Bagnell, and Jan Peters

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.126290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:02:45.762835Z digest=sha256:70584c3aa300fd44e054fbbc75822e55f4a3732292bdfba032030672447298dc

Observation f235896b-46c1-4921-9634-e50998ba371d · outbound

This paper cites Model-Ensemble Trust-Region Policy Optimization.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm Model-Ensemble Trust-Region Policy Optimization

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T20:02:45.769270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:02:45.769270Z digest=sha256:09f0da33ea4585bcddccb96b2236e922ed8a76d0aca4d429127df1877116e0fd

Observation 9a5326c9-2446-448c-92d4-6cf109c931b7 · outbound

This paper cites , 2022] Pawel Ladosz, Lilian Weng, Min- woo Kim, and Hyondong Oh.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm , 2022] Pawel Ladosz, Lilian Weng, Min- woo Kim, and Hyondong Oh

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.107074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:02:45.774118Z digest=sha256:d067208aaaf03c4f9693d4dd225e4d648be8727e713d867ffea4d8e999e46f9e

Observation 6a5e884e-7a78-4c31-a9f2-906831e5b99a · outbound

This paper cites Intrinsic Motivation in Model-based Reinforcement Learning: A Brief Review.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm Intrinsic Motivation in Model-based Reinforcement Learning: A Brief Review

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-11T20:02:45.915394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:02:45.779781Z digest=sha256:e8183eaf4fddc97a13b67cca1e63d3daa5bf28e49ac281a0f9804eae28ab8a79

Observation 278ad961-a352-4150-91e5-0bf82c1a9bd6 · outbound

This paper cites Algorithmic Framework for Model-based Deep Reinforcement Learning with Theoretical Guarantees.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm Algorithmic Framework for Model-based Deep Reinforcement Learning with Theoretical Guarantees

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T20:02:45.797201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:02:45.797201Z digest=sha256:cb8e23a80449cdcb596ace5a58b185a58fe83418d9f1d45f5bb25053aa89990d

Observation f7460c7f-3bf6-4d93-b407-66eb401f6c1f · outbound

This paper cites , 2018] Anusha Nagabandi, Gregory Kahn, Ronald S Fearing, and Sergey Levine.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm , 2018] Anusha Nagabandi, Gregory Kahn, Ronald S Fearing, and Sergey Levine

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.078816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:02:45.802267Z digest=sha256:656bf97e079ba67619c87517de42283c1ed61366f3ebe43f9da5e83647c0d075

Observation 643c7846-1ead-41ce-b151-86c8fc08c113 · outbound

This paper cites , 2017] Deepak Pathak, Pulkit Agrawal, Alexei A Efros, and Trevor Darrell.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm , 2017] Deepak Pathak, Pulkit Agrawal, Alexei A Efros, and Trevor Darrell

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.064087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:02:45.805967Z digest=sha256:7b0a3e47a4de5a68c6225ff0bf79dfdcece8067bf18e7bd7ed5fdf0ba8a92562

Observation 59f8fb34-c43f-4f75-a965-1be45d33603a · outbound

This paper cites , 2019] Deepak Pathak, Dhiraj Gandhi, and Abhinav Gupta.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm , 2019] Deepak Pathak, Dhiraj Gandhi, and Abhinav Gupta

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.048218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:02:45.810125Z digest=sha256:2a6915f8709c9b459fccf58bac112049d15a23798b8f9bdfded99761d89e3fc5

Observation 6e5d27a0-6282-4b9b-9e67-38bcefd2b87f · outbound

This paper cites , 2021] Yao Yao, Li Xiao, Zhicheng An, Wan- peng Zhang, and Dijun Luo.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm , 2021] Yao Yao, Li Xiao, Zhicheng An, Wan- peng Zhang, and Dijun Luo

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.030996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:02:45.813951Z digest=sha256:f5931d3272ec9f51a132808c877f943b480186e3d6760001656b734a661a2e28

Observation 4848ba8b-df00-4f91-bb80-6185fc03adc9 · outbound

This paper cites Modeling purpose- ful adaptive behavior with the principle of maximum causal entropy.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm Modeling purpose- ful adaptive behavior with the principle of maximum causal entropy

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.015578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:02:45.818390Z digest=sha256:ef9d303b668eaa24d4462d0a47271a4bb501f7211e4ca7d44a21cb5c4e5fc51c

Observation cd187864-74e3-43e8-ad71-1c1de22d8fc9 · outbound

This paper cites Intrinsic motivation and reinforcement learning.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm Intrinsic motivation and reinforcement learning

Reference 2002

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.331163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:02:45.698457Z digest=sha256:9befed508c099e6035e0e8a924482f7ef5445d7675f15780f9ae02b08473dd2a

Observation 41eb5a9b-bd06-49a4-9278-e764dc3e9c37 · outbound

This paper cites , 2022] Lukas Brunke, Melissa Greeff, Adam W Hall, Zhaocong Yuan, Siqi Zhou, Jacopo Panerati, and Angela P Schoellig.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm , 2022] Lukas Brunke, Melissa Greeff, Adam W Hall, Zhaocong Yuan, Siqi Zhou, Jacopo Panerati, and Angela P Schoellig

Reference 2013

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.319245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:02:45.702272Z digest=sha256:35217af1f98d62a90c806283b03bf54e0c413f28b14326a16e3de3090ce654a0

Observation f2e463cb-5094-4f04-aaaa-d901881a2eac · outbound

This paper cites Soft Actor-Critic Algorithms and Applications.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm Soft Actor-Critic Algorithms and Applications

Reference 2015

Resolution
unresolved
no resolver link, observed 2026-08-11T20:02:45.741095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:02:45.741095Z digest=sha256:0a2a3cb82543e9e7b5ce7051f0056dcf48eb90b8629f40b932a207fb97f63f50

Observation 45588b8e-38e7-4d7d-961a-8768d590bcda · outbound

This paper cites , 2018] Stefan Depeweg, Jose-Miguel Hernandez-Lobato, Finale Doshi-Velez, and Steffen Udluft.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm , 2018] Stefan Depeweg, Jose-Miguel Hernandez-Lobato, Finale Doshi-Velez, and Steffen Udluft

Reference 2016

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.268857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:02:45.719312Z digest=sha256:46a9b777fbc3c68099459760801f9a16f1b09ec58a876f89cf6abbca82b9cbd4

Observation 7c93570e-eced-4f8f-9682-b9aa755a03c3 · outbound

This paper cites Model-Based Value Estimation for Efficient Model-Free Reinforcement Learning.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm Model-Based Value Estimation for Efficient Model-Free Reinforcement Learning

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-11T20:02:45.731545Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:02:45.731545Z digest=sha256:82cedbf1f6cf2d1e298e45a3bb99639685c65d285959410bf1d7c48b5af9e8c1

Observation f06e91d1-0c3a-4c00-873e-36c1f4902579 · outbound

This paper cites , 2018] Kurtland Chua, Roberto Calandra, Rowan McAllister, and Sergey Levine.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm , 2018] Kurtland Chua, Roberto Calandra, Rowan McAllister, and Sergey Levine

Reference 2018

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.284785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:02:45.710787Z digest=sha256:aa6b8068774ef644f02b56d7c935e23da1164e4f9d317009621ae44158a8ec4a

Observation e0a1cbfa-b120-4af9-b12f-4dedd864262a · outbound

This paper cites , 2002] Peter Auer, Nicolo Cesa-Bianchi, and Paul Fischer.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm , 2002] Peter Auer, Nicolo Cesa-Bianchi, and Paul Fischer

Reference 2019

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.344031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:02:45.693655Z digest=sha256:a2240e071c26ea4601a43386d7be4f29643cadd784dd6da7c699961adac4c8ef

Observation 8d802246-0a33-4b64-8b21-221c48c55629 · outbound

This paper cites Reinforcement Learning and Control as Probabilistic Inference: Tutorial and Review.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm Reinforcement Learning and Control as Probabilistic Inference: Tutorial and Review

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-11T20:02:45.792466Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:02:45.792466Z digest=sha256:b28b82368bdf1667494bc1d3876f1d223c08cc76ae12e14961d723f13128da9f

Observation a28fbaf3-e78c-4031-abd6-d277e5351ab9 · outbound

This paper cites , 2018] Jacob Buckman, Danijar Hafner, George Tucker, Eugene Brevdo, and Honglak Lee.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm , 2018] Jacob Buckman, Danijar Hafner, George Tucker, Eugene Brevdo, and Honglak Lee

Reference 2022

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.305592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:02:45.706238Z digest=sha256:ded5d5186b3d56c4c7856ef08308bcd2b1fbffd6513c3548d47d2a5c71b7859a

Observation 3a72e50a-8791-4ccc-bc4b-fc8bed14c54e · outbound

This paper cites , 2021] Kimin Lee, Michael Laskin, Aravind Srinivas, and Pieter Abbeel.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm , 2021] Kimin Lee, Michael Laskin, Aravind Srinivas, and Pieter Abbeel

Reference 2023

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.092859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:02:45.786109Z digest=sha256:abc4f4d36b541363942a8f86aee4853c7a11cf44ffca7cb60750beda0acdbcb7

Pith citing papers

Observation 9670e234-2ab1-43de-8c13-e37aceaa7a1a · inbound

Morphology-Conditioned World Model for Cross-Embodiment Quadrupedal Locomotion cites this paper.

Morphology-Conditioned World Model for Cross-Embodiment Quadrupedal Locomotion Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:56:00.529175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T16:53:48.187942Z digest=sha256:039cf2fe83dccd10b44fe13ec8752c5cd18c52b2311bff7000e99c605914f3e2