Pith. sign in

Paper Citation Record · LEDGER

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm

As of 16 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 1 inbound Pith citation observation for arXiv:2412.06139.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.06139 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T20:02:45.818390Z

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-10T16:53:48.187942Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T07:56:00.523665Z

Reference resolution

28 of 28 outbound references displayed

  • verified exact1
  • verified fuzzy19
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1dd89f82-0b1b-4c64-80bd-bc3fb0dd824a · outbound

This paper cites A survey on intrinsic motivation in reinforcement learning.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm A survey on intrinsic motivation in reinforcement learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T20:02:45.685985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:02:45.685985Z digest=sha256:8f3d36a936e9f0aeb95a619959d6e7aaba97bb60b5e7439953d00b464c0e06c8

Observation 3613d47c-4572-450f-86a8-306a8719793e · outbound

This paper cites Learning and Policy Search in Stochastic Dynamical Systems with Bayesian Neural Networks.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm Learning and Policy Search in Stochastic Dynamical Systems with Bayesian Neural Networks

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T20:02:45.715371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:02:45.715371Z digest=sha256:3b635b780dd96269d0b5c0cfd381c5896bbcc37a5a061c678a7ca42b75915453

Observation 7db7953c-11f5-493b-948d-6c466888f356 · outbound

This paper cites , 2017] Prafulla Dhariwal, Christopher Hesse, Oleg Klimov, Alex Nichol, Matthias Plappert, Alec Radford, John Schulman, Szymon Sidor, Yuhuai Wu, and Peter Zhokhov.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm , 2017] Prafulla Dhariwal, Christopher Hesse, Oleg Klimov, Alex Nichol, Matthias Plappert, Alec Radford, John Schulman, Szymon Sidor, Yuhuai Wu, and Peter Zhokhov

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.245200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T20:02:45.726904Z digest=sha256:3d9ab6c01ab59880d2f283d805e3df7f828e0a6b85e5d29d202220807ca45e7f

Observation f257143d-5db1-4930-84e8-20560d8a57b7 · outbound

This paper cites A comprehensive survey on safe reinforcement learning.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm A comprehensive survey on safe reinforcement learning

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.215485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T20:02:45.736303Z digest=sha256:387d0bf528af2ae8a726adcff21c1f9ead2f15ca8e8aca9af6410791dbe2083e

Observation ecb6fd95-aeb9-46fd-b05d-dd83104fcdd4 · outbound

This paper cites an unresolved cited work.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-11T20:02:46.194137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T20:02:45.745359Z digest=sha256:6fe0a96e89e76f8aa4780122f3ea26612d3584090d6e98964085d11e577c5c98

Observation 832226f7-308c-429b-bf3e-2faeed70392c · outbound

This paper cites , 2016] Rein Houthooft, Xi Chen, Yan Duan, John Schulman, Filip De Turck, and Pieter Abbeel.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm , 2016] Rein Houthooft, Xi Chen, Yan Duan, John Schulman, Filip De Turck, and Pieter Abbeel

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.171607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T20:02:45.750534Z digest=sha256:55b2696fb279199f1b6346fe51817c02415e12d8f18b86f135ef9f558a4c8f3b

Observation 104d555d-73c1-4c02-bd3f-a654af097d37 · outbound

This paper cites , 2019] Michael Janner, Justin Fu, Marvin Zhang, and Sergey Levine.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm , 2019] Michael Janner, Justin Fu, Marvin Zhang, and Sergey Levine

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.154522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T20:02:45.754158Z digest=sha256:2d4c410bd23d67060d93961d73295d78ee65cc7dde2e3a6230d000f22824b602

Observation b8658f0c-fd8e-45d1-aeb6-80ff90e71e7f · outbound

This paper cites What uncertainties do we need in bayesian deep learn- ing for computer vision? Advances in neural informa- tion processing systems, 30,.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm What uncertainties do we need in bayesian deep learn- ing for computer vision? Advances in neural informa- tion processing systems, 30,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.140403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T20:02:45.757808Z digest=sha256:dd44055e0973fd0796a37386b7e6a4a5e58009e3a6eec2654b02672cd968b5c6

Observation 3dbbfe57-0df0-4f43-8f69-0378f032b92b · outbound

This paper cites , 2013] Jens Kober, J Andrew Bagnell, and Jan Peters.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm , 2013] Jens Kober, J Andrew Bagnell, and Jan Peters

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.126290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T20:02:45.762835Z digest=sha256:42ce60319e778ec62e3a1b02dfb92df196ba9cb7b8e4d09197404eec616b3ee8

Observation f235896b-46c1-4921-9634-e50998ba371d · outbound

This paper cites Model-Ensemble Trust-Region Policy Optimization.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm Model-Ensemble Trust-Region Policy Optimization

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T20:02:45.769270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:02:45.769270Z digest=sha256:5bc65ed33f3b563b3be861c16cf68fd71efcf9fe6ddbcbbdfce474c9464a0774

Observation 9a5326c9-2446-448c-92d4-6cf109c931b7 · outbound

This paper cites , 2022] Pawel Ladosz, Lilian Weng, Min- woo Kim, and Hyondong Oh.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm , 2022] Pawel Ladosz, Lilian Weng, Min- woo Kim, and Hyondong Oh

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.107074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T20:02:45.774118Z digest=sha256:0219aa7342b8ad16bd514c8a01d3224472027a3eb24d0a1226ce01b03d11039b

Observation 6a5e884e-7a78-4c31-a9f2-906831e5b99a · outbound

This paper cites Intrinsic Motivation in Model-based Reinforcement Learning: A Brief Review.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm Intrinsic Motivation in Model-based Reinforcement Learning: A Brief Review

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-11T20:02:45.915394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T20:02:45.779781Z digest=sha256:dcbf226ff498967f95615456702e8424875266fc46bebb26e9ba5ab1828fe395

Observation 278ad961-a352-4150-91e5-0bf82c1a9bd6 · outbound

This paper cites Algorithmic Framework for Model-based Deep Reinforcement Learning with Theoretical Guarantees.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm Algorithmic Framework for Model-based Deep Reinforcement Learning with Theoretical Guarantees

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T20:02:45.797201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:02:45.797201Z digest=sha256:440504306c4e942a42dc49addb112cc36a750b7133aea58648c9582007413ef7

Observation f7460c7f-3bf6-4d93-b407-66eb401f6c1f · outbound

This paper cites , 2018] Anusha Nagabandi, Gregory Kahn, Ronald S Fearing, and Sergey Levine.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm , 2018] Anusha Nagabandi, Gregory Kahn, Ronald S Fearing, and Sergey Levine

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.078816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T20:02:45.802267Z digest=sha256:2d147594d197577a7df8ec16703ddee44cbaf8e8e09869baad95bd60e45f9933

Observation 643c7846-1ead-41ce-b151-86c8fc08c113 · outbound

This paper cites , 2017] Deepak Pathak, Pulkit Agrawal, Alexei A Efros, and Trevor Darrell.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm , 2017] Deepak Pathak, Pulkit Agrawal, Alexei A Efros, and Trevor Darrell

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.064087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T20:02:45.805967Z digest=sha256:96a111258d709df34a499eb0e52a14fb58833939163067dd2a60c857add92a50

Observation 59f8fb34-c43f-4f75-a965-1be45d33603a · outbound

This paper cites , 2019] Deepak Pathak, Dhiraj Gandhi, and Abhinav Gupta.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm , 2019] Deepak Pathak, Dhiraj Gandhi, and Abhinav Gupta

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.048218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T20:02:45.810125Z digest=sha256:dc245720e46946059ad18dd0053eb96902ca904591c34a690a0e7cbca1928baa

Observation 6e5d27a0-6282-4b9b-9e67-38bcefd2b87f · outbound

This paper cites , 2021] Yao Yao, Li Xiao, Zhicheng An, Wan- peng Zhang, and Dijun Luo.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm , 2021] Yao Yao, Li Xiao, Zhicheng An, Wan- peng Zhang, and Dijun Luo

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.030996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T20:02:45.813951Z digest=sha256:372895c1a1e9c0aec13d4abf06a59aac8b4a6e6181348acf4a0e1a24463cebd0

Observation 4848ba8b-df00-4f91-bb80-6185fc03adc9 · outbound

This paper cites Modeling purpose- ful adaptive behavior with the principle of maximum causal entropy.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm Modeling purpose- ful adaptive behavior with the principle of maximum causal entropy

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.015578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T20:02:45.818390Z digest=sha256:4a82a4cea320bb253469a93ad80389733298147d89a7622a48ce020edcd7fe81

Observation cd187864-74e3-43e8-ad71-1c1de22d8fc9 · outbound

This paper cites Intrinsic motivation and reinforcement learning.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm Intrinsic motivation and reinforcement learning

Reference 2002

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.331163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T20:02:45.698457Z digest=sha256:59f249526b81b88df15c9b315311438aceb16c24273b66999d83b141bef2e404

Observation 41eb5a9b-bd06-49a4-9278-e764dc3e9c37 · outbound

This paper cites , 2022] Lukas Brunke, Melissa Greeff, Adam W Hall, Zhaocong Yuan, Siqi Zhou, Jacopo Panerati, and Angela P Schoellig.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm , 2022] Lukas Brunke, Melissa Greeff, Adam W Hall, Zhaocong Yuan, Siqi Zhou, Jacopo Panerati, and Angela P Schoellig

Reference 2013

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.319245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T20:02:45.702272Z digest=sha256:9b8cdedeebef44e43c99c56f569b2ce773a2a5dc59646104003340b1d50a7341

Observation f2e463cb-5094-4f04-aaaa-d901881a2eac · outbound

This paper cites Soft Actor-Critic Algorithms and Applications.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm Soft Actor-Critic Algorithms and Applications

Reference 2015

Resolution
unresolved
no resolver link, observed 2026-08-11T20:02:45.741095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:02:45.741095Z digest=sha256:6935d26a9cafb86b94828f4e1d93e87360e9d00fb6b6c97465800c2a046a548a

Observation 45588b8e-38e7-4d7d-961a-8768d590bcda · outbound

This paper cites , 2018] Stefan Depeweg, Jose-Miguel Hernandez-Lobato, Finale Doshi-Velez, and Steffen Udluft.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm , 2018] Stefan Depeweg, Jose-Miguel Hernandez-Lobato, Finale Doshi-Velez, and Steffen Udluft

Reference 2016

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.268857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T20:02:45.719312Z digest=sha256:ee1e7a7d9eae40b047ae6e5b55696023ce82092975e70aa9a5956a8cf1ebe621

Observation 7c93570e-eced-4f8f-9682-b9aa755a03c3 · outbound

This paper cites Model-Based Value Estimation for Efficient Model-Free Reinforcement Learning.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm Model-Based Value Estimation for Efficient Model-Free Reinforcement Learning

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-11T20:02:45.731545Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:02:45.731545Z digest=sha256:ac5849dc9801b8806374668c9d813d92d2824a5d875e25bbe5c8816652bdd527

Observation f06e91d1-0c3a-4c00-873e-36c1f4902579 · outbound

This paper cites , 2018] Kurtland Chua, Roberto Calandra, Rowan McAllister, and Sergey Levine.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm , 2018] Kurtland Chua, Roberto Calandra, Rowan McAllister, and Sergey Levine

Reference 2018

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.284785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T20:02:45.710787Z digest=sha256:bd2230aef21598d1253350dd50e3eccb7c00bb7344be08e8d11fa98fa3d38175

Observation e0a1cbfa-b120-4af9-b12f-4dedd864262a · outbound

This paper cites , 2002] Peter Auer, Nicolo Cesa-Bianchi, and Paul Fischer.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm , 2002] Peter Auer, Nicolo Cesa-Bianchi, and Paul Fischer

Reference 2019

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.344031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T20:02:45.693655Z digest=sha256:0daf527e107d764543c1cc8820a72d712cfd9913e53a6ca50b417bf76d6ef3a1

Observation 8d802246-0a33-4b64-8b21-221c48c55629 · outbound

This paper cites Reinforcement Learning and Control as Probabilistic Inference: Tutorial and Review.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm Reinforcement Learning and Control as Probabilistic Inference: Tutorial and Review

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-11T20:02:45.792466Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:02:45.792466Z digest=sha256:291304b49ce0adbe6139e7b6de55202e906c4397a966da7fba9055b92bbbdb4d

Observation a28fbaf3-e78c-4031-abd6-d277e5351ab9 · outbound

This paper cites , 2018] Jacob Buckman, Danijar Hafner, George Tucker, Eugene Brevdo, and Honglak Lee.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm , 2018] Jacob Buckman, Danijar Hafner, George Tucker, Eugene Brevdo, and Honglak Lee

Reference 2022

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.305592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T20:02:45.706238Z digest=sha256:ef03ae91942cdcb7f84fea0f0f534cd3fcfe95d8779ade97abb8332df6f03e76

Observation 3a72e50a-8791-4ccc-bc4b-fc8bed14c54e · outbound

This paper cites , 2021] Kimin Lee, Michael Laskin, Aravind Srinivas, and Pieter Abbeel.

Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm , 2021] Kimin Lee, Michael Laskin, Aravind Srinivas, and Pieter Abbeel

Reference 2023

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:02:46.092859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T20:02:45.786109Z digest=sha256:ca73f8f31cba98741c6a5dba95d842224b7c2cca10574d11717e0502b803a5cd

Pith citing papers

Observation 9670e234-2ab1-43de-8c13-e37aceaa7a1a · inbound

Toward Hardware-Agnostic Quadrupedal World Models via Morphology Conditioning cites this paper.

Toward Hardware-Agnostic Quadrupedal World Models via Morphology Conditioning Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:56:00.529175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T16:53:48.187942Z digest=sha256:e3a1686c578a4d5bf3d3893b718ea5b8fc715d1a3760f55f25c87f73ea72facf