Pith. sign in

Paper Citation Record · LEDGER

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation

As of 20 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 4 inbound Pith citation observations for arXiv:2412.12089.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.12089 v2

Coverage vector

measured 37 of 37 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T14:25:18.377948Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T11:03:31.098377Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-10T05:25:54.539497Z

Reference resolution

37 of 37 outbound references displayed

  • verified exact0
  • verified fuzzy18
  • unresolved19
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7f41b60c-134a-4edd-8cc0-81846c39d9dd · outbound

This paper cites an unresolved cited work.

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-11T14:25:19.120281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:25:18.291827Z digest=sha256:1457e617f5d0f503b65380409009840479e5c76f1949f268c54bab929fde7d26

Observation b419e46c-fec9-4f62-bd09-d8490d887d7b · outbound

This paper cites an unresolved cited work.

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-11T14:25:19.037570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:25:18.322804Z digest=sha256:59e46562d5a352797b3fcbe17e25f4ff948ff2d51ae4438d7571aa654d8df665

Observation 62afaa51-dc9b-4aa3-a093-68423ef2eef3 · outbound

This paper cites Simplifying Deep Temporal Difference Learning.

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation Simplifying Deep Temporal Difference Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T14:25:18.111207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:25:18.111207Z digest=sha256:0a7a79dc1b7d0f5449ce896e18ad6322915c4fafc5b8b7dd5d67995869e461a0

Observation f987b00d-d695-42eb-a626-7239356963ae · outbound

This paper cites We denote the goal state distribution ρ∗ if the task defines it.

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation We denote the goal state distribution ρ∗ if the task defines it

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:25:18.987155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:25:18.340367Z digest=sha256:56dbb1b55de5f3ac264d11b5b679a54c137223bf4eb3e2a998ed1e51127f74b7

Observation 8ac1a36d-75db-4e93-94ff-913231d78237 · outbound

This paper cites Distributed prioritized experience replay.

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation Distributed prioritized experience replay

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:25:19.362611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:25:18.134424Z digest=sha256:634ab4d6c18f26bb92f448390e874d5cad794500eed97759701027a3b0821a3e

Observation 28971864-133c-45bb-a2b5-a4c13ce62cbd · outbound

This paper cites Dojo: A Differentiable Physics Engine for Robotics.

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation Dojo: A Differentiable Physics Engine for Robotics

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T14:25:18.140983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:25:18.140983Z digest=sha256:f317c8dbffb6d1c879a33ac476a9ac69b22c051e4b8474988d8f7876ca255ccf

Observation 264b630d-494c-4f26-813e-0836c279dddc · outbound

This paper cites Path integrals and symmetry breaking for optimal control theory.

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation Path integrals and symmetry breaking for optimal control theory

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:25:19.332846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:25:18.146661Z digest=sha256:ede86646726a23320d845c00545b82ceae806acd9de5afb01a2038b1e2765210

Observation 302b6345-ee9e-47c0-9f5a-60077200ad9f · outbound

This paper cites Planning with spatial-temporal abstraction from point clouds for deformable object manipulation.

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation Planning with spatial-temporal abstraction from point clouds for deformable object manipulation

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:25:19.309051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:25:18.153676Z digest=sha256:dc09da17c629171a1013ab70bddef9d2c9e1d0649e14764fa9064ffb13236504

Observation 2d62166a-a1f3-403c-a653-f6e671ba0197 · outbound

This paper cites Empirical Design in Reinforcement Learning.

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation Empirical Design in Reinforcement Learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T14:25:18.175661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:25:18.175661Z digest=sha256:31a9cad2f3e918d00e1d5d0888f8512ec14e8001d2ef6bd2bf53fbfcafa20c6d

Observation 2cfd67f5-9273-43b1-b5b2-fdb0940599e8 · outbound

This paper cites Diffmimic: Efficient motion mimicking with differentiable physics.

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation Diffmimic: Efficient motion mimicking with differentiable physics

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:25:19.286345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:25:18.183285Z digest=sha256:f0d3be8c3d4ae86af7bf77d94448dbed7ace085b6ead015ecd09a31b184a0eb5

Observation 7d48e508-54e3-4987-9f99-ea6190c75cae · outbound

This paper cites High-Dimensional Continuous Control Using Generalized Advantage Estimation.

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation High-Dimensional Continuous Control Using Generalized Advantage Estimation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T14:25:18.204109Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:25:18.204109Z digest=sha256:3d904e76a587723431a0bed5e8f88b54953a1c7ef20a7387222ac0148f18ec35

Observation bc3d8159-d346-4001-b0d8-635058cbca9a · outbound

This paper cites Learning Deployable Locomotion Control via Differentiable Simulation.

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation Learning Deployable Locomotion Control via Differentiable Simulation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T14:25:18.214365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:25:18.214365Z digest=sha256:a988edd905f501fcfd2829018be26544486279229421a18405969c08e7a6e4f3

Observation 8c6e3295-80f7-4ebe-a84d-3c041583ffde · outbound

This paper cites Issues in using function approximation for reinforcement learning.

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation Issues in using function approximation for reinforcement learning

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:25:19.256385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:25:18.221798Z digest=sha256:e2a2f8a1a21159f6591643ae75a3dd1569a73b0262e70df24c5944bd89ebffcd

Observation 7d2fa325-137c-4857-baab-9ce6c67ad259 · outbound

This paper cites Do You Need the Entropy Reward (in Practice)?.

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation Do You Need the Entropy Reward (in Practice)?

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T14:25:18.230974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:25:18.230974Z digest=sha256:cd7ae24436b741a5db93325c1084cc5fe84791e6adbf0a4722ee61ebad83916d

Observation 951ab7bd-f575-4b3e-85f4-d69788a0d94b · outbound

This paper cites Back to Newton's Laws: Learning Vision-based Agile Flight via Differentiable Physics.

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation Back to Newton's Laws: Learning Vision-based Agile Flight via Differentiable Physics

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T14:25:18.240219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:25:18.240219Z digest=sha256:fa2c58efb8c730a4dd3a722e0caf9b5a9ec431f124abf40f2cdf4f1e5dbb3408

Observation 127e30f8-775d-4bfa-8afc-9edcfbf2e2a9 · outbound

This paper cites Also using implicit differentiation, Qiao et al.

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation Also using implicit differentiation, Qiao et al

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:25:19.213655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:25:18.257733Z digest=sha256:b2ece8c097d428305a6309a83dbd7b29b24e9a2e5fdcf7db5f2c9484b654c4bc

Observation 153a5a6f-abe2-44bc-9edc-24b09688c396 · outbound

This paper cites an unresolved cited work.

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-11T14:25:19.191840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:25:18.264993Z digest=sha256:c0ef722c900089559c1b380c9b234ef6222e84d9bcca2504c487417853646a5d

Observation a74e3df9-0876-4ec1-8a63-7d9fc1156d2a · outbound

This paper cites π[·] J(π) zt+H zt+H zt+1 zt+1 μt μt+H μt+1 σt+H σt+1 (V) (π) (π) (π) zt st at Actor MLP Actor Encoder Diff.

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation π[·] J(π) zt+H zt+H zt+1 zt+1 μt μt+H μt+1 σt+H σt+1 (V) (π) (π) (π) zt st at Actor MLP Actor Encoder Diff

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:25:19.167068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:25:18.275747Z digest=sha256:118548a5f79db8f5823973041c1e1c753ce96495c7ffe1b81bd1f8e6b380a98a

Observation af767c56-a068-4ac3-92c9-bd3bcf68ef78 · outbound

This paper cites We leave combining differentiable rendering (of RGB or depth image observations) with differentiable simulation, like in (Murthy et al., 2021), to future work.

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation We leave combining differentiable rendering (of RGB or depth image observations) with differentiable simulation, like in (Murthy et al., 2021), to future work

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:25:19.149850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:25:18.284274Z digest=sha256:564db1f32ff19821dab52ad935d4deb560849f66504bf4253184dbe795110d3c

Observation 6e55127b-4105-4733-9e4b-3ef72c23c688 · outbound

This paper cites In comparison, reverse-mode AD has time complexity O(M · L), and also benefits from GPU-acceleration.

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation In comparison, reverse-mode AD has time complexity O(M · L), and also benefits from GPU-acceleration

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:25:19.059741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:25:18.317911Z digest=sha256:92ce406aeef9c9ed58b8e05b1c23a51e91a268f3d8d6eadf1107e5e2e32ea5a6

Observation 5095f754-cc73-47a0-8c24-8b94b695fbf3 · outbound

This paper cites an unresolved cited work.

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-11T14:25:19.010612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:25:18.330699Z digest=sha256:f0b249879578e36338cf175f564ef689ebb79d6cbbbaeb938d0ae8309f508b87

Observation fe1795f4-95c2-4318-a270-af7b127cd438 · outbound

This paper cites Based on AllegroHand from (Makoviychuk et al., 2021).

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation Based on AllegroHand from (Makoviychuk et al., 2021)

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:25:18.966882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:25:18.352996Z digest=sha256:cefc5bba858ffc534402954110dbd533f4c64b136610374849afcd1e3838a6ce

Observation 411fd8d7-02b4-4982-8cc0-faa5d06b39b4 · outbound

This paper cites an unresolved cited work.

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-11T14:25:18.860908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:25:18.377948Z digest=sha256:9987a9a8c6f3c5993aa285a7ce19360e69ce943c1d69a8a236f303b55184b40d

Observation 2c3edd44-66ba-4fda-bdf0-f199542e77b4 · outbound

This paper cites Based on RollingPin from (Huang et al., 2020).

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation Based on RollingPin from (Huang et al., 2020)

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:25:18.926240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:25:18.360530Z digest=sha256:117596af33b7f7448f4d99d61375f0c9ee84dbc7d0d6c4c7e0fb306ec0dbd9e4

Observation e5f985f1-0f44-4c6c-866c-77c900e717fe · outbound

This paper cites All algorithms use the same actor and critic network architecture.

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation All algorithms use the same actor and critic network architecture

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:25:19.084085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:25:18.306073Z digest=sha256:1894f48a235a8a9e079edcb84b0c50952ff8a00d2375222512d301f3f5f9f174

Observation 5a257b74-960a-4cf5-9f25-a5aa275ccd7b · outbound

This paper cites Based on Flip from (Li et al., 2023a).

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation Based on Flip from (Li et al., 2023a)

Reference 130

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:25:18.877566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:25:18.372313Z digest=sha256:d9aaa0338a0ad6c4c0c17e9e4514e1b49623e062f87ccde1457ed5a9643b5bcd

Observation 786e052e-d36a-4945-b434-f1a7f32f84bf · outbound

This paper cites All algorithms use the same actor and critic network architecture.

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation All algorithms use the same actor and critic network architecture

Reference 256

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:25:19.099814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:25:18.298268Z digest=sha256:9e33b0bb4a5894cff317f4476c51d3a9d7ab1d5ef4274707542079c909072956

Observation b177be36-fbb5-41b3-809f-7e8229d41c3f · outbound

This paper cites Inspired by (Murthy et al., 2021; Hu et al., 2020).

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation Inspired by (Murthy et al., 2021; Hu et al., 2020)

Reference 1000

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:25:18.904096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:25:18.366344Z digest=sha256:6788ab0db26d1d3c6cb68e7a0ba13fd43006e2334f29304df565193910763498

Observation 00de969e-614a-418a-856a-243a103cb83b · outbound

This paper cites We review prior works on non-parallel differentiable simulators for robotics and control problems.

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation We review prior works on non-parallel differentiable simulators for robotics and control problems

Reference 2008

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:25:19.236317Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:25:18.249083Z digest=sha256:ae507308b6dae6d76e1b2b173ba33c902d4cc28e96d5a5157de7e728f1c29823

Observation e020e6fe-8290-414f-9db3-e592d4e3d8a6 · outbound

This paper cites Differentiable Implicit Soft-Body Physics.

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation Differentiable Implicit Soft-Body Physics

Reference 2014

Resolution
unresolved
no resolver link, observed 2026-08-11T14:25:18.191878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:25:18.191878Z digest=sha256:f06373bb53340a253ef16b9e1f993281e226339ec3d6554447efb52742173c0e

Observation 1f87b3bd-a73b-4c89-a02d-56d7803de230 · outbound

This paper cites Soft Actor-Critic Algorithms and Applications.

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation Soft Actor-Critic Algorithms and Applications

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-11T14:25:18.120507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:25:18.120507Z digest=sha256:06967fe572721d0ab80a401ab5204dd0025323f699b4ddf605768e6a55c36fca

Observation ffd102da-3bef-4727-be3f-12eb11cdbbf4 · outbound

This paper cites Impala: Scalable distributed deep-rl with importance weighted actor-learner architectures.

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation Impala: Scalable distributed deep-rl with importance weighted actor-learner architectures

Reference 2018

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:25:19.382852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:25:18.098857Z digest=sha256:a08cb29af1f946e7bc0104022607127113879bebd9facd40dbbb983b721a0b08

Observation 6428e1f0-4e4a-4f3f-a0c9-867b07190c09 · outbound

This paper cites Solving Rubik's Cube with a Robot Hand.

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation Solving Rubik's Cube with a Robot Hand

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-11T14:25:18.087081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:25:18.087081Z digest=sha256:104da2fa50bc4cb2bfa50e8737a6a144f177f5db1ca8fc3fccb5a5e8c8ade793

Observation 9e5d1978-2588-45de-a979-3fce921f4e09 · outbound

This paper cites Massively Parallel Methods for Deep Reinforcement Learning.

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation Massively Parallel Methods for Deep Reinforcement Learning

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-11T14:25:18.169445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:25:18.169445Z digest=sha256:ef0b16834d3f20a1810ab0b5e68cb6b560e7033b654b509248c8e58e06c3f45b

Observation f675e79f-52ef-4447-adeb-b288ebc72570 · outbound

This paper cites Model-Based Value Estimation for Efficient Model-Free Reinforcement Learning.

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation Model-Based Value Estimation for Efficient Model-Free Reinforcement Learning

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-11T14:25:18.104126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:25:18.104126Z digest=sha256:f1095116aa38fd6cd1aa4f6ac727a033b118d81148b32b4f46c8c7a50bfaaffe

Observation 41930944-0b66-4dbb-820f-caaede70a043 · outbound

This paper cites Layer Normalization.

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation Layer Normalization

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-11T14:25:18.093289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:25:18.093289Z digest=sha256:176e47cad24aceb4ccfd9774315fd0276a4161c8499f3c018a94976aa7f11c07

Observation 4baeafc8-f967-42ca-95cb-4912ec108a71 · outbound

This paper cites Decoupled Weight Decay Regularization.

Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation Decoupled Weight Decay Regularization

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-11T14:25:18.162971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:25:18.162971Z digest=sha256:2d0d735112e4dc4b3f898bd69afad4e9ca4a349c7082bc4f971dcb34a82f0e5c

Pith citing papers

Observation e35af45e-1f0b-448d-b13c-8b2845a3e1bb · inbound

Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks cites this paper.

Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T11:03:31.098377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:03:31.098377Z digest=sha256:aec05726d4ebffab051b0c1236d1ccba12dacd6af8e4ad89319ad102809cf7b5

Observation 0c47a33e-6ec9-4403-8350-7707d2faef05 · inbound

First Order Model-Based RL through Decoupled Backpropagation cites this paper.

First Order Model-Based RL through Decoupled Backpropagation Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T13:55:19.235569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:55:19.235569Z digest=sha256:11cc13ebb9da606e10e9d745a564b82158dc48a0487a667621dd3b183c161fce

Observation d80812c7-5ba6-474f-8198-3150962b2a74 · inbound

The HydroGym Reinforcement Learning Platform for Fluid Dynamics cites this paper.

The HydroGym Reinforcement Learning Platform for Fluid Dynamics Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-03T15:17:02.750342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:17:02.750342Z digest=sha256:b4ecd557d71c121e3cbe14b09ee06a0971909abbdd4a666b9ef8e8881c45c9e9

Observation 838e404a-bdbb-41a5-8f71-374de83cc229 · inbound

Does "Do Differentiable Simulators Give Better Policy Gradients?'' Give Better Policy Gradients? cites this paper.

Does "Do Differentiable Simulators Give Better Policy Gradients?'' Give Better Policy Gradients? Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-10T05:25:54.541004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T05:24:26.743531Z digest=sha256:477d71a15d687e0f035b4dd4b8cc758e338ed9e8b7157a8ef3cd157c7535dec1