Pith. sign in

Paper Citation Record · LEDGER

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures

As of 17 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 4 inbound Pith citation observations for arXiv:2504.17857.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.17857 v3

Coverage vector

measured 51 of 51 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:34:47.753469Z

measured 55 of 55 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T16:24:01.004152Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-28T20:52:37.979699Z

Reference resolution

51 of 51 outbound references displayed

  • verified exact0
  • verified fuzzy30
  • unresolved21
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e61ff5ab-66e9-4618-8597-bf3017b09f78 · outbound

This paper cites Boston Dynamics.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Boston Dynamics

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:34:48.411403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:34:47.538221Z digest=sha256:9607a0f6e0537ab52476274fb830f6be9f08d03307f3db37a9d5f5110ae312d9

Observation 7774cf65-9bda-43bf-b3b2-c004f6d83d8c · outbound

This paper cites [Online].

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures [Online]

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:34:48.396805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:34:47.543045Z digest=sha256:96ca1ed8f36f68b1b24406d379d3f9e601b90fc89804a9eace33d1dc2f7649a7

Observation ed60f90b-b5c2-4084-838e-0bc9aa3480b3 · outbound

This paper cites [Online].

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures [Online]

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:34:48.383449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:34:47.547576Z digest=sha256:4db39cc40616cb3f6358eed330b9df1c218c8d17248b1c8e734459826699c5bd

Observation 1c883275-e383-4160-b831-64994ac3c5b4 · outbound

This paper cites Boston Dynamics.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Boston Dynamics

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:34:48.370259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:34:47.552155Z digest=sha256:18aa6a54e93c058407fdcb85b7dddc1a5e73f0b55f0bcec3b563315097e05413

Observation 5d3dc23e-8b22-4b2f-aa6a-040e26202fd7 · outbound

This paper cites Boston Dynamics.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Boston Dynamics

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:34:48.356804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:34:47.556512Z digest=sha256:784593bb94558227e08dcc5eb847d82907aa7b238f7c50460ce6cd97bca430cb

Observation 4f047921-8809-4a90-adbe-1e7ff639c2b0 · outbound

This paper cites Domanico.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Domanico

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:34:48.343828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:34:47.560845Z digest=sha256:d7093d7147c54f39e9997daa9d6d0bec951dd168916b981c6372a5d628eee72c

Observation 73935ede-48d0-4ed3-9b23-039366f5e8e3 · outbound

This paper cites Mit cheetah 3: Design and control of a robust, dynamic quadruped robot,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Mit cheetah 3: Design and control of a robust, dynamic quadruped robot,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T10:34:47.565641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:34:47.565641Z digest=sha256:aaa6e5fe47e36986ef949b25786d994ec8071a454c88090949e047450d0e5706

Observation 1b6819f2-77d3-4612-91ec-bffebaf95991 · outbound

This paper cites Design of hyq–a hydraulically and electrically actuated quadruped robot,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Design of hyq–a hydraulically and electrically actuated quadruped robot,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:34:48.319891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:34:47.569930Z digest=sha256:06bb1a72a14cc7452b4cdb95ed442d691c9576b7701c05d146cc5e6b5326a8f3

Observation 2b648160-cc1b-41ae-a9c6-847b72d5b67e · outbound

This paper cites Anymal-a highly mobile and dynamic quadrupedal robot,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Anymal-a highly mobile and dynamic quadrupedal robot,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T10:34:47.574149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:34:47.574149Z digest=sha256:b5a06dbf78b96b8fa0aeb50b8a7601eccc1eabf461768da564d28743ed914ad6

Observation dc16cc9b-ca55-48c5-a45f-6bb86780305b · outbound

This paper cites Dynamic locomotion in the mit cheetah 3 through convex model-predictive control,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Dynamic locomotion in the mit cheetah 3 through convex model-predictive control,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T10:34:47.578428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:34:47.578428Z digest=sha256:8dc2bee49b604a20eb26c221e554671ae39879c0290f63678b2ce66b119d6c8f

Observation 46a0d6ca-f1ef-424f-b4c2-ab0797450502 · outbound

This paper cites Real-time motion planning of legged robots: A model predictive control approach,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Real-time motion planning of legged robots: A model predictive control approach,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T10:34:47.582616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:34:47.582616Z digest=sha256:7b71a866f8637c2c495788965de96d5ed821368961358748c60b042097d378e8

Observation b58594b9-4b87-4611-a85a-8d68afc8b126 · outbound

This paper cites Dynamic locomotion and whole-body control for quadrupedal robots,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Dynamic locomotion and whole-body control for quadrupedal robots,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T10:34:47.587107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:34:47.587107Z digest=sha256:48bdd0865056ef6efd7b26d5c38fbaf70b4bafdb008ffffd009f05ebb14d26c0

Observation 5b5cc935-673c-427b-88c1-6649c7f2e872 · outbound

This paper cites Learning agile and dynamic motor skills for legged robots,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Learning agile and dynamic motor skills for legged robots,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-16T10:34:47.591881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:34:47.591881Z digest=sha256:614afdfeef52840d3fcd025b92355e2f8fbee77db36e01c95059e8094b6390a5

Observation d20d3d53-af5c-4156-b085-43a4a629459c · outbound

This paper cites Rapid locomotion via reinforcement learning,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Rapid locomotion via reinforcement learning,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T10:34:47.596150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:34:47.596150Z digest=sha256:fa9cb465ea0a6e181a467abc67884c51436117b80da58923fa57f54e33c37953

Observation deff04f7-7ad5-44d7-b42c-ba39b2088ed7 · outbound

This paper cites Anymal parkour: Learning agile navigation for quadrupedal robots,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Anymal parkour: Learning agile navigation for quadrupedal robots,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T10:34:47.600414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:34:47.600414Z digest=sha256:f7c32f1d0ba38f66bb74de9106a610b0766cf278e8eb06aa264967b6a7c4d8cf

Observation 417664b8-92aa-4cbe-8dde-29c3f73ca7a1 · outbound

This paper cites Learning to walk in minutes using massively parallel deep reinforcement learning,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Learning to walk in minutes using massively parallel deep reinforcement learning,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T10:34:47.604641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:34:47.604641Z digest=sha256:dec464077eba6f7e9f26f703356cf7ce6c33c966cc38abc5d646b330759de0c0

Observation 5033f84a-1c45-410b-a37c-54d7e69efbef · outbound

This paper cites Concurrent training of a control policy and a state estimator for dynamic and robust legged locomotion,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Concurrent training of a control policy and a state estimator for dynamic and robust legged locomotion,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T10:34:47.608817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:34:47.608817Z digest=sha256:b7e12429b002f75a001c1e91fe51c96c67d78e341a6fa403147216fa61407b7d

Observation 48f4a069-af48-44b5-9cb2-9e58d9caf3db · outbound

This paper cites RMA: Rapid Motor Adaptation for Legged Robots.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures RMA: Rapid Motor Adaptation for Legged Robots

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T10:34:47.612791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:34:47.612791Z digest=sha256:adcf41f1e7bc2fc7dea90b06abf6bfc2fcc9f814ccec8faa7f54d1a34b9aee31

Observation 906b9c94-ca0f-4fde-bedb-4ccdfef34546 · outbound

This paper cites Actuator-Constrained Reinforcement Learning for High-Speed Quadrupedal Locomotion.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Actuator-Constrained Reinforcement Learning for High-Speed Quadrupedal Locomotion

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T10:34:47.617273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:34:47.617273Z digest=sha256:fc900bb63ef2509d66d925527b0693450fa4fe39bad6968720730e1fc4a4f461

Observation 39260b76-9595-4cac-b214-f5fdeba013b8 · outbound

This paper cites Robot learning from randomized simulations: A review,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Robot learning from randomized simulations: A review,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:34:48.226924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:34:47.622197Z digest=sha256:3fe4f9837f1436f7b0c52eeb12e8bbdffc9735d2e908ac0c3a95c7b6edf9fa51

Observation b2f1eeca-cc06-409a-9e90-d8f9432ae4ed · outbound

This paper cites Feedback control for cassie with deep reinforcement learning,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Feedback control for cassie with deep reinforcement learning,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:34:48.213573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:34:47.626732Z digest=sha256:4fa30c0c0bbaa41986269cfa0540f3f2ba0dad119dbf83397df80bf9da73e8da

Observation 946f0d13-1d8b-4e6c-8e7c-04d7f0aa05a7 · outbound

This paper cites Policy transfer via kinematic domain randomization and adaptation,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Policy transfer via kinematic domain randomization and adaptation,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:34:48.199252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:34:47.631040Z digest=sha256:d00d3ba2bbe0d7815540161ac343c15439264f9298d7b3f36267bca0c0c6d48c

Observation 9de06589-a0b0-4f0d-aa98-e66b38f52adb · outbound

This paper cites Reinforcement learning for robust parameterized locomotion control of bipedal robots,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Reinforcement learning for robust parameterized locomotion control of bipedal robots,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-16T10:34:47.635069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:34:47.635069Z digest=sha256:27d7a07c1b03a96a2d8c09dae10fa15c086652e27951b5810711ab05170b7003

Observation e417f7e4-5d7a-4a35-95b8-9331420e26f5 · outbound

This paper cites Sim2real transfer for reinforcement learning without dynamics randomization,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Sim2real transfer for reinforcement learning without dynamics randomization,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:34:48.176952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:34:47.639416Z digest=sha256:19e0b1308bf3d100ecee0e752c032f9264f4c1a48b3f6a0d1caae868fc6c3c9a

Observation 7536fa8d-0eff-4434-b1ce-27d3f1bfb68a · outbound

This paper cites Estimation of inertial parameters of rigid body links of manipulators,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Estimation of inertial parameters of rigid body links of manipulators,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:34:48.162613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:34:47.644491Z digest=sha256:0cb1075147fcf7cb1c3440a52be69bcc719df023bdd4a16438b98485170d5d15

Observation 74afa7bc-cb13-43c1-a9ab-46699c4ad9ed · outbound

This paper cites Geometric robot dynamic identification: A convex programming approach,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Geometric robot dynamic identification: A convex programming approach,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:34:48.148935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:34:47.649434Z digest=sha256:40e025e43604d0754849e86ba50dc00e03fe6abd611e5178ed1cd2e74a0e9596

Observation e986502e-b79d-4e7a-8b63-75c0cc8433e3 · outbound

This paper cites Pros and cons of gan evaluation measures,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Pros and cons of gan evaluation measures,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-16T10:34:47.653314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:34:47.653314Z digest=sha256:952a0c61b5a15f85a9dd0f960cef02fde8c5593b8beb359a24676229f1b559e6

Observation 81ced21a-467d-4979-acd4-3326aed28374 · outbound

This paper cites The wasserstein distance and approximation the- orems,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures The wasserstein distance and approximation the- orems,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:34:48.126874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:34:47.657562Z digest=sha256:8259202b2103768b35b11adc2bbd1fa0918a0d479cd3322759f526b818413292

Observation 6d0b83f9-3719-4d27-9414-8866faa2c88b · outbound

This paper cites Training generative neural networks via Maximum Mean Discrepancy optimization.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Training generative neural networks via Maximum Mean Discrepancy optimization

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-16T10:34:47.661621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:34:47.661621Z digest=sha256:6edef980949604b24d604095bc2751c49b23123d1ea78e4bf6c04a0e9de02589

Observation 6d31e23f-2c8c-4f11-837d-296c72fa2e93 · outbound

This paper cites Amp: Adversarial motion priors for stylized physics-based character con- trol,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Amp: Adversarial motion priors for stylized physics-based character con- trol,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-16T10:34:47.665981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:34:47.665981Z digest=sha256:f49902bbf4b6ea0a05e4dc806b515761a209e8f4d40144837ae9335bdcb041f3

Observation 7de43baa-1f93-4ac1-b8aa-971bd91a5440 · outbound

This paper cites Learning agile skills via adversarial imitation of rough partial demonstrations,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Learning agile skills via adversarial imitation of rough partial demonstrations,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:34:48.103810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:34:47.669939Z digest=sha256:d2d3a590327a406fb55c78cc40a202a67e08ce29f85277283083f598613ad320

Observation 9c1c3968-96e1-49ee-b1de-a143b6686173 · outbound

This paper cites Knowl- edge transfer across imaging modalities via simultaneous learning of adaptive autoencoders for high-fidelity mobile robot vision,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Knowl- edge transfer across imaging modalities via simultaneous learning of adaptive autoencoders for high-fidelity mobile robot vision,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:34:48.090179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:34:47.673984Z digest=sha256:8059e9044cfec021c61a18b99b38bc69e2716cf7f9281494625fb0882486cd10

Observation 2273d16f-747e-4771-94a7-a9911bdf7c8e · outbound

This paper cites an unresolved cited work.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:34:48.076505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:34:47.677924Z digest=sha256:f9f1df5fca80504e1f94b4ec245db33c1195a2f8ef5727bf089e6365d3ea7790

Observation 7b4ecbfe-a42b-410b-9fd9-cde075b119d1 · outbound

This paper cites About the nyquist frequency,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures About the nyquist frequency,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:34:48.062953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:34:47.682130Z digest=sha256:8ff8875e755c3d6e59721749e3aec7e12560d1a626387ef920d080542b0f7e60

Observation 429ececc-51bf-41b6-a2ce-6624231d878f · outbound

This paper cites Simulation-based design of dynamic controllers for humanoid balancing,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Simulation-based design of dynamic controllers for humanoid balancing,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:34:48.049548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:34:47.686263Z digest=sha256:6a79f2572ea16381a54c3a93e75d574aa3a12a093671006dcf01e21b56a540a8

Observation 10abcb00-3956-4279-8f7e-f767af486008 · outbound

This paper cites Sim-to-Real: Learning Agile Locomotion For Quadruped Robots.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Sim-to-Real: Learning Agile Locomotion For Quadruped Robots

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-16T10:34:47.691450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:34:47.691450Z digest=sha256:17b1b4a98c01c19b2f919bd218011a2436dfaae3d032fe97e73e4a8cdac97c80

Observation 4509a04b-f695-4743-acba-78c3c3bcf1c4 · outbound

This paper cites Sim-to-real via sim- to-sim: Data-efficient robotic grasping via randomized-to-canonical adaptation networks,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Sim-to-real via sim- to-sim: Data-efficient robotic grasping via randomized-to-canonical adaptation networks,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:34:48.035832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:34:47.695910Z digest=sha256:4bb91489509b9195313f636372c65a19dfc7a7981450107537cde597b61de63f

Observation bfb76f5e-d700-4dc6-a6ff-29b7ac6cc24d · outbound

This paper cites Learning inertial odometry for dynamic legged robot state estimation,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Learning inertial odometry for dynamic legged robot state estimation,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:34:48.022122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:34:47.699791Z digest=sha256:f812416bd42bc0676e28e9af9e0bd737f0fa4580f2ce3c4401073512894bad48

Observation 3961aa89-e1a9-4086-a0bf-379f9e5c9795 · outbound

This paper cites Starleth & co.: Design and control of legged robots with compliant actuation,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Starleth & co.: Design and control of legged robots with compliant actuation,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:34:48.008451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:34:47.703807Z digest=sha256:67ca8e700ec9254c1beaad5af211a7d92777383d94656118f700d2781b38a859

Observation 4497228d-2a20-4b47-9ed8-f3660a537a1a · outbound

This paper cites Mini cheetah: A platform for push- ing the limits of dynamic quadruped control,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Mini cheetah: A platform for push- ing the limits of dynamic quadruped control,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:34:47.994744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:34:47.708190Z digest=sha256:4e77ec0327d961e917d0ada6d6f8476c04b231d615c67c128940c7d4ae929e47

Observation 8e6a2a03-3b43-4c1b-ac64-a90bc8b10896 · outbound

This paper cites Tutorial cma-es: evolution strategies and covariance matrix adaptation,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Tutorial cma-es: evolution strategies and covariance matrix adaptation,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:34:47.981037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:34:47.712382Z digest=sha256:52d344125c314440194944851f2b623cdb172025d714b84decd0a68d792d4ef5

Observation 76afa3df-c89e-4642-a686-86da959f8d09 · outbound

This paper cites Hardware as Policy: Mechanical and Computational Co-Optimization using Deep Reinforcement Learning.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Hardware as Policy: Mechanical and Computational Co-Optimization using Deep Reinforcement Learning

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-16T10:34:47.716277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:34:47.716277Z digest=sha256:f863620465ab27abcf828c80a4a73c32a0365b858d3949eabe5faf01252c59db

Observation e607d970-232d-41ff-9a15-2775bd4149d5 · outbound

This paper cites Simulation aided co-design for robust robot optimization,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Simulation aided co-design for robust robot optimization,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:34:47.967050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:34:47.720532Z digest=sha256:edd614a48cb4a65fc448274d0af80328d71c228176dcebc314b95a217f27a8c1

Observation 6279070a-1439-48a6-8091-fb8832c85d54 · outbound

This paper cites Multiple task optimization with a mixture of controllers for motion generation,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Multiple task optimization with a mixture of controllers for motion generation,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:34:47.953762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:34:47.724837Z digest=sha256:89ba45d8078080dda6be2719aa9f5a69db419e3c79c7029cae45c7937944fef4

Observation 6e58ba1b-feab-4aec-860c-e158e77d7edf · outbound

This paper cites Kicking motion planning of nao robots based on cma-es,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Kicking motion planning of nao robots based on cma-es,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:34:47.939834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:34:47.728992Z digest=sha256:755cd4e77cf56292afebe6ff0353d34cfab8b646b6b0b095d38b412130d3482f

Observation c56abdfb-0489-4bd7-9572-4f3ed61b6120 · outbound

This paper cites Orbit: A unified simulation framework for interactive robot learning environments,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Orbit: A unified simulation framework for interactive robot learning environments,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:34:47.925595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:34:47.732837Z digest=sha256:9d298b20efef214e603c3bf4e26c665e90c8455c966f9cad6a42774f01d08ec6

Observation baa34c9d-8544-4d04-9c62-667c1b05e36f · outbound

This paper cites an unresolved cited work.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:34:47.912271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:34:47.736865Z digest=sha256:944ef2772fc79efb02be68a084ae7a30d574824e2575e549044e28ec9d62ba57

Observation 6981ee42-a225-40d3-8c70-094b3dae6763 · outbound

This paper cites Robotic Systems Lab.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Robotic Systems Lab

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:34:47.898554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:34:47.740872Z digest=sha256:595f8b3179d371911ce8ee4c3e199afa5ef8c8617052289d64000606c8d9050d

Observation 3dd07f4a-83f6-403f-8402-b50dcc8481eb · outbound

This paper cites A survey of actor-critic reinforcement learning: Standard and natural policy gradients,.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures A survey of actor-critic reinforcement learning: Standard and natural policy gradients,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:34:47.885185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:34:47.744956Z digest=sha256:2cd4d3b26792442f407511554d46a68f9ff8c74ff01c7290cc8d3c00d0ab56bb

Observation b2ba19a0-92fa-4c33-8bf2-01320d3adc1f · outbound

This paper cites Proximal Policy Optimization Algorithms.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Proximal Policy Optimization Algorithms

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-16T10:34:47.749086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:34:47.749086Z digest=sha256:569008973e450a2a81aeee7a239499d26cf3af0354a26bdec4e8cc49dcb0cd88

Observation 46ff918a-8c34-4863-bdd5-5b1a7408d52f · outbound

This paper cites Boston Dynamics.

High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures Boston Dynamics

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:34:47.870034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:34:47.753469Z digest=sha256:9b2286993f52e99c5c261ac5c7e50cc668031fffb5c61205555d2f28eef25fcb

Pith citing papers

Observation e639ffae-9b1f-45cd-9c1f-a2c2a1045d61 · inbound

Towards bridging the gap: Systematic sim-to-real transfer for diverse legged robots cites this paper.

Towards bridging the gap: Systematic sim-to-real transfer for diverse legged robots High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T16:24:01.004152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:24:01.004152Z digest=sha256:bd9ad10c73c8cbdc4980df761e45e9da85f4259bb4ac358062ad275d147c4030

Observation 6063e40d-8e20-4b72-ae92-a7416c55c41e · inbound

Isaac Lab: A GPU-Accelerated Simulation Framework for Multi-Modal Robot Learning cites this paper.

Isaac Lab: A GPU-Accelerated Simulation Framework for Multi-Modal Robot Learning High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:22:52.729902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-11T10:22:52.472354Z digest=sha256:8dc9ad2ba898dd835784eac7324e09c3a2743005493b13c9d65ffeb82bd3a966

Observation 838aa872-209c-4534-93c4-d353e38d1a9b · inbound

Differentiable Weightless Controllers: Learning Logic Circuits for Continuous Control cites this paper.

Differentiable Weightless Controllers: Learning Logic Circuits for Continuous Control High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures

Reference 1994

Resolution
unresolved
no resolver link, observed 2026-08-03T19:16:27.534748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T19:16:27.534748Z digest=sha256:6bf5bed2fc5eb3510387b1defe5fb9bdb0d8fa1b3b3ef72ae3e0d896448d403e

Observation b5ae5f6f-f69c-4512-a5bb-dc682dbf926c · inbound

Too Much of a Good Thing: When sim2real Efforts Impede Policy Learning (And What to Do About It) cites this paper.

Too Much of a Good Thing: When sim2real Efforts Impede Policy Learning (And What to Do About It) High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-06-28T20:52:37.981273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-28T18:19:58.499360Z digest=sha256:b42112c0151650e57262ce8f12d540ee4f68b31c9164148d4cf6ec543a1fcbec