Pith. sign in

Paper Citation Record · LEDGER

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization

As of 17 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 0 inbound Pith citation observations for arXiv:2507.10914.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.10914 v1

Coverage vector

measured 45 of 45 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:28:49.005957Z

measured 45 of 45 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

45 of 45 outbound references displayed

  • verified exact0
  • verified fuzzy41
  • unresolved4
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7bbbb9d1-ebfe-40e5-bfc5-0ad85bda4532 · outbound

This paper cites Optimal Algorithms for Online Convex Optimization with Multi- Point Bandit Feedback.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Optimal Algorithms for Online Convex Optimization with Multi- Point Bandit Feedback

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:55.939978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:44.566036Z digest=sha256:5cacf9887f83986562ef8963d91c74b9aed7aada33971cc55ccafe65cb9d63d6

Observation 38e30b3e-872a-4594-a97b-2d85dd73e694 · outbound

This paper cites Kakade, and Karan Singh.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Kakade, and Karan Singh

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:55.783691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:44.642567Z digest=sha256:1fc4fb29f8ab24df39b9290915b95f4a32aff406a8605a0945405507c07a9275

Observation c22ab43e-6208-49f9-b1dc-b1931389dcca · outbound

This paper cites Zico Kolter.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Zico Kolter

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:55.657831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:44.751125Z digest=sha256:cd5888374fdf399da12b20d3b273057c9a06939f9321aedd36e3c85f0e98838f

Observation aff5f864-1eb9-43af-b594-bcd1c0bd90fd · outbound

This paper cites On the model-based stochastic value gradient for continuous reinforcement learning.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization On the model-based stochastic value gradient for continuous reinforcement learning

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:55.545675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:44.810525Z digest=sha256:ee19f529ea53230783df95cf4095dd92429c8e143269ecc96158f4246cfa2b44

Observation 53d9cd46-22d8-41df-b8f0-1fb39d5ae242 · outbound

This paper cites Infinite-horizon policy-gradient estimation.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Infinite-horizon policy-gradient estimation

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:55.444040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:45.028275Z digest=sha256:c0d450b55fb4de48d35470b9900b90f0fc872b936df0d7d9e21d63b04637a5b9

Observation e4b7239a-3377-4d62-9c1c-a8916a0a3f8e · outbound

This paper cites A survey of iterative learning control.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization A survey of iterative learning control

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:55.300171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:45.113141Z digest=sha256:bb27c2efece881064d2fa766c3ca81dad4a096dbe16866cf8c69e2ba82b9cb2e

Observation b577cd09-eb72-4417-b0a6-512246ff2824 · outbound

This paper cites Difftune: Autotuning through autodifferentiation.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Difftune: Autotuning through autodifferentiation

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:55.184190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:45.204092Z digest=sha256:e0b3d8645286f88d1e802f6daf36252a0ba7bd572ca1854acad0352bd7b1b7df

Observation ae67067d-d100-466b-a27c-4fcbfdf4cd72 · outbound

This paper cites Differentiable simulation.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Differentiable simulation

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:55.049415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:45.322459Z digest=sha256:7dcdd2851ea9a929b0ad8ac60558bf9d03e9ec0fa5b0072f569c920ec33b62ef

Observation d8553229-0c08-4253-bc71-88d7b91af176 · outbound

This paper cites Zico Kolter.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Zico Kolter

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:54.944338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:45.418546Z digest=sha256:9a2d02177284e745398076a940be4dead5e2ab0caf46f8ff2c3c2ba9c0d57221

Observation 5c5787cc-f19a-4698-888a-776b4bb5798f · outbound

This paper cites Adaptive Regret for Control of Time-Varying Dynamics.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Adaptive Regret for Control of Time-Varying Dynamics

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:54.791007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:45.461848Z digest=sha256:7a00bf3c7fb43a66c48459944101c5d1dccfd872b0b8737d85f7b295fb570ec1

Observation 5834f0b8-a3a7-4312-8e75-2200fea83df7 · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T17:28:45.574042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:28:45.574042Z digest=sha256:4a44943d19868bbb56fc678f757bebc6e57f30eb640e9b7e4732502bcd1f9673

Observation 73c642e0-97fc-41aa-a307-2740153be11f · outbound

This paper cites Introduction to Online Convex Optimization.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Introduction to Online Convex Optimization

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:54.677312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:45.652824Z digest=sha256:fd761261267b728f66282118bd81743730ccab53f03079ec2843fa3cdf73aabe

Observation 77b70f49-23f9-48ea-94cb-9f2a4e0ca9ac · outbound

This paper cites Introduction to Online Control.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Introduction to Online Control

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T17:28:45.738890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:28:45.738890Z digest=sha256:4c75468f3ded162b80baace10841f88508bc28c634dd1ebebdb11dd3219b4a18

Observation 5cfb56cd-6617-4103-9c81-0fa50984dfa8 · outbound

This paper cites The Nonstochastic Control Problem.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization The Nonstochastic Control Problem

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:54.567972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:45.804696Z digest=sha256:011fb8f1d12a472c2bac5c4b164e901d6543139f79064c25f2e548519990c7bb

Observation d1cc8004-854f-416c-a8b6-72b72d6cb662 · outbound

This paper cites Ioannou and Jing Sun.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Ioannou and Jing Sun

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:54.451430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:45.934723Z digest=sha256:948102b3e3e80cd56de3fa8c6fec405dd3d7a9bb5edb3dcba06fff2d651e8502

Observation 08d63954-cf62-40d5-a8a4-fb0254a7c730 · outbound

This paper cites Scalable deep reinforcement learning for vision- based robotic manipulation.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Scalable deep reinforcement learning for vision- based robotic manipulation

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:54.314728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:46.031777Z digest=sha256:3bf486e60ae52bb952fe9f67b17616cdfbb677134f28d6ca39a542bd0a168d79

Observation c583a3a4-abc9-469c-a01e-3e08fe771a03 · outbound

This paper cites Kokotovic, and Ioannis Kanel- lakopoulos.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Kokotovic, and Ioannis Kanel- lakopoulos

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:54.190188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:46.081046Z digest=sha256:e60b24e15c7f70a7dab17c6fd6ca673842c8fb7ef6959dda44f3428f4a516c6e

Observation 37184144-6105-4486-9dfe-1e855195dbca · outbound

This paper cites Harris McClamroch.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Harris McClamroch

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:54.077934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:46.205280Z digest=sha256:cc93c00ec932f65fe6a22e77dd54227cb1c4e1b392c474d6f334648eb11e4348

Observation 0f4c16c9-2acb-4565-8d05-48fd25e588a0 · outbound

This paper cites Preiss, Emile Anand, Yingying Li, Yisong Yue, and Adam Wierman.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Preiss, Emile Anand, Yingying Li, Yisong Yue, and Adam Wierman

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:53.953338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:46.388185Z digest=sha256:aa016301df65055f7e4a1a6c78aeb9a213b0d0dce3ca1c79771dafefed6c62b9

Observation ea4a7654-2e64-4721-a344-5b33cddb8596 · outbound

This paper cites Preiss, Fengze Xie, Emile Anand, Soon-Jo Chung, Yisong Yue, and Adam Wierman.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Preiss, Fengze Xie, Emile Anand, Soon-Jo Chung, Yisong Yue, and Adam Wierman

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:53.832829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:46.489774Z digest=sha256:0b095f67d92014dc71039cc2389309c7e778bef0f348458120396150da486486

Observation 9d695f52-5c9f-4a72-b53e-6fd09c287e7b · outbound

This paper cites Universal adaptive control of nonlinear systems.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Universal adaptive control of nonlinear systems

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:53.702159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:46.584372Z digest=sha256:427c1c6a572eba54a75b6077f7914b09954b5a33ed98cc94847d4e1414d8ff51

Observation e5904d1f-e93d-4e61-b8bd-9c8e72e77f23 · outbound

This paper cites Preiss, Jedidiah Alindogan, Matthew Anderson, and Soon-Jo Chung.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Preiss, Jedidiah Alindogan, Matthew Anderson, and Soon-Jo Chung

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:53.596700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:46.693680Z digest=sha256:823a9c98adeac1cb3c727f747d6c8cdae0019a6e32578f14fa8ee4c629e359c2

Observation 2cc030d0-5a9a-464e-b021-2255c0064be2 · outbound

This paper cites Simple random search of static linear policies is competitive for reinforcement learning.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Simple random search of static linear policies is competitive for reinforcement learning

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:53.433582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:46.829555Z digest=sha256:2ac3720dfe76defa32415664407eb8b8af6e4a99386d967e24295b6e869e235a

Observation 2e280188-fd49-4dd3-81d9-0d0ee802a528 · outbound

This paper cites SymForce: Symbolic Com- putation and Code Generation for Robotics.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization SymForce: Symbolic Com- putation and Code Generation for Robotics

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:53.289204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:46.960052Z digest=sha256:6a8795ff614ad5a5e80d83b9816ec00490c26ee331f9e567891d505f02827a0e

Observation 3eacd447-915c-4609-9a7b-ac126b825d24 · outbound

This paper cites Minimum snap trajectory generation and control for quadrotors.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Minimum snap trajectory generation and control for quadrotors

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:53.108160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:47.053906Z digest=sha256:8aa4b93f1606f585e113a7c65a36fbc332b8943ab4b6d087c2b9d4a91c37b7bd

Observation 1dde1e08-914a-4afc-8bb2-9a5ee6373e94 · outbound

This paper cites Nonlinear and adap- tive intelligent control techniques for quadrotor UA V–a survey.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Nonlinear and adap- tive intelligent control techniques for quadrotor UA V–a survey

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:53.010365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:47.170933Z digest=sha256:afbaf432a4056b98a26dc0f62cfc0f80b77dc5577fe75a1337dd9eb03d86b606

Observation c6c51d56-7b6d-4a9d-89f4-5b90135c4d05 · outbound

This paper cites Pods: Policy op- timization via differentiable simulation.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Pods: Policy op- timization via differentiable simulation

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:52.833150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:47.277587Z digest=sha256:be93135ae17c8e3abd3c6115ab96427d19dbbdeba93a49b277ea0099cf931fd3

Observation ee19ca7c-0564-4b71-aa93-20579afcc2a4 · outbound

This paper cites Neural-fly enables rapid learning for agile flight in strong winds.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Neural-fly enables rapid learning for agile flight in strong winds

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:52.646712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:47.374923Z digest=sha256:081f9af0e774f6947655e98b55d7e4a3ffe3bd67050b416d6c70f38bd5db81c4

Observation 5acaeffb-f881-4f28-8c20-2c6f66a1c057 · outbound

This paper cites Policy gradient for continuing tasks in dis- counted Markov decision processes.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Policy gradient for continuing tasks in dis- counted Markov decision processes

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:52.452974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:47.469258Z digest=sha256:ade75a7c36fbfe3dbbebe2362302b0d1b2f3e4c4aaeb68e16b419293d8ae8da0

Observation 58efebf5-1364-46d0-a93a-113be0316e0f · outbound

This paper cites Preiss, Wolfgang H ¨onig, Gaurav S.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Preiss, Wolfgang H ¨onig, Gaurav S

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:52.270065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:47.569127Z digest=sha256:3aa7e80c0151faa8f610bffe10c0fba0ddd4169ea563f0222ff3893084161ee1

Observation 6caba319-dd72-4866-b2fe-179898315a63 · outbound

This paper cites SPNets: Differentiable Fluid Dynamics for Deep Neural Networks.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization SPNets: Differentiable Fluid Dynamics for Deep Neural Networks

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:51.939053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:47.694501Z digest=sha256:0e6aab749a70722ca1f2e595da5a8cabeb8dba10ec217429ef0e18828a385f96

Observation 16190867-7e1b-4c16-a4c8-a8e22c11461a · outbound

This paper cites Proximal Policy Optimization Algorithms.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Proximal Policy Optimization Algorithms

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T17:28:47.770195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:28:47.770195Z digest=sha256:d1e81decfe3061ace096ac1e40896d4d7de5633ff48f5481adc3d8d606cb5f65

Observation 98fceee1-c4fb-4cf0-add8-3175e8341357 · outbound

This paper cites Parameter-exploring policy gradients.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Parameter-exploring policy gradients

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:51.713345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:47.896497Z digest=sha256:91a73296bcdd0dc45ed7af53de1e6258f935b00dc69d86531c4e01880f8cfb9d

Observation a577770e-1d7a-4b6c-9662-c6c377d82bb2 · outbound

This paper cites Deterministic policy gradient algorithms.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Deterministic policy gradient algorithms

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:51.535018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:48.002435Z digest=sha256:1810305ff1d3fe8c0e6fe4dbb00b53145f851c4b750ff289f5359740e2c63222

Observation 84b807a0-1f22-4484-a30b-1502b9061660 · outbound

This paper cites Im- proper Learning for Non-Stochastic Control.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Im- proper Learning for Non-Stochastic Control

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:51.399197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:48.086917Z digest=sha256:6a887cee0e9ab6ad2a516c78666c54385ff25bb7bf286492fd011800c6227416

Observation 95620090-b966-4d60-b277-3ef2d0878c0c · outbound

This paper cites Slotine and W.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Slotine and W

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:51.256899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:48.154256Z digest=sha256:77645ddccbaaec7fe365427290429277f85e62ee3b3d7cea22c8b18bd8f8ce19

Observation 995c03e5-5e44-42f4-bd2f-de62dbb8c435 · outbound

This paper cites Do differentiable simulators give better policy gradients? In International Conference on Ma- chine Learning (ICML) , 2022.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Do differentiable simulators give better policy gradients? In International Conference on Ma- chine Learning (ICML) , 2022

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:51.073789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:48.280899Z digest=sha256:b201fbdbaf3a1d2df0cd3d5e707da72c97865129bc4a054a737453bc9907a5bf

Observation 8740c015-9516-4fb9-9d42-551529da676b · outbound

This paper cites an unresolved cited work.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:28:50.824812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:48.364930Z digest=sha256:5cd93be8bda531bc5b9892d78c345842419b18f11ee9e38b8c69fa003e856c09

Observation 95456459-8247-4628-a495-1f15e4ab6dc6 · outbound

This paper cites Williams.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Williams

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:50.611937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:48.485867Z digest=sha256:6e607213ec8e91145caa9f168a2e5a4e48fee79a97e92cf48e59bb298a56fa94

Observation fadda9a0-c1d0-46fb-986e-77ffa63e60a1 · outbound

This paper cites JAX- FEM: A differentiable GPU-accelerated 3D finite ele- ment solver for automatic inverse design and mechanistic data science.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization JAX- FEM: A differentiable GPU-accelerated 3D finite ele- ment solver for automatic inverse design and mechanistic data science

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:50.443986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:48.576728Z digest=sha256:f4e9a9122cf2ecd37fb8dc274026c206b1dbed653229c3732b115d76e54417c3

Observation 698615a6-472b-4027-afc8-8806fafb90ce · outbound

This paper cites Zavlanos.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Zavlanos

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:50.162035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:48.653726Z digest=sha256:b3b58a964053a87417866bdfaea495171f575f80d4c8cbd1ac04107ad388fe40

Observation 62a69f1d-b27c-427b-94b6-b8ed1ee28667 · outbound

This paper cites Zavlanos.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Zavlanos

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:49.969562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:48.767239Z digest=sha256:edc73d1f40b638dc9dc2e60be491b663b7195cfe711190e5f7afdf522de83142

Observation 2016b084-451e-4e5b-8ca3-cb7fff685b7b · outbound

This paper cites The system state comprises the car’s position p ∈ R2, body-frame velocity v ∈ R2, heading angle r ∈ so(2), angular velocity ω ∈ so(2), and steering angle ψ ∈ R.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization The system state comprises the car’s position p ∈ R2, body-frame velocity v ∈ R2, heading angle r ∈ so(2), angular velocity ω ∈ so(2), and steering angle ψ ∈ R

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:49.690094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:48.862950Z digest=sha256:a54cb5fef77f6108d4fc23e6736426e5b7927478907c78f7527abed3c6a849d9

Observation 49bd6700-1043-4dd8-a784-46dbf593ee1e · outbound

This paper cites Note that [vd t ]y = 0 for all desired trajectories.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Note that [vd t ]y = 0 for all desired trajectories

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:49.531718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:48.927624Z digest=sha256:7c27e59a84a3e498dea1952c330c474f8da61888622a11b70293653a1c9015c8

Observation 19a53eca-05fb-47c8-960b-c5bf7cd92319 · outbound

This paper cites The regularization weights were chosen empirically to be as small as possible while suppressing oscillations.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization The regularization weights were chosen empirically to be as small as possible while suppressing oscillations

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:49.257124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T17:28:49.005957Z digest=sha256:a6b6d0d700698a97aa957d2185249c72befab3ae2dbf23748e3a0020648f6154

Pith citing papers

No inbound Pith citation observations are available.