Pith. sign in

Paper Citation Record · LEDGER

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization

As of 19 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 0 inbound Pith citation observations for arXiv:2507.10914.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.10914 v1

Coverage vector

measured 45 of 45 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:28:49.005957Z

measured 45 of 45 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

45 of 45 outbound references displayed

  • verified exact0
  • verified fuzzy41
  • unresolved4
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7bbbb9d1-ebfe-40e5-bfc5-0ad85bda4532 · outbound

This paper cites Optimal Algorithms for Online Convex Optimization with Multi- Point Bandit Feedback.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Optimal Algorithms for Online Convex Optimization with Multi- Point Bandit Feedback

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:55.939978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:44.566036Z digest=sha256:3c2f39dc1211f51dfa87c82b8df8b4b55da2a59bd07b8bd9a98609423950d7e6

Observation 38e30b3e-872a-4594-a97b-2d85dd73e694 · outbound

This paper cites Kakade, and Karan Singh.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Kakade, and Karan Singh

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:55.783691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:44.642567Z digest=sha256:b4862c9b2b96d6a6077a39be21b83a6168bfc0ce832a933e4ae2e478529bb56e

Observation c22ab43e-6208-49f9-b1dc-b1931389dcca · outbound

This paper cites Zico Kolter.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Zico Kolter

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:55.657831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:44.751125Z digest=sha256:b30b93e04ebc7dc0dfdd93cdba299f03b9f13a9df8b709e7a101cf51b9dac9e1

Observation aff5f864-1eb9-43af-b594-bcd1c0bd90fd · outbound

This paper cites On the model-based stochastic value gradient for continuous reinforcement learning.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization On the model-based stochastic value gradient for continuous reinforcement learning

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:55.545675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:44.810525Z digest=sha256:b4e36c17ff7b464dfb606529d2cda5a7206df0d2ba0795dde4e2cdb1f972293f

Observation 53d9cd46-22d8-41df-b8f0-1fb39d5ae242 · outbound

This paper cites Infinite-horizon policy-gradient estimation.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Infinite-horizon policy-gradient estimation

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:55.444040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:45.028275Z digest=sha256:518eb3c3f5d4907e4f1d80e0fe60a6b94544cc7a7eba7565ae957bb2e878ca32

Observation e4b7239a-3377-4d62-9c1c-a8916a0a3f8e · outbound

This paper cites A survey of iterative learning control.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization A survey of iterative learning control

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:55.300171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:45.113141Z digest=sha256:7713c18cc5a8b5061aaa39ccae717456e56f63aa163a21620a8995f91d24c2c1

Observation b577cd09-eb72-4417-b0a6-512246ff2824 · outbound

This paper cites Difftune: Autotuning through autodifferentiation.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Difftune: Autotuning through autodifferentiation

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:55.184190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:45.204092Z digest=sha256:0583c3f738acd3e955e375642eeb538443fc664136eac968852e11eb784fd3ad

Observation ae67067d-d100-466b-a27c-4fcbfdf4cd72 · outbound

This paper cites Differentiable simulation.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Differentiable simulation

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:55.049415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:45.322459Z digest=sha256:44969d1ce41c5e871cf1f5f56ab7003d94932cfb60e1eb48252efdf2ac3ba388

Observation d8553229-0c08-4253-bc71-88d7b91af176 · outbound

This paper cites Zico Kolter.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Zico Kolter

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:54.944338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:45.418546Z digest=sha256:5f19af0984c28394f08592f0e57b2c7eec4c374653539fff219120873167c64e

Observation 5c5787cc-f19a-4698-888a-776b4bb5798f · outbound

This paper cites Adaptive Regret for Control of Time-Varying Dynamics.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Adaptive Regret for Control of Time-Varying Dynamics

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:54.791007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:45.461848Z digest=sha256:28994c91ba06ac1aa3f6d43336306a694739c46235510b945142093caf58be8d

Observation 5834f0b8-a3a7-4312-8e75-2200fea83df7 · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T17:28:45.574042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:28:45.574042Z digest=sha256:802035c0ee4bc1060feb6969514c89136f87ef628c5737d07b2d19068d295d74

Observation 73c642e0-97fc-41aa-a307-2740153be11f · outbound

This paper cites Introduction to Online Convex Optimization.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Introduction to Online Convex Optimization

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:54.677312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:45.652824Z digest=sha256:a22156e00fd95c60dc5c526f5684703fb5d7a4335c25985bb82fcbf4697dd85e

Observation 77b70f49-23f9-48ea-94cb-9f2a4e0ca9ac · outbound

This paper cites Introduction to Online Control.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Introduction to Online Control

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T17:28:45.738890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:28:45.738890Z digest=sha256:0e1c328ae8cc81d25d16fed933eac0f8b6c7582e131d88addfd04196493a08ac

Observation 5cfb56cd-6617-4103-9c81-0fa50984dfa8 · outbound

This paper cites The Nonstochastic Control Problem.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization The Nonstochastic Control Problem

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:54.567972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:45.804696Z digest=sha256:f43e4db0338a2e50d702cf1a8bf2d96525daafe48e79606210cbae54c5e4f575

Observation d1cc8004-854f-416c-a8b6-72b72d6cb662 · outbound

This paper cites Ioannou and Jing Sun.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Ioannou and Jing Sun

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:54.451430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:45.934723Z digest=sha256:18471eb51e9e8253155b29aaea73c3069491e9d7bfbc0c095ddc80b9133418bc

Observation 08d63954-cf62-40d5-a8a4-fb0254a7c730 · outbound

This paper cites Scalable deep reinforcement learning for vision- based robotic manipulation.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Scalable deep reinforcement learning for vision- based robotic manipulation

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:54.314728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:46.031777Z digest=sha256:9b25c47f8347900379d2fc635e57d128656d9c7da512aa01f2ba0f2043d270f3

Observation c583a3a4-abc9-469c-a01e-3e08fe771a03 · outbound

This paper cites Kokotovic, and Ioannis Kanel- lakopoulos.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Kokotovic, and Ioannis Kanel- lakopoulos

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:54.190188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:46.081046Z digest=sha256:876cc90a76a013d2c4433d3cfcce987a761a407a7191807b387f78d08d4d7bd7

Observation 37184144-6105-4486-9dfe-1e855195dbca · outbound

This paper cites Harris McClamroch.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Harris McClamroch

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:54.077934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:46.205280Z digest=sha256:e9a9749facef9f9e11a47f0f48e58bf663ab0f0ccc24b0e1193373608d2f05f1

Observation 0f4c16c9-2acb-4565-8d05-48fd25e588a0 · outbound

This paper cites Preiss, Emile Anand, Yingying Li, Yisong Yue, and Adam Wierman.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Preiss, Emile Anand, Yingying Li, Yisong Yue, and Adam Wierman

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:53.953338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:46.388185Z digest=sha256:a20e8a23bbebffa37e58f48bd5da74813e22c750698aa1dc56ac70f54d1d03fe

Observation ea4a7654-2e64-4721-a344-5b33cddb8596 · outbound

This paper cites Preiss, Fengze Xie, Emile Anand, Soon-Jo Chung, Yisong Yue, and Adam Wierman.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Preiss, Fengze Xie, Emile Anand, Soon-Jo Chung, Yisong Yue, and Adam Wierman

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:53.832829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:46.489774Z digest=sha256:1708071e99f06e34288d0fac040e7a1470a0700478ae0c3407380d837f38fb60

Observation 9d695f52-5c9f-4a72-b53e-6fd09c287e7b · outbound

This paper cites Universal adaptive control of nonlinear systems.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Universal adaptive control of nonlinear systems

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:53.702159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:46.584372Z digest=sha256:f3ca062ff018b93805c83bcd08f8c0d89d85ea796d9ce47b359e6a0317ba7103

Observation e5904d1f-e93d-4e61-b8bd-9c8e72e77f23 · outbound

This paper cites Preiss, Jedidiah Alindogan, Matthew Anderson, and Soon-Jo Chung.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Preiss, Jedidiah Alindogan, Matthew Anderson, and Soon-Jo Chung

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:53.596700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:46.693680Z digest=sha256:5490deb4b8e43798ab639fff088bd91a6fcc59d0b5cb8b7e434a987e04f226a2

Observation 2cc030d0-5a9a-464e-b021-2255c0064be2 · outbound

This paper cites Simple random search of static linear policies is competitive for reinforcement learning.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Simple random search of static linear policies is competitive for reinforcement learning

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:53.433582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:46.829555Z digest=sha256:6175e9ec2614a66d0de905e6cdf7b49f1613f1de5d2ea17d7e6e7f61684ce992

Observation 2e280188-fd49-4dd3-81d9-0d0ee802a528 · outbound

This paper cites SymForce: Symbolic Com- putation and Code Generation for Robotics.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization SymForce: Symbolic Com- putation and Code Generation for Robotics

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:53.289204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:46.960052Z digest=sha256:eeced916e85ebade03bdb4ae8a8e087ef21ad1603f44f6f00158ed5ed768b766

Observation 3eacd447-915c-4609-9a7b-ac126b825d24 · outbound

This paper cites Minimum snap trajectory generation and control for quadrotors.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Minimum snap trajectory generation and control for quadrotors

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:53.108160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:47.053906Z digest=sha256:8cf96ffd55317055555f48e6e4f3389285b75471d951900b2c6f20a5423fde37

Observation 1dde1e08-914a-4afc-8bb2-9a5ee6373e94 · outbound

This paper cites Nonlinear and adap- tive intelligent control techniques for quadrotor UA V–a survey.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Nonlinear and adap- tive intelligent control techniques for quadrotor UA V–a survey

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:53.010365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:47.170933Z digest=sha256:62b78a23a7083a77eb475e21cfd47640bb23ff9a2bcbe8ac4d38956e3d0703f5

Observation c6c51d56-7b6d-4a9d-89f4-5b90135c4d05 · outbound

This paper cites Pods: Policy op- timization via differentiable simulation.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Pods: Policy op- timization via differentiable simulation

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:52.833150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:47.277587Z digest=sha256:f0e8fe988ccebc5e4365dd26b3f097d07adeae89dda86e969dde9c752c2df86e

Observation ee19ca7c-0564-4b71-aa93-20579afcc2a4 · outbound

This paper cites Neural-fly enables rapid learning for agile flight in strong winds.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Neural-fly enables rapid learning for agile flight in strong winds

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:52.646712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:47.374923Z digest=sha256:0dcbeb3ca27d37fda947e16ceda61c9c0441c7dc8898f40fe19aa188ba86bb0e

Observation 5acaeffb-f881-4f28-8c20-2c6f66a1c057 · outbound

This paper cites Policy gradient for continuing tasks in dis- counted Markov decision processes.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Policy gradient for continuing tasks in dis- counted Markov decision processes

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:52.452974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:47.469258Z digest=sha256:13c33363899fa53f78f4fb68c4db5a0a3a94d4ec8d996fe3b7d307a220787a7a

Observation 58efebf5-1364-46d0-a93a-113be0316e0f · outbound

This paper cites Preiss, Wolfgang H ¨onig, Gaurav S.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Preiss, Wolfgang H ¨onig, Gaurav S

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:52.270065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:47.569127Z digest=sha256:72d4a6a78348fe14987cddce6c62393845803f00b9309b73f5d5a108fb930ca5

Observation 6caba319-dd72-4866-b2fe-179898315a63 · outbound

This paper cites SPNets: Differentiable Fluid Dynamics for Deep Neural Networks.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization SPNets: Differentiable Fluid Dynamics for Deep Neural Networks

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:51.939053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:47.694501Z digest=sha256:981f962e6e163321d4c426b7ecc05862edbf0898f9a5e928041aa95b55b3e41b

Observation 16190867-7e1b-4c16-a4c8-a8e22c11461a · outbound

This paper cites Proximal Policy Optimization Algorithms.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Proximal Policy Optimization Algorithms

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T17:28:47.770195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:28:47.770195Z digest=sha256:2ca96e6d47a45f185a9a257677c22b803fb57dec3a8a7aee76d58916ee211c63

Observation 98fceee1-c4fb-4cf0-add8-3175e8341357 · outbound

This paper cites Parameter-exploring policy gradients.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Parameter-exploring policy gradients

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:51.713345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:47.896497Z digest=sha256:0ebb5dab73da4763c4e70aed0158aae3c732b00e717a4dbd4e0c2ad20215a33b

Observation a577770e-1d7a-4b6c-9662-c6c377d82bb2 · outbound

This paper cites Deterministic policy gradient algorithms.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Deterministic policy gradient algorithms

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:51.535018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:48.002435Z digest=sha256:20fb4a3a367912c5baa2efe8d0c2f388f93e5f234d2bd95092037f2c5a0d45da

Observation 84b807a0-1f22-4484-a30b-1502b9061660 · outbound

This paper cites Im- proper Learning for Non-Stochastic Control.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Im- proper Learning for Non-Stochastic Control

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:51.399197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:48.086917Z digest=sha256:8edf81fdc447b4269f8673c8a8c41e7768ab9ab2398c434d9681d5a7d87e3b09

Observation 95620090-b966-4d60-b277-3ef2d0878c0c · outbound

This paper cites Slotine and W.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Slotine and W

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:51.256899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:48.154256Z digest=sha256:8280c5a86522ae0f8e1774f8b1d91ca6bd6040b0f358d7472ca67293e5adafa2

Observation 995c03e5-5e44-42f4-bd2f-de62dbb8c435 · outbound

This paper cites Do differentiable simulators give better policy gradients? In International Conference on Ma- chine Learning (ICML) , 2022.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Do differentiable simulators give better policy gradients? In International Conference on Ma- chine Learning (ICML) , 2022

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:51.073789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:48.280899Z digest=sha256:13f8c750bb4a55223aebc4771aec1c7f78c51df6d42036e1f79fb7f0d90a6ae5

Observation 8740c015-9516-4fb9-9d42-551529da676b · outbound

This paper cites an unresolved cited work.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:28:50.824812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:48.364930Z digest=sha256:747134b9c36d1d929eefef48f11ed2f84e64308c9734116c1f7cab3419f41b9e

Observation 95456459-8247-4628-a495-1f15e4ab6dc6 · outbound

This paper cites Williams.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Williams

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:50.611937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:48.485867Z digest=sha256:e7a30a95ef83bd3c45c4a2b4f87e354c64d353fd6e724adc5901eb93353d16ca

Observation fadda9a0-c1d0-46fb-986e-77ffa63e60a1 · outbound

This paper cites JAX- FEM: A differentiable GPU-accelerated 3D finite ele- ment solver for automatic inverse design and mechanistic data science.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization JAX- FEM: A differentiable GPU-accelerated 3D finite ele- ment solver for automatic inverse design and mechanistic data science

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:50.443986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:48.576728Z digest=sha256:af7273aa77e5d328bf90c0012c1a823d3e25458ccf43a01467ea628a41d347cc

Observation 698615a6-472b-4027-afc8-8806fafb90ce · outbound

This paper cites Zavlanos.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Zavlanos

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:50.162035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:48.653726Z digest=sha256:45a8c41953c66c22e2d81f58cc53e2b99a48cc2a3e7f00a661381af5c4c86222

Observation 62a69f1d-b27c-427b-94b6-b8ed1ee28667 · outbound

This paper cites Zavlanos.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Zavlanos

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:49.969562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:48.767239Z digest=sha256:d4548b370bce2899f7c03919057d71002dce2b36071d7d2106fb702f3f5dd11c

Observation 2016b084-451e-4e5b-8ca3-cb7fff685b7b · outbound

This paper cites The system state comprises the car’s position p ∈ R2, body-frame velocity v ∈ R2, heading angle r ∈ so(2), angular velocity ω ∈ so(2), and steering angle ψ ∈ R.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization The system state comprises the car’s position p ∈ R2, body-frame velocity v ∈ R2, heading angle r ∈ so(2), angular velocity ω ∈ so(2), and steering angle ψ ∈ R

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:49.690094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:48.862950Z digest=sha256:2f80c54e89762465f0dead9b2c9931e73cfc0eecc38298539b898bfec694c75c

Observation 49bd6700-1043-4dd8-a784-46dbf593ee1e · outbound

This paper cites Note that [vd t ]y = 0 for all desired trajectories.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Note that [vd t ]y = 0 for all desired trajectories

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:49.531718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:48.927624Z digest=sha256:a1933ea87ca787711879c37af4d77a2d87ed1847ebb2eede571fb45632170473

Observation 19a53eca-05fb-47c8-960b-c5bf7cd92319 · outbound

This paper cites The regularization weights were chosen empirically to be as small as possible while suppressing oscillations.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization The regularization weights were chosen empirically to be as small as possible while suppressing oscillations

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:49.257124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T17:28:49.005957Z digest=sha256:f79399caf6aef14e0f781c0162353d737d60d20fb15e506ca4fd94057f3de253

Pith citing papers

No inbound Pith citation observations are available.