Pith. sign in

Paper Citation Record · LEDGER

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches

As of 17 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 7 inbound Pith citation observations for arXiv:2505.03706.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.03706 v2

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T23:50:08.527402Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T06:45:38.442849Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T07:46:45.596273Z

Reference resolution

41 of 41 outbound references displayed

  • verified exact2
  • verified fuzzy20
  • unresolved19
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 40635779-7f77-48ac-8285-0b737af7a6e8 · outbound

This paper cites A historical perspective of adaptive control and learning,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches A historical perspective of adaptive control and learning,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:09.166600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:50:08.332465Z digest=sha256:6f243dec54f1302f0e456fce683b153c30bbfd18fa9740a14482630e18947318

Observation 79d64ad0-f90c-4a45-a8e8-623470f8671b · outbound

This paper cites Adaptive control and intersections with reinforce- ment learning,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Adaptive control and intersections with reinforce- ment learning,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:09.151575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:50:08.337313Z digest=sha256:f37faae265e900f10c90813eebdde18fcb746ae2501a72034582f9deca5b3901

Observation 1e1ca215-34cc-4ae8-bd92-4919ecda4fba · outbound

This paper cites Adaptive servomechanisms,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Adaptive servomechanisms,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:09.134123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:50:08.341131Z digest=sha256:98371afb38085620b1b7af5260106a08f19585a6d21d9b324e16ce047b491bf4

Observation 6754dfeb-50c7-4375-8cbf-623e35b59b05 · outbound

This paper cites an unresolved cited work.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-15T23:50:09.114356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:50:08.345879Z digest=sha256:bbd4b7fc90dfefdea485e6af6877970b5f6742a9efb45bf10ac2e726a652863f

Observation d1872603-5e5e-4a0a-a4e8-e36e3d04c776 · outbound

This paper cites On model-free adaptive control and its stability analysis,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches On model-free adaptive control and its stability analysis,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.351708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.351708Z digest=sha256:bcf41ce3155d6c3ac3e1d8da0556bf17cc3828ff59e558c0b00d9f1d052830ee

Observation a2795558-2437-4e4b-9395-12b01fdbd048 · outbound

This paper cites Adaptive control: Towards a complexity-based general theory,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Adaptive control: Towards a complexity-based general theory,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:09.087539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:50:08.357035Z digest=sha256:dac6edf7f358b5e2599d6cbee2124f26345961ea2e6d2179890069233e2d26bc

Observation d8923231-7d62-4183-8032-bcb0bc97e4de · outbound

This paper cites Reinforcement learning and adaptive dynamic programming for feedback control,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Reinforcement learning and adaptive dynamic programming for feedback control,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.363919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.363919Z digest=sha256:b7dc6db6c4e1427407a23b8bd553792521855c66289fecbc406a875ee17db64c

Observation 3a0b3cab-4804-4a66-abe0-8e87fb0f257f · outbound

This paper cites Value iteration and adaptive dynamic pro- gramming for data-driven adaptive optimal control design,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Value iteration and adaptive dynamic pro- gramming for data-driven adaptive optimal control design,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.369044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.369044Z digest=sha256:00dee4e89a2b05efa232b0a82fed00fb795f316fb75051f4b791c18ab51c7caa

Observation 8090c747-8c58-41cc-9560-edd4635f196c · outbound

This paper cites Certainty equivalence is efficient for linear quadratic control,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Certainty equivalence is efficient for linear quadratic control,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:09.048285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:50:08.373945Z digest=sha256:81ec218be820f0dcfd1a1d00e3ec7f9f9195abd698d6aa43780b80d8412645aa

Observation 4fa18f40-493e-411b-8341-725c61ef617c · outbound

This paper cites Almost surely √ T regret bound for adaptive LQR,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Almost surely √ T regret bound for adaptive LQR,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:09.033011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:50:08.379323Z digest=sha256:c7b4d14338b354588b26976eecf679b2f085154c34441a09fb463041afc3b9fe

Observation daaa9b28-d4ef-45d3-8697-379a637fde1b · outbound

This paper cites Naive exploration is optimal for online LQR,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Naive exploration is optimal for online LQR,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:09.015119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:50:08.383232Z digest=sha256:704b847c8d53ad3606198eeef61e9442d5cd3b9342510abb75a3b6002a6f6204

Observation 7a8461ba-5587-4f5b-bfc8-85b41f51998a · outbound

This paper cites Learning linear-quadratic regu- lators efficiently with only √ T regret,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Learning linear-quadratic regu- lators efficiently with only √ T regret,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:08.999150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:50:08.387634Z digest=sha256:d9bf8a99ad91619fae22e7600ad0bdfb3193d30a25dbd2c7edff9b996c90457e

Observation 348c57c0-fa5d-4ce6-b5ec-3ba56bfb08ad · outbound

This paper cites Any-Time Regret-Guaranteed Algorithm for Control of Linear Quadratic Systems.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Any-Time Regret-Guaranteed Algorithm for Control of Linear Quadratic Systems

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.392249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.392249Z digest=sha256:cc1751e75d09f70d42ac353ec0afb37b86d1de9f303099ca3e20ed80bbd8fa3e

Observation c7654087-cec3-456b-823c-d1be04f43458 · outbound

This paper cites Adaptive control by regulation-triggered batch least squares,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Adaptive control by regulation-triggered batch least squares,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:08.981265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:50:08.397500Z digest=sha256:39f56dbd371234b88077e6045043cfac0354bdb277e51f1dbac9a12d452a6b1c

Observation c84d02ea-3bf1-48b7-8f0a-a174a266f252 · outbound

This paper cites Robustness of Online Identification-based Policy Iteration to Noisy Data.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Robustness of Online Identification-based Policy Iteration to Noisy Data

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-15T23:50:08.675938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:50:08.402209Z digest=sha256:53ed84411e4354b5c32447c673ca394e500c90b8ac8da08d6c0d473cfdc1bf42

Observation 5e14a8e5-93cd-4d30-8066-96a8a8b50635 · outbound

This paper cites Failures of adaptive control theory and their resolu- tion,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Failures of adaptive control theory and their resolu- tion,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:08.961342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:50:08.407502Z digest=sha256:7407df0db43f9b058c2cbb9110b996694c1a1d3c0e959a36408892773f399497

Observation 796cf92e-fe24-4d46-88d7-4c93094a36af · outbound

This paper cites Toward a theoretical foundation of policy optimization for learning control policies,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Toward a theoretical foundation of policy optimization for learning control policies,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:08.942729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:50:08.412487Z digest=sha256:8da0dba6a97ab2deb2b677ecc18d1df96cae3a6638cfa993377ccae548ac22e5

Observation 5135def5-221d-4347-bd5e-403192e40d49 · outbound

This paper cites Global convergence of policy gradient methods for the linear quadratic regulator,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Global convergence of policy gradient methods for the linear quadratic regulator,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:08.922626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:50:08.417713Z digest=sha256:3818f1dc0c7a06e19092b54bd93e0b0f3e8f531773f82bcce3a371a1bd153c07

Observation d707ee08-2849-42bd-b133-ac35137a74ac · outbound

This paper cites Convergence and sample complexity of gradient methods for the model-free linear quadratic regulator problem,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Convergence and sample complexity of gradient methods for the model-free linear quadratic regulator problem,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:08.905794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:50:08.422090Z digest=sha256:bf5536925135e84260c946ab01cbbd07bc1a1e676f63578666924044e7c3e675

Observation f79ea46b-8c5e-446b-be6e-9408fab373e3 · outbound

This paper cites Global convergence of policy gradient primal-dual methods for risk-constrained LQRs,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Global convergence of policy gradient primal-dual methods for risk-constrained LQRs,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:08.888349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:50:08.426472Z digest=sha256:afead1b5b0df2d52f82a2f9d29bb40608f27f4709a37f9427f611c70224b74a6

Observation d52bcd24-bd87-42ec-8238-a01305777e41 · outbound

This paper cites Data-Enabled Policy Optimization for Direct Adaptive Learning of the LQR.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Data-Enabled Policy Optimization for Direct Adaptive Learning of the LQR

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.430560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.430560Z digest=sha256:56a1b2c9fc7d34ece4b0b39b1ce31ed5f18f13aa7978a04936100f28989d45e5

Observation d6121874-dc5e-4ae3-b462-0d7ebd8caf1d · outbound

This paper cites Integration of adaptive control and reinforcement learning for real-time control and learning,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Integration of adaptive control and reinforcement learning for real-time control and learning,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:08.872404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:50:08.435104Z digest=sha256:84da4a3c4458319b3791d303744680c925ee726fc53014e14c446e329c887369

Observation 8a1a81c1-a71c-4b3c-a4b8-a2bac70e6bed · outbound

This paper cites Behavioral systems theory in data-driven analysis, signal processing, and control,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Behavioral systems theory in data-driven analysis, signal processing, and control,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.439339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.439339Z digest=sha256:e48ae2312e529c897fddd03f9db4549b86f5126657d7f254b946f64148a3d8e3

Observation f570d65e-6893-4076-8c94-cbd937175c08 · outbound

This paper cites Data-enabled predictive control: In the shallows of the DeePC,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Data-enabled predictive control: In the shallows of the DeePC,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:08.847891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:50:08.443588Z digest=sha256:0de4cb95f8b790475790bc337536e5bd6516b9f7dadf59d1b69eb84827ecc938

Observation 818c93c5-e96d-49fa-9dab-c9ad41cd42bd · outbound

This paper cites Formulas for data-driven control: Stabilization, optimality, and robustness,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Formulas for data-driven control: Stabilization, optimality, and robustness,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.447741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.447741Z digest=sha256:c2a467cde492c734f67c21f6795fcbcac256dbd74f61904dc08f289195f6cde7

Observation 5ec8e207-be91-42ab-9bd8-ddde9a111726 · outbound

This paper cites On the certainty-equivalence ap- proach to direct data-driven lqr design,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches On the certainty-equivalence ap- proach to direct data-driven lqr design,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:08.821067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:50:08.451654Z digest=sha256:eec8057c34fa6551623cd565001b8e0d218046bda70f7d288ce4414c7009d8da

Observation 858d59ee-37f8-43b7-b344-d1ff981ef716 · outbound

This paper cites On the role of regularization in direct data-driven LQR control,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches On the role of regularization in direct data-driven LQR control,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:08.806460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:50:08.456028Z digest=sha256:e788ab3206db166fabc81e35fe048b65766c12adfe583e4792a3f8b64581d826

Observation 2fa8fff5-26c4-41b6-8d06-67e9f97c5b64 · outbound

This paper cites Harnessing uncertainty for a separation principle in direct data-driven predictive control,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Harnessing uncertainty for a separation principle in direct data-driven predictive control,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.461355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.461355Z digest=sha256:e37da01a8be47b507561a2c0cf5733f78fbc8938f47f2a63f3c5d7ea4c60f5d8

Observation fb0afc44-f3a2-4c0f-a157-39274face45e · outbound

This paper cites Data-enabled policy optimization for the linear quadratic regulator,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Data-enabled policy optimization for the linear quadratic regulator,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.466469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.466469Z digest=sha256:a484b69ca32ce3a72f817ca6073e8a4ec6e00d9af2ea6efd03bce8121ed29ab1

Observation 3aaee222-5001-4e26-a4b5-d4f173986bc7 · outbound

This paper cites Direct Adaptive Control of Grid-Connected Power Converters via Output-Feedback Data-Enabled Policy Optimization.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Direct Adaptive Control of Grid-Connected Power Converters via Output-Feedback Data-Enabled Policy Optimization

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.471807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.471807Z digest=sha256:82829e6be5313f31ed88d1dec46424e9ca349c1be76dcb0af2eca042b2ffc4e7

Observation 39141363-9792-455c-a12d-dd50f425848a · outbound

This paper cites Unified aeroelastic flutter and loads control via data-enabled policy optimiza- tion,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Unified aeroelastic flutter and loads control via data-enabled policy optimiza- tion,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:08.771986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:50:08.477278Z digest=sha256:786bf5f0f189d99a1f747fa9be45401a6a33a5aa8ce5b314025e00578998b975

Observation 769d165f-8ce3-44ba-bc7b-86e719bf7405 · outbound

This paper cites An Adaptive Data-Enabled Policy Optimization Approach for Autonomous Bicycle Control.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches An Adaptive Data-Enabled Policy Optimization Approach for Autonomous Bicycle Control

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.482040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.482040Z digest=sha256:02ebaf68d7a994fd6dab51de169cf092ad08248480125956894663e47f7619d9

Observation 046c9ba2-af68-4cb8-b29c-06052b16fea5 · outbound

This paper cites An iterative technique for the computation of the steady state gains for the discrete optimal regulator,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches An iterative technique for the computation of the steady state gains for the discrete optimal regulator,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.487421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.487421Z digest=sha256:a3d3dbb05e21dfafe2532b0aafdf3c39d174eb5e4d3610973d95f0c9bdfb9293

Observation 0498b88e-123e-497d-b064-109e6af19fe3 · outbound

This paper cites Regularization for Covariance Parameterization of Direct Data-Driven LQR Control.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Regularization for Covariance Parameterization of Direct Data-Driven LQR Control

Reference 34

Resolution
verified exact
local_arxiv, observed 2026-08-15T23:50:08.603614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:50:08.492398Z digest=sha256:d682de034b214a1a7344899e5f9ec669599cb13d64e95fe22ea37bd878a590c4

Observation e9bcef5c-0a59-4723-a4f4-1515719219ca · outbound

This paper cites On the sample com- plexity of the linear quadratic regulator,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches On the sample com- plexity of the linear quadratic regulator,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.497648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.497648Z digest=sha256:001e1183966f8358a6482a157e39ed8000b2ac9561fb7dd205ecc692c867bd57

Observation ea4545a8-3e5b-48f3-9fa6-0b576276e096 · outbound

This paper cites an unresolved cited work.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.503348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.503348Z digest=sha256:a86bd8dbc7e23ec8760a1eb549ab69099a65ed981ab8393cd4ac7813fe3fea94

Observation 44bd1052-5b6d-4799-a175-07fd004fd393 · outbound

This paper cites A note on persistency of excitation,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches A note on persistency of excitation,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.508573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.508573Z digest=sha256:4e6a9d0ce3e73ed336adccb7dbff3a87b46089815ccbef0ad80dc914ca5840a0

Observation 5f762dbb-cd0f-45bf-b362-50bbd39dca49 · outbound

This paper cites The informativity approach: To data-driven analysis and control,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches The informativity approach: To data-driven analysis and control,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.513091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.513091Z digest=sha256:7cc334b9c24b70a52b1ae1ae95c7f1a4b2eb11b3ca9aed080429192b6846af99

Observation 611fa837-4fbe-4d4f-ba9c-56f369aee887 · outbound

This paper cites A quantitative notion of persistency of excitation and the robust fundamental lemma,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches A quantitative notion of persistency of excitation and the robust fundamental lemma,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:08.709079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:50:08.517946Z digest=sha256:a7dce6c3e2715c71eb113021b06cfe5cd3a5008d1544c2bbd27d4ab64cf732a8

Observation c7215d5a-ae1b-4e40-9a12-5562565f8ae0 · outbound

This paper cites LQR through the Lens of First Order Methods: Discrete-time Case.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches LQR through the Lens of First Order Methods: Discrete-time Case

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.522490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.522490Z digest=sha256:c1573bccdd81f7db4e74c8d4f6014eefdc34bb7b2a54eb470924e8f462db507a

Observation 786554f1-64b4-453d-a84c-6c3e2fce4fd9 · outbound

This paper cites Noise Sensitivity of the Semidefinite Programs for Direct Data-Driven LQR.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Noise Sensitivity of the Semidefinite Programs for Direct Data-Driven LQR

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.527402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.527402Z digest=sha256:e019a719ef7c48f0e77b135061f05aa3540e939c21edca8ff2151077c77eeb92

Pith citing papers

Observation 01be5d16-edf7-4620-b66e-5ae49c8aea9b · inbound

Stability of Certainty-Equivalent Adaptive LQR for Linear Systems with Unknown Time-Varying Parameters cites this paper.

Stability of Certainty-Equivalent Adaptive LQR for Linear Systems with Unknown Time-Varying Parameters Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-18T00:00:31.571050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-17T23:57:37.299127Z digest=sha256:803e1eefe19e4f632c670ccd36155f098b289793139e2dd7036b87810bb0417e

Observation d8f09bf2-027c-43d8-ad5a-45320c53b051 · inbound

Sample-Efficient Model-Free Policy Gradient Methods for Stochastic LQR via Robust Linear Regression cites this paper.

Sample-Efficient Model-Free Policy Gradient Methods for Stochastic LQR via Robust Linear Regression Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T02:38:53.954364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-17T02:35:57.661008Z digest=sha256:e17945a8244347967f9439d3c9927955b7b41d667d3ac6d1bfd4b80d8c173cf4

Observation fea2e6a9-c1f7-462e-8c3c-11b0ed6ba680 · inbound

Sample-Efficient Model-Free Policy Gradient Methods for Stochastic LQR via Robust Linear Regression cites this paper.

Sample-Efficient Model-Free Policy Gradient Methods for Stochastic LQR via Robust Linear Regression Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches

Reference 5014

Resolution
unresolved
no resolver link, observed 2026-08-04T06:45:38.442849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:45:38.442849Z digest=sha256:4e42221936ccf00c56e65dec8403748e073fb620f990a8c1acc1e3b552761565

Observation 969faada-9e6f-422a-932d-b8aaccc294d4 · inbound

Global Convergence of Policy Gradient Methods for ReLU Controllers in Linear Quadratic Regulation cites this paper.

Global Convergence of Policy Gradient Methods for ReLU Controllers in Linear Quadratic Regulation Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:36:15.031170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-08T11:25:06.725383Z digest=sha256:95e9b42ae5df279a93e0e8049f09b872449136617c11207eba5c5b363982eea7

Observation dd5455a6-cb58-4481-8c66-79f057739f43 · inbound

Direct Data-Driven Linear Quadratic Tracking via Policy Optimization cites this paper.

Direct Data-Driven Linear Quadratic Tracking via Policy Optimization Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-05-20T19:48:57.395173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-05-20T19:47:30.108992Z digest=sha256:01581ea4164a2cd6e3c052c5d1eeb85612ac9fd2a669ccdb477ffd0c97cae7c7

Observation 5d1047a6-a758-4c11-b5b9-f20a3b89c917 · inbound

A Data-Enabled Primal-Dual Approach for Policy Learning with SDP Formulations cites this paper.

A Data-Enabled Primal-Dual Approach for Policy Learning with SDP Formulations Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-02T07:46:45.597817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-02T07:39:47.486139Z digest=sha256:3342385699b0d4901620bb5e9227020d325f636e59a55cec01b955cdfa5596bc

Observation fb1bd19f-a6e3-440c-8130-5a04fd783a0a · inbound

Adaptive Linear Quadratic Control of Unknown Linear Time-Varying Systems via Policy Gradient Methods cites this paper.

Adaptive Linear Quadratic Control of Unknown Linear Time-Varying Systems via Policy Gradient Methods Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches

Reference 39

Resolution
unresolved
no resolver link, observed 2026-07-12T03:48:44.349867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T03:48:44.349867Z digest=sha256:0d213b09d2884e775eb4cba1bf3f95d066a34ab8618bf6fd8ba774a7e263606f