Pith. sign in

Paper Citation Record · LEDGER

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches

As of 17 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 7 inbound Pith citation observations for arXiv:2505.03706.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.03706 v2

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T23:50:08.527402Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T06:45:38.442849Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T07:46:45.596273Z

Reference resolution

41 of 41 outbound references displayed

  • verified exact2
  • verified fuzzy20
  • unresolved19
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 40635779-7f77-48ac-8285-0b737af7a6e8 · outbound

This paper cites A historical perspective of adaptive control and learning,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches A historical perspective of adaptive control and learning,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:09.166600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:50:08.332465Z digest=sha256:9e753f3d3d80cbac0b25daf70a0a0cd9b63953bf7b17d16f72da2d2a02f2bd46

Observation 79d64ad0-f90c-4a45-a8e8-623470f8671b · outbound

This paper cites Adaptive control and intersections with reinforce- ment learning,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Adaptive control and intersections with reinforce- ment learning,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:09.151575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:50:08.337313Z digest=sha256:8b90452b5c9230b6f11731c3b242ac16e00fc6689138c158af91227a6aca5b0f

Observation 1e1ca215-34cc-4ae8-bd92-4919ecda4fba · outbound

This paper cites Adaptive servomechanisms,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Adaptive servomechanisms,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:09.134123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:50:08.341131Z digest=sha256:97b51e1c9d48b075329e33ec7715fbcb7c0a3770a5fc648a665bc2f09940a523

Observation 6754dfeb-50c7-4375-8cbf-623e35b59b05 · outbound

This paper cites an unresolved cited work.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-15T23:50:09.114356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:50:08.345879Z digest=sha256:4fe1bff200058cd6c8e24c2fd7aa3a22f8a17c9d8c8bf3b16039c0669208d883

Observation d1872603-5e5e-4a0a-a4e8-e36e3d04c776 · outbound

This paper cites On model-free adaptive control and its stability analysis,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches On model-free adaptive control and its stability analysis,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.351708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.351708Z digest=sha256:bcf41ce3155d6c3ac3e1d8da0556bf17cc3828ff59e558c0b00d9f1d052830ee

Observation a2795558-2437-4e4b-9395-12b01fdbd048 · outbound

This paper cites Adaptive control: Towards a complexity-based general theory,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Adaptive control: Towards a complexity-based general theory,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:09.087539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:50:08.357035Z digest=sha256:d6ecf79e4b042a7f0f2f1f9c6085c604244821a083fce4915a830125426ae8c2

Observation d8923231-7d62-4183-8032-bcb0bc97e4de · outbound

This paper cites Reinforcement learning and adaptive dynamic programming for feedback control,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Reinforcement learning and adaptive dynamic programming for feedback control,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.363919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.363919Z digest=sha256:b7dc6db6c4e1427407a23b8bd553792521855c66289fecbc406a875ee17db64c

Observation 3a0b3cab-4804-4a66-abe0-8e87fb0f257f · outbound

This paper cites Value iteration and adaptive dynamic pro- gramming for data-driven adaptive optimal control design,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Value iteration and adaptive dynamic pro- gramming for data-driven adaptive optimal control design,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.369044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.369044Z digest=sha256:00dee4e89a2b05efa232b0a82fed00fb795f316fb75051f4b791c18ab51c7caa

Observation 8090c747-8c58-41cc-9560-edd4635f196c · outbound

This paper cites Certainty equivalence is efficient for linear quadratic control,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Certainty equivalence is efficient for linear quadratic control,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:09.048285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:50:08.373945Z digest=sha256:ca6ab13715a374051c3b814828fd61e81e586f282b40f4b74a5b9546def58d2f

Observation 4fa18f40-493e-411b-8341-725c61ef617c · outbound

This paper cites Almost surely √ T regret bound for adaptive LQR,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Almost surely √ T regret bound for adaptive LQR,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:09.033011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:50:08.379323Z digest=sha256:dab9808d16060b3e483e3dd30eb0fa721f257270d9222f87a0035eb9b00db3a5

Observation daaa9b28-d4ef-45d3-8697-379a637fde1b · outbound

This paper cites Naive exploration is optimal for online LQR,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Naive exploration is optimal for online LQR,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:09.015119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:50:08.383232Z digest=sha256:1c1d58d20f1fab3f286066492a356ab5b77773544e08825f03c17f2dcb30b70f

Observation 7a8461ba-5587-4f5b-bfc8-85b41f51998a · outbound

This paper cites Learning linear-quadratic regu- lators efficiently with only √ T regret,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Learning linear-quadratic regu- lators efficiently with only √ T regret,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:08.999150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:50:08.387634Z digest=sha256:3e3a9b1ad6bc1f010bb6d716f754304a0addfedd5d117745e2c2a94fed26d08c

Observation 348c57c0-fa5d-4ce6-b5ec-3ba56bfb08ad · outbound

This paper cites Any-Time Regret-Guaranteed Algorithm for Control of Linear Quadratic Systems.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Any-Time Regret-Guaranteed Algorithm for Control of Linear Quadratic Systems

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.392249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.392249Z digest=sha256:71bc170a9f191c8bb25c2e913cd1e82faf56b5953b731c68ab42c17e80af02b0

Observation c7654087-cec3-456b-823c-d1be04f43458 · outbound

This paper cites Adaptive control by regulation-triggered batch least squares,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Adaptive control by regulation-triggered batch least squares,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:08.981265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:50:08.397500Z digest=sha256:0a66bd59f7e806677359c20c1edbcf199732103c8092e26af461a32100e390df

Observation c84d02ea-3bf1-48b7-8f0a-a174a266f252 · outbound

This paper cites Robustness of Online Identification-based Policy Iteration to Noisy Data.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Robustness of Online Identification-based Policy Iteration to Noisy Data

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-15T23:50:08.675938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:50:08.402209Z digest=sha256:c596e4a25a43c4fc96cb874d8cfd6fd8f9ba7892e10e98f3768026b9ec592c41

Observation 5e14a8e5-93cd-4d30-8066-96a8a8b50635 · outbound

This paper cites Failures of adaptive control theory and their resolu- tion,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Failures of adaptive control theory and their resolu- tion,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:08.961342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:50:08.407502Z digest=sha256:6bb1bdc18dfe54544e303e4fcd4e6521431ce7fc42670dc8e29a30547bb27e4d

Observation 796cf92e-fe24-4d46-88d7-4c93094a36af · outbound

This paper cites Toward a theoretical foundation of policy optimization for learning control policies,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Toward a theoretical foundation of policy optimization for learning control policies,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:08.942729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:50:08.412487Z digest=sha256:186a8d22704e1cf750dfca02f5750412f5089963b08a3b02db8102ac0162d9a6

Observation 5135def5-221d-4347-bd5e-403192e40d49 · outbound

This paper cites Global convergence of policy gradient methods for the linear quadratic regulator,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Global convergence of policy gradient methods for the linear quadratic regulator,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:08.922626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:50:08.417713Z digest=sha256:9bd70d26bc1d5ffd4bfd2b7d61b7bbca00b43876583c30d574be255299d7dcb0

Observation d707ee08-2849-42bd-b133-ac35137a74ac · outbound

This paper cites Convergence and sample complexity of gradient methods for the model-free linear quadratic regulator problem,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Convergence and sample complexity of gradient methods for the model-free linear quadratic regulator problem,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:08.905794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:50:08.422090Z digest=sha256:07cc39351d6a6cfab0cb267a9cbc3d46f516e148724705b95928d2a65626edee

Observation f79ea46b-8c5e-446b-be6e-9408fab373e3 · outbound

This paper cites Global convergence of policy gradient primal-dual methods for risk-constrained LQRs,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Global convergence of policy gradient primal-dual methods for risk-constrained LQRs,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:08.888349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:50:08.426472Z digest=sha256:eb631ec75d7bf8fe9e117747e258235c196118de660ae0df5b4bf4eaef8fc5dd

Observation d52bcd24-bd87-42ec-8238-a01305777e41 · outbound

This paper cites Data-Enabled Policy Optimization for Direct Adaptive Learning of the LQR.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Data-Enabled Policy Optimization for Direct Adaptive Learning of the LQR

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.430560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.430560Z digest=sha256:56a1b2c9fc7d34ece4b0b39b1ce31ed5f18f13aa7978a04936100f28989d45e5

Observation d6121874-dc5e-4ae3-b462-0d7ebd8caf1d · outbound

This paper cites Integration of adaptive control and reinforcement learning for real-time control and learning,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Integration of adaptive control and reinforcement learning for real-time control and learning,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:08.872404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:50:08.435104Z digest=sha256:ede8d4f67797d267eb366a08e3b02982d3d369fc01f1caf17400aae3e3dfce5d

Observation 8a1a81c1-a71c-4b3c-a4b8-a2bac70e6bed · outbound

This paper cites Behavioral systems theory in data-driven analysis, signal processing, and control,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Behavioral systems theory in data-driven analysis, signal processing, and control,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.439339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.439339Z digest=sha256:e48ae2312e529c897fddd03f9db4549b86f5126657d7f254b946f64148a3d8e3

Observation f570d65e-6893-4076-8c94-cbd937175c08 · outbound

This paper cites Data-enabled predictive control: In the shallows of the DeePC,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Data-enabled predictive control: In the shallows of the DeePC,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:08.847891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:50:08.443588Z digest=sha256:a3e534ba3502902e5599d6dbc0910fffbefd252960a24501c8fa0ae8561265c5

Observation 818c93c5-e96d-49fa-9dab-c9ad41cd42bd · outbound

This paper cites Formulas for data-driven control: Stabilization, optimality, and robustness,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Formulas for data-driven control: Stabilization, optimality, and robustness,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.447741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.447741Z digest=sha256:c2a467cde492c734f67c21f6795fcbcac256dbd74f61904dc08f289195f6cde7

Observation 5ec8e207-be91-42ab-9bd8-ddde9a111726 · outbound

This paper cites On the certainty-equivalence ap- proach to direct data-driven lqr design,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches On the certainty-equivalence ap- proach to direct data-driven lqr design,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:08.821067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:50:08.451654Z digest=sha256:edb06320a7edc52523980db28683e5df1b7384a926b7ee6c9a3c3d8b75320d55

Observation 858d59ee-37f8-43b7-b344-d1ff981ef716 · outbound

This paper cites On the role of regularization in direct data-driven LQR control,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches On the role of regularization in direct data-driven LQR control,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:08.806460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:50:08.456028Z digest=sha256:230376b067ffd35e05c1c0a6378923233fe48511bc30c5e6bdb21fea95524811

Observation 2fa8fff5-26c4-41b6-8d06-67e9f97c5b64 · outbound

This paper cites Harnessing uncertainty for a separation principle in direct data-driven predictive control,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Harnessing uncertainty for a separation principle in direct data-driven predictive control,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.461355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.461355Z digest=sha256:e37da01a8be47b507561a2c0cf5733f78fbc8938f47f2a63f3c5d7ea4c60f5d8

Observation fb0afc44-f3a2-4c0f-a157-39274face45e · outbound

This paper cites Data-enabled policy optimization for the linear quadratic regulator,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Data-enabled policy optimization for the linear quadratic regulator,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.466469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.466469Z digest=sha256:a484b69ca32ce3a72f817ca6073e8a4ec6e00d9af2ea6efd03bce8121ed29ab1

Observation 3aaee222-5001-4e26-a4b5-d4f173986bc7 · outbound

This paper cites Direct Adaptive Control of Grid-Connected Power Converters via Output-Feedback Data-Enabled Policy Optimization.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Direct Adaptive Control of Grid-Connected Power Converters via Output-Feedback Data-Enabled Policy Optimization

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.471807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.471807Z digest=sha256:624b2c61fcc6bc1e71bbcdb70efdbff52928bcbd662b291f107753503ece42c1

Observation 39141363-9792-455c-a12d-dd50f425848a · outbound

This paper cites Unified aeroelastic flutter and loads control via data-enabled policy optimiza- tion,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Unified aeroelastic flutter and loads control via data-enabled policy optimiza- tion,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:08.771986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:50:08.477278Z digest=sha256:50cb0165d40c2c60bfaec351bd7690f069e16aeca8212e8223eb254b3026bd21

Observation 769d165f-8ce3-44ba-bc7b-86e719bf7405 · outbound

This paper cites An Adaptive Data-Enabled Policy Optimization Approach for Autonomous Bicycle Control.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches An Adaptive Data-Enabled Policy Optimization Approach for Autonomous Bicycle Control

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.482040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.482040Z digest=sha256:02ebaf68d7a994fd6dab51de169cf092ad08248480125956894663e47f7619d9

Observation 046c9ba2-af68-4cb8-b29c-06052b16fea5 · outbound

This paper cites An iterative technique for the computation of the steady state gains for the discrete optimal regulator,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches An iterative technique for the computation of the steady state gains for the discrete optimal regulator,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.487421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.487421Z digest=sha256:a3d3dbb05e21dfafe2532b0aafdf3c39d174eb5e4d3610973d95f0c9bdfb9293

Observation 0498b88e-123e-497d-b064-109e6af19fe3 · outbound

This paper cites Regularization for Covariance Parameterization of Direct Data-Driven LQR Control.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Regularization for Covariance Parameterization of Direct Data-Driven LQR Control

Reference 34

Resolution
verified exact
local_arxiv, observed 2026-08-15T23:50:08.603614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:50:08.492398Z digest=sha256:98b5579c69b385d1baa22fcb63eefec3183e4b5702eb9c586a2c60845ddfe36e

Observation e9bcef5c-0a59-4723-a4f4-1515719219ca · outbound

This paper cites On the sample com- plexity of the linear quadratic regulator,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches On the sample com- plexity of the linear quadratic regulator,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.497648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.497648Z digest=sha256:001e1183966f8358a6482a157e39ed8000b2ac9561fb7dd205ecc692c867bd57

Observation ea4545a8-3e5b-48f3-9fa6-0b576276e096 · outbound

This paper cites an unresolved cited work.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.503348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.503348Z digest=sha256:a86bd8dbc7e23ec8760a1eb549ab69099a65ed981ab8393cd4ac7813fe3fea94

Observation 44bd1052-5b6d-4799-a175-07fd004fd393 · outbound

This paper cites A note on persistency of excitation,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches A note on persistency of excitation,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.508573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.508573Z digest=sha256:4e6a9d0ce3e73ed336adccb7dbff3a87b46089815ccbef0ad80dc914ca5840a0

Observation 5f762dbb-cd0f-45bf-b362-50bbd39dca49 · outbound

This paper cites The informativity approach: To data-driven analysis and control,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches The informativity approach: To data-driven analysis and control,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.513091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.513091Z digest=sha256:7cc334b9c24b70a52b1ae1ae95c7f1a4b2eb11b3ca9aed080429192b6846af99

Observation 611fa837-4fbe-4d4f-ba9c-56f369aee887 · outbound

This paper cites A quantitative notion of persistency of excitation and the robust fundamental lemma,.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches A quantitative notion of persistency of excitation and the robust fundamental lemma,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:50:08.709079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:50:08.517946Z digest=sha256:8073eb2e26849b22a3f87cdf29f52cac6a2d09261d2f71c97896a3214d4d7af4

Observation c7215d5a-ae1b-4e40-9a12-5562565f8ae0 · outbound

This paper cites LQR through the Lens of First Order Methods: Discrete-time Case.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches LQR through the Lens of First Order Methods: Discrete-time Case

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.522490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.522490Z digest=sha256:c1573bccdd81f7db4e74c8d4f6014eefdc34bb7b2a54eb470924e8f462db507a

Observation 786554f1-64b4-453d-a84c-6c3e2fce4fd9 · outbound

This paper cites Noise Sensitivity of the Semidefinite Programs for Direct Data-Driven LQR.

Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches Noise Sensitivity of the Semidefinite Programs for Direct Data-Driven LQR

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:08.527402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:08.527402Z digest=sha256:e019a719ef7c48f0e77b135061f05aa3540e939c21edca8ff2151077c77eeb92

Pith citing papers

Observation 01be5d16-edf7-4620-b66e-5ae49c8aea9b · inbound

Stability of Certainty-Equivalent Adaptive LQR for Linear Systems with Unknown Time-Varying Parameters cites this paper.

Stability of Certainty-Equivalent Adaptive LQR for Linear Systems with Unknown Time-Varying Parameters Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-18T00:00:31.571050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-17T23:57:37.299127Z digest=sha256:5327e63a5c796c7202fb9c59bed2db8a11ed7d8f90054478cd806b0d964a1736

Observation d8f09bf2-027c-43d8-ad5a-45320c53b051 · inbound

Sample-Efficient Model-Free Policy Gradient Methods for Stochastic LQR via Robust Linear Regression cites this paper.

Sample-Efficient Model-Free Policy Gradient Methods for Stochastic LQR via Robust Linear Regression Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T02:38:53.954364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-17T02:35:57.661008Z digest=sha256:76fd9e2b66ef6a680451acb0f0d45a98007a3aee25cc972ec007a351cfda0c4d

Observation fea2e6a9-c1f7-462e-8c3c-11b0ed6ba680 · inbound

Sample-Efficient Model-Free Policy Gradient Methods for Stochastic LQR via Robust Linear Regression cites this paper.

Sample-Efficient Model-Free Policy Gradient Methods for Stochastic LQR via Robust Linear Regression Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches

Reference 5014

Resolution
unresolved
no resolver link, observed 2026-08-04T06:45:38.442849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:45:38.442849Z digest=sha256:4e42221936ccf00c56e65dec8403748e073fb620f990a8c1acc1e3b552761565

Observation 969faada-9e6f-422a-932d-b8aaccc294d4 · inbound

Global Convergence of Policy Gradient Methods for ReLU Controllers in Linear Quadratic Regulation cites this paper.

Global Convergence of Policy Gradient Methods for ReLU Controllers in Linear Quadratic Regulation Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:36:15.031170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-08T11:25:06.725383Z digest=sha256:720029a8e674d123afd05d6afd8a6c3464f2f67aa8b6118259cfc5992b5a6760

Observation dd5455a6-cb58-4481-8c66-79f057739f43 · inbound

Direct Data-Driven Linear Quadratic Tracking via Policy Optimization cites this paper.

Direct Data-Driven Linear Quadratic Tracking via Policy Optimization Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-05-20T19:48:57.395173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-05-20T19:47:30.108992Z digest=sha256:916808e47ef6885d300ac38d96e0bbeac19c840ff8add6f2b68290cb122c62a6

Observation 5d1047a6-a758-4c11-b5b9-f20a3b89c917 · inbound

A Data-Enabled Primal-Dual Approach for Policy Learning with SDP Formulations cites this paper.

A Data-Enabled Primal-Dual Approach for Policy Learning with SDP Formulations Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-02T07:46:45.597817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-07-02T07:39:47.486139Z digest=sha256:da966c4afb1e64dcfda3345e5950a9c5eae8949b2567bd4fc397b2c9825f9a84

Observation fb1bd19f-a6e3-440c-8130-5a04fd783a0a · inbound

Adaptive Linear Quadratic Control of Unknown Linear Time-Varying Systems via Policy Gradient Methods cites this paper.

Adaptive Linear Quadratic Control of Unknown Linear Time-Varying Systems via Policy Gradient Methods Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches

Reference 39

Resolution
unresolved
no resolver link, observed 2026-07-12T03:48:44.349867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T03:48:44.349867Z digest=sha256:0d213b09d2884e775eb4cba1bf3f95d066a34ab8618bf6fd8ba774a7e263606f