Pith. sign in

Paper Citation Record · LEDGER

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems

As of 9 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 5 inbound Pith citation observations for arXiv:2507.00358.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.00358 v2

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:33:38.321579Z

measured 54 of 54 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T11:22:20.635669Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T17:18:43.032059Z

Reference resolution

49 of 49 outbound references displayed

  • verified exact0
  • verified fuzzy34
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 53c011f9-07ed-4843-baeb-28990310a5df · outbound

This paper cites Sublinear Regret for a Class of Continuous-Time Linear-Quadratic Reinforcement Learning Problems.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Sublinear Regret for a Class of Continuous-Time Linear-Quadratic Reinforcement Learning Problems

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:32.491006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:32.491006Z digest=sha256:a113a069decca420bfe4d34fa46bf3d655120314b0fcbe50ac4c3f853f41d785

Observation a416a5af-1070-4a81-87d0-f759c5104c62 · outbound

This paper cites an unresolved cited work.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:33:50.572667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:32.598128Z digest=sha256:9754bb669457a9b2690fd6fe68678464062a64fe5bd3ba70841c6b0784f4d67c

Observation 3319b20d-c984-487a-8f7e-7b9980ff7326 · outbound

This paper cites Yong and X.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Yong and X

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:50.346689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:32.682350Z digest=sha256:a04a4a5a182c8d4f0ea4831b459f3c1d0e767c13c3d9952cdab5db5b205978ae

Observation a0260570-6ce0-4368-bfca-79e54bd1e914 · outbound

This paper cites Stochastic linear quadratic regulators with indefinite control weight costs,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Stochastic linear quadratic regulators with indefinite control weight costs,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:50.100535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:32.785792Z digest=sha256:3e3fb535f2c64908cba230295422d55300e2f53a8fa3e16367ddd0cd796a26ee

Observation 9a56b895-bbdd-4b7b-9b39-ad595502bcb5 · outbound

This paper cites Well-posedness and attainability of indefinite stochastic linear quadratic control in infinite time horizon,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Well-posedness and attainability of indefinite stochastic linear quadratic control in infinite time horizon,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:49.849387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:32.875427Z digest=sha256:84181f92c00110bf64cd4ba09462c37aea960625a6b0bf9ceefe79c5906d3a58

Observation 56b4a220-ed46-4e8d-bf98-140e0303d25c · outbound

This paper cites Linear matrix inequalities, riccati equations, and indefinite stochastic linear quadratic controls,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Linear matrix inequalities, riccati equations, and indefinite stochastic linear quadratic controls,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:49.628482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:32.970981Z digest=sha256:9c710f7168cf7fdd6ebf5e2c07bb4cce04d7928989b5c5e679c92b6bf74f8d70

Observation 6a310869-e25f-46c5-b6e2-e8ea4743ab12 · outbound

This paper cites Solvability and asymptotic behavior of generalized riccati equations arising in indefinite stochastic lq controls,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Solvability and asymptotic behavior of generalized riccati equations arising in indefinite stochastic lq controls,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:49.359104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:33.097423Z digest=sha256:a87bf427d4c891a304869f3fdc5e50d9e8a9875b1f9d1ed2df78f08f252fe061

Observation 5304aed6-de98-4bb2-9c9f-dd2410cfa39f · outbound

This paper cites A primal-dual semi-definite programming approach to linear quadratic control,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems A primal-dual semi-definite programming approach to linear quadratic control,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:47.792609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:33.187398Z digest=sha256:34b5b69718d834e4cc5922830da9653d1c60b709809099e319e9a56677daf92e

Observation c4910c53-0ea6-4635-b660-f8dc3fdc3498 · outbound

This paper cites Stabilization control for itˆ o stochastic system with indefinite state and control weight costs,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Stabilization control for itˆ o stochastic system with indefinite state and control weight costs,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:47.001170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:33.296546Z digest=sha256:4a25a55f45802b11a6ccd5058c46196df6233d450be4df715e306c236c2fd172

Observation 0cc3dcd9-378d-4d3e-a63d-a3fe1fe2e24e · outbound

This paper cites Optimal regulators for a class of nonlinear stochastic systems,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Optimal regulators for a class of nonlinear stochastic systems,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:46.604066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:33.397836Z digest=sha256:61d9ef73a6e43a8c75bc22c7fd24285156d4a5122f6d66dfd73f4d4d206338bd

Observation 224c8670-cdc7-48ce-bd8a-0e48d8985ff7 · outbound

This paper cites Indefinite linear-quadratic optimal control of mean-field stochas- tic differential equation with jump diffusion: an equivalent cost functional method,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Indefinite linear-quadratic optimal control of mean-field stochas- tic differential equation with jump diffusion: an equivalent cost functional method,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:46.262692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:33.501366Z digest=sha256:3dad6613bcae1edd270e9d7bb4ef84e51dcd4f34f88fd83cb1084afccac37689

Observation 9ffde277-e768-4d2b-a6fa-a3f77730792b · outbound

This paper cites Stochastic linear quadratic optimal control problems with regime- switching jumps in infinite horizon,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Stochastic linear quadratic optimal control problems with regime- switching jumps in infinite horizon,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:45.935693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:33.595650Z digest=sha256:998d572aa6ea4da345bfef97a195b8074733058da5731af95f617420eba3f58c

Observation c30d62f0-e968-4f7c-925b-97b18f70c6c6 · outbound

This paper cites On estimating the expected return on the market: An exploratory investiga- tion,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems On estimating the expected return on the market: An exploratory investiga- tion,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:45.607054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:33.693784Z digest=sha256:a10757c716b02d57f578b708e180a0c07cd330467a37fe71f220e6529d7b5919

Observation 22553b20-340f-4dc7-840b-ee89e39fa168 · outbound

This paper cites an unresolved cited work.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:33:45.269404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:33.776096Z digest=sha256:84b173b52494dbb21a4cbaaf5e52ab5eb3486e5a2572dfbc2c95f5b64ef67b79

Observation 7081615b-f0b1-426c-86b4-783149d67788 · outbound

This paper cites Rustem and M.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Rustem and M

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:45.008450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:33.870263Z digest=sha256:6d36458aa0889feec30d26317255849d0067f86a06095cbe2447bd2803d16b59

Observation 37632502-7169-4d11-845d-733f8f20b593 · outbound

This paper cites an unresolved cited work.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:33:44.729693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:33.944653Z digest=sha256:ffa6134a21aa2cdcf9a1a7cbddeebfa2f04875f3d937cfd996bfcaa485a79c35

Observation c75ffc6c-70b2-4f62-ae92-8ac9ea4ad5bf · outbound

This paper cites A survey on intrinsic motivation in reinforcement learning.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems A survey on intrinsic motivation in reinforcement learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:34.035769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:34.035769Z digest=sha256:badc96ea9e4852fb01cd37b4dec54165f6d8fa7b0da700c1b8883c439adda247

Observation 5ec4e6e4-16a0-4143-a74f-b77f89cebd07 · outbound

This paper cites Formal theory of creativity, fun, and intrinsic motivation (1990–2010),.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Formal theory of creativity, fun, and intrinsic motivation (1990–2010),

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:44.443994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:34.126427Z digest=sha256:399ae845a371ed519e03e266993cb3dc1300c4087741c107bb05ae220b7dc636

Observation 86e125e1-8cb6-47dd-acbd-2a38acffa9ff · outbound

This paper cites Surprise-Based Intrinsic Motivation for Deep Reinforcement Learning.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Surprise-Based Intrinsic Motivation for Deep Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:34.257469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:34.257469Z digest=sha256:f4b6ec7eb7f751c430aa9745833fa8ba1be27bc77b9f6279fe499f2438cf660a

Observation 347fd813-d80b-41d9-a067-a87eebf867ef · outbound

This paper cites Curiosity-driven exploration by self- supervised prediction,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Curiosity-driven exploration by self- supervised prediction,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:44.157257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:34.353252Z digest=sha256:3cf1b86b60291152e254a9d94a9a4d208b141890bc56a8fec2a14b059943e0f5

Observation 84763f26-e404-4299-83e2-4f9271c8ff3d · outbound

This paper cites Large-Scale Study of Curiosity-Driven Learning.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Large-Scale Study of Curiosity-Driven Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:34.448173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:34.448173Z digest=sha256:8e3ef43656dc3e69de2f8fb6acbf5a360c091cbe42fcb5a129a90a5c3dad187f

Observation 10d21d05-a695-4b4e-9f32-1cfd6bd8fa23 · outbound

This paper cites Exploration by Random Network Distillation.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Exploration by Random Network Distillation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:34.568573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:34.568573Z digest=sha256:d7735a081442827caa20e0b5814c9dc3f41748d45e622dddfbbe97a5587d7f08

Observation d9708eae-0a06-4e27-8b0a-39fdd77d889f · outbound

This paper cites Randomized prior functions for deep reinforcement learning,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Randomized prior functions for deep reinforcement learning,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:43.818381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:34.642461Z digest=sha256:03f940af906db55b3f9aabfac08ac5ccf6cb916606c4c18e0c398e99c8452c1c

Observation bc9ef8ca-a6c4-49c3-b588-e076d4af1322 · outbound

This paper cites Fast active learning for pure exploration in reinforcement learning,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Fast active learning for pure exploration in reinforcement learning,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:43.572233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:34.727222Z digest=sha256:5515c98e37cada9d81af5466c5a8b7b3285eb627df1386a4e058bc7c184365ab

Observation e0d6309e-1719-4bf5-a7d2-7193a8dae6a2 · outbound

This paper cites # exploration: A study of count-based exploration for deep reinforcement learning,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems # exploration: A study of count-based exploration for deep reinforcement learning,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:43.310875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:34.830998Z digest=sha256:aa3a6d8a295ef1fb838b697075e18c739ddb93d309faafd1227838fd6d6155d5

Observation 16a5722b-ccca-4208-bb58-b7412a6a7e36 · outbound

This paper cites Go-Explore: a New Approach for Hard-Exploration Problems.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Go-Explore: a New Approach for Hard-Exploration Problems

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:34.943251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:34.943251Z digest=sha256:038ff11e73bb37ba79b8d579bdfc1900b394d40719c0e7b0f43fcdae3b94d546

Observation d28035f0-aadb-4915-83d3-90f2e7e7cc3d · outbound

This paper cites First return, then explore,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems First return, then explore,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:43.039615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:35.051690Z digest=sha256:81ac053bd8b94d709252dcdad676e6026078906bcafa85b1912bf61e212f2e49

Observation b25f14d2-0ef2-4fc4-98fc-5b3db13d6108 · outbound

This paper cites Leveraging Demonstrations for Deep Reinforcement Learning on Robotics Problems with Sparse Rewards.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Leveraging Demonstrations for Deep Reinforcement Learning on Robotics Problems with Sparse Rewards

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:35.132039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:35.132039Z digest=sha256:04c28cc6f0390386c1d361f92bed59307bea0f2181cd84067d18fcfa9d38a4ad

Observation 3f31a89c-dfa3-46ad-a787-31beb912227c · outbound

This paper cites A comprehensive survey on safe reinforcement learning,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems A comprehensive survey on safe reinforcement learning,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:42.698606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:35.225247Z digest=sha256:81fc1eefdc53208e147f89ceb6ef30ce72862d29d0976b9fee6b790089f4c9d7

Observation 4a86cb93-5458-470a-acfa-5d08381dd836 · outbound

This paper cites Trial without Error: Towards Safe Reinforcement Learning via Human Intervention.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Trial without Error: Towards Safe Reinforcement Learning via Human Intervention

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:35.315253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:35.315253Z digest=sha256:7f037d932c8549ec7da433b630bf8966fded739f65792285c565080370eedcfc

Observation 2e800d42-c001-4948-8579-cb69a9048ba5 · outbound

This paper cites Policy gradient in continuous time,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Policy gradient in continuous time,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:42.482119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:35.459696Z digest=sha256:98621bbd2581e683683cedc50426737c020f750348112c86e4cb4b6b4ed9ecf5

Observation 12a82e21-f152-4421-bd38-9659339fcf81 · outbound

This paper cites Making deep Q-learning methods robust to time dis- cretization,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Making deep Q-learning methods robust to time dis- cretization,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:42.225469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:35.546997Z digest=sha256:dbb86cc7032897967b2e2aa2c47853c36105321efd58477893fb3faf553cf57b

Observation 9bf35128-1d6c-40f9-8dff-3e1a08431cdc · outbound

This paper cites Time discretization-invariant safe action repetition for policy gradient methods,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Time discretization-invariant safe action repetition for policy gradient methods,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:42.037479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:35.719240Z digest=sha256:bd863bf757603e889c045bb65c78f1c9d6f5c7092a00b49a1ceda5608f8bf334

Observation 6a3f0064-a959-44be-bc37-060185985191 · outbound

This paper cites Optimal scheduling of entropy regularizer for continuous-time linear-quadratic reinforcement learning,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Optimal scheduling of entropy regularizer for continuous-time linear-quadratic reinforcement learning,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:41.858896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:35.850763Z digest=sha256:c79fdbbf87e7bd80f1687512557e733c775a569d21491eb31aedfa57b83c8a10

Observation a8021655-0c66-493c-bb20-04aee29c2885 · outbound

This paper cites Indefinite stochastic riccati equations,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Indefinite stochastic riccati equations,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:41.714543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:36.049474Z digest=sha256:9cf30eff04ff45a9e5cf6fe1abd94e86c95e509141665091d442d04994e45f2b

Observation 7accc150-d4ad-4931-80bc-e33c5cef8d0a · outbound

This paper cites Existence of solutions to a class of indefinite stochastic riccati equations,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Existence of solutions to a class of indefinite stochastic riccati equations,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:41.550446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:36.115610Z digest=sha256:a0fbb65bf07283e227aeaf24bb9b753e7cda59842a733ea53e2d813afbb3a4df

Observation ccc0b981-160e-4ae4-9d31-cb9faf4049c9 · outbound

This paper cites Reinforcement learning in continuous time and space: A stochastic control approach,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Reinforcement learning in continuous time and space: A stochastic control approach,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:41.370950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:36.290103Z digest=sha256:5da504b63acae31a425e83d284b640c4f9f593808a09ee4c4df67e278e1f1422

Observation 8f088e2d-74bb-468c-8c19-ee4d41f8cfcc · outbound

This paper cites Exploration-exploitation trade-off for continuous-time episodic reinforcement learning with linear-convex models.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Exploration-exploitation trade-off for continuous-time episodic reinforcement learning with linear-convex models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:36.415315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:36.415315Z digest=sha256:2d9a8d1b61354a944c0eb0cc9258eb6e023cb5559e26ff156ca0e504817db97a

Observation dd631aed-b86c-4aaf-a250-0ecdc1143c55 · outbound

This paper cites Policy gradient and actor-critic learning in continuous time and space: Theory and algorithms,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Policy gradient and actor-critic learning in continuous time and space: Theory and algorithms,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:41.177508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:36.559039Z digest=sha256:2facc10c03d2444fbcbba33e31b156d76930714be755730988a6d01dbc71b502

Observation 99628674-cdfd-489c-8807-1efccfe2d68d · outbound

This paper cites Policy evaluation and temporal-difference learning in continuous time and space: A martingale approach,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Policy evaluation and temporal-difference learning in continuous time and space: A martingale approach,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:41.048499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:36.759306Z digest=sha256:fd345fcd3f7b79688bbf06585f86d5e83fc3b5fc68678089b02b82ffb556c2d5

Observation d2f9f109-02ae-4cbe-a0d1-57cad82d45d9 · outbound

This paper cites A stochastic approximation method,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems A stochastic approximation method,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:40.850349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:36.893967Z digest=sha256:fb1068c6975d03ea101e7542f8637b483418732bac118d15223266f00ee58f73

Observation bddc39d1-6174-4d91-ba31-6f58c63fe5d9 · outbound

This paper cites Stochastic approximation,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Stochastic approximation,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:40.619413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:37.102211Z digest=sha256:8f875fdad4bfaf364e4897f8fc2fcd9e8d216954793a71e152a2fc0fc815ffa4

Observation 52b5835a-44b2-4069-ac9a-03d5f5a4f485 · outbound

This paper cites an unresolved cited work.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:33:40.425098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:37.259548Z digest=sha256:c80238024d28604c33ac17d3631cb772321730bf3b4613b112fa3028e5312b82

Observation ddca0fb6-7710-483a-9b99-e3eade5a629b · outbound

This paper cites An overview of stochastic approximation,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems An overview of stochastic approximation,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:40.198447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:37.457228Z digest=sha256:f5f8996578790467eab6dd0de322644da505e6dd9cc8774418da55aa07346f15

Observation ac6aff8d-77e2-43be-a347-3e08a1c4bc21 · outbound

This paper cites A stochastic approximation algorithm with varying bounds,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems A stochastic approximation algorithm with varying bounds,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:39.861209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:37.676106Z digest=sha256:6722e69c02ba09be32c237054b13c36180585bc359bca7d5370396611780b816

Observation a05ae417-2113-4264-88f5-45e8b3888bc7 · outbound

This paper cites General bounds and finite-time improvement for the Kiefer-Wolfowitz stochastic approximation algorithm,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems General bounds and finite-time improvement for the Kiefer-Wolfowitz stochastic approximation algorithm,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:39.528730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:37.874175Z digest=sha256:cf3096f1532e215e923c44f133435b3750a24bbc84aead98b4f2969579606b01

Observation 3d5c3919-a8a0-4fce-80c2-6332c5d319a7 · outbound

This paper cites A convergence theorem for non negative almost supermartin- gales and some applications,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems A convergence theorem for non negative almost supermartin- gales and some applications,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:39.185660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:38.010679Z digest=sha256:1c1bf5a2cf7a9ded99685550da273b8cdf29e57bf0e21e1b916c1011c84dc4be

Observation 9b5aaec9-6d3e-4734-8818-9053c94d0707 · outbound

This paper cites an unresolved cited work.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:33:38.840537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:33:38.138033Z digest=sha256:be34f6b40e604762ae4164fe55f874f53ed989577e613d49ed2c6a1ee6763b56

Observation 8fcc4cb7-4208-41ba-83b1-a123ac4ebeb2 · outbound

This paper cites Regret of exploratory policy improvement and $q$-learning.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Regret of exploratory policy improvement and $q$-learning

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:38.321579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:38.321579Z digest=sha256:c0c6d9a8e94805a868fb8ca543c3978a9c8a27710a75fb716bf3a0b1c1e88a76

Pith citing papers

Observation ec665d63-6095-424a-ab99-5b351f53c8d3 · inbound

ART for Diffusion Sampling: A Reinforcement Learning Approach to Timestep Schedule cites this paper.

ART for Diffusion Sampling: A Reinforcement Learning Approach to Timestep Schedule Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-16T10:42:45.321033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T10:42:23.768313Z digest=sha256:8c680197403ffa60f25dd369e0868f50a6e940dc73693ff117573f3077db67b5

Observation 9818ebf3-6f70-41c9-8d27-84dac68bed2a · inbound

Amortized Guidance for Image Inpainting with Pretrained Diffusion Models cites this paper.

Amortized Guidance for Image Inpainting with Pretrained Diffusion Models Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T19:37:52.603822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-14T19:33:27.641167Z digest=sha256:1f38e1fa68b64fa27bd2a5777f4cb246a407b1bc5249a2af8f2ea780e8b6620d

Observation dd9a3f69-c1e9-47d8-9199-dfe392315231 · inbound

ART for Diffusion Sampling: Continuous-Time Control and Actor-Critic Learning cites this paper.

ART for Diffusion Sampling: Continuous-Time Control and Actor-Critic Learning Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-03T17:18:43.033802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-03T17:14:04.821073Z digest=sha256:ccfdfbf5b6744eeb41ad8970cf9c68f10aa906d495eafed2a52b12a4d90c0cdd

Observation 10076fa3-b771-462d-89e9-b6081230cc75 · inbound

ART for Diffusion Sampling: Continuous-Time Control and Actor-Critic Learning cites this paper.

ART for Diffusion Sampling: Continuous-Time Control and Actor-Critic Learning Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-12T08:25:48.021715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T08:25:48.021715Z digest=sha256:fd798d557c66a836fbceb75eb854c8a6d635c18aaaeaf51bd14eac0ba7ce0796

Observation 2a8504d5-6aa5-4220-9480-8210bd94ac96 · inbound

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies cites this paper.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.635669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.635669Z digest=sha256:6b60cb22b57d742bc25f310b818fede27ac64fbf706367e1ebc21c82a1ac2ef8