Pith. sign in

Paper Citation Record · LEDGER

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems

As of 13 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 5 inbound Pith citation observations for arXiv:2507.00358.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.00358 v2

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:33:38.321579Z

measured 54 of 54 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T11:22:20.635669Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T17:18:43.032059Z

Reference resolution

49 of 49 outbound references displayed

  • verified exact0
  • verified fuzzy34
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 53c011f9-07ed-4843-baeb-28990310a5df · outbound

This paper cites Sublinear Regret for a Class of Continuous-Time Linear-Quadratic Reinforcement Learning Problems.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Sublinear Regret for a Class of Continuous-Time Linear-Quadratic Reinforcement Learning Problems

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:32.491006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:32.491006Z digest=sha256:0514e3556c38bcf8c852cdb0bd1b7e7a3a61ac60875c5d3ca32e1738dd64f7a3

Observation a416a5af-1070-4a81-87d0-f759c5104c62 · outbound

This paper cites an unresolved cited work.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:33:50.572667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:32.598128Z digest=sha256:4012b04bbc518f9af1039ca683fd8410d6bb271599b11c50c274d8d4419479db

Observation 3319b20d-c984-487a-8f7e-7b9980ff7326 · outbound

This paper cites Yong and X.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Yong and X

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:50.346689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:32.682350Z digest=sha256:95c2d7d17358d37df4bfdd095641a1590df85a132de98a77a7b84ceb3b447883

Observation a0260570-6ce0-4368-bfca-79e54bd1e914 · outbound

This paper cites Stochastic linear quadratic regulators with indefinite control weight costs,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Stochastic linear quadratic regulators with indefinite control weight costs,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:50.100535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:32.785792Z digest=sha256:350e3d97bd4c3fa707c66931a907bc4a0b5ed0a33291397626e09dee4f40677c

Observation 9a56b895-bbdd-4b7b-9b39-ad595502bcb5 · outbound

This paper cites Well-posedness and attainability of indefinite stochastic linear quadratic control in infinite time horizon,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Well-posedness and attainability of indefinite stochastic linear quadratic control in infinite time horizon,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:49.849387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:32.875427Z digest=sha256:587680cf3f9394c42f60c6c694993abac31958f5d7804b088f08e66d94865666

Observation 56b4a220-ed46-4e8d-bf98-140e0303d25c · outbound

This paper cites Linear matrix inequalities, riccati equations, and indefinite stochastic linear quadratic controls,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Linear matrix inequalities, riccati equations, and indefinite stochastic linear quadratic controls,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:49.628482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:32.970981Z digest=sha256:538999d6c92cd0077ad4308c71854973532be69b713ac089addb63eb0d3f7be7

Observation 6a310869-e25f-46c5-b6e2-e8ea4743ab12 · outbound

This paper cites Solvability and asymptotic behavior of generalized riccati equations arising in indefinite stochastic lq controls,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Solvability and asymptotic behavior of generalized riccati equations arising in indefinite stochastic lq controls,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:49.359104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:33.097423Z digest=sha256:cceafb2a1c0d9d9e6116dccd43e924bb24eda4b141a78231592c4043189a5738

Observation 5304aed6-de98-4bb2-9c9f-dd2410cfa39f · outbound

This paper cites A primal-dual semi-definite programming approach to linear quadratic control,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems A primal-dual semi-definite programming approach to linear quadratic control,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:47.792609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:33.187398Z digest=sha256:4198d8a408f714ea32af9ba478161239387b5bdf8dc2d3472d16cbf888b12fe0

Observation c4910c53-0ea6-4635-b660-f8dc3fdc3498 · outbound

This paper cites Stabilization control for itˆ o stochastic system with indefinite state and control weight costs,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Stabilization control for itˆ o stochastic system with indefinite state and control weight costs,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:47.001170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:33.296546Z digest=sha256:f04c4d7337c434b65eff824c07c720e2208c7a00bf3571f579c37649a351936b

Observation 0cc3dcd9-378d-4d3e-a63d-a3fe1fe2e24e · outbound

This paper cites Optimal regulators for a class of nonlinear stochastic systems,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Optimal regulators for a class of nonlinear stochastic systems,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:46.604066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:33.397836Z digest=sha256:c6341a601a889abdedec12e4ba4286af0595759fdb3e5353eb96305864290392

Observation 224c8670-cdc7-48ce-bd8a-0e48d8985ff7 · outbound

This paper cites Indefinite linear-quadratic optimal control of mean-field stochas- tic differential equation with jump diffusion: an equivalent cost functional method,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Indefinite linear-quadratic optimal control of mean-field stochas- tic differential equation with jump diffusion: an equivalent cost functional method,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:46.262692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:33.501366Z digest=sha256:0230e93114f80ac0de53f2e8f5512ce13b84d43ab45200c392da521bebc18357

Observation 9ffde277-e768-4d2b-a6fa-a3f77730792b · outbound

This paper cites Stochastic linear quadratic optimal control problems with regime- switching jumps in infinite horizon,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Stochastic linear quadratic optimal control problems with regime- switching jumps in infinite horizon,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:45.935693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:33.595650Z digest=sha256:e9af485518050caac4780a7950d4990d5bebb034bfaec51859cc50267fed6b65

Observation c30d62f0-e968-4f7c-925b-97b18f70c6c6 · outbound

This paper cites On estimating the expected return on the market: An exploratory investiga- tion,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems On estimating the expected return on the market: An exploratory investiga- tion,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:45.607054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:33.693784Z digest=sha256:25975e91734aef52e43364b2c486e0b977388c68faac09095a9ef2b92e262b6e

Observation 22553b20-340f-4dc7-840b-ee89e39fa168 · outbound

This paper cites an unresolved cited work.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:33:45.269404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:33.776096Z digest=sha256:c265e00a782cf7497ae27afffa8e5d128a62f6428133a0f833b98ff14f3867f6

Observation 7081615b-f0b1-426c-86b4-783149d67788 · outbound

This paper cites Rustem and M.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Rustem and M

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:45.008450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:33.870263Z digest=sha256:4e034c71c2e2d2728dda7b4aaf4bf35653004173776c40b98d57bc9ffc82b337

Observation 37632502-7169-4d11-845d-733f8f20b593 · outbound

This paper cites an unresolved cited work.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:33:44.729693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:33.944653Z digest=sha256:2d8c9204f2f51cce334f3ef3be04a262a0ac93a730575a1baad61415723bad25

Observation c75ffc6c-70b2-4f62-ae92-8ac9ea4ad5bf · outbound

This paper cites A survey on intrinsic motivation in reinforcement learning.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems A survey on intrinsic motivation in reinforcement learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:34.035769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:34.035769Z digest=sha256:7e8794f1630a1f005aa2d6dbc7b25350ebe529b62dd6277e40b5e9c73d36e45e

Observation 5ec4e6e4-16a0-4143-a74f-b77f89cebd07 · outbound

This paper cites Formal theory of creativity, fun, and intrinsic motivation (1990–2010),.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Formal theory of creativity, fun, and intrinsic motivation (1990–2010),

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:44.443994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:34.126427Z digest=sha256:a00dd7d53208a9edccf6db21a9e13622b175cfaf959727ff54e86c89f22a1bde

Observation 86e125e1-8cb6-47dd-acbd-2a38acffa9ff · outbound

This paper cites Surprise-Based Intrinsic Motivation for Deep Reinforcement Learning.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Surprise-Based Intrinsic Motivation for Deep Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:34.257469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:34.257469Z digest=sha256:3a628ff05e1d6dda568b3a31317323deae08ba8ca09808ba478de84d21812df8

Observation 347fd813-d80b-41d9-a067-a87eebf867ef · outbound

This paper cites Curiosity-driven exploration by self- supervised prediction,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Curiosity-driven exploration by self- supervised prediction,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:44.157257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:34.353252Z digest=sha256:6e7637341ef65ebd2e75095160dd309fa3ff98b8dffaad7e58dd4a550180cc36

Observation 84763f26-e404-4299-83e2-4f9271c8ff3d · outbound

This paper cites Large-Scale Study of Curiosity-Driven Learning.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Large-Scale Study of Curiosity-Driven Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:34.448173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:34.448173Z digest=sha256:7a04269efba3c15d1c2e943e1c2ebb9c8d2f08ebaf778c3ea6e2b0cb5be3b571

Observation 10d21d05-a695-4b4e-9f32-1cfd6bd8fa23 · outbound

This paper cites Exploration by Random Network Distillation.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Exploration by Random Network Distillation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:34.568573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:34.568573Z digest=sha256:67e466b5770963fb828e9a3d63fff85d31dcb975396b084bbacc7a1c16dc84dc

Observation d9708eae-0a06-4e27-8b0a-39fdd77d889f · outbound

This paper cites Randomized prior functions for deep reinforcement learning,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Randomized prior functions for deep reinforcement learning,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:43.818381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:34.642461Z digest=sha256:168f47e0c64d9a4f8d3255dd2510b663930c3e3b97aa2705d04b2b97d07b2f9a

Observation bc9ef8ca-a6c4-49c3-b588-e076d4af1322 · outbound

This paper cites Fast active learning for pure exploration in reinforcement learning,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Fast active learning for pure exploration in reinforcement learning,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:43.572233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:34.727222Z digest=sha256:6df861471a5daf0979cc121ba1b791409b915562ff060bbca1e03d14395d75b4

Observation e0d6309e-1719-4bf5-a7d2-7193a8dae6a2 · outbound

This paper cites # exploration: A study of count-based exploration for deep reinforcement learning,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems # exploration: A study of count-based exploration for deep reinforcement learning,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:43.310875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:34.830998Z digest=sha256:5754123196279479d1b56c6e8d4b5d820fc97285ab7cabbd6676b93a162b322c

Observation 16a5722b-ccca-4208-bb58-b7412a6a7e36 · outbound

This paper cites Go-Explore: a New Approach for Hard-Exploration Problems.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Go-Explore: a New Approach for Hard-Exploration Problems

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:34.943251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:34.943251Z digest=sha256:d792624b97357d2fecc16ee9b75a514ffe3049bd4c226167313c48e4f8039221

Observation d28035f0-aadb-4915-83d3-90f2e7e7cc3d · outbound

This paper cites First return, then explore,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems First return, then explore,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:43.039615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:35.051690Z digest=sha256:58a442cde2752f25abc6329b981c4f6dae3bbf8a568d1f1643ce1a40b67aa649

Observation b25f14d2-0ef2-4fc4-98fc-5b3db13d6108 · outbound

This paper cites Leveraging Demonstrations for Deep Reinforcement Learning on Robotics Problems with Sparse Rewards.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Leveraging Demonstrations for Deep Reinforcement Learning on Robotics Problems with Sparse Rewards

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:35.132039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:35.132039Z digest=sha256:6b19fe90de84e83e9ef18b94801a5e873bd11b289d346e4ea5a6f2c4b867392d

Observation 3f31a89c-dfa3-46ad-a787-31beb912227c · outbound

This paper cites A comprehensive survey on safe reinforcement learning,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems A comprehensive survey on safe reinforcement learning,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:42.698606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:35.225247Z digest=sha256:a2a8102a3fef794ec541a4d95c129f3b19774c911cc4dd15c6877f9ccca99079

Observation 4a86cb93-5458-470a-acfa-5d08381dd836 · outbound

This paper cites Trial without Error: Towards Safe Reinforcement Learning via Human Intervention.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Trial without Error: Towards Safe Reinforcement Learning via Human Intervention

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:35.315253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:35.315253Z digest=sha256:d1e080e2ed5b8fd74fce6f9899782d6b758ee7f4091da20503864feb67d4e445

Observation 2e800d42-c001-4948-8579-cb69a9048ba5 · outbound

This paper cites Policy gradient in continuous time,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Policy gradient in continuous time,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:42.482119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:35.459696Z digest=sha256:e43d8d1f77ee161e11b2f5d515322bbf6e5fec119ed79aacc6b63ed8b28dd6a6

Observation 12a82e21-f152-4421-bd38-9659339fcf81 · outbound

This paper cites Making deep Q-learning methods robust to time dis- cretization,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Making deep Q-learning methods robust to time dis- cretization,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:42.225469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:35.546997Z digest=sha256:6ef2af195627408a7c69ad20effaf08d6449c3ecc2d230e9c8ed09579cc57ac9

Observation 9bf35128-1d6c-40f9-8dff-3e1a08431cdc · outbound

This paper cites Time discretization-invariant safe action repetition for policy gradient methods,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Time discretization-invariant safe action repetition for policy gradient methods,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:42.037479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:35.719240Z digest=sha256:71c97690183fdb3395d582281f302aabfd8485a2abb1ed3415dbfb73d1090aad

Observation 6a3f0064-a959-44be-bc37-060185985191 · outbound

This paper cites Optimal scheduling of entropy regularizer for continuous-time linear-quadratic reinforcement learning,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Optimal scheduling of entropy regularizer for continuous-time linear-quadratic reinforcement learning,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:41.858896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:35.850763Z digest=sha256:c80a3629b613e24809cd8018e731b9d5cd5735c2c4196125df210f4f9af20dac

Observation a8021655-0c66-493c-bb20-04aee29c2885 · outbound

This paper cites Indefinite stochastic riccati equations,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Indefinite stochastic riccati equations,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:41.714543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:36.049474Z digest=sha256:06d0586eb3e9d6d556af1da0af02dccafe72ed1ff9f6fc13311f074a4823a960

Observation 7accc150-d4ad-4931-80bc-e33c5cef8d0a · outbound

This paper cites Existence of solutions to a class of indefinite stochastic riccati equations,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Existence of solutions to a class of indefinite stochastic riccati equations,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:41.550446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:36.115610Z digest=sha256:8871b9bda23cdfe95bcc9b4374b6b543e192cb27a90e3297ad3c4a285cd613bd

Observation ccc0b981-160e-4ae4-9d31-cb9faf4049c9 · outbound

This paper cites Reinforcement learning in continuous time and space: A stochastic control approach,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Reinforcement learning in continuous time and space: A stochastic control approach,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:41.370950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:36.290103Z digest=sha256:7735dcc223fb32fd2c71dc366a2eb554c0c24f73240bf609eb6126cda428e6fe

Observation 8f088e2d-74bb-468c-8c19-ee4d41f8cfcc · outbound

This paper cites Exploration-exploitation trade-off for continuous-time episodic reinforcement learning with linear-convex models.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Exploration-exploitation trade-off for continuous-time episodic reinforcement learning with linear-convex models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:36.415315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:36.415315Z digest=sha256:60a8fa07008298184e03db04a1fcff6ceedabbf12a8e1d79a4a2d7d14c563f5d

Observation dd631aed-b86c-4aaf-a250-0ecdc1143c55 · outbound

This paper cites Policy gradient and actor-critic learning in continuous time and space: Theory and algorithms,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Policy gradient and actor-critic learning in continuous time and space: Theory and algorithms,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:41.177508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:36.559039Z digest=sha256:b270b1cd3ff66de14fcdb7111f7fa4ddeb450ccd0e2426d6784bb89f6814b673

Observation 99628674-cdfd-489c-8807-1efccfe2d68d · outbound

This paper cites Policy evaluation and temporal-difference learning in continuous time and space: A martingale approach,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Policy evaluation and temporal-difference learning in continuous time and space: A martingale approach,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:41.048499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:36.759306Z digest=sha256:6ca935bcbdf46bcb5f11b2e2c2d30e1d98d9ba882976011052f7d15f1ea65ea1

Observation d2f9f109-02ae-4cbe-a0d1-57cad82d45d9 · outbound

This paper cites A stochastic approximation method,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems A stochastic approximation method,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:40.850349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:36.893967Z digest=sha256:cc663c17f6f76f8a5cfd2297950051f11ed089320178fe748e4ad7b5cafca5d4

Observation bddc39d1-6174-4d91-ba31-6f58c63fe5d9 · outbound

This paper cites Stochastic approximation,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Stochastic approximation,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:40.619413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:37.102211Z digest=sha256:b474d413a2416ed6b724c6011ebbbb6d704741c78c6dd4fd9d9b72c9c86ea8d1

Observation 52b5835a-44b2-4069-ac9a-03d5f5a4f485 · outbound

This paper cites an unresolved cited work.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:33:40.425098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:37.259548Z digest=sha256:0ba9e065527376c949078e2a366fe5d4a0f4a23388f4fe3d9e04d4887ba51a81

Observation ddca0fb6-7710-483a-9b99-e3eade5a629b · outbound

This paper cites An overview of stochastic approximation,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems An overview of stochastic approximation,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:40.198447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:37.457228Z digest=sha256:6ae93eb6c1e9964ca48f4d98a8e474e966104d0082820ac84bbc4911d27c7a88

Observation ac6aff8d-77e2-43be-a347-3e08a1c4bc21 · outbound

This paper cites A stochastic approximation algorithm with varying bounds,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems A stochastic approximation algorithm with varying bounds,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:39.861209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:37.676106Z digest=sha256:a262034a7b696414741e4c4a86ae997c8ffe174ce356b79e346abe10b00ca86b

Observation a05ae417-2113-4264-88f5-45e8b3888bc7 · outbound

This paper cites General bounds and finite-time improvement for the Kiefer-Wolfowitz stochastic approximation algorithm,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems General bounds and finite-time improvement for the Kiefer-Wolfowitz stochastic approximation algorithm,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:39.528730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:37.874175Z digest=sha256:e94029a011b8799f107c6c980a268e8d8dc88c0733733aef316413f7e8ddad91

Observation 3d5c3919-a8a0-4fce-80c2-6332c5d319a7 · outbound

This paper cites A convergence theorem for non negative almost supermartin- gales and some applications,.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems A convergence theorem for non negative almost supermartin- gales and some applications,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:33:39.185660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:38.010679Z digest=sha256:546ba045d1e5e491fbe51781fdf06def78c1962cae757c548ec4c502a0758229

Observation 9b5aaec9-6d3e-4734-8818-9053c94d0707 · outbound

This paper cites an unresolved cited work.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:33:38.840537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T21:33:38.138033Z digest=sha256:dcf99300f29c8fb1a169977f5d39ca5b2590931ee42c6d013940b9eb59843a5e

Observation 8fcc4cb7-4208-41ba-83b1-a123ac4ebeb2 · outbound

This paper cites Regret of exploratory policy improvement and $q$-learning.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Regret of exploratory policy improvement and $q$-learning

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:38.321579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:38.321579Z digest=sha256:61fa1527eaf1b917c9dddf1057e0c7a1d583bdebd7df81da5866b81c6a35bd73

Pith citing papers

Observation ec665d63-6095-424a-ab99-5b351f53c8d3 · inbound

ART for Diffusion Sampling: A Reinforcement Learning Approach to Timestep Schedule cites this paper.

ART for Diffusion Sampling: A Reinforcement Learning Approach to Timestep Schedule Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-16T10:42:45.321033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T10:42:23.768313Z digest=sha256:aa1a00d57d83b893e858d46b906e5d603dd2b2577cf9c162a3b9dd76b8ad4b46

Observation 9818ebf3-6f70-41c9-8d27-84dac68bed2a · inbound

Amortized Guidance for Image Inpainting with Pretrained Diffusion Models cites this paper.

Amortized Guidance for Image Inpainting with Pretrained Diffusion Models Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T19:37:52.603822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-14T19:33:27.641167Z digest=sha256:a287e4592a946598a7bc013d6a507132979cf1556746c96ca599b237445515bf

Observation dd9a3f69-c1e9-47d8-9199-dfe392315231 · inbound

ART for Diffusion Sampling: Continuous-Time Control and Actor-Critic Learning cites this paper.

ART for Diffusion Sampling: Continuous-Time Control and Actor-Critic Learning Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-03T17:18:43.033802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-07-03T17:14:04.821073Z digest=sha256:67c725617fbb730b7e1fb212078e4b1e21252e77bb9cf4416f3a29be4380c1bf

Observation 10076fa3-b771-462d-89e9-b6081230cc75 · inbound

ART for Diffusion Sampling: Continuous-Time Control and Actor-Critic Learning cites this paper.

ART for Diffusion Sampling: Continuous-Time Control and Actor-Critic Learning Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-12T08:25:48.021715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T08:25:48.021715Z digest=sha256:273753f89dfd34b98cae8111311f3d2cd0aef8ac0f7321286e9778241a7cd335

Observation 2a8504d5-6aa5-4220-9480-8210bd94ac96 · inbound

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies cites this paper.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.635669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.635669Z digest=sha256:8560b293f867ab4a5e5c3140aa5ccacc192707d2b47705dc8853cad9417bed64