Pith. sign in

Paper Citation Record · LEDGER

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems

As of 19 August 2026, this Paper Citation Record lists 29 of 29 outbound references and 2 inbound Pith citation observations for arXiv:2209.07059.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2209.07059 v5

Coverage vector

measured 29 of 29 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-24T11:25:12.822765Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T15:52:37.070457Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T05:30:53.567417Z

Reference resolution

29 of 29 outbound references displayed

  • verified exact2
  • verified fuzzy21
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c21db4c5-ae56-4c38-b755-d15e2c0a2151 · outbound

This paper cites Second order elliptic equations and elliptic systems, volume 174 of Trans- lations of Mathematical Monographs.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Second order elliptic equations and elliptic systems, volume 174 of Trans- lations of Mathematical Monographs

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.217603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:fabd64a99fbee35eb978908743c4b7d162e1b9fec41d273c608ad3377f097378

Observation 9eb77077-ee9a-4560-ac47-0d357e567849 · outbound

This paper cites Learning equilibrium mean-variance strategy.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Learning equilibrium mean-variance strategy

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.153180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:8b26c7c2f77155928a9ff5d0bc9ab296f11ebb9a1bfbd24625dae1f52d01628c

Observation d94b767f-a519-42a9-ace3-fc64780a318a · outbound

This paper cites an unresolved cited work.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-05-24T11:26:09.163985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:5e85f9c387aacc3121298c007edb2fc76072405f050abf75ace7c7b6e76e1c19

Observation 394a5215-505d-431b-969b-a03390425b50 · outbound

This paper cites Exploratory LQG mean field games with entropy regularization.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Exploratory LQG mean field games with entropy regularization

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.240206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:b67f58cd73a4fc3558485c86586592ddbcc6671de8d552f539f0cd025343dafe

Observation 98050124-d32b-48e2-acad-658b994783a8 · outbound

This paper cites Taming the noise in reinforcement learning via soft updates.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Taming the noise in reinforcement learning via soft updates

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.246705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:2e718ecb5a68af2451bd8d4958e0b74e7bc0947ea0c5a400f6dd5efa47bb567c

Observation 3c103ba0-2e30-4dfc-bf0c-e4a2d8d2ebb2 · outbound

This paper cites Trudinger.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Trudinger

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.202064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:4e8e5f9300fcd276ff4bfdc91a8c4de2b73567049588b6d987aef91deb20ade0

Observation b0c7577d-06ce-4d79-9ff3-d08d2312b5e4 · outbound

This paper cites Entropy regularization for mean field games with learning.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Entropy regularization for mean field games with learning

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.146960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:b9e8163b2c4173841f37ebcad93f22220ea76fb1b420eb810c4ce6af82271092

Observation 2531fded-2d8f-40ba-818a-776c6ef5cdfb · outbound

This paper cites Reinforcement learning with deep energy-based policies.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Reinforcement learning with deep energy-based policies

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.232513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:c79be3902705157a2941ab4cf4fc3e41311d30f43a5aa70d5814aae28984037e

Observation 6d969779-70cf-402b-b603-9dc3e556422f · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.224900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:8c23da8ef828f19b596bd6f286ba7fa4ea63e3a8ad38eae026cedca3d0b49fb4

Observation 7ee46a17-1f7d-4d20-8550-129021c20016 · outbound

This paper cites Jacka and Aleksandar Mijatovi´ c.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Jacka and Aleksandar Mijatovi´ c

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.167727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:f6329b900687363f53c48d47487f43e4d31223db8f65d151c9b55785b7f192d6

Observation df8a86c2-e2bf-4f59-a05e-b4d533172a61 · outbound

This paper cites an unresolved cited work.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-05-24T11:26:09.236857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:1f0c84d930ef114e73aae85518c1c0530bee5c0603596dd2a2fbec669b5552e8

Observation 385fdfac-bede-4493-a528-e9c26c281144 · outbound

This paper cites an unresolved cited work.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-05-24T11:26:09.176502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:c4c61a121987ecb83bc2750e93d3c56d5ba2829e2461ca298f72ec685a2c7e05

Observation 4811f6d9-e3bd-4642-8d3c-0ad973a3c75b · outbound

This paper cites q-learning in continuous time.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems q-learning in continuous time

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.232890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:1a7138e793be1324de0f444b678c7cb26be3f064fb22ca92c8186a0ba450de04

Observation 94e53293-c1d9-43b9-aaa9-d3ef331ce335 · outbound

This paper cites an unresolved cited work.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-05-24T11:26:09.242907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:6d5c08fca82c1b1ed1b6194dd8dd65a16077085c353c73fbbd714e983b9377f3

Observation c105a17b-02b9-49df-96a5-eacf875b0f3c · outbound

This paper cites Kerimkulov, D.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Kerimkulov, D

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.236215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:27d6a8d6e1c75277f5444fdf24b9774f96acd160f1b02d0f537620be6e895934

Observation abb087c2-99ac-48ff-9021-700f50d94391 · outbound

This paper cites Exponential convergence and stability of Howard’s policy improvement algorithm for controlled diffusions.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Exponential convergence and stability of Howard’s policy improvement algorithm for controlled diffusions

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.220497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:791936c4eb180644ca1e3f25cc79b8f304d38cf4abd8b8ee55223217273c46b0

Observation cf0bef36-12f0-446c-a26d-b991c1564ab5 · outbound

This paper cites Policy iterations for reinforcement learning problems in continuous time and space—fundamental theory and methods.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Policy iterations for reinforcement learning problems in continuous time and space—fundamental theory and methods

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.186718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:c70e3ee8cf81ab43efab170d992dde0ce600f83e2664d788951eba3d11776a96

Observation 184c3cd6-3214-45db-a1a2-7f4fa125436b · outbound

This paper cites Value Iteration in Continuous Actions, States and Time.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Value Iteration in Continuous Actions, States and Time

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-24T11:26:09.005905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:079c7903c1fd21862a42975ab5b977e8de5cd1726677ebbbcc7a7f35a23a63f3

Observation f554acd8-3220-49de-8d23-195d447c70f1 · outbound

This paper cites Higher chain formula proved by combinatorics.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Higher chain formula proved by combinatorics

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.239670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:700c86cb6cac2b2fff479e162b11a77e3d0af8f53e12aff01b2d43fab6cb7c5e

Observation 8144c54c-9df6-4d9c-9a4d-10050f192096 · outbound

This paper cites an unresolved cited work.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-05-24T11:26:09.215950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:4272547a283e2010bad5da8bd7363ad7ebc5a1fb0c963d338dda10dadc360e07

Observation 556a6338-a26e-4640-aa46-ebcbd6355afc · outbound

This paper cites Regularity and stability of feedback relaxed controls.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Regularity and stability of feedback relaxed controls

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.213788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:302c0490b292f905ffff1a95f85d0bbfc922362d1186e16e645c28b95ebe870a

Observation 3b8c6086-367b-4fde-b741-f5e766181509 · outbound

This paper cites an unresolved cited work.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-05-24T11:26:09.206427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:9e138162ed812525818eb1e9ad29149d933d31720f1c72a27cf3789ea15861e4

Observation 3c9b46e0-4dc8-42bc-96db-63ed5a8c38f4 · outbound

This paper cites Policy iteration for the deterministic control problems -- a viscosity approach.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Policy iteration for the deterministic control problems -- a viscosity approach

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-24T11:26:08.991619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:4e3e7ae0c1e38f927a6c94f6457e440d8679e870da7cf08bf01151fa75ef25fa

Observation db841bae-6f80-44e9-8320-57b46e065a59 · outbound

This paper cites Exploratory hjb equations and their convergence.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Exploratory hjb equations and their convergence

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.229034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:84fd0144c96fe881ce1fdb95e6526938d612ffcf964280ec7f077194df51df7d

Observation 7282fa5a-4398-4db2-8cd0-e1b8769f71dd · outbound

This paper cites Continuous-time reinforcement learning control: A review of theoretical results, insights on performance, and needs for new designs.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Continuous-time reinforcement learning control: A review of theoretical results, insights on performance, and needs for new designs

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.160372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:36b6a61dfde83fbc931ad0c6cdf4ab78ccaa8656479ec7ac4e0add08d0819ba0

Observation a6e86ab9-f662-4571-b749-8bf30913d064 · outbound

This paper cites Reinforcement learning in continuous time and space: a stochastic control approach.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Reinforcement learning in continuous time and space: a stochastic control approach

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.243588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:03a0883f50155cb6b97fe46bf6ba3d131f84b7e8c60e5cf05b3e7f15329c1ddf

Observation 07ab9f59-b436-4839-b75f-6bbf4dbcb1b4 · outbound

This paper cites Continuous-time mean-variance portfolio selection: a reinforcement learning framework.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Continuous-time mean-variance portfolio selection: a reinforcement learning framework

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.211783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:e0dd52514ee1e578fbf31cf5289ca594b2520f66cd91382a0d76e1c0480190aa

Observation a9de8540-d1ad-4efa-b18c-664f5e3b8aa6 · outbound

This paper cites Ziebart, J.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Ziebart, J

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.224593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:9ffc0ba235937aecfb7a81edfad8d20eddb8adf63242113da52e0e4d338062d6

Observation 9011900a-b31c-4c85-a773-828d553a3769 · outbound

This paper cites Ziebart, Andrew L.

Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems Ziebart, Andrew L

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T11:26:09.207878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-24T11:25:12.822765Z digest=sha256:bd5370dca913a1b4e6106fa51505a09eb10de8b402f77e04fcbc39a061bcaf16

Pith citing papers

Observation 515f29ec-a540-42de-864a-95f813e05948 · inbound

Continuous-time optimal investment with portfolio constraints: a reinforcement learning approach cites this paper.

Continuous-time optimal investment with portfolio constraints: a reinforcement learning approach Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-11T15:52:37.070457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:52:37.070457Z digest=sha256:22e22a6e326a2f5de0ecc90a9464b3a127fca280625734a9f0634dd93b846654

Observation d2e2057b-20b1-4630-8632-f97c71617ab8 · inbound

Continuous Policy and Value Iteration for Stochastic Control Problems and Its Convergence cites this paper.

Continuous Policy and Value Iteration for Stochastic Control Problems and Its Convergence Convergence of Policy Iteration for Entropy-Regularized Stochastic Control Problems

Reference 8510

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:30:53.573835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:30:53.529461Z digest=sha256:1fb936bddbaf09dcdb58f05f5ca5fcfc9e519e7b29bbb222c546fbd5d2b0b24b