Pith. sign in

Paper Citation Record · LEDGER

Operator Splitting for Convex Constrained Markov Decision Processes

As of 17 August 2026, this Paper Citation Record lists 62 of 62 outbound references and 1 inbound Pith citation observation for arXiv:2412.14002.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.14002 v1

Coverage vector

measured 62 of 62 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T12:40:45.727688Z

measured 63 of 63 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-01T01:40:08.230114Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T12:55:44.116744Z

Reference resolution

62 of 62 outbound references displayed

  • verified exact4
  • verified fuzzy42
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7d5debb6-f0c9-42d9-8d71-e9b5c137e1ba · outbound

This paper cites Mastering the game of go without human knowledge,.

Operator Splitting for Convex Constrained Markov Decision Processes Mastering the game of go without human knowledge,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.474791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.523266Z digest=sha256:94acd880fa39a2391662efd23ed6707040f7920a67bd7f24a2172729a230c7ba

Observation a14a017c-7b4c-46bb-93e6-842395d858bc · outbound

This paper cites Magnetic control of tokamak plasmas through deep reinforcement learning,.

Operator Splitting for Convex Constrained Markov Decision Processes Magnetic control of tokamak plasmas through deep reinforcement learning,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.463803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.527737Z digest=sha256:8043c38382507e6018d3b023902185a6472a237f7367ebed2406711399ae5d17

Observation fc2e3662-6eba-4b80-8a36-fd86e8f01db8 · outbound

This paper cites an unresolved cited work.

Operator Splitting for Convex Constrained Markov Decision Processes Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.531734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.531734Z digest=sha256:d6291feaa9bccd9f0891bc8d07ee303afca5494b88cf4b6b8622d05a94ebfb6a

Observation d5ac92ef-5561-4ba3-9ad0-9402d472a519 · outbound

This paper cites Altman, Constrained Markov decision processes.

Operator Splitting for Convex Constrained Markov Decision Processes Altman, Constrained Markov decision processes

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.535342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.535342Z digest=sha256:9ebf112b039ac0b403c1da51d6d20c961c7af075bc21d31f4e629d91bf437e2e

Observation 5a515a8e-f66a-489c-a58b-7e289c4e5f3f · outbound

This paper cites Policy gradients with variance related risk criteria,.

Operator Splitting for Convex Constrained Markov Decision Processes Policy gradients with variance related risk criteria,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.438344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.539032Z digest=sha256:3db765826e6eb8c9361d58c8576149d1fdf802828fd06eaa7c0b6941c9ad1680

Observation 9ec6acb3-8b02-4766-91c3-e51e28f664db · outbound

This paper cites Risk-constrained reinforcement learning with percentile risk criteria,.

Operator Splitting for Convex Constrained Markov Decision Processes Risk-constrained reinforcement learning with percentile risk criteria,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.426469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.542889Z digest=sha256:652af2454feb1c587a3465d1cbd59e81e0d49e447640d9b45c8b570a06679264

Observation ee91af8a-5923-494f-a011-6aca5f6ab9d7 · outbound

This paper cites Control and optimization meet the smart power grid: Scheduling of power demands for optimal energy management,.

Operator Splitting for Convex Constrained Markov Decision Processes Control and optimization meet the smart power grid: Scheduling of power demands for optimal energy management,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.416377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.546898Z digest=sha256:6563a574aa4aade8f11e0dcd59a18ac9df0c060b2d11015b8b7d431bf01e91a6

Observation 542b87d3-b08d-435d-be6e-0e07797a0d49 · outbound

This paper cites Constrained policy optimization,.

Operator Splitting for Convex Constrained Markov Decision Processes Constrained policy optimization,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.406379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.550549Z digest=sha256:96ca076d072da127651fa54f8c56171b330991ae3b57af32f70e69f316b9dbff

Observation 45b38f5c-3b29-440f-81cb-573b57a0950d · outbound

This paper cites Dynamic programming equations for dis- counted constrained stochastic control,.

Operator Splitting for Convex Constrained Markov Decision Processes Dynamic programming equations for dis- counted constrained stochastic control,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.396802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.553944Z digest=sha256:0af2dbccce1c1c9e68e6732aee4f996b7db9e2a3857f2c9c4f6c99a7b2ecf85e

Observation ad5d1b4d-e489-46fd-b2bf-0ace75ad2016 · outbound

This paper cites Dynamic programming in constrained Markov decision processes,.

Operator Splitting for Convex Constrained Markov Decision Processes Dynamic programming in constrained Markov decision processes,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.387474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.557316Z digest=sha256:38d93220f55d82e112664e0753f4c256a923660faf44cdf85eca9539992b1377

Observation cd8afd8c-185d-45f4-bb8d-cdd29e24ae5f · outbound

This paper cites A Gradient-Aware Search Algorithm for Constrained Markov Decision Processes.

Operator Splitting for Convex Constrained Markov Decision Processes A Gradient-Aware Search Algorithm for Constrained Markov Decision Processes

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-08-11T12:40:45.942537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.560934Z digest=sha256:e2031838a1304314108ac6e546ef05ca4d9f1770e65c209353932114e253ea26

Observation 81bb3222-c4a7-4975-9bfa-60d1a820a59d · outbound

This paper cites Natural policy gradient primal-dual method for constrained Markov decision processes,.

Operator Splitting for Convex Constrained Markov Decision Processes Natural policy gradient primal-dual method for constrained Markov decision processes,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.376767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.564782Z digest=sha256:1354965cfa71cd7c54aa72240f2f209cd38639bcb2d2f19ef0b49c0e68c2c197

Observation a42db668-8b00-4eca-88ab-0631211d5210 · outbound

This paper cites Learning policies with zero or bounded constraint violation for constrained MDPs,.

Operator Splitting for Convex Constrained Markov Decision Processes Learning policies with zero or bounded constraint violation for constrained MDPs,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.366953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.568209Z digest=sha256:6e6bc406b8439b64bcab222ee71e30c9a4e81681fbd6a21c8eb58f1eb12e8573

Observation e5e4fd73-681f-4d25-8067-da8a5f2ce359 · outbound

This paper cites State Augmented Constrained Reinforcement Learning: Overcoming the Limitations of Learning with Rewards.

Operator Splitting for Convex Constrained Markov Decision Processes State Augmented Constrained Reinforcement Learning: Overcoming the Limitations of Learning with Rewards

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-08-11T12:40:45.928287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.571142Z digest=sha256:4fe85924346ba43e09bb300359da84d2f41546985f3b5d23cfd2eff158d22488

Observation 059a8990-aadc-44e1-b9e0-bbe49e2e94f9 · outbound

This paper cites Constrained MDPs and the reward hypothesis.

Operator Splitting for Convex Constrained Markov Decision Processes Constrained MDPs and the reward hypothesis

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.357018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.574448Z digest=sha256:f5b5af467ff7c7937a4afc3f92ba345a9bab7eb8474742568c01dd1aaff561db

Observation dc886c82-2731-4f9d-a3a9-9c3e88d9d4ec · outbound

This paper cites Two “well-known.

Operator Splitting for Convex Constrained Markov Decision Processes Two “well-known

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.346626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.576948Z digest=sha256:53fbf115f03061a17e7aca0ec9411ab113f0c7a96dbb5a61c21ec5f428b7a255

Observation 94358415-df0b-4c2c-8470-c8a9c5a99ca8 · outbound

This paper cites Algorithm for constrained Markov decision process with linear convergence,.

Operator Splitting for Convex Constrained Markov Decision Processes Algorithm for constrained Markov decision process with linear convergence,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.336444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.579723Z digest=sha256:aec432b02f1ce1d591339aba64988a93b7e34fc7252682865705f7662309b10e

Observation df37e178-a510-4759-9e23-c8ca8325bc49 · outbound

This paper cites Faster Algorithm and Sharper Analysis for Constrained Markov Decision Process.

Operator Splitting for Convex Constrained Markov Decision Processes Faster Algorithm and Sharper Analysis for Constrained Markov Decision Process

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-11T12:40:45.912662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.582528Z digest=sha256:4149421882742c363ee05f9a2d094cd5dc2b08270450d2202bd2ece5ae88d8fa

Observation 436de9d0-9ff3-4e8f-8b7a-e28e86f689d4 · outbound

This paper cites Cancellation-Free Regret Bounds for Lagrangian Approaches in Constrained Markov Decision Processes.

Operator Splitting for Convex Constrained Markov Decision Processes Cancellation-Free Regret Bounds for Lagrangian Approaches in Constrained Markov Decision Processes

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.585747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.585747Z digest=sha256:200d0a1c467f98c7aede5be304a03d8eeeceb2dba29871f9217d17cb3d8c6c6b

Observation e7dc5fbf-0b77-45e9-bc34-8fb93dbf9c2d · outbound

This paper cites Last-Iterate Convergent Policy Gradient Primal-Dual Methods for Constrained MDPs.

Operator Splitting for Convex Constrained Markov Decision Processes Last-Iterate Convergent Policy Gradient Primal-Dual Methods for Constrained MDPs

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.589096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.589096Z digest=sha256:17ff7632e59303148436bd2ed90937948a1d1c0f4208e833a4885e07c148eee8

Observation d9b00ec9-7e4a-45d9-b6c9-35dac966afef · outbound

This paper cites Reload: Reinforcement learning with optimistic ascent- descent for last-iterate convergence in constrained MDPs,.

Operator Splitting for Convex Constrained Markov Decision Processes Reload: Reinforcement learning with optimistic ascent- descent for last-iterate convergence in constrained MDPs,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.324278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.592853Z digest=sha256:cc1c1fcb1ca15f307382e8bdb2e1bf47856dd58726915450200b9e00184b2d9a

Observation 15e44324-8a48-4e21-aec3-1d4ff6d76fea · outbound

This paper cites Ipo: Interior-point policy optimization under constraints,.

Operator Splitting for Convex Constrained Markov Decision Processes Ipo: Interior-point policy optimization under constraints,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.313087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.596272Z digest=sha256:44d68c9ce595bbd59ccd4706e4b6c0585858ab51d5acc2fd6c60c8404fed86b3

Observation d9ca5f64-4792-4ad5-8b81-cca65e75d970 · outbound

This paper cites Projection-Based Constrained Policy Optimization.

Operator Splitting for Convex Constrained Markov Decision Processes Projection-Based Constrained Policy Optimization

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.599554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.599554Z digest=sha256:0e1ce4ac802d2c1d6aab3705ca1cad781a4aad33ab8883426893fa8f15206e9b

Observation 0bc2fc3e-f15d-42a0-b7ac-9ed2ce28805e · outbound

This paper cites Reward is enough for convex MDPs,.

Operator Splitting for Convex Constrained Markov Decision Processes Reward is enough for convex MDPs,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.302622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.604198Z digest=sha256:1942ed06ce1b1bb56523d3815029c11550a8a4e1e29904bd92e0eb7fbef71430

Observation 3b81aac5-0a43-47d3-8586-bddbc21ce90a · outbound

This paper cites Apprenticeship learning via inverse reinforce- ment learning,.

Operator Splitting for Convex Constrained Markov Decision Processes Apprenticeship learning via inverse reinforce- ment learning,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.292322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.607868Z digest=sha256:4f3d654273cc700389bfe92fab4e8feb41b0d593fcc54352832db838431c1dbe

Observation 41a27b59-5d10-424e-9af1-3f570162db0b · outbound

This paper cites Provably efficient maximum entropy exploration,.

Operator Splitting for Convex Constrained Markov Decision Processes Provably efficient maximum entropy exploration,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.281858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.611007Z digest=sha256:e8b28e5385acd570bb956015294184f91b63dec01aea7182f1d005ddac7d0810

Observation b9945c27-18a9-4bb5-8a98-148d484fa784 · outbound

This paper cites Diversity is All You Need: Learning Skills without a Reward Function.

Operator Splitting for Convex Constrained Markov Decision Processes Diversity is All You Need: Learning Skills without a Reward Function

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.614529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.614529Z digest=sha256:e9847def367aedee98e0d3a9cfe0ca48b811d1805f8fa3b4ccd1ca1fdfd6c238

Observation a405dd85-26bc-424e-9b46-9cece872a9ad · outbound

This paper cites Policy-based primal-dual methods for convex constrained Markov decision processes,.

Operator Splitting for Convex Constrained Markov Decision Processes Policy-based primal-dual methods for convex constrained Markov decision processes,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.271500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.618422Z digest=sha256:7a3a2a7bf61cd9b29aa08c2e5be6531b6d294b7707a26edf5ce1ad7178b71e5a

Observation 5d238fcc-519d-47b8-97fe-741bdfe0b226 · outbound

This paper cites Variational policy gradient method for reinforcement learning with general utilities,.

Operator Splitting for Convex Constrained Markov Decision Processes Variational policy gradient method for reinforcement learning with general utilities,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.260720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.621795Z digest=sha256:dc99a569f33e332d2a5e27a158bcba9c5b3c84a7af4fddd54db590fb92232cc4

Observation 5c4a64fb-80a3-4e21-babd-fb34144bf916 · outbound

This paper cites Reinforcement learning with convex constraints,.

Operator Splitting for Convex Constrained Markov Decision Processes Reinforcement learning with convex constraints,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.249377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.625100Z digest=sha256:3daf2e582cd856fd302cedb987d9e3f72d872f837d248a9a47c4a3a944ce6467

Observation 6452146a-e617-4068-b5a7-af6bf91dbf54 · outbound

This paper cites A simple reward-free approach to constrained reinforcement learning,.

Operator Splitting for Convex Constrained Markov Decision Processes A simple reward-free approach to constrained reinforcement learning,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.238752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.628643Z digest=sha256:ef406bf76b4a7156bdc3cbc78b37d475c23d4dc7610fb0e6327fc4d786112a50

Observation 002b4718-6274-4d5d-8a35-5130c8ea7d9c · outbound

This paper cites Bauschke and P.

Operator Splitting for Convex Constrained Markov Decision Processes Bauschke and P

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.227584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.632070Z digest=sha256:eb033b626623c353e9691a0124a19ed7b3adcd47c58b1ed509371b730e598c20

Observation aae25587-1f49-4ccd-bdf5-ae98d45aed63 · outbound

This paper cites Distributed optimization and statistical learning via the alternating direction method of multipliers,.

Operator Splitting for Convex Constrained Markov Decision Processes Distributed optimization and statistical learning via the alternating direction method of multipliers,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.635420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.635420Z digest=sha256:33da43edf795ba4fd20a70c7330568ab53f0b1502cf6222d3ab6296d28a37ddb

Observation 8694d196-c0b1-433a-91ec-0f717a1e6517 · outbound

This paper cites A note on the equivalence of operator splitting methods,.

Operator Splitting for Convex Constrained Markov Decision Processes A note on the equivalence of operator splitting methods,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.210336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.638587Z digest=sha256:ae8caed8d4e2dc46199b2c00dcdf09c80d14049fcb0423f8df40f30c86beb9a4

Observation 729dc1db-fde2-477b-9cc8-c21276720462 · outbound

This paper cites Provably efficient algorithms for multi-objective competitive RL,.

Operator Splitting for Convex Constrained Markov Decision Processes Provably efficient algorithms for multi-objective competitive RL,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.198587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.641928Z digest=sha256:cbf10b4c807a2b7ccf59535d6de19cf0a4238e3e1ad8f2d2536c3e793dd30dd3

Observation 148017e3-2317-4d51-86ef-164c340ffebb · outbound

This paper cites A splitting method for optimal control,.

Operator Splitting for Convex Constrained Markov Decision Processes A splitting method for optimal control,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.187573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.645327Z digest=sha256:9183e98d3e37c8c1b0b139a85449f3b8f0686f471af6584a72c1c62e0d7c7428

Observation 335a47a6-b86e-4328-95d5-b970d8ec748d · outbound

This paper cites A unified view of entropy-regularized Markov decision processes.

Operator Splitting for Convex Constrained Markov Decision Processes A unified view of entropy-regularized Markov decision processes

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.648460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.648460Z digest=sha256:e582f099d22d0672c457f3effa7640fd6196dd2cfff358469dad3251e53445e2

Observation 34500308-5375-46e9-a658-cf89b33ce04e · outbound

This paper cites A theory of regularized Markov decision processes,.

Operator Splitting for Convex Constrained Markov Decision Processes A theory of regularized Markov decision processes,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.176808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.651895Z digest=sha256:f2e354f57177e3bd44fe09949792cc984c0eb7e28467591c7d8f8f342144de44

Observation 47314e45-c403-49b2-9885-921c406adbef · outbound

This paper cites Dynamic programming through the lens of semismooth Newton-type methods,.

Operator Splitting for Convex Constrained Markov Decision Processes Dynamic programming through the lens of semismooth Newton-type methods,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.165676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.655087Z digest=sha256:7d17bc61358034ffb37f63baca4eb80a112dd8c0dd69177c9c47d786c054fe34

Observation f226bed8-e293-433e-ae8c-f2f142ea9ed2 · outbound

This paper cites From optimization to control: quasi policy iteration,.

Operator Splitting for Convex Constrained Markov Decision Processes From optimization to control: quasi policy iteration,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.658357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.658357Z digest=sha256:d8ce150b0508d2231871624ca9470396c430116aa495b73ae1233716481a0842

Observation c85f91e9-7369-41ab-9e12-03177f016048 · outbound

This paper cites On the minimal displacement vector of the Douglas–Rachford operator,.

Operator Splitting for Convex Constrained Markov Decision Processes On the minimal displacement vector of the Douglas–Rachford operator,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.154673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.661823Z digest=sha256:03e7f75422747199fcefd4beb4f5bba33c75d75493c5137f7a3fc936706d5835

Observation 845b7a3d-417a-45ab-a633-b0261fba5ace · outbound

This paper cites On the Douglas–Rachford algorithm for solving possibly inconsistent optimization problems,.

Operator Splitting for Convex Constrained Markov Decision Processes On the Douglas–Rachford algorithm for solving possibly inconsistent optimization problems,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.143746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.664885Z digest=sha256:ba29fda45a7a71e8f9fe9fcef462de48694583c5e64b5de49567e4f61ac743c3

Observation d1af1847-d264-4ecb-b59f-461b404f105a · outbound

This paper cites Infeasibility detection in alternating direction method of multipliers for convex quadratic programs,.

Operator Splitting for Convex Constrained Markov Decision Processes Infeasibility detection in alternating direction method of multipliers for convex quadratic programs,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.133546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.668212Z digest=sha256:cfee84c7b30c49a0c3a98201dfce14ad8a59f2ec2a60f16d6e4583b77af26561

Observation 38476989-5207-4d52-9083-c9d5b7edbc04 · outbound

This paper cites A New Use of Douglas-Rachford Splitting and ADMM for Identifying Infeasible, Unbounded, and Pathological Conic Programs.

Operator Splitting for Convex Constrained Markov Decision Processes A New Use of Douglas-Rachford Splitting and ADMM for Identifying Infeasible, Unbounded, and Pathological Conic Programs

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-08-11T12:40:45.774770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.672091Z digest=sha256:835f3ae159f99dd591b7d1fd2a68eb05869631ac3ccf1d1db6ace601c8f44eac

Observation c19741f1-0603-4eff-8ff8-a0f75088959b · outbound

This paper cites Infeasibility detection in the alternating direction method of multipliers for convex optimization,.

Operator Splitting for Convex Constrained Markov Decision Processes Infeasibility detection in the alternating direction method of multipliers for convex optimization,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.122245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.675613Z digest=sha256:d4268fcb62992277ecba2fce75c7d7df21c72dfdf5846ff4d4796f34bb139e8f

Observation 2d88bf23-f813-4fe2-943b-38f62e751c1b · outbound

This paper cites an unresolved cited work.

Operator Splitting for Convex Constrained Markov Decision Processes Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-11T12:40:46.108724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.678725Z digest=sha256:f49bf25dcd0b3676cae557267e3fde63709fd654576d22499cc544757091f5c2

Observation 1b6545ba-014c-4585-a7dd-adf545063e20 · outbound

This paper cites On the Douglas—Rachford splitting method and the proximal point algorithm for maximal monotone operators,.

Operator Splitting for Convex Constrained Markov Decision Processes On the Douglas—Rachford splitting method and the proximal point algorithm for maximal monotone operators,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.096953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.681633Z digest=sha256:e0fe33feb8b5c642936f80e3d9306462f997b84600b4bb99f33cb66b96d07618

Observation c69d66ec-44ed-4ae4-aa7c-882db7784a41 · outbound

This paper cites On the convergence of the coordinate descent method for convex differentiable minimization,.

Operator Splitting for Convex Constrained Markov Decision Processes On the convergence of the coordinate descent method for convex differentiable minimization,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.084852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.684600Z digest=sha256:d021171e0038cf4b0f03c025f6fd3e8a57bbb512214139b6035cc79095a8a72b

Observation b49bbd16-32d3-4d0c-a4da-e95a873a9931 · outbound

This paper cites Nocedal and S.

Operator Splitting for Convex Constrained Markov Decision Processes Nocedal and S

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.072601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.687334Z digest=sha256:6fa397e7d75ee9db7af2af87095613d6699f90e0f4adb6bd5921211d1269dd47

Observation 9e4d7071-9aac-4c60-99b2-5ec01d98fb67 · outbound

This paper cites Operator-splitting methods for monotone affine variational inequalities, with a parallel application to optimal control,.

Operator Splitting for Convex Constrained Markov Decision Processes Operator-splitting methods for monotone affine variational inequalities, with a parallel application to optimal control,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.060999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.690317Z digest=sha256:62add7936adc06a43833e93ca58cccd219ae9c58bc0cf10ab81f07c21258ffb4

Observation a1b6c6ae-34c2-4999-90b9-0e8a26dfd4b6 · outbound

This paper cites Parallel alternating direction multiplier decomposition of convex programs,.

Operator Splitting for Convex Constrained Markov Decision Processes Parallel alternating direction multiplier decomposition of convex programs,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.693417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.693417Z digest=sha256:159567d481334f944fc5fcc4a8a4651b61220d2a111cc4a259b8612b602b7fc2

Observation 5dbb4e01-7ed9-4674-8b68-5ef2c559d5f2 · outbound

This paper cites Natural Actor- Critic Algorithms,.

Operator Splitting for Convex Constrained Markov Decision Processes Natural Actor- Critic Algorithms,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.041883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.696698Z digest=sha256:ca707972fecd0ab61af441e75950ad51076fef6e6557bb4f80ab0461cce06e88

Observation c93d3625-2ce4-4e7d-a5b2-e990d6dbc021 · outbound

This paper cites Pytorch: An imperative style, high-performance deep learning library,.

Operator Splitting for Convex Constrained Markov Decision Processes Pytorch: An imperative style, high-performance deep learning library,

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.699652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.699652Z digest=sha256:e5cf46d4e51d3e6205ea9c213a06a8615b4e8d196f1fbbcb87c32602d21a2ace

Observation 88f42b45-ae06-4a60-8f92-525e4980a090 · outbound

This paper cites PID accelerated value iteration algorithm,.

Operator Splitting for Convex Constrained Markov Decision Processes PID accelerated value iteration algorithm,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.022465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.702427Z digest=sha256:3983900cd69aae234851c9ddcf98434e45cb691746e415f48bcac723ebd5601f

Observation e5599083-65e5-438d-9d0e-c179ab9b09e5 · outbound

This paper cites Scalable first-order methods for robust MDPs,.

Operator Splitting for Convex Constrained Markov Decision Processes Scalable first-order methods for robust MDPs,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.008711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.705410Z digest=sha256:0dc23fa5997fd34cc40d73132c6698e135ed273920ea66691eac3ab8334e4b20

Observation 39b06f7f-f487-4906-9d07-fdd4f6b374f1 · outbound

This paper cites Integrating a partial model into model free reinforcement learning.,.

Operator Splitting for Convex Constrained Markov Decision Processes Integrating a partial model into model free reinforcement learning.,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:45.996527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.708307Z digest=sha256:4a528f433aaa3f153527a836ad860f33b34ac027f78519d0d4b0caf861576ec0

Observation 63fdc30f-6f3c-4af2-a99d-59108ef8521e · outbound

This paper cites Gurobi Optimizer Reference Manual,.

Operator Splitting for Convex Constrained Markov Decision Processes Gurobi Optimizer Reference Manual,

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.711404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.711404Z digest=sha256:c028c8dfeaab8d9c049a739687d8940a3af73ff22215d6f70e7f10a80cb1cdf5

Observation 259e0a59-fbdb-4c88-a3c2-6483e6ec897c · outbound

This paper cites Safe policies for reinforcement learning via primal-dual methods,.

Operator Splitting for Convex Constrained Markov Decision Processes Safe policies for reinforcement learning via primal-dual methods,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:45.978868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.714654Z digest=sha256:5ef0031a70b800b28d584ffaac4882ad5ff9e9556b4c159889bbbc0d60a3b7b0

Observation 2729d892-f1e5-4e18-b5ff-3355062fa1bd · outbound

This paper cites Conic optimization via operator splitting and homogeneous self-dual embedding,.

Operator Splitting for Convex Constrained Markov Decision Processes Conic optimization via operator splitting and homogeneous self-dual embedding,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:45.968594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:40:45.718057Z digest=sha256:7fe52560bb0fc8d1e5f1ab7390e576f40b3d87bca94fe6bd6129b194a6794f65

Observation d61bd2c1-8f1d-44a1-b1e3-8e9084a80edd · outbound

This paper cites Reward Constrained Policy Optimization.

Operator Splitting for Convex Constrained Markov Decision Processes Reward Constrained Policy Optimization

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.721211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.721211Z digest=sha256:cd85e9b48519793d9bfd2a2ef99a90333df5da75de709d88fb724364d1d5e95b

Observation 992f0858-c3e3-459b-ab05-cc8006d66e68 · outbound

This paper cites Markov decision processes,.

Operator Splitting for Convex Constrained Markov Decision Processes Markov decision processes,

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.724621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.724621Z digest=sha256:8b5810a95da2a879d718f8a57fbb7a6e940119b8ef5397dd6cfa1c320f4f6812

Observation 32b6a08f-3524-4360-8b11-de9e3173d032 · outbound

This paper cites an unresolved cited work.

Operator Splitting for Convex Constrained Markov Decision Processes Unresolved cited work

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.727688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.727688Z digest=sha256:128a598cd9acbb516e21ea0a9b5ec4440b0d782aecee937e65225ad902c40504

Pith citing papers

Observation c91b5c64-12f6-42f0-b676-61bc626b43da · inbound

Joint Chance Constrained Safe-Optimal Control cites this paper.

Joint Chance Constrained Safe-Optimal Control Operator Splitting for Convex Constrained Markov Decision Processes

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-07-01T12:55:44.118969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-01T01:40:08.230114Z digest=sha256:7126b3a737d7e966df57d85f89df59e7bc789c87b569773bf2388126ae695090