Pith. sign in

Paper Citation Record · LEDGER

Operator Splitting for Convex Constrained Markov Decision Processes

As of 16 August 2026, this Paper Citation Record lists 62 of 62 outbound references and 1 inbound Pith citation observation for arXiv:2412.14002.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.14002 v1

Coverage vector

measured 62 of 62 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T12:40:45.727688Z

measured 63 of 63 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-01T01:40:08.230114Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T12:55:44.116744Z

Reference resolution

62 of 62 outbound references displayed

  • verified exact4
  • verified fuzzy42
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7d5debb6-f0c9-42d9-8d71-e9b5c137e1ba · outbound

This paper cites Mastering the game of go without human knowledge,.

Operator Splitting for Convex Constrained Markov Decision Processes Mastering the game of go without human knowledge,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.474791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.523266Z digest=sha256:094843abefd28cec6923b522770a2c647977b8899ce40c1abe24ef14f0cdf617

Observation a14a017c-7b4c-46bb-93e6-842395d858bc · outbound

This paper cites Magnetic control of tokamak plasmas through deep reinforcement learning,.

Operator Splitting for Convex Constrained Markov Decision Processes Magnetic control of tokamak plasmas through deep reinforcement learning,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.463803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.527737Z digest=sha256:2366e705c86d1dbd4cc754b8a2c45e6d5d365ed71111fc39be076120a25de6c2

Observation fc2e3662-6eba-4b80-8a36-fd86e8f01db8 · outbound

This paper cites an unresolved cited work.

Operator Splitting for Convex Constrained Markov Decision Processes Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.531734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.531734Z digest=sha256:497f3ac960992632756ffd1977633efbb5879b25d6c520aff720cd4012ced91b

Observation d5ac92ef-5561-4ba3-9ad0-9402d472a519 · outbound

This paper cites Altman, Constrained Markov decision processes.

Operator Splitting for Convex Constrained Markov Decision Processes Altman, Constrained Markov decision processes

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.535342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.535342Z digest=sha256:7795146e4a0ff8a1b23f86e8243e670ded75f7802f21eea5d5bf7ca6771f2175

Observation 5a515a8e-f66a-489c-a58b-7e289c4e5f3f · outbound

This paper cites Policy gradients with variance related risk criteria,.

Operator Splitting for Convex Constrained Markov Decision Processes Policy gradients with variance related risk criteria,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.438344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.539032Z digest=sha256:cfa7cbe3ef6fbe72d049faa541185489d7133a9d0d7385014fea57df58d0d57b

Observation 9ec6acb3-8b02-4766-91c3-e51e28f664db · outbound

This paper cites Risk-constrained reinforcement learning with percentile risk criteria,.

Operator Splitting for Convex Constrained Markov Decision Processes Risk-constrained reinforcement learning with percentile risk criteria,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.426469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.542889Z digest=sha256:7c11ef20d3835df3f4a6c30bb00019ae465ee483bd2f3711f6b589c989aec01c

Observation ee91af8a-5923-494f-a011-6aca5f6ab9d7 · outbound

This paper cites Control and optimization meet the smart power grid: Scheduling of power demands for optimal energy management,.

Operator Splitting for Convex Constrained Markov Decision Processes Control and optimization meet the smart power grid: Scheduling of power demands for optimal energy management,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.416377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.546898Z digest=sha256:e750e35bb9841b0b63a7203c3619407cb97cb77014df6c469d29bd3fc01854fd

Observation 542b87d3-b08d-435d-be6e-0e07797a0d49 · outbound

This paper cites Constrained policy optimization,.

Operator Splitting for Convex Constrained Markov Decision Processes Constrained policy optimization,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.406379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.550549Z digest=sha256:0e0dd7cfa3d6998f0c274f07f9dbf03271b2e2005fd64e7216e90b9153267d86

Observation 45b38f5c-3b29-440f-81cb-573b57a0950d · outbound

This paper cites Dynamic programming equations for dis- counted constrained stochastic control,.

Operator Splitting for Convex Constrained Markov Decision Processes Dynamic programming equations for dis- counted constrained stochastic control,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.396802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.553944Z digest=sha256:76ea90a01f723d7a84f158ed13bcfa81c287db1a0ce37ecb1a864986dba76dcc

Observation ad5d1b4d-e489-46fd-b2bf-0ace75ad2016 · outbound

This paper cites Dynamic programming in constrained Markov decision processes,.

Operator Splitting for Convex Constrained Markov Decision Processes Dynamic programming in constrained Markov decision processes,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.387474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.557316Z digest=sha256:06764cb88fd9c2c74141e9c191ce506a49840a3fa806ffa1476a8faa9167a333

Observation cd8afd8c-185d-45f4-bb8d-cdd29e24ae5f · outbound

This paper cites A Gradient-Aware Search Algorithm for Constrained Markov Decision Processes.

Operator Splitting for Convex Constrained Markov Decision Processes A Gradient-Aware Search Algorithm for Constrained Markov Decision Processes

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-08-11T12:40:45.942537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.560934Z digest=sha256:d5f0da1e38690430c10b9f8e475e8c41ebee70141b4909a3a50503b7eb3552b8

Observation 81bb3222-c4a7-4975-9bfa-60d1a820a59d · outbound

This paper cites Natural policy gradient primal-dual method for constrained Markov decision processes,.

Operator Splitting for Convex Constrained Markov Decision Processes Natural policy gradient primal-dual method for constrained Markov decision processes,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.376767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.564782Z digest=sha256:481fc96dbb8ab82919782c64a7ea3fd010f7439c60cf1dc398452bb374c6b37c

Observation a42db668-8b00-4eca-88ab-0631211d5210 · outbound

This paper cites Learning policies with zero or bounded constraint violation for constrained MDPs,.

Operator Splitting for Convex Constrained Markov Decision Processes Learning policies with zero or bounded constraint violation for constrained MDPs,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.366953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.568209Z digest=sha256:fe25c6f1923503d9984ac0ed0639858f0ccf929ed7b12bb0a5d9908cb2df096d

Observation e5e4fd73-681f-4d25-8067-da8a5f2ce359 · outbound

This paper cites State Augmented Constrained Reinforcement Learning: Overcoming the Limitations of Learning with Rewards.

Operator Splitting for Convex Constrained Markov Decision Processes State Augmented Constrained Reinforcement Learning: Overcoming the Limitations of Learning with Rewards

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-08-11T12:40:45.928287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.571142Z digest=sha256:40a09860948e12d5e894a1025453a47983c43638b2d9de33ddf6050689bf5626

Observation 059a8990-aadc-44e1-b9e0-bbe49e2e94f9 · outbound

This paper cites Constrained MDPs and the reward hypothesis.

Operator Splitting for Convex Constrained Markov Decision Processes Constrained MDPs and the reward hypothesis

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.357018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.574448Z digest=sha256:e3e59cf605bf134556f9d4da659e62432a62a8e97afcad8dbb714ed66a2f18d1

Observation dc886c82-2731-4f9d-a3a9-9c3e88d9d4ec · outbound

This paper cites Two “well-known.

Operator Splitting for Convex Constrained Markov Decision Processes Two “well-known

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.346626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.576948Z digest=sha256:741fa6e87057ae2722df0632207f85faf8419e9f8de2ff72af4bf6b31956c6bb

Observation 94358415-df0b-4c2c-8470-c8a9c5a99ca8 · outbound

This paper cites Algorithm for constrained Markov decision process with linear convergence,.

Operator Splitting for Convex Constrained Markov Decision Processes Algorithm for constrained Markov decision process with linear convergence,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.336444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.579723Z digest=sha256:f5035dbf3edd4480f6ba4764b637578052ebc0d5c6dffdfe840c4b08b17f92de

Observation df37e178-a510-4759-9e23-c8ca8325bc49 · outbound

This paper cites Faster Algorithm and Sharper Analysis for Constrained Markov Decision Process.

Operator Splitting for Convex Constrained Markov Decision Processes Faster Algorithm and Sharper Analysis for Constrained Markov Decision Process

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-11T12:40:45.912662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.582528Z digest=sha256:d5f3792e541ea4bca46a7509572f319b4253a5f34aa91d6306875532f054954c

Observation 436de9d0-9ff3-4e8f-8b7a-e28e86f689d4 · outbound

This paper cites Cancellation-Free Regret Bounds for Lagrangian Approaches in Constrained Markov Decision Processes.

Operator Splitting for Convex Constrained Markov Decision Processes Cancellation-Free Regret Bounds for Lagrangian Approaches in Constrained Markov Decision Processes

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.585747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.585747Z digest=sha256:4d5b6365d7e341bf101221cb869ae59fbeb6bbd4eda2cf6bf7903e6a816a22ea

Observation e7dc5fbf-0b77-45e9-bc34-8fb93dbf9c2d · outbound

This paper cites Last-Iterate Convergent Policy Gradient Primal-Dual Methods for Constrained MDPs.

Operator Splitting for Convex Constrained Markov Decision Processes Last-Iterate Convergent Policy Gradient Primal-Dual Methods for Constrained MDPs

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.589096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.589096Z digest=sha256:4fa4cea5b45ddb3728879395ef510d8aea3afd512d8ee79fc2345f4ccc09c54c

Observation d9b00ec9-7e4a-45d9-b6c9-35dac966afef · outbound

This paper cites Reload: Reinforcement learning with optimistic ascent- descent for last-iterate convergence in constrained MDPs,.

Operator Splitting for Convex Constrained Markov Decision Processes Reload: Reinforcement learning with optimistic ascent- descent for last-iterate convergence in constrained MDPs,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.324278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.592853Z digest=sha256:e576d4bdcbd44a87fc0f3358ea1df3f8c73dcc70042ff360cc88ad11d9418885

Observation 15e44324-8a48-4e21-aec3-1d4ff6d76fea · outbound

This paper cites Ipo: Interior-point policy optimization under constraints,.

Operator Splitting for Convex Constrained Markov Decision Processes Ipo: Interior-point policy optimization under constraints,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.313087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.596272Z digest=sha256:f93fd4d7b3a327b07af139b9948a1f8c1c9452546d756ab846c13da815d6f611

Observation d9ca5f64-4792-4ad5-8b81-cca65e75d970 · outbound

This paper cites Projection-Based Constrained Policy Optimization.

Operator Splitting for Convex Constrained Markov Decision Processes Projection-Based Constrained Policy Optimization

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.599554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.599554Z digest=sha256:5afa0b24498e39d6ffeaa2fd00bd21fa17d278406e4811cf276aca14fd3edd64

Observation 0bc2fc3e-f15d-42a0-b7ac-9ed2ce28805e · outbound

This paper cites Reward is enough for convex MDPs,.

Operator Splitting for Convex Constrained Markov Decision Processes Reward is enough for convex MDPs,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.302622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.604198Z digest=sha256:d32d67e2b33bcb0419461d4c6cc3781d639fca58da351304ed603180e53aa9c2

Observation 3b81aac5-0a43-47d3-8586-bddbc21ce90a · outbound

This paper cites Apprenticeship learning via inverse reinforce- ment learning,.

Operator Splitting for Convex Constrained Markov Decision Processes Apprenticeship learning via inverse reinforce- ment learning,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.292322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.607868Z digest=sha256:ff4035beb122582a91bbd7844f33d41a56c203857dbc6243b8cb3522fc705617

Observation 41a27b59-5d10-424e-9af1-3f570162db0b · outbound

This paper cites Provably efficient maximum entropy exploration,.

Operator Splitting for Convex Constrained Markov Decision Processes Provably efficient maximum entropy exploration,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.281858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.611007Z digest=sha256:64cd9b680ce8e30174a800e46fedc7d24d00f8156477389523b306c61ee46306

Observation b9945c27-18a9-4bb5-8a98-148d484fa784 · outbound

This paper cites Diversity is All You Need: Learning Skills without a Reward Function.

Operator Splitting for Convex Constrained Markov Decision Processes Diversity is All You Need: Learning Skills without a Reward Function

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.614529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.614529Z digest=sha256:5b549d96d0e5b34ab310075d9c9f8f82468eb018361d22f1f9e491e3677c37da

Observation a405dd85-26bc-424e-9b46-9cece872a9ad · outbound

This paper cites Policy-based primal-dual methods for convex constrained Markov decision processes,.

Operator Splitting for Convex Constrained Markov Decision Processes Policy-based primal-dual methods for convex constrained Markov decision processes,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.271500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.618422Z digest=sha256:a4591bf0422b4c8d53d5226bbb4db28d5c323179cf4f8e639773a1c2cac43cb0

Observation 5d238fcc-519d-47b8-97fe-741bdfe0b226 · outbound

This paper cites Variational policy gradient method for reinforcement learning with general utilities,.

Operator Splitting for Convex Constrained Markov Decision Processes Variational policy gradient method for reinforcement learning with general utilities,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.260720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.621795Z digest=sha256:2e86cff5dad3393d0478bad191e342fdacafc3f890487f0bcb784b13d06b5899

Observation 5c4a64fb-80a3-4e21-babd-fb34144bf916 · outbound

This paper cites Reinforcement learning with convex constraints,.

Operator Splitting for Convex Constrained Markov Decision Processes Reinforcement learning with convex constraints,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.249377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.625100Z digest=sha256:b1273bb7544dad65220ed5ff072a3a39a0a3bd607118eab372c905d3afc3f0a4

Observation 6452146a-e617-4068-b5a7-af6bf91dbf54 · outbound

This paper cites A simple reward-free approach to constrained reinforcement learning,.

Operator Splitting for Convex Constrained Markov Decision Processes A simple reward-free approach to constrained reinforcement learning,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.238752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.628643Z digest=sha256:6d2d1455a8c32c165b10787ecfcfee31369f83109eee3e981d138dc9f59db458

Observation 002b4718-6274-4d5d-8a35-5130c8ea7d9c · outbound

This paper cites Bauschke and P.

Operator Splitting for Convex Constrained Markov Decision Processes Bauschke and P

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.227584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.632070Z digest=sha256:3abcbf15b2db448c6a43393ef71602aded048a0cf743a68732dffcfc2f58a3a7

Observation aae25587-1f49-4ccd-bdf5-ae98d45aed63 · outbound

This paper cites Distributed optimization and statistical learning via the alternating direction method of multipliers,.

Operator Splitting for Convex Constrained Markov Decision Processes Distributed optimization and statistical learning via the alternating direction method of multipliers,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.635420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.635420Z digest=sha256:a72cb6bed6189dafdc63b1098e3d2aaf88ef1d9aacfc155807e155407f116cf3

Observation 8694d196-c0b1-433a-91ec-0f717a1e6517 · outbound

This paper cites A note on the equivalence of operator splitting methods,.

Operator Splitting for Convex Constrained Markov Decision Processes A note on the equivalence of operator splitting methods,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.210336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.638587Z digest=sha256:6f384dd7ed7f5ce9136040b59498b1e158102077cfe153ee5b597542dc969eb2

Observation 729dc1db-fde2-477b-9cc8-c21276720462 · outbound

This paper cites Provably efficient algorithms for multi-objective competitive RL,.

Operator Splitting for Convex Constrained Markov Decision Processes Provably efficient algorithms for multi-objective competitive RL,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.198587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.641928Z digest=sha256:17891351e23c1f3d16ed8af02462bb2ca398d31c5699022e0de56b2ce43ecfba

Observation 148017e3-2317-4d51-86ef-164c340ffebb · outbound

This paper cites A splitting method for optimal control,.

Operator Splitting for Convex Constrained Markov Decision Processes A splitting method for optimal control,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.187573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.645327Z digest=sha256:57ec4c18fe05d457b66cf79f3c576b926139d75688b94a1ac3b1f05ff2dd0f5b

Observation 335a47a6-b86e-4328-95d5-b970d8ec748d · outbound

This paper cites A unified view of entropy-regularized Markov decision processes.

Operator Splitting for Convex Constrained Markov Decision Processes A unified view of entropy-regularized Markov decision processes

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.648460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.648460Z digest=sha256:f5836919e4af8eb8dc39b4e8a7dfd445eec9195e647a1e973c9b8bbad7996160

Observation 34500308-5375-46e9-a658-cf89b33ce04e · outbound

This paper cites A theory of regularized Markov decision processes,.

Operator Splitting for Convex Constrained Markov Decision Processes A theory of regularized Markov decision processes,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.176808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.651895Z digest=sha256:8323360220a1c2452f63e23425693b2653446369c8d71fa6af119ef84d298aef

Observation 47314e45-c403-49b2-9885-921c406adbef · outbound

This paper cites Dynamic programming through the lens of semismooth Newton-type methods,.

Operator Splitting for Convex Constrained Markov Decision Processes Dynamic programming through the lens of semismooth Newton-type methods,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.165676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.655087Z digest=sha256:4f09d4e551dbf04935e62b45e28223da4bf3966aea3708be8834239a65e431a8

Observation f226bed8-e293-433e-ae8c-f2f142ea9ed2 · outbound

This paper cites From optimization to control: quasi policy iteration,.

Operator Splitting for Convex Constrained Markov Decision Processes From optimization to control: quasi policy iteration,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.658357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.658357Z digest=sha256:6efacd27b3481f148c7a2364c1492bcbc73c52f0c89844ce0b3f5ffcea21ca49

Observation c85f91e9-7369-41ab-9e12-03177f016048 · outbound

This paper cites On the minimal displacement vector of the Douglas–Rachford operator,.

Operator Splitting for Convex Constrained Markov Decision Processes On the minimal displacement vector of the Douglas–Rachford operator,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.154673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.661823Z digest=sha256:0bf0867517e45aaf66951c1bb12853bb4a9451ce62549178071a44f68fd2ba34

Observation 845b7a3d-417a-45ab-a633-b0261fba5ace · outbound

This paper cites On the Douglas–Rachford algorithm for solving possibly inconsistent optimization problems,.

Operator Splitting for Convex Constrained Markov Decision Processes On the Douglas–Rachford algorithm for solving possibly inconsistent optimization problems,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.143746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.664885Z digest=sha256:2d64c5888637374b53eb127c9e5b7cc76e838675b4e6259da52d62fc64f6582f

Observation d1af1847-d264-4ecb-b59f-461b404f105a · outbound

This paper cites Infeasibility detection in alternating direction method of multipliers for convex quadratic programs,.

Operator Splitting for Convex Constrained Markov Decision Processes Infeasibility detection in alternating direction method of multipliers for convex quadratic programs,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.133546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.668212Z digest=sha256:101e012e55098378be1fec04ba544f1b48c80d7fd0dcfd82e38b1a0ac2be4b89

Observation 38476989-5207-4d52-9083-c9d5b7edbc04 · outbound

This paper cites A New Use of Douglas-Rachford Splitting and ADMM for Identifying Infeasible, Unbounded, and Pathological Conic Programs.

Operator Splitting for Convex Constrained Markov Decision Processes A New Use of Douglas-Rachford Splitting and ADMM for Identifying Infeasible, Unbounded, and Pathological Conic Programs

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-08-11T12:40:45.774770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.672091Z digest=sha256:33aef1a12626c6d3831fbe75d91cf016dc39d419ad16b183f5fe6dee2b3821dc

Observation c19741f1-0603-4eff-8ff8-a0f75088959b · outbound

This paper cites Infeasibility detection in the alternating direction method of multipliers for convex optimization,.

Operator Splitting for Convex Constrained Markov Decision Processes Infeasibility detection in the alternating direction method of multipliers for convex optimization,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.122245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.675613Z digest=sha256:3affc67d91c6e81b315fd03dfc0382fe32cb3809e33232cc07ee0574e6fe5b49

Observation 2d88bf23-f813-4fe2-943b-38f62e751c1b · outbound

This paper cites an unresolved cited work.

Operator Splitting for Convex Constrained Markov Decision Processes Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-11T12:40:46.108724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.678725Z digest=sha256:b0bdc1be2a1573979f6ab9e6a234ad185e96aa27ef83db50f66b9bf54476d00d

Observation 1b6545ba-014c-4585-a7dd-adf545063e20 · outbound

This paper cites On the Douglas—Rachford splitting method and the proximal point algorithm for maximal monotone operators,.

Operator Splitting for Convex Constrained Markov Decision Processes On the Douglas—Rachford splitting method and the proximal point algorithm for maximal monotone operators,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.096953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.681633Z digest=sha256:e2af60f2d0604292a270a4afa0602515a0fc799f40bcec637909463e65abf857

Observation c69d66ec-44ed-4ae4-aa7c-882db7784a41 · outbound

This paper cites On the convergence of the coordinate descent method for convex differentiable minimization,.

Operator Splitting for Convex Constrained Markov Decision Processes On the convergence of the coordinate descent method for convex differentiable minimization,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.084852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.684600Z digest=sha256:1107875be09e9a57b50b40891251e411cebcb7041b49b91c94b96b5bbf1c61cd

Observation b49bbd16-32d3-4d0c-a4da-e95a873a9931 · outbound

This paper cites Nocedal and S.

Operator Splitting for Convex Constrained Markov Decision Processes Nocedal and S

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.072601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.687334Z digest=sha256:6cb688124260690a3a1a04b4d1ef880ab0d4373e8b8476433438a5cb9e18c50f

Observation 9e4d7071-9aac-4c60-99b2-5ec01d98fb67 · outbound

This paper cites Operator-splitting methods for monotone affine variational inequalities, with a parallel application to optimal control,.

Operator Splitting for Convex Constrained Markov Decision Processes Operator-splitting methods for monotone affine variational inequalities, with a parallel application to optimal control,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.060999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.690317Z digest=sha256:015d46772254a256877bf3c7d19d0a02457a3a44516c0c880eca3b1304360e86

Observation a1b6c6ae-34c2-4999-90b9-0e8a26dfd4b6 · outbound

This paper cites Parallel alternating direction multiplier decomposition of convex programs,.

Operator Splitting for Convex Constrained Markov Decision Processes Parallel alternating direction multiplier decomposition of convex programs,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.693417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.693417Z digest=sha256:469c392880128248a4a7ea41f4051873938f18ba4f2d32984ac1c72d63ad37ae

Observation 5dbb4e01-7ed9-4674-8b68-5ef2c559d5f2 · outbound

This paper cites Natural Actor- Critic Algorithms,.

Operator Splitting for Convex Constrained Markov Decision Processes Natural Actor- Critic Algorithms,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.041883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.696698Z digest=sha256:eff9a3d965527f3728f6b961fa53140dcb6c98f34a9b795a832ad226c96e7127

Observation c93d3625-2ce4-4e7d-a5b2-e990d6dbc021 · outbound

This paper cites Pytorch: An imperative style, high-performance deep learning library,.

Operator Splitting for Convex Constrained Markov Decision Processes Pytorch: An imperative style, high-performance deep learning library,

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.699652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.699652Z digest=sha256:0e06258a94241bd00f3f7a2fc7e58a6a1f940e64d0dc3e6a5f32a4c918122425

Observation 88f42b45-ae06-4a60-8f92-525e4980a090 · outbound

This paper cites PID accelerated value iteration algorithm,.

Operator Splitting for Convex Constrained Markov Decision Processes PID accelerated value iteration algorithm,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.022465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.702427Z digest=sha256:f02eafe4b42bb721ed3619242ba4be504f521f648f6c0a2e97dd20ad5ac34354

Observation e5599083-65e5-438d-9d0e-c179ab9b09e5 · outbound

This paper cites Scalable first-order methods for robust MDPs,.

Operator Splitting for Convex Constrained Markov Decision Processes Scalable first-order methods for robust MDPs,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.008711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.705410Z digest=sha256:b07d86614f6d30baf090945181187eada4778abceddab9351d9130fd268d10a2

Observation 39b06f7f-f487-4906-9d07-fdd4f6b374f1 · outbound

This paper cites Integrating a partial model into model free reinforcement learning.,.

Operator Splitting for Convex Constrained Markov Decision Processes Integrating a partial model into model free reinforcement learning.,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:45.996527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.708307Z digest=sha256:87d4f0f653f6f68ed886dae0e784f1796faf2714646c395924bb364c1cd225a6

Observation 63fdc30f-6f3c-4af2-a99d-59108ef8521e · outbound

This paper cites Gurobi Optimizer Reference Manual,.

Operator Splitting for Convex Constrained Markov Decision Processes Gurobi Optimizer Reference Manual,

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.711404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.711404Z digest=sha256:a69dcec3c89bb9a96a9118e3463b6b22a010ca1b5c5207949a2a2d68b64543e8

Observation 259e0a59-fbdb-4c88-a3c2-6483e6ec897c · outbound

This paper cites Safe policies for reinforcement learning via primal-dual methods,.

Operator Splitting for Convex Constrained Markov Decision Processes Safe policies for reinforcement learning via primal-dual methods,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:45.978868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.714654Z digest=sha256:869300e67c13d9cc8d583054c9f9981da994cb56d387f46c929fe11c98fd2ccf

Observation 2729d892-f1e5-4e18-b5ff-3355062fa1bd · outbound

This paper cites Conic optimization via operator splitting and homogeneous self-dual embedding,.

Operator Splitting for Convex Constrained Markov Decision Processes Conic optimization via operator splitting and homogeneous self-dual embedding,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:45.968594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T12:40:45.718057Z digest=sha256:ce3fd5b583f0343a14e798ff03529a708c2882b883734734e85265b7c905ee3e

Observation d61bd2c1-8f1d-44a1-b1e3-8e9084a80edd · outbound

This paper cites Reward Constrained Policy Optimization.

Operator Splitting for Convex Constrained Markov Decision Processes Reward Constrained Policy Optimization

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.721211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.721211Z digest=sha256:e385b2905f65b3f0718e6884683ea18dd99c1b6cb75330c141fe1f2c9d3959c1

Observation 992f0858-c3e3-459b-ab05-cc8006d66e68 · outbound

This paper cites Markov decision processes,.

Operator Splitting for Convex Constrained Markov Decision Processes Markov decision processes,

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.724621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.724621Z digest=sha256:2fc641b1bd6578972d7738e4ba6379b8911957dcf1ae6977a113c400c5030c42

Observation 32b6a08f-3524-4360-8b11-de9e3173d032 · outbound

This paper cites an unresolved cited work.

Operator Splitting for Convex Constrained Markov Decision Processes Unresolved cited work

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.727688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.727688Z digest=sha256:9ffd193b4e380fc9597210a3e8ce582814c177028c34dc37d98bdf48d883ada1

Pith citing papers

Observation c91b5c64-12f6-42f0-b676-61bc626b43da · inbound

Joint Chance Constrained Safe-Optimal Control cites this paper.

Joint Chance Constrained Safe-Optimal Control Operator Splitting for Convex Constrained Markov Decision Processes

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-07-01T12:55:44.118969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-07-01T01:40:08.230114Z digest=sha256:0f1a21e7a18a362df6be594dd94a4cd1706e6f6e1ea4d272addb23e5fd20a083