Pith. sign in

Paper Citation Record · LEDGER

Beyond Slater's Condition in Online CMDPs with Stochastic and Adversarial Constraints

As of 10 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 0 inbound Pith citation observations for arXiv:2509.20114.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.20114 v3

Coverage vector

measured 15 of 15 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T15:22:10.261211Z

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

15 of 15 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a60c335d-cb89-4eb5-ac27-02fd3efae8f2 · outbound

This paper cites Lemma B.1.For anyδ∈(0,1)and for anyq∈ T t∈[T] b∆t(Pt), Algorithm 1 attains: TX t=1 bℓ⊤ t (bqt −q)≤L ln |X| 2|A| η +η|X||A|T+ ηLln L δ γ , with probability at least1−δ.

Beyond Slater's Condition in Online CMDPs with Stochastic and Adversarial Constraints Lemma B.1.For anyδ∈(0,1)and for anyq∈ T t∈[T] b∆t(Pt), Algorithm 1 attains: TX t=1 bℓ⊤ t (bqt −q)≤L ln |X| 2|A| η +η|X||A|T+ ηLln L δ γ , with probability at least1−δ

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T15:22:10.258453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T15:22:10.258453Z digest=sha256:9b559583c76170680179f1052332dced155e972d8a0d525cdfa46cda984206d2

Observation 569961bc-66eb-4e7a-83c8-76ffeda4b0ac · outbound

This paper cites Aviv Rosenberg and Yishay Mansour.

Beyond Slater's Condition in Online CMDPs with Stochastic and Adversarial Constraints Aviv Rosenberg and Yishay Mansour

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T15:22:10.233504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T15:22:10.233504Z digest=sha256:94c8014c41c87171ff97ad89cf12accd814efb589b68c56c4a446f65cac2c7a0

Observation 4373585a-8186-4eab-be09-5558d2862fe7 · outbound

This paper cites Learning Constrained Markov Decision Processes With Non-stationary Rewards and Constraints.

Beyond Slater's Condition in Online CMDPs with Stochastic and Adversarial Constraints Learning Constrained Markov Decision Processes With Non-stationary Rewards and Constraints

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T15:22:10.240141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T15:22:10.240141Z digest=sha256:5bf42ca88fc982f972396ee965b25a5380d0cc87c364cc2cd4c83709064e06e2

Observation 6735a635-8dc9-48ae-a763-32af00432d41 · outbound

This paper cites The authors analyze two approaches, both providing sub- linear regret and cumulative constraint violation.

Beyond Slater's Condition in Online CMDPs with Stochastic and Adversarial Constraints The authors analyze two approaches, both providing sub- linear regret and cumulative constraint violation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T15:22:10.249378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T15:22:10.249378Z digest=sha256:288879999f0038fda006c682f571993499c46a8cd61fa523408ed77faddb04bd

Observation cc09dbf4-8760-4d98-a818-b2658c2f5f88 · outbound

This paper cites This algorithm achieves eO(T 3 4 ) regret and guarantees that the cumulative constraint violation remains below a certain threshold with a given probability.

Beyond Slater's Condition in Online CMDPs with Stochastic and Adversarial Constraints This algorithm achieves eO(T 3 4 ) regret and guarantees that the cumulative constraint violation remains below a certain threshold with a given probability

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T15:22:10.252298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T15:22:10.252298Z digest=sha256:a32c729d3a44a88e898b2ea475cc27bc2556dd46e1065238104360c3cc29cdcc

Observation 203f5f2a-5e65-467f-b2fe-3ac8e50fc338 · outbound

This paper cites The first best-of-both- worlds algorithm for online learning in episodic CMDPs was proposed by Stradi et al.

Beyond Slater's Condition in Online CMDPs with Stochastic and Adversarial Constraints The first best-of-both- worlds algorithm for online learning in episodic CMDPs was proposed by Stradi et al

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T15:22:10.255540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T15:22:10.255540Z digest=sha256:863102e20a6050aa258fda2211d9221760d78d42e0f1fc4a003e93fd692c61b9

Observation b2b2ef56-77a9-4708-bbfd-04238ebcf7d2 · outbound

This paper cites In the stochastic setting, Algorithm 1 guarantees with probability at least 1−16δ: Vt ≤18L|X| r 2t|A|ln 2mT|X||A| δ ∀t∈[T].

Beyond Slater's Condition in Online CMDPs with Stochastic and Adversarial Constraints In the stochastic setting, Algorithm 1 guarantees with probability at least 1−16δ: Vt ≤18L|X| r 2t|A|ln 2mT|X||A| δ ∀t∈[T]

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T15:22:10.261211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T15:22:10.261211Z digest=sha256:1e22106463ba594402d99bf98564240e0a427fd1a5509376f265826a5177f43e

Observation 02f1a4b9-f96a-4213-bc20-cf82780c2fd4 · outbound

This paper cites URLhttps://proceedings.neurips.cc/paper/2019/file/ a0872cc5b5ca4cc25076f3d868e1bdf8-Paper.pdf.

Beyond Slater's Condition in Online CMDPs with Stochastic and Adversarial Constraints URLhttps://proceedings.neurips.cc/paper/2019/file/ a0872cc5b5ca4cc25076f3d868e1bdf8-Paper.pdf

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T15:22:10.237090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T15:22:10.237090Z digest=sha256:ba1eeb5e093bfbfe60b9ee675bd75baf85c5b998abe9530282bc39a597f1cfbc

Observation 34b9996a-d294-40e2-9f7e-92c20fe929b6 · outbound

This paper cites Mohammad Gheshlaghi Azar, Ian Osband, and R´ emi Munos.

Beyond Slater's Condition in Online CMDPs with Stochastic and Adversarial Constraints Mohammad Gheshlaghi Azar, Ian Osband, and R´ emi Munos

Reference 2008

Resolution
unresolved
no resolver link, observed 2026-08-04T15:22:09.694108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T15:22:09.694108Z digest=sha256:5f3eeb20b48b237d22987971391c4c6ce5ce07aefcc064f9e2b2cbf91be2cab6

Observation 5b8aad86-3091-4962-acdd-1bab78d4805c · outbound

This paper cites the algorithm receives the complete loss/reward information.

Beyond Slater's Condition in Online CMDPs with Stochastic and Adversarial Constraints the algorithm receives the complete loss/reward information

Reference 2009

Resolution
unresolved
no resolver link, observed 2026-08-04T15:22:10.246588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T15:22:10.246588Z digest=sha256:2f9120ead056a898ec484a449692c116f7e2f56a4c8f0cf212dc2f57257eeb9d

Observation 2c4ea8fa-317f-45f2-b760-dfa6fba25a97 · outbound

This paper cites 13 Contents 1 Introduction 1 1.1 Original Contributions.

Beyond Slater's Condition in Online CMDPs with Stochastic and Adversarial Constraints 13 Contents 1 Introduction 1 1.1 Original Contributions

Reference 2013

Resolution
unresolved
no resolver link, observed 2026-08-04T15:22:10.243720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T15:22:10.243720Z digest=sha256:ccb32eb26d9dc334a0fced542405c237d72796c55b7657ce2f133719aa931660

Observation ba09c85e-9df3-46d6-a408-475eb95d0af7 · outbound

This paper cites Gergely Neu, Andras Antos, Andr´ as Gy¨ orgy, and Csaba Szepesv´ ari.

Beyond Slater's Condition in Online CMDPs with Stochastic and Adversarial Constraints Gergely Neu, Andras Antos, Andr´ as Gy¨ orgy, and Csaba Szepesv´ ari

Reference 2015

Resolution
unresolved
no resolver link, observed 2026-08-04T15:22:09.982015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T15:22:09.982015Z digest=sha256:6a2f6cd6aec28a6f9a90a62e41990af27da847968a41c5e4f813d4f119be0935

Observation 86531a45-37d1-4ab5-8b3b-a6267f30a5c9 · outbound

This paper cites Online Learning: A Modern Introduction Using Convex Optimization.

Beyond Slater's Condition in Online CMDPs with Stochastic and Adversarial Constraints Online Learning: A Modern Introduction Using Convex Optimization

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-04T15:22:10.230161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T15:22:10.230161Z digest=sha256:b0054ee66b59ac73c0ffc1542928548cd5d28a15be8a95eeb5f654547ce42229

Observation 571673aa-a36d-4155-abbd-62165f56dc9a · outbound

This paper cites Exploration-Exploitation in Constrained MDPs.

Beyond Slater's Condition in Online CMDPs with Stochastic and Adversarial Constraints Exploration-Exploitation in Constrained MDPs

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-04T15:22:09.724181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T15:22:09.724181Z digest=sha256:82884bb3683e17e30bc4d8c6e0a78f501f15451872329083c82dee6660a56470

Observation 054b12b5-6740-4e04-8b42-a24eaf4f41b6 · outbound

This paper cites Safe reinforcement learning on autonomous vehicles.

Beyond Slater's Condition in Online CMDPs with Stochastic and Adversarial Constraints Safe reinforcement learning on autonomous vehicles

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-04T15:22:09.865668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T15:22:09.865668Z digest=sha256:2e345272a96c35155b32784af853cf8d5f1562cafe30e7201cb5ed783f1264eb

Pith citing papers

No inbound Pith citation observations are available.