Pith. sign in

Paper Citation Record · LEDGER

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning

As of 21 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 0 inbound Pith citation observations for arXiv:2506.14125.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.14125 v1

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:01:08.600180Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

49 of 49 outbound references displayed

  • verified exact1
  • verified fuzzy37
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5be75a55-5d9c-4f6e-a65e-ce76b800917e · outbound

This paper cites Constrained policy optimization.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Constrained policy optimization

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.256897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.396381Z digest=sha256:5bdc47998f2354f380f19b0e614733d4b529df95598ad8609a2e3ced234cabe8

Observation 05f4c639-5402-4d0c-9e39-eb0f1b9d5967 · outbound

This paper cites Safe reinforcement learning via shielding, 2017.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Safe reinforcement learning via shielding, 2017

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.243510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.401174Z digest=sha256:1871ae97f3f65f8673b0f47c1556c2011898b289ef002e4eb8dae0acefbc2a6d

Observation 1d364716-2cde-46eb-9fbd-d48e60177199 · outbound

This paper cites Asymptotic properties of constrained markov decision processes.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Asymptotic properties of constrained markov decision processes

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.229000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.405566Z digest=sha256:fa3bab1175be76e59cd6e1385f9f044b967749099cd4d02c76b11d05e4db6d7d

Observation 4c69894e-990d-4215-be53-24a5235ab0b3 · outbound

This paper cites Deep reinforcement learning for demand response in distribution networks.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Deep reinforcement learning for demand response in distribution networks

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.215622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.410077Z digest=sha256:82e7a857319a68b05e96fe96d3f97ebce7642bf765cc3ad54141a7c5f546e4d7

Observation 7e3fd9b3-e9a8-407e-a3c0-638221a2a80f · outbound

This paper cites Dynamic allocations for multi-product distribution.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Dynamic allocations for multi-product distribution

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.202202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.414520Z digest=sha256:a9af42051fd8830138c59095a5ee175efc0285b6be76249567da8dd34e05e732

Observation ba30aaeb-4974-4d4b-90e1-479690a6b11b · outbound

This paper cites Deliveries in an inventory/routing problem using stochastic dynamic programming.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Deliveries in an inventory/routing problem using stochastic dynamic programming

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.188578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.418841Z digest=sha256:cfe66bdd1baa2504c8212b62c87ef25bda473a68b81f17d3eaa33e1858ab7b63

Observation 0e48722e-60d5-4a7f-be11-af9e8544ea56 · outbound

This paper cites Resource constrained deep reinforcement learning.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Resource constrained deep reinforcement learning

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.175242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.423605Z digest=sha256:67614a5309ba4f7f8f0efd6053d949ff6d1584c8ea5841df03ed10069b06b394

Observation 7910ef68-b089-4069-b9cd-2697d104341e · outbound

This paper cites Duality between density function and value function with applications in constrained optimal control and Markov Decision Process.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Duality between density function and value function with applications in constrained optimal control and Markov Decision Process

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-08-15T20:01:08.759910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.427952Z digest=sha256:234338de0900d878c6a12f11bfb0101ac36fa46b0d4dfdab9d976597af737962

Observation 12714306-cb76-499b-9d89-0877225c6f86 · outbound

This paper cites A tutorial on kernel density estimation and recent advances.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning A tutorial on kernel density estimation and recent advances

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.432595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.432595Z digest=sha256:6357cc248bf8302caefe1e95bee34ad22115c6bd360ac95cd659e97e54bb881e

Observation 7995c69f-1974-48c2-9d0b-468dcc0d49a7 · outbound

This paper cites Supervised fuzzy reinforcement learning for robot navigation.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Supervised fuzzy reinforcement learning for robot navigation

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.153865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.436678Z digest=sha256:1c66a22bb4969cb9615d3cf3212a032fbbfcf7442e0a70ac8416d9da7d0050a0

Observation 41d4d7a9-e90b-488a-b1c9-a3cff33419b5 · outbound

This paper cites A comprehensive survey on safe reinforcement learning.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning A comprehensive survey on safe reinforcement learning

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.142245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.440861Z digest=sha256:8aa793aa4f89a9d0cddc47de68e97dc83b6b2545e618754f0bfd6dee9acab4dd

Observation ba56ffb1-40e2-42e6-9208-98fb0ce15ca4 · outbound

This paper cites Fuzzy q-learning.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Fuzzy q-learning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.130738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.445200Z digest=sha256:0ca88872cd4dfe74da3022f20fb6c68c41042bb5b7e149ddaee626c14eecd1df

Observation de379194-4561-4849-b674-b63d3bf68da1 · outbound

This paper cites Deep reinforcement learning for robotic manipulation with asynchronous off-policy updates.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Deep reinforcement learning for robotic manipulation with asynchronous off-policy updates

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.119336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.449306Z digest=sha256:bf99789b2e8e8b60b52e28b461e641aa86821caad5003f80ae72d624d0c39078

Observation 14b093ad-91a2-433a-9ab2-e4d41e255575 · outbound

This paper cites A Review of Safe Reinforcement Learning: Methods, Theory and Applications.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning A Review of Safe Reinforcement Learning: Methods, Theory and Applications

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.453440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.453440Z digest=sha256:2405952d7a25619709ee356dced80749021cf6c441b9ab997751f8ffb5e0b16d

Observation 88062e48-5473-4e2f-9530-f51adaf64026 · outbound

This paper cites Learning to Walk in the Real World with Minimal Human Effort.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Learning to Walk in the Real World with Minimal Human Effort

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.458090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.458090Z digest=sha256:80c9737f830a511cd6fbb598f087c49a28d731b52dd9c5ad48669346860335a7

Observation e3590007-a40a-422f-b7af-b26cb61e6b52 · outbound

This paper cites Hierarchical reinforcement learning for scarce medical resource allocation with imperfect information.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Hierarchical reinforcement learning for scarce medical resource allocation with imperfect information

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.107394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.462602Z digest=sha256:fe927316e28ea9f05c57907dd093ef5373c4890c7f0a70624962517691c57551

Observation 5e5fb747-f99f-4985-bc58-65a0d267616a · outbound

This paper cites Logically-Constrained Reinforcement Learning.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Logically-Constrained Reinforcement Learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.466900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.466900Z digest=sha256:00ece10fc5f164075bb48c3e7b79b2a8a05b128f0de428a86f5a7525c91ee280

Observation 82d539d9-c5d4-48c0-8671-88ab9ef72d1a · outbound

This paper cites Deep reinforcement learning with temporal logics.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Deep reinforcement learning with temporal logics

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.094339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.471573Z digest=sha256:41ed9faa054d5d457a646bcfaa0eb4974b4bf97e94bead4cc15ad3f3f752217b

Observation 89b97076-1a54-444e-a2d5-24564f2123b0 · outbound

This paper cites Achieving sustainable supply chains through energy justice.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Achieving sustainable supply chains through energy justice

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.081515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.475627Z digest=sha256:b554c2b6b07de94ea3aab2259f9f08ef7b36eb5236f76dc74485e1a40954a7ca

Observation 202b9f2b-0152-4143-9156-db1170c6accc · outbound

This paper cites Line: Logical query reasoning over hierarchical knowledge graphs.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Line: Logical query reasoning over hierarchical knowledge graphs

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.068577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.479956Z digest=sha256:f2b4eba190ecffbda10c8b4f1cfe70f00644751f933a46fe71a3f62c041d206b

Observation 8cdda88b-b8b6-45c6-9a1a-92f9d296e6da · outbound

This paper cites America's strategy to secure the supply chain for a robust clean energy transition.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning America's strategy to secure the supply chain for a robust clean energy transition

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.055482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.484102Z digest=sha256:d99b2379c9ac1090bf84c408376b0d4697158d9a583c6627dd3a9a1cd315102d

Observation 2084936b-8d76-41c1-99e3-3d9a89687a1e · outbound

This paper cites Community-based operations research.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Community-based operations research

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.042539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.488583Z digest=sha256:14dd8566ae9c1fc4c0983b132c85046957b5c5487caababb3ff3524280210f11

Observation 302212df-a0a1-43b2-b284-e203a09e1a88 · outbound

This paper cites On bayesian index policies for sequential resource allocation.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning On bayesian index policies for sequential resource allocation

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.029257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.492684Z digest=sha256:7c0b67258745df1a4fc9077644d52701c3de159a8de0f7a260740d1a8dd9b88e

Observation 578cea17-3554-4b64-aeaa-41ceb3b279b5 · outbound

This paper cites Between steps: Intermediate relaxations between big-m and convex hull formulations.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Between steps: Intermediate relaxations between big-m and convex hull formulations

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.016571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.496825Z digest=sha256:edc6baf6841ebc80e7f55a2d04484a8627c7b2f8afed9290f4eecd0321b3e59b

Observation c825ee03-52fb-44b7-bb5f-73b2dfd2e703 · outbound

This paper cites Augmenting Neural Networks with First-order Logic.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Augmenting Neural Networks with First-order Logic

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.500522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.500522Z digest=sha256:41144d4924453507df316922ef6dc030b766c83796a2be84f519aee0c943ccc5

Observation 6264bf78-637d-498b-91e9-5fc8dca917e0 · outbound

This paper cites Deep Reinforcement Learning for Efficient and Fair Allocation of Health Care Resources.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Deep Reinforcement Learning for Efficient and Fair Allocation of Health Care Resources

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.504544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.504544Z digest=sha256:4b8b33ac1a63762219dbeea4a3480dcfb0a399cad5cab2b7838723c768e3cf44

Observation a607d748-7a27-4367-9403-1aa91efb3936 · outbound

This paper cites Sequential resource allocation for nonprofit operations.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Sequential resource allocation for nonprofit operations

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:09.002610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.508841Z digest=sha256:3e36ecad7a2f484f0a2a3ce09bb48850f5ae2eb4227ab0a8bbd194f6d1391bce

Observation de61594b-3e94-43e1-b4d1-acbcbb0dfcc0 · outbound

This paper cites Clara: A constrained reinforcement learning based resource allocation framework for network slicing.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Clara: A constrained reinforcement learning based resource allocation framework for network slicing

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.990043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.513144Z digest=sha256:b8cea41811b921d8d0ef37144a64a2341f149169e8441a46b304b7ae6f53a870

Observation e973397e-f8a5-414d-9a06-5aebaba24179 · outbound

This paper cites Adaptive sequential surveillance with network and temporal dependence.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Adaptive sequential surveillance with network and temporal dependence

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.978102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.517342Z digest=sha256:b16120e241c6cddd39d79b4c8c1f39fe7d7ee7f7aa1fa6997c44486ac262d91c

Observation 80454cf3-ccc6-4a8d-b12b-b88fdd8c7b20 · outbound

This paper cites Ethical resource allocation in policing: Why policing requires a different approach from healthcare.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Ethical resource allocation in policing: Why policing requires a different approach from healthcare

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.966400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.521491Z digest=sha256:b302a9639b64f26f75a57dc17a70c5a2ce226a639bcbc3da00e29c1ec41130ae

Observation acc5a0ff-d3ef-48f5-ba61-29a15f77d610 · outbound

This paper cites A primal dual formulation for deep learning with constraints.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning A primal dual formulation for deep learning with constraints

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.954186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.525778Z digest=sha256:be5ac151a27798219dc286998c11476768a6ec06dd83df5d1c06958424cb2ced

Observation b7567450-0cc4-47cc-a7f3-979a69769425 · outbound

This paper cites Deep reinforcement learning approach for capacitated supply chain optimization under demand uncertainty.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Deep reinforcement learning approach for capacitated supply chain optimization under demand uncertainty

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.942113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.529994Z digest=sha256:f880f1b56c40c56022370a5e7efb3cc0a725c931887dc0693b64094030a48a57

Observation 92c40bbb-04a0-47f4-9bd7-c59d718e438c · outbound

This paper cites Combining fuzzy logic and reinforcement learning for resource management in edge computing.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Combining fuzzy logic and reinforcement learning for resource management in edge computing

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.928679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.534280Z digest=sha256:7d49f77d1f7ce3ff2e08cfab7780f98ec95dfa7532a29216077549a16c9eccbb

Observation 1dfc4132-eedb-4c26-8214-c3b69eec8098 · outbound

This paper cites Fairness of the distribution of public medical and health resources.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Fairness of the distribution of public medical and health resources

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.915694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.538381Z digest=sha256:a203cb75aae7f018b63b4cb7ac01c34a036b63585b272817dc8345f8c523a453

Observation 27eb12f5-e496-440d-a687-bebfd19efc97 · outbound

This paper cites Density constrained reinforcement learning.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Density constrained reinforcement learning

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.902621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.542300Z digest=sha256:c11214d94f3a7e1f07886084324ec1e1fa2c2a0cdef92d61526b5328b3d3935d

Observation a0b9784c-736f-40ce-88d7-21c4d9e15602 · outbound

This paper cites A dual to lyapunov's stability theorem.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning A dual to lyapunov's stability theorem

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.889428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.547582Z digest=sha256:91e194d6c14bd5f582d9ac6d477dd7c8ffc5d72dd820507859d60555c48a8ff7

Observation 38398143-cf84-4274-b7c1-9790afab83c2 · outbound

This paper cites Benchmarking Safe Exploration in Deep Reinforcement Learning.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Benchmarking Safe Exploration in Deep Reinforcement Learning

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.876651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.551759Z digest=sha256:d4e666168bd0948bfb78320fe17a0f438fa3dc7afcb9342b4456524558eecd49

Observation 2f1b0077-17c1-4454-9540-4155843d11e5 · outbound

This paper cites Query2box: Reasoning over Knowledge Graphs in Vector Space using Box Embeddings.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Query2box: Reasoning over Knowledge Graphs in Vector Space using Box Embeddings

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.555712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.555712Z digest=sha256:3c5543fb70415902224b6fb5ceaf450efc8562f17916d807210af2095339ad23

Observation 133b02cf-616d-496b-90e0-9e10b66e26d4 · outbound

This paper cites Apprenticeship learning using linear programming.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Apprenticeship learning using linear programming

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.863196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.560069Z digest=sha256:067a7e800a0a4f2f938d69bcf1fb56ac5c68d600f845526ab4a97d26e2175a8c

Observation c9d9f0ca-6b92-4be6-b98f-a1c9cd5c8665 · outbound

This paper cites The impacts of the covid-19 traffic light system on staff in tertiary education in new zealand.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning The impacts of the covid-19 traffic light system on staff in tertiary education in new zealand

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.848622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.564114Z digest=sha256:4f9607095907286241a3b71fa33f49d44778547bbd87133ae6a3bcb922749585

Observation fe46283b-7b35-440a-9ffd-8a00c8395dcc · outbound

This paper cites Reward Constrained Policy Optimization.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Reward Constrained Policy Optimization

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.568016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.568016Z digest=sha256:154632aff1b8f8bf9ce714661b367ff3dc3428a887a83c8df3804429851603e8

Observation a1bfcaee-d873-4fde-ad75-fd784cc5df72 · outbound

This paper cites Improved big-m reformulation for generalized disjunctive programs.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Improved big-m reformulation for generalized disjunctive programs

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.833735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.572065Z digest=sha256:2a5606839fc245fb0710749ed0fc66793ab4c281d19cb66ebc9c0460c076fb9b

Observation 9222da3e-c6b5-4756-a7ef-3f0fdf7b288c · outbound

This paper cites Disjunctive programming techniques for the optimization of process systems with discontinuous investment costs- multiple size regions.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Disjunctive programming techniques for the optimization of process systems with discontinuous investment costs- multiple size regions

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.820024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.576006Z digest=sha256:be36189c6bb304b8024d733c2ecec91465dbcdfe500f38f6fe1951adc3e39279

Observation ea06e07e-2147-4ac1-b132-57ad133fc53d · outbound

This paper cites Dynamic shielding for reinforcement learning in black-box environments, 2022.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Dynamic shielding for reinforcement learning in black-box environments, 2022

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.806624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.580100Z digest=sha256:f2bf43b71ddcb78a06dcae5a9225961999b93ba48acb0e49c028c3ff673d7204

Observation d1e43745-4fb0-4697-8826-476cac56d1cc · outbound

This paper cites Off-Policy Primal-Dual Safe Reinforcement Learning.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Off-Policy Primal-Dual Safe Reinforcement Learning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.584153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.584153Z digest=sha256:195f591d4395dd89b1779fc566928d171972b7609d6644e89e877e9eabfd63fc

Observation 311564d5-b7e6-4960-aa54-f66b851479f6 · outbound

This paper cites Projection-Based Constrained Policy Optimization.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Projection-Based Constrained Policy Optimization

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.588304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.588304Z digest=sha256:67a2e5aa1ddee0b577e6c58f42f015d725575088de824573ac0eaef0043cc792

Observation eb4687ca-97ae-4ace-86ec-c1b1c18c5817 · outbound

This paper cites Learning density-based correlated equilibria for markov games.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Learning density-based correlated equilibria for markov games

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.794529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.592377Z digest=sha256:114fd330bdba712f2f3f32b4e873545a3b0ff15103c619c41d3645d404b34399

Observation 96cf58b7-de99-4668-9312-c989ccffde74 · outbound

This paper cites The energy injustice of hydropower: Development, resettlement, and social exclusion at the hongjiang and wanmipo hydropower stations in china.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning The energy injustice of hydropower: Development, resettlement, and social exclusion at the hongjiang and wanmipo hydropower stations in china

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:01:08.782022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T20:01:08.596286Z digest=sha256:36de585887f85033936e05242e3db1fc02b2b7e963e4a00ca9f381c11393f3a5

Observation 824a5941-4552-4a73-bb4b-6e728d9cc9b9 · outbound

This paper cites write newline.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning write newline

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.600180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.600180Z digest=sha256:b37ec337b5428f774d6f0d1bfe4db1b4bde9158cf3c112e90887a275b570b208

Pith citing papers

No inbound Pith citation observations are available.