Pith. sign in

Paper Citation Record · LEDGER

Fairness Aware Reinforcement Learning via Proximal Policy Optimization

As of 13 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 1 inbound Pith citation observation for arXiv:2502.03953.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.03953 v2

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T00:13:05.031194Z

measured 50 of 50 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-17T20:18:09.847453Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-17T20:20:11.884602Z

Reference resolution

49 of 49 outbound references displayed

  • verified exact4
  • verified fuzzy6
  • unresolved38
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7597e3d7-fef0-43f3-9430-07525c09888f · outbound

This paper cites , " * write output.state after.block = add.period write newline.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization , " * write output.state after.block = add.period write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-09T00:13:04.509708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T00:13:04.509708Z digest=sha256:e33198bb63b849df77e9010ef18ac13aab801cd7968651cf955115c65ce0a5ed

Observation 1b39a62c-c7b0-48fd-a713-1c07ca7aff02 · outbound

This paper cites write newline.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-09T00:13:04.536407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T00:13:04.536407Z digest=sha256:3e09ae744b6edcdf4b85b8a7b785c79fa87216f5d5ff48f7391562f1a131bf7c

Observation 508edbb0-ab7f-4e53-ae51-068116ce3f63 · outbound

This paper cites an unresolved cited work.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:13:05.599971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:04.581538Z digest=sha256:8901bde7ecafec4db3f9fdbdd2fa26ffa0582c4aa7d0dfd18399691e5a9e2052

Observation 8f65295e-7d16-41b9-bdf0-d7a0d1fbd831 · outbound

This paper cites an unresolved cited work.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:13:05.590380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:04.608274Z digest=sha256:fe2bc97f6f5ac66d615fd7b7671bfb52467c260869f896fbab8fe771ab51dbfd

Observation f3420c49-b8c2-4d55-b441-59691b39884f · outbound

This paper cites an unresolved cited work.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:13:05.580453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:04.635898Z digest=sha256:cc91c9d35148e597d6af62cf367e018e5035aeb572223acf493e3fd7eeea07c0

Observation 47cd44f1-fd6a-48ef-9964-5c73a1448f9f · outbound

This paper cites an unresolved cited work.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:13:05.570238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:04.639856Z digest=sha256:a0113ed41de486471e4fab3e1db8b774976ddc1c78b77772275b3d13a8130c6e

Observation 183f86e7-f135-4a1b-9280-cc8a8ba06751 · outbound

This paper cites Nash Social Welfare Approximation for Strategic Agents.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Nash Social Welfare Approximation for Strategic Agents

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-08-09T00:13:05.185852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:04.643745Z digest=sha256:ee699424ef788f73dd6882c91b6fbcd9ee69645779194970e6063dd31824540d

Observation 0a4dca87-eb4a-4c47-a646-21d6d88bf00b · outbound

This paper cites an unresolved cited work.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:13:05.560018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:04.647952Z digest=sha256:a447de3b9e549173b17d5a5d0cc3d875c2989b0de843f8f7dd58a4d3bedf3817

Observation d4cb1313-22e5-4e6a-a3bc-723998701963 · outbound

This paper cites an unresolved cited work.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:13:05.549627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:04.651475Z digest=sha256:d939829ece2f65e9f19b9b0433f08173029a24126a85f13c7663c7a4ee350c1b

Observation 886acf7d-8c19-493f-b847-5e762a01c86d · outbound

This paper cites an unresolved cited work.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:13:05.538864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:04.655207Z digest=sha256:ad0c3802f720ee67f25881294b7b54864eadc30f4f52af3f0f0d6287b57d17e5

Observation 8e220efc-ca9d-4f2f-b160-5cd639fc6342 · outbound

This paper cites an unresolved cited work.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:13:05.529142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:04.658730Z digest=sha256:33b0b723cb2963e80fbb9b6e2e4c125868764c39a53bd64fbd6a42156e4ae76d

Observation 3ee215b1-f047-491b-96d4-ef0dc7600734 · outbound

This paper cites an unresolved cited work.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:13:05.518759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:04.662813Z digest=sha256:c8c32fa740dbbf4b2b217a155206e8df4621851e2e41ce9fdeddd4285554e8a3

Observation 29b71bbe-443a-4523-a24f-dbd01b2acc40 · outbound

This paper cites Fair prediction with disparate impact: A study of bias in recidivism prediction instruments.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Fair prediction with disparate impact: A study of bias in recidivism prediction instruments

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-09T00:13:04.666275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T00:13:04.666275Z digest=sha256:da52d175741d34243be1983eb4f30c812721c83e0dd55fdb5a6e7faae406700a

Observation 8c3b6044-3f91-4539-9314-0999155f23ed · outbound

This paper cites an unresolved cited work.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:13:05.508381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:04.669849Z digest=sha256:399b57bfb5c88a8034a3d927899d05d993c8fb7bccf25840008280e5e051733a

Observation 7910e40e-45fb-42b9-9e46-5f1d495fbf73 · outbound

This paper cites an unresolved cited work.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:13:05.499152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:04.673386Z digest=sha256:790c2b680460b01d41086556c828b553452d3aa6e80ca6b2ceb50f3a09c612c2

Observation ea2f92e5-86d8-4594-aac7-1f8fb91cde44 · outbound

This paper cites an unresolved cited work.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:13:05.489757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:04.676463Z digest=sha256:c966ba2f9e21d20c7309dc376e52d76c8a34703ede0a2cc4148550026366e0df

Observation 3708b520-e229-46f9-8c9d-a537023f0080 · outbound

This paper cites W.; and Livingston Jr., J.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization W.; and Livingston Jr., J

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:13:05.480620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:04.679667Z digest=sha256:7ff0f6e96ff7a128a6bb3946285da9fb361940c32cfa2ed35b99ea9f588c1624

Observation a0b31bf6-c804-4ff7-aaf2-ad8cb10df59b · outbound

This paper cites an unresolved cited work.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:13:05.450228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:04.682829Z digest=sha256:0cc6525be1b327e7e5ae1f0e981180693f8db6e1c55cf2027c5bc58eb648019c

Observation 08f9265f-c87a-436e-9cb2-3fd59b8bc785 · outbound

This paper cites an unresolved cited work.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:13:05.436317Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:04.686098Z digest=sha256:f0bc664ef7bb113e3653af3552f148be30512c652275ba556590c698b155c6d8

Observation 2a98032c-1840-4697-830d-9528dd773592 · outbound

This paper cites Optimally Interpolating between Ex-Ante Fairness and Welfare.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Optimally Interpolating between Ex-Ante Fairness and Welfare

Reference 20

Resolution
metadata mismatch
local_arxiv, observed 2026-08-09T00:13:05.162244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:04.689565Z digest=sha256:dd031cdfdcb7fe87d6e2dd451946a987dde969662943d045b26064ebc2eb1c70

Observation fbef346e-264f-47bd-80bc-3c1606a7652c · outbound

This paper cites an unresolved cited work.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:13:05.426007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:04.693308Z digest=sha256:2a86de5e8c8989ecd96cb915e196aba7e2e88cd12be8b0281a60670ee2a340c5

Observation 262d7826-d633-4f65-bf7a-a46161dea322 · outbound

This paper cites an unresolved cited work.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:13:05.415287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:04.696442Z digest=sha256:151671c2363a78266b52c10254e21e1b30b6a7fde4c25406c775a88222d801c2

Observation 591858d6-ec90-428a-8346-1dd1757de137 · outbound

This paper cites an unresolved cited work.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:13:05.405046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:04.699796Z digest=sha256:edd0f4e155124f8ab3e73c5bf14599bab6cd39b237733685028c18d339908f22

Observation 803d01c5-db61-41e9-a73d-d4b657411c0e · outbound

This paper cites H.; and Roth, A.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization H.; and Roth, A

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:13:05.383677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:04.703550Z digest=sha256:254e7b6d5f153d1f4ec0911ca1714670e3e30445150d1c3100a7264319798784

Observation bb018550-560e-409e-ad43-8b5d5e30235f · outbound

This paper cites an unresolved cited work.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:13:05.370401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:04.706524Z digest=sha256:a21d0877538c0eca62324210dace9fae4907d84040961a2e22cc44c0d6a4c796

Observation 68144b3e-1dfc-45f7-b3fc-a35c78a6810e · outbound

This paper cites an unresolved cited work.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:13:05.360303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:04.709845Z digest=sha256:7baafbcebaa3d9dcefc6e3edd6f253f5ea69337b8ebd41bf85ac3a4a535e8054

Observation d57b8993-ea49-4889-b597-7354d1567394 · outbound

This paper cites an unresolved cited work.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:13:05.344719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:04.713327Z digest=sha256:e318f6d1e39a6ada71a5e7540bb4105bd0b3dc60b9fd1c534295cb332b65d7da

Observation 438f5e04-d6ed-40cc-8f2a-4f078ce75852 · outbound

This paper cites an unresolved cited work.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:13:05.334722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:04.717740Z digest=sha256:720aec637f7d35e85f1fa79c68e7a6415385e2e3987b8d193c90b23bd0f6f9c4

Observation 1def4555-5cbd-474e-96c0-49466a99b47d · outbound

This paper cites Counterfactual Fairness.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Counterfactual Fairness

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-09T00:13:04.721799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T00:13:04.721799Z digest=sha256:44aebf983911c619996212540d494b0892957c3f0836ff6ccaddb07728db43ca

Observation 0aaf7d0c-f1ed-42cb-86af-901a3920b12d · outbound

This paper cites Z.; Perolat, J.; Hughes, E.; and et al.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Z.; Perolat, J.; Hughes, E.; and et al

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:13:05.325295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:04.754547Z digest=sha256:a39e06e87fbfd5620c3d5a91f0f8c970d9fad8842d37adb3326ab582333cb262

Observation f8336578-1446-40da-a104-aee51eee71a9 · outbound

This paper cites an unresolved cited work.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:13:05.316266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:04.834950Z digest=sha256:c85f5221df7fcb8b383607c2a250140b8e658d07f2d1675d8cf1666207f667b7

Observation 2dd158f2-9a01-4611-9001-00ff0a89fddf · outbound

This paper cites an unresolved cited work.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:13:05.307187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:04.902048Z digest=sha256:6afcdcf709db917366e059191b6ab3e8126216b4faa8bf4c451e628603d67966

Observation 7a6d240c-09e2-4fe7-8a7e-077884603a48 · outbound

This paper cites an unresolved cited work.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:13:05.297526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:04.973164Z digest=sha256:a839478a03aaf0aeabdcb643ee71799c9fd3f27ed95e4ede80f23d189a01eec4

Observation bc8959ff-6d59-41f5-a1f8-aafb0082366d · outbound

This paper cites J.; Markakis, E.; Mossel, E.; and Saberi, A.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization J.; Markakis, E.; Mossel, E.; and Saberi, A

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:13:05.286556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:04.976308Z digest=sha256:3aa0822fd47f8badd6212289d3f1e74d6cd1e804968bb1a305306a09b3b2931e

Observation 54abf6cd-5086-490d-bde9-e34d55c28fbf · outbound

This paper cites Calibrated Fairness in Bandits.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Calibrated Fairness in Bandits

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-09T00:13:04.979176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T00:13:04.979176Z digest=sha256:967964a6f51cd0cdad5891a2e4db856669050f56ec6ef1b3043917dcd981b4e4

Observation d33d5e36-f48f-46df-9401-7dd9e2063cba · outbound

This paper cites an unresolved cited work.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:13:05.276562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:04.982595Z digest=sha256:262c2805f84a90be08f3a4380c5b39565345f3d9abdd89bccf64786313a22dbb

Observation 881ae253-2501-4fe6-8164-55382b2ef5b8 · outbound

This paper cites an unresolved cited work.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:13:05.266239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:04.985684Z digest=sha256:fb1f82c2c2508af7dcd2b7d7eac31a864f7bcd2d761a964d53ae9fdb13e28ab4

Observation 05a5baf9-692a-4457-bbaa-00d30fbd9296 · outbound

This paper cites an unresolved cited work.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:13:05.255522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:04.988874Z digest=sha256:74e739a955c72362a0cb8c26970a4c8b54a149a722d5bff68033bacbd2aee656

Observation a029b85f-4714-439d-8549-04b2987ada85 · outbound

This paper cites Fairness in Reinforcement Learning: A Survey.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Fairness in Reinforcement Learning: A Survey

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-08-09T00:13:05.127956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:04.991911Z digest=sha256:3f18322f665b3126ab9a834dd82e6e2ff360d6059a6df3088bf5e48212dddd2c

Observation 28546539-8cbe-4447-9a99-38343387c414 · outbound

This paper cites Trust Region Policy Optimization.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Trust Region Policy Optimization

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-09T00:13:04.996057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T00:13:04.996057Z digest=sha256:dba4c99a1fff14294d4605e72893622067010c725eb7e36702a95756f34c73fb

Observation a302619e-4928-4495-89c4-724093671588 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Proximal Policy Optimization Algorithms

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-09T00:13:04.999624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T00:13:04.999624Z digest=sha256:8668ac945a31e7fd864217ad4723ea21bc44ada15a28cfc184c6e0e5427cfc38

Observation b69f0019-07cf-488a-b7d0-5dc1cbf7a1f4 · outbound

This paper cites an unresolved cited work.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:13:05.244237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:05.003100Z digest=sha256:c686b68b7b73c225dd6e27dff4d7e7faf3fff9b8486e9fde84b1c662eef8cb9e

Observation 94a05665-28dd-4d56-aa2d-0704a66731da · outbound

This paper cites S.; and Barto, A.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization S.; and Barto, A

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:13:05.231786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:05.006349Z digest=sha256:ba45015e9da3515a2d0354900708b04469ac90b55db878f6f85bedc0d3fa77f0

Observation ca76107d-9aa8-441a-bd6e-577d8dccf96a · outbound

This paper cites A.; Eisenstein, L.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization A.; Eisenstein, L

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:13:05.219510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:05.010270Z digest=sha256:a0ac31d214e193cd8300809d180488133d6fa3b9aec47ebc28f8253c63ec4055

Observation 177bfb8f-973a-48b4-9fac-d8625bb50d52 · outbound

This paper cites Algorithms for Fairness in Sequential Decision Making.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Algorithms for Fairness in Sequential Decision Making

Reference 45

Resolution
verified exact
local_arxiv, observed 2026-08-09T00:13:05.094774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:05.013838Z digest=sha256:3dfdc6d3f09533467a2e08aca4a4d99147a3f97d1daa141836a030e92f5d495c

Observation 214cd1ec-bb67-418d-908e-c93526566bda · outbound

This paper cites an unresolved cited work.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:13:05.207432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:05.017739Z digest=sha256:30b9b9209287df7cd77b04133e211372726cfc3e8a1cfebafbac6caefe0b0faf

Observation a613fb61-476b-488e-a02c-87230457f205 · outbound

This paper cites an unresolved cited work.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:13:05.196659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:05.021664Z digest=sha256:f919154af49d868500ebe8591e489376ea26bda9f820ff11be32ab346bba4a31

Observation e4350450-d4ba-44ce-8319-6cf82725fe99 · outbound

This paper cites Penalized Proximal Policy Optimization for Safe Reinforcement Learning.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Penalized Proximal Policy Optimization for Safe Reinforcement Learning

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-09T00:13:05.026156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T00:13:05.026156Z digest=sha256:8c474efddbb79f5df67480a5d81ba69df3d01ec4ceeeb00fd3960afdff77e66b

Observation 1deed8e9-4385-4b60-94ac-f008b69173df · outbound

This paper cites Learning Fair Policies in Decentralized Cooperative Multi-Agent Reinforcement Learning.

Fairness Aware Reinforcement Learning via Proximal Policy Optimization Learning Fair Policies in Decentralized Cooperative Multi-Agent Reinforcement Learning

Reference 49

Resolution
verified exact
local_arxiv, observed 2026-08-09T00:13:05.069862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-09T00:13:05.031194Z digest=sha256:17a58bef22c4f7e738b6e8508307dbf2f1b16c57cf6b310ee410537fe9e3e9dc

Pith citing papers

Observation 209a152e-7232-46c1-a36a-bb8ea635365e · inbound

AdaFair-MARL: Enforcing Adaptive Fairness Constraints in Multi-Agent Reinforcement Learning cites this paper.

AdaFair-MARL: Enforcing Adaptive Fairness Constraints in Multi-Agent Reinforcement Learning Fairness Aware Reinforcement Learning via Proximal Policy Optimization

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-17T20:20:11.886396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-17T20:18:09.847453Z digest=sha256:0ba4d1a92f63e044844f8a4775d7b4f6b375c1a25463d008198505048b1ecaaf