Pith. sign in

Paper Citation Record · LEDGER

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals

As of 9 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 0 inbound Pith citation observations for arXiv:2505.23124.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.23124 v2

Coverage vector

measured 34 of 34 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:03:58.099350Z

measured 34 of 34 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

34 of 34 outbound references displayed

  • verified exact3
  • verified fuzzy17
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 42a2f4df-6928-40ea-ac1f-51e7a0af0e8f · outbound

This paper cites write newline.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:55.250411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:03:55.250411Z digest=sha256:a3a2a805f6b29abab799103b7355c3d212534adc989b4de87faaea13712a2065

Observation a1866754-f6e1-4bcd-b6e2-16da248e316a · outbound

This paper cites Nearly-Optimal Bandit Learning in Stackelberg Games with Side Information.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Nearly-Optimal Bandit Learning in Stackelberg Games with Side Information

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:55.339861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:03:55.339861Z digest=sha256:5a8e3d33a97e7621aa5f7ad9774a2f84d1f160091b76eaeca90a57d6bbd0829f

Observation 85dd2d88-cd1c-4e5f-83dc-4d8ff2ba6a96 · outbound

This paper cites Principal-agent reward shaping in mdps.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Principal-agent reward shaping in mdps

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:02.419844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:03:55.435050Z digest=sha256:6f0555cce0db7e82c7de88a16e2c9bbcc01b4f7be319c47736e9cbbef7558b79

Observation 31ec152e-7f95-4de3-9410-d8b0dcc2f3ce · outbound

This paper cites Optimal rates and efficient algorithms for online bayesian persuasion.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Optimal rates and efficient algorithms for online bayesian persuasion

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:02.249370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:03:55.503853Z digest=sha256:032224e03bc7ec12a3b7fa0c762594d34fea58db9b53559544cf8143658d6d97

Observation e75ba6e5-2a03-43fd-81b8-e31b967ab5eb · outbound

This paper cites and Dewatripont, M.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals and Dewatripont, M

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:02.040640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:03:55.582634Z digest=sha256:83b9fb3e3f5b16f374171c1bbf08a63421b66702e9d17c848f8333b2b047f0f7

Observation e656ab1b-c7b1-4b92-af6d-9f8fb6360012 · outbound

This paper cites Robustness and linear contracts.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Robustness and linear contracts

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:01.815171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:03:55.666991Z digest=sha256:d34a6876267a35ce207c5ed2f81a776ce714872c4a013cad0f9817864a01cb7f

Observation deeec76f-974e-49f7-bb12-1389c2d328c4 · outbound

This paper cites Learning to Price Homogeneous Data.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Learning to Price Homogeneous Data

Reference 7

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T13:03:58.698148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:03:55.758735Z digest=sha256:82b6af4eb6ce146a84b07e85055533dff813cedbbc2fc9524aa4f95b5e031e59

Observation e6bfac6a-ee86-4ec2-b333-2015fa4f6c22 · outbound

This paper cites Estimating and Incentivizing Imperfect-Knowledge Agents with Hidden Rewards.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Estimating and Incentivizing Imperfect-Knowledge Agents with Hidden Rewards

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:55.914058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:03:55.914058Z digest=sha256:3d442c7ee0f3a2fc7755d4ac66b074fd8aae8510089e8bd438f6da4e9eb9dd1e

Observation 8535712b-1ea0-4538-81e9-31e4fd8f4e20 · outbound

This paper cites Repeated Principal-Agent Games with Unobserved Agent Rewards and Perfect-Knowledge Agents.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Repeated Principal-Agent Games with Unobserved Agent Rewards and Perfect-Knowledge Agents

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:55.983445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:03:55.983445Z digest=sha256:6f2fbb2bd582ee89cdc0f77d07e5289e5fdcca1a13fb7be4d660cd74348fa76d

Observation 8e4ca8b3-98c5-4937-a88c-dd3c77d65b30 · outbound

This paper cites Multi-agent combinatorial contracts.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Multi-agent combinatorial contracts

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:01.591506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:03:56.039245Z digest=sha256:bf6a36ad56c0e939fe04bee35116ad865e584829ed8cff237bc2efd138e86dd6

Observation 16e8b822-2992-4ad2-8965-c4a651ab64e9 · outbound

This paper cites Simple versus optimal contracts.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Simple versus optimal contracts

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:01.376280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:03:56.146040Z digest=sha256:00c7c13bd46adc2564f006a819681f7f3d4d3e886d60471f31475c92cd0734ce

Observation e09c02dd-e28a-4d9f-beaf-7da48169562b · outbound

This paper cites Combinatorial contracts.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Combinatorial contracts

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:01.119324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:03:56.272790Z digest=sha256:029b634af655a7093ec500a79bcf66d4a8ba803afa284696df1c1d572b179552

Observation 0237acb8-d9f4-451f-a689-4ecfe3e760c8 · outbound

This paper cites Multi-agent contracts.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Multi-agent contracts

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:00.898863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:03:56.341986Z digest=sha256:4755fddcda53a56cb03203ab1794c037f3fc4eb39a34ff28d159ef8f9d2288c0

Observation 80fd2c91-5f50-47ce-8596-18b9156c8bfc · outbound

This paper cites an unresolved cited work.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:04:00.672928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:03:56.427122Z digest=sha256:f9f567598f89db5571bfac68aa9d56486361d0e028b8a32ef7ee6b855b1cda36

Observation 4551f37c-0668-4da8-8dd6-4f47d3a50c72 · outbound

This paper cites and Miller, R.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals and Miller, R

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:00.460405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:03:56.556800Z digest=sha256:dd2e7af345e142e25b51be8fe2795703a8cde93e102c5f6368656738aad8ae8a

Observation 35e206c3-a181-450d-b2ad-6d617ab3d55a · outbound

This paper cites an unresolved cited work.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:04:00.292546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:03:56.620275Z digest=sha256:c8c4f18cf35492d885ba309dbd885dd93bdb51fb838671c1620da2e10f2b770f

Observation f857fabb-98e6-4fb2-ba7b-a98ad9c76a32 · outbound

This paper cites Z., and Balcan, M.-F.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Z., and Balcan, M.-F

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:00.108777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:03:56.703746Z digest=sha256:93babffd52a44e6059b05c46df2323cfe778a7805a6faee728d00fd0b5cf49c7

Observation 8d7285fc-1858-434a-a66d-ee1e967d5c94 · outbound

This paper cites an unresolved cited work.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:03:59.945634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:03:56.761059Z digest=sha256:ae8bfc1427aaa0b022251be4f32a7edf32067c549254bfffdf0f85e49a9be325

Observation f64ebdb7-15f4-40aa-80a5-bd24a182ceb1 · outbound

This paper cites Moral hazard and observability.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Moral hazard and observability

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:59.744244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:03:56.862182Z digest=sha256:1c4f773a02706a148f724e2265c07f03b8ad832a5a0b9710624ef949625f15c1

Observation cdc8f525-ca14-423d-ab26-0679c422614c · outbound

This paper cites J., Rand, D.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals J., Rand, D

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:59.634972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:03:56.943018Z digest=sha256:b56a5469150f0b5b94686308e735291f2c2e6579a9ca672cc4939df7551a46f0

Observation 03d54429-0d98-4eb9-9244-45a7f9a2108d · outbound

This paper cites and Siddiq, A.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals and Siddiq, A

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:59.465278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:03:57.022280Z digest=sha256:c37f4a3359321de1a96fef5615f76fd549519679e6e10146867aea1ad5f3c604

Observation 653aad28-cde3-4c75-aac6-2d938902be0c · outbound

This paper cites and Szepesv \'a ri, C.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals and Szepesv \'a ri, C

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:57.089780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:03:57.089780Z digest=sha256:b4e120071f63e6ff17ba3b097423899b961cbc5cecfcb4615c7eea0947429e1a

Observation af64217d-881b-43b3-a166-02718cfa7c29 · outbound

This paper cites Principal-Agent Bandit Games with Self-Interested and Exploratory Learning Agents.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Principal-Agent Bandit Games with Self-Interested and Exploratory Learning Agents

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:03:58.582546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:03:57.161602Z digest=sha256:2690f3b295b65fba146486115d68d3b808eb46a0c3b19551f762c565f2cedf2c

Observation ea1f350f-e5be-44dd-9dd4-09217f4cc620 · outbound

This paper cites On the Limitations and Possibilities of Nash Regret Minimization in Zero-Sum Matrix Games under Noisy Feedback.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals On the Limitations and Possibilities of Nash Regret Minimization in Zero-Sum Matrix Games under Noisy Feedback

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:57.238622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:03:57.238622Z digest=sha256:3ea9be1ab999a89e61c4f30d2d46965f8685bd538fe90131879dcbb2b9d65a27

Observation 7419b645-f73e-46af-a77b-5d6f5fff10c8 · outbound

This paper cites Efficient Near-Optimal Algorithm for Online Shortest Paths in Directed Acyclic Graphs with Bandit Feedback Against Adaptive Adversaries.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Efficient Near-Optimal Algorithm for Online Shortest Paths in Directed Acyclic Graphs with Bandit Feedback Against Adaptive Adversaries

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:57.337942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:03:57.337942Z digest=sha256:01de7c05cbbbed21320606044a07a843b227312f010faa0dc1c92e1659fdc8db

Observation 6c3b199b-d782-4d64-97a5-f9ccfbeb5c11 · outbound

This paper cites and Nair, H.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals and Nair, H

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:59.332062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:03:57.446306Z digest=sha256:c08ce34561b15a751e0561fe21d6732464e93ef7d8c2f9c68be480409d9a2596

Observation 121ef90f-431f-43c6-b3b1-cdb6fade7362 · outbound

This paper cites T., and Narasimhan, C.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals T., and Narasimhan, C

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:59.209606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:03:57.504728Z digest=sha256:5093fa8b8ca58f6e6c88981223f7e6f91e8051efbabc710c675cb377c2a55826

Observation 37e4a06d-bbbf-426b-8140-18d4a4623f49 · outbound

This paper cites and Slivkins, A.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals and Slivkins, A

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:59.098326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:03:57.608075Z digest=sha256:aacfa2d4e9eb44f641c013d03a81fa745ad25b345cb9395ab055f4749a2c062b

Observation b49c2a9e-a8ba-4958-a2e2-8ebcfc773255 · outbound

This paper cites Learning to Mitigate Externalities: the Coase Theorem with Hindsight Rationality.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Learning to Mitigate Externalities: the Coase Theorem with Hindsight Rationality

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:03:58.428319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:03:57.665267Z digest=sha256:a3fbf735f2dbd3a817ebc295d58bea1ad7e9a41e69593c7d1ebbd700debf96e2

Observation 62bd797d-15ec-4e52-ab61-0ffe285344b3 · outbound

This paper cites Incentivized Learning in Principal-Agent Bandit Games.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Incentivized Learning in Principal-Agent Bandit Games

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:03:58.299182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:03:57.746639Z digest=sha256:31e3106acc7567183cf54fb76d944b6a6bfc99750742c47637d9b9ea071b5830

Observation b4cd5186-42a8-4770-86eb-aa33b367ed34 · outbound

This paper cites an unresolved cited work.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:03:58.990757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:03:57.842207Z digest=sha256:df0db089f5514a064d53e1f831f4b03de002706dd8ded2c57a2d59e06ed2b0d6

Observation 46d66d4f-ee05-4395-9569-8465266007b3 · outbound

This paper cites Contractual Reinforcement Learning: Pulling Arms with Invisible Hands.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Contractual Reinforcement Learning: Pulling Arms with Invisible Hands

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:57.940291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:03:57.940291Z digest=sha256:e68b7a6527c071e7934bfd5dacc5ee4fcff70ad59e8912ce218a484ba7f3e27c

Observation a6b3d551-4122-4822-b98e-5645994c4db2 · outbound

This paper cites The Sample Complexity of Online Contract Design.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals The Sample Complexity of Online Contract Design

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:58.056440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:03:58.056440Z digest=sha256:b279690c8e391f38c9d38a1d8d59c2c2783e5f718f3b9c0195f179e6b83365ef

Observation ecfdb087-a3f5-4bbc-b31f-7059bb66d8bb · outbound

This paper cites and Seldin, Y.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals and Seldin, Y

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:58.881888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:03:58.099350Z digest=sha256:9be7338bcd664c89ab693ca1e45dad156083033db67bfef02d88fd8677d79443

Pith citing papers

No inbound Pith citation observations are available.