Pith. sign in

Paper Citation Record · LEDGER

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals

As of 23 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 0 inbound Pith citation observations for arXiv:2505.23124.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.23124 v2

Coverage vector

measured 34 of 34 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:03:58.099350Z

measured 34 of 34 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

34 of 34 outbound references displayed

  • verified exact3
  • verified fuzzy17
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 42a2f4df-6928-40ea-ac1f-51e7a0af0e8f · outbound

This paper cites write newline.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:55.250411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:03:55.250411Z digest=sha256:5b205205cf6caefdfce0b36604582b885e3bb245295325a676ad3153bd0094ef

Observation a1866754-f6e1-4bcd-b6e2-16da248e316a · outbound

This paper cites Nearly-Optimal Bandit Learning in Stackelberg Games with Side Information.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Nearly-Optimal Bandit Learning in Stackelberg Games with Side Information

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:55.339861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:03:55.339861Z digest=sha256:5de196e5b4012aff0c802607c6f0669d5fbffed255b3cccd4df1c4a59735e121

Observation 85dd2d88-cd1c-4e5f-83dc-4d8ff2ba6a96 · outbound

This paper cites Principal-agent reward shaping in mdps.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Principal-agent reward shaping in mdps

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:02.419844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:03:55.435050Z digest=sha256:63c9f3e69e74aaec7bf2ffbef4bf83f0121417dbae612cbf5dba992604db5b31

Observation 31ec152e-7f95-4de3-9410-d8b0dcc2f3ce · outbound

This paper cites Optimal rates and efficient algorithms for online bayesian persuasion.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Optimal rates and efficient algorithms for online bayesian persuasion

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:02.249370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:03:55.503853Z digest=sha256:ed53b26fd1941a6c9f23f5fc4806cec5f2ecdbf9385709c7e359f8c737f57b0b

Observation e75ba6e5-2a03-43fd-81b8-e31b967ab5eb · outbound

This paper cites and Dewatripont, M.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals and Dewatripont, M

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:02.040640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:03:55.582634Z digest=sha256:4e2a15b469eebe7d8cb76c6df456b5d63f6e435b468d5ff8ddacc898905b5ed5

Observation e656ab1b-c7b1-4b92-af6d-9f8fb6360012 · outbound

This paper cites Robustness and linear contracts.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Robustness and linear contracts

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:01.815171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:03:55.666991Z digest=sha256:baaf9a6b4075b5b16b26fa85c56c4c01584f1cadd3ed403c75354d58834689e3

Observation deeec76f-974e-49f7-bb12-1389c2d328c4 · outbound

This paper cites Learning to Price Homogeneous Data.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Learning to Price Homogeneous Data

Reference 7

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T13:03:58.698148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:03:55.758735Z digest=sha256:bc516c5d727fdc82f1913feb661547598a4d9f6ada9935e74a5968a4ec7715a3

Observation e6bfac6a-ee86-4ec2-b333-2015fa4f6c22 · outbound

This paper cites Estimating and Incentivizing Imperfect-Knowledge Agents with Hidden Rewards.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Estimating and Incentivizing Imperfect-Knowledge Agents with Hidden Rewards

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:55.914058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:03:55.914058Z digest=sha256:57638689b99a06a97f2931d5ec477beeb18b4d3a4adfae277acbdd8082687fad

Observation 8535712b-1ea0-4538-81e9-31e4fd8f4e20 · outbound

This paper cites Repeated Principal-Agent Games with Unobserved Agent Rewards and Perfect-Knowledge Agents.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Repeated Principal-Agent Games with Unobserved Agent Rewards and Perfect-Knowledge Agents

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:55.983445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:03:55.983445Z digest=sha256:eb020a867ecc7c93166660823af9eaed2998e312ba9d9aaec6cf65e707c20572

Observation 8e4ca8b3-98c5-4937-a88c-dd3c77d65b30 · outbound

This paper cites Multi-agent combinatorial contracts.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Multi-agent combinatorial contracts

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:01.591506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:03:56.039245Z digest=sha256:70b927b3715993d63d28c529b4eade96aec1ab78cf3c54e9edd2b43326011793

Observation 16e8b822-2992-4ad2-8965-c4a651ab64e9 · outbound

This paper cites Simple versus optimal contracts.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Simple versus optimal contracts

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:01.376280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:03:56.146040Z digest=sha256:8847639d0409c1378fab312ccd1be39f51c81dc68b2ff06ae57121b943db8c39

Observation e09c02dd-e28a-4d9f-beaf-7da48169562b · outbound

This paper cites Combinatorial contracts.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Combinatorial contracts

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:01.119324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:03:56.272790Z digest=sha256:cfc208d91ef719414f04df4dcea6aee388832ea81ec4543dc48a2f8bb2bad593

Observation 0237acb8-d9f4-451f-a689-4ecfe3e760c8 · outbound

This paper cites Multi-agent contracts.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Multi-agent contracts

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:00.898863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:03:56.341986Z digest=sha256:bf64269b5dcb0ddce4d8cfa04bb7db1fe0eaa08a530a7951898a538f8e723161

Observation 80fd2c91-5f50-47ce-8596-18b9156c8bfc · outbound

This paper cites an unresolved cited work.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:04:00.672928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:03:56.427122Z digest=sha256:e00a80c5ae57b6b8b56f1e069e0d11e87ea94418be5bcce0d7ea415f19959395

Observation 4551f37c-0668-4da8-8dd6-4f47d3a50c72 · outbound

This paper cites and Miller, R.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals and Miller, R

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:00.460405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:03:56.556800Z digest=sha256:c75e05d2576194ce8f38f5aa3fcc64f45d01c1854f613679c661651f05a4e2fe

Observation 35e206c3-a181-450d-b2ad-6d617ab3d55a · outbound

This paper cites an unresolved cited work.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:04:00.292546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:03:56.620275Z digest=sha256:d8376cfdf1904471307659b86b55f4f217bc5f010a2847ed6f209b5077635db7

Observation f857fabb-98e6-4fb2-ba7b-a98ad9c76a32 · outbound

This paper cites Z., and Balcan, M.-F.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Z., and Balcan, M.-F

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:00.108777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:03:56.703746Z digest=sha256:eb3e73a8641c4059fdf767441115f4a9d0eb2e4d31103500b326ad52be0822fa

Observation 8d7285fc-1858-434a-a66d-ee1e967d5c94 · outbound

This paper cites an unresolved cited work.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:03:59.945634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:03:56.761059Z digest=sha256:cdbe9295500991e60e9ddc61ddef36d0809a9b0fe16fc8b462d475b8dadad376

Observation f64ebdb7-15f4-40aa-80a5-bd24a182ceb1 · outbound

This paper cites Moral hazard and observability.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Moral hazard and observability

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:59.744244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:03:56.862182Z digest=sha256:d10f982be6d702fd6f39b3668c165271e97a7a1f058d7efda27f02e6dc1d1e3d

Observation cdc8f525-ca14-423d-ab26-0679c422614c · outbound

This paper cites J., Rand, D.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals J., Rand, D

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:59.634972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:03:56.943018Z digest=sha256:e45491ac675808bf16738d7964e33c000f730302b8b2e6736d8e2b54085491b8

Observation 03d54429-0d98-4eb9-9244-45a7f9a2108d · outbound

This paper cites and Siddiq, A.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals and Siddiq, A

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:59.465278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:03:57.022280Z digest=sha256:dfeaa3776509a5f1e9dc6252b6febf54a95267eb33b6548488c0c6980cfa09a3

Observation 653aad28-cde3-4c75-aac6-2d938902be0c · outbound

This paper cites and Szepesv \'a ri, C.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals and Szepesv \'a ri, C

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:57.089780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:03:57.089780Z digest=sha256:193c82c243cbd7835fbefd2999458c37e5d473d4791e0100577bfbfa4a8e2d18

Observation af64217d-881b-43b3-a166-02718cfa7c29 · outbound

This paper cites Principal-Agent Bandit Games with Self-Interested and Exploratory Learning Agents.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Principal-Agent Bandit Games with Self-Interested and Exploratory Learning Agents

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:03:58.582546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:03:57.161602Z digest=sha256:b7f90d751891e83067f2a15b470526cc9b3bde8a9d0cf218c5e6a6207add734f

Observation ea1f350f-e5be-44dd-9dd4-09217f4cc620 · outbound

This paper cites On the Limitations and Possibilities of Nash Regret Minimization in Zero-Sum Matrix Games under Noisy Feedback.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals On the Limitations and Possibilities of Nash Regret Minimization in Zero-Sum Matrix Games under Noisy Feedback

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:57.238622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:03:57.238622Z digest=sha256:fb5a50f6c0f564398ca7a0f6979884d25322037b40d3486f98eb5326d5510948

Observation 7419b645-f73e-46af-a77b-5d6f5fff10c8 · outbound

This paper cites Efficient Near-Optimal Algorithm for Online Shortest Paths in Directed Acyclic Graphs with Bandit Feedback Against Adaptive Adversaries.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Efficient Near-Optimal Algorithm for Online Shortest Paths in Directed Acyclic Graphs with Bandit Feedback Against Adaptive Adversaries

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:57.337942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:03:57.337942Z digest=sha256:a241f9b2de2f123fdfbda0e68b96e912cbe20e5d21269db32b449c777eac146d

Observation 6c3b199b-d782-4d64-97a5-f9ccfbeb5c11 · outbound

This paper cites and Nair, H.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals and Nair, H

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:59.332062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:03:57.446306Z digest=sha256:b4ad3ba1ae2e2e7ab6bda4f0d4858fb60e06fd0a61bfc6ecacf937e307773481

Observation 121ef90f-431f-43c6-b3b1-cdb6fade7362 · outbound

This paper cites T., and Narasimhan, C.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals T., and Narasimhan, C

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:59.209606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:03:57.504728Z digest=sha256:d7bec1e074eb4aa0c21955db5b6f6dbb83438ab5246caed12102124eee9c8e6e

Observation 37e4a06d-bbbf-426b-8140-18d4a4623f49 · outbound

This paper cites and Slivkins, A.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals and Slivkins, A

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:59.098326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:03:57.608075Z digest=sha256:cd5719dc5da9f293c949fc037551ae01201229949dd93523490fdb5f1583be45

Observation b49c2a9e-a8ba-4958-a2e2-8ebcfc773255 · outbound

This paper cites Learning to Mitigate Externalities: the Coase Theorem with Hindsight Rationality.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Learning to Mitigate Externalities: the Coase Theorem with Hindsight Rationality

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:03:58.428319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:03:57.665267Z digest=sha256:34cc69ab2deba6650272926535f8054a4f030e0dddeeb956e698c9f31a816e4e

Observation 62bd797d-15ec-4e52-ab61-0ffe285344b3 · outbound

This paper cites Incentivized Learning in Principal-Agent Bandit Games.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Incentivized Learning in Principal-Agent Bandit Games

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:03:58.299182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:03:57.746639Z digest=sha256:6b0d304fba4ef2a39456ff1b2b1659d0cdfdd5492588bd6509777ddd28fe70bf

Observation b4cd5186-42a8-4770-86eb-aa33b367ed34 · outbound

This paper cites an unresolved cited work.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:03:58.990757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:03:57.842207Z digest=sha256:6d2288fb264543a36efea11dedf68d61ce41b8fe49bb7612181c0b1f3127bdea

Observation 46d66d4f-ee05-4395-9569-8465266007b3 · outbound

This paper cites Contractual Reinforcement Learning: Pulling Arms with Invisible Hands.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Contractual Reinforcement Learning: Pulling Arms with Invisible Hands

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:57.940291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:03:57.940291Z digest=sha256:86a2cd7f803b1b9cf1126b22c8e9605e4277e290de9d759438c8b6ededd79e04

Observation a6b3d551-4122-4822-b98e-5645994c4db2 · outbound

This paper cites The Sample Complexity of Online Contract Design.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals The Sample Complexity of Online Contract Design

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:58.056440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:03:58.056440Z digest=sha256:b0dba48ff170095e06e873899882c6e802e35cd8a09b97c1707a20c5b7e9aeba

Observation ecfdb087-a3f5-4bbc-b31f-7059bb66d8bb · outbound

This paper cites and Seldin, Y.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals and Seldin, Y

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:58.881888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:03:58.099350Z digest=sha256:030c106500bcef534fb883b46283bc014030129f38252487863b95fe7e0c7431

Pith citing papers

No inbound Pith citation observations are available.