Pith. sign in

Paper Citation Record · LEDGER

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals

As of 7 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 0 inbound Pith citation observations for arXiv:2505.23124.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.23124 v2

Coverage vector

measured 34 of 34 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:03:58.099350Z

measured 34 of 34 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

34 of 34 outbound references displayed

  • verified exact3
  • verified fuzzy17
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 42a2f4df-6928-40ea-ac1f-51e7a0af0e8f · outbound

This paper cites write newline.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:55.250411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:03:55.250411Z digest=sha256:a3a2a805f6b29abab799103b7355c3d212534adc989b4de87faaea13712a2065

Observation a1866754-f6e1-4bcd-b6e2-16da248e316a · outbound

This paper cites Nearly-Optimal Bandit Learning in Stackelberg Games with Side Information.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Nearly-Optimal Bandit Learning in Stackelberg Games with Side Information

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:55.339861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:03:55.339861Z digest=sha256:5a8e3d33a97e7621aa5f7ad9774a2f84d1f160091b76eaeca90a57d6bbd0829f

Observation 85dd2d88-cd1c-4e5f-83dc-4d8ff2ba6a96 · outbound

This paper cites Principal-agent reward shaping in mdps.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Principal-agent reward shaping in mdps

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:02.419844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:03:55.435050Z digest=sha256:daaa87974284fa17332cbb3d5b97521deffc8eb633512f1d0abd4884fc9c5be4

Observation 31ec152e-7f95-4de3-9410-d8b0dcc2f3ce · outbound

This paper cites Optimal rates and efficient algorithms for online bayesian persuasion.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Optimal rates and efficient algorithms for online bayesian persuasion

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:02.249370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:03:55.503853Z digest=sha256:a039e77c8455280f0cb3fa180dcc7cd69ec60682bc6f0fa598713ba12a38abd3

Observation e75ba6e5-2a03-43fd-81b8-e31b967ab5eb · outbound

This paper cites and Dewatripont, M.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals and Dewatripont, M

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:02.040640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:03:55.582634Z digest=sha256:56156c35b0aee6c01878f8605811835171c776e8ab9d6ac156f7314a64a86129

Observation e656ab1b-c7b1-4b92-af6d-9f8fb6360012 · outbound

This paper cites Robustness and linear contracts.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Robustness and linear contracts

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:01.815171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:03:55.666991Z digest=sha256:270583efa3de3d06e56366b1792fb35b2f5ac1b798b940b16406331e0fe005fa

Observation deeec76f-974e-49f7-bb12-1389c2d328c4 · outbound

This paper cites Learning to Price Homogeneous Data.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Learning to Price Homogeneous Data

Reference 7

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T13:03:58.698148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:03:55.758735Z digest=sha256:0d716cbe0cea278236c039613124bfe7a71e4b9122605a8d8e2bf569538b44da

Observation e6bfac6a-ee86-4ec2-b333-2015fa4f6c22 · outbound

This paper cites Estimating and Incentivizing Imperfect-Knowledge Agents with Hidden Rewards.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Estimating and Incentivizing Imperfect-Knowledge Agents with Hidden Rewards

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:55.914058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:03:55.914058Z digest=sha256:3d442c7ee0f3a2fc7755d4ac66b074fd8aae8510089e8bd438f6da4e9eb9dd1e

Observation 8535712b-1ea0-4538-81e9-31e4fd8f4e20 · outbound

This paper cites Repeated Principal-Agent Games with Unobserved Agent Rewards and Perfect-Knowledge Agents.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Repeated Principal-Agent Games with Unobserved Agent Rewards and Perfect-Knowledge Agents

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:55.983445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:03:55.983445Z digest=sha256:c11c7f5e872e1710f449c7922fb341ce49d9367129f6c02f59d17423574527cd

Observation 8e4ca8b3-98c5-4937-a88c-dd3c77d65b30 · outbound

This paper cites Multi-agent combinatorial contracts.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Multi-agent combinatorial contracts

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:01.591506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:03:56.039245Z digest=sha256:f6ddd1057513ed578d9090a9c7185b740b577b5d9c8604e2d49e4c1b6c63c0d4

Observation 16e8b822-2992-4ad2-8965-c4a651ab64e9 · outbound

This paper cites Simple versus optimal contracts.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Simple versus optimal contracts

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:01.376280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:03:56.146040Z digest=sha256:021345e35ee8eb73476eede75cae6308225141708b4a6993896a6094d089e295

Observation e09c02dd-e28a-4d9f-beaf-7da48169562b · outbound

This paper cites Combinatorial contracts.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Combinatorial contracts

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:01.119324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:03:56.272790Z digest=sha256:1597e54b6692933875232b12781d43b854508bfa736a96227c6f12c8dae7abe5

Observation 0237acb8-d9f4-451f-a689-4ecfe3e760c8 · outbound

This paper cites Multi-agent contracts.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Multi-agent contracts

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:00.898863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:03:56.341986Z digest=sha256:63bdffab3c0287a9e2f8e1cda81c4ee6ca8915dca5c1a4d1d8fc4a0c3abcf0dc

Observation 80fd2c91-5f50-47ce-8596-18b9156c8bfc · outbound

This paper cites an unresolved cited work.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:04:00.672928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:03:56.427122Z digest=sha256:493b27384576ced0ae84f423e438d4eae8023047b89a49ee95b36514652700ef

Observation 4551f37c-0668-4da8-8dd6-4f47d3a50c72 · outbound

This paper cites and Miller, R.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals and Miller, R

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:00.460405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:03:56.556800Z digest=sha256:ad63fb738661bc8feb513afef2eedbdca6953599954c3bdac8a7726ed0262fee

Observation 35e206c3-a181-450d-b2ad-6d617ab3d55a · outbound

This paper cites an unresolved cited work.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:04:00.292546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:03:56.620275Z digest=sha256:ca5c071940c95ed12a32fc1c3fde9bd688143cbfc187b95152a170a535583bf1

Observation f857fabb-98e6-4fb2-ba7b-a98ad9c76a32 · outbound

This paper cites Z., and Balcan, M.-F.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Z., and Balcan, M.-F

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:04:00.108777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:03:56.703746Z digest=sha256:7bc6ac7c210f4a4b97858cda720085256630f4bdee00bd744b5ebc654fffb51e

Observation 8d7285fc-1858-434a-a66d-ee1e967d5c94 · outbound

This paper cites an unresolved cited work.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:03:59.945634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:03:56.761059Z digest=sha256:9f95ed61a0d72f0a7a42b46a53756e5e8929c3f3fd9094289bbdb4892b15e2d1

Observation f64ebdb7-15f4-40aa-80a5-bd24a182ceb1 · outbound

This paper cites Moral hazard and observability.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Moral hazard and observability

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:59.744244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:03:56.862182Z digest=sha256:de0bdafb5de33759850238d7efc1c87df7c7501057084e8bc0a18a591ec521e0

Observation cdc8f525-ca14-423d-ab26-0679c422614c · outbound

This paper cites J., Rand, D.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals J., Rand, D

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:59.634972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:03:56.943018Z digest=sha256:e10740330ff8c76ff122447ab21afca419f689409b3cd1abd5e31748887ee276

Observation 03d54429-0d98-4eb9-9244-45a7f9a2108d · outbound

This paper cites and Siddiq, A.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals and Siddiq, A

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:59.465278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:03:57.022280Z digest=sha256:aa5280c24784843efae03ecb2beaef723147e58f1ff8baa2c430d464d42ac104

Observation 653aad28-cde3-4c75-aac6-2d938902be0c · outbound

This paper cites and Szepesv \'a ri, C.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals and Szepesv \'a ri, C

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:57.089780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:03:57.089780Z digest=sha256:b4e120071f63e6ff17ba3b097423899b961cbc5cecfcb4615c7eea0947429e1a

Observation af64217d-881b-43b3-a166-02718cfa7c29 · outbound

This paper cites Principal-Agent Bandit Games with Self-Interested and Exploratory Learning Agents.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Principal-Agent Bandit Games with Self-Interested and Exploratory Learning Agents

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:03:58.582546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:03:57.161602Z digest=sha256:46a70148fa7132248e78cc59e7ef84227c9c5ebc51692a65f8f6f2dd37ba6ebf

Observation ea1f350f-e5be-44dd-9dd4-09217f4cc620 · outbound

This paper cites On the Limitations and Possibilities of Nash Regret Minimization in Zero-Sum Matrix Games under Noisy Feedback.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals On the Limitations and Possibilities of Nash Regret Minimization in Zero-Sum Matrix Games under Noisy Feedback

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:57.238622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:03:57.238622Z digest=sha256:d99eb8d709385cd57b0362b76ea0fb17bd14395b5616eae5624f1d1e0b1e80a7

Observation 7419b645-f73e-46af-a77b-5d6f5fff10c8 · outbound

This paper cites Efficient Near-Optimal Algorithm for Online Shortest Paths in Directed Acyclic Graphs with Bandit Feedback Against Adaptive Adversaries.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Efficient Near-Optimal Algorithm for Online Shortest Paths in Directed Acyclic Graphs with Bandit Feedback Against Adaptive Adversaries

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:57.337942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:03:57.337942Z digest=sha256:01de7c05cbbbed21320606044a07a843b227312f010faa0dc1c92e1659fdc8db

Observation 6c3b199b-d782-4d64-97a5-f9ccfbeb5c11 · outbound

This paper cites and Nair, H.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals and Nair, H

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:59.332062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:03:57.446306Z digest=sha256:a3742d653922eb65f28394101efdb9444d4f0ae8d0fa463e79ea072dedab5d1e

Observation 121ef90f-431f-43c6-b3b1-cdb6fade7362 · outbound

This paper cites T., and Narasimhan, C.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals T., and Narasimhan, C

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:59.209606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:03:57.504728Z digest=sha256:382dfbae92e764b61dfc006be8778271ca69039df479670b8e3b38a6104f91bf

Observation 37e4a06d-bbbf-426b-8140-18d4a4623f49 · outbound

This paper cites and Slivkins, A.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals and Slivkins, A

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:59.098326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:03:57.608075Z digest=sha256:085fd4f3c4089941a8302efb940bc1d160dd040aa16c48ead133f34f812992d1

Observation b49c2a9e-a8ba-4958-a2e2-8ebcfc773255 · outbound

This paper cites Learning to Mitigate Externalities: the Coase Theorem with Hindsight Rationality.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Learning to Mitigate Externalities: the Coase Theorem with Hindsight Rationality

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:03:58.428319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:03:57.665267Z digest=sha256:1706f817a32c2834a3dd3e533cd88b2c6f0cfb925d67c8472628ac4ed42ffd19

Observation 62bd797d-15ec-4e52-ab61-0ffe285344b3 · outbound

This paper cites Incentivized Learning in Principal-Agent Bandit Games.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Incentivized Learning in Principal-Agent Bandit Games

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:03:58.299182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:03:57.746639Z digest=sha256:27293e10750a423eadb83faaed6dde733ef57411794eef3090eaa86e753004e6

Observation b4cd5186-42a8-4770-86eb-aa33b367ed34 · outbound

This paper cites an unresolved cited work.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:03:58.990757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:03:57.842207Z digest=sha256:e82553a4af3574409c845aa61bee25ae65c43e694ac268486d9460bf73b37869

Observation 46d66d4f-ee05-4395-9569-8465266007b3 · outbound

This paper cites Contractual Reinforcement Learning: Pulling Arms with Invisible Hands.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals Contractual Reinforcement Learning: Pulling Arms with Invisible Hands

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:57.940291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:03:57.940291Z digest=sha256:e68b7a6527c071e7934bfd5dacc5ee4fcff70ad59e8912ce218a484ba7f3e27c

Observation a6b3d551-4122-4822-b98e-5645994c4db2 · outbound

This paper cites The Sample Complexity of Online Contract Design.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals The Sample Complexity of Online Contract Design

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:58.056440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:03:58.056440Z digest=sha256:b279690c8e391f38c9d38a1d8d59c2c2783e5f718f3b9c0195f179e6b83365ef

Observation ecfdb087-a3f5-4bbc-b31f-7059bb66d8bb · outbound

This paper cites and Seldin, Y.

Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals and Seldin, Y

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:03:58.881888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:03:58.099350Z digest=sha256:7cdb451331627682795857d9b9e0ac0b54bcafbcb471ddb4e05005a9bab599f9

Pith citing papers

No inbound Pith citation observations are available.