Pith. sign in

Paper Citation Record · LEDGER

Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games

As of 20 August 2026, this Paper Citation Record lists 18 of 18 outbound references and 0 inbound Pith citation observations for arXiv:2510.24515.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2510.24515 v2

Coverage vector

measured 18 of 18 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T15:49:22.378838Z

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

18 of 18 outbound references displayed

  • verified exact2
  • verified fuzzy13
  • unresolved3
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 66b07687-4b4c-42fa-a703-60781bbf25e7 · outbound

This paper cites The team orienteering problem,.

Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games The team orienteering problem,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:49:22.669018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T15:49:22.301713Z digest=sha256:c675fb754533acf1f7fbe7bcc696017ee8911261345818109eb34e2da32fd639

Observation 3e00d25a-5dcd-4dc9-a7f2-51fe14d2c274 · outbound

This paper cites Multi- robot scheduling for environmental monitoring as a team orienteering problem,.

Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games Multi- robot scheduling for environmental monitoring as a team orienteering problem,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:49:22.654603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T15:49:22.306673Z digest=sha256:0c861c8186dc0346f6dbd5b98da227792055f777aa94ea02e00271acd1d7a639

Observation 173b04ed-e48b-4dbd-bb33-c23bfb0dbbc4 · outbound

This paper cites A spatio-temporal representation for the orienteering problem with time-varying profits,.

Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games A spatio-temporal representation for the orienteering problem with time-varying profits,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:49:22.640597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T15:49:22.311063Z digest=sha256:f49f15be0182b0e4e042e6c1d6708a38dd33cb4e07452af8adcb690c262baf83

Observation 4543c375-f62d-4c83-8294-82afcf7d960b · outbound

This paper cites Orienteering Problem: A survey of recent variants, solution approaches and applications,.

Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games Orienteering Problem: A survey of recent variants, solution approaches and applications,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:49:22.624805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T15:49:22.315435Z digest=sha256:092344f12503eacc32b14594eff1af0f305c122a764a51b8091785e9972cac73

Observation e3b76725-514f-4b4d-8751-4792665916e9 · outbound

This paper cites Learn to solve the min-max multiple traveling salesmen problem with reinforcement learning,.

Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games Learn to solve the min-max multiple traveling salesmen problem with reinforcement learning,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:49:22.609743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T15:49:22.320084Z digest=sha256:7a55b1129c86f25cfa8cfd75ddbecdee9eda84cb9c6892cf9bf4df7d926ce9c0

Observation a93edeb8-854f-43f1-a770-fc79719d11b9 · outbound

This paper cites Competing for the most profitable tour: The orienteering interdiction game.

Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games Competing for the most profitable tour: The orienteering interdiction game

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-15T15:49:22.456866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T15:49:22.324530Z digest=sha256:b289417797a9e1e7e8518c17acc730afc495af52a5070a9063094a68468e194f

Observation 3bf35bf7-251b-472f-8e9f-813795920025 · outbound

This paper cites DIRECT: A scalable approach for route guidance in Selfish Orienteering Problems,.

Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games DIRECT: A scalable approach for route guidance in Selfish Orienteering Problems,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:49:22.594651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T15:49:22.329350Z digest=sha256:a9b28b8bbca4af0305663b90cf7826d194d93dbe21f7cf6796be77255acb61a2

Observation be888651-ed16-4863-95d2-58cee8539732 · outbound

This paper cites Prize collecting multiagent orienteer- ing: Price of anarchy bounds and solution methods,.

Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games Prize collecting multiagent orienteer- ing: Price of anarchy bounds and solution methods,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:49:22.579784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T15:49:22.334145Z digest=sha256:528f220e2edde46f85cdbf808d158367c94f2ab65ebbf10eb8121b9f0649d060

Observation ec52a542-0848-49d4-a561-abf4f513fee8 · outbound

This paper cites Collaborative dynamic scheduling in a self-organizing manufacturing system using multi-agent reinforcement learning,.

Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games Collaborative dynamic scheduling in a self-organizing manufacturing system using multi-agent reinforcement learning,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:49:22.564650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T15:49:22.338900Z digest=sha256:a0071e6fc70a0328129ee683df93531a73485af9ad1f6c32e01243ced17eec41

Observation 21345f24-b39c-497d-b30f-a0524eaf49fc · outbound

This paper cites Optimizing task scheduling in human- robot collaboration with deep multi-agent reinforcement learning,.

Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games Optimizing task scheduling in human- robot collaboration with deep multi-agent reinforcement learning,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:49:22.550337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T15:49:22.343056Z digest=sha256:fd615393c67428e60eecdeabad95dd84e3e606a2b7fddbe0a5ca30ec8b4ef753

Observation 15a68cac-e63c-4ace-96e9-16ea94ed60cd · outbound

This paper cites A dynamic task assignment model for aviation emergency rescue based on multi- agent reinforcement learning,.

Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games A dynamic task assignment model for aviation emergency rescue based on multi- agent reinforcement learning,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:49:22.535739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T15:49:22.347261Z digest=sha256:6e5d90f31f7dd9481e95f12a05dc9eb35881fa4ce71974a8a073728a8799693c

Observation 544bfbf2-e66d-4aa4-b7bd-78c90ca65377 · outbound

This paper cites Graph Attention Multi-Agent Fleet Autonomy for Advanced Air Mobility.

Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games Graph Attention Multi-Agent Fleet Autonomy for Advanced Air Mobility

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-08-15T15:49:22.434802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T15:49:22.351355Z digest=sha256:50e2cd02da176f5f2c6c2c1eb39ab9754a8f0fca9c2751ee04d3b23247c920d6

Observation fea77039-023c-4e29-9f7a-d85a3953bf64 · outbound

This paper cites Consensus-based decentralized auctions for robust task allocation,.

Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games Consensus-based decentralized auctions for robust task allocation,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T15:49:22.356869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:49:22.356869Z digest=sha256:23f44460cbe3235d8fedfab1fdc62e2d9c1c853a748e64739fc13ddf6e687a88

Observation 2a6967ef-ab77-44bc-9b7b-31c331827c8d · outbound

This paper cites Optimal cost-sharing in general resource selection games,.

Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games Optimal cost-sharing in general resource selection games,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:49:22.512368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T15:49:22.361724Z digest=sha256:ee671924427feb14cbdae385cd82537204e653c6714d944b90718968da8aab1a

Observation b88d9933-c6fa-4966-988f-7a3b2d65354f · outbound

This paper cites Nash q-learning for general-sum stochastic games,.

Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games Nash q-learning for general-sum stochastic games,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:49:22.496874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T15:49:22.365750Z digest=sha256:f3d6ca300583b3db7efcbf2e380e41e7688e3cebaeaa671433da49191e992ee8

Observation 543bf2e6-9c22-40a5-9b9d-132ef7e07659 · outbound

This paper cites Multiagent Systems: Algorithmic, Game-Theoretic, and Logical Foundations,.

Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games Multiagent Systems: Algorithmic, Game-Theoretic, and Logical Foundations,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:49:22.481311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T15:49:22.369687Z digest=sha256:4af9d10ebacc92a8a67a62d0d0df6d9bf9c90e8a47a997e19d7086ddc4dcd918

Observation cdeba35a-1efa-4ffe-afcd-fa9a50b79ee1 · outbound

This paper cites Stabilizing Transformers for Reinforcement Learning.

Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games Stabilizing Transformers for Reinforcement Learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T15:49:22.373877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:49:22.373877Z digest=sha256:3a09dc8d18680f5d3530457e3ffa723d1529775f008a3a66e48f31390c08de91

Observation 9332d9e6-77ba-4d08-9693-a3037e297754 · outbound

This paper cites Orienteering problem: A survey of recent variants, solution approaches and applications,.

Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games Orienteering problem: A survey of recent variants, solution approaches and applications,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T15:49:22.378838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:49:22.378838Z digest=sha256:58a7dc38a2b2f19f4079686542e4dbc3d65fa5cdf7de02ab70de79d1db6ef825

Pith citing papers

No inbound Pith citation observations are available.