Pith. sign in

Paper Citation Record · LEDGER

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork

As of 9 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 0 inbound Pith citation observations for arXiv:2605.24423.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.24423 v1

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-30T13:41:33.048760Z

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

17 of 17 outbound references displayed

  • verified exact1
  • verified fuzzy8
  • unresolved1
  • parse uncertain0
  • malformed identifier3
  • metadata mismatch4

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2994aa2f-ea39-486b-9bbb-c0ae1fbbb9ca · outbound

This paper cites org/CorpusID:258845718.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork org/CorpusID:258845718

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T01:05:50.377952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:2f1ee13dc16fcdb5743c938b6bd436611e2598e461dc68ffe4be66b9b30c1fd9

Observation 265e1385-03b7-47fa-b88a-ab2b6308995f · outbound

This paper cites Gaussian Error Linear Units (GELUs).

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Gaussian Error Linear Units (GELUs)

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-06-30T13:44:40.646681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:b49c805e59979dd1ca37005e94d9aa4a0f670dee7f3e77130d8227fae8d2f7cd

Observation 8f643b94-2f42-4010-9287-9968208ae998 · outbound

This paper cites Population Based Training of Neural Networks.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Population Based Training of Neural Networks

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-06-30T13:44:40.655360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:787b381ae1b26f25813dc252b1952647a1c344cd8952bbddc47e76a5ff8a3e5d

Observation 16824615-afb9-47c1-9d37-92f47a96f050 · outbound

This paper cites org/CorpusID:235313679.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork org/CorpusID:235313679

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T01:05:50.379622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:30bd4b92773f611a2b6e278d7a10eeab00320c7d1ec16e01baaaf209550ff90c

Observation a8845882-9b53-45fb-b7da-d223bc92992c · outbound

This paper cites Lee, K.-H., Nachum, O., Yang, M.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Lee, K.-H., Nachum, O., Yang, M

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T01:05:50.376310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:da5448caa6094b494fdfbaadf8d3abc1152dcb0c5c369aaaca3f1c69db141e07

Observation 9d35b37b-59e1-43eb-9bb7-aa9eedff9f76 · outbound

This paper cites A Survey of In-Context Reinforcement Learning.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork A Survey of In-Context Reinforcement Learning

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T13:44:40.652995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:b0835fbab1deb4686a06c15be98091ff6cd75d7f1b9ac07e04c4c1668da355ee

Observation fdddb94a-055c-40ba-a3a0-bf0bb16f9ce6 · outbound

This paper cites Nikulin, A., Zisman, I., Zemtsov, A., and Kurenkov, V.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Nikulin, A., Zisman, I., Zemtsov, A., and Kurenkov, V

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T01:05:50.381274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:323cc9e7fba54dc094059b7458fb500e85c8a512e4c0694ec8eeafde5aaf6ada

Observation 3cf610d3-6094-4450-b28e-6c46046c07d0 · outbound

This paper cites Papoudakis, G., Christianos, F., Sch¨afer, L., and Albrecht, S.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Papoudakis, G., Christianos, F., Sch¨afer, L., and Albrecht, S

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T01:05:50.384709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:65abfc3ce20f3d2a9f5915492356713f97c8d60327c2f7a517361bb0c604f12f

Observation 094f98ac-c233-4892-81d0-2936901bb134 · outbound

This paper cites Rahman, A., Fosong, E., Carlucho, I., and Albrecht, S.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Rahman, A., Fosong, E., Carlucho, I., and Albrecht, S

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T01:05:50.382977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:f0dada814b7a1dade4fb6a1392cd39485369d46cae67b74bd04dbc4983f3bddc

Observation 8c372c56-e4e3-46d5-9d5f-629fc0451f94 · outbound

This paper cites Rahman, M., Cui, J., and Stone, P.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Rahman, M., Cui, J., and Stone, P

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T01:05:50.386230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:b4213aebae6a9e51679aac9669c25b459c1e30d74b96b878cf238f4424ada2cc

Observation 68c6b7eb-5f46-4fe6-a84f-fd68489ca1ed · outbound

This paper cites Proximal Policy Optimization Algorithms.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Proximal Policy Optimization Algorithms

Reference 11

Resolution
metadata mismatch
local_arxiv, observed 2026-06-30T13:44:40.650034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:401e2e102d1992d92e7f7292f2a38349f6da715c534f3399f50ae3bf7ff37c90

Observation 1622cf25-36bc-4b78-ac98-ee0b0c320d51 · outbound

This paper cites Game-Theoretic Multiagent Reinforcement Learning.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Game-Theoretic Multiagent Reinforcement Learning

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T13:44:40.658304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:1cdfc8299b745bdf739eacb0879c87fb2e6525126e76353cc1de3394f61faa8c

Observation 297e9e37-ef01-4313-b3fc-9ccc8ac17698 · outbound

This paper cites 13 Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork A.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork 13 Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork A

Reference 13

Resolution
malformed identifier
raw_fallback, observed 2026-07-09T01:05:50.369574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:2e36e6e2d3e3f696c65fafc5904860190a04ef1ba8d20ac0b82e0f1662a9f9ac

Observation 257aa6cb-219e-4e42-9005-ebe973523059 · outbound

This paper cites an unresolved cited work.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Unresolved cited work

Reference 14

Resolution
malformed identifier
raw_fallback, observed 2026-07-09T01:05:50.371255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:4389f1d5e22b2bbedf660e58f9b19c418d5bacd4dbc89da71e00108b758a08e0

Observation b726f50b-16e6-439b-8866-9d280035ccb8 · outbound

This paper cites an unresolved cited work.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-07-09T01:05:50.372914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:e50426eef7b05f277a512689f1113c1eacab2a378511c3f563a4bfb277d5a489

Observation 1d53bfb5-76d2-4fab-af37-51755952b90b · outbound

This paper cites These layers perform channel-wise feature transformation without spatial mixing.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork These layers perform channel-wise feature transformation without spatial mixing

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T01:05:50.374630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:fa827413f5e75b97fb10acb10dea5d165124bcef99c76daf0aae5c68e218907e

Observation c178aa6a-d8fe-46b7-a1db-799a383de70a · outbound

This paper cites with prior.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork with prior

Reference 17

Resolution
malformed identifier
raw_fallback, observed 2026-07-09T01:05:50.367729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:ff6b3ca8ffde116fbf964962df69688b59c1917e566483843e303ffceac837bf

Pith citing papers

No inbound Pith citation observations are available.