Pith. sign in

Paper Citation Record · LEDGER

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork

As of 9 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 0 inbound Pith citation observations for arXiv:2605.24423.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.24423 v1

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-30T13:41:33.048760Z

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

17 of 17 outbound references displayed

  • verified exact1
  • verified fuzzy8
  • unresolved1
  • parse uncertain0
  • malformed identifier3
  • metadata mismatch4

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2994aa2f-ea39-486b-9bbb-c0ae1fbbb9ca · outbound

This paper cites org/CorpusID:258845718.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork org/CorpusID:258845718

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T01:05:50.377952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:d238ae3cd416e7c05a8c8e2cbb39e52b3f5de4f88d0de0465bf388589fcf19bb

Observation 265e1385-03b7-47fa-b88a-ab2b6308995f · outbound

This paper cites Gaussian Error Linear Units (GELUs).

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Gaussian Error Linear Units (GELUs)

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-06-30T13:44:40.646681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:66ff8465751a8d57bc6c5303e42cb9a53734ee7dd313af733a8b32c13400f394

Observation 8f643b94-2f42-4010-9287-9968208ae998 · outbound

This paper cites Population Based Training of Neural Networks.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Population Based Training of Neural Networks

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-06-30T13:44:40.655360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:77502bb47e267005c28a034f9cca74b85f6d1ff7640f5b8631878f52992f0d2b

Observation 16824615-afb9-47c1-9d37-92f47a96f050 · outbound

This paper cites org/CorpusID:235313679.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork org/CorpusID:235313679

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T01:05:50.379622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:8c632f003b35421f2a59b941c42af97f7a521ca65eede29b24db904cc601e636

Observation a8845882-9b53-45fb-b7da-d223bc92992c · outbound

This paper cites Lee, K.-H., Nachum, O., Yang, M.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Lee, K.-H., Nachum, O., Yang, M

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T01:05:50.376310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:a905c97f2cd0a7c9d71153ee5e74ff28355673dfe205bb55a011cee41e3582d2

Observation 9d35b37b-59e1-43eb-9bb7-aa9eedff9f76 · outbound

This paper cites A Survey of In-Context Reinforcement Learning.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork A Survey of In-Context Reinforcement Learning

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T13:44:40.652995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:51aea1be72c2e050e3adf4d26c2cb30c44ac6ec93b73bbd21a7bfc8d13b72ae1

Observation fdddb94a-055c-40ba-a3a0-bf0bb16f9ce6 · outbound

This paper cites Nikulin, A., Zisman, I., Zemtsov, A., and Kurenkov, V.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Nikulin, A., Zisman, I., Zemtsov, A., and Kurenkov, V

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T01:05:50.381274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:885294b9620a391d4062df0594c9d9782ba6bc269bcb14f4a4b209c7683428d4

Observation 3cf610d3-6094-4450-b28e-6c46046c07d0 · outbound

This paper cites Papoudakis, G., Christianos, F., Sch¨afer, L., and Albrecht, S.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Papoudakis, G., Christianos, F., Sch¨afer, L., and Albrecht, S

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T01:05:50.384709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:69bb3786a5281c21ddd2b99b5c41e52a5f3ebd08e43adcd3418ca82c55df676e

Observation 094f98ac-c233-4892-81d0-2936901bb134 · outbound

This paper cites Rahman, A., Fosong, E., Carlucho, I., and Albrecht, S.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Rahman, A., Fosong, E., Carlucho, I., and Albrecht, S

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T01:05:50.382977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:ce8a966849225a9cbb859c8f50b750d248380d60eb34aa9d66eeea36c2ea2529

Observation 8c372c56-e4e3-46d5-9d5f-629fc0451f94 · outbound

This paper cites Rahman, M., Cui, J., and Stone, P.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Rahman, M., Cui, J., and Stone, P

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T01:05:50.386230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:594cc914823d5a795a96e3acb14ecf72b499f3c3cb560ee35e3c90d208c81c4a

Observation 68c6b7eb-5f46-4fe6-a84f-fd68489ca1ed · outbound

This paper cites Proximal Policy Optimization Algorithms.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Proximal Policy Optimization Algorithms

Reference 11

Resolution
metadata mismatch
local_arxiv, observed 2026-06-30T13:44:40.650034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:25580b3132ee8263f3b3e6694390fe5b761692a3865ea01bf025e4c5aec3b53a

Observation 1622cf25-36bc-4b78-ac98-ee0b0c320d51 · outbound

This paper cites Game-Theoretic Multiagent Reinforcement Learning.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Game-Theoretic Multiagent Reinforcement Learning

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T13:44:40.658304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:4e05d8bdc24c2c5438a19608c72925fccaa2ecaacbf499e30fb08a01cb629cc7

Observation 297e9e37-ef01-4313-b3fc-9ccc8ac17698 · outbound

This paper cites 13 Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork A.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork 13 Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork A

Reference 13

Resolution
malformed identifier
raw_fallback, observed 2026-07-09T01:05:50.369574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:0cc02089dfcc7c38a6bef0e3405a33b63cbe703e04ecfe562ee72df327a648df

Observation 257aa6cb-219e-4e42-9005-ebe973523059 · outbound

This paper cites an unresolved cited work.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Unresolved cited work

Reference 14

Resolution
malformed identifier
raw_fallback, observed 2026-07-09T01:05:50.371255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:2d48ecef83979c654737f8f7ce36e49bc1fb9e01fffa2e3c5a44ed90efc33cd4

Observation b726f50b-16e6-439b-8866-9d280035ccb8 · outbound

This paper cites an unresolved cited work.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-07-09T01:05:50.372914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:7887ff75765a768f0cfbf3894678132b9dd3ca1b819c4d2e22a7d0937f5b4faf

Observation 1d53bfb5-76d2-4fab-af37-51755952b90b · outbound

This paper cites These layers perform channel-wise feature transformation without spatial mixing.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork These layers perform channel-wise feature transformation without spatial mixing

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T01:05:50.374630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:303d8f73ade5e1ec985bd5f6270fb16e41480d579527350bd59063a8e0eacab4

Observation c178aa6a-d8fe-46b7-a1db-799a383de70a · outbound

This paper cites with prior.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork with prior

Reference 17

Resolution
malformed identifier
raw_fallback, observed 2026-07-09T01:05:50.367729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:4d5f1c2032cc988ea9e787c617f0bba618e6afdabe3c371d59a292f06047340a

Pith citing papers

No inbound Pith citation observations are available.