Pith. sign in

Paper Citation Record · LEDGER

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork

As of 22 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 0 inbound Pith citation observations for arXiv:2605.24423.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.24423 v1

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-30T13:41:33.048760Z

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

17 of 17 outbound references displayed

  • verified exact1
  • verified fuzzy8
  • unresolved1
  • parse uncertain0
  • malformed identifier3
  • metadata mismatch4

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2994aa2f-ea39-486b-9bbb-c0ae1fbbb9ca · outbound

This paper cites org/CorpusID:258845718.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork org/CorpusID:258845718

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T01:05:50.377952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:bc06b31559a993f356180e75b7461790e0131a978d591804809cebd4f2261cf5

Observation 265e1385-03b7-47fa-b88a-ab2b6308995f · outbound

This paper cites Gaussian Error Linear Units (GELUs).

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Gaussian Error Linear Units (GELUs)

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-06-30T13:44:40.646681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:cc8714159921b51746d402956c388a7a6152d4cb28a06f16c2adabac7f97612c

Observation 8f643b94-2f42-4010-9287-9968208ae998 · outbound

This paper cites Population Based Training of Neural Networks.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Population Based Training of Neural Networks

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-06-30T13:44:40.655360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:3c86dce8500a2d61abe1e4a9faacc657277afe45592029a77715ebd06d2d57f4

Observation 16824615-afb9-47c1-9d37-92f47a96f050 · outbound

This paper cites org/CorpusID:235313679.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork org/CorpusID:235313679

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T01:05:50.379622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:648ff994a41f200edd70900ce7604a405f293653721b3f62a5e90f2946e807bc

Observation a8845882-9b53-45fb-b7da-d223bc92992c · outbound

This paper cites Lee, K.-H., Nachum, O., Yang, M.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Lee, K.-H., Nachum, O., Yang, M

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T01:05:50.376310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:f75e8e06a944d2dbcf030428c82eca7e90bbf21ae44fb3bb5cee0e104391c5d7

Observation 9d35b37b-59e1-43eb-9bb7-aa9eedff9f76 · outbound

This paper cites A Survey of In-Context Reinforcement Learning.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork A Survey of In-Context Reinforcement Learning

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T13:44:40.652995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:7100e268e93255820865751d3b57462d5dcaa9b3cf627fcd1794482bd61a85ae

Observation fdddb94a-055c-40ba-a3a0-bf0bb16f9ce6 · outbound

This paper cites Nikulin, A., Zisman, I., Zemtsov, A., and Kurenkov, V.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Nikulin, A., Zisman, I., Zemtsov, A., and Kurenkov, V

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T01:05:50.381274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:bfd4f9dea99bbdf0177c211e9acebac7df676cc02631017faf960d13462395f3

Observation 3cf610d3-6094-4450-b28e-6c46046c07d0 · outbound

This paper cites Papoudakis, G., Christianos, F., Sch¨afer, L., and Albrecht, S.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Papoudakis, G., Christianos, F., Sch¨afer, L., and Albrecht, S

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T01:05:50.384709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:0a8de32bf1f83bae92ca679dea3b4ea35fc80b3d0fac610beef1295f82af784d

Observation 094f98ac-c233-4892-81d0-2936901bb134 · outbound

This paper cites Rahman, A., Fosong, E., Carlucho, I., and Albrecht, S.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Rahman, A., Fosong, E., Carlucho, I., and Albrecht, S

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T01:05:50.382977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:e5e05185d31cd3ff302ea57b8c4906c1d2e582c0566697df12293eb96b95c291

Observation 8c372c56-e4e3-46d5-9d5f-629fc0451f94 · outbound

This paper cites Rahman, M., Cui, J., and Stone, P.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Rahman, M., Cui, J., and Stone, P

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T01:05:50.386230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:771bbd5739fa20f538808e719772b061243d18fd941dab8462cde7c970bff39a

Observation 68c6b7eb-5f46-4fe6-a84f-fd68489ca1ed · outbound

This paper cites Proximal Policy Optimization Algorithms.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Proximal Policy Optimization Algorithms

Reference 11

Resolution
metadata mismatch
local_arxiv, observed 2026-06-30T13:44:40.650034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:b2a5d6b9c473f21ee3111816b1c5c58054d54e19fc411e0a10226fc135135abd

Observation 1622cf25-36bc-4b78-ac98-ee0b0c320d51 · outbound

This paper cites Game-Theoretic Multiagent Reinforcement Learning.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Game-Theoretic Multiagent Reinforcement Learning

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T13:44:40.658304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:07c62a61286a672fa33ae59e0861f9f7b0bd2a4b53b32b1e353ebc03059d8520

Observation 297e9e37-ef01-4313-b3fc-9ccc8ac17698 · outbound

This paper cites 13 Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork A.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork 13 Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork A

Reference 13

Resolution
malformed identifier
raw_fallback, observed 2026-07-09T01:05:50.369574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:199599f5addcc72d96d56742875dfdeeea28239fccca0d0f1dc216eec68a4982

Observation 257aa6cb-219e-4e42-9005-ebe973523059 · outbound

This paper cites an unresolved cited work.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Unresolved cited work

Reference 14

Resolution
malformed identifier
raw_fallback, observed 2026-07-09T01:05:50.371255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:745d3d5df9025fa05493a307d25723a2b34f06f4211c858dedd76ad390eea391

Observation b726f50b-16e6-439b-8866-9d280035ccb8 · outbound

This paper cites an unresolved cited work.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-07-09T01:05:50.372914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:7f9c91ea953fbe1cae75a80310f8753a752e79d850db387d689520455a9e214e

Observation 1d53bfb5-76d2-4fab-af37-51755952b90b · outbound

This paper cites These layers perform channel-wise feature transformation without spatial mixing.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork These layers perform channel-wise feature transformation without spatial mixing

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T01:05:50.374630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:b5bc918fcfe62cb2374d0af851c57ed59c1a79152dc4bfe43dafaa3b4f99dc95

Observation c178aa6a-d8fe-46b7-a1db-799a383de70a · outbound

This paper cites with prior.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork with prior

Reference 17

Resolution
malformed identifier
raw_fallback, observed 2026-07-09T01:05:50.367729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:91cfc7bff99691958b3c6025cf297363e14c062aa2d1f1eac8bc42ebf769d46c

Pith citing papers

No inbound Pith citation observations are available.