Pith. sign in

Paper Citation Record · LEDGER

In-Context Reinforcement Learning via Communicative World Models

As of 7 August 2026, this Paper Citation Record lists 62 of 62 outbound references and 0 inbound Pith citation observations for arXiv:2508.06659.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.06659 v2

Coverage vector

measured 62 of 62 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T22:39:48.597389Z

measured 62 of 62 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

62 of 62 outbound references displayed

  • verified exact2
  • verified fuzzy13
  • unresolved47
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c9f2c5e1-3ea8-4e09-9cfd-c0fdbeb8ac6e · outbound

This paper cites , " * write output.state after.block = add.period write newline.

In-Context Reinforcement Learning via Communicative World Models , " * write output.state after.block = add.period write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T22:39:45.037158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:39:45.037158Z digest=sha256:29b7c53d4e13deda47732b7e1076cbdac263c264f971feac4c8fae8430e813e8

Observation 82fdc10e-4870-4ef2-bf71-0379e17b066f · outbound

This paper cites write newline.

In-Context Reinforcement Learning via Communicative World Models write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T22:39:45.101891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:39:45.101891Z digest=sha256:df79ad010723d4b40e1782c4a90a6dcdb25dc56f1882b417427d4b45b6baa073

Observation 82b4842c-e2f0-4ddf-838c-e2a8edb2e121 · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:49.224061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:45.253014Z digest=sha256:0e34e4bb60a83aa4f3d52d766613e2d24dd732b93c3c36241b03e1ddf86d9d06

Observation b6edb23e-ade9-49a3-9f3c-e1ec74cada93 · outbound

This paper cites J.; Leary, C.; Maclaurin, D.; Necula, G.; Paszke, A.; Vander P las, J.; Wanderman- M ilne, S.; and Zhang, Q.

In-Context Reinforcement Learning via Communicative World Models J.; Leary, C.; Maclaurin, D.; Necula, G.; Paszke, A.; Vander P las, J.; Wanderman- M ilne, S.; and Zhang, Q

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:39:49.214555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:45.370556Z digest=sha256:7912ad0a904f2ef8a2e447f0c5b6c2d9263fd72d14f3e271972841436dbdc9f5

Observation 937d04a6-05b2-47ed-8a4b-5ca56e0fd702 · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:49.204825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:45.485644Z digest=sha256:80fa128a0d9f5bdd918026ca4fdff02e8538ca5279df1c261c9fe5337d278f69

Observation 9e1b90c1-a3fb-4b2c-873d-3dbe8d6210ef · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:49.195059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:45.562696Z digest=sha256:c3a3a5c963f8bb051266a4dafe75978e84b994104606b6fe9ff716d724d905ba

Observation 1257ab98-e5a1-4c32-8c23-1b5baaa16ec8 · outbound

This paper cites S.; and Terry, J.

In-Context Reinforcement Learning via Communicative World Models S.; and Terry, J

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:39:49.186166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:45.640886Z digest=sha256:bf948f4d8a6d5859fb2cb70d124b0ef49424b26777c30647b0782c3e4154e433

Observation 05771e7c-4670-47fd-a69c-d8986891c441 · outbound

This paper cites Empirical Evaluation of Gated Recurrent Neural Networks on Sequence Modeling.

In-Context Reinforcement Learning via Communicative World Models Empirical Evaluation of Gated Recurrent Neural Networks on Sequence Modeling

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T22:39:45.712710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:39:45.712710Z digest=sha256:ad7118b4da3db2e0027893701dbd34089718b5bab719fbb23a0289b0001322c2

Observation 9d7da363-b97d-4495-8556-6a44f67099e0 · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:49.176992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:45.792918Z digest=sha256:9000c25fb97421e7d8b6db46b41821bda29d9e063d66eff1669aedd7863d7e47

Observation 7cad8035-5f6e-4e94-9732-9b1201312a73 · outbound

This paper cites P.; and Sobel, J.

In-Context Reinforcement Learning via Communicative World Models P.; and Sobel, J

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:39:49.166369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:45.903437Z digest=sha256:0638beee7fc764e28a9776dcb7eb9ab66f4465341ef58b04dc05ade4de66329f

Observation 717c0461-847e-4870-b145-0085a4bffc64 · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:49.155969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:45.986911Z digest=sha256:a513011c2fbdbaa621b3425bc9da80c1692df3c6d95d87309ff210fbbfb29789

Observation 2496343b-ce7a-4c18-852b-1d02dc244c1e · outbound

This paper cites L.; Sutskever, I.; and Abbeel, P.

In-Context Reinforcement Learning via Communicative World Models L.; Sutskever, I.; and Abbeel, P

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:39:49.145676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:46.070952Z digest=sha256:27fd3b67e6c33bd1f055e1f48ea1068e82f0ccee48b70170a612dd1793c85188

Observation bbb60da9-630b-47f2-864d-7d7b0059d273 · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:49.137426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:46.151557Z digest=sha256:2c93029b85091a8e17bfb3218aa7870d984e4fe4762fbdda9fa4bdeaef24bf04

Observation fd767ba3-6308-4ea2-a26c-cea15cfdf23c · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:49.128829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:46.254845Z digest=sha256:f3babc27b615706f8797e5accbb8752b4c8709300b91e5990a24d0d1f35ea25b

Observation 2e5df821-dc60-47ef-9234-739d5ff940af · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:49.119192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:46.341860Z digest=sha256:4f199fd2c17f096b9e23dcb8d4c57070f86aa5c69c797395c9fc7368043f9921

Observation 4822acc7-c3c1-4d05-acd1-1d0ec8854985 · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:49.107968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:46.440911Z digest=sha256:7ff5fdc1f757d4f5b58ae6c64886eba26b2ba1057b02b891d27bd86c1039eb9b

Observation d357e248-b66d-4ec0-885c-9224ae6f501e · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:49.097975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:46.530566Z digest=sha256:43c7dcebd994c07e8c546c60bd88ef3dac6b6ccacceea9db536c1fe21c8cfcf9

Observation 3dbe7702-97f6-4ec7-8e72-2fd99fbf8532 · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:49.089130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:46.636882Z digest=sha256:d291a3085138de2683f6fec8d4a58e1967c7834342925bac8e7e3aefc03db86d

Observation e0d384ef-5cb6-43df-a8e2-d4d19a0a6061 · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:49.080649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:46.753166Z digest=sha256:3a45256edd7a843249d0348779cd1b0f2d83bd2f6242ffde46556ab359149162

Observation 46f10cdb-564a-4bf7-b2f1-c6f25dd8d169 · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:49.071754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:46.837047Z digest=sha256:627fbf109e0d840994651d33e6a1d5ff93611f39ba5e35bd2e6fb54c1a9f835c

Observation 632b3fb6-3c27-49c8-9fbb-aa8d7ee0cab3 · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:49.061487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:46.940472Z digest=sha256:f2724d06d67a595f521f810c1c483778b88fb58a45671b65ee41ef1a84247cad

Observation 5fd79b3b-feee-468c-8513-78f47033fe46 · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:49.052641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:47.059583Z digest=sha256:f6c20ce7b6c7490923790b8a45e4c4da110c8a319ff46681415f7f5f5908143d

Observation 968b59cc-bf7d-44f5-a21c-4b96de6d6429 · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:49.043265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:47.173166Z digest=sha256:9d121d9df953c75c0da664c60a234b9273741e0cd8996c66b62443bdf5306aae

Observation 8107a4c5-b134-4a93-a911-390829b305ee · outbound

This paper cites v.; Modayil, J.; and Silver, D.

In-Context Reinforcement Learning via Communicative World Models v.; Modayil, J.; and Silver, D

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:39:49.033604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:47.176719Z digest=sha256:610416e0eb512c3e99ff3f8a4a81a42a3d5e3579217e268d37d41294ab1b9591

Observation d0257c05-e59e-49c9-a34d-fbb3f8c02361 · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:49.023234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:47.216761Z digest=sha256:7aad26caa8e0c11ef3fb13390eab8a6169d88896b8483978ebd63025280c3531

Observation 34e1c847-a035-4dce-b3fc-27378ab1d01e · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:49.013860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:47.344823Z digest=sha256:1c73cf87538537a55235c23dfaa82ac16b530a98884eb6028ebd85015316d38f

Observation 1f263870-0123-4e51-9b81-042bc8f85e26 · outbound

This paper cites P.; Littman, M.

In-Context Reinforcement Learning via Communicative World Models P.; Littman, M

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:39:49.004302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:47.422842Z digest=sha256:fc8aa099ebc169079b1f70468c1f45e7d8c7b9af6c592780406aeb0b255ffba4

Observation 5496bdbc-3c7c-427d-9a05-ea5762b963bc · outbound

This paper cites J.; Zhang, C.; and Slivkins, A.

In-Context Reinforcement Learning via Communicative World Models J.; Zhang, C.; and Slivkins, A

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:39:48.994516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:47.550840Z digest=sha256:6c56d6b48128f90f56b015191bd47d4f8befec3d1154bdcfc0b81ee3a2fe80bc

Observation 13dfe71c-19a1-444f-bcac-ce8589b3469a · outbound

This paper cites S.; Filos, A.; Brooks, E.; maxime gazeau; Sahni, H.; Singh, S.; and Mnih, V.

In-Context Reinforcement Learning via Communicative World Models S.; Filos, A.; Brooks, E.; maxime gazeau; Sahni, H.; Singh, S.; and Mnih, V

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:39:48.983579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:47.678196Z digest=sha256:7e9e0d1632bea5ab2740fd1383b8bb09b5d28283d8178451af4b8bb4a85794de

Observation 6814470f-7c6e-4d0d-aa41-f0d99aa115ac · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:48.973002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:47.775691Z digest=sha256:d31ebfdf9e850da931b0af6a6c2c7a52f7f099e60c5aac6e6391bb6065cd0173

Observation 57495216-2c67-4909-b418-11691882a433 · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:48.963595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:47.909998Z digest=sha256:6d965ecf18dd78a6b8c51650a0227dda313a64dfbba6233df1a327cf75a7f23a

Observation d7b40c88-fa18-4b58-80fa-ddcb9460eabb · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:48.954717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:48.014528Z digest=sha256:179fc5a433cd1b71e0c0947ac1f6777dfb37d994f56cc00f70efca043db134c7

Observation 759b882a-82ff-4b0a-935f-51874d2a2cfd · outbound

This paper cites Reinforcement Learning with Physics-Informed Symbolic Program Priors for Zero-Shot Wireless Indoor Navigation.

In-Context Reinforcement Learning via Communicative World Models Reinforcement Learning with Physics-Informed Symbolic Program Priors for Zero-Shot Wireless Indoor Navigation

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-08-05T22:39:48.706648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:48.097418Z digest=sha256:f7c5cf65fe8b46f0973c71c7969bfd7c8e695094a6de7a068369b08b264a04b5

Observation f247a299-ce7a-469b-b946-2d188b7439b8 · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:48.944439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:48.184090Z digest=sha256:b86d4284c4390ceaa2b4965a26ffaf4f991ec0598f10528e86e52c3b3b9a3f80

Observation 895a7f7d-4b4a-4d52-9de0-ac2a67659c29 · outbound

This paper cites Meta Stackelberg Game: Robust Federated Learning against Adaptive and Mixed Poisoning Attacks.

In-Context Reinforcement Learning via Communicative World Models Meta Stackelberg Game: Robust Federated Learning against Adaptive and Mixed Poisoning Attacks

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-05T22:39:48.291594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:39:48.291594Z digest=sha256:b22c4734956c4d8a7c1fcb5eeb39e2b781032dfa0f63f6b971b78ff39bee50fd

Observation 55ee3ead-8daa-4b54-8c5f-ecb2def60d42 · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:48.934449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:48.377691Z digest=sha256:e1a022f98010d0322e56753b9278feab9a0333ca5071c91d30650e847f5866a2

Observation 3a91f180-76ab-49fe-8b11-21d3f109bc7c · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:48.921774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:48.494699Z digest=sha256:87059865d46c10aec797f5561b81689efd492a81401679584c7f9c1bd18466dc

Observation e7b4880d-d47e-4c72-9213-efaede1ad5c9 · outbound

This paper cites Symbiotic Game and Foundation Models for Cyber Deception Operations in Strategic Cyber Warfare.

In-Context Reinforcement Learning via Communicative World Models Symbiotic Game and Foundation Models for Cyber Deception Operations in Strategic Cyber Warfare

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-05T22:39:48.504727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:39:48.504727Z digest=sha256:b312311e5007339d66e139897278476d69749ef1dbfda2d1bc0659eca9538433

Observation bd75d4af-314b-4241-ab98-1b8af8166b0c · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:48.910511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:48.529741Z digest=sha256:2a56d3d779e5489d486c6caaac4dd83e66bb648e514ce2b631e13163b292ead5

Observation 6939a6c7-6c5b-44c4-b73f-eccca38d9fc3 · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:48.901451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:48.533017Z digest=sha256:21f27effb16b9aee47a6011212b8db450a0be057a6d407887d0938c38cf67b83

Observation 90c37970-0813-4ca4-b93f-84bff36370f1 · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:48.890773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:48.535344Z digest=sha256:098479827a6291a9b6ac096ab1cc85b41e4990e63d73652490797a403bc222ee

Observation 3f613e76-579e-4b98-9919-1b338cde599c · outbound

This paper cites A.; Veness, J.; Bellemare, M.

In-Context Reinforcement Learning via Communicative World Models A.; Veness, J.; Bellemare, M

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:39:48.880743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:48.538125Z digest=sha256:c918ee1da61c2082302d7abf9283ee3660758d4abc8e87690cdc8bf7e81d6c09

Observation 6de6714d-47c2-47d4-8036-ac7410f5e057 · outbound

This paper cites A Survey of In-Context Reinforcement Learning.

In-Context Reinforcement Learning via Communicative World Models A Survey of In-Context Reinforcement Learning

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-05T22:39:48.540944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:39:48.540944Z digest=sha256:387dd23ac67e81272975984270a76c6da97efb17acd55eea7158c58053259b59

Observation 1f541c4c-778b-421f-b35b-057cd76ae471 · outbound

This paper cites Model-Agnostic Meta-Policy Optimization via Zeroth-Order Estimation: A Linear Quadratic Regulator Perspective.

In-Context Reinforcement Learning via Communicative World Models Model-Agnostic Meta-Policy Optimization via Zeroth-Order Estimation: A Linear Quadratic Regulator Perspective

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-08-05T22:39:48.661001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:48.543747Z digest=sha256:294be5c27876fd33611165c56d3c65773319e1b187e0fcf142902617353418e9

Observation 3b28e50e-105d-4fe1-8c90-fdcc0ccf15e7 · outbound

This paper cites NAVIX: Scaling MiniGrid Environments with JAX.

In-Context Reinforcement Learning via Communicative World Models NAVIX: Scaling MiniGrid Environments with JAX

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-05T22:39:48.546970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:39:48.546970Z digest=sha256:abd8ec8c4b8a30d9c7ec07cd0f82c63e21c020abf4c15e2e2e6deff60c5c91d0

Observation ed6db75a-4179-42c5-9aa2-ba125c89440c · outbound

This paper cites C.; Hambro, E.; Kirk, R.; Henaff, M.; and Raileanu, R.

In-Context Reinforcement Learning via Communicative World Models C.; Hambro, E.; Kirk, R.; Henaff, M.; and Raileanu, R

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:39:48.869810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:48.549714Z digest=sha256:2f7600a142275f7227235de5372323450034bc952e0057c23d597c6eca1af38e

Observation 5e5acaff-42d3-4db7-a822-ba0fb1ec2bc5 · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:48.860721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:48.552687Z digest=sha256:31477422520e0e7d86fdf6befe01768dfae26cb794a60e507dae86ba85578578

Observation 405c2cfa-8f0d-44ea-a359-f183ff4dbd0a · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:48.850632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:48.555638Z digest=sha256:8160d697173abb0daff89169790fe4e72322f47c18a1b3fc94cfdd124e757528

Observation 2caa0435-33b6-4952-afd2-e545f674f057 · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:48.841136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:48.558557Z digest=sha256:7da964fbc16a096e90078e64f20e0ee447ebee5289d0c5890f756c1d848e5f59

Observation cd0a460a-15d1-4419-8d40-2e2cf4b26fe7 · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:48.830161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:48.560952Z digest=sha256:eb8f03c3c87b60ecd9fa50f6174ddd31c086df8b9db475c7b4afaabe12fe8bab

Observation fa3466d6-f038-413f-844e-84c3a261f477 · outbound

This paper cites High-Dimensional Continuous Control Using Generalized Advantage Estimation.

In-Context Reinforcement Learning via Communicative World Models High-Dimensional Continuous Control Using Generalized Advantage Estimation

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-05T22:39:48.563655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:39:48.563655Z digest=sha256:ac57e0e7516537072f7da7deb615739f369a0c27222401028e9facca50726d33

Observation f404ca42-36cf-435b-a183-7b79ceb75c4c · outbound

This paper cites Proximal Policy Optimization Algorithms.

In-Context Reinforcement Learning via Communicative World Models Proximal Policy Optimization Algorithms

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-05T22:39:48.566736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:39:48.566736Z digest=sha256:fcb104611e92ebdad2b4ab2e846cc115575728edd953eb7f89fc9e7f67a09d4a

Observation d10855f4-9a21-4fcb-8dac-c51e2b33a7fe · outbound

This paper cites D.; Bachman, P.; and Courville, A.

In-Context Reinforcement Learning via Communicative World Models D.; Bachman, P.; and Courville, A

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:39:48.820719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:48.569515Z digest=sha256:6bd873ad2be389896dcfd2d4f7ee0875819a8e7dc7dac5c4df81dc37710aa1b3

Observation e282ff5e-fd73-459c-8400-b4111a6ede16 · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:48.810322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:48.572741Z digest=sha256:b41b28ae6b9029050361fe9a4177546580d134130cd1d44ea61752a751150fcc

Observation 5f931f02-27f7-4854-b926-dce4e171d81f · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:48.798385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:48.576513Z digest=sha256:74a79a82afed134cdf4cc7847878ad05c53838b3a71a3b42316044c263350548

Observation 09164379-eb46-4af5-bc37-74b6f3622377 · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:48.786670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:48.579690Z digest=sha256:5a4067b832275e7bfa964f8a7fdb03b7de7664b067f1f8fd69c64c546821f062

Observation 099d7c70-1b13-4b25-adf4-2f6c552c23ae · outbound

This paper cites N.; Kaiser, .; and Polosukhin, I.

In-Context Reinforcement Learning via Communicative World Models N.; Kaiser, .; and Polosukhin, I

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-05T22:39:48.582743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:39:48.582743Z digest=sha256:d05ac37d80276e382438e3eb828ebef68908726acef990707f063031c7c50ed6

Observation f28c99fb-7021-4d68-b689-5c45c6e08821 · outbound

This paper cites H.; Daneshmand, H.; and Zhang, S.

In-Context Reinforcement Learning via Communicative World Models H.; Daneshmand, H.; and Zhang, S

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:39:48.766949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:48.585289Z digest=sha256:cfec47d3bbe5ccdb7d9e86f1e809d3eade16b635979e1a1b9a37fed9adc7f097

Observation c9069b21-03e1-4432-800d-5fc3eaba7029 · outbound

This paper cites X.; Kurth-Nelson, Z.; Tirumala, D.; Soyer, H.; Leibo, J.

In-Context Reinforcement Learning via Communicative World Models X.; Kurth-Nelson, Z.; Tirumala, D.; Soyer, H.; Leibo, J

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:39:48.756783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:48.588447Z digest=sha256:56beaf4959a23d5de4b6d46c3682babc4f44f7d08a97f730d3bb6bbb6cd970dc

Observation e68a9297-65d8-4b39-a67f-08488ec2e4ec · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:48.747073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:48.591590Z digest=sha256:ed630d783368a2049263f73548216afdec34d80460d77d01751a39cc1a75e93f

Observation 6cdda72d-588d-43db-9b6e-80ffbf7722d0 · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 61

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:48.736593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:48.594243Z digest=sha256:10975b19d2188c4f333bbc067eb53fad33fd90e947d2afd74ad1023823768a11

Observation e7562ec4-0e0e-4b33-bd76-6490be2ae0f6 · outbound

This paper cites an unresolved cited work.

In-Context Reinforcement Learning via Communicative World Models Unresolved cited work

Reference 62

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:39:48.726669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T22:39:48.597389Z digest=sha256:50fc1db638ead8049ac71f8501adbf249157094549a85b450ef6d55fd10a7525

Pith citing papers

No inbound Pith citation observations are available.