Pith. sign in

Paper Citation Record · LEDGER

Training Small LLMs as Spatial Multi-Agent Policies

As of 7 August 2026, this Paper Citation Record lists 82 of 82 outbound references and 0 inbound Pith citation observations for arXiv:2608.01425.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.01425 v1

Coverage vector

measured 82 of 82 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T00:17:19.162842Z

measured 82 of 82 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

82 of 82 outbound references displayed

  • verified exact1
  • verified fuzzy33
  • unresolved48
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a97d8b63-9663-4fa8-87f7-6a9852f58212 · outbound

This paper cites Voyager: An Open-Ended Embodied Agent with Large Language Models.

Training Small LLMs as Spatial Multi-Agent Policies Voyager: An Open-Ended Embodied Agent with Large Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.689848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.689848Z digest=sha256:c9010db1a353cf056ba2b4bb84d92abaf7f03a38fe492d771302e31c36c0cf99

Observation aff42b47-9994-4087-9ba5-d6dac4fa6c1e · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.696932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.696932Z digest=sha256:e2397915b108f1cd83db46194573e93ca3201e3d2eada71669b4db460a777cb4

Observation 2de7eabb-a2ba-4075-943c-a1239c942983 · outbound

This paper cites arXiv preprint arXiv:2601.17152 , year=.

Training Small LLMs as Spatial Multi-Agent Policies arXiv preprint arXiv:2601.17152 , year=

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.703643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.703643Z digest=sha256:83f502d7f5933a062d06fc910f9949f5f2b64f5e1087df6a04cd6467cbd453b5

Observation b12b963c-b795-4405-aed9-6c0bba901d64 · outbound

This paper cites OpenReview , year=.

Training Small LLMs as Spatial Multi-Agent Policies OpenReview , year=

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.709484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.709484Z digest=sha256:0f1604eadd04bcb027ddd145efb8fe4dad87b6fb40b7ecb586f8583b0f90bbfd

Observation 6c284024-588c-4f79-8f89-4e5bd08b8576 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Training Small LLMs as Spatial Multi-Agent Policies Advances in Neural Information Processing Systems , volume=

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.714665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.714665Z digest=sha256:9bbff8f2e9e4133ac75647c7344164939714186aff1a1b056325e7210825960a

Observation 4c865463-b048-4ba8-bae9-50add8251865 · outbound

This paper cites arXiv preprint arXiv:2510.01586 , year=.

Training Small LLMs as Spatial Multi-Agent Policies arXiv preprint arXiv:2510.01586 , year=

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.720447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.720447Z digest=sha256:ed1eade9911b0c10fbc0df0a47a20eb329504c3ead19292e2168b41bed85f680

Observation 0576f503-603f-4e3f-8182-8163a01bb468 · outbound

This paper cites Closed-Loop Vision-Language Planning for Multi-Agent Coordination.

Training Small LLMs as Spatial Multi-Agent Policies Closed-Loop Vision-Language Planning for Multi-Agent Coordination

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.726467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.726467Z digest=sha256:22cbd19f1d557e6585b4a30c1cbac4add8644fd66196004ce2c7e4455b12e55a

Observation 6b6ad325-012c-4c22-9c49-247357ae114d · outbound

This paper cites UC Santa Barbara , year=.

Training Small LLMs as Spatial Multi-Agent Policies UC Santa Barbara , year=

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.868559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.732211Z digest=sha256:5183b2162ca98c57ed00c883baa848d28ef91348f403862dfe146b076a902c34

Observation e94b610e-713a-4c07-825a-0edb073fcbda · outbound

This paper cites Understanding R1-Zero-Like Training: A Critical Perspective.

Training Small LLMs as Spatial Multi-Agent Policies Understanding R1-Zero-Like Training: A Critical Perspective

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.737411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.737411Z digest=sha256:8db26c82d489a8a621e6bbc7e7d457fdfcf6436c98f88c74e77988329b1f1a13

Observation 2f03ed0d-1683-4da8-af96-da53025443df · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Training Small LLMs as Spatial Multi-Agent Policies Advances in Neural Information Processing Systems , volume=

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.743725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.743725Z digest=sha256:ffcf4b39cd804094ec51652865f6b19643a1a3a27f0abaa6ad4d8640f20069c1

Observation 1a588016-8a39-4dff-bb02-d5a245d662a8 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Training Small LLMs as Spatial Multi-Agent Policies Proximal Policy Optimization Algorithms

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.749113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.749113Z digest=sha256:dc9a8dac7647f540c9037ce5a6b14c353c4fcc6a1ab78293fbf04484fad12613

Observation 8521a6a6-5341-4500-a4d6-f3c3b96eabde · outbound

This paper cites International Conference on Machine Learning , pages=.

Training Small LLMs as Spatial Multi-Agent Policies International Conference on Machine Learning , pages=

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.840127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.758844Z digest=sha256:fdd666e6c613b1aa2537b535892952b1196f40ab4c85b965400ce02c88121f94

Observation 4840f87f-4e23-461d-8e09-ce62ef5e7172 · outbound

This paper cites NeurIPS Workshop on Foundation Models for Decision Making , year=.

Training Small LLMs as Spatial Multi-Agent Policies NeurIPS Workshop on Foundation Models for Decision Making , year=

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.822807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.769822Z digest=sha256:4945e40752f70eb1997442854292b70bd2c447b27ae71ae9cf870ce59bc0b622

Observation 281d48e2-fd5f-4f85-bb87-24c1eb599cd7 · outbound

This paper cites Proceedings of the AAAI Conference on Artificial Intelligence , year=.

Training Small LLMs as Spatial Multi-Agent Policies Proceedings of the AAAI Conference on Artificial Intelligence , year=

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.804134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.779453Z digest=sha256:78fa1041f23a288f8affa859fce85211ee1c3f17d0899846879ddac6a6ad1752

Observation 35d66fd2-3521-47fd-abe6-b9fa4f4ff948 · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year=.

Training Small LLMs as Spatial Multi-Agent Policies Advances in Neural Information Processing Systems (NeurIPS) , year=

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.787311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.785160Z digest=sha256:de5dcaef2ef840653b08949d7becd83b74d4d2af44ddbac92d3a3d8dddd1e0e8

Observation bb4a014f-d93e-4aeb-b37f-2f5256887198 · outbound

This paper cites Proceedings of the AAAI Conference on Artificial Intelligence , volume=.

Training Small LLMs as Spatial Multi-Agent Policies Proceedings of the AAAI Conference on Artificial Intelligence , volume=

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.769709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.790147Z digest=sha256:815fa72507241a9aaebd5248b9ae5fdb92c953a8718cb5cfc24552759ac97826

Observation 7d236f9d-0376-402b-bc1c-dfcefd257f76 · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year=.

Training Small LLMs as Spatial Multi-Agent Policies Advances in Neural Information Processing Systems (NeurIPS) , year=

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.751698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.795288Z digest=sha256:fa9576a9d5be8c287b5a4d9b8f66ff4b94e465a7b98bf410814ec615ead928a2

Observation b31dce2a-ab47-4d36-8973-bd6747d3c956 · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year=.

Training Small LLMs as Spatial Multi-Agent Policies Advances in Neural Information Processing Systems (NeurIPS) , year=

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.733834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.800200Z digest=sha256:005e2f0260d2dcc37549f09bc58022f6d7bc7f6c73dc79c0c61a83c04b9ebe4c

Observation d2e3b472-7bea-484b-b0fd-562c5707ebb5 · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year=.

Training Small LLMs as Spatial Multi-Agent Policies Advances in Neural Information Processing Systems (NeurIPS) , year=

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.717022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.805288Z digest=sha256:80eeb42d6ada748a2b40b2c33ea76719145e1ace568483d5cfed2954966adcad

Observation b3a1fb5d-c5e6-4ec0-b7c9-70c168ad9ac4 · outbound

This paper cites Springer , year=.

Training Small LLMs as Spatial Multi-Agent Policies Springer , year=

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.695508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.810810Z digest=sha256:e396c3b89c076ceb71c97b356773d1a5e40ea6034343904a46d8a041298db352

Observation 6e57958b-43fb-4984-a586-4443d56da4d3 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Training Small LLMs as Spatial Multi-Agent Policies LoRA: Low-Rank Adaptation of Large Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.816275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.816275Z digest=sha256:16a874693998a8b6997956f00ec1b780f2111daa19bf142fa0539c72317dfdd1

Observation c1a833ca-92fa-4c3f-b1a5-ad0d3cad9a9d · outbound

This paper cites International Conference on Machine Learning , year=.

Training Small LLMs as Spatial Multi-Agent Policies International Conference on Machine Learning , year=

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.677463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.822297Z digest=sha256:a8e0155e714c2d9c7b83469c1d7030d48e7215925cab03f174e0da1774d0bed8

Observation 351da279-b29a-4b49-b18b-d5b61739b9eb · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Training Small LLMs as Spatial Multi-Agent Policies Advances in Neural Information Processing Systems , volume=

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.828752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.828752Z digest=sha256:37458b93ddeb7e9c57d2aa3f3efda85b370a8601299b02ea71c44b451db8b647

Observation 23479c98-ff00-4dfa-992c-b82043a4e229 · outbound

This paper cites MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework.

Training Small LLMs as Spatial Multi-Agent Policies MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.833955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.833955Z digest=sha256:552c7740af4d44ad16992d82665d22d93e5c65bf7dd8d4a3a44b05a1a9f70c33

Observation 3434c501-f3ac-42d2-b063-930a559fc8ce · outbound

This paper cites ACL , year=.

Training Small LLMs as Spatial Multi-Agent Policies ACL , year=

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.647587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.839710Z digest=sha256:b77a872e2147b2ef6aa7ad51bd6a54c2420b7416cbcdfaa206f76ef9ddddfd8a

Observation 70af42e8-0a44-4878-9741-25ad902fdfc7 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Training Small LLMs as Spatial Multi-Agent Policies Advances in Neural Information Processing Systems , volume=

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.844676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.844676Z digest=sha256:7ca7e14f0e761c0fbf02201076b41fa87d8fcc79771a658267343173dd5c4ee0

Observation d67e673f-7d17-42f8-a067-bb02a803aea5 · outbound

This paper cites and Shen, Yelong and Wallis, Phillip and Allen-Zhu, Zeyuan and Li, Yuanzhi and Wang, Shean and Wang, Lu and Chen, Weizhu , booktitle=.

Training Small LLMs as Spatial Multi-Agent Policies and Shen, Yelong and Wallis, Phillip and Allen-Zhu, Zeyuan and Li, Yuanzhi and Wang, Shean and Wang, Lu and Chen, Weizhu , booktitle=

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.850417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.850417Z digest=sha256:34fb1c6129ed5db2f06abbe52c5616fc95a122f1d779f28c29775badce504735

Observation 4dfda6af-2954-4ccc-9a9c-78698920f64f · outbound

This paper cites Ghost in the Minecraft: Generally Capable Agents for Open-World Environments via Large Language Models with Text-based Knowledge and Memory.

Training Small LLMs as Spatial Multi-Agent Policies Ghost in the Minecraft: Generally Capable Agents for Open-World Environments via Large Language Models with Text-based Knowledge and Memory

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.854928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.854928Z digest=sha256:1711037ad95d616e92952489c40097e574b9fcc9e1d589c1b3c145d448f0d490

Observation a1e8b022-ed96-4af7-9f0e-770c5e5c6d99 · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:17:20.607884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.859747Z digest=sha256:155fe717bc7e956fb229b98695186d4a2c9cbf8f02ad1ae4bd5152b0c876e51c

Observation a286020b-7751-4734-ac4b-c6e17a7d02e5 · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:17:20.590631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.864434Z digest=sha256:3188f131ed4635cbee23cdbdbc43b2af5e7c4b20a3a855c9bb39db90cc8bae0f

Observation 6c585948-2051-4a80-a9e5-1eaa5c4b91e6 · outbound

This paper cites International Conference on Learning Representations (ICLR) , year=.

Training Small LLMs as Spatial Multi-Agent Policies International Conference on Learning Representations (ICLR) , year=

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.572774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.869706Z digest=sha256:2c835fe3e5674a6575373f33ba5425910415aff430efc0e12f878c3ee76118ea

Observation 80db720c-85ad-4f7f-a485-d8af810fac3f · outbound

This paper cites Conference on Robot Learning (CoRL) , year=.

Training Small LLMs as Spatial Multi-Agent Policies Conference on Robot Learning (CoRL) , year=

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.875905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.875905Z digest=sha256:ab816bad3ac42c9ab0719eea41fc123391c82dbea7508f1191af4d53b20484ab

Observation 0f972238-1c77-4327-b00b-f931b383e25b · outbound

This paper cites Conference on Robot Learning (CoRL) , year=.

Training Small LLMs as Spatial Multi-Agent Policies Conference on Robot Learning (CoRL) , year=

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.880997Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.880997Z digest=sha256:9cce29dfd4838ba125e95a74be6f89068a7d1d5bcd9fe486e021e1addc072873

Observation 3eefc640-8de1-4b56-8967-db89ca05c890 · outbound

This paper cites IEEE International Conference on Robotics and Automation (ICRA) , year=.

Training Small LLMs as Spatial Multi-Agent Policies IEEE International Conference on Robotics and Automation (ICRA) , year=

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.529731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.886656Z digest=sha256:3c0eebed906e851f3518420ca6ba44f124f3e17ccd5f73c75bc111123ad31c11

Observation 6ee843cd-c5da-4d20-8dae-407dbf1297c5 · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.891696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.891696Z digest=sha256:bf47a25ccb4e5bbe3110394fecec7fef15c03dbd5bafaf367701c4b64c274aed

Observation 50abe3fe-4824-40bc-8362-7c5e3f8c01c0 · outbound

This paper cites Efficient Guided Generation for Large Language Models.

Training Small LLMs as Spatial Multi-Agent Policies Efficient Guided Generation for Large Language Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.899014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.899014Z digest=sha256:8b5bd76765d711e8caf38618e04ddb84e35f81922d6b4952940ceb6bc6d98acb

Observation 3b668109-28ad-42f5-9ff9-993a5fa4ad31 · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year=.

Training Small LLMs as Spatial Multi-Agent Policies Advances in Neural Information Processing Systems (NeurIPS) , year=

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.501153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.904272Z digest=sha256:4412137b5a53f3999d0e3443f59cb27b971d6db9fe84c28991af26ec7b59f68e

Observation b8ed0b0c-df6a-4c95-8538-5e662421c4d8 · outbound

This paper cites 2025 , note=.

Training Small LLMs as Spatial Multi-Agent Policies 2025 , note=

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.480863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.910855Z digest=sha256:8aaed29d72009caca2ce4da7ab283d31cb6a3c2141aa49d75552176a060732de

Observation cbe7a7b1-ded3-432c-998d-8f591ead5018 · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:17:20.462829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.916313Z digest=sha256:83befddc3119931087ec2ceab2297d746df163f09345b3f7e08610ea2af7be39

Observation ea4d3b2b-2173-45cd-8185-ea230b9d4968 · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:17:20.446121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.921996Z digest=sha256:156d4cccc631109b82f9ba7d96cb2dd3aac1be943c7d90f2152f42b9e6e2c63f

Observation 099afb6d-1268-4920-b234-5036b0fe6c98 · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:17:20.430371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.927241Z digest=sha256:8c9cc62e876a7ed698b77463eaea9ee75de55970af998e5dfaebf94d11636a72

Observation b12cd2aa-80fc-4ab2-9d5f-1df7132057b6 · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year=.

Training Small LLMs as Spatial Multi-Agent Policies Advances in Neural Information Processing Systems (NeurIPS) , year=

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.413440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.933257Z digest=sha256:a4b4a6741fe3d80c513a83666c16c5303ea83b6a27cfffc54fc25079dc3f8793

Observation 892fa440-2f8a-4f92-ba0f-2d8734c00959 · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:17:20.394338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.938232Z digest=sha256:b6ea4c6fc3c56ca4941c32cb78ff64bcff594a8ac8ddbd974bf74d38bc7bc1a0

Observation 18e5a08a-154f-424b-b87e-54c41f148893 · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.943167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.943167Z digest=sha256:73a1e6e7d4a74f782b99f161f38e7a8839d6d53d2b3816e474de3f87b6a6823c

Observation 7f693115-a410-42fa-b2f0-661c294ac217 · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.948504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.948504Z digest=sha256:892e931f1ad202789c0cba22d8fc132b9bbf9700416b5958af89454513f7ae1f

Observation c093510d-5867-4870-bafd-7ca458ccfe5b · outbound

This paper cites and Chandar, Sarath and Burch, Neil and Lanctot, Marc and Song, H.

Training Small LLMs as Spatial Multi-Agent Policies and Chandar, Sarath and Burch, Neil and Lanctot, Marc and Song, H

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.355411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.953219Z digest=sha256:92e291c89f469e96813ca5523f37964ca0326eb68ff3f138a5613114ef87f7fd

Observation d7dbceaf-98ba-4b84-9951-587a4e13783d · outbound

This paper cites Proceedings of the 16th International Conference on Autonomous Agents and Multiagent Systems (AAMAS) , year=.

Training Small LLMs as Spatial Multi-Agent Policies Proceedings of the 16th International Conference on Autonomous Agents and Multiagent Systems (AAMAS) , year=

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.338289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.958388Z digest=sha256:c4d493db8706323a2694271ad746d3e1552b6b5048e2bfbe0addb35c65b8b824

Observation 90bb8644-b115-4a79-8631-2a3c3bf2a99c · outbound

This paper cites and Griffiths, Thomas L.

Training Small LLMs as Spatial Multi-Agent Policies and Griffiths, Thomas L

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.321830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.963478Z digest=sha256:d9aa418953d3091192b72378f0bc2d71e3c635b2dd8810b07ee8418bee7ea16d

Observation a0eccb0c-3296-4040-a020-e778a140476b · outbound

This paper cites Proceedings of the 36th Annual ACM Symposium on User Interface Software and Technology (UIST) , year=.

Training Small LLMs as Spatial Multi-Agent Policies Proceedings of the 36th Annual ACM Symposium on User Interface Software and Technology (UIST) , year=

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.968298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.968298Z digest=sha256:55f28e9566c1ae67ed02c2c3f0c5012c90de45904d0c824557b0129cebd5b72f

Observation 7908b1a1-b13f-4add-aed4-8534b8cded8e · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:17:20.296262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.973189Z digest=sha256:8fd3fd2798dce04c3d6ae17e1900ee3cc2f11326c7825c409b847620edad450c

Observation 96fa44f0-7e4e-40a4-b57e-ca9d028302f4 · outbound

This paper cites and Precup, Doina and Singh, Satinder , journal=.

Training Small LLMs as Spatial Multi-Agent Policies and Precup, Doina and Singh, Satinder , journal=

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.280455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.977900Z digest=sha256:e32710d7acc3617044a198db9fc973f7cde45419a15b260acd3bb54af9a1a612

Observation 0525e42e-14f9-42fb-904c-03edd38e1c91 · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:17:20.264604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.983484Z digest=sha256:ed671023e685d66f2874e8c84aa296bdb7124b34813e844cf66cbd4cd064085a

Observation 5eda816b-5dbe-453d-9aa5-5a7dfa8698d1 · outbound

This paper cites Proceedings of the AAAI Conference on Artificial Intelligence , volume=.

Training Small LLMs as Spatial Multi-Agent Policies Proceedings of the AAAI Conference on Artificial Intelligence , volume=

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.245258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.988541Z digest=sha256:1560f1e764989ba463083fba658f9cfb693f33fcb83d9a1e914336a6eeaed028

Observation 3b71affc-c92d-4308-96a8-428307d4bdca · outbound

This paper cites International Conference on Learning Representations (ICLR) , year=.

Training Small LLMs as Spatial Multi-Agent Policies International Conference on Learning Representations (ICLR) , year=

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.227093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.993534Z digest=sha256:7a96b66ffaa1dc3279328ad19518808aac83eb5f939be605a7fd82ea240782fd

Observation ed4071c0-a341-418b-95a3-59bb87720136 · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:17:20.209083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.999934Z digest=sha256:f4564f241661f43a9d9c946a6cde13160e948f4f0d4f17c1884056d84f0ac530

Observation e0b86ce6-fe95-4fe9-a2fb-5ff0681e48a5 · outbound

This paper cites Findings of the Association for Computational Linguistics: EACL , year=.

Training Small LLMs as Spatial Multi-Agent Policies Findings of the Association for Computational Linguistics: EACL , year=

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.192653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.020776Z digest=sha256:184419802221477cae054ca7e25e2dd7383c12fdc89d13c7a06504bd828a9a13

Observation 045eb470-05de-40f5-82db-752649c50f95 · outbound

This paper cites Planning with Macro-Actions in Decentralized.

Training Small LLMs as Spatial Multi-Agent Policies Planning with Macro-Actions in Decentralized

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.177195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.026518Z digest=sha256:04e3226a30243ef2df2e60adaa4d8e7bd5fd3c78b443ec3ff4f802d297ba8eff

Observation 208abd0b-4dc1-49f1-9c62-0b32969abecb · outbound

This paper cites Proceedings of the 3rd Conference on Robot Learning (CoRL) , series =.

Training Small LLMs as Spatial Multi-Agent Policies Proceedings of the 3rd Conference on Robot Learning (CoRL) , series =

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.161210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.031859Z digest=sha256:9bd247437912e1b527eec35e5aa97c4f2d663e10d98ecad1118d18d29b7cadac

Observation 6fa2ef9f-167e-4585-883b-af4bfbf96553 · outbound

This paper cites Melting Pot 2.0.

Training Small LLMs as Spatial Multi-Agent Policies Melting Pot 2.0

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:19.036828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:19.036828Z digest=sha256:4a83889708bcb62ed01b64ea3c0e5608b6eef334f05141e69cce4bcf8ba4b380

Observation 4446c246-3a87-4cab-8a68-bba161cd121a · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 66

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:17:20.144954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.042331Z digest=sha256:76c4dcf8839559c5a95e1ccb5630d7a2f0092da2ede446f61ce07b11b866fc5f

Observation ca5ef68b-b704-4e09-b427-a00fb561974a · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 67

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:17:20.129244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.047580Z digest=sha256:3a4d33e39d6f39f2aa38dbda25a7f0419cf1313a052e027a7e782d7c5c18d09b

Observation 37a23397-6dbb-4921-98ba-c35f6f86013b · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 68

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:17:20.112323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.052882Z digest=sha256:d44c46ebe7da8f2283241cc1c8d92e97586fb09fcdb6ada76c497a97614c62f9

Observation be1ee015-6b03-40fa-8258-0f241e4d48e2 · outbound

This paper cites K.; Griffiths, T.

Training Small LLMs as Spatial Multi-Agent Policies K.; Griffiths, T

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.095697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.058640Z digest=sha256:cd937918b793a14505ed179f48529023b1281530f0c1a67d3ed01fdf59da331f

Observation f51176cd-a51d-4833-92ac-0f9649304b43 · outbound

This paper cites Beyond Scalar Rewards: Dense Feedback for LLM Policy Synthesis in Sequential Social Dilemmas.

Training Small LLMs as Spatial Multi-Agent Policies Beyond Scalar Rewards: Dense Feedback for LLM Policy Synthesis in Sequential Social Dilemmas

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:19.064587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:19.064587Z digest=sha256:3b117df3508fa1e29c4dfc4b8546bf454a180598691cbb859eb4e8d8a4abbc66

Observation 074dcf2b-4495-4183-9722-abe3ed2c6eb1 · outbound

This paper cites J.; Shen, Y.; Wallis, P.; Allen-Zhu, Z.; Li, Y.; Wang, S.; Wang, L.; and Chen, W.

Training Small LLMs as Spatial Multi-Agent Policies J.; Shen, Y.; Wallis, P.; Allen-Zhu, Z.; Li, Y.; Wang, S.; Wang, L.; and Chen, W

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.079512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.069699Z digest=sha256:f195d2ff90cffcd4a0bd39519cdd50369c120c718dc31cff9c89163bc3b67abe

Observation 6e6c1218-4dcb-43f3-96ef-eef2c6215eec · outbound

This paper cites Shapley-Coop: Credit Assignment for Emergent Cooperation in Self-Interested LLM Agents.

Training Small LLMs as Spatial Multi-Agent Policies Shapley-Coop: Credit Assignment for Emergent Cooperation in Self-Interested LLM Agents

Reference 72

Resolution
verified exact
local_arxiv, observed 2026-08-06T00:17:19.440446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.076571Z digest=sha256:60b7554e3bdefbf941e380c3559d6027fce204d5995a584b84fa85093ccc34ed

Observation c3d2e64b-13ae-4f2b-9695-637cb1ae01da · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 73

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:17:20.063475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.082763Z digest=sha256:03031b875641b6eb60f4e1b3695e624ddc11f077cdc1b875d1d99d23d581a768

Observation 5f08f481-6680-4341-8cc4-ff7a8559f6f3 · outbound

This paper cites Z.; Phillips, M.; Tuyls, K.; Du \'e \ n ez-Guzm \'a n, E.

Training Small LLMs as Spatial Multi-Agent Policies Z.; Phillips, M.; Tuyls, K.; Du \'e \ n ez-Guzm \'a n, E

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.047296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.088260Z digest=sha256:115b8dedf5d82470a86fbdb9e392c2b5c92ccc0799a08dba9ee3700cee839f9c

Observation 4e82f720-32cc-46fc-bd4a-9162abb3a677 · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:19.093339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:19.093339Z digest=sha256:86c6496548be2a1c1726b6d425b416298f0472fdeeceffc0e3d6384b7c721e2c

Observation 7b0df47a-78fc-41e5-9a9d-ef91e1a0e6ac · outbound

This paper cites Z.; Zambaldi, V.; Lanctot, M.; Marecki, J.; and Graepel, T.

Training Small LLMs as Spatial Multi-Agent Policies Z.; Zambaldi, V.; Lanctot, M.; Marecki, J.; and Graepel, T

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.030609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.098155Z digest=sha256:e7787983102257dd0d3360591043b01f65c38a3883d1fb2aa02155e51b896ff0

Observation 4ef82487-1800-488d-8341-243b57cc5ee5 · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 77

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:17:20.013919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.103067Z digest=sha256:a5ae3e1ba33c29bc5c0d32f9b468bb39151af4cd6084e476f7843df16f5500aa

Observation 9fb159ba-b3df-4cd0-b1b3-dfc0ae6029be · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:19.109264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:19.109264Z digest=sha256:05e6153b17ce756b3234bd55f1e01bf75b7b52b613cea83563489fb710254361

Observation 3b9ef1fc-705d-4fcb-b573-359e092ebda6 · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 79

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:17:19.996489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.115120Z digest=sha256:065c1081500ff6971caf5f9a2e5f932371b3d07e5dd99dac98011f7d7946a417

Observation 393301f8-4665-4d83-9de7-38482675db31 · outbound

This paper cites S.; Rios, M.; Fonseca, Y.; Giraldo, L.

Training Small LLMs as Spatial Multi-Agent Policies S.; Rios, M.; Fonseca, Y.; Giraldo, L

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:19.975414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.120501Z digest=sha256:503430d43caffc7a372a35c9a65b99f894005f0011db8c88c0054e912760ee2b

Observation c5f27482-1d88-405c-9225-5ccec37cbaf9 · outbound

This paper cites Z.; Zambaldi, V.; Beattie, C.; Tuyls, K.; and Graepel, T.

Training Small LLMs as Spatial Multi-Agent Policies Z.; Zambaldi, V.; Beattie, C.; Tuyls, K.; and Graepel, T

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:19.958505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.126320Z digest=sha256:31280056d99fd004c48392bf02beacefe8d5e75c66ba325c9ed7e72505803f32

Observation a032285e-e12f-460b-8ab8-5d5fd80d6392 · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:19.131642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:19.131642Z digest=sha256:ae5f42a42d93e56f1e72a8a1b6019153eeed0a1e0fda1cbdeb924bd641744fca

Observation 23aee3d0-7425-4fee-b340-f63ba4738826 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Training Small LLMs as Spatial Multi-Agent Policies DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:19.136940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:19.136940Z digest=sha256:3442aa66d2eba3bcf0f49bd980068df6bb4044954450bb02920cafc64d70afb9

Observation bf7d81a3-36db-414c-ad10-6dcefc65d953 · outbound

This paper cites S.; Precup, D.; and Singh, S.

Training Small LLMs as Spatial Multi-Agent Policies S.; Precup, D.; and Singh, S

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:19.941061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.141764Z digest=sha256:d130ec167e4fae16fc43a1976c3be92b59f0af839fec5dc1af18f3e6aba1716f

Observation a35f7338-3833-459b-9157-b94ed8552733 · outbound

This paper cites On the Planning Abilities of Large Language Models : A Critical Investigation.

Training Small LLMs as Spatial Multi-Agent Policies On the Planning Abilities of Large Language Models : A Critical Investigation

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:19.146958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:19.146958Z digest=sha256:7146dbc86a5a0dfbcfecde6a3ed6fc64bb7466278b1f0875b787e12e9589109b

Observation 177a0a35-6a89-47ef-9c81-af255d2df0de · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 86

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:17:19.924618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.152230Z digest=sha256:f3c945ae3545073cf55ede13f6c49aff91a3dc5d95d3e3c2cab8152a0c1058a9

Observation 34d0d686-e323-4f85-ad92-f58c18a118ed · outbound

This paper cites Low-Rank Agent-Specific Adaptation (LoRASA) for Multi-Agent Policy Learning.

Training Small LLMs as Spatial Multi-Agent Policies Low-Rank Agent-Specific Adaptation (LoRASA) for Multi-Agent Policy Learning

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:19.157771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:19.157771Z digest=sha256:e445aea0daa0281ddf7f0bb0c0170437c106d9c10695e1a99857e73b80cbe506

Observation 747efd9e-4ff6-476a-82c9-b01992e752e3 · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 88

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:17:19.906827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.162842Z digest=sha256:fff6ea8d168fb08aedfabc8886c47bcc268789f5741c64b075924978c93beb08

Pith citing papers

No inbound Pith citation observations are available.