Pith. sign in

Paper Citation Record · LEDGER

Robust Autonomy Emerges from Self-Play

As of 21 August 2026, this Paper Citation Record lists 79 of 79 outbound references and 25 inbound Pith citation observations for arXiv:2502.03349.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.03349 v1

Coverage vector

measured 79 of 79 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T05:08:33.099836Z

measured 104 of 104 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 25 of 25 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T12:06:02.123061Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

79 of 79 outbound references displayed

  • verified exact4
  • verified fuzzy48
  • unresolved26
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation c560996d-532e-4dfa-9a22-f531aa64e2e3 · outbound

This paper cites write newline.

Robust Autonomy Emerges from Self-Play write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-09T05:08:31.280463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T05:08:31.280463Z digest=sha256:dc6b4e32a94fbd13988cd104f618a2ff4ff8511cc2b24fe20d6dfa4db41da0fb

Observation bc938c67-d263-409a-84ff-2964299347bb · outbound

This paper cites PyTorch 2 : Faster machine learning through dynamic python bytecode transformation and graph compilation.

Robust Autonomy Emerges from Self-Play PyTorch 2 : Faster machine learning through dynamic python bytecode transformation and graph compilation

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:37.304921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:31.288911Z digest=sha256:041178ae73605b1c1330526e97a65b3e2c544b63f067c8786ca0b838523c7a37

Observation d2da4e02-6da2-47cc-a3ff-e13254fa9dc0 · outbound

This paper cites Dota 2 with Large Scale Deep Reinforcement Learning.

Robust Autonomy Emerges from Self-Play Dota 2 with Large Scale Deep Reinforcement Learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-09T05:08:31.300900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T05:08:31.300900Z digest=sha256:ca9ed933c2eb31b035a7608ccee09c003d636ad832b9dbd2bb674b33c735613f

Observation d57ff388-c54f-4f81-83a9-cbdd7f4f6e71 · outbound

This paper cites and Sandholm, T.

Robust Autonomy Emerges from Self-Play and Sandholm, T

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-09T05:08:31.307317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T05:08:31.307317Z digest=sha256:53eab98910dd4afd0c30851e8dbe14db5b989054926e6ce847f972763c540403

Observation a2cbb4db-98b9-4960-bde6-656b0acaecf1 · outbound

This paper cites H., Vora, S., Liong, V.

Robust Autonomy Emerges from Self-Play H., Vora, S., Liong, V

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:37.287964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:31.316655Z digest=sha256:cd8cd80afc7a527639df7b5c0e32bbad26e33ce0d8f40dd2433b576680014844

Observation 4fc03a67-4b04-4823-adc6-d38e1115946c · outbound

This paper cites NuPlan: A closed-loop ML-based planning benchmark for autonomous vehicles.

Robust Autonomy Emerges from Self-Play NuPlan: A closed-loop ML-based planning benchmark for autonomous vehicles

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-09T05:08:31.322144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T05:08:31.322144Z digest=sha256:18cb9046534e903129c21f9a4807a9858cfa830e014044e749f74e0db2ec248f

Observation f38f5efc-e40f-4542-ae90-d7c37106e588 · outbound

This paper cites MP3 : A unified model to map, perceive, predict and plan.

Robust Autonomy Emerges from Self-Play MP3 : A unified model to map, perceive, predict and plan

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:37.230258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:31.327925Z digest=sha256:2db1a3c50c51dbb069bea300cff82c4fac063dae0157741131d7a582d34a2b84

Observation 5135be55-04ea-4c97-9b18-634343eaf5b0 · outbound

This paper cites a henb \.

Robust Autonomy Emerges from Self-Play a henb \

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:37.197066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:31.335494Z digest=sha256:3e1ee908ee1f3c5ac7909a3a1703456e2ad963dff5e3fac82ead49a9b44c71b9

Observation 979f5309-117a-4e10-a3cc-7d0547270b34 · outbound

This paper cites End-to-end Autonomous Driving: Challenges and Frontiers.

Robust Autonomy Emerges from Self-Play End-to-end Autonomous Driving: Challenges and Frontiers

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-09T05:08:31.344975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T05:08:31.344975Z digest=sha256:ec534a2fc7103390308fa8ca115062b310261c8218b03697a396ebf713d5d844

Observation f6390ece-7529-405a-bca8-7de7c3567bd4 · outbound

This paper cites TransFuser : Imitation with transformer-based sensor fusion for autonomous driving.

Robust Autonomy Emerges from Self-Play TransFuser : Imitation with transformer-based sensor fusion for autonomous driving

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:37.178189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:31.430170Z digest=sha256:3d5ac334119de876f40e658cd25c758bd8aae5340ee0a427131771b402c1a005

Observation d4b7e2e8-011f-4a45-900b-f441fd3c00d2 · outbound

This paper cites Parting with misconceptions about learning-based vehicle motion planning.

Robust Autonomy Emerges from Self-Play Parting with misconceptions about learning-based vehicle motion planning

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:37.075055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:31.489514Z digest=sha256:847d585a7b894e8ed5ab40977d3a5d3a8e8181a59a70944e7bc86ce66bf4b7cd

Observation 64c7d806-2ca5-48b9-9130-4f7da7271df4 · outbound

This paper cites an unresolved cited work.

Robust Autonomy Emerges from Self-Play Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-09T05:08:37.057545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:31.562737Z digest=sha256:e21850a4faada48a1170f2654fba2d4e3eaeec9a13f57ece358bd389810d20f8

Observation b7db3ee7-d395-47ed-a420-e83738cb4532 · outbound

This paper cites and Koltun, V.

Robust Autonomy Emerges from Self-Play and Koltun, V

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:36.998204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:31.662951Z digest=sha256:3094535727ff67cccf68d3237be4bffacb2d74b980d0d88d7df50cbe727fac32

Observation 02d02b73-ce5c-4b52-ad87-38fddd4028be · outbound

This paper cites M., and Koltun, V.

Robust Autonomy Emerges from Self-Play M., and Koltun, V

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:36.855683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:31.760252Z digest=sha256:5c908e45fcecae2a90600c2b3b6de04d0cf446a4429ad058cb2692314e86bf58

Observation 1742629c-563e-43f3-bb73-3bbf3f6e444b · outbound

This paper cites The Llama 3 Herd of Models.

Robust Autonomy Emerges from Self-Play The Llama 3 Herd of Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-09T05:08:31.765716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T05:08:31.765716Z digest=sha256:02e01df9677df9ea9e9d5a6aadb06f30144f1b3c635208d4eea22b8bc8cd74e8

Observation d6be0833-dff3-4506-b24a-b71665e5f733 · outbound

This paper cites R., Zhou, Y., Yang, Z., Chouard, A., Sun, P., Ngiam, J., Vasudevan, V., McCauley, A., Shlens, J., and Anguelov, D.

Robust Autonomy Emerges from Self-Play R., Zhou, Y., Yang, Z., Chouard, A., Sun, P., Ngiam, J., Vasudevan, V., McCauley, A., Shlens, J., and Anguelov, D

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:36.839015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:31.770999Z digest=sha256:c2a72ed040fc751374fe5a93bc150431ab9abacbd208c47b3ae9e2ac68f4c012

Observation 7301a986-253a-4e1c-ba90-947469b7d0ef · outbound

This paper cites an unresolved cited work.

Robust Autonomy Emerges from Self-Play Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-09T05:08:31.775962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T05:08:31.775962Z digest=sha256:ccc46d937c46c5dd5ad521d0eeab97cec4e42be9add71b45eff30b686e39620d

Observation 7c3e9edf-ee13-4ffe-b4af-ac859ba159d2 · outbound

This paper cites Establishing a crash rate benchmark using large-scale naturalistic human ridehail data.

Robust Autonomy Emerges from Self-Play Establishing a crash rate benchmark using large-scale naturalistic human ridehail data

Reference 18

Resolution
malformed identifier
doi_truncated, observed 2026-08-09T05:08:33.219886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:31.780939Z digest=sha256:5735c948ed3298285a9d9ac94ef30fdab58f30c37851116ba542cab4cd33e998

Observation adb09096-11da-4dab-9ffc-b2af6d3ac3a1 · outbound

This paper cites D., Frey, E., Raichuk, A., Girgin, S., Mordatch, I., and Bachem, O.

Robust Autonomy Emerges from Self-Play D., Frey, E., Raichuk, A., Girgin, S., Mordatch, I., and Bachem, O

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:36.821983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:31.787321Z digest=sha256:4f23ca5d4be233ce7e0651a68ce28a0b9890c28835024f6b5d761e27cfc5460d

Observation 9b209c5e-15b2-4520-8f06-d315c231556b · outbound

This paper cites L., and Baxter, J.

Robust Autonomy Emerges from Self-Play L., and Baxter, J

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:36.805844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:31.792592Z digest=sha256:fa662bb05187b6494cea5a7dcd5f0f958fbcaf1f22a013cc854d2e2e40daed00

Observation 69322343-643a-40bb-9239-144990206207 · outbound

This paper cites D., Agarwal, R., Roelofs, R., Lu, Y., Montali, N., Mougin, P., Yang, Z., White, B., Faust, A., McAllister, R., Anguelov, D., and Sapp, B.

Robust Autonomy Emerges from Self-Play D., Agarwal, R., Roelofs, R., Lu, Y., Montali, N., Mougin, P., Yang, Z., White, B., Faust, A., McAllister, R., Anguelov, D., and Sapp, B

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:36.785332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:31.798825Z digest=sha256:5429a07528c733d87bc421f91fd0f8fc2b15a2d34812b51de7d6d7442a66bf32

Observation 2e30f332-cac3-41e4-84d9-85a71342b334 · outbound

This paper cites SceneDM: Scene-level Multi-agent Trajectory Generation with Consistent Diffusion Models.

Robust Autonomy Emerges from Self-Play SceneDM: Scene-level Multi-agent Trajectory Generation with Consistent Diffusion Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-09T05:08:31.803663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T05:08:31.803663Z digest=sha256:86f2e7a1c802fe601604bdad87504f3c3aee8e26b11a192a9d4455e8d4a53069

Observation 1a1ccbed-5cb6-43a3-87fe-800073ea520d · outbound

This paper cites From prediction to planning with goal conditioned lane graph traversals.

Robust Autonomy Emerges from Self-Play From prediction to planning with goal conditioned lane graph traversals

Reference 23

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-09T05:08:34.697741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:31.808533Z digest=sha256:a80982f8fcc5e9758f4e6a0cbe44dfccb77fdefbb3f7ca92cd16e962b9730fcc

Observation 1beda1b5-a086-4ae8-93e3-3653f3563a4a · outbound

This paper cites A., Bernstein, D.

Robust Autonomy Emerges from Self-Play A., Bernstein, D

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:36.763465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:31.813694Z digest=sha256:00ffedcc3cfe916218e931f2dc1952f185a6a1c9e02498728ed8fb086fdb6fd2

Observation 98a6f3ff-2d90-41f8-aa2f-1b030b972d20 · outbound

This paper cites Reimagining an autonomous vehicle.

Robust Autonomy Emerges from Self-Play Reimagining an autonomous vehicle

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-09T05:08:31.819460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T05:08:31.819460Z digest=sha256:f928f492bba5b1669164ba4dd7730e62b4aeacc0848f2712d3c45454285a61dd

Observation d822a172-e9fa-49d7-ab85-d45f95de21dc · outbound

This paper cites GAIA-1: A Generative World Model for Autonomous Driving.

Robust Autonomy Emerges from Self-Play GAIA-1: A Generative World Model for Autonomous Driving

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-09T05:08:31.824252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T05:08:31.824252Z digest=sha256:07baa7abfd40d40b00d5866bbde763e5a8b2f522ff27f5956b8dad34b02c89e7

Observation 6523d3fe-7f3d-4149-ab23-edbc1ed4a575 · outbound

This paper cites an unresolved cited work.

Robust Autonomy Emerges from Self-Play Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-09T05:08:31.860092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T05:08:31.860092Z digest=sha256:f396c9a72390b9afdd9016d8883cfffa8c12f2d26c05a2d5aba08c586cab5925

Observation a6cd2b5e-eb20-4354-adf0-421726636f69 · outbound

This paper cites and Pavone, M.

Robust Autonomy Emerges from Self-Play and Pavone, M

Reference 28

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-09T05:08:34.150562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:31.901651Z digest=sha256:85b0576f5732268297d041d7d0012219be8d1fc7f1df12c7e5039068a378edde

Observation 2adf9542-c9af-4021-8e89-d2891c9df044 · outbound

This paper cites Population Based Training of Neural Networks.

Robust Autonomy Emerges from Self-Play Population Based Training of Neural Networks

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-09T05:08:31.941551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T05:08:31.941551Z digest=sha256:82459b2a705247bcd5e14ec6f587dec767db88589305afe741c6974c58b32d1d

Observation 03b86e64-be43-4188-9054-c9ef0b664f03 · outbound

This paper cites M., Dunning, I., Marris, L., Lever, G., Castañeda, A.

Robust Autonomy Emerges from Self-Play M., Dunning, I., Marris, L., Lever, G., Castañeda, A

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-09T05:08:32.057143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T05:08:32.057143Z digest=sha256:6963c5b4c0e7fa57b55de8bc7def4138b9f48d0ba00ddf762c8d3225d958fc6c

Observation 911f755b-7d74-4189-926e-bac5b0d6d35a · outbound

This paper cites Hidden biases of end-to-end driving models.

Robust Autonomy Emerges from Self-Play Hidden biases of end-to-end driving models

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:36.570025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:32.104948Z digest=sha256:fcd6e5c0e50e3a3277e831beae6064def59fb472e81fb85149623db0335d50b9

Observation 70b12475-9228-4962-8074-2b3e9e245365 · outbound

This paper cites Autonomy 2.0: Why is self-driving always 5 years away?.

Robust Autonomy Emerges from Self-Play Autonomy 2.0: Why is self-driving always 5 years away?

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-09T05:08:32.144768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T05:08:32.144768Z digest=sha256:81e2e6a039ea10a80a3400634b2ebebde5fb169e06f5decdc2e7e1d190896c57

Observation 8dac1b9b-b608-42eb-b7fa-be8b07b672a2 · outbound

This paper cites Champion-level drone racing using deep reinforcement learning.

Robust Autonomy Emerges from Self-Play Champion-level drone racing using deep reinforcement learning

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:36.470284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:32.162560Z digest=sha256:948799ad7db960739db8b9e97d9f73b63b42427574e9b038a9d455edb3438814

Observation ca8ef44d-86ef-4ec5-92d3-fe2c2218987a · outbound

This paper cites General lane-changing model mobil for car-following models.

Robust Autonomy Emerges from Self-Play General lane-changing model mobil for car-following models

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:36.339360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:32.168744Z digest=sha256:87ed6ed114c56ace737a7dae751ce9fcccd9ad8df1d2a98d2bb53a670da65043

Observation f8059699-dc48-4ac2-af47-0d5f04a34666 · outbound

This paper cites Enhanced intelligent driver model to access the impact of driving strategies on traffic capacity.

Robust Autonomy Emerges from Self-Play Enhanced intelligent driver model to access the impact of driving strategies on traffic capacity

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:36.320111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:32.175264Z digest=sha256:0e848bb1978687ce0782b2aeab709d765d25aa2a061e5fc6624390dcd6125b82

Observation 278f60fc-900d-49b2-9391-1c14f31608bd · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Robust Autonomy Emerges from Self-Play Adam: A Method for Stochastic Optimization

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-09T05:08:32.184692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T05:08:32.184692Z digest=sha256:595673a07c0ec90d8302ef19fb16424b127427b936d14be3d988802f2d887e2d

Observation ef899268-aa00-41da-a4c5-39f6b246efb8 · outbound

This paper cites Learning quadrupedal locomotion over challenging terrain.

Robust Autonomy Emerges from Self-Play Learning quadrupedal locomotion over challenging terrain

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:36.302113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:32.190340Z digest=sha256:389a848ecd336a1ccbc0e14a02cfd32e0f570dc4894837bf1060945a60aea9b4

Observation 743119ef-d968-4b2e-99fc-9e82467ed1e7 · outbound

This paper cites Metadrive: Composing diverse driving scenarios for generalizable reinforcement learning.

Robust Autonomy Emerges from Self-Play Metadrive: Composing diverse driving scenarios for generalizable reinforcement learning

Reference 38

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-09T05:08:33.931627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:32.195248Z digest=sha256:a6bb69b6d6f5fcfa5141a6bf122abb52da042c3c45006d6a898880248d56a613

Observation 47277fe1-4ab4-4e40-86f2-258a9b9e00ed · outbound

This paper cites Imitation is not enough: Robustifying imitation with reinforcement learning for challenging driving scenarios.

Robust Autonomy Emerges from Self-Play Imitation is not enough: Robustifying imitation with reinforcement learning for challenging driving scenarios

Reference 39

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-09T05:08:33.617331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:32.200506Z digest=sha256:0fa976b0fb66a89e5cbfcbe47ecad63a8541dfbd0558645b1324447b1d050f4d

Observation 43c8fe92-ee9d-42f4-a0f2-203352bc3f32 · outbound

This paper cites Isaac gym: High performance GPU based physics simulation for robot learning.

Robust Autonomy Emerges from Self-Play Isaac gym: High performance GPU based physics simulation for robot learning

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:36.283094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:32.206460Z digest=sha256:c836b079e956b3b3da43c1dc0060b4143533e84924b5e0686978b603f9430101

Observation 5712b127-e18f-45f5-9ba3-a1e59141203b · outbound

This paper cites Z., Whiteson, S., White, B., and Anguelov, D.

Robust Autonomy Emerges from Self-Play Z., Whiteson, S., White, B., and Anguelov, D

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:36.262953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:32.238463Z digest=sha256:ef5fb525824ffd901273c53a267c990a54f704ed52e9debb80eb58e6096fc803

Observation bd217bb1-d4d5-4626-9807-68132cad3ce6 · outbound

This paper cites Driving policy transfer via modularity and abstraction.

Robust Autonomy Emerges from Self-Play Driving policy transfer via modularity and abstraction

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:36.243432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:32.245689Z digest=sha256:86be46b2d23f9ae31fcefd49b78e99dc1fe1a4a0dc91c1afdd18749981dd1df4

Observation 419a879a-8999-48b1-97b1-40896a3a6a78 · outbound

This paper cites D., and Yanagisawa, M.

Robust Autonomy Emerges from Self-Play D., and Yanagisawa, M

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:36.224796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:32.251134Z digest=sha256:90d9047e4e1660941b4ef93d9228cf50cdd8f62249e093bd125af45f35604918

Observation c076c960-2818-44fc-ad46-961bdbde9f3e · outbound

This paper cites S., and Sapp, B.

Robust Autonomy Emerges from Self-Play S., and Sapp, B

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:36.073939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:32.255968Z digest=sha256:e9f7aa4f85aa9ebf659c550623e170750ad9d001ecafbb344ef59eeaca10ecee

Observation 0b2923a5-782c-4498-ac13-2dce5bd14bf8 · outbound

This paper cites Neural scene graphs for dynamic scenes.

Robust Autonomy Emerges from Self-Play Neural scene graphs for dynamic scenes

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:35.923307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:32.262086Z digest=sha256:cd85f852e7f7c21686b48ec25f5691c81f38f368ef9906a229ab410af4227acf

Observation 556fe3c0-92ce-4e07-8f9f-870612e84298 · outbound

This paper cites D., Hennes, D., Tarassov, E., Strub, F., de Boer, V., Muller, P., Connor, J.

Robust Autonomy Emerges from Self-Play D., Hennes, D., Tarassov, E., Strub, F., de Boer, V., Muller, P., Connor, J

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-09T05:08:32.267490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T05:08:32.267490Z digest=sha256:74486b74cc365b51a7be6c304f0481bef2b172167974fca6f98532482d074b71

Observation 313ef77b-7ea2-4d64-a713-652a6b113587 · outbound

This paper cites Megaverse: Simulating embodied agents at one million experiences per second.

Robust Autonomy Emerges from Self-Play Megaverse: Simulating embodied agents at one million experiences per second

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:35.830489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:32.273152Z digest=sha256:8da80193b220fc853b83a9a15db4ca4b641a4dc54cf2c6853a257116f68de72b

Observation b156634d-56be-43cc-9511-8f87e69ab864 · outbound

This paper cites PDexPBT : Scaling up dexterous manipulation for hand-arm systems with population based training.

Robust Autonomy Emerges from Self-Play PDexPBT : Scaling up dexterous manipulation for hand-arm systems with population based training

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:35.814928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:32.285725Z digest=sha256:f84a127ec59c05f0f48698e20e7cfd5945c67a08c584f584d7d882fb6d716139

Observation 2913390b-16b9-4245-8924-a9d0e6e2856f · outbound

This paper cites B., and Fidler, S.

Robust Autonomy Emerges from Self-Play B., and Fidler, S

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:35.800133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:32.365954Z digest=sha256:1d20cbce59d63fde790ff569662e486489319b21064cae2ea217d6dc4bcec934

Observation 13adea34-0abd-4ef1-935a-8da5468ed2d9 · outbound

This paper cites Asymmetric self-play for automatic goal discovery in robotic manipulation.

Robust Autonomy Emerges from Self-Play Asymmetric self-play for automatic goal discovery in robotic manipulation

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-09T05:08:32.433492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T05:08:32.433492Z digest=sha256:38f1b81775f21c9f9e5c3437003b8f3ccaef5ef7e1f7107e0ec806af7ae5dc81

Observation 86133d92-ea25-41bd-bcb2-3aadd043a929 · outbound

This paper cites A simple yet effective method for simulating realistic multi-agent behaviors.

Robust Autonomy Emerges from Self-Play A simple yet effective method for simulating realistic multi-agent behaviors

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:35.785118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:32.470788Z digest=sha256:76233b1a0eae3754ce117cfb52c7644b177fda94b2baa6cfbb6f8156a122b822

Observation 4a590291-3890-46e9-8703-167f750dad8a · outbound

This paper cites Stable-baselines3: Reliable reinforcement learning implementations.

Robust Autonomy Emerges from Self-Play Stable-baselines3: Reliable reinforcement learning implementations

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:35.768403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:32.514949Z digest=sha256:0a1a9d453783448327edf271d6a31a1b31f151585bcf0f21dee9a5cbb1c09973

Observation 53975d24-fb29-4b31-b39e-1613ce43bfaa · outbound

This paper cites S., Akata, Z., and Geiger, A.

Robust Autonomy Emerges from Self-Play S., Akata, Z., and Geiger, A

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:35.751815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:32.556107Z digest=sha256:0be0d1f190f90926560e2c7a20b06f864a3ca97f5769e3ec6ccfccdf35bd1fea

Observation 4833ac4b-fac7-4e44-a69e-622ca3863320 · outbound

This paper cites Learning to walk in minutes using massively parallel deep reinforcement learning.

Robust Autonomy Emerges from Self-Play Learning to walk in minutes using massively parallel deep reinforcement learning

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:35.729327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:32.604900Z digest=sha256:2170d9dd114118e16b44e095bd1eeb94288fa9659af76c0bdee8a35869ea871a

Observation ce0d0d04-a84f-4348-aec5-c18be5763990 · outbound

This paper cites Prioritized experience replay.

Robust Autonomy Emerges from Self-Play Prioritized experience replay

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:35.698364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:32.609554Z digest=sha256:7ba12b3df297212cf766f04d5123be86418e59bade5be393a1afc911dafb4a0a

Observation 16e4460a-caf9-475e-8332-fc5cb91c0da9 · outbound

This paper cites Urban driver: Learning to drive from real-world demonstrations using policy gradients.

Robust Autonomy Emerges from Self-Play Urban driver: Learning to drive from real-world demonstrations using policy gradients

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:35.481238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:32.613919Z digest=sha256:d75c7e5b3fd52f05ecc9f2139d6f7f6f277511dcb0615674aad47a861ca51538

Observation d893577c-fc7c-4482-b928-ce7b94a8332f · outbound

This paper cites I., and Abbeel, P.

Robust Autonomy Emerges from Self-Play I., and Abbeel, P

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:35.402447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:32.618526Z digest=sha256:9acfbdba03e2ab60e7e4cb275ec1780a95ea5b22d838837fb3f2d528094b135a

Observation c64ce62c-9ee7-41b2-9655-ab83c59b89b3 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Robust Autonomy Emerges from Self-Play Proximal Policy Optimization Algorithms

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-09T05:08:32.623266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T05:08:32.623266Z digest=sha256:7fa1036eb9ff5c1ab9f0a196d9add3441e769ec8a3c1d15314a7d5c0973d1fd3

Observation 68cfe68f-a755-457e-96a9-bf706128fd0b · outbound

This paper cites Large batch simulation for deep reinforcement learning.

Robust Autonomy Emerges from Self-Play Large batch simulation for deep reinforcement learning

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:35.385229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:32.628727Z digest=sha256:ed47fe39482b6aad9cd2f757bf2ffa28ae8f9245d815aad1e360f82847a5865b

Observation 890c8775-cd6d-4eca-ae57-d56e52a051ab · outbound

This paper cites G., Xie, Z., Sarkar, B., Szot, A., Wijmans, E., Koltun, V., Batra, D., and Fatahalian, K.

Robust Autonomy Emerges from Self-Play G., Xie, Z., Sarkar, B., Szot, A., Wijmans, E., Koltun, V., Batra, D., and Fatahalian, K

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:35.366424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:32.633421Z digest=sha256:603ac3a27c098feed82e7b4f019435a2b0e93648794544be0fe40a7b1572a974

Observation f5ab747f-b612-40c5-824b-38437f505a33 · outbound

This paper cites Motion transformer with global intention localization and local movement refinement.

Robust Autonomy Emerges from Self-Play Motion transformer with global intention localization and local movement refinement

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:35.347608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:32.638000Z digest=sha256:dbfdb966981ac9a11e279bd7928a5790f55e353c82697b523b16604d85116c6a

Observation 63ebb15c-ddd6-400a-bf9e-739cfa3bed2d · outbound

This paper cites P., Hui, F., Sifre, L., van den Driessche, G., Graepel, T., and Hassabis, D.

Robust Autonomy Emerges from Self-Play P., Hui, F., Sifre, L., van den Driessche, G., Graepel, T., and Hassabis, D

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-09T05:08:32.642587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T05:08:32.642587Z digest=sha256:37c210abc0765298458b151bf9be00cd66a6f0ac42b5d57f5322274924cb1068

Observation 32a976c8-3cd8-41af-9244-2a03b6b17b23 · outbound

This paper cites A general reinforcement learning algorithm that masters chess, shogi, and Go through self-play.

Robust Autonomy Emerges from Self-Play A general reinforcement learning algorithm that masters chess, shogi, and Go through self-play

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-09T05:08:32.647978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T05:08:32.647978Z digest=sha256:e29fef61a007e75e46dc1670f7e69c5a1b30f55471544a268527044cc1844bc0

Observation da09e55e-097e-4b12-8edf-5693f2252bfe · outbound

This paper cites Overview of motor vehicle traffic crashes in 2021.

Robust Autonomy Emerges from Self-Play Overview of motor vehicle traffic crashes in 2021

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:35.330293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:32.653022Z digest=sha256:67c68b5b10650861b75b9b81dbbaa2b01492212c672d089ae67b35029909fb7d

Observation 607780d5-162b-40ea-96c2-38a102ab9a93 · outbound

This paper cites TrafficSim : Learning to simulate realistic multi-agent behaviors.

Robust Autonomy Emerges from Self-Play TrafficSim : Learning to simulate realistic multi-agent behaviors

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:35.304392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:32.658574Z digest=sha256:77aed6dd22b6e8d85fd9fbe68b81ab2937c7a6b7ba3d4c229677d3a77d20927c

Observation 5a452fba-fc03-4af1-a85c-8ee22f5662f6 · outbound

This paper cites Repaint: Knowledge transfer in deep reinforcement learning.

Robust Autonomy Emerges from Self-Play Repaint: Knowledge transfer in deep reinforcement learning

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:35.275560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:32.663773Z digest=sha256:9d51c03747d1c3d5036a9d94e5bba32fbf1e8e00d8cf71710cc070a75f81af3f

Observation 7c49851a-5ab6-451a-bc12-c6992396d76b · outbound

This paper cites Congested traffic states in empirical observations and microscopic simulations.

Robust Autonomy Emerges from Self-Play Congested traffic states in empirical observations and microscopic simulations

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-09T05:08:32.669511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T05:08:32.669511Z digest=sha256:74efbab2e3bb48aa61268fa5b7a4f03b0281dfb190883240a071ddb1bee7a788

Observation 0ff84da0-9e2a-4710-8cc9-db24e3fa0e4c · outbound

This paper cites Nocturne: a scalable driving benchmark for bringing multi-agent learning one step closer to the real world.

Robust Autonomy Emerges from Self-Play Nocturne: a scalable driving benchmark for bringing multi-agent learning one step closer to the real world

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:35.163734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:32.674034Z digest=sha256:dff8e439bb4f0457e18a3b5c2f25c31b86cd73c97cb7580697c76b6aec57428e

Observation be7dfae8-de9a-4e6a-9126-dd2951eb45b9 · outbound

This paper cites u l c ehre, C ., Wang, Z., Pfaff, T., Wu, Y., Ring, R., Yogatama, D., W \.

Robust Autonomy Emerges from Self-Play u l c ehre, C ., Wang, Z., Pfaff, T., Wu, Y., Ring, R., Yogatama, D., W \

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-09T05:08:32.679149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T05:08:32.679149Z digest=sha256:e85dc04e255918f446e6ad8c146c9a9fe3649386839620d860ab6364b4d1a458

Observation 11c8efa5-31dd-4716-989c-986ecf884f0f · outbound

This paper cites and Zhen, H.

Robust Autonomy Emerges from Self-Play and Zhen, H

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:35.043471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:32.683801Z digest=sha256:607e65047569ea2f423b2804bf1775190642c4fa548537e782bd17f7a2473334

Observation ee420ee1-6631-4c52-b8a3-8d2e7ea42baf · outbound

This paper cites Self-play reinforcement learning guides protein engineering.

Robust Autonomy Emerges from Self-Play Self-play reinforcement learning guides protein engineering

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:35.022127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:32.715561Z digest=sha256:a19c842ba58f282e381a0806b0d2c5bbfdc0e56b42dc0beab3eb1bf542cce86b

Observation 01c17366-be68-4d96-8b93-5b0da930d68b · outbound

This paper cites Multiverse Transformer: 1st Place Solution for Waymo Open Sim Agents Challenge 2023.

Robust Autonomy Emerges from Self-Play Multiverse Transformer: 1st Place Solution for Waymo Open Sim Agents Challenge 2023

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-09T05:08:32.786796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T05:08:32.786796Z digest=sha256:621e850c2971449a03154a97b292bbb3c730c92e6a8809a1027d7b393e41fdc1

Observation cd49ece2-1ad6-4d39-9754-093a3bd43bee · outbound

This paper cites DD-PPO: learning near-perfect pointgoal navigators from 2.5 billion frames.

Robust Autonomy Emerges from Self-Play DD-PPO: learning near-perfect pointgoal navigators from 2.5 billion frames

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:35.001362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:32.836104Z digest=sha256:152cab33612b1881714ed7e3794f87f81e956bc7e687f61844d4afffbdcb987a

Observation d0dab807-a513-484c-90cd-39afce3b0573 · outbound

This paper cites Bits: Bi-level imitation for traffic simulation.

Robust Autonomy Emerges from Self-Play Bits: Bi-level imitation for traffic simulation

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:34.978009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:32.916636Z digest=sha256:48657b30315d3fdecca2b5fc1f59f59bbab3915e9f279745220549ee4a5777bf

Observation af82bec9-6b07-4187-a4d1-a298dac05127 · outbound

This paper cites Diffusion-ES: Gradient-free Planning with Diffusion for Autonomous Driving and Zero-Shot Instruction Following.

Robust Autonomy Emerges from Self-Play Diffusion-ES: Gradient-free Planning with Diffusion for Autonomous Driving and Zero-Shot Instruction Following

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-09T05:08:32.980823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T05:08:32.980823Z digest=sha256:61aa7a84bc436210cc24005b0c43a97fef72607bcf33f0ed5a1f8d44c792a98d

Observation 267c91bb-6aae-4230-affe-eec5fe8f101e · outbound

This paper cites J., and Urtasun, R.

Robust Autonomy Emerges from Self-Play J., and Urtasun, R

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:34.958563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:33.086473Z digest=sha256:6654c73deea67712b482c605b40a0555dcee716d97cedf04273a2da5fe94863d

Observation a1c4d79b-43f8-42f8-9d07-6d4cc806eb3a · outbound

This paper cites an unresolved cited work.

Robust Autonomy Emerges from Self-Play Unresolved cited work

Reference 77

Resolution
unresolved
raw_fallback, observed 2026-08-09T05:08:34.941341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:33.090791Z digest=sha256:ee2656ee928bc6a96cde327f3ace02f52acbaf52ed5b437a5354719f5c84913a

Observation 8f88b43a-8a67-4605-b079-c560bb1bc3b6 · outbound

This paper cites Learning realistic traffic agents in closed-loop.

Robust Autonomy Emerges from Self-Play Learning realistic traffic agents in closed-loop

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:34.924219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:33.094973Z digest=sha256:8751517d00df1c4fb507a11dc1239abf51f3ea4e1d1a9f0599691cfa9a754fc1

Observation 9c44ad32-e149-4f97-b62d-684951e89a01 · outbound

This paper cites Guided conditional diffusion for controllable traffic simulation.

Robust Autonomy Emerges from Self-Play Guided conditional diffusion for controllable traffic simulation

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T05:08:34.905940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-09T05:08:33.099836Z digest=sha256:3176f1e88d2778c1930bf01f5f5b7f2dc5ed3b487a2b9d39d604ca1a5021cda0

Pith citing papers

Observation 86e3f049-2f6d-4614-b10b-77efe135a7be · inbound

Knowledge Integration Strategies in Autonomous Vehicle Prediction and Planning: A Comprehensive Survey cites this paper.

Knowledge Integration Strategies in Autonomous Vehicle Prediction and Planning: A Comprehensive Survey Robust Autonomy Emerges from Self-Play

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T20:44:01.716477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T20:44:01.716477Z digest=sha256:c11c388d495c109b3144dd64c217a7896c0bbc16c1d8f7b52d7828deb1544256

Observation 7b5e9c16-b0ac-4ae5-b037-9702b09565d1 · inbound

Learning Through Retrospection: Improving Trajectory Prediction for Automated Driving with Error Feedback cites this paper.

Learning Through Retrospection: Improving Trajectory Prediction for Automated Driving with Error Feedback Robust Autonomy Emerges from Self-Play

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-16T12:06:02.123061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:06:02.123061Z digest=sha256:44bfbf8db7a6a4e2bb27b40eb3a200fa8d537212f6eef34abeb4d4054f8db405

Observation c266118a-6eb4-4c4e-bef6-301b6503384c · inbound

Not All Rollouts are Useful: Down-Sampling Rollouts in LLM Reinforcement Learning cites this paper.

Not All Rollouts are Useful: Down-Sampling Rollouts in LLM Reinforcement Learning Robust Autonomy Emerges from Self-Play

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-22T18:46:57.129172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-22T18:46:11.571831Z digest=sha256:889f942ea426af500730d42f4c51c3dda43a5108dba859831e9f1132a8562e54

Observation 6236b8c8-c3ae-4f4f-9491-e21250568153 · inbound

CaRL: Learning Scalable Planning Policies with Simple Rewards cites this paper.

CaRL: Learning Scalable Planning Policies with Simple Rewards Robust Autonomy Emerges from Self-Play

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-16T10:40:50.279677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:40:50.279677Z digest=sha256:210a4823d5ad4bdda572b670d6fc5dda089342ad07d09039bd2b692682b9537f

Observation 8156596f-6c89-4e73-88de-d91278db7eb1 · inbound

Multiple-Frequencies Population-Based Training cites this paper.

Multiple-Frequencies Population-Based Training Robust Autonomy Emerges from Self-Play

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T11:22:12.059284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:22:12.059284Z digest=sha256:b41506cfa6d3b3a106c23f4b2183ed4f257bd026bb17492820e775237bfe8190

Observation 10e5fbc3-e30a-4a61-9136-346a2b2fdfad · inbound

Reinforced Refinement with Self-Aware Expansion for End-to-End Autonomous Driving cites this paper.

Reinforced Refinement with Self-Aware Expansion for End-to-End Autonomous Driving Robust Autonomy Emerges from Self-Play

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T04:46:48.599598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:46:48.599598Z digest=sha256:8b2ff256e6482c7e894f610580528ef6251ed6e068cda55d11648e1b439d79da

Observation ff55e5b3-f230-4397-8a20-57ca98e3f08f · inbound

ReSim: Reliable World Simulation for Autonomous Driving cites this paper.

ReSim: Reliable World Simulation for Autonomous Driving Robust Autonomy Emerges from Self-Play

Reference 122

Resolution
verified exact
arxiv_id, observed 2026-05-19T09:32:16.639793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-19T09:28:00.597160Z digest=sha256:ddc7c687cfdf53cd5ed983fe4bc3c31ef9f596403854d72c6dd09df68210c181

Observation a77c0a0b-3c20-429b-a4f1-ba3465a6445f · inbound

Long-term Traffic Simulation with Interleaved Autoregressive Motion and Scenario Generation cites this paper.

Long-term Traffic Simulation with Interleaved Autoregressive Motion and Scenario Generation Robust Autonomy Emerges from Self-Play

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T19:15:25.299216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:15:25.299216Z digest=sha256:38626363f9025c6dea3e5be34965edfe94b4fc9284c76b6c4ba065dd14916437

Observation 465aad7e-8b2d-4b01-8aac-a190d7c6edc4 · inbound

Do LLM Modules Generalize? A Study on Motion Generation for Autonomous Driving cites this paper.

Do LLM Modules Generalize? A Study on Motion Generation for Autonomous Driving Robust Autonomy Emerges from Self-Play

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T11:30:39.719000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:30:39.719000Z digest=sha256:b60d14cba5d62bb89ab46327b9953ae6e1f17b60cc037f950db45de2e5a3e910

Observation dd036bb1-edd3-4b43-a406-918dbef118f8 · inbound

Do LLM Modules Generalize? A Study on Motion Generation for Autonomous Driving cites this paper.

Do LLM Modules Generalize? A Study on Motion Generation for Autonomous Driving Robust Autonomy Emerges from Self-Play

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-05T11:30:39.785533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:30:39.785533Z digest=sha256:d2c14ce26750148a5d606b06ca3c62b3b04bb163c19c69562465bc2610b4135e

Observation a83e19fe-3e0e-4029-9efd-23a47a025ff1 · inbound

Zero-Human Demonstration End-to-end Autonomous Driving with Trajectory Scorer cites this paper.

Zero-Human Demonstration End-to-end Autonomous Driving with Trajectory Scorer Robust Autonomy Emerges from Self-Play

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-04T07:54:30.759844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T07:54:30.759844Z digest=sha256:cbaf2cbece82173f09158568fa7c0b863ea2d47ed86fc288ed11cbe0dd6a1e82

Observation 0b292628-4c79-4062-89ff-98aba0300a7a · inbound

SimScale: Learning to Drive via Real-World Simulation at Scale cites this paper.

SimScale: Learning to Drive via Real-World Simulation at Scale Robust Autonomy Emerges from Self-Play

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-17T04:34:01.801151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-17T04:33:03.629533Z digest=sha256:adf292d12d7d6f69169ab55a32e9bc7b9d7bc8991abed7f5bc217ba233d70443

Observation 48f2fab1-111f-4781-8bc5-41e4693a69a1 · inbound

Toward Efficient and Robust Behavior Models for Multi-Agent Driving Simulation cites this paper.

Toward Efficient and Robust Behavior Models for Multi-Agent Driving Simulation Robust Autonomy Emerges from Self-Play

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-17T01:01:24.381990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-17T00:58:51.811689Z digest=sha256:b7d1d50141c3c3c06ccfce33f508fe576ab8bd7a505687d5cad79e09d6e5fb4f

Observation 8d07f791-8d78-44ee-ac19-7839504b674b · inbound

Pseudo-Expert Regularized Offline RL for End-to-End Autonomous Driving in Photorealistic Closed-Loop Environments cites this paper.

Pseudo-Expert Regularized Offline RL for End-to-End Autonomous Driving in Photorealistic Closed-Loop Environments Robust Autonomy Emerges from Self-Play

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:48:32.670260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-16T20:45:30.382235Z digest=sha256:ddf81c733ad3f13cd2497500acec742bb0dd527dcc04da51afd914fe72084da7

Observation 34b42deb-fcd3-4d6f-a907-b5cf663ef41b · inbound

Goal-Oriented Reactive Simulation for Closed-Loop Trajectory Prediction cites this paper.

Goal-Oriented Reactive Simulation for Closed-Loop Trajectory Prediction Robust Autonomy Emerges from Self-Play

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:58:26.236918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T00:55:01.446865Z digest=sha256:d588b042aa217b072d0a764e30b1756a7fcf133af30699406557aff869518dd7

Observation 36c708b8-8da3-4d96-a343-ca6193000eda · inbound

Fail2Drive: Benchmarking Closed-Loop Driving Generalization cites this paper.

Fail2Drive: Benchmarking Closed-Loop Driving Generalization Robust Autonomy Emerges from Self-Play

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:06:00.614667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T17:19:01.704384Z digest=sha256:17d6d51618210ff21fb83d9139c433608a618fcd08cca40893ce46b45b236a28

Observation 00dd2f59-8b17-49e0-bc5c-22d13bee7b60 · inbound

Artificial Intelligence for Modeling and Simulation of Mixed Automated and Human Traffic cites this paper.

Artificial Intelligence for Modeling and Simulation of Mixed Automated and Human Traffic Robust Autonomy Emerges from Self-Play

Reference 110

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T09:26:02.355386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T16:00:59.662003Z digest=sha256:4bc578a4c4566f14bb27e8d0ad46777a815bea327755b54ea6ed6aac6f849ea3

Observation a82d4539-80ce-4b82-9ab5-8fd6a6737eea · inbound

A High-Throughput Compute-Efficient POMDP Hide-And-Seek-Engine (HASE) for Multi-Agent Operations cites this paper.

A High-Throughput Compute-Efficient POMDP Hide-And-Seek-Engine (HASE) for Multi-Agent Operations Robust Autonomy Emerges from Self-Play

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:36:26.174818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-07T10:27:55.337253Z digest=sha256:8daececde23eacf50c768a75a8558ba46f45ace94f3487c8a9ae6a84a388f419

Observation 8b6896ce-6007-417b-b44f-ea1ff0f21a0a · inbound

Beyond Self-Play: Hierarchical Reasoning for Continuous Motion in Closed-Loop Traffic Simulation cites this paper.

Beyond Self-Play: Hierarchical Reasoning for Continuous Motion in Closed-Loop Traffic Simulation Robust Autonomy Emerges from Self-Play

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-12T01:56:14.768801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T01:55:14.466724Z digest=sha256:be69b7a9aa8ce20600b269b178e7f264600f9f09304a3e4a9830d0bb1166fae2

Observation 4c5e5b1a-1677-40ad-a51f-428cb3753b4b · inbound

TerraTransfer: Learning End-to-End Driving Policies Without Expert Demonstrations cites this paper.

TerraTransfer: Learning End-to-End Driving Policies Without Expert Demonstrations Robust Autonomy Emerges from Self-Play

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-07-03T18:38:49.981009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-27T02:27:51.946679Z digest=sha256:47818eb51bb053119ca66480ea8d0847a9eabf26ddd00d899c5a8a11021af467

Observation c9834175-3cba-4559-963f-fd1af52c1532 · inbound

TerraTransfer: Learning End-to-End Driving Policies Without Expert Demonstrations cites this paper.

TerraTransfer: Learning End-to-End Driving Policies Without Expert Demonstrations Robust Autonomy Emerges from Self-Play

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T11:08:36.440196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:08:36.440196Z digest=sha256:3d0303db729e998d7dec8c3068ffce5f56453fd14768118deffa92b03d7848d4

Observation 6860ca74-a2ef-4fba-a27f-a32fd183fdf0 · inbound

Human-like autonomy emerges from self-play and a pinch of human data cites this paper.

Human-like autonomy emerges from self-play and a pinch of human data Robust Autonomy Emerges from Self-Play

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-03T14:28:31.997601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-27T07:01:12.737217Z digest=sha256:595cd1d5d23eb3bbca6290a166f9863b300e94a9ee32ba88ec7e3b0853c04ca6

Observation e1586130-8cf0-4e99-baa7-5925c3bd030b · inbound

Reward-Conditioned Attention: How Reward Design Shapes What Autonomous Driving Agents See cites this paper.

Reward-Conditioned Attention: How Reward Design Shapes What Autonomous Driving Agents See Robust Autonomy Emerges from Self-Play

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-04T17:09:58.515865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-25T23:57:34.962907Z digest=sha256:872a27283007692f8c14dc8a6333b3e54f6a3f26d62177851ae74ae9c638aec6

Observation af1a91b1-d965-4fd1-b8cc-c41be498a61f · inbound

EvoFlock: evolved inverse design of multi-agent motion cites this paper.

EvoFlock: evolved inverse design of multi-agent motion Robust Autonomy Emerges from Self-Play

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-06-25T19:58:18.375526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-25T19:54:26.458543Z digest=sha256:c8d6c12211e61cc0b19a7cf529010aad454818046b208e561764d9a843517ee4

Observation dccaaadd-7715-4fd8-a161-d1a2e932297b · inbound

Is Inter-Seed Cross-Play Enough? Evaluating the Robustness of Zero-Shot Coordination Algorithms to Implementation Details cites this paper.

Is Inter-Seed Cross-Play Enough? Evaluating the Robustness of Zero-Shot Coordination Algorithms to Implementation Details Robust Autonomy Emerges from Self-Play

Reference 188

Resolution
unresolved
no resolver link, observed 2026-08-05T15:25:40.157105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:25:40.157105Z digest=sha256:ebd5f50807e6524c70359d0a2b69fd628b5c61ac69fae8824cd63a82a821bc28