Pith. sign in

Paper Citation Record · LEDGER

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models

As of 15 August 2026, this Paper Citation Record lists 87 of 87 outbound references and 0 inbound Pith citation observations for arXiv:2608.06729.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.06729 v1

Coverage vector

measured 87 of 87 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T14:33:40.117731Z

measured 87 of 87 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

87 of 87 outbound references displayed

  • verified exact0
  • verified fuzzy7
  • unresolved80
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ae7e2e93-53ed-489e-8bf6-92381f8214a2 · outbound

This paper cites PaliGemma: A versatile 3B VLM for transfer.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models PaliGemma: A versatile 3B VLM for transfer

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.148006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.148006Z digest=sha256:ce144e5987363b0e9d0f7d8f8cfe79701cd8fa7c84c2b0e588b2acb7d1d0ce7c

Observation 1763a70b-8f94-4838-aea3-9f4acd2bf58d · outbound

This paper cites Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.153666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.153666Z digest=sha256:be751bc6b95626ce4a1c211b5b7d73be25eb875527004544bf624c9b19463f24

Observation d0729d75-683b-4540-9d75-9442a5284b0f · outbound

This paper cites IEEE Robotics and Automation Letters , year=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models IEEE Robotics and Automation Letters , year=

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.158941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.158941Z digest=sha256:eee86c53a4fcab39a82ea810fb972405baf872ce70502c92d17b5da17a7c19e3

Observation 7c7b2dd4-08cb-4a5a-bfa4-2e82a53d2a33 · outbound

This paper cites What Matters in Building Vision-Language-Action Models for Generalist Robots.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models What Matters in Building Vision-Language-Action Models for Generalist Robots

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.183731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.183731Z digest=sha256:56e0d9f18ead7f85e56b53d19ee13bffbc725222fac7fd6e6cc1d99c5379357b

Observation ea6c5f2f-8764-428a-a367-1c6c15cd2cb9 · outbound

This paper cites HybridVLA: Collaborative Diffusion and Autoregression in a Unified Vision-Language-Action Model.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models HybridVLA: Collaborative Diffusion and Autoregression in a Unified Vision-Language-Action Model

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.268148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.268148Z digest=sha256:84027e12557fecd73d60457e8c88cd10a34ef70df818117ea56ff5d5a921f1c8

Observation 0cfc02c2-dbc6-46fb-9324-08180790e94d · outbound

This paper cites an unresolved cited work.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.283509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.283509Z digest=sha256:9ce5a6bc3c8e33ad0087699dfecbe3ce1ad23cfbad106ff121b2a7f906a1d793

Observation efaa64dd-ad82-4936-bd72-302bd4c7eb17 · outbound

This paper cites RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.289124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.289124Z digest=sha256:3ef5b9f5ac9b8aaaa74ae5035b79d40fe611979e03407dd9ef7be557ee9633ca

Observation ca3c8834-8946-4a8f-87ad-cadb67b3be5c · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models OpenVLA: An Open-Source Vision-Language-Action Model

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.293367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.293367Z digest=sha256:af247dfa5d1086f4cea657a4a2c04a20b0dcf87d90c40605396db99c5df6455e

Observation a159239b-43ff-4994-88b1-4d2cf859d8aa · outbound

This paper cites Conference on Robot Learning , pages=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Conference on Robot Learning , pages=

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.297762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.297762Z digest=sha256:763af5cd0413951f9b812523b2572e929aa6b37ce6f2815b6df206a4353db63f

Observation 91dea7aa-eeec-4f30-98dc-51c2f9f7398f · outbound

This paper cites RT-1: Robotics Transformer for Real-World Control at Scale.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models RT-1: Robotics Transformer for Real-World Control at Scale

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.381582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.381582Z digest=sha256:b2bd808b9da95039e98e221cc82d9e95b03a5b7e6eaacced31b395a0d52a77e9

Observation f9e79d84-e7ae-474c-a7ec-c7e8ea9ee1c9 · outbound

This paper cites RT-H: Action Hierarchies Using Language.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models RT-H: Action Hierarchies Using Language

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.488119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.488119Z digest=sha256:9343d4a7202dcc25160ab349a6d07c374085589c8921090833eed8c2dec27f4b

Observation 353f80cf-62d5-4ade-8b96-1a9c810582fa · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models LLaMA: Open and Efficient Foundation Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.492031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.492031Z digest=sha256:1bcf1956dc9bbe724ea6217b27f756e0e15a831e9f476112505b8e050e206d36

Observation 8cc9df09-01a3-4ae3-ac50-3f0d41b3d231 · outbound

This paper cites an unresolved cited work.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.497050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.497050Z digest=sha256:51dfa832267dc852afcf74e43df8d040794d76e5a939f38a2965f9c8474b1934

Observation 8a0c119c-abc1-4766-bacc-09239d225052 · outbound

This paper cites Advances in neural information processing systems , volume=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Advances in neural information processing systems , volume=

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.501329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.501329Z digest=sha256:c452b680ae0ba14e47acbc36b55f7c59d14667badd3b08b148e0dfe034d87efe

Observation 4d22abc2-52d9-4713-8225-26bfc05b6e51 · outbound

This paper cites 2023 , journal =.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models 2023 , journal =

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.505225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.505225Z digest=sha256:d1589150f0380a08b9da2733858daf8756ca67a268bdcbd9f4208b2c0b5dd000

Observation 99f222f3-f23f-475d-a98b-c0c742f15921 · outbound

This paper cites Advances in neural information processing systems , volume=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Advances in neural information processing systems , volume=

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.619247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.619247Z digest=sha256:22cc9fe78d16c19adc80e79d99146f89515df98c36edeaa9c6687f080e3494ca

Observation 8408abc6-ef5b-44b8-a463-a4cd5cf473bc · outbound

This paper cites an unresolved cited work.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-15T14:33:41.423444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T14:33:38.683398Z digest=sha256:5bc064ea008f5a306978d97d275269d754e80f3fbfbe22d2f21e4138b719c3a6

Observation 418df141-abb1-4f02-91a3-f3b5585fedc6 · outbound

This paper cites Computer Vision -.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Computer Vision -

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:33:41.413259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T14:33:38.686536Z digest=sha256:642aaf4ebda82842842020b166ae0efbf36a296b8f102d395bc29229eaddc929

Observation 025c6806-f067-4b91-8594-41212a13192a · outbound

This paper cites SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.691558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.691558Z digest=sha256:d4b07e2238ecf1f81a2d20119b45a5bfa6da2ffbba6b013e57c6c60b2f41b90c

Observation 684f71ff-c72e-4af1-9211-ad0792094b98 · outbound

This paper cites GeoVLA: Empowering 3D Representations in Vision-Language-Action Models.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models GeoVLA: Empowering 3D Representations in Vision-Language-Action Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.696655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.696655Z digest=sha256:b7aa9fa65976e55798ebea39c0f643895c99e4d3d7b442f6c489d477739feaad

Observation 2c0dd2ac-bf53-4dcb-afb7-27fb5a82e3bc · outbound

This paper cites IEEE Robotics and Automation Letters , volume=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models IEEE Robotics and Automation Letters , volume=

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.700126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.700126Z digest=sha256:ca570110a49f8ba24b0e6f02a8039c96fe25319cfe03236429c3dd31e9f2684d

Observation a2e0bae6-d0d5-4e6e-8386-f08a305fa272 · outbound

This paper cites arXiv preprint arXiv:2506.07961 , year=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models arXiv preprint arXiv:2506.07961 , year=

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.794118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.794118Z digest=sha256:42a2ec4a8fb73f8da4ae5ba76eb06f1ca9f811b87cafb2e5ba887b70d4bd4c46

Observation 3bed30ce-dd15-4cf7-9159-d8450e8655fc · outbound

This paper cites arXiv preprint arXiv:2506.22242 , year=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models arXiv preprint arXiv:2506.22242 , year=

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.870725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.870725Z digest=sha256:46482c181644c85acc077308f8a4421676cecaf0d5059be9cdc55aa033714f24

Observation ffa41d8a-b812-453b-b475-eec3757968b8 · outbound

This paper cites arXiv preprint arXiv:2510.17439 , year=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models arXiv preprint arXiv:2510.17439 , year=

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.875114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.875114Z digest=sha256:41d1648c60c409295c506ca16b655d7541d2886663658011b05866fd97fe6564

Observation e9beff1d-e16c-4b87-acc8-b52f80ebefbe · outbound

This paper cites 3D-VLA: A 3D Vision-Language-Action Generative World Model.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models 3D-VLA: A 3D Vision-Language-Action Generative World Model

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.878597Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.878597Z digest=sha256:7d850b463e436d493c3440179087431e15938a28197a4e93ac5833f64a07309a

Observation 8a05c435-c8a2-49b0-93f9-5a3ec9b93b0a · outbound

This paper cites arXiv preprint arXiv:2507.00416 , year=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models arXiv preprint arXiv:2507.00416 , year=

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.882575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.882575Z digest=sha256:9a5645ee4cfcbd34d23f3b8abfe46d31c8eb6b69868c860f11496fdbf58f8d61

Observation 801b711f-8a16-467e-b694-942cdcf552df · outbound

This paper cites AlphaMath Almost Zero: Process Supervision without Process.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models AlphaMath Almost Zero: Process Supervision without Process

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.971255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.971255Z digest=sha256:0b776a36954d83b4bf45fb492f0e6f2887b500644f95f24d0322480243756a51

Observation 974d7a01-3d29-4984-aff4-c142f277c92a · outbound

This paper cites Findings of the Association for Computational Linguistics: ACL 2025 , pages=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Findings of the Association for Computational Linguistics: ACL 2025 , pages=

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.975423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.975423Z digest=sha256:f2c06af48a063e6998bbc214df2ea93f2b6b1953373bba15b1bc17bae338bcdf

Observation 7f19ee55-8d2b-410b-bf90-83d05d1f441c · outbound

This paper cites Self-Consistency Improves Chain of Thought Reasoning in Language Models.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Self-Consistency Improves Chain of Thought Reasoning in Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:38.979347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:38.979347Z digest=sha256:f73bd6816f1acf13ef24cf5075b7efb4aac28c8cf5d90190ccd8007f6536e6e5

Observation 6e95a505-16e4-4cda-bfd2-331d85b41f05 · outbound

This paper cites Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing , pages=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing , pages=

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:33:41.389106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T14:33:38.983031Z digest=sha256:f60faf5dbb554dce48b5191562eab4905f8a3b45f5a408a1ca83ca9f7a8c7efa

Observation 921a7363-0122-470c-b298-c0b43d6364cd · outbound

This paper cites Large Language Monkeys: Scaling Inference Compute with Repeated Sampling.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Large Language Monkeys: Scaling Inference Compute with Repeated Sampling

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.033472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.033472Z digest=sha256:3730d4c8d9ce00c3c2b57b3dd943f1c03868d151827d482f22a6ad7481a9d574

Observation 1bcaec27-b361-476f-b672-df4542556946 · outbound

This paper cites Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.091894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.091894Z digest=sha256:6c6ee542b27d763c45403d0457d62d7b7be8921ab4ff94d46fa36c82e1870a9d

Observation dd585a45-c7b8-4582-8fed-cfd5017f5521 · outbound

This paper cites RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.096472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.096472Z digest=sha256:8049d5ca3dc0e7c83ee83a8b250c21a763c58e1faed976bb81341b8c523de213

Observation 6fcb863b-8d34-49f3-bdb6-5348083076a1 · outbound

This paper cites Verifier-free Test-Time Sampling for Vision-Language-Action Models.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Verifier-free Test-Time Sampling for Vision-Language-Action Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.100502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.100502Z digest=sha256:81152099fcde7181a147a37ff5482b7dfc31db1a96be5774e1b9fcb0479f1f08

Observation 28bfe46d-e9f3-482b-9162-6da73869dcc0 · outbound

This paper cites arXiv preprint arXiv:2510.10975 , year=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models arXiv preprint arXiv:2510.10975 , year=

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.104348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.104348Z digest=sha256:a75601785113ba18eb2ed640318a6cc6e5f88f3816ece5521dbe6dabd5888195

Observation 24dccad6-a362-4f62-9ccb-e2b96c7d4114 · outbound

This paper cites arXiv preprint arXiv:2601.00675 , year=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models arXiv preprint arXiv:2601.00675 , year=

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.160032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.160032Z digest=sha256:dc2a4664f9e5590f9151981c70e959b3f7f3449d68e39d376a270ef198728587

Observation b008e761-fc45-4e5e-b572-177f675dd105 · outbound

This paper cites DROID: A Large-Scale In-The-Wild Robot Manipulation Dataset.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models DROID: A Large-Scale In-The-Wild Robot Manipulation Dataset

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.203775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.203775Z digest=sha256:4a6746648aebe7d7cee0e973fdd1c1a5c07851afca4cc550506240a4de4436b2

Observation 0111b435-3b3b-44a4-88d2-9353815a56a2 · outbound

This paper cites Conference on Robot Learning , pages=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Conference on Robot Learning , pages=

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.209322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.209322Z digest=sha256:35cece832f21f6d13b112645d31bfdf960a34998dfa8875d831c51fc59192a34

Observation 87c4d7ed-efee-4e1d-9445-c5caa928a19d · outbound

This paper cites RoboMIND: Benchmark on Multi-embodiment Intelligence Normative Data for Robot Manipulation.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models RoboMIND: Benchmark on Multi-embodiment Intelligence Normative Data for Robot Manipulation

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.212850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.212850Z digest=sha256:e207b52019f5527c5522f400d3036b91feacb80e99e2bab82ecd8da35bb75eb9

Observation 80dd6fb4-2975-4007-b080-c3fba28930fb · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Advances in Neural Information Processing Systems , volume=

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.262864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.262864Z digest=sha256:c115b63ac0bbc647d7d5e3e948d056defe4b757bbaece01ed5c3922936f25f5e

Observation ec6d4e48-53ec-4841-b49e-765388534aab · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.287949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.287949Z digest=sha256:cdd793a6fb583f2597abea028ee193595449163f8c312efe68b97198cc5f8e40

Observation bc59b65e-895c-4e78-b10c-449d389761e1 · outbound

This paper cites $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.292896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.292896Z digest=sha256:affd518b066d112fa03b160b165d1b3ccc84cfa0a79263185ccae73c5c00e97b

Observation ffea7ef0-1caf-4aaa-a27e-7decbd33ab69 · outbound

This paper cites Proceedings of the Computer Vision and Pattern Recognition Conference , pages=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Proceedings of the Computer Vision and Pattern Recognition Conference , pages=

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.297696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.297696Z digest=sha256:c5c26ef68cb8980c0d2ad75e08d99ff37ad35e1672216e41ccf97d4dbff18365

Observation 23d72755-8f67-4d36-bd1b-4d2bfaa0ebd8 · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.301691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.301691Z digest=sha256:202884315dee839d83a62407cd3f07e394e72636b750c0e5c28153bd3cee518b

Observation 01f7e032-03b0-4ca9-9032-83978cc94e92 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Advances in Neural Information Processing Systems , volume=

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.353990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.353990Z digest=sha256:292028e4a96d39f4f6c840d15f4a9eb24bc5dbb30c23c5ddbad78d068d951c7a

Observation 6cb7f7ad-b15e-40c1-8bf0-485001dbee02 · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.358576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.358576Z digest=sha256:f2c9b5680aee512b229190ff7f61fd4b2b8aaa1ccdc80fbc73b49da69835f30b

Observation b4ffd6d4-740b-4a24-8bbf-a2c81d120ec7 · outbound

This paper cites ZoeDepth: Zero-shot Transfer by Combining Relative and Metric Depth.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models ZoeDepth: Zero-shot Transfer by Combining Relative and Metric Depth

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.362769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.362769Z digest=sha256:582ebf0d22be1cc5889139638367b03ce50761455c8dde88322c502060a75d53

Observation 62f9fa2f-2332-47c7-b234-eb9ccb9992b9 · outbound

This paper cites SafeVLA: Towards Safety Alignment of Vision-Language-Action Model via Constrained Learning.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models SafeVLA: Towards Safety Alignment of Vision-Language-Action Model via Constrained Learning

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.367292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.367292Z digest=sha256:8a51157159b72cf103b4510aeca8000ac8483d67b862a799bd75972ec4cf1d39

Observation 8c311ac6-cc58-4af1-b442-d92bed8e7ea4 · outbound

This paper cites Evaluating Real-World Robot Manipulation Policies in Simulation.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Evaluating Real-World Robot Manipulation Policies in Simulation

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.371997Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.371997Z digest=sha256:a5b9808f34b5f644a8a66d6a20fe72351b06adf0105fce68894774ac3e4ed8f6

Observation 5b344481-fde6-48e3-b820-939cd989bafd · outbound

This paper cites DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.375454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.375454Z digest=sha256:3d4fdd2562a8b6ee96276e88ee090de567fba1f2cb58e83c31773d692719bb8d

Observation 686c96fe-47bf-4e27-b40e-8108637e46a7 · outbound

This paper cites International conference on machine learning , pages=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models International conference on machine learning , pages=

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.379418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.379418Z digest=sha256:01af5641cdf54bfbdb370afaae653395a24d2628b8c3418090c32d9773affed3

Observation b6ddcaa8-62ba-48f8-aa9e-2e228d718057 · outbound

This paper cites Proceedings of the IEEE/CVF international conference on computer vision , pages=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Proceedings of the IEEE/CVF international conference on computer vision , pages=

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.392414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.392414Z digest=sha256:7cc84c6bdf443cff8b8aa344b014b1fd0c361e8705cc8f05d7822ae38e5f45b5

Observation 60ca2448-e9bd-4003-ae0f-6ce16ba6df92 · outbound

This paper cites , author=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models , author=

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.473885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.473885Z digest=sha256:5b7fc07d2a661802e633ec9ca201670bc15b1b0ecbe8e11026e0e66820d699b8

Observation f1c1451d-d63e-46f4-b1f5-b44bd5a73354 · outbound

This paper cites Octo: An Open-Source Generalist Robot Policy.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Octo: An Open-Source Generalist Robot Policy

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.537822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.537822Z digest=sha256:69403a96ad01c831b508598f68387daa8afae61491682a0a4d410784567e03fc

Observation 0a34adf9-9825-40be-a7b7-a21df0bfa7c2 · outbound

This paper cites MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.561942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.561942Z digest=sha256:76ae5bcdafb7defe044f61603c979a1feb85f7e78942e0e214f3adfcdb9d7012

Observation 4da556c2-42cf-425a-b1a7-af0af792fa5f · outbound

This paper cites 2024 IEEE International Conference on Robotics and Automation (ICRA) , pages=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models 2024 IEEE International Conference on Robotics and Automation (ICRA) , pages=

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.566446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.566446Z digest=sha256:0058b19ac6b904d53ab96334eb89c6cc27cf93ed4b7e666681169789c057dbab

Observation f762a73a-42df-4efb-b706-eb386f54e39d · outbound

This paper cites Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:33:41.311095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T14:33:39.570305Z digest=sha256:0c53b1efa5392e67a87b93f77b4e120a6e9fbeb1e188046c4c599ed7f8f566e6

Observation abbfdb0f-96c0-4d4e-9908-ca777fba6274 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.573805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.573805Z digest=sha256:3c9742724cff4bea01e271d634dc77a3f912e863d2b25d3b8284e08a592f3a96

Observation 3de06352-e39b-44b0-92c2-f68f8bc7ddab · outbound

This paper cites Advances in neural information processing systems , volume=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Advances in neural information processing systems , volume=

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.578402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.578402Z digest=sha256:e6fb8280e88a98cf7537e5fd94c6bce91fb9d91ec2ce09b7b577c6cb9d70edcc

Observation b334f3a3-4c84-4115-bc12-0c502d664b1e · outbound

This paper cites European conference on computer vision , pages=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models European conference on computer vision , pages=

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.634116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.634116Z digest=sha256:5832439640a365ab8454dcca413932f1324569131efd3b1916db7308bf5c7d60

Observation 9f3f74d0-5ab8-4336-87b5-c1b280691917 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Advances in Neural Information Processing Systems , volume=

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.728583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.728583Z digest=sha256:7ac6636713f940e82aaf6b8d21de22b6cecf3e0e9f60e0b1b48e989c1daa828e

Observation bcad87bf-fb90-4a15-a5fc-3386623c3a3e · outbound

This paper cites 9th Annual Conference on Robot Learning , year=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models 9th Annual Conference on Robot Learning , year=

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:33:41.280516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T14:33:39.732138Z digest=sha256:465d1c43554c5e8552ee2b1fd596018ce15f3455abee34c9be9f589a1a0c156c

Observation 9e5e6f1f-6f55-43a1-aec8-30ede2ab26a8 · outbound

This paper cites Depth Anything 3: Recovering the Visual Space from Any Views.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Depth Anything 3: Recovering the Visual Space from Any Views

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.735865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.735865Z digest=sha256:3f9f8a033536cd66e59eddfc90c8d66e26093d3e9cb235e2c94fcc1fb70e08b1

Observation 0c33db0d-3f55-4324-83ce-01cc86ad7d2f · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.739135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.739135Z digest=sha256:932b5afa9574b18043e07d42d8cf8a2e648a5ac0c96116ace4ddd73b4ee6fd0d

Observation 10296197-c141-4541-a43b-bd0dfb0f63f6 · outbound

This paper cites IEEE Robotics and Automation Letters , volume=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models IEEE Robotics and Automation Letters , volume=

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:33:41.270224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T14:33:39.744353Z digest=sha256:d67487237925d0e1141d1fc5f81d66ed24acf60913cf1a1e9bf2150a630a11ab

Observation b3c62d0e-f7be-4c67-9049-a55541de67a0 · outbound

This paper cites arXiv preprint arXiv:2603.12942 , year=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models arXiv preprint arXiv:2603.12942 , year=

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.748139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.748139Z digest=sha256:2e5951b1798d720307e3a4279c55609573d5441dc6b9f0f36834b8f187101d80

Observation 217fca06-6500-4992-a60c-3739e205c05c · outbound

This paper cites arXiv preprint arXiv:2511.09516 , year=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models arXiv preprint arXiv:2511.09516 , year=

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.751081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.751081Z digest=sha256:141ee38f9c112d29551c9867d1d813be5454c5715a342b3ab86fae7f4f96c2be

Observation e3608715-e71f-4952-9ffe-bc43eae14b7b · outbound

This paper cites Spatial Memory for Out-of-Vision Manipulation in Vision-Language-Action.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Spatial Memory for Out-of-Vision Manipulation in Vision-Language-Action

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.754892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.754892Z digest=sha256:362ed3d5dd7b6bc62783ac7c8b995a68a7fff5f97ba09403ae744b8f7be2becf

Observation e6a6648e-c7a3-4fc3-ba67-1a971de1217d · outbound

This paper cites AgiBot World Colosseo: A Large-scale Manipulation Platform for Scalable and Intelligent Embodied Systems.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models AgiBot World Colosseo: A Large-scale Manipulation Platform for Scalable and Intelligent Embodied Systems

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.871201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.871201Z digest=sha256:975d34d603b0e3faf4604e457b400c9add497ffaac0275c06170d2cfb341e24b

Observation da6697d6-8d2a-4f8a-acd7-b65ac646bc6c · outbound

This paper cites CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.959379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.959379Z digest=sha256:423b228c15e148f5d90a9b6d2fb4f68e0e76248b8f55ba9a736aca959c4a27ee

Observation be4a9eeb-d7cb-430f-b815-adb4031cd6e6 · outbound

This paper cites DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.964543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.964543Z digest=sha256:0c55fce4f9ad1e88494766e6072f460d02a5913bbf94d1a302e2b06b267c8462

Observation b8fa5fcd-f5e0-416d-8c1c-162408133c7e · outbound

This paper cites The International Journal of Robotics Research , volume=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models The International Journal of Robotics Research , volume=

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.968623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.968623Z digest=sha256:f489d8a946a9f87a5529e3a503628f4e3e2d492f54517efe540fe9ce42ca2d1d

Observation 50c141b7-f62e-4f38-8685-38709cd06859 · outbound

This paper cites International Conference on Learning Representations , volume=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models International Conference on Learning Representations , volume=

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.972759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.972759Z digest=sha256:cd1159a52f2f6632213537c50a6cefe0b8dbf967ec0625f6585e452f4c8b257a

Observation 70e15055-bb60-4e35-8792-896d4f1fd471 · outbound

This paper cites arXiv preprint arXiv:2601.17885 , year=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models arXiv preprint arXiv:2601.17885 , year=

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.976530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.976530Z digest=sha256:fb74a70fb2e8fd48b4c4f9943973de1a8c54a3030cd291057639b3d181a1d447

Observation cd0a5095-a81e-4042-b2cc-7c11d0bf3421 · outbound

This paper cites arXiv preprint arXiv:2603.03596 , year=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models arXiv preprint arXiv:2603.03596 , year=

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.980491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.980491Z digest=sha256:cb1dee039da81924122e21b8f920509ff8f399dbeb2de2f4d6939184a728d067

Observation 7356e200-f8de-419a-9f72-b57d7124bfe1 · outbound

This paper cites International Conference on Machine Learning , pages=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models International Conference on Machine Learning , pages=

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:33:41.245491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T14:33:39.984092Z digest=sha256:681dae77520792965b1a515b57148c37cf1421d81a95a1201ab86113634e5af5

Observation 1a06449e-2bd5-4e54-b807-016c520003fb · outbound

This paper cites 2011 10th IEEE international symposium on mixed and augmented reality , pages=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models 2011 10th IEEE international symposium on mixed and augmented reality , pages=

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.998452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.998452Z digest=sha256:803f7860df02653a3e7f0242f483020a163c7bc6353a16767b6f2854c89e637f

Observation 282bb211-58b9-49b5-a321-fc7623c404df · outbound

This paper cites Genie Envisioner: A Unified World Foundation Platform for Robotic Manipulation.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Genie Envisioner: A Unified World Foundation Platform for Robotic Manipulation

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:40.052053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:40.052053Z digest=sha256:fbd595c26c708eb3126c4c61c9bdc539772949195e4b203cf9b174cffd2dd111

Observation cbcab0b7-be21-459d-a1dd-d7d2f6d4e54c · outbound

This paper cites RynnVLA-002: A Unified Vision-Language-Action and World Model.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models RynnVLA-002: A Unified Vision-Language-Action and World Model

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:40.086994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:40.086994Z digest=sha256:a4183eab01a9254c9fe67f9cc1fb654abc880d9e3339ad353b325c045db613be

Observation ba421f0d-dcf6-4f75-bc15-eb4da3008fa7 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Advances in Neural Information Processing Systems , volume=

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:33:41.226904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T14:33:40.091060Z digest=sha256:813bb70eea0f50b883a2bde0a4c700d974545b94976e3ec8049faca1b766a357

Observation d3b3487e-4528-4459-846e-406b4becffbb · outbound

This paper cites International Conference on Learning Representations , year=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models International Conference on Learning Representations , year=

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:40.094674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:40.094674Z digest=sha256:d6c2ca63f5d8e92aba951a20615b410369a81f9e6b0c7070a72a66ff36b4f278

Observation c44050e2-1ab4-4ddd-b3f5-5cad689ec8ce · outbound

This paper cites Classifier-Free Diffusion Guidance.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Classifier-Free Diffusion Guidance

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:40.098346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:40.098346Z digest=sha256:f439b43368c53255ee1ad834b258f7a5df949563214f7e3410754dc5e90d15b1

Observation c4c54e4c-464a-488a-93a5-16ccbfcb7628 · outbound

This paper cites IEEE Robotics and Automation Letters , volume=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models IEEE Robotics and Automation Letters , volume=

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:40.102244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:40.102244Z digest=sha256:12a7ef0d5e651f24460f75356f5a6409f307095e0847ae8303e3814c9d9f57ec

Observation 623d2989-3534-45f8-b3ce-72f24b64da33 · outbound

This paper cites Mask World Model: Predicting What Matters for Robust Robot Policy Learning.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Mask World Model: Predicting What Matters for Robust Robot Policy Learning

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:40.105970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:40.105970Z digest=sha256:0138eb6984c4fca62d91a6c966111ebbd29630baf74d4e9a4f252272c1d1864e

Observation c8fe6e67-3b19-4e38-9dad-4a20fba9c767 · outbound

This paper cites Transactions on Machine Learning Research Journal , year=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Transactions on Machine Learning Research Journal , year=

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:40.109759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:40.109759Z digest=sha256:291d9bd93efba4d7bc60112bdf86348853ce63836ff7c7d649ab1f0618333520

Observation 6c3c7702-35ce-4323-8246-6f111d3673ff · outbound

This paper cites Proceedings of the IEEE/CVF international conference on computer vision , pages=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Proceedings of the IEEE/CVF international conference on computer vision , pages=

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:40.113681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:40.113681Z digest=sha256:f1bc2a4ab5d7fbbbb625a1f1fbe733fba160b0f2363848fa5342e468081f1e59

Observation dd029b01-e1d7-4015-8164-7d463853202e · outbound

This paper cites Proceedings of the IEEE/CVF international conference on computer vision , pages=.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models Proceedings of the IEEE/CVF international conference on computer vision , pages=

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:40.117731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:40.117731Z digest=sha256:c59ef665586abd18c2480e5fbacb80164b813dfbcfbf4ce4aed1011323614bb3

Pith citing papers

No inbound Pith citation observations are available.