Pith. sign in

Paper Citation Record · LEDGER

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

As of 5 August 2026, this Paper Citation Record lists 100 of 114 outbound references and 86 inbound Pith citation observations for arXiv:2604.15483.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2604.15483 v2

Coverage vector

measured 100 of 114 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-10T11:42:34.409651Z

measured 186 of 186 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 86 of 86 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T14:32:56.296925Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T19:47:32.631418Z

Reference resolution

100 of 114 outbound references displayed

  • verified exact52
  • verified fuzzy41
  • unresolved1
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch6

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5499286a-509c-4982-adc0-53ad82ea09ac · outbound

This paper cites RT-1: Robotics Transformer for Real-World Control at Scale.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities RT-1: Robotics Transformer for Real-World Control at Scale

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:41:14.065009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:5fb2230c4d2ffd514134112eba6c7318b83d7802824b867454715e7e9defb4bb

Observation c6ae61e6-946d-4f38-9d44-596b9a264696 · outbound

This paper cites A generalist agent.Transactions on Machine Learning Research (TMLR).

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities A generalist agent.Transactions on Machine Learning Research (TMLR)

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.932739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:89b708334d5dcd47a7bdbc9932521684cdffadff2f8388df7990dd93162690e6

Observation fcfcb28f-966e-4e09-b895-54868e59762c · outbound

This paper cites Octo: An Open-Source Generalist Robot Policy.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Octo: An Open-Source Generalist Robot Policy

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:26:16.174119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:cd2115e3cb0fde1f438055d467886f22ac6e63b0c1f24d8c07032366d41dd374

Observation e38a98d7-e623-47a8-aae8-d1a2593f8008 · outbound

This paper cites Rdt-1b: a diffusion foundation model for bimanual manipulation.International Conference on Learning Representations (ICLR).

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Rdt-1b: a diffusion foundation model for bimanual manipulation.International Conference on Learning Representations (ICLR)

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.794889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:162e06f47742a46bd5b904cd2ad885994eb79267e53d2752cb41a580d0e86ef8

Observation 944475f2-d358-4f95-a303-79d8275edf6f · outbound

This paper cites Scaling proprioceptive-visual learning with hetero- geneous pre-trained transformers.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Scaling proprioceptive-visual learning with hetero- geneous pre-trained transformers

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.926536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:99cc227684b68489ea66ebe5cb690467bfb18cccb12d8dea49b4ed3a694dfc43

Observation 9a8437df-aebd-4291-b13b-c5f816856dc0 · outbound

This paper cites A Careful Examination of Large Behavior Models for Multitask Dexterous Manipulation.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities A Careful Examination of Large Behavior Models for Multitask Dexterous Manipulation

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:32:56.991484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:182b74a4fac0f0fb6fad2a5e650ffdc1a00c0b91fa414f302e49a4ad02e442a0

Observation 28736f9c-9414-4d93-ac38-179d1c6e67bc · outbound

This paper cites Rt-2: Vision-language- action models transfer web knowledge to robotic con- trol.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Rt-2: Vision-language- action models transfer web knowledge to robotic con- trol

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.847745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:17cea32131db883c82b66bb5b42aed3bf0c8773bd713861e04e0fdafc053e2c8

Observation eb215997-e7fc-4089-99b8-1ae54a5b496b · outbound

This paper cites an unresolved cited work.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-05-19T13:27:19.928415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:d22b5cb2198b2c72d6a2cd4a1bd5c9108ff846eabb0ba5d738932140e20f1c06

Observation b604166b-9fe6-48fa-a0bc-23f15e456bc3 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities OpenVLA: An Open-Source Vision-Language-Action Model

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:46:37.168344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:2718de01e8f6791f26c959780bb3347cd873e195362dbb76853a46f4859f271e

Observation 9057856d-2265-4867-a7ad-2520a3eb247a · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-10T12:38:24.665044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:f76dc6570eaa6c30982da8d5499e5f265d6b660663aae743db22de974a6c8dfe

Observation 66051aed-f4ae-4b45-afb0-4257e111006a · outbound

This paper cites TinyVLA: Towards Fast, Data-Efficient Vision-Language-Action Models for Robotic Manipulation.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities TinyVLA: Towards Fast, Data-Efficient Vision-Language-Action Models for Robotic Manipulation

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-17T16:12:26.208161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:5ffcddb08340c23beabe0de9e9124891e6da8fb18e3df7cf13fd7cd3f270b165

Observation fcbf9ec7-030e-40b3-bb66-c11208c6e426 · outbound

This paper cites 3D-VLA: A 3D Vision-Language-Action Generative World Model.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities 3D-VLA: A 3D Vision-Language-Action Generative World Model

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-13T18:18:27.368966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:42f51b5cd57f957bb1c6f7778f36d4f2b580ba0a5af38dc67663fe9d13364229

Observation 4f68c41f-49b4-493b-8f9d-4d969afe6aa1 · outbound

This paper cites Gemini Robotics: Bringing AI into the Physical World.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Gemini Robotics: Bringing AI into the Physical World

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:30:26.569831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:4ef89f78b20f0f20e4a58eabf571c326ea257487b03cc5177f0fba0c2a39dff0

Observation c9c3b61f-9144-46a2-9087-e5645ac06dda · outbound

This paper cites In9th Annual Conference on Robot Learning.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities In9th Annual Conference on Robot Learning

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.874030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:d42ae8b05cd2d84b6e40c514b34fdd00093467705bc8a71ab9808006734c6efe

Observation 6c123208-e091-4ff9-aeb9-fde21e206a6f · outbound

This paper cites X-VLA: Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities X-VLA: Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-12T14:57:47.957311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:64dc6e497ca9dfeb2b623c98c86d10d5607d612c576e008477960f67578069dd

Observation 9e668534-bb55-48eb-9e1e-b083db0bdd75 · outbound

This paper cites Galaxea Open-World Dataset and G0 Dual-System VLA Model.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Galaxea Open-World Dataset and G0 Dual-System VLA Model

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:21.873536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:e75a275ad301f4a26dd9a37971f43f3097a51723dded7b8ddf3a071afe4d0de5

Observation ffa84160-6c3f-42fe-b48a-bd5fa675caf1 · outbound

This paper cites Vision- language foundation models as effective robot imitators.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Vision- language foundation models as effective robot imitators

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.863425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:732b1a5da12962d4efe2ff8a059cda7c71cd31fee8b6914c4d587a569177135b

Observation 8f585254-030f-4ae2-bd7e-fda12d0e6807 · outbound

This paper cites CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:21.844783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:11e02342cb4188b998b76ea0df23051082159bc8b7dbb6fab9a54060f4bd9a95

Observation 533f4318-9037-4260-9f0f-edb371e37449 · outbound

This paper cites Spatialvla: Explor- ing spatial representations for visual-language-action model.Robotics: Science and Systems (RSS).

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Spatialvla: Explor- ing spatial representations for visual-language-action model.Robotics: Science and Systems (RSS)

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.918162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:d3ce2da0d12a8e8bb24dfc2c3d0d3cc5ccba576d8e8c66239d449cf28366f58f

Observation bb4e3ce4-d2f9-4665-8998-3e164266b013 · outbound

This paper cites GR00T N1: An Open Foundation Model for Generalist Humanoid Robots.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities GR00T N1: An Open Foundation Model for Generalist Humanoid Robots

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-10T19:09:10.448691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:1c7f98c9bc71a801791d5075be61080b30b8a60a01df4f114270956ac3f01b31

Observation 96c23fb0-d691-425f-ba27-8cbbeb9f9e16 · outbound

This paper cites Robotic Control via Embodied Chain-of-Thought Reasoning.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Robotic Control via Embodied Chain-of-Thought Reasoning

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-15T08:14:32.442267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:185777e960dfb2b683b5cfe19cfa79940a8d406662405d5e4193347131e6b6b7

Observation c30e9247-9d82-4efa-b738-de353b9756fb · outbound

This paper cites AgiBot World Colosseo: A Large-scale Manipulation Platform for Scalable and Intelligent Embodied Systems.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities AgiBot World Colosseo: A Large-scale Manipulation Platform for Scalable and Intelligent Embodied Systems

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:09:24.919963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:5ab65b4671cc28905e0025d4f86fc4d50e41a431462d91fe1e023d2addc0f031

Observation a3ede2b6-4afd-416e-8adc-8a04621e0239 · outbound

This paper cites ChatVLA: Unified Multimodal Understanding and Robot Control with Vision-Language-Action Model.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities ChatVLA: Unified Multimodal Understanding and Robot Control with Vision-Language-Action Model

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T11:45:21.885175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:3bfa8aca3b055f8d3b71994ac51fdd28076340ab49f52ade624b3fc3f5e35dbf

Observation 42a253e7-08fd-46b7-92e9-47f0ed12cb64 · outbound

This paper cites Cosmos policy: Fine-tuning video models for visuomotor control and planning.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Cosmos policy: Fine-tuning video models for visuomotor control and planning

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.884467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:a33afadce8bc5ac2575f6dfd76a81115d82692e0430eadfa0e096c837d5f798d

Observation a2a80532-2ae8-4288-a36e-b2394014d663 · outbound

This paper cites mimic- video: Video-action models for generalizable robot con- trol beyond vlas.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities mimic- video: Video-action models for generalizable robot con- trol beyond vlas

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.915806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:6a9ed1ad92387b895df2de03d72a517b1099b12d7f46890a331b8f40ad1e4bf5

Observation e81a3363-8f73-45ed-aacd-177c9aa7aac0 · outbound

This paper cites World Action Models are Zero-shot Policies.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities World Action Models are Zero-shot Policies

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:18:16.286837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:f4d26ab9f25cde7107221203cf541d9b9352016fcf41f59a369bab8f391413af

Observation b33fd404-56e6-47bd-9634-d1eac626d3ce · outbound

This paper cites Unleashing large-scale video genera- tive pre-training for visual robot manipulation.Interna- tional Conference on Learning Representations (ICLR).

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Unleashing large-scale video genera- tive pre-training for visual robot manipulation.Interna- tional Conference on Learning Representations (ICLR)

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.855703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:55325fe79e0912f8d050734f87f77cf509fc4adf4e01bc1b97fa86e411d342e6

Observation c96b2652-88f5-40c4-ba0c-abfc23f58041 · outbound

This paper cites GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-12T01:09:34.215368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:566b75bfd9e4aacf354a0d23e7e9d482340890da218a274d4c780e4797099b0c

Observation 67e2f47d-e8c6-4c96-a102-b00519c818f7 · outbound

This paper cites TraceVLA: Visual Trace Prompting Enhances Spatial-Temporal Awareness for Generalist Robotic Policies.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities TraceVLA: Visual Trace Prompting Enhances Spatial-Temporal Awareness for Generalist Robotic Policies

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-15T18:27:23.210785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:45dc9146510b37419c0bcd189b9b64a519e7e5c3499d4a101182c569afc0bfaa

Observation 28e81c60-0052-4540-903c-e2c2128232b4 · outbound

This paper cites Memer: Scaling up memory for robot control via experience retrieval.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Memer: Scaling up memory for robot control via experience retrieval

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:21.550622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:436eff625ec6bd96fb7175843ea961c2415c32fe379a942d3d6396790926adde

Observation efd91af7-1d1c-46d8-9e8f-b2d60d17133e · outbound

This paper cites MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:43:24.547911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:db12071d9cf486e230a5e8207cfb5a449947639cba6f28b37b17450fe83c34e9

Observation e81bab52-6c14-496a-b7bf-07524e74e897 · outbound

This paper cites Onetwovla: A unified vision-language-action model with adaptive reasoning.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Onetwovla: A unified vision-language-action model with adaptive reasoning

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:21.539915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:a98af6870237f0d7d72a3e35b75fb0ecc51024ba87ba41913d82a0772eec81c2

Observation d2a233ed-afcc-4663-802e-2791381a4b97 · outbound

This paper cites SAM2Act: Integrating Visual Foundation Model with A Memory Architecture for Robotic Manipulation.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities SAM2Act: Integrating Visual Foundation Model with A Memory Architecture for Robotic Manipulation

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:21.545683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:5746dd2117854706fad84ca14abe42e8dba426ead03cf4cdf0d1a3519776a020

Observation bf3c97a3-2827-4025-ad5d-4a78dc8d82d3 · outbound

This paper cites CronusVLA: Towards Efficient and Robust Manipulation via Multi-Frame Vision-Language-Action Modeling, October 2025.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities CronusVLA: Towards Efficient and Robust Manipulation via Multi-Frame Vision-Language-Action Modeling, October 2025

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:21.617737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:b7ee205451ab0201228574ac874bc13c3d2c9fb6289c6bd3c7c44c84bbfb3b14

Observation 422b40d1-a1d7-417e-8621-fe5408f3c735 · outbound

This paper cites TA-VLA: Elucidating the Design Space of Torque-aware Vision-Language-Action Models.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities TA-VLA: Elucidating the Design Space of Torque-aware Vision-Language-Action Models

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:21.588337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:893de8975b21354b3d81bd42072021304be9dc1deba5b98f5c2e9498fe7024d8

Observation 470c9e1a-2b3a-40f9-adca-f407efdbd1d0 · outbound

This paper cites arXiv:2510.04246 [cs].

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities arXiv:2510.04246 [cs]

Reference 36

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T11:45:21.571497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:1e7214f9f6b6f73b0bb0bb452bdebe564e57bc73d2c857869be900403b7c5be3

Observation 785090fb-e912-420e-a3ff-17a2802b5957 · outbound

This paper cites Ren, Haohuan Wang, Jiaming Tang, Kyle Stachowicz, Karan Dhabalia, Michael Equi, Quan Vuong, Jost Tobias Springenberg, Sergey Levine, Chelsea Finn, and Danny Driess.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Ren, Haohuan Wang, Jiaming Tang, Kyle Stachowicz, Karan Dhabalia, Michael Equi, Quan Vuong, Jost Tobias Springenberg, Sergey Levine, Chelsea Finn, and Danny Driess

Reference 37

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T11:45:21.593581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:b629a3847b3411f0198758f6ed7b5b3cc6cc084f0559639b5817d8e7d2bfb100

Observation dce42d12-6bcc-4f23-a58c-fe986fdac7c2 · outbound

This paper cites Do as i can, not as i say: Grounding language in robotic affordances.Conference on Robot Learning (CoRL).

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Do as i can, not as i say: Grounding language in robotic affordances.Conference on Robot Learning (CoRL)

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.930443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:243114a96b2426ddd02096e360234d665ea3d8d9530047fc1ff13f0d101aa621

Observation 6ee35dbb-e89e-4b80-9096-a5ca8f2eab6d · outbound

This paper cites Code as policies: Language model programs for embod- ied control.IEEE International Conference on Robotics and Automation (ICRA).

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Code as policies: Language model programs for embod- ied control.IEEE International Conference on Robotics and Automation (ICRA)

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.834287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:905adf4742b728bceadcb2a2a6e01b761a275b6f5f34122061497e67c5318f2f

Observation cbc1d3c1-e7c9-43dc-8af6-2acac9628246 · outbound

This paper cites Hi Robot: Open-Ended Instruction Following with Hierarchical Vision-Language-Action Models.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Hi Robot: Open-Ended Instruction Following with Hierarchical Vision-Language-Action Models

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-15T22:53:37.470313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:f8b34e51661cfa46748cc092bfc3e147629317d23f9a5612920744eef326ad2e

Observation 876aed88-3356-470b-b908-3c84643cb433 · outbound

This paper cites Cot-vla: Visual chain-of-thought reasoning for vision-language-action models.Computer Vision and Pattern Recognition (CVPR).

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Cot-vla: Visual chain-of-thought reasoning for vision-language-action models.Computer Vision and Pattern Recognition (CVPR)

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.829709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:80fb3a688c6602bfb6713722bb1fed624149c659f89adf00f956d272d7c5c96a

Observation 65aee9bc-3046-4953-9ebb-83d506062aa6 · outbound

This paper cites 3, 8, 10.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities 3, 8, 10

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.871663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:3f4132ae7b872af560a8358e3b38c18de508fd7158f651a137be590a70572367

Observation ed6fc1aa-5a18-4638-9451-f87cd753eeb2 · outbound

This paper cites Latent Action Pretraining from Videos.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Latent Action Pretraining from Videos

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:21.748452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:fdd79d09df5ec3a457966bd620db9a8ac8b5576951e8c113e7447cce79d66686

Observation ceb2ac52-d72c-4a16-8e45-da1bf676de5f · outbound

This paper cites Bo Liu, Yifeng Zhu, Chongkai Gao, Yihao Feng, Qiang Liu, Yuke Zhu, and Peter Stone.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Bo Liu, Yifeng Zhu, Chongkai Gao, Yihao Feng, Qiang Liu, Yuke Zhu, and Peter Stone

Reference 44

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T11:45:21.738304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:9f70df297b72509066f581193dc53e7bb9ef8797ba46e59147e30d78f7e1a03f

Observation ffff8023-6043-44a8-a223-8e36d802c67a · outbound

This paper cites Emergence of human to robot transfer in vision-language-action models.arXiv preprint arXiv:2512.22414.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Emergence of human to robot transfer in vision-language-action models.arXiv preprint arXiv:2512.22414

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:21.682819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:8c76c16ce2ffb06f5e9b2b320e14981e1349e339ea570bcf7c3257bef62bc169

Observation 7618d134-1a2b-43f8-b9b8-d4b8d192b5c5 · outbound

This paper cites Latbot: Distilling universal latent actions for vision-language-action models.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Latbot: Distilling universal latent actions for vision-language-action models

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:21.717805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:6f34ecb6a5283c89d0cd00690e7675cc54ad9dce48165d0f6763a5eb4d210e41

Observation 05f8b45d-88af-4951-92b5-3285be8b5e24 · outbound

This paper cites Egovla: Learning vision- language-action models from egocentric human videos.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Egovla: Learning vision- language-action models from egocentric human videos

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.840032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:7a2599cbd2213019a05f495970ec1c1f3bda1e3d091d4a78cb28c05121b3b6f6

Observation 5b3a441a-bf1a-4da0-905a-6e077ac80f7e · outbound

This paper cites EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-21T04:32:58.962571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:b6e614d0c7f02dca9b628a30d5afb47fb5e8b2f1092d494bc987b1462c42878e

Observation e5a7c4a5-fd37-4082-9757-759d3ce5720e · outbound

This paper cites Being-h0.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Being-h0

Reference 49

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T11:45:21.688378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:71b726f528146a4fcf8976897a4dc202903c8d6e41bd7e973a90ce8d825ce621

Observation 160350f8-1e29-4239-9865-8e825fa59622 · outbound

This paper cites Clap: Contrastive latent action pretraining for learn- ing vision-language-action models from human videos.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Clap: Contrastive latent action pretraining for learn- ing vision-language-action models from human videos

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.869089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:bcb9ddb3e3743ca1909c11be7bd85d28ea239a944a0dd25c643c4f9f24c26c7f

Observation 10b32a56-4677-41ba-a6d8-86c04296739d · outbound

This paper cites Clap: Contrastive latent action pretraining for learning vision-language-action models from human videos.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Clap: Contrastive latent action pretraining for learning vision-language-action models from human videos

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:21.722753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:13f101749d3b782a7e9c01b6685483b695fde75a5d8efc113b476d4d41980e83

Observation 849a9f66-883f-42de-b293-8ff9b5fac114 · outbound

This paper cites $\pi^{*}_{0.6}$: a VLA That Learns From Experience.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities $\pi^{*}_{0.6}$: a VLA That Learns From Experience

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:21.702816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:8ca9a1b042b5c32ad226d961285427074a882b786a78dbe9f3a8070a5f6d56f1

Observation a6bac5bf-d19d-453b-b1a2-575a2262b002 · outbound

This paper cites RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:21.728091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:b89ccf94f27a462fb179465d724abcf29dddbee280198eabd2fa208461f96f40

Observation c99354fd-9f99-4b5d-937d-1ca158176505 · outbound

This paper cites arXiv preprint arXiv:2511.00091 , year=.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities arXiv preprint arXiv:2511.00091 , year=

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:21.796763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:96c36bf9e1c91f24f71195449020a00c6322b0eab677b5bf6c336886e1e1126f

Observation 1676fdf2-44b7-4a67-b8d4-1f2f34f2c31c · outbound

This paper cites R3m: A universal visual representation for robot manipulation.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities R3m: A universal visual representation for robot manipulation

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.899939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:619d29cd3da4ab26a8c334f4cceeb1cf014ff401f39d24d64e98cfdfe0595305

Observation a3b03959-b8de-46c2-a7d7-3dc2f285600a · outbound

This paper cites Vip: Towards universal visual reward and representation via value-implicit pre-training.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Vip: Towards universal visual reward and representation via value-implicit pre-training

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.902399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:711945cb8ad4243148a1618d3b410b626d4c2b42b96327a39bba47e4cd328e93

Observation efb8432f-ceea-4659-a5c2-ba7500d588de · outbound

This paper cites Masked Visual Pre-training for Motor Control.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Masked Visual Pre-training for Motor Control

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:21.877887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:3cf57f543af4168e11baf1361d97903d89a4e0eda0f3827608e8833cae79bbe1

Observation 5cbe4a0f-6ef7-4cd0-9d09-3ab84961b08a · outbound

This paper cites Robotic Offline RL from Internet Videos via Value-Function Pre-Training.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Robotic Offline RL from Internet Videos via Value-Function Pre-Training

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:21.855813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:72847d6d5a452a74389e35120428f74df4a9c4930a68a2eabfa87fa2736fbcf2

Observation b4dc4b58-1754-4e82-a377-276116123533 · outbound

This paper cites Manipulator-Independent Representations for Visual Imitation.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Manipulator-Independent Representations for Visual Imitation

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:21.850260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:baa65803819c424d9c557b1c53fd20a1747b086890e78145fa4b78cdee3e178e

Observation 3e2bdfed-ba71-44c9-b28a-3b63434956e1 · outbound

This paper cites Visual Affordance Prediction for Guiding Robot Exploration.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Visual Affordance Prediction for Guiding Robot Exploration

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:21.889438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:b116682cbe95b4feb889d616eddd834262adce88a0a5fc8b28654f43f2c92ccf

Observation 2e3d698f-6101-4c33-b357-6faa78168c6a · outbound

This paper cites Dexterous manipulation policies from rgb human videos via 3d hand-object trajectory reconstruction.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Dexterous manipulation policies from rgb human videos via 3d hand-object trajectory reconstruction

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:21.763758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:d9a8b772be51724b38de0abb89cbbca1eb728e3e7e5feecbab12da67a7761f3e

Observation f4d34dfa-19c1-477f-9345-0c24cea73599 · outbound

This paper cites Videodex: Learning dexterity from internet videos.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Videodex: Learning dexterity from internet videos

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.904883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:c82422cbc4e3545d3bd9a36bef3c6e60ec45b9b96b5c343b6afa0ee602018c5a

Observation b9a8d842-c514-489b-b46c-2870c3cbda38 · outbound

This paper cites Zero-Shot Robot Manipulation from Passive Human Videos.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Zero-Shot Robot Manipulation from Passive Human Videos

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:21.893947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:fd90b11e23753187411d8e7901e868db829456b7bb1eeaf7c31496ddfe2b2372

Observation ac26a394-dba4-4306-99d3-bcbce89c6db1 · outbound

This paper cites Human-to-Robot Imitation in the Wild.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Human-to-Robot Imitation in the Wild

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:21.823777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:8f9817b321643abae227287d77398f8030727b8333bceb03be1ccf31fc00bee9

Observation a4256272-88ab-451e-8352-c1f71bdf81bd · outbound

This paper cites Affordances from human videos as a versatile representation for robotics.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Affordances from human videos as a versatile representation for robotics

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.858460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:95f32737f49b3c214e496174e506d325d95c04650d5ab84f3e88b5dd6270d091

Observation 7b940472-fe87-45f1-b657-1af6d0d1eb29 · outbound

This paper cites Egomimic: Scaling imitation learning via egocentric video.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Egomimic: Scaling imitation learning via egocentric video

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.910836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:7a3dbbc129a927659c8c79aaa7afc80709005925945de2ec77c4ca613c4750ef

Observation c135fd5b-9626-4e83-ac98-8beff8d750f5 · outbound

This paper cites Learning adaptive dexterous grasping from single demonstrations.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Learning adaptive dexterous grasping from single demonstrations

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.843417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:8517926ebbe10e12b3b6385206ba6922b34305e171bd8750ec7eb5e0e1167423

Observation a40188c9-05e3-4609-badf-ce3e9941749f · outbound

This paper cites Track2act: Predicting point tracks from internet videos enables generalizable robot manipulation.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Track2act: Predicting point tracks from internet videos enables generalizable robot manipulation

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.889487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:ca944745ea441b3ccd910005f013f39a015de4eb5342a9f4e78082fed838f2b9

Observation 1a492ee5-9642-4c1c-9f4a-e577c67a94d6 · outbound

This paper cites Robotap: Tracking arbitrary points for few-shot visual imitation.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Robotap: Tracking arbitrary points for few-shot visual imitation

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.787320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:1d4019c15b43da9bd3cc68a2a89617c30501404b972d1cbfbfbf1c849595ff58

Observation 4321d485-7c21-43f5-9950-0d3ba56657e6 · outbound

This paper cites Any-point Trajectory Modeling for Policy Learning.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Any-point Trajectory Modeling for Policy Learning

Reference 70

Resolution
verified exact
arxiv_id, observed 2026-05-16T23:32:51.209901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:9bf456c8d80993b6fc6a6d5d9900f60611658b5f720896a8a5d11dd503e7a15d

Observation 684b281b-b9de-4f04-9776-4faa3f304c70 · outbound

This paper cites Rt-trajectory: Robotic task generalization via hindsight trajectory sketches.International Conference on Learning Representations (ICLR).

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Rt-trajectory: Robotic task generalization via hindsight trajectory sketches.International Conference on Learning Representations (ICLR)

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.808360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:87090067bb5584ba88be82927c7955d7881212a4a600d69735b7b928e1ec3fba

Observation 818e9e97-5c2c-440d-bd1d-49328ee03cf3 · outbound

This paper cites Dall-e-bot: Introducing web-scale diffusion models to robotics.IEEE Robotics and Automation Letters, 8(7): 3956–3963.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Dall-e-bot: Introducing web-scale diffusion models to robotics.IEEE Robotics and Automation Letters, 8(7): 3956–3963

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.866565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:2dc0b10a4769420612aa45ffa1390880b0808212d3d1022071b112be74c581c9

Observation c0ef0fc2-307a-41ce-9fd5-e5f940d02a43 · outbound

This paper cites CACTI: A Framework for Scalable Multi-Task Multi-Scene Visual Imitation Learning.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities CACTI: A Framework for Scalable Multi-Task Multi-Scene Visual Imitation Learning

Reference 73

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T11:45:21.599029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:53e0afaf885b68870574c79e73ae2bd0113cdfc5fe7de4dbbf5e809f4988abea

Observation e7b9427a-bca5-4cfa-b7e9-60af368e5a20 · outbound

This paper cites GenAug: Retargeting behaviors to unseen situations via Generative Augmentation.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities GenAug: Retargeting behaviors to unseen situations via Generative Augmentation

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:21.612020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:84c308cee9af63f49adc4fd311c941b7ae8681d4d6ca59fcec2f05d46e4df076

Observation 456a735d-28c4-475c-8a45-32245624e926 · outbound

This paper cites Scaling Robot Learning with Semantically Imagined Experience.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Scaling Robot Learning with Semantically Imagined Experience

Reference 75

Resolution
verified exact
arxiv_id, observed 2026-05-17T18:59:10.716066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:6dffe00ed47159a342f6efe76df8516447f9b84b113528c219dddc6ea45acc8d

Observation 35314711-93ca-455e-b754-e5bb313e6b82 · outbound

This paper cites Open-World Object Manipulation using Pre-trained Vision-Language Models.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Open-World Object Manipulation using Pre-trained Vision-Language Models

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:21.565939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:871ab5372e0e8c9aa1b91b79c0125dfc8e6c7a24cee18f9ceedc5c88eac09757

Observation b1fc5cff-c7aa-4a21-8359-5e0af40926b8 · outbound

This paper cites Sajjadi, et al.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Sajjadi, et al

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.792963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:c0cef070aab187f94607fa11b7b177778a31aa13d3aa73f937e78448207f7b59

Observation 5071beb4-bfb4-46c9-af44-ce52d7ca7919 · outbound

This paper cites Vima: General robot manipulation with multimodal prompts.Interna- tional Conference on Machine Learning (ICML).

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Vima: General robot manipulation with multimodal prompts.Interna- tional Conference on Machine Learning (ICML)

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.860992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:a8f4c36be338ec9f22a20ef43c6bb5670ae40bc344febef194e94a7ca850c09a

Observation ad8f95d8-fcf3-4dcd-a44e-dcd64e68e0fa · outbound

This paper cites Open X-Embodiment: Robotic Learning Datasets and RT-X Models.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Open X-Embodiment: Robotic Learning Datasets and RT-X Models

Reference 79

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:23:25.503833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:798649047431025fe90fa5fa671bfefbb5b5c3c95c06b11237fc0340dc1d6ff6

Observation b8df06fc-4467-43eb-ad16-9c2e01f5ced0 · outbound

This paper cites Data analogies enable efficient cross-embodiment transfer.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Data analogies enable efficient cross-embodiment transfer

Reference 80

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:21.582513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:8ab7406883689ae6608080431a3cd509c46445a7c7775282339d3e890dde0106

Observation 21c74a01-7cef-43ff-9bb3-9a8781849e16 · outbound

This paper cites Scaling cross-embodied learning: One policy for manipulation, navigation, locomotion and aviation.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Scaling cross-embodied learning: One policy for manipulation, navigation, locomotion and aviation

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.791250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:c0e9345389b424191196a135a3e3c9812dc3bf0fc5d515153e6c197411edc013

Observation b758beb3-a8de-499a-839d-8611baaeb51e · outbound

This paper cites Pushing the Limits of Cross-Embodiment Learning for Manipulation and Navigation.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Pushing the Limits of Cross-Embodiment Learning for Manipulation and Navigation

Reference 82

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:21.622949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:f4185670bcfd1d95930e6df60d327c4f30f010a2a50b7066c3e0036ff1cce3f2

Observation 7630a27e-a534-4e17-b555-a89629b89340 · outbound

This paper cites Lap: Language-action pre-training enables zero-shot cross-embodiment transfer.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Lap: Language-action pre-training enables zero-shot cross-embodiment transfer

Reference 83

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:21.813385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:e41afac6ad497e4749718c4aadc93a4c4a9c26b7895f3bbcf6fbdc7de139cb06

Observation 4eed333c-bb80-4749-9168-6c2c72ff8772 · outbound

This paper cites Enhancing generalization in vision-language-action models by preserving pretrained representations.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Enhancing generalization in vision-language-action models by preserving pretrained representations

Reference 84

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:21.818836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:0e0520729372dbbb0eabaef579dc34a28fd15523054b10915c284f34a4d49d02

Observation 672df567-1336-4126-99f4-4abc4d95c0bf · outbound

This paper cites Towards embodiment scaling laws in robot locomotion.Conference on Robot Learning (CoRL).

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Towards embodiment scaling laws in robot locomotion.Conference on Robot Learning (CoRL)

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.797224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:afb4742c3deb6da5d1740c1737422fe565265db640aacac9d948506d97abadc6

Observation b9824886-c32f-4ae7-a6cd-f72c74928edf · outbound

This paper cites Scaling Cross-Embodiment World Models for Dexterous Manipulation.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Scaling Cross-Embodiment World Models for Dexterous Manipulation

Reference 86

Resolution
verified exact
arxiv_id, observed 2026-07-22T01:22:15.500710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:faaf9f7320c6067fc1522b9186535c7fdae7d93517dc0596bf6267b26d1e2d01

Observation 224c84cc-584b-4018-b362-3827b57a187b · outbound

This paper cites Universal manipulation interface: In- the-wild robot teaching without in-the-wild robots.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Universal manipulation interface: In- the-wild robot teaching without in-the-wild robots

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.907983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:682f4c4e4ac6fcac58dc3bf5aa0266f0e38346caabcce7bf89803d29ea83961d

Observation 24cc32e7-9fe7-499f-8c04-c0d0cd274faa · outbound

This paper cites Visual imitation made easy.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Visual imitation made easy

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.876191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:783d6e43f891b17b98500a422b192fa2cf069e544cffb223bdaf80027a2f2ab0

Observation 785fc0f3-168b-4faa-9db3-d19a88ac8959 · outbound

This paper cites Zero-shot visual imitation.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Zero-shot visual imitation

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.837546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:76fa5c08e00c83e00ac00b6c670b853a543c7b8c6cfa90adfb3cabc60cb828b2

Observation 86a198e7-de8b-4178-9f99-a53ecb9a41de · outbound

This paper cites Actionable models: Unsupervised offline reinforcement learning of robotic skills.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Actionable models: Unsupervised offline reinforcement learning of robotic skills

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.879308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:b1e4364bdb08103f49129f9227436593a259c739dd84529f6ead07933ccedb94

Observation 15d28256-1bc6-4260-a19a-63b42d4ca2ac · outbound

This paper cites RoboCat: A Self-Improving Generalist Agent for Robotic Manipulation.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities RoboCat: A Self-Improving Generalist Agent for Robotic Manipulation

Reference 91

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:21.769418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:59b61600857d59042481595ca6b1ddf976e4f444d5df8789a86182f0450ab1ab

Observation fd01362e-981d-43c5-828d-e3b4de84f4de · outbound

This paper cites Goal representations for instruction following: A semi-supervised language interface to control.Con- ference on Robot Learning (CoRL).

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Goal representations for instruction following: A semi-supervised language interface to control.Con- ference on Robot Learning (CoRL)

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.921299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:1e91e73fc384681361d9a6e67c91241fccfcebb47e7f773c13320d8d1c6ce57f

Observation 35bd5f53-1139-40c5-89e1-43a4dd77d48c · outbound

This paper cites Visual reinforce- ment learning with imagined goals.Advances in neural information processing systems, 31.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Visual reinforce- ment learning with imagined goals.Advances in neural information processing systems, 31

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.886800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:03f7f3e6057e8332c11c7f7fc6579ebc19407e5cc7dcf51881d20aa3310f4c9e

Observation e9a933bb-e341-4be1-af94-4e3e85b9de1f · outbound

This paper cites Hierarchical foresight: Self-supervised learning of long-horizon tasks via visual subgoal generation.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Hierarchical foresight: Self-supervised learning of long-horizon tasks via visual subgoal generation

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.913476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:42744c64e483372607d91bda48547d52cf3e4e67c1d2f4535f750acc8ad0ca21

Observation d94f16ba-6759-4add-b488-4efaf5b5b6d1 · outbound

This paper cites Zero-Shot Robotic Manipulation with Pretrained Image-Editing Diffusion Models.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Zero-Shot Robotic Manipulation with Pretrained Image-Editing Diffusion Models

Reference 95

Resolution
verified exact
arxiv_id, observed 2026-05-16T05:54:59.300665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:15ef4705d77904f9ca845f4839c8c766f9f7166a9c376d5d5388689ff267034e

Observation f10fbf37-5c92-40d8-8359-eba1102a5937 · outbound

This paper cites Learning to act from ac- tionless videos through dense correspondences.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Learning to act from ac- tionless videos through dense correspondences

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.894767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:2f7894ab444bfffab43ced84bce3c333c7173a36d1f384a447a6e5be42a9bc1a

Observation 41077997-830f-409e-a871-634ed3952f06 · outbound

This paper cites Uniskill: Imitating human videos via cross-embodiment skill rep- resentations.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Uniskill: Imitating human videos via cross-embodiment skill rep- resentations

Reference 97

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.924026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:366b13c93df109cfaa45fa74e3fc5dc4b0422f36d57f04837a86f4a501b78db7

Observation b17f9ab7-1ed0-4818-af75-c660d8d8a048 · outbound

This paper cites Dreamitate: Real-world visuomotor policy learning via video generation.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Dreamitate: Real-world visuomotor policy learning via video generation

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.850086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:de4e18392d33611c80d2faa8a20508b16f1f2e03e828d0480456a6f153f49f96

Observation 50c9fb8c-6dd9-4d0c-812a-eb86ab25a33c · outbound

This paper cites Foreact: Steering your vla with efficient visual foresight planning.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Foreact: Steering your vla with efficient visual foresight planning

Reference 99

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:21.698336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:9efb5b8ba632b3d204db46b1302e6c8296117daa19d9abcb9e8bbc709cb33621

Observation f8da7cee-b043-4403-bfbf-5d9d90cb5862 · outbound

This paper cites Tenenbaum, Leslie Kael- bling, Andy Zeng, and Jonathan Tompson.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities Tenenbaum, Leslie Kael- bling, Andy Zeng, and Jonathan Tompson

Reference 100

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:27:19.881805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:387d47e16f5a5ab579379f73bb99a4cd6378f40271843378c87e5b21c7e00a42

Pith citing papers

Observation 57decbdb-b9c4-47cf-8717-c42536e4f690 · inbound

DiscreteRTC: Discrete Diffusion Policies are Natural Asynchronous Executors cites this paper.

DiscreteRTC: Discrete Diffusion Policies are Natural Asynchronous Executors ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-05-11T22:46:12.108198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T02:30:30.319810Z digest=sha256:bd0b3f715aa10bc1c7a954bc3210f8c836048c8a51be69a7aabfc1147f7d7939

Observation 0f15a919-d0a6-4039-8113-c466d04409f3 · inbound

DiscreteRTC: Discrete Diffusion Policies are Natural Asynchronous Executors cites this paper.

DiscreteRTC: Discrete Diffusion Policies are Natural Asynchronous Executors ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-07-01T08:35:33.472337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T08:32:26.160039Z digest=sha256:26a03301da17a0f1d5db8e8e6f5188f53914259a3628a10373c6ed20b0af51dd

Observation 61684c59-2a70-42fd-b434-52dced25ab07 · inbound

LaST-R1: Reinforcing Robotic Manipulation via Adaptive Physical Latent Reasoning cites this paper.

LaST-R1: Reinforcing Robotic Manipulation via Adaptive Physical Latent Reasoning ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-05-12T10:31:29.925016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-07T05:47:17.494531Z digest=sha256:c185084e7adc99df446f108bee831dd995058290d7ea560d6ddff727f48d97a0

Observation 0a98a4f9-2ed1-4ec1-8137-e582c3942887 · inbound

LaST-R1: Reinforcing Robotic Manipulation via Adaptive Physical Latent Reasoning cites this paper.

LaST-R1: Reinforcing Robotic Manipulation via Adaptive Physical Latent Reasoning ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-05-11T22:17:01.686696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T03:00:26.352130Z digest=sha256:9562b4ea7f587451535e05608aa1df2a01fd5f02935896fec10f5ba3136749c9

Observation 25651082-a28c-4413-949e-8b438d6539a0 · inbound

RLDX-1 Technical Report cites this paper.

RLDX-1 Technical Report ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-12T10:51:30.995242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-07T16:06:10.595164Z digest=sha256:c8592f75e636694b10a2a4227e2da47fb6d31525ebc1acaa324c67ae77abb795

Observation ffdc0da3-2d4b-457f-a807-1e7852b1af44 · inbound

RLDX-1 Technical Report cites this paper.

RLDX-1 Technical Report ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-09T06:05:35.161958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T19:02:19.756092Z digest=sha256:2a19bfdc2f1a5c88c8a5a1119ba8af97f13a3032fda006c2dc9048bf540e465f

Observation d36c808c-c977-4def-bab4-da222c716490 · inbound

Kintsugi: Learning Policies by Repairing Executable Knowledge Bases cites this paper.

Kintsugi: Learning Policies by Repairing Executable Knowledge Bases ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-12T06:31:28.553591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-12T04:11:13.607637Z digest=sha256:8506a541c85df5887e56dc83c1109625df6ad708de2251f2bea17e31735ddf29

Observation f7083cac-59b6-48c6-8be3-5b8265d714a2 · inbound

The Agent Use of Agent Beings: Agent Cybernetics Is the Missing Science of Foundation Agents cites this paper.

The Agent Use of Agent Beings: Agent Cybernetics Is the Missing Science of Foundation Agents ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-05-12T05:41:25.913020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-12T05:03:58.419364Z digest=sha256:3c7e3fadaf71bb3acb72a907482d5ecc75c6daea75c7c877721aed37dba5a5b4

Observation f2eebd49-ca7a-4385-b735-ea6654994c30 · inbound

HarmoWAM: Harmonizing Generalizable and Precise Manipulation via Adaptive World Action Models cites this paper.

HarmoWAM: Harmonizing Generalizable and Precise Manipulation via Adaptive World Action Models ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-05-12T07:21:24.085346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-12T03:28:18.730751Z digest=sha256:de2ecc08450549d8e21379ab4fcbdb4a7f4057dede122ac5edcb19ebf7a8ecd9

Observation 0ce0b116-9906-4b0a-9bc8-8515b95a5f5d · inbound

Engagement Process: Rethinking the Temporal Interface of Action and Observation cites this paper.

Engagement Process: Rethinking the Temporal Interface of Action and Observation ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-05-13T01:42:04.094932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T01:37:31.252556Z digest=sha256:fcac1aacc52119a509a1802ba35d9c630e5b1353c0cca4da4aa5e9634d889e46

Observation 2d2d474e-1747-4359-88dd-4dbe05c6e600 · inbound

Engagement Process: Rethinking the Temporal Interface of Action and Observation cites this paper.

Engagement Process: Rethinking the Temporal Interface of Action and Observation ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-07-01T13:45:46.507182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T22:40:05.658212Z digest=sha256:29076430ea34be7d3df07e8c0cc4c282eff8e3386e212374c007a33a97dd9801

Observation 243e0eb2-d1f2-4541-8b35-6d1bc3c1ff45 · inbound

World Action Models: The Next Frontier in Embodied AI cites this paper.

World Action Models: The Next Frontier in Embodied AI ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 85

Resolution
verified exact
local_arxiv, observed 2026-05-13T05:02:17.798098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T05:01:16.802019Z digest=sha256:399e495ef8e705043d51608f261b2c35d35164dc60ca7e7445f3e4f9d10410d2

Observation d0bfa7a3-35e4-41d2-beaf-cf461f22e23a · inbound

WorldVLN: Autoregressive World Action Model for Aerial Vision-Language Navigation cites this paper.

WorldVLN: Autoregressive World Action Model for Aerial Vision-Language Navigation ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-05-20T17:48:48.697975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-20T17:47:32.953903Z digest=sha256:71f470c065036ebb3fe0566d4d5c20d743eda42b066b8af491964a536f45ad3c

Observation 4a6f0dbe-0af6-403e-8a2f-e6136809a796 · inbound

Learning Bilevel Policies over Symbolic World Models for Long-Horizon Planning cites this paper.

Learning Bilevel Policies over Symbolic World Models for Long-Horizon Planning ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-05-20T18:38:52.800872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-20T18:34:53.530617Z digest=sha256:29949726a05e52ac818d3ee062f96199d32658e83bb1177b6e5cfd15463fecfa

Observation 62a0f0a6-08b6-4aaf-91d4-fa43621fe939 · inbound

PAPO-VLA: Planning-Aware Policy Optimization for Vision-Language-Action Models cites this paper.

PAPO-VLA: Planning-Aware Policy Optimization for Vision-Language-Action Models ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-05-20T05:13:03.409152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-20T05:12:26.907425Z digest=sha256:d964582cb8662c13c4ed6e53633918b202ae55fdcdec34e2d3f8eefcbd1a62a0

Observation 451087b1-17e1-4a95-9cee-3975be5e5356 · inbound

Safe and Steerable Geometric Motion Policies for Robotic Dexterous Manipulation cites this paper.

Safe and Steerable Geometric Motion Policies for Robotic Dexterous Manipulation ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-05-22T08:34:45.122947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-22T08:33:58.052653Z digest=sha256:d0ba4d280fc7b640f213f06bf3c458076952e613c59ad974313a5975e235a964

Observation 4ac5759b-b8b5-4356-9759-047272e538ae · inbound

Efficient Agentic Reasoning Through Self-Regulated Simulative Planning cites this paper.

Efficient Agentic Reasoning Through Self-Regulated Simulative Planning ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-05-22T06:34:41.010252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-22T06:33:36.846345Z digest=sha256:be055c7c1d52344ffd7856cb7e7c4ad2a2aa6f56d02be28bdf7750cf9c8bc627

Observation d40a8543-9797-45e5-a69a-ee0055a5ddb2 · inbound

Action with Visual Primitives cites this paper.

Action with Visual Primitives ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-05-22T05:31:08.057225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-22T05:26:55.722814Z digest=sha256:600cfa92442528ad4b6661964d755717881627888afff9a1fd228f0612f7f799

Observation 2d65b839-22fb-4bfc-8f31-622746408066 · inbound

Imagine2Real: Towards Zero-shot Humanoid-Object Interaction via Video Generative Priors cites this paper.

Imagine2Real: Towards Zero-shot Humanoid-Object Interaction via Video Generative Priors ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-22T06:04:39.515026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-22T06:02:40.043871Z digest=sha256:4e8d26f65335688fa8a57380b96493da5ab4dc54d021eed359b62335b254a02d

Observation db434718-f1b1-4a7b-8a3b-a495ee45103e · inbound

Imagine2Real: Towards Zero-shot Humanoid-Object Interaction via Video Generative Priors cites this paper.

Imagine2Real: Towards Zero-shot Humanoid-Object Interaction via Video Generative Priors ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-25T06:00:23.034862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-25T05:59:52.760486Z digest=sha256:0dd5e9309974918648827325c43f336291e9a0df39426c2c3e9a69cf53fe2c07

Observation 6b5ba20f-db10-4bc9-9265-ca0979bc5a51 · inbound

QuoVLA: Quotient Space for Vision-Language-Action Models cites this paper.

QuoVLA: Quotient Space for Vision-Language-Action Models ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-06-30T12:34:39.488469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T12:09:12.124995Z digest=sha256:c024746bc3f5c26ab6c056f90cde174f073569ca730376014c3309ba5f6b3de7

Observation ecc78742-a561-4145-83ea-7fe9e7683a4b · inbound

Scaling World-Model Reinforcement Learning Through Diffusion Policy Optimization cites this paper.

Scaling World-Model Reinforcement Learning Through Diffusion Policy Optimization ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-06-29T22:54:01.433052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T22:46:27.179341Z digest=sha256:47ed8246606f225c3e6a1ee8fc1e910e144f058a4af5a40610d6ab168a08a6c1

Observation 77170b98-4ed4-43b6-906b-8cbedf936175 · inbound

FineVLA: Fine-Grained Instruction Alignment for Steerable Vision-Language-Action Policies cites this paper.

FineVLA: Fine-Grained Instruction Alignment for Steerable Vision-Language-Action Policies ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-06-29T17:23:44.921783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T17:19:04.477325Z digest=sha256:d997b3133e4dd3890d5c4f26c3b61f0002d58d7ece1dfb770b6969840c227c99

Observation c1ab1ea2-4b5c-456a-828b-232da0e09fbe · inbound

PrimitiveVLA: Learning Reusable Motion Primitives for Efficient and Generalizable Robotic Manipulation cites this paper.

PrimitiveVLA: Learning Reusable Motion Primitives for Efficient and Generalizable Robotic Manipulation ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-06-29T11:53:24.300660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T11:46:14.257386Z digest=sha256:a85a81ae9e059d2874894230b4833365284f706aa715c34a025b294d5fd7f7c0

Observation 14a07dd2-e445-4936-80da-7be1f2fb767a · inbound

PrimitiveVLA: Learning Reusable Motion Primitives for Efficient and Generalizable Robotic Manipulation cites this paper.

PrimitiveVLA: Learning Reusable Motion Primitives for Efficient and Generalizable Robotic Manipulation ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T12:57:54.284324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:57:54.284324Z digest=sha256:4f050b467adbe4530233a834ab61a4120da01b9753c687d1f2d8ff45561f66e4

Observation 0113793f-5209-4316-8a02-31e5a98c4263 · inbound

Wall-OSS-0.5 Technical Report cites this paper.

Wall-OSS-0.5 Technical Report ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 81

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:26:00.503450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T22:35:00.258436Z digest=sha256:120452e868f5848bfaf2af2983b564723104ecd9057b21ae0cb94c75d2af7bf8

Observation 787b2246-765f-4eb9-afc3-6cbb54dedd0e · inbound

Make Your VLA More Robust Without More Data By Interleaving Motion Planning cites this paper.

Make Your VLA More Robust Without More Data By Interleaving Motion Planning ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-07-01T21:06:13.499645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T17:31:55.850609Z digest=sha256:46638deae7bd9337d29f98b43e2434119dda0a73239cf10c443694738dd99c96

Observation 7b656867-0b6f-46dd-bca3-a01aacd4eb35 · inbound

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models cites this paper.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-07-02T03:26:28.487849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:4366488a4240ef38d9464dc1a44256575fcd685c6890f4c66b637f6a8801faa7

Observation 3d16579e-7890-4fa4-b3b8-9db81188f035 · inbound

World-Language-Action Model for Unified World Modeling, Language Reasoning, and Action Synthesis cites this paper.

World-Language-Action Model for Unified World Modeling, Language Reasoning, and Action Synthesis ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-07-02T13:46:58.925625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T01:05:26.382826Z digest=sha256:897b5cc1a70007cfdb6bfcbbfd7513ea9c2fecae13529d8520565415443c7d09

Observation 89941633-26df-4a4a-9c0d-35d79d69ec6d · inbound

AffordanceVLA: A Vision-Language-Action Model Empowering Action Generation through Affordance-Aware Understanding cites this paper.

AffordanceVLA: A Vision-Language-Action Model Empowering Action Generation through Affordance-Aware Understanding ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-07-02T13:16:59.196143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T01:23:02.576098Z digest=sha256:88f9c624ef119848b2229a295db3548691cced89f98e73994d9bafdd3e4060bf

Observation 1799288a-3892-4e66-a616-fd81c480550a · inbound

VoLo: A Physical Orchestrator for Open-Vocabulary Long-Horizon Manipulation cites this paper.

VoLo: A Physical Orchestrator for Open-Vocabulary Long-Horizon Manipulation ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-07-02T19:07:18.099164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T21:40:00.330510Z digest=sha256:04f1d6bb2a7ac16792947f3d614274dc93379a212203fa686776fdd40d08fe57

Observation 079a8326-2807-43bb-841d-2885aa08e94e · inbound

Reinforcement Learning for Flow-Matching Policies with Density Transport cites this paper.

Reinforcement Learning for Flow-Matching Policies with Density Transport ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-07-02T22:27:25.787364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T18:55:02.040180Z digest=sha256:fd211bd282b36370ace783b71d94d4affa3dc4c1b53c6b831a1c877883f9eb51

Observation 5fea3549-5fac-4c09-94b0-626a642f8b04 · inbound

MemoryVLA++: Temporal Modeling via Memory and Imagination in Vision-Language-Action Models cites this paper.

MemoryVLA++: Temporal Modeling via Memory and Imagination in Vision-Language-Action Models ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 65

Resolution
verified exact
local_arxiv, observed 2026-07-03T01:57:32.211561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T16:13:19.238535Z digest=sha256:7e7f0d105c75e6d2813d31ba60a8f4953060a1d39a556f5a7e9f3aaa5cc576b0

Observation 7be85c3c-433c-4379-b032-3e1e94b6382b · inbound

What Matters in Orchestrating Robot Policies: A Systematic Study of Hierarchical VLA Agents cites this paper.

What Matters in Orchestrating Robot Policies: A Systematic Study of Hierarchical VLA Agents ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-07-03T04:47:38.581329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T13:36:56.552395Z digest=sha256:6a104add1c9987137fbf9df8c22e29187042a77425a377e5b06c75374f8c0a5e

Observation f32c9224-030d-48ef-91b6-a01e306a6a1f · inbound

Hierarchical Policies from Verbal and Egocentric Human Signals for Natural Human-Robot Interaction cites this paper.

Hierarchical Policies from Verbal and Egocentric Human Signals for Natural Human-Robot Interaction ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-07-03T04:57:38.577014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T13:31:02.640975Z digest=sha256:add40b37d1f3feaf7607519ede5d51c281d24c987adae68e0122935ab6bb0a80

Observation ab57231d-b6a9-4b2d-b208-8b0863178a7b · inbound

Next Forcing: Causal World Modeling with Multi-Chunk Prediction cites this paper.

Next Forcing: Causal World Modeling with Multi-Chunk Prediction ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 45

Resolution
verified exact
local_arxiv, observed 2026-07-03T04:57:38.510774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T13:31:53.905704Z digest=sha256:682a0f5ee23d2a7888fb71ecfe24526c0626cffb62f79fed42c134ae301865f7

Observation 4fefe77e-421c-439c-9cde-aa421f39dc1e · inbound

Learning What to Say to Your VLA: Mostly Harmless Vision Language Action Model Steering cites this paper.

Learning What to Say to Your VLA: Mostly Harmless Vision Language Action Model Steering ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-07-03T10:27:56.307007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T10:02:35.406713Z digest=sha256:81c5b261b9ff7190b3c2c7ad2617ef386c33f1aa3729c0de78acfc233255cbfd

Observation 7119e630-13c4-4492-8669-737aa8201686 · inbound

APT: Action Expert Pretraining Improves Instruction Generalization of Vision-Language-Action Policies cites this paper.

APT: Action Expert Pretraining Improves Instruction Generalization of Vision-Language-Action Policies ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 45

Resolution
verified exact
local_arxiv, observed 2026-07-03T10:48:02.725096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T09:49:56.894300Z digest=sha256:95bd9c46b15825de6550a4eb286e862627401b9d70bc7fa513f7f7eb79d94f95

Observation 9663f9d9-c3e0-4893-9eb9-25e951ae176c · inbound

World Pilot: Steering Vision-Language-Action Models with World-Action Priors cites this paper.

World Pilot: Steering Vision-Language-Action Models with World-Action Priors ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 60

Resolution
verified exact
local_arxiv, observed 2026-07-03T11:18:03.343910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T09:40:02.137152Z digest=sha256:43f37fe491d6f59acd9f745d13b6f0b424032a2546e90bf04ecdf67633550053

Observation 79dbe04c-ac9c-4d99-a7b3-1d5549ac0bb3 · inbound

Action-Effect Memory Pretraining for Robot Manipulation cites this paper.

Action-Effect Memory Pretraining for Robot Manipulation ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-07-03T11:48:04.875396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T09:22:17.435427Z digest=sha256:cd45af7e3bb980cc0656fdc1fe9a434503de48681c74e3f3bce21fe463701133

Observation dc9a6292-a9a3-42ce-b843-fdab21230eaf · inbound

AIR-VLA+: Decoupling Movement and Manipulation via Cascaded Dual-Action Decoders with Asymmetric MoE for Aerial Robots cites this paper.

AIR-VLA+: Decoupling Movement and Manipulation via Cascaded Dual-Action Decoders with Asymmetric MoE for Aerial Robots ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-07-03T14:58:32.743690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T06:51:11.482211Z digest=sha256:ce6f7c307c84ae4706b0f5ffe71c121cf7ea0390d66fe3fd88ac895864e6c86d

Observation 708b250d-3e07-4dbe-92b1-5cffa60e4259 · inbound

Hy-Embodied-0.5-VLA: From Vision-Language-Action Models to a Real-World Robot Learning Stack cites this paper.

Hy-Embodied-0.5-VLA: From Vision-Language-Action Models to a Real-World Robot Learning Stack ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-02T11:31:37.930748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:31:37.930748Z digest=sha256:399abd1395cbc2858919084b3fd1d0b0074176d5f9ebabd5edd99e2be576f15d

Observation 2438e615-e33c-48f1-b7a9-49e4d2257771 · inbound

X-Tokenizer: A Multimodal Action Tokenizer for Vision-Language-Action Pretraining cites this paper.

X-Tokenizer: A Multimodal Action Tokenizer for Vision-Language-Action Pretraining ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-06-30T11:24:37.806023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T11:23:36.941801Z digest=sha256:296e05546b3ad321a6bd6b9abae71a8724b0b153d6a097c7076139cef743019e

Observation 3a412ad3-eade-46ad-b5fd-968b912f5a90 · inbound

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation cites this paper.

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-07-03T21:08:58.816613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T00:49:13.291897Z digest=sha256:cd4fe47af52caf40a5ed46751513f1863c09a837c8d44737173f01e6c2cf3847

Observation 63973cf2-a6f5-4812-8b1c-48c1fa76a1e0 · inbound

Frequency-Aware Flow Matching for Continuous and Consistent Robotic Action Generation cites this paper.

Frequency-Aware Flow Matching for Continuous and Consistent Robotic Action Generation ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-07-04T03:59:34.161509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T17:17:45.015151Z digest=sha256:a735ca24e24cbaa73f1ee6825ac9a829eb37e4084860a083fc2f476eb1afa455

Observation eb877b50-db85-4245-85b8-9ad3d9377be5 · inbound

World Action Models: A Survey cites this paper.

World Action Models: A Survey ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 72

Resolution
verified exact
local_arxiv, observed 2026-07-04T04:09:35.198384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T17:11:12.686936Z digest=sha256:def75f86d28373d478a9cf1fab6b77e217248c0240122fd7586216eb55052b0f

Observation 44c93694-da74-4ef4-b232-b2adf46b9b41 · inbound

NAC: Neural Action Codec for Vision-Language-Action Models cites this paper.

NAC: Neural Action Codec for Vision-Language-Action Models ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-07-04T06:59:37.896084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T13:59:53.484305Z digest=sha256:05fb3c59e0d6c2014dd32eed743766a264c7d28b7304cf08537e2b90654c27d1

Observation 01fd0bd0-88ff-4cf9-90b5-f3651faed181 · inbound

Wh0: Generative World Models as Scalable Sources of Egocentric Human Hand Manipulation Data cites this paper.

Wh0: Generative World Models as Scalable Sources of Egocentric Human Hand Manipulation Data ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-07-04T08:19:44.728525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T11:47:54.951618Z digest=sha256:f69fc0156d6f23f980ab35685b2c932386ffa9b479c63f1a221619c9bcb92afa

Observation a91f0f35-9c1f-477b-9135-315fa0974a0a · inbound

OpenHLM: An Empirical Recipe for Whole-Body Humanoid Loco-Manipulation cites this paper.

OpenHLM: An Empirical Recipe for Whole-Body Humanoid Loco-Manipulation ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-07-04T08:29:41.805860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T11:38:43.339469Z digest=sha256:eb2eeefbe7a52a0eaf372591541bb97229b19851b617e5f38745ae47910127b8

Observation fa4efcb9-cd4a-47f9-9459-190e1604c4fa · inbound

HoloAgent-0: A Unified Embodied Agent Framework with 3D Spatial Memory cites this paper.

HoloAgent-0: A Unified Embodied Agent Framework with 3D Spatial Memory ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 59

Resolution
verified exact
local_arxiv, observed 2026-07-04T10:49:46.507698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T08:27:15.889885Z digest=sha256:d7e12797cc4324bf82bf8942558849b42502181c77d7a1fbe691f6a675e2ad30

Observation a9c9dd9f-2ca2-4d31-a134-ced100b1c873 · inbound

A Watermark for Vision-Language-Action and World Action Models cites this paper.

A Watermark for Vision-Language-Action and World Action Models ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-07-04T11:49:50.333461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T07:41:44.703750Z digest=sha256:b40c37f0e171113e13b314a51dbaf8a3925fcb29d8c83aa13ff7661b7af28be3

Observation 75a30022-1d2e-4224-9a65-922c7b2fdde2 · inbound

LaST-HD: Learning Latent Physical Reasoning from Scalable Human Data for Robot Manipulation cites this paper.

LaST-HD: Learning Latent Physical Reasoning from Scalable Human Data for Robot Manipulation ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-07-04T11:29:50.730912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T07:58:17.225491Z digest=sha256:fe6ff9361a2ea01a035f9bb8cf46ac2966b6a742595a928113684b8623eabd70

Observation e31cb185-2741-43a8-8a42-460c0c23f51f · inbound

TuringViT: Making SOTA Vision Transformers Accessible to All cites this paper.

TuringViT: Making SOTA Vision Transformers Accessible to All ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-07-04T16:29:57.881393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T00:29:41.291832Z digest=sha256:84f3a453a06a49a7bd4e87547e7f2db0f54281df50e0faf61af007b0750b029c

Observation 601dd6f0-2236-4606-8861-b720a6319309 · inbound

TuringViT: Making SOTA Vision Transformers Accessible to All cites this paper.

TuringViT: Making SOTA Vision Transformers Accessible to All ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-06-29T15:03:32.209393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T05:32:26.746776Z digest=sha256:aab609b950149d4ac199c15e93f8e2341a852e2b0f13b910804c23dae08d8ffb

Observation 045a9c6b-55fd-4325-92fe-75a11d2b8744 · inbound

World Value Models for Robotic Manipulation cites this paper.

World Value Models for Robotic Manipulation ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-07-04T17:40:00.329559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-25T23:33:50.854261Z digest=sha256:16b4c8b10387ea924e9e5ad153b2585b21e56ce8d2d145572f8ff037969d231b

Observation 091f9252-0af3-42dd-8fe9-7f94793a7d87 · inbound

InSight: Self-Guided Skill Acquisition via Steerable VLAs cites this paper.

InSight: Self-Guided Skill Acquisition via Steerable VLAs ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-07-04T16:49:58.320870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T00:10:51.721485Z digest=sha256:708cb7ee40121c7083314b3d1d77cd646668853ac6854c06ea6ccb77b9192e2f

Observation 8cedc5d9-bc42-488a-9671-05e061037ed4 · inbound

Chronos: A Physics-Informed Full-History Framework for Non-Markovian Long-Horizon Manipulation cites this paper.

Chronos: A Physics-Informed Full-History Framework for Non-Markovian Long-Horizon Manipulation ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-06-30T15:04:46.821190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T05:12:47.183078Z digest=sha256:35a1c066e543aad07df43e199be390c6c62c17ce5fdc905900255dcea0a08b70

Observation 5d46660f-23cc-4ad7-9c17-53d96a41762d · inbound

Z-1: Efficient Reinforcement Learning for Vision-Language-Action Models cites this paper.

Z-1: Efficient Reinforcement Learning for Vision-Language-Action Models ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-07-01T10:45:43.005556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-07-01T05:05:59.690016Z digest=sha256:2b4fb9e10fdf050e3dc36a85500c3b45e5f2fb83509ce66ad1356557755ece38

Observation 35d9f03d-203d-47b2-87e5-a8e799ff64de · inbound

ABot-M0.5: Unified Mobility-and-Manipulation World Action Model cites this paper.

ABot-M0.5: Unified Mobility-and-Manipulation World Action Model ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 54

Resolution
verified exact
local_arxiv, observed 2026-07-02T14:27:03.150804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-02T14:24:23.187164Z digest=sha256:206fecf66c1d69350f6beeb4d21859a8bad677ee03f02ee4f5717a3fa4c9638b

Observation 9ab43b31-98c4-42dd-842d-00e5ad6305ca · inbound

ABot-M0.5: Unified Mobility-and-Manipulation World Action Model cites this paper.

ABot-M0.5: Unified Mobility-and-Manipulation World Action Model ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 54

Resolution
unresolved
no resolver link, observed 2026-07-12T09:22:08.000379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:22:08.000379Z digest=sha256:0107ad8aa5cf5ca0d4a5bd0c8c77077aaf409c46ef0ebefbfd2f51fa2de4d306

Observation 64bce90d-0fcd-44b0-8b6f-c3092caddef1 · inbound

TACO: TActile World Model as a Self-COrrector forScalable VLA Post-Training cites this paper.

TACO: TActile World Model as a Self-COrrector forScalable VLA Post-Training ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-12T06:41:57.276146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:41:57.276146Z digest=sha256:70b820ce32bc261fa19491d10118e674a3157eef527bed3aceb565856fb3b8f7

Observation 77f5b74a-e87a-4d66-98ab-3129af903266 · inbound

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI cites this paper.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:aa7392dd537c743ac8e4191dbb0f09e5d444637d475b9375c1ab41f781ebd7c2

Observation 53441317-a28a-4ec6-8565-9b94650f7860 · inbound

RoboDojo: A Unified Sim-and-Real Benchmark for Comprehensive Evaluation of Generalist Robot Manipulation Policies cites this paper.

RoboDojo: A Unified Sim-and-Real Benchmark for Comprehensive Evaluation of Generalist Robot Manipulation Policies ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-11T19:13:23.494763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:13:23.494763Z digest=sha256:8842b739b42f215522eae44f6c01a4270f903e0d233d03fb26011bd51d9acd35

Observation 4e8396e2-e721-4662-8c20-fde71e4f42f9 · inbound

WAM-TTT: Steering World-Action Models by Watching Human Play at Test Time cites this paper.

WAM-TTT: Steering World-Action Models by Watching Human Play at Test Time ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-13T06:48:14.554799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T06:48:14.554799Z digest=sha256:8fdd03371b87ca67da84695b95d91418fdf434856fa1406c8a6890bc2485218a

Observation 3702ca00-90f7-46be-84df-25508cdcf39b · inbound

TouchWorld: A Predictive and Reactive Tactile Foundation Model for Dexterous Manipulation cites this paper.

TouchWorld: A Predictive and Reactive Tactile Foundation Model for Dexterous Manipulation ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-07-09T15:56:19.492580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-09T15:54:12.479002Z digest=sha256:2f72106d5fec264579a1ded60e0f85f8ea4ec8bd109f9c042524c3aa9edbbf06

Observation b77aa7ec-a2a0-4dde-86db-7f2dbdbc4045 · inbound

TouchWorld: A Predictive and Reactive Tactile Foundation Model for Dexterous Manipulation cites this paper.

TouchWorld: A Predictive and Reactive Tactile Foundation Model for Dexterous Manipulation ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-07-10T19:47:32.632659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-10T19:42:37.215600Z digest=sha256:dcfe8576679472511a8086d4b872a86cd49ad611cfd43ae9697efa0c09b48781

Observation ff0d25d9-7332-45c1-90f1-7b854959f683 · inbound

APIVOT: Adaptive Planning with Interleaved Vision-Language Thoughts cites this paper.

APIVOT: Adaptive Planning with Interleaved Vision-Language Thoughts ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-07-10T13:37:06.892676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-10T13:33:17.574108Z digest=sha256:e8baf1fa6326d0ce1c3c92660a1901f4fbd009296740e4ff253e2a9c125ad3ad

Observation a25d479e-4054-4ac3-88f9-d0e92f01adb1 · inbound

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio cites this paper.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-07-10T12:47:05.541892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:636dd2143eeb3ab115f12653586723284c28e040f03cfa21c47dfaf89c695793

Observation d279077d-57a1-4350-9989-debe36983d17 · inbound

Native Video-Action Pretraining for Generalizable Robot Control cites this paper.

Native Video-Action Pretraining for Generalizable Robot Control ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 76

Resolution
verified exact
local_arxiv, observed 2026-07-10T04:16:48.646230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-10T04:12:29.153764Z digest=sha256:73ca50c2640b1d23b118a296196053333c99e2fc5b21169647ca19ab33805efd

Observation 620f651e-dbd9-44f5-904a-03fc0f8512ea · inbound

Native Video-Action Pretraining for Generalizable Robot Control cites this paper.

Native Video-Action Pretraining for Generalizable Robot Control ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-02T07:53:40.223709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:53:40.223709Z digest=sha256:fad82f92b7a8e9cdd83aa0f11d53b8f2d06c5e231c9c3f7ac4e333ba0299112f

Observation 56105b3a-da43-42de-8f60-8bab3f4aaa17 · inbound

EgoSteer: A Full-Stack System Towards Steerable Dexterous Manipulation from Egocentric Videos cites this paper.

EgoSteer: A Full-Stack System Towards Steerable Dexterous Manipulation from Egocentric Videos ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-14T17:30:54.988498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T17:30:54.988498Z digest=sha256:e26c5e7d739be6548b2d646c0f91e45eef4b10b48ce18e08c861d8f6e77b1296

Observation 4de6156f-d7a6-413e-ba7e-6642d9778cf0 · inbound

Worlds in One Demo: A Synthetic Data Engine for Learning Open-World Mobile Manipulation cites this paper.

Worlds in One Demo: A Synthetic Data Engine for Learning Open-World Mobile Manipulation ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:59.248578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:06:59.248578Z digest=sha256:82d20eabad6a260efd283e763f08d50e219effbd7073d00821fbeeeac40d292f

Observation ba7bc83f-654d-4bc5-9c77-c9215aa66c3f · inbound

GigaWorld-Policy-0.5: A Faster and Stronger WAM Empowered by AutoResearch cites this paper.

GigaWorld-Policy-0.5: A Faster and Stronger WAM Empowered by AutoResearch ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T03:16:52.982239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:16:52.982239Z digest=sha256:1eeca4d2773e6b2940a2154dfc45a988ee951569b9f45e6aa044d2c813d2aa11

Observation 3a66bdc8-57aa-487b-9a80-58a44eb08f65 · inbound

RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination cites this paper.

RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T03:16:53.054573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:16:53.054573Z digest=sha256:66e4d8b6e5e0433268e07839d5f92067293e2fbc0712aa75df63f83193a2e721

Observation 9e49428c-fea9-4777-b44d-2bf49e7839a9 · inbound

FoMoVLA: Bridging Visual Foresight and Motion Guidance for Vision-Language-Action Models cites this paper.

FoMoVLA: Bridging Visual Foresight and Motion Guidance for Vision-Language-Action Models ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T01:15:51.043766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:15:51.043766Z digest=sha256:d9ff3575ee67b47a5c5bfb4a4b7dccd1921e2b86831266301eab84fb2ef8e81d

Observation f093673d-4b63-4894-9d1d-2b3ffc21433c · inbound

Towards Human-like Physical Intelligence: Lifelong Vision-Language-Action Learning for Robotic Manipulation cites this paper.

Towards Human-like Physical Intelligence: Lifelong Vision-Language-Action Learning for Robotic Manipulation ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-02T00:57:43.044566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:57:43.044566Z digest=sha256:7adaa583c494f18ad9465fe0f89f097bb8195b6baa9603da19d1f3f6bb95fc5a

Observation 220fe4b1-7681-4f30-9224-08474d676fed · inbound

BadWAM: When World-Action Models Dream Right but Act Wrong cites this paper.

BadWAM: When World-Action Models Dream Right but Act Wrong ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T23:54:07.073670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:54:07.073670Z digest=sha256:9139149daa0deb6f50d87497400b0b234300702bb59f8f87a78a86e101712baa

Observation a937599b-d17a-48bd-9544-644636988650 · inbound

Xiaomi-Robotics-1: Scaling Vision-Language-Action Models with over 100K Hours of Real-World Trajectories cites this paper.

Xiaomi-Robotics-1: Scaling Vision-Language-Action Models with over 100K Hours of Real-World Trajectories ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T00:04:01.761066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:04:01.761066Z digest=sha256:75cda9a4e344b3d6a767a97ffbb8b8de285664b3030552894b1d12ef75a02c43

Observation 641b1cfb-5099-4f9d-bc1e-f2486f0fb389 · inbound

RoboInter1.5: A Holistic Intermediate Representation Suite for Embodied World Modeling and Robotic Manipulation cites this paper.

RoboInter1.5: A Holistic Intermediate Representation Suite for Embodied World Modeling and Robotic Manipulation ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-01T14:39:36.522700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T14:39:36.522700Z digest=sha256:78055954c47e906780e7b90064652beeca24d5a50e66047b44de9a648de808f4

Observation 569038f2-9972-435f-bd96-ea0b6b62907f · inbound

WorldScape Policy 2.0: Empowering Steerable World Action Modeling with Reasoning-Augmented Memory cites this paper.

WorldScape Policy 2.0: Empowering Steerable World Action Modeling with Reasoning-Augmented Memory ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T14:13:38.379148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:13:38.379148Z digest=sha256:079e05264da1dec562fb6def8c35a35cef123f6de94007c6406e0cc3116e480a

Observation e77c9d92-d849-4dac-b3c8-70c07648c137 · inbound

Data Pyramid for Embodied Manipulation cites this paper.

Data Pyramid for Embodied Manipulation ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 151

Resolution
unresolved
no resolver link, observed 2026-07-31T06:18:55.616561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:18:55.616561Z digest=sha256:c1f32a32a511e4c0a89fc7ff9335843096bcaf0a8d6a9b2f109f8204492a0bf5

Observation d5132b7a-b8b3-4cb1-958f-bdf3f2af0e3a · inbound

CheckVLA: Execution-Time Verification with Action-Conditioned World Model for Long-Horizon Mobile Manipulation cites this paper.

CheckVLA: Execution-Time Verification with Action-Conditioned World Model for Long-Horizon Mobile Manipulation ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-30T21:05:52.691663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-30T21:05:52.691663Z digest=sha256:969635023f4549463264a39a595951a3a097e6a4e8cc32df5c0338283caa7a1f

Observation 84ff1724-9a32-4d09-9f9e-eecbfad0ae6e · inbound

ST-WAM: Semantic-Temporal World Action Model for Robust Manipulation under Visual Distribution Shifts cites this paper.

ST-WAM: Semantic-Temporal World Action Model for Robust Manipulation under Visual Distribution Shifts ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-03T15:49:34.780788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T15:49:34.780788Z digest=sha256:b1ab03b39b80940f73abb024a793f2f322577b792f0bc810b093010334afc7aa

Observation 83ab3ffb-07f9-4034-bfe5-34ba90136c3c · inbound

WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning cites this paper.

WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T03:34:03.301615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:34:03.301615Z digest=sha256:4b3a53e8ab2056bd60c8ab97f2abac80ccdf5c79602051446495ab22fe53e031

Observation 24d59c5a-fde3-4cda-8689-f5693f3400a9 · inbound

Learning Panorama-Aware VLA for Mobile Manipulation with Whole-Body Teleoperation cites this paper.

Learning Panorama-Aware VLA for Mobile Manipulation with Whole-Body Teleoperation ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T10:39:54.003907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T10:39:54.003907Z digest=sha256:00eaa8121e8e926c9af9aeceab99a39302372ccd5651fda60169a82373f00371

Observation a169bc19-149c-410c-8a59-274aed69d5d5 · inbound

LiLa-WAM: Lightweight Latent Reasoning World-Action Model for Robotic Manipulation cites this paper.

LiLa-WAM: Lightweight Latent Reasoning World-Action Model for Robotic Manipulation ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T14:32:56.296925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:32:56.296925Z digest=sha256:cf3c5e1e5525460ba638271e649ebb7ef818afec3954a7b6ebc9a91dde6452b8