Pith. sign in

Paper Citation Record · LEDGER

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI

As of 23 August 2026, this Paper Citation Record lists 100 of 169 outbound references and 2 inbound Pith citation observations for arXiv:2607.04426.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.04426 v1

Coverage vector

measured 100 of 169 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-11T19:16:57.396710Z

measured 102 of 102 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T15:17:24.978338Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-10T21:18:37.447702Z

Reference resolution

100 of 169 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved100
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 141927ee-73c4-4369-bbef-f41d36144431 · outbound

This paper cites A survey of embodied ai: From simulators to research tasks.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI A survey of embodied ai: From simulators to research tasks

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:96becbcfd474f292e030430fe7eb73d2bf602ca41fc303cd7fbeb70638d05024

Observation 9c1edee8-2682-463f-a42f-df7bb3e171f7 · outbound

This paper cites A Survey on Robotics with Foundation Models: toward Embodied AI.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI A Survey on Robotics with Foundation Models: toward Embodied AI

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:85d36817c35402d6fcce37432d11646c2f8dbc5a00c5ce2fed62479feae76e68

Observation 60a72b50-6c7b-4375-a943-681ffe53f303 · outbound

This paper cites RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:fcae67b1186a82dc89b67b6bee34c27df1e9affcbd07e8924a5978c39c54a2b5

Observation 58165d18-7887-405b-baa3-2afbd455deb6 · outbound

This paper cites InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:5ca0cf3af3d439f74cb7f04877aba0d0fbd20d4614ad5a075f5018cb0e580e4b

Observation 28cf012f-3d06-4319-b46c-5ef29d0e95ec · outbound

This paper cites Qwen3-VL Technical Report.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Qwen3-VL Technical Report

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:b4211b19052004ff8769265214b1954131fdbee060ca5ff7f316968496fd94d4

Observation 5533c13a-09e7-414d-89d0-1ff840ab15ab · outbound

This paper cites Qwen3 Technical Report.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Qwen3 Technical Report

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:0c0acbb0ad98c87e75fa0f71347578a0deb35959015fb38b3387fdf4686c3b3a

Observation 2ef9bf37-7c1e-41a8-bc78-7a46816425ac · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Gemini: A Family of Highly Capable Multimodal Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:9c92827573aa878ae126ca38e3542419d4396a058a2024ec8e62a73e06a6ddde

Observation bda8b1a1-daa3-4a73-9f9d-525f9dd8159b · outbound

This paper cites Qwen-RobotNav Technical Report: A Scalable Navigation Model Designed for an Agentic Navigation System.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Qwen-RobotNav Technical Report: A Scalable Navigation Model Designed for an Agentic Navigation System

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:5108832b3595eb1b43debe2f5e5563dce8ae86e524e96a6a58a66f7550eda76f

Observation 6ba5691b-b952-4833-bc33-a6304beed28a · outbound

This paper cites Embodied navigation foundation model.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Embodied navigation foundation model

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:24560dc216da4e38f05b1a843ad2f0256802ec91c97b546fda95220c852ec134

Observation 53966804-dddf-43d0-bed7-fdefd305dd56 · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:a6514a340bcb70c93a951fb648ff2ee8fb208a1e44ee2a182258b02951e62b87

Observation a9fa1942-49dc-429a-bcd5-2b4629eb556d · outbound

This paper cites World Action Models are Zero-shot Policies.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI World Action Models are Zero-shot Policies

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:083ba0e3547c5908a42a605441cd27c421660693637c0f0e8c614e8e0c9ad5e4

Observation 1c4aa590-11c8-4afe-b1f8-4c74f3172d39 · outbound

This paper cites RoboReward: General-purpose vision-language reward models for robotics.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI RoboReward: General-purpose vision-language reward models for robotics

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:b82b3341ecc8248a900ab8037d5378e02ac921c1cf66cd63e24537a4f6dc4721

Observation 59ebd359-cb5b-44bd-a2ae-eacdf6469054 · outbound

This paper cites Robometer: Scaling General-Purpose Robotic Reward Models via Trajectory Comparisons.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Robometer: Scaling General-Purpose Robotic Reward Models via Trajectory Comparisons

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:c5080373ac81c4ffe5d16b87c3a521001f81e71db030d47edf0e7f87cd9a590d

Observation bd073da7-9d4e-4522-bdcb-dcf59df884ce · outbound

This paper cites an unresolved cited work.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:2814505ede36b696704d395fd72dba7b335c4f8ede5e08b7d901a48fce05c939

Observation 504a61aa-e041-4261-bd8e-a3b130c7d4ec · outbound

This paper cites an unresolved cited work.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:d0f8baa1d0256f2977511ac428d9580f696d4f2f353fd8fcb40490ba9b98d240

Observation 4059a9bd-021a-41ec-b698-9c681b8e03ad · outbound

This paper cites an unresolved cited work.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:7a26d642072de46794d8cd9a4257bd7e1afa58da23ad7770c2d73893c902c0fe

Observation a11f0199-c5a2-48cd-8264-5b744e947274 · outbound

This paper cites $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:efc807819ee5dc2bd5e24bb38defb32a6ffd2f624fb1ad473b26d4f799d99727

Observation d8024768-a443-4f5e-8af0-e68f88e5c6cd · outbound

This paper cites Interleave-vla: Enhancing robot manipulation with interleaved image-text instructions, 2025.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Interleave-vla: Enhancing robot manipulation with interleaved image-text instructions, 2025

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:fb2d97755a14be4a27264056a5a39b80c7294ec2a267b4ec18d8204d0ce6fcbc

Observation 9b4a8d58-da77-4c2d-9b77-64e09f539bfa · outbound

This paper cites Do as i can, not as i say: Grounding language in robotic affordances.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Do as i can, not as i say: Grounding language in robotic affordances

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:5bee04eeef2ff5b6442b8e8a35ac28ccad45f69cda1c68e4690a0464a08a72e9

Observation f70ceaa5-d7ca-4acb-8659-72802253413f · outbound

This paper cites PaLM-E: An embodied multimodal language model.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI PaLM-E: An embodied multimodal language model

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:5ea27de9e3444b7b9dad2f9757bb550937c544b103251bbf73568ae32344926d

Observation b058d83b-7b2b-4987-b783-11ed6293f832 · outbound

This paper cites Code as policies: Language model programs for embodied control.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Code as policies: Language model programs for embodied control

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:7b9bf7a4542f7c74e28442aa89a183da8ae9249ac1a3f7a65289da1a054963a2

Observation 648bb15a-ea49-4431-8936-94d9c26d43f5 · outbound

This paper cites Voxposer: Composable 3d value maps for robotic manipulation with language models.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Voxposer: Composable 3d value maps for robotic manipulation with language models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:28fc6a59c41c542553a68d590877b1d8da10fc857e450de406e05f039833230a

Observation 2cc9a582-2bab-45f3-9454-e90d5452ef79 · outbound

This paper cites RoboCodeX: Multimodal code generation for robotic behavior synthesis.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI RoboCodeX: Multimodal code generation for robotic behavior synthesis

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:decc1a94f5cff9cc02232e0ec8fa165b287e4a80e98556fb17921aa2a9e6ac4a

Observation 5e62c089-d830-47c1-8049-ff8edd7b22ab · outbound

This paper cites RoboAgent: Chaining Basic Capabilities for Embodied Task Planning.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI RoboAgent: Chaining Basic Capabilities for Embodied Task Planning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:b257561bc2e6d325516274b5e3bbec887c726c60aa496eff3665d13fdfcee243

Observation 4a918928-a2f2-482a-b288-595ea9092593 · outbound

This paper cites GR00T N1: An Open Foundation Model for Generalist Humanoid Robots.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI GR00T N1: An Open Foundation Model for Generalist Humanoid Robots

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:c82c3024c9684222531b2b39063370e73fd3a24b270339c2b2f7ea0eebc9cd40

Observation ada86e16-2b55-4088-a0a5-0fe7598a8668 · outbound

This paper cites Gr00t n1.5: An improved open foundation model for generalist humanoid robots.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Gr00t n1.5: An improved open foundation model for generalist humanoid robots

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:d6a52f3f8b6f20d73605b96143479c9531f57ba6118a2934aac73159aa74574d

Observation 9766699a-0bce-4b34-b608-09ec4ea3ec5a · outbound

This paper cites Qwen-VLA: Unifying Vision-Language-Action Modeling across Tasks, Environments, and Robot Embodiments.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Qwen-VLA: Unifying Vision-Language-Action Modeling across Tasks, Environments, and Robot Embodiments

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:82530bedf3c0048a8b641092188c15015f69b4b8c055663162aaf49d0a929b1c

Observation d342d7b5-fe9e-46dd-997f-5ec135fef29f · outbound

This paper cites Qwen-RobotWorld Technical Report: Unifying Embodied World Modeling through Language-Conditioned Video Generation.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Qwen-RobotWorld Technical Report: Unifying Embodied World Modeling through Language-Conditioned Video Generation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:76b7b28bbd4b6b112a95e384da1da0b30b862dfe28c4f4ed64f16779e399f053

Observation 56ec2478-6d79-44df-9007-85ba28ba5160 · outbound

This paper cites ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:b517e651256b875f4eb72b1b823cc55a6df624aef034a2d185427ae0ee9f3fcf

Observation 8bb52d65-6ea9-4646-81c9-beee96f73e00 · outbound

This paper cites Rynnbrain: Open embodied foundation models.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Rynnbrain: Open embodied foundation models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:9d6a12c30b65d738b4e5e5d5316e73a76469e5b3d9f289a56f528eda9e21fc7d

Observation 2e3ebbcd-57e3-4dbe-8636-e49b96634871 · outbound

This paper cites Cosmos 3: Omnimodal World Models for Physical AI.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Cosmos 3: Omnimodal World Models for Physical AI

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:cd71cb7df0071626d41329d443d557cea788bfdc81110d2bbdbfbf5d13b38121

Observation ff92efcc-6e25-4e7c-a4aa-2ecb99a70e42 · outbound

This paper cites Ace-brain-0: Spatial intelligence as a shared scaffold for universal embodiments.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Ace-brain-0: Spatial intelligence as a shared scaffold for universal embodiments

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:ab7bf31d0706ac9dfad50bad214e142208f11d7ab17971f0b1560befac8da964

Observation d107ad8e-e58f-4504-99cf-7b901ab047d2 · outbound

This paper cites Pelican-Unify 1.0: A Unified Embodied Intelligence Model for Understanding, Reasoning, Imagination and Action.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Pelican-Unify 1.0: A Unified Embodied Intelligence Model for Understanding, Reasoning, Imagination and Action

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:5b342c99fa5da7dace84e57376447d6e2ffcbbdc90bc62da901c5aa7bc921054

Observation 52869e84-c7ed-44fc-8db0-d9ee91b7629b · outbound

This paper cites HY-Embodied-0.5: Embodied Foundation Models for Real-World Agents.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI HY-Embodied-0.5: Embodied Foundation Models for Real-World Agents

Reference 34

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:ea8d0d3f903c051b5b8fc7c29f4f73f7d79d9ff8789c99d879d878ac4488c2ed

Observation 77f5b74a-e87a-4d66-98ab-3129af903266 · outbound

This paper cites ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:5cece9c54340d604a18721d1320fb7f334880b67b6ade2d6cec20c05bfb79e03

Observation ceda9816-273d-4627-b6d5-b6722cf28e35 · outbound

This paper cites $\pi^{*}_{0.6}$: a VLA That Learns From Experience.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI $\pi^{*}_{0.6}$: a VLA That Learns From Experience

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:e88b2b998170e0a0f0de9670e4bfcaf353dd12e494e64e70ad696bac7247f91f

Observation 40513af3-cfa8-4723-8f19-3c1572883530 · outbound

This paper cites Embodied-R1.5: Evolving Physical Intelligence via Embodied Foundation Models.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Embodied-R1.5: Evolving Physical Intelligence via Embodied Foundation Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:f575299cd925e3d2224e705d1406dd7152fea38b785f44af9798c2fdd7bddbd8

Observation 7093daec-0540-48bc-a710-2b00f05698e1 · outbound

This paper cites MolmoAct2: Action Reasoning Models for Real-world Deployment.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI MolmoAct2: Action Reasoning Models for Real-world Deployment

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:0925e91baad0e997201ad975369e06ffaa8970c91d1c49850199bce726fcf50b

Observation a1a78ed6-52c2-4e16-9127-fcb0dc4d66a5 · outbound

This paper cites Qwen-RobotManip Technical Report: Alignment Unlocks Scale for Robotic Manipulation Foundation Models.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Qwen-RobotManip Technical Report: Alignment Unlocks Scale for Robotic Manipulation Foundation Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:44968fde14b78a2b210b29f467d4de1c33515dd8e36a92236ec3d352b7562cc6

Observation 5af40899-3c4d-49ba-83e9-a1d7e7a60108 · outbound

This paper cites Abot-n0: Technical report on the vla foundation model for versatile embodied navigation.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Abot-n0: Technical report on the vla foundation model for versatile embodied navigation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:5f911fb28af379ce805c95f77a96c5f914117dc0fcea7d52a8283af10188fbae

Observation faaff22f-dfa0-47c9-8ab1-15d592b83ca9 · outbound

This paper cites Gemini Robotics-ER 1.6: Powering real-world robotics tasks through enhanced em- bodied reasoning.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Gemini Robotics-ER 1.6: Powering real-world robotics tasks through enhanced em- bodied reasoning

Reference 41

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:56840965d5207cc4a1a74af8821bab43d10ddb0ab096f81d0f5b0883f89eea33

Observation fcf9ce38-f87e-4df2-a044-3f947ac31eaf · outbound

This paper cites Introducing helix 02: Full-body autonomy, 2026.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Introducing helix 02: Full-body autonomy, 2026

Reference 42

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:fb532e4c74c11167382107bea44a12268294d8d409592ddcbc27de3b3eaa1587

Observation e8e959ef-3a8e-47aa-9716-2f0f78c9795c · outbound

This paper cites Model merging in llms, mllms, and beyond: Methods, theories, applications, and opportunities.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Model merging in llms, mllms, and beyond: Methods, theories, applications, and opportunities

Reference 43

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:d80cbe045e78c84cf65ee742d4835f73125db1d0e15c3f60026457e615e25b86

Observation c77a7c9f-8fed-43b9-914a-4b149fa1fc46 · outbound

This paper cites Qwen2.5-VL Technical Report.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Qwen2.5-VL Technical Report

Reference 44

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:fd59c693a42a9c4104d071a90af20029ea05b6ecd35bd5ee4688156857dddc78

Observation c260d960-5da1-4142-80a2-d5343894d2a4 · outbound

This paper cites Improved baselines with visual instruction tuning.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Improved baselines with visual instruction tuning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:c9df1e1494550df0cc255451acb30811a762e54225a28f0b7e45dd4973f81aea

Observation f058b581-ad60-4015-b802-a7787c392e00 · outbound

This paper cites Gpt-4o system card.https://openai.com/index/gpt-4o-system-card/, 2025.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Gpt-4o system card.https://openai.com/index/gpt-4o-system-card/, 2025

Reference 46

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:c6504f08f9494ee54c65c859977e2793f707607a8944e3b97859dbbcf976234c

Observation c7c7f3ab-af64-4758-81aa-258786bb90f7 · outbound

This paper cites Claude sonnet 4.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Claude sonnet 4

Reference 47

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:61f9e7496912f3d08e45ee930f621a38772519c11ed82fd76550702c3a242631

Observation b8cdf4f8-057f-405f-a32c-8e1666e9c794 · outbound

This paper cites SpatialVLM: Endowing vision-language models with spatial reasoning capabilities.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI SpatialVLM: Endowing vision-language models with spatial reasoning capabilities

Reference 48

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:0541f13e3cbb27f8cfc42465ef6b51742be26451db78a86a177df762ffdfef31

Observation 151f5f20-ec80-4fa0-a285-f1b30583eee6 · outbound

This paper cites RoboPoint: A vision-language model for spatial affordance prediction for robotics.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI RoboPoint: A vision-language model for spatial affordance prediction for robotics

Reference 49

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:1cc02b9dcd059e08cae27f5b4172441f7d6981eba74ddfd7c2efeeeeda819742

Observation 1932c905-de85-4bc3-bd57-fff28e24d2d1 · outbound

This paper cites RoboRefer: Towards spatial referring with reasoning in vision-language models for robotics.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI RoboRefer: Towards spatial referring with reasoning in vision-language models for robotics

Reference 50

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:96196831fa247dcc5105f52ffef511d7115efe4a26a5a8a55e92bd169948e161

Observation 74304a63-58e8-45f6-8b5c-011c187abc7a · outbound

This paper cites Robobrain: A unified brain model for robotic manipulation from abstract to concrete.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Robobrain: A unified brain model for robotic manipulation from abstract to concrete

Reference 51

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:c5d425e183f2acc6638c1afeae57f11436f1b42dc92db12f113377f816e4e56e

Observation ee474def-3ff1-4243-ac9b-1e9352afacc3 · outbound

This paper cites Robobrain 2.5: Depth in sight, time in mind.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Robobrain 2.5: Depth in sight, time in mind

Reference 52

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:7fe500e9a3d74b1fd088fc8e2f099f84fa4f50b6ec593af7ccd390bc210e6533

Observation ed0ca22f-7515-4283-99ad-b1a05dfdc1ec · outbound

This paper cites MiMo-Embodied: X-Embodied Foundation Model Technical Report.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI MiMo-Embodied: X-Embodied Foundation Model Technical Report

Reference 53

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:5ccba74000d47db1ef6e5a11bb4e2940cbe46421555a31a1b4ca266e4be6c98a

Observation 5ca3a554-d397-484d-9748-597a412c8479 · outbound

This paper cites Visual Embodied Brain: Let Multimodal Large Language Models See, Think, and Control in Spaces.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Visual Embodied Brain: Let Multimodal Large Language Models See, Think, and Control in Spaces

Reference 54

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:1480d62a1d74a5f5e38d8e2be0f28cd829a036e6d539bc3bc558d98cf0c8997d

Observation eddaea49-160f-4f46-b2b9-b8d3487ef1b8 · outbound

This paper cites RoboBrain 2.0 Technical Report.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI RoboBrain 2.0 Technical Report

Reference 55

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:e264d136479def72cd0ea52dcbf497272a5ee3ff14bf73217228731f4ab98ce6

Observation d583d91a-6417-459d-9fa6-2ebbc8e914b1 · outbound

This paper cites Vlaser: Vision-language-action model with synergistic embodied reasoning.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Vlaser: Vision-language-action model with synergistic embodied reasoning

Reference 56

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:0431ca47be21cedd55cab310cdcaab56c9e4b12589e4a1f4b87802cafb86bdf6

Observation 54080e52-ce8c-4bf5-8b82-f3fd26ad96ca · outbound

This paper cites Pelican-vl 1.0: A foundation brain model for embodied intelligence.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Pelican-vl 1.0: A foundation brain model for embodied intelligence

Reference 57

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:af18457eda352668cd002d43b63b8e9f5774731f6ca88c1c4a49d77845ce5dfc

Observation 1e99e69a-0033-4827-92e4-02b8173ea7d9 · outbound

This paper cites RT-1: Robotics transformer for real-world control at scale.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI RT-1: Robotics transformer for real-world control at scale

Reference 58

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:d37414f4b6af9daf0e74b4b8f4a4264188d65269a6cbaf95328a54b44d3e0aa8

Observation 04acb0ba-45d5-4c32-9959-c7cc41d7fe34 · outbound

This paper cites RT-2: Vision-language-action models transfer web knowledge to robotic control.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI RT-2: Vision-language-action models transfer web knowledge to robotic control

Reference 59

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:8c5fae3e17543bbb4f4bde278fb9db1ae6697617b01ad1bc17bf3cb36c95c131

Observation 9f8adba5-4552-4e8f-ab10-8e7466f112e5 · outbound

This paper cites Octo: An open-source generalist robot policy.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Octo: An open-source generalist robot policy

Reference 60

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:a29faf666e9a96e5323fa76b0be54c09dc235721d058098ad65e0c99eec8ecce

Observation 90fc8551-ae27-4c8b-9644-2bb77415fc4c · outbound

This paper cites OpenVLA: An open-source vision-language-action model.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI OpenVLA: An open-source vision-language-action model

Reference 61

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:f723ff22597d73007d61582ce6e058c8d7095b11aeeef53a0849c6e036d1d186

Observation 8ae275de-1280-41ce-b6aa-e200627afc7c · outbound

This paper cites Open X-Embodiment: Robotic learning datasets and RT-X models.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Open X-Embodiment: Robotic learning datasets and RT-X models

Reference 62

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:de3f8655a7d188b3a8149ffdbc82d8011daa94f450fb059de3a72cbe4993863c

Observation d418b3a4-033b-43db-bbc6-a9b3be2dc434 · outbound

This paper cites RT-H: Action Hierarchies Using Language.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI RT-H: Action Hierarchies Using Language

Reference 63

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:a84d8c089f600143246223cf31b72fd992762a91936da436e187522f9057130c

Observation 4a7f19e4-3639-49ab-bc9d-9727accbdbb9 · outbound

This paper cites Eo-1: Interleaved vision-text- action pretraining for general robot control.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Eo-1: Interleaved vision-text- action pretraining for general robot control

Reference 64

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:0a9391f7cc419655293121ce9a7c99c9001da46090fb9dc9897ba4fe631ccc69

Observation bd9eebd2-127e-4c50-a7ef-6a4a573ee6a4 · outbound

This paper cites CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation

Reference 65

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:670d8b61dcaddfba6576a1c2109946b0fece0e2c5177296719ddae6ea0ed2ec9

Observation 8f03d772-fc24-432a-adbe-9a25ee457fe9 · outbound

This paper cites GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation

Reference 66

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:e18bfddf22c8c3db5008f791c931db1834711162d8c54c6409452959a7ecea21

Observation 60b77533-9dca-4462-84ea-b589a9826d6f · outbound

This paper cites GR-3 Technical Report.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI GR-3 Technical Report

Reference 67

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:9490ba21789afe43296deb75554e7c0ea83cec1c1d86257b7469e64e63ffbb83

Observation 64cc7f85-a32d-412a-8536-7999f23524a6 · outbound

This paper cites Gaze-Regularized Vision-Language-Action Models for Robotic Manipulation.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Gaze-Regularized Vision-Language-Action Models for Robotic Manipulation

Reference 68

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:892dc5a55226056785ba48657df4216abd0bb4c5a05a7227aa2e12a12d7efdf1

Observation 00cda2d9-a497-42fa-bc7c-0f2268f891c9 · outbound

This paper cites Vla-jepa: Enhancing vision-language-action model with latent world model, 2026.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Vla-jepa: Enhancing vision-language-action model with latent world model, 2026

Reference 69

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:ef0eb3895e1422453061b9ebcd6f7b9e18e4d48b38eef2e34e6e3d269520b883

Observation 28fb3364-9f6d-4b9c-bb42-b091c7ca314a · outbound

This paper cites Cosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planning.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Cosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planning

Reference 70

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:9e943d839e2417d5fa66e92252a807812b06b8c39e526b5c1f6c3b4b445db00f

Observation e10d4955-22cc-4d6d-8715-a21740fe01b3 · outbound

This paper cites Gigaworld-policy: An efficient action-centered world–action model.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Gigaworld-policy: An efficient action-centered world–action model

Reference 71

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:1eec368c17b9ca89303ededbe5f3b1bb3f20d867b61a5eeb3ad7afee67def95f

Observation 06d68dd9-aab4-4a75-ac01-3b64f32c5d02 · outbound

This paper cites Unified 4D World Action Modeling from Video Priors with Asynchronous Denoising.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Unified 4D World Action Modeling from Video Priors with Asynchronous Denoising

Reference 72

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:db593133017e80557686a10ac0c94b4b7269b43fd5e02e7833ca2a649fe19c8e

Observation 7f27a67c-4155-423a-ad86-593274337c6b · outbound

This paper cites Fast-WAM: Do World Action Models Need Test-time Future Imagination?.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Fast-WAM: Do World Action Models Need Test-time Future Imagination?

Reference 73

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:c02a42bd710832da04f42c7bdcf61a2c11f04daf515ac0c5628441ebda608a31

Observation dc0dd1f3-11d3-4dd1-a5cf-5e46182382d5 · outbound

This paper cites Dreamvla: a vision-language-action model dreamed with comprehensive world knowledge.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Dreamvla: a vision-language-action model dreamed with comprehensive world knowledge

Reference 74

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:d264cdf6ca12006b3d94440265744bd0db8ec10f71c5d208403d80dd1963a2e7

Observation 44eefec1-bbc7-475f-ba4b-1f4a4598dea3 · outbound

This paper cites Being-H0.7: A Latent World-Action Model from Egocentric Videos.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Being-H0.7: A Latent World-Action Model from Egocentric Videos

Reference 75

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:628de606f0828711d8cceb87b9a2e1c2d6e29662926133fda373c97667ae86a8

Observation a247e9bc-0198-4fa9-bb9f-c4e5dc329c98 · outbound

This paper cites RynnVLA-002: A Unified Vision-Language-Action and World Model.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI RynnVLA-002: A Unified Vision-Language-Action and World Model

Reference 76

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:618b156e44c9069b3125e3833ed29d1e6873f2d2f45110f0535ce317ba1f21a0

Observation 17b0e8c8-93d0-4fd3-81e9-bb74d27b390a · outbound

This paper cites ABot-M0.5: Unified Mobility-and-Manipulation World Action Model.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI ABot-M0.5: Unified Mobility-and-Manipulation World Action Model

Reference 77

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:1ab0c08a6fbf60cf0a214fd87ba7d289d43f2cb59162609b3a1bdcdb4da18967

Observation 864dbc8a-3285-493e-9047-ba9db8359042 · outbound

This paper cites Vision-and-language navigation: Interpreting visually-grounded navigation instructions in real environments.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Vision-and-language navigation: Interpreting visually-grounded navigation instructions in real environments

Reference 78

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:755cc9eb0ed8109a6c9626b01dcf23dcabc3c0451178bcfb01079f173231bd8f

Observation 9afdb9af-1361-45b3-874b-699a4d12c1c2 · outbound

This paper cites Think global, act local: Dual-scale graph transformer for vision-and-language navigation.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Think global, act local: Dual-scale graph transformer for vision-and-language navigation

Reference 79

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:f16b8dd35efeb157f5907bcb770bf24cb842a8432bd3e91d4e0add34f728d46f

Observation 962e6fd2-29f4-4a41-83a8-5a31522491e1 · outbound

This paper cites Beyond the nav-graph: Vision- and-language navigation in continuous environments.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Beyond the nav-graph: Vision- and-language navigation in continuous environments

Reference 80

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:3e89dcb64a946a32f37c305796e6f2d26b54f663cd32bf20e55293665720fbcf

Observation c34dd25f-b779-4a85-b762-077e359d9d4e · outbound

This paper cites Vision-and-language navigation with foundation models: A survey.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Vision-and-language navigation with foundation models: A survey

Reference 81

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:a92b5fec720fa2b5837d61dfa87cb0a266f905e3bbbe47f7026619d5b6417169

Observation 3555cc1c-fdbb-4dee-bd49-a9e4733fe0df · outbound

This paper cites Navgpt: Explicit reasoning in vision-and-language navigation with large language models.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Navgpt: Explicit reasoning in vision-and-language navigation with large language models

Reference 82

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:855b78a240b98c25a763d3f12c2a0bba84988367e4878f180e02fcaa5955cfc8

Observation 96c27ba5-01f0-46d9-8c3d-54894c5a31b1 · outbound

This paper cites NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation

Reference 83

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:0db11d9640a901261aa35c2db744db33009dc720ee480c96b7ac7d17dca1a080

Observation 60e3694c-c7e2-4522-8943-3ba9f774cccf · outbound

This paper cites NaVILA: Legged Robot Vision-Language-Action Model for Navigation.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI NaVILA: Legged Robot Vision-Language-Action Model for Navigation

Reference 84

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:5c63e7414a13a4bebfa6dcfc6c000711f78ff5702586b437406c7910967161a2

Observation 761b0e6b-5041-4c56-a8a9-cded259122ca · outbound

This paper cites Uni-NaVid: A Video-based Vision-Language-Action Model for Unifying Embodied Navigation Tasks.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Uni-NaVid: A Video-based Vision-Language-Action Model for Unifying Embodied Navigation Tasks

Reference 85

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:5935ce5ce567132f09398b005ca9d0307fb3f012ab351464cb5337aff719ea92

Observation eb84062f-aa1c-4a85-998a-0afa452ba9d4 · outbound

This paper cites StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 86

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:1b14ae7c3aecee2682889869f8ab0b126f4eb7dbd4e0f336810b7c6f91f6c823

Observation 3dd353ec-0325-4b1e-a3be-8dbbf91a841d · outbound

This paper cites Octonav: Towards generalist embodied navigation.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Octonav: Towards generalist embodied navigation

Reference 87

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:a30763d6dbf739de736959954670fab4daee79caf715f60fc7ece61095aef3e3

Observation 4b54c8f8-097f-43fa-8587-5fb801c461e4 · outbound

This paper cites Agentvln: Towards agentic vision-and-language navigation.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Agentvln: Towards agentic vision-and-language navigation

Reference 88

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:306dc4f812a28623e3655deb79b96950fe7b3c44801c284837e69d58d2a86926

Observation 56c67dc1-15b3-4d69-9282-3dd02bd454cf · outbound

This paper cites Learning goal-oriented language-guided navigation with self-improving demonstrations at scale.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Learning goal-oriented language-guided navigation with self-improving demonstrations at scale

Reference 89

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:e098e6c8c41755cada9e0c38bd4638d6a5405bac8577853905d7f4ffa266d194

Observation b095cb67-ab07-492c-a945-66d2ae287efe · outbound

This paper cites Endowing embodied agents with spatial reasoning capabilities for vision-and-language navigation.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Endowing embodied agents with spatial reasoning capabilities for vision-and-language navigation

Reference 90

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:e081d4caa6c1c196ed3fb12ec74d6d649c70dfdcb48493e5dc5d8d6f9e378ef7

Observation 78ec595b-d2f1-43c8-92c3-d324399c5d5c · outbound

This paper cites Modernising Reinforcement Learning-Based Navigation for Embodied Semantic Scene Graph Generation.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Modernising Reinforcement Learning-Based Navigation for Embodied Semantic Scene Graph Generation

Reference 91

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:c6027eebb62b09b89c4f2c8c1a7bf429c620dc4cfa9d37d5d65b8cea683ea431

Observation ef660308-8d90-414f-a3a0-d47002b0ef05 · outbound

This paper cites Sontakke, Jesse Zhang, S´ ebastien M.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Sontakke, Jesse Zhang, S´ ebastien M

Reference 92

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:84cc8d4dda1c49106d5202d10fc6e6b27cb10c8fdab9e57e55180a212e3eb074

Observation 5210083a-0d52-47c4-8368-140dff9f04df · outbound

This paper cites VIP: Towards universal visual reward and representation via value-implicit pre-training.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI VIP: Towards universal visual reward and representation via value-implicit pre-training

Reference 93

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:b5074fe73abae1eedd8cf90107a4d5adf55dfd2f966b63e85364ad3b93b565ab

Observation 0da25528-39c7-47c1-8aec-daaa5b5ee79f · outbound

This paper cites Vision language models are in-context value learners.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Vision language models are in-context value learners

Reference 94

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:a6b964a43e97d1a3c604f3493ff025674492379eb4e3fe008db1ddf25e1bfb37

Observation ffbd59dc-0b7b-4f66-98bd-2681944d382f · outbound

This paper cites Vision-language models are zero-shot reward models for reinforcement learning.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Vision-language models are zero-shot reward models for reinforcement learning

Reference 95

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:17ea3200392f42ca22de5dee0c677e4db584d80ad26626ac14eba3c8705f5a90

Observation 98b95458-49aa-49e0-901c-a1c5c83564b1 · outbound

This paper cites A Survey on Explainable Reinforcement Learning: Concepts, Algorithms, Challenges.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI A Survey on Explainable Reinforcement Learning: Concepts, Algorithms, Challenges

Reference 96

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:84f65ca5c4db0538bb2418a1f4aca92818d7d6e1206b901cbd6627de4303163a

Observation f6b3acb4-899d-4215-be04-9d528e4b7861 · outbound

This paper cites LIV: Language-image representations and rewards for robotic control.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI LIV: Language-image representations and rewards for robotic control

Reference 97

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:6eef95847acfde26de231610b299d429997bc1bcf2c1da45f992cf10e95b27f9

Observation de4c356b-7ed9-4a73-838d-4f50e6e14b46 · outbound

This paper cites Rank2Reward: Learning shaped reward functions from passive video.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Rank2Reward: Learning shaped reward functions from passive video

Reference 98

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:6dbfe5c8b176983a97d92cfac9804f8d6d771938488f01cb12fc0a52f9c83bc6

Observation 0e60e437-ce27-4d9d-be2a-21ebdd005052 · outbound

This paper cites Lim, Jesse Thomason, Erdem Biyik, and Jesse Zhang.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI Lim, Jesse Thomason, Erdem Biyik, and Jesse Zhang

Reference 99

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:8acb465b3625d30276b42053e7e7a466cfc53fd0674725f011a9203e53e15c7e

Observation b89c608b-c6f5-45a1-88db-14a7e3f1d15d · outbound

This paper cites SARM: Stage-aware reward modeling for long horizon robot manipulation.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI SARM: Stage-aware reward modeling for long horizon robot manipulation

Reference 100

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:0c4d2d5562075a1700c04a6efa56964806e89a28d4fdcd7c9e7b5427b3541378

Pith citing papers

Observation 552d0570-b757-4edb-867b-e7857617e555 · inbound

OC-VLA++: Monocular Geometry-Guided Cross-View Consistency for Viewpoint-Robust Robotic Manipulation cites this paper.

OC-VLA++: Monocular Geometry-Guided Cross-View Consistency for Viewpoint-Robust Robotic Manipulation ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T15:17:24.978338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:17:24.978338Z digest=sha256:bf2c50c15595d38f8aa72d293e02334741e02e23b5383ff3e1522ae2a0e1f4c8

Observation 7f6474f4-f9c4-455f-9091-189374bf72e2 · inbound

Capek 0.5: An Execution-Centric Vision-Language Model for Embodied Intelligence cites this paper.

Capek 0.5: An Execution-Centric Vision-Language Model for Embodied Intelligence ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-10T21:18:37.517943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-10T21:18:33.164594Z digest=sha256:7c9d1896efb2cdfac57178d3045cf139963ace4a50aa1a33e21262abc475944a