Pith. sign in

Paper Citation Record · LEDGER

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents

As of 21 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 15 inbound Pith citation observations for arXiv:2505.23450.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.23450 v2

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:49:35.893945Z

measured 56 of 56 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T04:27:31.616619Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-08T12:04:50.484951Z

Reference resolution

41 of 41 outbound references displayed

  • verified exact0
  • verified fuzzy11
  • unresolved30
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 07b3175a-981f-4d7e-a52e-2d6e42b38bd8 · outbound

This paper cites Qwen2.5-VL Technical Report.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents Qwen2.5-VL Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T12:49:31.647116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:49:31.647116Z digest=sha256:5704518b9bb17f05c8fef135c519d731e16bdcb6a0aeed2b56b58d5d21eee2c4

Observation 37f4e001-5881-4f48-9366-1e7cfdf4aabb · outbound

This paper cites RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:49:31.708324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:49:31.708324Z digest=sha256:78180f6e547b8c39cb0ecbce04c4eea4e0cbc9a2d3f595da897d8003ce5f778f

Observation fb419ca2-0df5-4ead-af57-62c7f4432c95 · outbound

This paper cites Do as i can, not as i say: Grounding language in robotic affordances.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents Do as i can, not as i say: Grounding language in robotic affordances

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:49:38.421389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:49:31.766548Z digest=sha256:b722d95005e8a83fbdd3f8fb7cf5bb52d0e5533739897150ec82dd8675e9356d

Observation 7839d018-0eab-4cfd-b081-89523ba87169 · outbound

This paper cites Scar: Refining skill chaining for long-horizon robotic manipulation via dual regularization.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents Scar: Refining skill chaining for long-horizon robotic manipulation via dual regularization

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:49:38.262669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:49:31.826481Z digest=sha256:e91bd7ef1e84db5a54369c8c0f83e7c52f26dfecc0c6ee6b20ea91fb0836ea49

Observation f0dc7b5e-0f89-464f-b3b6-bf3684a46dfb · outbound

This paper cites Diffusion policy: Visuomotor policy learning via action diffusion.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents Diffusion policy: Visuomotor policy learning via action diffusion

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T12:49:31.901414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:49:31.901414Z digest=sha256:9497791e5d2a908fe70d579e3b35eb2549080565fd65a88cb872caf9e853a780

Observation 868b6e62-1111-499d-aa96-81f401789321 · outbound

This paper cites Palm-e: An embodied multimodal language model.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents Palm-e: An embodied multimodal language model

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T12:49:31.982629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:49:31.982629Z digest=sha256:5beaf0406cfe5bd6c97e10abd99babf0970ee27b1fb497a1ec905c315ac1ef11

Observation 1af3a251-9a8a-4eb6-a97d-cfd41ff38f99 · outbound

This paper cites Reflective Planning: Vision-Language Models for Multi-Stage Long-Horizon Robotic Manipulation.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents Reflective Planning: Vision-Language Models for Multi-Stage Long-Horizon Robotic Manipulation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T12:49:31.987931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:49:31.987931Z digest=sha256:9da355c26011fd2edde60d56cf1174fa0544c00b54411ee4abd132e60cbe5346

Observation 69a2d934-f88a-43f7-9938-6699fabae562 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T12:49:31.993601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:49:31.993601Z digest=sha256:c5307b7f676f9675a550cfed34deddc023ea480aaf554a4779174585fe82b4dc

Observation 950a3357-adb5-444b-879c-65bb47f46f7b · outbound

This paper cites MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T12:49:32.054170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:49:32.054170Z digest=sha256:38b0214e5641ef8c8b2d47adf2bcc10f38aa9400839364b4d56c0bf89f171c01

Observation 32fb7241-ac6c-4eae-9f67-c9be919cf253 · outbound

This paper cites Lora: Low-rank adaptation of large language models.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents Lora: Low-rank adaptation of large language models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T12:49:32.166252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:49:32.166252Z digest=sha256:934f822833fd0e1a6ca50e7a35963eb8db2af60e4afd8f91fe9d3797cd1d4b61

Observation 749944ae-5b24-4ac8-9989-584ca442e286 · outbound

This paper cites Inner Monologue: Embodied Reasoning through Planning with Language Models.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents Inner Monologue: Embodied Reasoning through Planning with Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T12:49:32.305377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:49:32.305377Z digest=sha256:f3f462140adf8891ff4dee578caa85953ab0052561d449efdf8fc27a06921905

Observation 8c8ee4b2-0c9a-46ba-9f6d-887327888b86 · outbound

This paper cites V oxposer: Composable 3d value maps for robotic manipulation with language models.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents V oxposer: Composable 3d value maps for robotic manipulation with language models

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:49:38.075939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:49:32.461049Z digest=sha256:9f875d9f20d2e299cb3a1f67f9dd36d3d31adb9de5876822d5290de8511b5d50

Observation a8eca4fa-6151-461e-b687-e19a2e5942c6 · outbound

This paper cites VIMA: General Robot Manipulation with Multimodal Prompts.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents VIMA: General Robot Manipulation with Multimodal Prompts

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T12:49:32.585908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:49:32.585908Z digest=sha256:fe4cf19c9d3e5601f705934bd1574a532d2262baf3eab4fd7d7dcff4c2342b5d

Observation 0a339f41-1d52-4630-a4a1-123bcc0afd5f · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents OpenVLA: An Open-Source Vision-Language-Action Model

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:49:32.701971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:49:32.701971Z digest=sha256:dd3a31b461b15631002ff9b2d46cb110f34c1afe8415139f33ce70dd41e2b5ef

Observation 701e24d9-11d1-42db-a3e1-a9c93656eebe · outbound

This paper cites An Atomic Skill Library Construction Method for Data-Efficient Embodied Manipulation.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents An Atomic Skill Library Construction Method for Data-Efficient Embodied Manipulation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T12:49:32.866537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:49:32.866537Z digest=sha256:cfd5000acf6941e5e3758c0dcd331754969464de283eb3d422ac6959334470c9

Observation d57de53c-332a-43a2-ac21-3ca6e1c1b8d9 · outbound

This paper cites Code as policies: Language model programs for embodied control.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents Code as policies: Language model programs for embodied control

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T12:49:33.003910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:49:33.003910Z digest=sha256:43e31fec05222888074ae9715d0ef6ff586afd59839cb6637e6ad4979f06d1ad

Observation 217b65f9-7d49-441f-89af-aa5eceb772a7 · outbound

This paper cites Libero: Benchmarking knowledge transfer for lifelong robot learning.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents Libero: Benchmarking knowledge transfer for lifelong robot learning

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:49:37.869354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:49:33.106104Z digest=sha256:2350aae55113e3708dad6aab0fc64d42e81e811f831f685db13cbecc89c4b0ad

Observation 167d7aa0-df7f-4f1c-9dde-fb1bc9423f04 · outbound

This paper cites Robomamba: Efficient vision-language-action model for robotic reasoning and manipulation.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents Robomamba: Efficient vision-language-action model for robotic reasoning and manipulation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:49:37.642776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:49:33.245568Z digest=sha256:2905eccce40ac0ba918c7b96a24de417af6ae42cb240be518a13ee9e55580c0f

Observation 2949a8d8-733e-4756-8152-d8aa7c1911d5 · outbound

This paper cites GSON: A Group-based Social Navigation Framework with Large Multimodal Model.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents GSON: A Group-based Social Navigation Framework with Large Multimodal Model

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T12:49:33.358579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:49:33.358579Z digest=sha256:529e608b8e27646e66df2d957ae0438ac5de369342b3e3191bdc73ce896ddd59

Observation 132d569b-376f-4681-88eb-edca5e24113c · outbound

This paper cites MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T12:49:33.499902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:49:33.499902Z digest=sha256:957ec6ed52dbb480c22f5f780078a4c8c3a8ec5a533c59357ef33b8a274d09cc

Observation e6a5dee1-f574-4996-9181-e235de475e1f · outbound

This paper cites Data-efficient hierarchical reinforce- ment learning.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents Data-efficient hierarchical reinforce- ment learning

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:49:37.466185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:49:33.572404Z digest=sha256:dcf0eff6c5328c49701c5f193aef18a73b3e0ed9c330a301142f0599d1ee4c37

Observation 3f7bcbdf-bbbe-469d-a038-7a1993752dac · outbound

This paper cites Open x-embodiment: Robotic learning datasets and rt-x models: Open x-embodiment collaboration 0.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents Open x-embodiment: Robotic learning datasets and rt-x models: Open x-embodiment collaboration 0

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T12:49:33.724173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:49:33.724173Z digest=sha256:4f1c21af9c242aefb70092034ad72612491077ec846472bde31ba2d7c8465a8c

Observation cca39330-68b0-4ad6-b9c9-ea1328258d4d · outbound

This paper cites SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:49:33.799678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:49:33.799678Z digest=sha256:06b0494546049563c82cccaef73b041ec1254750e20eaa40a6037a6a825c8fee

Observation 445bbfe5-ae23-4d83-a8aa-97c74c6f4c31 · outbound

This paper cites VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T12:49:33.934161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:49:33.934161Z digest=sha256:03bb6285773c69fd8a5495c39ad6c8ebeb2717c0f1ff930da625e035c22b9c38

Observation 676af8f2-772d-4eaf-b204-2295bfb959a6 · outbound

This paper cites Vlm-social-nav: Socially aware robot navigation through scoring using vision-language models.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents Vlm-social-nav: Socially aware robot navigation through scoring using vision-language models

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:49:37.266362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:49:34.025327Z digest=sha256:4aac4ef6381f4f1d35da776077b0164fec016c1c19a267d9bc3065dbbb2a13a7

Observation 028e56a5-8ae6-46a8-b8c9-e8600267bf04 · outbound

This paper cites Between mdps and semi-mdps: A framework for temporal abstraction in reinforcement learning.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents Between mdps and semi-mdps: A framework for temporal abstraction in reinforcement learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T12:49:34.111610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:49:34.111610Z digest=sha256:1ee6d707889ca65ed5d6a470f9a7d280461ed648c9fe5b133c0a4a26be0dece4

Observation 7637e2df-5957-4ea1-ad5a-9cef938465fd · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents Gemini: A Family of Highly Capable Multimodal Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T12:49:34.174528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:49:34.174528Z digest=sha256:9340ee460010ec05f09075d31e42658bde1afd40fb0235598f4cf2c81c2dbb7c

Observation 6e9c6d8c-0a5e-4424-951c-df25ac59003e · outbound

This paper cites Octo: An Open-Source Generalist Robot Policy.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents Octo: An Open-Source Generalist Robot Policy

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T12:49:34.236703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:49:34.236703Z digest=sha256:61a7892e55481fbe87f61a312b8dce651c046bdd9437298bba24e41acc898d0e

Observation 749fe4ba-84ee-499d-b40c-a4addda8c2c6 · outbound

This paper cites A Survey on Post-training of Large Language Models.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents A Survey on Post-training of Large Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T12:49:34.413285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:49:34.413285Z digest=sha256:1984c5075ea3e62b246742e7fe9e177827ad01d93e9c5911ca0ba45dc4ac4373

Observation 212ee4e1-3406-4cd8-a49a-4d42d5e645ef · outbound

This paper cites Omnijarvis: Unified vision-language-action tokenization enables open-world instruction following agents.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents Omnijarvis: Unified vision-language-action tokenization enables open-world instruction following agents

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:49:37.079557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:49:34.533852Z digest=sha256:fb1bd118b102d3db254a7e960f5e1ae09b5297a825da473ca6be6d22b1019915

Observation 391d3068-1dcb-4f7b-8eb2-89a9e8058095 · outbound

This paper cites AutoGen: Enabling Next-Gen LLM Applications via Multi-Agent Conversation.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents AutoGen: Enabling Next-Gen LLM Applications via Multi-Agent Conversation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:49:34.687637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:49:34.687637Z digest=sha256:209a9f84670693e1755bbffba76b4ac372cfea79478ce9b85dd966234cd83185

Observation 1eb1c284-96dc-4b22-a9aa-7791cb7f84c0 · outbound

This paper cites Dpmpc-planner: A real-time uav trajectory planning framework for complex static environments with dynamic obstacles.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents Dpmpc-planner: A real-time uav trajectory planning framework for complex static environments with dynamic obstacles

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:49:34.838379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:49:34.838379Z digest=sha256:5320c71fac4d701c76f7a1c54ae7a46c815f26cd525df70340f1b5f2f75af2b8

Observation 7bb82601-f268-427d-8dc2-a3a5a9a4f35c · outbound

This paper cites Robomm: All-in-one multimodal large model for robotic manipulation.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents Robomm: All-in-one multimodal large model for robotic manipulation

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T12:49:34.935143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:49:34.935143Z digest=sha256:026f4bc6d746676def459ecc9ff6fb26fc4a414d9e79a3922bac9bfa14fe2213

Observation 5f38e9c5-ef59-41a1-90ec-b3355cc26ac4 · outbound

This paper cites Deer-vla: Dynamic inference of multimodal large language models for efficient robot execution.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents Deer-vla: Dynamic inference of multimodal large language models for efficient robot execution

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:49:36.864478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:49:35.028019Z digest=sha256:4653f42bd44bf5a27390558602ef685212c009498328c7354b101cd10aa38647

Observation e74ad650-c13b-4690-98bb-0100dd5465d8 · outbound

This paper cites Robotic Control via Embodied Chain-of-Thought Reasoning.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents Robotic Control via Embodied Chain-of-Thought Reasoning

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T12:49:35.156044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:49:35.156044Z digest=sha256:0f6982b4945abe1ad881279ba7d571a707ad19ce41acb38ae5f8e8e2e8aa81cf

Observation ec8f3e0b-5897-46b3-bcc2-cead05f6f881 · outbound

This paper cites NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T12:49:35.263383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:49:35.263383Z digest=sha256:aac70740b468e0c3cb429008f2619a448b1c7ff8c9cd13561070b755e5e48197

Observation 7e64bab8-3716-4a6a-9c8a-640cef8291ea · outbound

This paper cites Unifying modern ai with robotics: Survey on mdps with diffusion and foundation models.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents Unifying modern ai with robotics: Survey on mdps with diffusion and foundation models

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:49:36.712774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:49:35.384041Z digest=sha256:80d01d1f3ad07e4701799375909c04c77abbe5e92cac0d0683d2c58074a99f7a

Observation e7a02642-da1b-4d23-9c3f-4eed442f0a08 · outbound

This paper cites TraceVLA: Visual Trace Prompting Enhances Spatial-Temporal Awareness for Generalist Robotic Policies.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents TraceVLA: Visual Trace Prompting Enhances Spatial-Temporal Awareness for Generalist Robotic Policies

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T12:49:35.479145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:49:35.479145Z digest=sha256:1e5d469e61bf6f28f1932ed5e820bffc817aac01b9ce5205e4cc8f64d32cdff4

Observation 15ee6829-a0a9-4ab3-a46f-971d4b4f73f7 · outbound

This paper cites ChatVLA: Unified Multimodal Understanding and Robot Control with Vision-Language-Action Model.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents ChatVLA: Unified Multimodal Understanding and Robot Control with Vision-Language-Action Model

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T12:49:35.637019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:49:35.637019Z digest=sha256:744154023c7491c18e7523527a3f957b4f21c565706548e9e910d25643a77233

Observation d417ec4b-bbe2-4a0e-8ce9-2c15c21a19d2 · outbound

This paper cites Hierarchical planning for long-horizon manipulation with geometric and symbolic scene graphs.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents Hierarchical planning for long-horizon manipulation with geometric and symbolic scene graphs

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T12:49:35.718105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:49:35.718105Z digest=sha256:74ba1972dd8144a6a0a966d19c0a927d7dcf75864ad0b16da5b88ea0f7ff312a

Observation cc0d619b-b6af-49f6-8e72-f4f78749d0fe · outbound

This paper cites Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents.

Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:49:36.532530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:49:35.893945Z digest=sha256:97316e2b628fa15f8099567c82b47d61529e6fde71cf725ada9763243f2f4291

Pith citing papers

Observation c1fc3edf-4d67-46db-b8cd-f60af0b31192 · inbound

Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey cites this paper.

Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents

Reference 157

Resolution
verified exact
arxiv_id, observed 2026-05-17T20:28:15.994977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-17T20:28:15.818016Z digest=sha256:b63afdb5c8c2eccfe4cc0442d860efed948e5f460cae892935750733ca91c8b7

Observation 40837d7a-c78d-4c57-b691-259da894a6a5 · inbound

ThermoAct:Thermal-Aware Vision-Language-Action Models for Robotic Perception and Decision-Making cites this paper.

ThermoAct:Thermal-Aware Vision-Language-Action Models for Robotic Perception and Decision-Making Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:48:24.847308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-15T00:46:03.360770Z digest=sha256:b57172655ef7a1a663442961d8788f4fbec321b041f5f98b1a312ed5d608bec1

Observation c7a1923a-aedf-429e-820d-4135e15fe09a · inbound

A Physical Agentic Loop for Language-Guided Grasping with Execution-State Monitoring cites this paper.

A Physical Agentic Loop for Language-Guided Grasping with Execution-State Monitoring Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T00:20:55.310763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T18:34:06.704114Z digest=sha256:40a16435477f651c0b94b6a4cbe44c9f013787b18b11e207b4b15e796c172f29

Observation 307446db-3a34-4c0c-be65-ae56f1fc4a17 · inbound

Vision-and-Language Navigation for UAVs: Progress, Challenges, and a Research Roadmap cites this paper.

Vision-and-Language Navigation for UAVs: Progress, Challenges, and a Research Roadmap Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents

Reference 185

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:10:29.738474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T13:48:08.135538Z digest=sha256:5103cc36c0eefa054cb958c482c89a798b0653b8166e540ef5054a567b1b4ed9

Observation f75cafd3-bc3a-44a0-a7f7-acfc82c34efa · inbound

Bridging Values and Behavior: A Hierarchical Framework for Proactive Embodied Agents cites this paper.

Bridging Values and Behavior: A Hierarchical Framework for Proactive Embodied Agents Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:06:29.508278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-07T07:46:25.550238Z digest=sha256:50886d4d25704ee68a2a7f565177eb1575515cf6d922d81b12b941522aea9782

Observation b23c1271-9265-47b0-8160-d64f0aae333b · inbound

Towards Long-horizon Embodied Agents with Tool-Aligned Vision-Language-Action Models cites this paper.

Towards Long-horizon Embodied Agents with Tool-Aligned Vision-Language-Action Models Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-14T18:37:35.447650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-14T18:35:22.595183Z digest=sha256:69107974cda565182afcec183d4c6e5413a79dee1ad6146002e7f90169e02308

Observation b3a95fe3-ab54-49b9-8bec-b2c7f7c0eee9 · inbound

Completion at the Boundary (CaB): Deployable Switching with Completion-Aware Control under Limited Calibration cites this paper.

Completion at the Boundary (CaB): Deployable Switching with Completion-Aware Control under Limited Calibration Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-06-28T22:42:46.699503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-28T22:38:27.562345Z digest=sha256:514e587d375b64fc892d550ea4613f94da488f4a5ec4da4599a8a8b2c40703be

Observation 1d1e3e86-c332-4ff1-9812-7d813c40fa2e · inbound

VoLo: A Physical Orchestrator for Open-Vocabulary Long-Horizon Manipulation cites this paper.

VoLo: A Physical Orchestrator for Open-Vocabulary Long-Horizon Manipulation Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents

Reference 71

Resolution
verified exact
arxiv_id, observed 2026-07-02T19:07:18.172070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-27T21:40:00.330510Z digest=sha256:6b631e9bbc97fcc7dc54c33a7219a95f712ec61eb2434d7c58db3378e7024a70

Observation 9b9c7d2b-f8c5-41e7-8ec1-6f0e0bb9081a · inbound

Diagnosing Semantic Handoff Failures in Agent-Orchestrated Vision-Language-Action Skill Composition cites this paper.

Diagnosing Semantic Handoff Failures in Agent-Orchestrated Vision-Language-Action Skill Composition Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-07-08T12:04:50.487184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-07-08T12:00:14.336135Z digest=sha256:258f4f59a19058bf57f003818abd97e8d34f12cbbdb897d2caa912665563b53a

Observation 2a7face0-0a48-4b3e-87d8-62c078f4c5fe · inbound

Diagnosing Semantic Handoff Failures in Agent-Orchestrated Vision-Language-Action Skill Composition cites this paper.

Diagnosing Semantic Handoff Failures in Agent-Orchestrated Vision-Language-Action Skill Composition Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T08:21:39.024057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:21:39.024057Z digest=sha256:1eda9c3e00a08edecd7edaf3d9e585cdd3183242e9948781856aedb965c218f9

Observation 72f1dbf5-7b8a-457e-9303-25f70d223ca7 · inbound

ABot-AgentOS: A General Robotic Agent OS with Lifelong Multi-modal Memory cites this paper.

ABot-AgentOS: A General Robotic Agent OS with Lifelong Multi-modal Memory Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents

Reference 108

Resolution
unresolved
no resolver link, observed 2026-07-14T12:26:27.446079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:26:27.446079Z digest=sha256:78d3ad9f3abf3a86de36fb00443925e5e6ece48fa1d6e348d500778af0c94c32

Observation 24f1332b-af98-404c-8af8-8e8a48fb8632 · inbound

ABot-AgentOS: A General Robotic Agent OS with Lifelong Multi-modal Memory cites this paper.

ABot-AgentOS: A General Robotic Agent OS with Lifelong Multi-modal Memory Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents

Reference 108

Resolution
unresolved
no resolver link, observed 2026-08-02T07:21:35.267367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:21:35.267367Z digest=sha256:02092623a5c161601780fa1b7545febf5017cb85b536c2652a8c6b8007ba3f40

Observation bc029bc7-ec60-4a7f-83a7-c706913c876e · inbound

Learning Robust Execution in Robotic Manipulation with Agentic Reinforcement Learning cites this paper.

Learning Robust Execution in Robotic Manipulation with Agentic Reinforcement Learning Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T03:39:46.108100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:39:46.108100Z digest=sha256:9d1cbecd4408dd0bddadea9968e2eb9e42e78e188afdd152a7623324a357e75e

Observation 7d238bda-fd2b-48eb-8ff7-73b90a11b6d8 · inbound

SpatialCLI: Learning to Reason With Spatial Tools, Then Without Them cites this paper.

SpatialCLI: Learning to Reason With Spatial Tools, Then Without Them Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-01T03:06:07.841166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:06:07.841166Z digest=sha256:ee889dd58b5d8b0195cad41cb931c7f253fd8a535d6105c7e461bb5e1a25e9fb

Observation d82b96bd-144b-40c4-b02f-1daf503d4f4b · inbound

SpatialCLI: Learning to Reason With Spatial Tools, Then Without Them cites this paper.

SpatialCLI: Learning to Reason With Spatial Tools, Then Without Them Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-05T04:27:31.616619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T04:27:31.616619Z digest=sha256:843dad41f907cbe0f8acebac9aabc2256fc47e12a460e7a9be234e9cc7bb8032