Pith. sign in

Paper Citation Record · LEDGER

RationalVLA: A Rational Vision-Language-Action Model with Dual System

As of 7 August 2026, this Paper Citation Record lists 56 of 56 outbound references and 5 inbound Pith citation observations for arXiv:2506.10826.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.10826 v2

Coverage vector

measured 56 of 56 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:24:48.618223Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T20:32:46.126998Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-17T20:28:16.362582Z

Reference resolution

56 of 56 outbound references displayed

  • verified exact0
  • verified fuzzy41
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 387c9046-046c-4967-8cae-7428aa95c4dd · outbound

This paper cites Embodied intelligence toward future smart manufacturing in the era of ai foundation model,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Embodied intelligence toward future smart manufacturing in the era of ai foundation model,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.872776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:44.868625Z digest=sha256:654e98be6b29d7f6dbe80769a872503fc87db224a3cb3008bfb748f47dfc0a13

Observation d685da79-2213-4a04-a6c2-1c8e28914acf · outbound

This paper cites Rt-2: Vision-language-action models transfer web knowl- edge to robotic control,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Rt-2: Vision-language-action models transfer web knowl- edge to robotic control,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.862585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:44.912237Z digest=sha256:d9298c4538f135ae2f82761e22e7f00873d019683fcb783fe64bf56b45b22601

Observation 49c92e94-8b8a-4f63-8383-505cad7793eb · outbound

This paper cites Vision-language foundation models as effective robot imitators,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Vision-language foundation models as effective robot imitators,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.853354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:44.970763Z digest=sha256:cd7fd79bb69d97ed4a862301eef64c44c50e87c4d2beff56bd540163bb622244

Observation 191dc47f-f927-4e3e-b15b-08cb5dbe399d · outbound

This paper cites Quar-vla: Vision-language-action model for quadruped robots,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Quar-vla: Vision-language-action model for quadruped robots,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.843683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:45.022103Z digest=sha256:4cb6097f12490d3416377eae58eb3eb07cd9d2c7296522b654c40018877923b6

Observation 119901bc-e5c1-484e-a364-42d78c87e8ae · outbound

This paper cites Visual instruction tuning,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Visual instruction tuning,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:45.113110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:45.113110Z digest=sha256:31410fe625f32edfa0230d7ae7b03bb4daf943474e125df0aac911760c3063f2

Observation a2de6888-5f24-43f8-92d0-79d4ea827f92 · outbound

This paper cites GPT-4o System Card.

RationalVLA: A Rational Vision-Language-Action Model with Dual System GPT-4o System Card

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:45.193155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:45.193155Z digest=sha256:276720250bb4fde6b1e25059c4e8c8c1abfb3ca67528e432465f2a70e85145bf

Observation d6b8a8dc-f977-48bc-99ab-4b9d9a293255 · outbound

This paper cites Lisa: Reasoning segmentation via large language model,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Lisa: Reasoning segmentation via large language model,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.827431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:45.265895Z digest=sha256:1dac27cf02bf651caf0b7dd2dfe8c768e4629a942d375fb989d2ebcc583a88e9

Observation ae4f1e19-fdd8-4370-ad1e-98d726ddfd71 · outbound

This paper cites Deepseek-vl: Towards real-world vision-language understanding,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Deepseek-vl: Towards real-world vision-language understanding,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.816933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:45.321252Z digest=sha256:a357a84ebabfd7fbae41d6c8c11c21dae1043e62619de1db698d0cd440dda924

Observation 1033c4e6-fda2-44ac-98e2-541e5b8a0e55 · outbound

This paper cites Cobra: Extending mamba to multi-modal large language model for efficient inference,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Cobra: Extending mamba to multi-modal large language model for efficient inference,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.806802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:45.403055Z digest=sha256:cdc9cbb25e7083ad5e953d26f2849046a04066c728da714c92f58307087b5262

Observation 4f1d83fe-9482-4a2c-9ac0-6dc25704c33f · outbound

This paper cites Seeing far and clearly: Mitigating hallucina- tions in mllms with attention causal decoding,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Seeing far and clearly: Mitigating hallucina- tions in mllms with attention causal decoding,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.796964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:45.487310Z digest=sha256:96981c8c992d5dc4daee5ba89d13c57c7d1bf21f4e706745de5401a39b82f989

Observation b5ac0b5a-9979-48d6-a27e-d01bb338344f · outbound

This paper cites Rt-1: 11 Robotics transformer for real-world control at scale,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Rt-1: 11 Robotics transformer for real-world control at scale,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.787391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:45.554345Z digest=sha256:22206ae4d49e1bbed5691435e8547adfcdc3c353594c35052b7cf2eee0d0511c

Observation ebbe92f7-785d-4da9-a2d8-949e31587153 · outbound

This paper cites Germ: A generalist robotic model with mixture-of-experts for quadruped robot,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Germ: A generalist robotic model with mixture-of-experts for quadruped robot,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.777271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:45.590200Z digest=sha256:e6f1a035cb28fb87aafcb15fa07775f353ea8de1545e2bc6fb274df30f7a7417

Observation ea117410-28f8-4ea0-a23c-64f356e1d9d4 · outbound

This paper cites Octo: An open- source generalist robot policy,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Octo: An open- source generalist robot policy,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.766987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:45.667528Z digest=sha256:f944f82e42a2d3112dbf28a79cade7707aa2f52936b4b6c39fdb365aea66868b

Observation 80517fcf-175f-4148-b169-177e2975468e · outbound

This paper cites MoRE: Unlocking Scalability in Reinforcement Learning for Quadruped Vision-Language-Action Models.

RationalVLA: A Rational Vision-Language-Action Model with Dual System MoRE: Unlocking Scalability in Reinforcement Learning for Quadruped Vision-Language-Action Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:45.741915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:45.741915Z digest=sha256:e91ba3513a7cf5c86d42111eb0324a5e515e2cd6cd85f185c3416e64ef50aa53

Observation f131b99b-d0ac-40ac-8297-7f1dd8473089 · outbound

This paper cites Accelerating vision-language-action model inte- grated with action chunking via parallel decoding,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Accelerating vision-language-action model inte- grated with action chunking via parallel decoding,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:45.805186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:45.805186Z digest=sha256:be131f6ec3abcb30bc21591ac40a0c472ddc2e3b11fee29a3f0ece236dd118de

Observation c4f4ca83-a28a-4d09-ada4-2c9cff80f355 · outbound

This paper cites 3d diffuser actor: Policy diffusion with 3d scene representations,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System 3d diffuser actor: Policy diffusion with 3d scene representations,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.757176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:45.882089Z digest=sha256:1b01fe8ef00e01e08055d5a1b2dea66ad454024fd2de49733672eb40f48a33b8

Observation 21c901ab-49cf-4d2e-9f55-8085eff1325c · outbound

This paper cites Calvin: A benchmark for language-conditioned policy learning for long-horizon robot manipulation tasks,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Calvin: A benchmark for language-conditioned policy learning for long-horizon robot manipulation tasks,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:45.970657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:45.970657Z digest=sha256:ec4b13cfb342e415aaa6dc47ab29bf881dd527e26141a3115c57c9903fe4e3e1

Observation 3d798a09-016e-4b4a-badf-e51fd4ed9225 · outbound

This paper cites ChatVLA-2: Vision-Language-Action Model with Open-World Embodied Reasoning from Pretrained Knowledge.

RationalVLA: A Rational Vision-Language-Action Model with Dual System ChatVLA-2: Vision-Language-Action Model with Open-World Embodied Reasoning from Pretrained Knowledge

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:46.040172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:46.040172Z digest=sha256:70365382180d6ab77dc549f1959ab1c750312523a28a1dd10b0ba8e2c2bf3675

Observation ca1709af-eec4-4714-a99f-1d3a1b4158ec · outbound

This paper cites Dynamic neural networks: A survey,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Dynamic neural networks: A survey,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.692542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:46.112873Z digest=sha256:c9de9683a8e8f5a6d226509d7d98cf225e9b7c9d0f15d67c383c0691e5c33ef5

Observation c75630a4-3e9e-442c-ba4b-4c2fa5580d78 · outbound

This paper cites Gsva: Generalized segmentation via multimodal large language models,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Gsva: Generalized segmentation via multimodal large language models,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.510967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:46.188389Z digest=sha256:3187854be90ea9c7d403dfb69d009853ec538119242e33821bd060aa9686ee3b

Observation c8133683-a472-459c-9c0b-c0c7c3506d3d · outbound

This paper cites A multimodal robust recognition method for grasping objects with robot flexible grippers,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System A multimodal robust recognition method for grasping objects with robot flexible grippers,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.380767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:46.242550Z digest=sha256:fcf4526f4d7ea48f1d9883c1fc7323d7bbd76c6bdb41c004a0159275f64eeb7a

Observation 8dea092c-2df3-4ec7-8a5e-43745a7c72a0 · outbound

This paper cites Learning generalizable vision-tactile robotic grasping strategy for deformable objects via transformer,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Learning generalizable vision-tactile robotic grasping strategy for deformable objects via transformer,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.187612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:46.320504Z digest=sha256:2b203f7a236b433d8deea7465cec408d9dcffd9a3884e283bbe5308243c2de59

Observation 2e35c844-d577-4243-837b-9dfdf08756db · outbound

This paper cites Dih-tele: Dexterous in-hand teleoperation framework for learning mul- tiobjects manipulation with tactile sensing,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Dih-tele: Dexterous in-hand teleoperation framework for learning mul- tiobjects manipulation with tactile sensing,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:53.968689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:46.408258Z digest=sha256:73da0a7f6379d8e475da270ace02a955fd96c55f34fef55b7148db4453374ba4

Observation c2a83e1e-9fc5-4800-a7dc-7957fb7ebd44 · outbound

This paper cites Efficient grasp detection network with gaussian-based grasp representation for robotic manipulation,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Efficient grasp detection network with gaussian-based grasp representation for robotic manipulation,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:46.486156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:46.486156Z digest=sha256:28fae94a8f5b6a5ec571b2950e9274710924396313d6111dc7507f1017066e48

Observation 1d67e1e7-5647-4986-99a7-78a4812fb4b2 · outbound

This paper cites Language conditioned imitation learning over unstructured data,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Language conditioned imitation learning over unstructured data,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:53.685780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:46.561868Z digest=sha256:05289030f1cc9bb78ae6578ae089abfad9a2c49fddd94fb5aff15ce3b6ee05b3

Observation ea9e6e9d-3174-467e-a56a-82cd197b0bf3 · outbound

This paper cites What matters in language conditioned robotic imitation learning over unstructured data,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System What matters in language conditioned robotic imitation learning over unstructured data,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:46.649628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:46.649628Z digest=sha256:bd19423defea31751bf60584538ea1eb57e7b5c9d5b6bd82da143d7569fd3d5a

Observation b4de731c-37d6-46ed-a747-53b8704be40b · outbound

This paper cites Diffusion policy: Visuomotor policy learning via ac- tion diffusion,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Diffusion policy: Visuomotor policy learning via ac- tion diffusion,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:46.728769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:46.728769Z digest=sha256:bf8e6574303cefa6ec742331e9feee3e8d75533e4f132292a11ee5be0daac656

Observation 286f726c-5063-499a-80de-a221ea9ee758 · outbound

This paper cites 3d diffusion policy,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System 3d diffusion policy,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:53.460891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:46.790181Z digest=sha256:086ea0160772961579bdadd1783df7af958dd71ebc5ee038b316c2084bf28540

Observation d3216e62-bf3c-4d88-a14c-33c0df62c9bc · outbound

This paper cites Consistency policy: Accelerated visuomotor policies via consistency distillation,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Consistency policy: Accelerated visuomotor policies via consistency distillation,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:53.259184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:46.844026Z digest=sha256:8eddc46fde8bbd4f4a3ba25ea5cc4b73bdcaf5c3112636df9be25347e3947607

Observation b0c7a529-c8b4-4042-90cb-ee5bdb5e4d72 · outbound

This paper cites RT-trajectory: Robotic task generalization via hindsight trajectory sketches,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System RT-trajectory: Robotic task generalization via hindsight trajectory sketches,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:53.060753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:46.898184Z digest=sha256:fbb909ea7e1021f08d5dfdb4f8d5a56deb6293a51ecf552e13c0dc99a302fcf5

Observation b8533a63-2a44-4ea7-8caf-241a80272b5d · outbound

This paper cites Sara-rt: Scaling up robotics transformers with self-adaptive robust attention,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Sara-rt: Scaling up robotics transformers with self-adaptive robust attention,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:52.862970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:46.957774Z digest=sha256:ba33dd653347927ba9895221d09ce690f96c8c65aef6acaad1273248fb947afb

Observation 50e6b248-cc1c-436b-81e9-660457f5e1dc · outbound

This paper cites Inner monologue: Em- bodied reasoning through planning with language models,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Inner monologue: Em- bodied reasoning through planning with language models,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:52.691441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:47.022196Z digest=sha256:f0258df1b828eaaad1bf75195162876eb2184b58f3757fc109e68969cf4456e1

Observation f9abede6-c0b4-4513-b0cb-c0f4cf76bf20 · outbound

This paper cites Unleashing large-scale video generative pre-training for visual robot manipulation,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Unleashing large-scale video generative pre-training for visual robot manipulation,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:52.427299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:47.082933Z digest=sha256:1c7ff017b7af82455e908f6cf0ec5ef79de525dcba081dd1e84cddd3380c9340

Observation 2eda2f3a-330a-4827-826a-0b95e2ac8715 · outbound

This paper cites Video language planning,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Video language planning,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:52.151150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:47.145939Z digest=sha256:2b4a7a79bbcc31da3839828c429d6df1332242cd4a8aa454b21696548e77de24

Observation 9c3534bc-6442-43b5-a17d-b74bc742c8bd · outbound

This paper cites Robomamba: Efficient vision-language-action model for robotic reasoning and manipulation,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Robomamba: Efficient vision-language-action model for robotic reasoning and manipulation,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:52.039687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:47.193783Z digest=sha256:ed4832e426aea5cf0b27fddd7d0d48f7685d6fe1626bca52ab72554018ffd7ce

Observation 2a951cf5-e9f0-4ce4-bf9b-c545e4f26e67 · outbound

This paper cites Vlas: Vision-language-action model with speech instructions for customized robot manipulation,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Vlas: Vision-language-action model with speech instructions for customized robot manipulation,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:51.882889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:47.248717Z digest=sha256:5f55ee0f0ee5639cd1b4f71a27e7b3678f15a3724b121ef8e79a8daffd59ffdb

Observation cd995e3c-fb3a-43e7-a78a-355efaaefe90 · outbound

This paper cites Dual-arm robotic fabric manipulation with quasi-static and dynamic primitives for rapid garment flattening,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Dual-arm robotic fabric manipulation with quasi-static and dynamic primitives for rapid garment flattening,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:51.696294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:47.317109Z digest=sha256:432d459bbc04edefedfe7bea498638b0cc891a2b4ecf035a4a2cbf23bbee1c53

Observation cf663e19-9a37-4eee-b411-84fde9e53b8c · outbound

This paper cites From llms to actions: Latent codes as bridges in hierarchical robot control,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System From llms to actions: Latent codes as bridges in hierarchical robot control,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:51.511411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:47.373690Z digest=sha256:c9d7e0e0dd8ec6d0bc9dcd27e9df9b772b65bbfe211004a37b934e8b79c0671d

Observation 0c424af0-ad07-4135-855e-f574d34dd158 · outbound

This paper cites Hirt: Enhancing robotic control with hierarchical robot transformers,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Hirt: Enhancing robotic control with hierarchical robot transformers,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:51.318105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:47.438243Z digest=sha256:98e2e0b2e325ac1793f7bf7afc696247636ae220512e0d0ac9ed890651feefec

Observation de8c6ffc-a6fe-4607-b527-41a9ad8022d9 · outbound

This paper cites Towards synergistic, generalized, and efficient dual-system for robotic manipulation,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Towards synergistic, generalized, and efficient dual-system for robotic manipulation,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:51.113011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:47.485857Z digest=sha256:3288655ad7061b83d1bfbc87d1d975a69ad36e4cf74fc0d90fe6b4ae014505d1

Observation 095963b2-6f1c-4e10-bcb7-0c9684fb567d · outbound

This paper cites A Dual Process VLA: Efficient Robotic Manipulation Leveraging VLM.

RationalVLA: A Rational Vision-Language-Action Model with Dual System A Dual Process VLA: Efficient Robotic Manipulation Leveraging VLM

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:47.538512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:47.538512Z digest=sha256:afa31d720673c82a4ee76f6663c22e9b4a75d81b32ff1d321e56b456a694dc49

Observation 8132137a-43a1-49da-b5f8-88d0bfcb8855 · outbound

This paper cites Openvla: An open-source vision-language-action model,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Openvla: An open-source vision-language-action model,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:50.939829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:47.589923Z digest=sha256:00a169a42257dbea3e78b8da08d664413af813f4fb7d6c09eb9642dc3e666330

Observation 6d5ba01f-60b7-4566-a5f4-f9693fc67b41 · outbound

This paper cites Gr00t n1: An open foundation model for generalist humanoid robots,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Gr00t n1: An open foundation model for generalist humanoid robots,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:50.721443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:47.631099Z digest=sha256:2fde061b5e18b44de4a59b29637bc95364c4316a590c52a1664e44e95d0a8ab3

Observation 7ac11666-e4db-4399-898c-10e4ab4007ec · outbound

This paper cites OpenHelix: A Short Survey, Empirical Analysis, and Open-Source Dual-System VLA Model for Robotic Manipulation.

RationalVLA: A Rational Vision-Language-Action Model with Dual System OpenHelix: A Short Survey, Empirical Analysis, and Open-Source Dual-System VLA Model for Robotic Manipulation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:47.684400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:47.684400Z digest=sha256:d0451493c578c9e0be9c87eae9fdbc9934d31b082135441d815a663d32b55d20

Observation 55622c64-f47f-4ad1-8a5a-7fecc9d9c88b · outbound

This paper cites Hierarchical reinforcement learning with model guidance for mobile manipulation,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Hierarchical reinforcement learning with model guidance for mobile manipulation,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:50.469853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:47.745003Z digest=sha256:8f776ceca969f187529218428260afbb72aac04fad8efc7352cc9e7cca48c0b9

Observation 797b14a9-7840-4dba-ace9-80fa9de8d5a5 · outbound

This paper cites Interactive imitation learning of bimanual movement primitives,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Interactive imitation learning of bimanual movement primitives,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:50.246959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:47.802533Z digest=sha256:acb7cbe672a9b83e4ca021b56b12e47939f18a5dc785faa081ec38da88be62a1

Observation 75f51d1f-f10b-4d7d-a507-1f7565cbca43 · outbound

This paper cites Navigating beyond in- structions: Vision-and-language navigation in obstructed environments,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Navigating beyond in- structions: Vision-and-language navigation in obstructed environments,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:50.037859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:47.872273Z digest=sha256:e3fa6b4797b289b8f035d09bf0885c25cb6d5a94725e75670b1beee4f46f9a48

Observation dd672fc8-f90e-4b37-a440-925b000f4f7f · outbound

This paper cites BadNAVer: Exploring Jailbreak Attacks On Vision-and-Language Navigation.

RationalVLA: A Rational Vision-Language-Action Model with Dual System BadNAVer: Exploring Jailbreak Attacks On Vision-and-Language Navigation

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:47.924165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:47.924165Z digest=sha256:c9457eb73e91851b1c572e1116cd31fcd4326b201a05abc49bd9161ddc2eecef

Observation 23b0c067-e096-479b-8d13-fe55ba550797 · outbound

This paper cites Safety bounds in human robot interaction: A survey,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Safety bounds in human robot interaction: A survey,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:49.819340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:47.963015Z digest=sha256:dab0e7005b549f910fb3d4438091291ea9ef3514198d4e988e6f6d63a19746b5

Observation 3d25a37a-ca7d-463e-953a-df8b123260e3 · outbound

This paper cites Gpt-4 technical report,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Gpt-4 technical report,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:48.002210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:48.002210Z digest=sha256:6f9f084bf333e54fc92583a88ae74fd39f753ad781a03738f517f88ae5237b41

Observation c5238311-00ae-44b5-a152-7a29b06dce08 · outbound

This paper cites Pybullet, a python module for physics simulation for games, robotics and machine learning,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Pybullet, a python module for physics simulation for games, robotics and machine learning,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:49.607146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:48.076903Z digest=sha256:ef42476c1c91a8cfc08b2d2d15b8d0630304cdd863a6919dbcdabdf85baab37b

Observation 2f4c3a85-1148-411e-a693-e5cf933c4902 · outbound

This paper cites LoRA: Low-rank adaptation of large language models,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System LoRA: Low-rank adaptation of large language models,

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:48.225606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:48.225606Z digest=sha256:2739b7061be30f6a615c9bc4e5772165dc7b48907d29cfd0d4475edf64892331

Observation d08f8b2b-a536-4225-9c91-cd89043e73a5 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Learning transferable visual models from natural language supervision,

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:48.386974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:48.386974Z digest=sha256:6d57f98f881ccd88e4a0ef2547910d0a350db3e34d65946c4ab3bb4ae7a204f9

Observation b6a9d7b3-7121-4f33-80c7-924bcbd4d6ea · outbound

This paper cites Investigating the catastrophic forgetting in multimodal large language models,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Investigating the catastrophic forgetting in multimodal large language models,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:49.387945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:48.482359Z digest=sha256:71143b1de42fc84e08f7d942199dd1f650bf0a9524041bc6e04bbc440d87de2c

Observation 7bfc2d50-168d-4edf-9806-fad476749dbf · outbound

This paper cites Allava: Harnessing gpt4v-synthesized data for lite vision-language models,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Allava: Harnessing gpt4v-synthesized data for lite vision-language models,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:49.134933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:48.559986Z digest=sha256:c502eecf3a2a5ce787e392bf841b305101c54a9fd5acf2cacde32b9dbcef2aa7

Observation 87327f2a-f89b-43fb-814e-4764dd3287a2 · outbound

This paper cites Llama: Open and efficient foundation language models,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Llama: Open and efficient foundation language models,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:48.954179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:48.618223Z digest=sha256:1d5d8fb731b4e655271d9b58f8fad02384e6d4304c0584eec30a3288e1123f6e

Pith citing papers

Observation aa00de08-9610-42d6-9bd4-d8f86cb31bc2 · inbound

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver cites this paper.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver RationalVLA: A Rational Vision-Language-Action Model with Dual System

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:46.126998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:46.126998Z digest=sha256:07d16bec74f483b35b45bdbe4f72d0bc0271fa748080cf4b3a8500decc167c23

Observation 393d89ca-2901-43e6-b271-35aaeadcf5f6 · inbound

Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey cites this paper.

Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey RationalVLA: A Rational Vision-Language-Action Model with Dual System

Reference 135

Resolution
verified exact
arxiv_id, observed 2026-05-17T20:28:16.364814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T20:28:15.818016Z digest=sha256:1eeceb31c7c60f3263046168751314f4528ad0eb1742f97318be893025819712

Observation 790f1a97-6d41-44c6-ad6f-6be6bc7ff1da · inbound

RoboMemArena: A Comprehensive and Challenging Robotic Memory Benchmark cites this paper.

RoboMemArena: A Comprehensive and Challenging Robotic Memory Benchmark RationalVLA: A Rational Vision-Language-Action Model with Dual System

Reference 33

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T06:56:26.744197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-12T03:50:18.706396Z digest=sha256:3cffaaca26c01b7c44eb5b4cfa0b3c4f4d17323def49675336ebca2ee270c243

Observation 88b85e26-b952-4e1d-b7d9-cadf72149d98 · inbound

Dual-Process Atomic Skill Learning: Decoupling Semantic Reasoning and Real-Time Control cites this paper.

Dual-Process Atomic Skill Learning: Decoupling Semantic Reasoning and Real-Time Control RationalVLA: A Rational Vision-Language-Action Model with Dual System

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-14T10:23:25.158961Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T10:23:25.158961Z digest=sha256:af44da8de30ce1aa7d4c2071cdae8f32a2435744c744b9f038a98d1f86cf94f1

Observation 6512388d-82c6-4906-865d-8e13e2cf7f6e · inbound

Token-Wise Latent Streaming from Slow Reasoners to Fast Planners for Dynamic Vision Language Navigation cites this paper.

Token-Wise Latent Streaming from Slow Reasoners to Fast Planners for Dynamic Vision Language Navigation RationalVLA: A Rational Vision-Language-Action Model with Dual System

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-01T19:56:57.341989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T19:56:57.341989Z digest=sha256:e13f9e6cf2338c1157ffccd2a1027d1de3548bdbdac4aacfea3bbe0f7cad4924