Pith. sign in

Paper Citation Record · LEDGER

RationalVLA: A Rational Vision-Language-Action Model with Dual System

As of 7 August 2026, this Paper Citation Record lists 56 of 56 outbound references and 5 inbound Pith citation observations for arXiv:2506.10826.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.10826 v2

Coverage vector

measured 56 of 56 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:24:48.618223Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T20:32:46.126998Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-17T20:28:16.362582Z

Reference resolution

56 of 56 outbound references displayed

  • verified exact0
  • verified fuzzy41
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 387c9046-046c-4967-8cae-7428aa95c4dd · outbound

This paper cites Embodied intelligence toward future smart manufacturing in the era of ai foundation model,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Embodied intelligence toward future smart manufacturing in the era of ai foundation model,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.872776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:44.868625Z digest=sha256:89f7dbfc80f1ec923540a69ac95b8e8335dbd9438afd3620d80da4d109506205

Observation d685da79-2213-4a04-a6c2-1c8e28914acf · outbound

This paper cites Rt-2: Vision-language-action models transfer web knowl- edge to robotic control,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Rt-2: Vision-language-action models transfer web knowl- edge to robotic control,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.862585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:44.912237Z digest=sha256:1a339bff357b53a28de62853b49c2c70c45387cfc89cdfa448619dab6c064cb1

Observation 49c92e94-8b8a-4f63-8383-505cad7793eb · outbound

This paper cites Vision-language foundation models as effective robot imitators,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Vision-language foundation models as effective robot imitators,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.853354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:44.970763Z digest=sha256:406fba25df5f8de112e5ee8ca0dad3c6d07286edfd7286d55f5820d44b50e03c

Observation 191dc47f-f927-4e3e-b15b-08cb5dbe399d · outbound

This paper cites Quar-vla: Vision-language-action model for quadruped robots,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Quar-vla: Vision-language-action model for quadruped robots,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.843683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:45.022103Z digest=sha256:bbb42a1cf5d298f9fa89d7e67bb6fa21018341cf04fd9c3877219b50d3c3bf43

Observation 119901bc-e5c1-484e-a364-42d78c87e8ae · outbound

This paper cites Visual instruction tuning,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Visual instruction tuning,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:45.113110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:45.113110Z digest=sha256:045b865b04bec23b3bd5e328b68a0bd57ea1733137812a0fa0903308414a7e51

Observation a2de6888-5f24-43f8-92d0-79d4ea827f92 · outbound

This paper cites GPT-4o System Card.

RationalVLA: A Rational Vision-Language-Action Model with Dual System GPT-4o System Card

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:45.193155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:45.193155Z digest=sha256:84a049b61122eddd5eb02158edbfe3af3d0fd0bdd5e96b57516544ef215afc11

Observation d6b8a8dc-f977-48bc-99ab-4b9d9a293255 · outbound

This paper cites Lisa: Reasoning segmentation via large language model,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Lisa: Reasoning segmentation via large language model,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.827431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:45.265895Z digest=sha256:ad2403fb6041e74c0b7dd6ab1552dce8cf5bd647f663f403058d95240fb5e89b

Observation ae4f1e19-fdd8-4370-ad1e-98d726ddfd71 · outbound

This paper cites Deepseek-vl: Towards real-world vision-language understanding,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Deepseek-vl: Towards real-world vision-language understanding,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.816933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:45.321252Z digest=sha256:437245ddd5f6c42e2c9131e0fb9f65eb50cd3d618a0d88a4b0916bef1be8f6ec

Observation 1033c4e6-fda2-44ac-98e2-541e5b8a0e55 · outbound

This paper cites Cobra: Extending mamba to multi-modal large language model for efficient inference,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Cobra: Extending mamba to multi-modal large language model for efficient inference,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.806802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:45.403055Z digest=sha256:61cd4e4384619390dd1a6262d3f8c8878c1900b427c0d4e58689344b3465d6ca

Observation 4f1d83fe-9482-4a2c-9ac0-6dc25704c33f · outbound

This paper cites Seeing far and clearly: Mitigating hallucina- tions in mllms with attention causal decoding,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Seeing far and clearly: Mitigating hallucina- tions in mllms with attention causal decoding,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.796964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:45.487310Z digest=sha256:53115f2ac8abc255411861bb70c796c48498b88f322085bbeba088fcebac2149

Observation b5ac0b5a-9979-48d6-a27e-d01bb338344f · outbound

This paper cites Rt-1: 11 Robotics transformer for real-world control at scale,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Rt-1: 11 Robotics transformer for real-world control at scale,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.787391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:45.554345Z digest=sha256:c17455eb81b5dc0091d88462a0a3c74eee2f39bb15f19fc49b95a8b392e97341

Observation ebbe92f7-785d-4da9-a2d8-949e31587153 · outbound

This paper cites Germ: A generalist robotic model with mixture-of-experts for quadruped robot,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Germ: A generalist robotic model with mixture-of-experts for quadruped robot,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.777271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:45.590200Z digest=sha256:5d40cbe0e55a7fe349edc5b2240eb7e0985e3f4d338e62fcad719a3d5149899b

Observation ea117410-28f8-4ea0-a23c-64f356e1d9d4 · outbound

This paper cites Octo: An open- source generalist robot policy,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Octo: An open- source generalist robot policy,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.766987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:45.667528Z digest=sha256:10e473cd4c2e2e6ea126a093605662bfd1ba1cee8e01b938afe5e3204efffe8b

Observation 80517fcf-175f-4148-b169-177e2975468e · outbound

This paper cites MoRE: Unlocking Scalability in Reinforcement Learning for Quadruped Vision-Language-Action Models.

RationalVLA: A Rational Vision-Language-Action Model with Dual System MoRE: Unlocking Scalability in Reinforcement Learning for Quadruped Vision-Language-Action Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:45.741915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:45.741915Z digest=sha256:473324b76527f50dc8cea065905e14f9cc44fb8caa72c93b9883b5de5cedfac5

Observation f131b99b-d0ac-40ac-8297-7f1dd8473089 · outbound

This paper cites Accelerating vision-language-action model inte- grated with action chunking via parallel decoding,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Accelerating vision-language-action model inte- grated with action chunking via parallel decoding,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:45.805186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:45.805186Z digest=sha256:2f8e75c6c9c00d3e58618f15da98a0b066fdcaa1162b25eebcd536b1a1517c9c

Observation c4f4ca83-a28a-4d09-ada4-2c9cff80f355 · outbound

This paper cites 3d diffuser actor: Policy diffusion with 3d scene representations,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System 3d diffuser actor: Policy diffusion with 3d scene representations,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.757176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:45.882089Z digest=sha256:15aba936afc2813acc12f12208a8d9d5052d28fd26b88e0d182501600f68e7e4

Observation 21c901ab-49cf-4d2e-9f55-8085eff1325c · outbound

This paper cites Calvin: A benchmark for language-conditioned policy learning for long-horizon robot manipulation tasks,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Calvin: A benchmark for language-conditioned policy learning for long-horizon robot manipulation tasks,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:45.970657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:45.970657Z digest=sha256:6a025b00a2298ef65dc35d3b08241715690429e9985c8d945694b25a3dcac805

Observation 3d798a09-016e-4b4a-badf-e51fd4ed9225 · outbound

This paper cites ChatVLA-2: Vision-Language-Action Model with Open-World Embodied Reasoning from Pretrained Knowledge.

RationalVLA: A Rational Vision-Language-Action Model with Dual System ChatVLA-2: Vision-Language-Action Model with Open-World Embodied Reasoning from Pretrained Knowledge

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:46.040172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:46.040172Z digest=sha256:68040ed3282a783110f4cd6eddf9d36216ab64bdb6b96be52104f779bfb3592b

Observation ca1709af-eec4-4714-a99f-1d3a1b4158ec · outbound

This paper cites Dynamic neural networks: A survey,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Dynamic neural networks: A survey,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.692542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:46.112873Z digest=sha256:a5d2f9a5f8ba1d031d673687ac535cf668c5ef5b4d64fe3fbffd2a40dd4f52ce

Observation c75630a4-3e9e-442c-ba4b-4c2fa5580d78 · outbound

This paper cites Gsva: Generalized segmentation via multimodal large language models,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Gsva: Generalized segmentation via multimodal large language models,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.510967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:46.188389Z digest=sha256:49fe896b1a225b957c1fb10aa8300b13533270ccc1d89e9d4498eeb21230ea1e

Observation c8133683-a472-459c-9c0b-c0c7c3506d3d · outbound

This paper cites A multimodal robust recognition method for grasping objects with robot flexible grippers,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System A multimodal robust recognition method for grasping objects with robot flexible grippers,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.380767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:46.242550Z digest=sha256:f9a5f08d86789a296b23838f2ea943908ebf27ea4304a022f02ca899a3ea2f5e

Observation 8dea092c-2df3-4ec7-8a5e-43745a7c72a0 · outbound

This paper cites Learning generalizable vision-tactile robotic grasping strategy for deformable objects via transformer,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Learning generalizable vision-tactile robotic grasping strategy for deformable objects via transformer,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.187612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:46.320504Z digest=sha256:78e56c9a640a2f9c36bca4786ace6172faa1fe29ccf037afdfe49300143d2daa

Observation 2e35c844-d577-4243-837b-9dfdf08756db · outbound

This paper cites Dih-tele: Dexterous in-hand teleoperation framework for learning mul- tiobjects manipulation with tactile sensing,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Dih-tele: Dexterous in-hand teleoperation framework for learning mul- tiobjects manipulation with tactile sensing,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:53.968689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:46.408258Z digest=sha256:769fd60a066f9934a5cf68ec2aa1b78c94882b375c20b037a34d80b12206e797

Observation c2a83e1e-9fc5-4800-a7dc-7957fb7ebd44 · outbound

This paper cites Efficient grasp detection network with gaussian-based grasp representation for robotic manipulation,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Efficient grasp detection network with gaussian-based grasp representation for robotic manipulation,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:46.486156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:46.486156Z digest=sha256:d9816538493e941a32427fac1b81ab28c2a9355592785c4835953e60009e6957

Observation 1d67e1e7-5647-4986-99a7-78a4812fb4b2 · outbound

This paper cites Language conditioned imitation learning over unstructured data,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Language conditioned imitation learning over unstructured data,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:53.685780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:46.561868Z digest=sha256:5323c8c2223c624e4cdd1eebd07d06be35bf55bc0d59e4e273e648c6d6bb986d

Observation ea9e6e9d-3174-467e-a56a-82cd197b0bf3 · outbound

This paper cites What matters in language conditioned robotic imitation learning over unstructured data,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System What matters in language conditioned robotic imitation learning over unstructured data,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:46.649628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:46.649628Z digest=sha256:4d21b1eed4f9601360f68766e03f6da927c83650877ba12c23d8f039bcc91ddf

Observation b4de731c-37d6-46ed-a747-53b8704be40b · outbound

This paper cites Diffusion policy: Visuomotor policy learning via ac- tion diffusion,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Diffusion policy: Visuomotor policy learning via ac- tion diffusion,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:46.728769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:46.728769Z digest=sha256:cad646297d2558abf3418c1bb00b0b56896e012e52d780d80d322089fcd75a76

Observation 286f726c-5063-499a-80de-a221ea9ee758 · outbound

This paper cites 3d diffusion policy,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System 3d diffusion policy,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:53.460891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:46.790181Z digest=sha256:5c64fb1fb4ac109e709dd9973b6bd279d83479c3ade9a009ec9c31198f4453f7

Observation d3216e62-bf3c-4d88-a14c-33c0df62c9bc · outbound

This paper cites Consistency policy: Accelerated visuomotor policies via consistency distillation,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Consistency policy: Accelerated visuomotor policies via consistency distillation,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:53.259184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:46.844026Z digest=sha256:786c8d40ac825eedd28811c9ed03a564f75a6be697674d601dae1bfb713b22b6

Observation b0c7a529-c8b4-4042-90cb-ee5bdb5e4d72 · outbound

This paper cites RT-trajectory: Robotic task generalization via hindsight trajectory sketches,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System RT-trajectory: Robotic task generalization via hindsight trajectory sketches,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:53.060753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:46.898184Z digest=sha256:a7febb151f39786aa49c0d6245a479bbb74812c58a1d6123b705510d39b2befd

Observation b8533a63-2a44-4ea7-8caf-241a80272b5d · outbound

This paper cites Sara-rt: Scaling up robotics transformers with self-adaptive robust attention,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Sara-rt: Scaling up robotics transformers with self-adaptive robust attention,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:52.862970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:46.957774Z digest=sha256:118d252597ed45a0af7bef63c4a6f19798fa39d79872d143c0480f95f126fe97

Observation 50e6b248-cc1c-436b-81e9-660457f5e1dc · outbound

This paper cites Inner monologue: Em- bodied reasoning through planning with language models,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Inner monologue: Em- bodied reasoning through planning with language models,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:52.691441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:47.022196Z digest=sha256:85ce1bee79960967d6dfc295acc68024bc4b300f2074cefd5317d70e32e018ad

Observation f9abede6-c0b4-4513-b0cb-c0f4cf76bf20 · outbound

This paper cites Unleashing large-scale video generative pre-training for visual robot manipulation,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Unleashing large-scale video generative pre-training for visual robot manipulation,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:52.427299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:47.082933Z digest=sha256:5146ab41b02082ecbc8118c3b03a4f5a65704da57d1d8a5176a6c16412633806

Observation 2eda2f3a-330a-4827-826a-0b95e2ac8715 · outbound

This paper cites Video language planning,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Video language planning,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:52.151150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:47.145939Z digest=sha256:61f898290d3ff9259a1b9feba52a8df003dcaf2bd2167e58783ba683ca0ddba7

Observation 9c3534bc-6442-43b5-a17d-b74bc742c8bd · outbound

This paper cites Robomamba: Efficient vision-language-action model for robotic reasoning and manipulation,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Robomamba: Efficient vision-language-action model for robotic reasoning and manipulation,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:52.039687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:47.193783Z digest=sha256:25a4a5a488a8e20a3121d60695fe2060a2d472f1e5fe3ce5ccda32d677599af9

Observation 2a951cf5-e9f0-4ce4-bf9b-c545e4f26e67 · outbound

This paper cites Vlas: Vision-language-action model with speech instructions for customized robot manipulation,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Vlas: Vision-language-action model with speech instructions for customized robot manipulation,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:51.882889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:47.248717Z digest=sha256:ffc4047f6b995c1e48524f51528b4252da698a1bfa52de47d80b91474895990f

Observation cd995e3c-fb3a-43e7-a78a-355efaaefe90 · outbound

This paper cites Dual-arm robotic fabric manipulation with quasi-static and dynamic primitives for rapid garment flattening,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Dual-arm robotic fabric manipulation with quasi-static and dynamic primitives for rapid garment flattening,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:51.696294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:47.317109Z digest=sha256:fc8fd39b8b22a7644f47c604424921b306f316b0bbe6f08cd26ba42bc2343421

Observation cf663e19-9a37-4eee-b411-84fde9e53b8c · outbound

This paper cites From llms to actions: Latent codes as bridges in hierarchical robot control,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System From llms to actions: Latent codes as bridges in hierarchical robot control,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:51.511411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:47.373690Z digest=sha256:f001854d7ed83b894f92acd20aa9a548959c4e9ab12e03baf849487dc15d61d4

Observation 0c424af0-ad07-4135-855e-f574d34dd158 · outbound

This paper cites Hirt: Enhancing robotic control with hierarchical robot transformers,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Hirt: Enhancing robotic control with hierarchical robot transformers,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:51.318105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:47.438243Z digest=sha256:020c86dcc9d1a75d760602768764136758970666246a80ef917145c94421baee

Observation de8c6ffc-a6fe-4607-b527-41a9ad8022d9 · outbound

This paper cites Towards synergistic, generalized, and efficient dual-system for robotic manipulation,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Towards synergistic, generalized, and efficient dual-system for robotic manipulation,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:51.113011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:47.485857Z digest=sha256:f382671b029157cd3131f70a5ddbabcfe0b1b965539d050d5588d1ed90557124

Observation 095963b2-6f1c-4e10-bcb7-0c9684fb567d · outbound

This paper cites A Dual Process VLA: Efficient Robotic Manipulation Leveraging VLM.

RationalVLA: A Rational Vision-Language-Action Model with Dual System A Dual Process VLA: Efficient Robotic Manipulation Leveraging VLM

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:47.538512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:47.538512Z digest=sha256:48f8b8790172ce5704f3fa0a242f694a552c83a1feef6c544cce7c18130fa9cb

Observation 8132137a-43a1-49da-b5f8-88d0bfcb8855 · outbound

This paper cites Openvla: An open-source vision-language-action model,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Openvla: An open-source vision-language-action model,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:50.939829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:47.589923Z digest=sha256:fb9a97891bfa1a96e03d1a57f1a43f44a6e169fc7d8b56b1accb05b71304eb5b

Observation 6d5ba01f-60b7-4566-a5f4-f9693fc67b41 · outbound

This paper cites Gr00t n1: An open foundation model for generalist humanoid robots,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Gr00t n1: An open foundation model for generalist humanoid robots,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:50.721443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:47.631099Z digest=sha256:a345bda2b098a4fc8bdccf8fc3283c090ae598f6d6159a3e25c15ec9c647dbf5

Observation 7ac11666-e4db-4399-898c-10e4ab4007ec · outbound

This paper cites OpenHelix: A Short Survey, Empirical Analysis, and Open-Source Dual-System VLA Model for Robotic Manipulation.

RationalVLA: A Rational Vision-Language-Action Model with Dual System OpenHelix: A Short Survey, Empirical Analysis, and Open-Source Dual-System VLA Model for Robotic Manipulation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:47.684400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:47.684400Z digest=sha256:cf34d6f5b780ed551e0fb19dd9d2a74dd56fdb653fa78fe075991b3f9230f21e

Observation 55622c64-f47f-4ad1-8a5a-7fecc9d9c88b · outbound

This paper cites Hierarchical reinforcement learning with model guidance for mobile manipulation,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Hierarchical reinforcement learning with model guidance for mobile manipulation,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:50.469853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:47.745003Z digest=sha256:fea72c7e96403cf704c30d614ec845fe478658a2c775adaea5ef2c29a6782ed2

Observation 797b14a9-7840-4dba-ace9-80fa9de8d5a5 · outbound

This paper cites Interactive imitation learning of bimanual movement primitives,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Interactive imitation learning of bimanual movement primitives,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:50.246959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:47.802533Z digest=sha256:1dfa70e8b6c9848a3dec6880b8e3fc2d413f7bcaf8d5ede5ee7b49a495c140dc

Observation 75f51d1f-f10b-4d7d-a507-1f7565cbca43 · outbound

This paper cites Navigating beyond in- structions: Vision-and-language navigation in obstructed environments,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Navigating beyond in- structions: Vision-and-language navigation in obstructed environments,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:50.037859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:47.872273Z digest=sha256:d22e38644a1cd2c3eba72d22e1871bad77682c63929c61bf4ee815cd6e700494

Observation dd672fc8-f90e-4b37-a440-925b000f4f7f · outbound

This paper cites BadNAVer: Exploring Jailbreak Attacks On Vision-and-Language Navigation.

RationalVLA: A Rational Vision-Language-Action Model with Dual System BadNAVer: Exploring Jailbreak Attacks On Vision-and-Language Navigation

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:47.924165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:47.924165Z digest=sha256:ca58999d6cc05471b2f3a5f23b11c46bffdc037caee270cf9458e5437a932083

Observation 23b0c067-e096-479b-8d13-fe55ba550797 · outbound

This paper cites Safety bounds in human robot interaction: A survey,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Safety bounds in human robot interaction: A survey,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:49.819340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:47.963015Z digest=sha256:7d92a3a7710147878b49b992c64188047876cbf4890d8042b782678583e708a3

Observation 3d25a37a-ca7d-463e-953a-df8b123260e3 · outbound

This paper cites Gpt-4 technical report,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Gpt-4 technical report,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:48.002210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:48.002210Z digest=sha256:d48bec0aabfac917c137d32f2b77415411642d6f64fcb6c243dfd66f87547d17

Observation c5238311-00ae-44b5-a152-7a29b06dce08 · outbound

This paper cites Pybullet, a python module for physics simulation for games, robotics and machine learning,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Pybullet, a python module for physics simulation for games, robotics and machine learning,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:49.607146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:48.076903Z digest=sha256:e0f8ac1847831cca77c95747171e47e78a7e41391b007bd14a5709d4d9ba72f7

Observation 2f4c3a85-1148-411e-a693-e5cf933c4902 · outbound

This paper cites LoRA: Low-rank adaptation of large language models,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System LoRA: Low-rank adaptation of large language models,

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:48.225606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:48.225606Z digest=sha256:5aedc15f7a14f513088722646f27d121581726ded3de1eb57021cc5aa20a22d0

Observation d08f8b2b-a536-4225-9c91-cd89043e73a5 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Learning transferable visual models from natural language supervision,

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:48.386974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:48.386974Z digest=sha256:19eaec7457116bd9258e556a70b549ca7d52205f9206449d6dad65b801d52d3b

Observation b6a9d7b3-7121-4f33-80c7-924bcbd4d6ea · outbound

This paper cites Investigating the catastrophic forgetting in multimodal large language models,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Investigating the catastrophic forgetting in multimodal large language models,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:49.387945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:48.482359Z digest=sha256:e920069b689b3f57a524a67acf546c95acf6c2627b51b93a735a59f098005f57

Observation 7bfc2d50-168d-4edf-9806-fad476749dbf · outbound

This paper cites Allava: Harnessing gpt4v-synthesized data for lite vision-language models,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Allava: Harnessing gpt4v-synthesized data for lite vision-language models,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:49.134933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:48.559986Z digest=sha256:1c7175e834afe0055f5020d8394026cfd58bef7eda25c9e42d463b16f0971bc2

Observation 87327f2a-f89b-43fb-814e-4764dd3287a2 · outbound

This paper cites Llama: Open and efficient foundation language models,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Llama: Open and efficient foundation language models,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:48.954179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:24:48.618223Z digest=sha256:99baabcb931fc840e865c36af8ed99b4cc04bf31fce03ecb91093bbfefcaf850

Pith citing papers

Observation aa00de08-9610-42d6-9bd4-d8f86cb31bc2 · inbound

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver cites this paper.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver RationalVLA: A Rational Vision-Language-Action Model with Dual System

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:46.126998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:46.126998Z digest=sha256:7606201e1555677c0d5f620f7a222aa1641686a63b30e519e164535be34375b7

Observation 393d89ca-2901-43e6-b271-35aaeadcf5f6 · inbound

Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey cites this paper.

Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey RationalVLA: A Rational Vision-Language-Action Model with Dual System

Reference 135

Resolution
verified exact
arxiv_id, observed 2026-05-17T20:28:16.364814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T20:28:15.818016Z digest=sha256:8921b6368cf9874772e262a1ab35e5270de7a0977da2f5e6e6641d0c1bfbd4f2

Observation 790f1a97-6d41-44c6-ad6f-6be6bc7ff1da · inbound

RoboMemArena: A Comprehensive and Challenging Robotic Memory Benchmark cites this paper.

RoboMemArena: A Comprehensive and Challenging Robotic Memory Benchmark RationalVLA: A Rational Vision-Language-Action Model with Dual System

Reference 33

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T06:56:26.744197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-12T03:50:18.706396Z digest=sha256:a402cfdc85001cbd396e843f464ef756bbf6ec0a962ebd20e12060d61efaaa9f

Observation 88b85e26-b952-4e1d-b7d9-cadf72149d98 · inbound

Dual-Process Atomic Skill Learning: Decoupling Semantic Reasoning and Real-Time Control cites this paper.

Dual-Process Atomic Skill Learning: Decoupling Semantic Reasoning and Real-Time Control RationalVLA: A Rational Vision-Language-Action Model with Dual System

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-14T10:23:25.158961Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T10:23:25.158961Z digest=sha256:e89738d689dbe2092ac2c1aeae1f140eccf3eeb554bc883c6c14ecdde4095668

Observation 6512388d-82c6-4906-865d-8e13e2cf7f6e · inbound

Token-Wise Latent Streaming from Slow Reasoners to Fast Planners for Dynamic Vision Language Navigation cites this paper.

Token-Wise Latent Streaming from Slow Reasoners to Fast Planners for Dynamic Vision Language Navigation RationalVLA: A Rational Vision-Language-Action Model with Dual System

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-01T19:56:57.341989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T19:56:57.341989Z digest=sha256:e4aee4879322fe54d5d398af995ac7dc3c4869dfb18f4fdad64de18e25c1cf48