Pith. sign in

Paper Citation Record · LEDGER

RoboMamba: Efficient Vision-Language-Action Model for Robotic Reasoning and Manipulation

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 22 inbound Pith citation observations for arXiv:2406.04339.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.04339 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 22 of 22 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:42:22.491193Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T15:05:47.993970Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 02702282-2022-4da4-96d1-3542817a2aa2 · inbound

A Survey of Mamba cites this paper.

A Survey of Mamba RoboMamba: Efficient Vision-Language-Action Model for Robotic Reasoning and Manipulation

Reference 121

Resolution
verified exact
arxiv_id, observed 2026-05-23T22:13:30.932811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-23T22:09:19.917854Z digest=sha256:a9988b7b3be2b4ad367a575731e45ce0284c527e02b08007bd1ed81002ad54d4

Observation 290b5786-1f73-4067-a976-d11cdcb7f124 · inbound

What Matters in Building Vision-Language-Action Models for Generalist Robots cites this paper.

What Matters in Building Vision-Language-Action Models for Generalist Robots RoboMamba: Efficient Vision-Language-Action Model for Robotic Reasoning and Manipulation

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T21:37:50.751231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-17T21:37:50.617813Z digest=sha256:9a1ee6fe5cfd82df81475f46d1eaa218c8b40915806d1251ec2895904c92ef9b

Observation 2363350b-bbe2-4354-b5c1-d3eb7f646721 · inbound

HybridVLA: Collaborative Diffusion and Autoregression in a Unified Vision-Language-Action Model cites this paper.

HybridVLA: Collaborative Diffusion and Autoregression in a Unified Vision-Language-Action Model RoboMamba: Efficient Vision-Language-Action Model for Robotic Reasoning and Manipulation

Reference 45

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T22:00:48.981558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T22:00:48.667428Z digest=sha256:9af84f1a9d3e2402fc61b3ab2eee318f4bc6c61d1e78ba51522c1e23532e7428

Observation 4028d45d-1f3b-4f46-8544-e66684b8a3ec · inbound

EfficientVLA: Training-Free Acceleration and Compression for Vision-Language-Action Models cites this paper.

EfficientVLA: Training-Free Acceleration and Compression for Vision-Language-Action Models RoboMamba: Efficient Vision-Language-Action Model for Robotic Reasoning and Manipulation

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:22.491193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:22.491193Z digest=sha256:528fa7ea2b6ca260ec98dec072ae02574d8ce5c5e074c393534cc721744bac32

Observation 38cca183-d0f7-48ce-8f4d-94170850d9d4 · inbound

AntiGrounding: Lifting Robotic Actions into VLM Representation Space for Decision Making cites this paper.

AntiGrounding: Lifting Robotic Actions into VLM Representation Space for Decision Making RoboMamba: Efficient Vision-Language-Action Model for Robotic Reasoning and Manipulation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T00:57:29.206479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:57:29.206479Z digest=sha256:e1ef143521bc7dbb78b1c623454f56e0ce9b375a94c1056907f6350900a17da9

Observation 4afea0c2-b45a-4512-ad4e-b59c05750672 · inbound

AC-DiT: Adaptive Coordination Diffusion Transformer for Mobile Manipulation cites this paper.

AC-DiT: Adaptive Coordination Diffusion Transformer for Mobile Manipulation RoboMamba: Efficient Vision-Language-Action Model for Robotic Reasoning and Manipulation

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T20:45:29.955616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:45:29.955616Z digest=sha256:cd7bcfbd1bf38a4404476939df39c271cc5716152c37f3e39d80271e42e92e28

Observation 09e0e0a6-017d-4cc1-ae66-daf950206a4f · inbound

RwoR: Generating Robot Demonstrations from Human Hand Collection for Policy Learning without Robot cites this paper.

RwoR: Generating Robot Demonstrations from Human Hand Collection for Policy Learning without Robot RoboMamba: Efficient Vision-Language-Action Model for Robotic Reasoning and Manipulation

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T20:04:38.171370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:04:38.171370Z digest=sha256:0c00e490dfabe815f8bdfd0eb572ddfafda49039c33a344f3162902f230de7af

Observation c3e0b10e-500c-40ed-9d4b-64a657555ab4 · inbound

VisionThink: Smart and Efficient Vision Language Model via Reinforcement Learning cites this paper.

VisionThink: Smart and Efficient Vision Language Model via Reinforcement Learning RoboMamba: Efficient Vision-Language-Action Model for Robotic Reasoning and Manipulation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:57.134511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:33:57.134511Z digest=sha256:bf9ea1944a5d91a9b4beef382fea073326e6ba97a4de80049dc02c50ac740ddf

Observation 827e18ad-5ac8-4119-9e93-12a9df8da143 · inbound

Source Component Shift Adaptation via Offline Decomposition and Online Mixing Approach cites this paper.

Source Component Shift Adaptation via Offline Decomposition and Online Mixing Approach RoboMamba: Efficient Vision-Language-Action Model for Robotic Reasoning and Manipulation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-05T20:35:45.475365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:35:45.475365Z digest=sha256:37fcbe6a314b2cd828be88866fdd5ebd060e4b3a1878cc4c1cb7e94042668f3b

Observation 1272386a-855e-4b88-b10d-489b933db069 · inbound

Leveraging OS-Level Primitives for Robotic Action Management cites this paper.

Leveraging OS-Level Primitives for Robotic Action Management RoboMamba: Efficient Vision-Language-Action Model for Robotic Reasoning and Manipulation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-05T20:38:42.321331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:38:42.321331Z digest=sha256:35e023b3553e418b6438a98d5c22b31f1b4ae154ec78d5e7559aec36cf3d23e0

Observation 279c0bc5-afb6-48cf-adfa-27dc11b302c6 · inbound

CogVLA: Cognition-Aligned Vision-Language-Action Model via Instruction-Driven Routing & Sparsification cites this paper.

CogVLA: Cognition-Aligned Vision-Language-Action Model via Instruction-Driven Routing & Sparsification RoboMamba: Efficient Vision-Language-Action Model for Robotic Reasoning and Manipulation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:32.706108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:42:32.706108Z digest=sha256:83963f2909a86ec771b28cb4e2f69d1bdac9d25d10df8878db704f14f531a790

Observation 232d7b16-fcef-45a4-8b47-f1d67965b293 · inbound

HeiSD: Hybrid Speculative Decoding for Embodied Vision-Language-Action Models with Kinematic Awareness cites this paper.

HeiSD: Hybrid Speculative Decoding for Embodied Vision-Language-Action Models with Kinematic Awareness RoboMamba: Efficient Vision-Language-Action Model for Robotic Reasoning and Manipulation

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-15T09:19:54.429268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T09:15:50.963123Z digest=sha256:0e81cfb5a6acf1091af74807c34a14db15dd8de7ecedd171bc620c883b5f43aa

Observation 0d2c3558-3b71-41ab-8934-d827233a0666 · inbound

The Hyperscale Lottery: How State-Space Models Have Sacrificed Edge Efficiency cites this paper.

The Hyperscale Lottery: How State-Space Models Have Sacrificed Edge Efficiency RoboMamba: Efficient Vision-Language-Action Model for Robotic Reasoning and Manipulation

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:11:01.047812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T17:46:11.859634Z digest=sha256:bbee4dc41343adc6593bec9cbe86a65848a461b63130b7362c2ceea45c7f82f7

Observation 3deb1b49-17d2-4a9d-89ec-32c44af5808c · inbound

Characterizing Vision-Language-Action Models across XPUs: Constraints and Acceleration for On-Robot Deployment cites this paper.

Characterizing Vision-Language-Action Models across XPUs: Constraints and Acceleration for On-Robot Deployment RoboMamba: Efficient Vision-Language-Action Model for Robotic Reasoning and Manipulation

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:16:52.201546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T03:04:46.460069Z digest=sha256:065b1548865e343e35717336b93ed3ba0ec7068309a44804527b3a0bad88be9c

Observation eb6e3af0-41c6-4784-a60f-2bba7d547084 · inbound

AT-VLA: Adaptive Tactile Injection for Enhanced Feedback Reaction in Vision-Language-Action Models cites this paper.

AT-VLA: Adaptive Tactile Injection for Enhanced Feedback Reaction in Vision-Language-Action Models RoboMamba: Efficient Vision-Language-Action Model for Robotic Reasoning and Manipulation

Reference 33

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T04:25:57.892889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-11T01:24:34.761241Z digest=sha256:d15bbc7766a8fee53433e9e1c8cc87e699932702bd41179704637c017ebdaf4f

Observation 63dccea2-d187-4cec-b06b-c4ded16157da · inbound

AT-VLA: Adaptive Tactile Injection for Enhanced Feedback Reaction in Vision-Language-Action Models cites this paper.

AT-VLA: Adaptive Tactile Injection for Enhanced Feedback Reaction in Vision-Language-Action Models RoboMamba: Efficient Vision-Language-Action Model for Robotic Reasoning and Manipulation

Reference 33

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T23:23:51.545108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T23:21:10.571364Z digest=sha256:e374250286cfbadb79f1fb358478bf0d35edbc5979115f49eb48e0363b8e292b

Observation c44723ea-3401-4271-946b-0bc903555288 · inbound

CrossVLA: Cross-Paradigm Post-Training and Inference Optimization for Vision-Language-Action Models cites this paper.

CrossVLA: Cross-Paradigm Post-Training and Inference Optimization for Vision-Language-Action Models RoboMamba: Efficient Vision-Language-Action Model for Robotic Reasoning and Manipulation

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-22T08:11:17.224854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-22T08:07:51.697353Z digest=sha256:b1bb3d5c26f9b1da3be035ac56415b838261b10d94a190c99975a56706f20086

Observation 14abd914-4aa0-4f37-9e5b-5391753079be · inbound

CrossVLA: Cross-Paradigm Post-Training and Inference Optimization for Vision-Language-Action Models cites this paper.

CrossVLA: Cross-Paradigm Post-Training and Inference Optimization for Vision-Language-Action Models RoboMamba: Efficient Vision-Language-Action Model for Robotic Reasoning and Manipulation

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-01T15:05:47.995575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-30T17:48:02.228941Z digest=sha256:c6d100ebf4c3afbb9031b24b862bfe18cbaf3ff995457115e87f39d8e0bb7583

Observation 5e7ca5f2-56ce-463e-8057-42436297d75f · inbound

CoDex: Learning Compositional Dexterous Functional Manipulation without Demonstrations cites this paper.

CoDex: Learning Compositional Dexterous Functional Manipulation without Demonstrations RoboMamba: Efficient Vision-Language-Action Model for Robotic Reasoning and Manipulation

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:55:41.904539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T04:58:31.739325Z digest=sha256:30e87b17f11d71ae3834f3549db2a619d7b41932c378cdc3351e5b35764cfcbb

Observation bc8f7130-84a0-4d70-9e1a-6150da353e7a · inbound

EDAR: Learning Environment-Dependent Action Representations for Robotic Manipulation cites this paper.

EDAR: Learning Environment-Dependent Action Representations for Robotic Manipulation RoboMamba: Efficient Vision-Language-Action Model for Robotic Reasoning and Manipulation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-14T05:40:47.306935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:40:47.306935Z digest=sha256:7aaab96d5ae0cfd77beff84f77ca47c6392284ce4d4b570a14b1e7522c58ccdd

Observation e630e1e5-c998-4b01-b933-51fe81955606 · inbound

CoTinyVLA: Chain-of-Thought Distillation for a Sub-Billion-Parameter Vision-Language-Action Model cites this paper.

CoTinyVLA: Chain-of-Thought Distillation for a Sub-Billion-Parameter Vision-Language-Action Model RoboMamba: Efficient Vision-Language-Action Model for Robotic Reasoning and Manipulation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T02:21:24.234162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:21:24.234162Z digest=sha256:09dca9cadedb295989560be27e4157a21754bd02df28e02361fed1f3fa8ba0f8

Observation 0ab7cd4f-ce11-4b07-ac97-2d39d3e53784 · inbound

Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills cites this paper.

Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills RoboMamba: Efficient Vision-Language-Action Model for Robotic Reasoning and Manipulation

Reference 154

Resolution
unresolved
no resolver link, observed 2026-08-04T19:45:34.954818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:45:34.954818Z digest=sha256:f860289f3a42cb922b1764a1bf5292b368409bbacc3ec3211fa866a52307220c