Pith. sign in

Paper Citation Record · LEDGER

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing

As of 18 August 2026, this Paper Citation Record lists 19 of 19 outbound references and 1 inbound Pith citation observation for arXiv:2501.06919.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.06919 v1

Coverage vector

measured 19 of 19 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T20:53:25.904376Z

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T16:55:52.464200Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T16:55:52.507941Z

Reference resolution

19 of 19 outbound references displayed

  • verified exact3
  • verified fuzzy12
  • unresolved4
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 13bf16d3-9abb-4768-8577-b732e9119b1e · outbound

This paper cites CoHRT: A Collaboration System for Human-Robot Teamwork.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing CoHRT: A Collaboration System for Human-Robot Teamwork

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-08-10T20:53:26.058696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T20:53:25.813626Z digest=sha256:3890a4d2fa041ace36caac1d4edd7d572a3ff0461ccd72e9729a42c55b28e822

Observation 0793366e-5b30-460e-8eea-015e77817a7b · outbound

This paper cites Efficient human-robot interaction using deep learning with mask r-cnn: Detection, recogni- tion, tracking and segmentation,.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing Efficient human-robot interaction using deep learning with mask r-cnn: Detection, recogni- tion, tracking and segmentation,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:53:26.250728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T20:53:25.819650Z digest=sha256:352d50df27518ab911ea051ab049abd48fd3bcd03851fad0d3c148f2746cfeab

Observation 2c2eb4b2-a388-4377-819b-3ca11356e95a · outbound

This paper cites Co-speech gestures for human-robot collaboration,.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing Co-speech gestures for human-robot collaboration,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:53:26.235107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T20:53:25.824651Z digest=sha256:d2974f85c5466e647278491a7f565894044ca0ae05742a732135534b58fbb08b

Observation ea796569-05d7-4081-9c00-2fc8b86309a2 · outbound

This paper cites Evaluating fluency in human–robot collaboration,.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing Evaluating fluency in human–robot collaboration,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:53:26.218303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T20:53:25.829743Z digest=sha256:354547b7ca124c37af0269af2150efc6c0e8df9a800284ae66250778bf4ca984

Observation a1ffe04b-6156-40bf-808f-c8689c6259f4 · outbound

This paper cites Get smart: Collaborative goal setting with cognitively assistive robots,.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing Get smart: Collaborative goal setting with cognitively assistive robots,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:53:26.201277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T20:53:25.835049Z digest=sha256:1a8c05bdb6ee4ad7283195f89d06e9d8111f440a886806b721023ee04b6da2ce

Observation cde0db74-043e-4d70-bc05-981cbfca3447 · outbound

This paper cites Personalized Speech Emotion Recognition in Human-Robot Interaction using Vision Transformers.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing Personalized Speech Emotion Recognition in Human-Robot Interaction using Vision Transformers

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-10T20:53:26.037858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T20:53:25.839934Z digest=sha256:eb72935290b7cb6d89d673bdf5303d1440be17db3009852dc572ac50aba0bb76

Observation 2047b4e6-c6d3-435e-9949-37e284acff5a · outbound

This paper cites Cobottouch: Ar-based interface with fingertip-worn tactile display for immersive operation/control of collaborative robots,.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing Cobottouch: Ar-based interface with fingertip-worn tactile display for immersive operation/control of collaborative robots,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:53:26.185918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T20:53:25.845529Z digest=sha256:35c1ac439f0184d1d85eb72708a87678b95aabf5453a3a4fc74b1124e22d6eec

Observation cbc0984e-bfc0-4fcb-8403-7c1ff1edd902 · outbound

This paper cites Coboguider: Haptic potential fields for safe human-robot interaction,.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing Coboguider: Haptic potential fields for safe human-robot interaction,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:53:26.170827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T20:53:25.850129Z digest=sha256:fd6d9fb8c0256d9bca9502cee9a6aee4340c23fd3b3ec9406ead2ddfd8ff4894

Observation f25950b5-b919-42ef-82b4-a4babc5ac192 · outbound

This paper cites Development of multi-robotic arm system for sorting system using computer vision,.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing Development of multi-robotic arm system for sorting system using computer vision,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:53:26.155779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T20:53:25.854967Z digest=sha256:edfed70d3d439609d8c93bb44d0903e961202793a247bdb9677c3937b9579c6a

Observation e3b0fcea-c1d7-4a27-a58f-4c5c71b42ec9 · outbound

This paper cites Edsinger and C.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing Edsinger and C

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:53:26.139264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T20:53:25.859874Z digest=sha256:1e5a4a5c890ffffb675a4e5ec40e12af8b0b712356cfcaf5f4634305b2532535

Observation 55d7aa78-631c-406e-b6be-58d6b4e26e90 · outbound

This paper cites GPT-4 Technical Report.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing GPT-4 Technical Report

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T20:53:25.864372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:53:25.864372Z digest=sha256:f52ece0ee6ab6f146031a095ad5a125975f8c0e6f36b0b4244a112f903c91824

Observation 4f1678b1-628d-4285-af50-84c66bb7bda0 · outbound

This paper cites Cliport: What and where pathways for robotic manipulation,.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing Cliport: What and where pathways for robotic manipulation,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:53:26.123680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T20:53:25.869195Z digest=sha256:1a0b6c8093f7cde7ba4a9b630c095e4cc3059e57b24d63f2a7d0eb6938ba43ab

Observation 14c7e728-3a75-436b-9d65-cdf6aa305cc5 · outbound

This paper cites RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T20:53:25.875000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:53:25.875000Z digest=sha256:fea3338da6d698d28b02ed5e8bff70612d8fb8f2565657413a4342297ebdeade

Observation ecbed312-3952-4e53-9ad4-660f064cee68 · outbound

This paper cites Bi-vla: Vision-language-action model- based system for bimanual robotic dexterous manipulations,.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing Bi-vla: Vision-language-action model- based system for bimanual robotic dexterous manipulations,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:53:26.106474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T20:53:25.879955Z digest=sha256:c7160e68a94d0e45827f27e6f823678c70e94c4cf19551d8575f76c06b584eb9

Observation 012067d1-2d31-42e9-bb5a-3cf0c6e1da09 · outbound

This paper cites PaLM-E: An Embodied Multimodal Language Model.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing PaLM-E: An Embodied Multimodal Language Model

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T20:53:25.884499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:53:25.884499Z digest=sha256:240f432d5fb13be7963bb77953a550d3350506f59c1d2e638d07d85f2c93f926

Observation cbff08d4-2f3f-499d-ac83-642690ab3f80 · outbound

This paper cites Industry 6.0: New Generation of Industry driven by Generative AI and Swarm of Heterogeneous Robots.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing Industry 6.0: New Generation of Industry driven by Generative AI and Swarm of Heterogeneous Robots

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-10T20:53:25.966496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T20:53:25.889372Z digest=sha256:d92b764cade19b2f7a89870438203779dc7e7ee4b1793bb8ccf8a64055a93eb6

Observation 77e11055-2ecf-4e79-b656-ee85f29c4c40 · outbound

This paper cites Robust speech recognition via large-scale weak super- vision,.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing Robust speech recognition via large-scale weak super- vision,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:53:26.090247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T20:53:25.894419Z digest=sha256:c958a1c94767c6f5984ae74400dc4a15e217a0104293f988510818da276f169d

Observation 744eccea-26a4-4e3f-be0e-15055fcf683a · outbound

This paper cites New and improved embedding model,.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing New and improved embedding model,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:53:26.075342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T20:53:25.899687Z digest=sha256:03e10db9f1a610fcac2cc4139807b1effae5e2005aee6018eca9821f6cde5c63

Observation f525b77c-4d80-4d84-9f21-38135b79698e · outbound

This paper cites The Faiss library.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing The Faiss library

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T20:53:25.904376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:53:25.904376Z digest=sha256:859dd000ed1ebf71971eb69647f163d62dcf78c718b815d7f44d3e3305893a38

Pith citing papers

Observation 4f17d2a5-9b3a-4fde-80d5-82575cc8754c · inbound

Robotic Manipulation via Imitation Learning: Taxonomy, Evolution, Benchmark, and Challenges cites this paper.

Robotic Manipulation via Imitation Learning: Taxonomy, Evolution, Benchmark, and Challenges Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing

Reference 148

Resolution
verified exact
local_arxiv, observed 2026-08-05T16:55:52.513539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T16:55:52.464200Z digest=sha256:45a5f60f51e6500d9974f05dc79a0a22446fbb543b9e3e17192bcd1186a7c19d