Pith. sign in

Paper Citation Record · LEDGER

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing

As of 19 August 2026, this Paper Citation Record lists 19 of 19 outbound references and 1 inbound Pith citation observation for arXiv:2501.06919.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.06919 v1

Coverage vector

measured 19 of 19 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T20:53:25.904376Z

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T16:55:52.464200Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T16:55:52.507941Z

Reference resolution

19 of 19 outbound references displayed

  • verified exact3
  • verified fuzzy12
  • unresolved4
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 13bf16d3-9abb-4768-8577-b732e9119b1e · outbound

This paper cites CoHRT: A Collaboration System for Human-Robot Teamwork.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing CoHRT: A Collaboration System for Human-Robot Teamwork

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-08-10T20:53:26.058696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T20:53:25.813626Z digest=sha256:1250f323828d92512f995691f40593155e2085892337360fcf5c06a0f5fa50c3

Observation 0793366e-5b30-460e-8eea-015e77817a7b · outbound

This paper cites Efficient human-robot interaction using deep learning with mask r-cnn: Detection, recogni- tion, tracking and segmentation,.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing Efficient human-robot interaction using deep learning with mask r-cnn: Detection, recogni- tion, tracking and segmentation,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:53:26.250728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T20:53:25.819650Z digest=sha256:b5e5c971f19423e288e8bfbacffe496e552c0eec32794526511e95aa0ac64e5c

Observation 2c2eb4b2-a388-4377-819b-3ca11356e95a · outbound

This paper cites Co-speech gestures for human-robot collaboration,.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing Co-speech gestures for human-robot collaboration,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:53:26.235107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T20:53:25.824651Z digest=sha256:187f868471c9627b0a73f503bdafb22300744508f75eee22ffdce816543ddbac

Observation ea796569-05d7-4081-9c00-2fc8b86309a2 · outbound

This paper cites Evaluating fluency in human–robot collaboration,.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing Evaluating fluency in human–robot collaboration,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:53:26.218303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T20:53:25.829743Z digest=sha256:dc806463dfda34e66879ebdeff675b88623b8cae2d252466c0991e75a5b7ab96

Observation a1ffe04b-6156-40bf-808f-c8689c6259f4 · outbound

This paper cites Get smart: Collaborative goal setting with cognitively assistive robots,.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing Get smart: Collaborative goal setting with cognitively assistive robots,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:53:26.201277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T20:53:25.835049Z digest=sha256:d3696bd9115e96f55d3cd18419b0e18ac20e861ce665742fd0a3d1dcb684ea62

Observation cde0db74-043e-4d70-bc05-981cbfca3447 · outbound

This paper cites Personalized Speech Emotion Recognition in Human-Robot Interaction using Vision Transformers.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing Personalized Speech Emotion Recognition in Human-Robot Interaction using Vision Transformers

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-10T20:53:26.037858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T20:53:25.839934Z digest=sha256:542982a4d9d40fa54bf67b19d8f3db41a5f83addbf898c77e9a2ce377ced9651

Observation 2047b4e6-c6d3-435e-9949-37e284acff5a · outbound

This paper cites Cobottouch: Ar-based interface with fingertip-worn tactile display for immersive operation/control of collaborative robots,.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing Cobottouch: Ar-based interface with fingertip-worn tactile display for immersive operation/control of collaborative robots,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:53:26.185918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T20:53:25.845529Z digest=sha256:c83400c931aa337d3fc7167688d92756c1efd778c42edbe508890806636365f5

Observation cbc0984e-bfc0-4fcb-8403-7c1ff1edd902 · outbound

This paper cites Coboguider: Haptic potential fields for safe human-robot interaction,.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing Coboguider: Haptic potential fields for safe human-robot interaction,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:53:26.170827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T20:53:25.850129Z digest=sha256:cf07a9f9ff92ecf499bcf8ce3de0036f8f596c178fa6f8adb7055ecaecdf32a1

Observation f25950b5-b919-42ef-82b4-a4babc5ac192 · outbound

This paper cites Development of multi-robotic arm system for sorting system using computer vision,.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing Development of multi-robotic arm system for sorting system using computer vision,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:53:26.155779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T20:53:25.854967Z digest=sha256:f2b24ee700044756c32eec2359bb50fd3955aa2627d790afdd62dd1c1f27d881

Observation e3b0fcea-c1d7-4a27-a58f-4c5c71b42ec9 · outbound

This paper cites Edsinger and C.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing Edsinger and C

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:53:26.139264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T20:53:25.859874Z digest=sha256:3762eb66f275f55e0d2c55b34728c09fceaa3248110a9c05b501520f7567c287

Observation 55d7aa78-631c-406e-b6be-58d6b4e26e90 · outbound

This paper cites GPT-4 Technical Report.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing GPT-4 Technical Report

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T20:53:25.864372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:53:25.864372Z digest=sha256:e6b48432068507adf5aa381cde591be7a2adba760b2515da7ca3b384f444554b

Observation 4f1678b1-628d-4285-af50-84c66bb7bda0 · outbound

This paper cites Cliport: What and where pathways for robotic manipulation,.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing Cliport: What and where pathways for robotic manipulation,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:53:26.123680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T20:53:25.869195Z digest=sha256:831550109607e365fc2c2ac743458e8cb475f6892c2b49ad1c36be1cfbe626ce

Observation 14c7e728-3a75-436b-9d65-cdf6aa305cc5 · outbound

This paper cites RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T20:53:25.875000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:53:25.875000Z digest=sha256:2e1049d7fad1bfcbd13f9c7570ef30145f4d86bd0473574412942b71aed81170

Observation ecbed312-3952-4e53-9ad4-660f064cee68 · outbound

This paper cites Bi-vla: Vision-language-action model- based system for bimanual robotic dexterous manipulations,.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing Bi-vla: Vision-language-action model- based system for bimanual robotic dexterous manipulations,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:53:26.106474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T20:53:25.879955Z digest=sha256:dfc4e673cfe9ac47fcdd56fe7b0ad19b821f4459f430b13b3172b4c0212a996b

Observation 012067d1-2d31-42e9-bb5a-3cf0c6e1da09 · outbound

This paper cites PaLM-E: An Embodied Multimodal Language Model.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing PaLM-E: An Embodied Multimodal Language Model

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T20:53:25.884499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:53:25.884499Z digest=sha256:193e0758e46d9c32385cfcae757112b53a38c22ff69b29eea09302b0b16caa8f

Observation cbff08d4-2f3f-499d-ac83-642690ab3f80 · outbound

This paper cites Industry 6.0: New Generation of Industry driven by Generative AI and Swarm of Heterogeneous Robots.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing Industry 6.0: New Generation of Industry driven by Generative AI and Swarm of Heterogeneous Robots

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-10T20:53:25.966496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T20:53:25.889372Z digest=sha256:434d3f4e3c2ecd5f8d20a7315b45059b2f9a6e062b98cb69e3b51c1c51548f03

Observation 77e11055-2ecf-4e79-b656-ee85f29c4c40 · outbound

This paper cites Robust speech recognition via large-scale weak super- vision,.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing Robust speech recognition via large-scale weak super- vision,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:53:26.090247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T20:53:25.894419Z digest=sha256:a454dbdf9d7d119c84692ae942b0065b5b61325a5be863f8f2c3cece4d021497

Observation 744eccea-26a4-4e3f-be0e-15055fcf683a · outbound

This paper cites New and improved embedding model,.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing New and improved embedding model,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:53:26.075342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T20:53:25.899687Z digest=sha256:5010935512ee9b5bdb0873c39e328ea381db6dffbf001fd36ffc2763e53ad64a

Observation f525b77c-4d80-4d84-9f21-38135b79698e · outbound

This paper cites The Faiss library.

Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing The Faiss library

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T20:53:25.904376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:53:25.904376Z digest=sha256:bb29bae28a2dc84b71bb613372d139d488a8c9d60043f60efee31f7feaf5fe4f

Pith citing papers

Observation 4f17d2a5-9b3a-4fde-80d5-82575cc8754c · inbound

Robotic Manipulation via Imitation Learning: Taxonomy, Evolution, Benchmark, and Challenges cites this paper.

Robotic Manipulation via Imitation Learning: Taxonomy, Evolution, Benchmark, and Challenges Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing

Reference 148

Resolution
verified exact
local_arxiv, observed 2026-08-05T16:55:52.513539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T16:55:52.464200Z digest=sha256:87ca36f4dc8a53b741d2574bb11d6a8418a594eff38bafea0514709a9e8e4933