Pith. sign in

Paper Citation Record · LEDGER

RationalVLA: A Rational Vision-Language-Action Model with Dual System

As of 8 August 2026, this Paper Citation Record lists 56 of 56 outbound references and 5 inbound Pith citation observations for arXiv:2506.10826.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.10826 v2

Coverage vector

measured 56 of 56 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:24:48.618223Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T20:32:46.126998Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-17T20:28:16.362582Z

Reference resolution

56 of 56 outbound references displayed

  • verified exact0
  • verified fuzzy41
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 387c9046-046c-4967-8cae-7428aa95c4dd · outbound

This paper cites Embodied intelligence toward future smart manufacturing in the era of ai foundation model,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Embodied intelligence toward future smart manufacturing in the era of ai foundation model,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.872776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:44.868625Z digest=sha256:64950c9882ac3d86c689978efa326173c2fc6e7afe5206517a6a8549b4ebee5d

Observation d685da79-2213-4a04-a6c2-1c8e28914acf · outbound

This paper cites Rt-2: Vision-language-action models transfer web knowl- edge to robotic control,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Rt-2: Vision-language-action models transfer web knowl- edge to robotic control,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.862585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:44.912237Z digest=sha256:2005322341e18d65eaff6e6d0d88abc45cfe530175319696e28474443552a396

Observation 49c92e94-8b8a-4f63-8383-505cad7793eb · outbound

This paper cites Vision-language foundation models as effective robot imitators,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Vision-language foundation models as effective robot imitators,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.853354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:44.970763Z digest=sha256:4c9f578158a1b0b8bccd9d6de7540827773f29b8ac51ab6e294c5403286fd4dd

Observation 191dc47f-f927-4e3e-b15b-08cb5dbe399d · outbound

This paper cites Quar-vla: Vision-language-action model for quadruped robots,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Quar-vla: Vision-language-action model for quadruped robots,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.843683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:45.022103Z digest=sha256:7f0042e0c67ec909cbddb86dd85c8d857a1353d1e6ebe9a76ba4c9f600ba2d50

Observation 119901bc-e5c1-484e-a364-42d78c87e8ae · outbound

This paper cites Visual instruction tuning,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Visual instruction tuning,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:45.113110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:45.113110Z digest=sha256:045b865b04bec23b3bd5e328b68a0bd57ea1733137812a0fa0903308414a7e51

Observation a2de6888-5f24-43f8-92d0-79d4ea827f92 · outbound

This paper cites GPT-4o System Card.

RationalVLA: A Rational Vision-Language-Action Model with Dual System GPT-4o System Card

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:45.193155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:45.193155Z digest=sha256:84a049b61122eddd5eb02158edbfe3af3d0fd0bdd5e96b57516544ef215afc11

Observation d6b8a8dc-f977-48bc-99ab-4b9d9a293255 · outbound

This paper cites Lisa: Reasoning segmentation via large language model,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Lisa: Reasoning segmentation via large language model,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.827431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:45.265895Z digest=sha256:c324ecfd1afe9d53162c4618e5c1f5f71ac696774c032d78a96026a780710210

Observation ae4f1e19-fdd8-4370-ad1e-98d726ddfd71 · outbound

This paper cites Deepseek-vl: Towards real-world vision-language understanding,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Deepseek-vl: Towards real-world vision-language understanding,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.816933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:45.321252Z digest=sha256:6f4cc0753cd265268d0073904643d6a3afbcba32a5270933564f2ad5d233a89b

Observation 1033c4e6-fda2-44ac-98e2-541e5b8a0e55 · outbound

This paper cites Cobra: Extending mamba to multi-modal large language model for efficient inference,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Cobra: Extending mamba to multi-modal large language model for efficient inference,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.806802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:45.403055Z digest=sha256:9385e2f5fffd8f6e2281051004b5372e9faec969f32c3e3a91352682bf0b073c

Observation 4f1d83fe-9482-4a2c-9ac0-6dc25704c33f · outbound

This paper cites Seeing far and clearly: Mitigating hallucina- tions in mllms with attention causal decoding,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Seeing far and clearly: Mitigating hallucina- tions in mllms with attention causal decoding,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.796964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:45.487310Z digest=sha256:62e9b538d17f79bcb0e3cfd16d7c4c8b596e27f4ddba1ada88ee7943430e8708

Observation b5ac0b5a-9979-48d6-a27e-d01bb338344f · outbound

This paper cites Rt-1: 11 Robotics transformer for real-world control at scale,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Rt-1: 11 Robotics transformer for real-world control at scale,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.787391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:45.554345Z digest=sha256:9e08c1aa7b063fff17132bc79269a6952f6f56ea32579ca0ecadb9ab186a5f17

Observation ebbe92f7-785d-4da9-a2d8-949e31587153 · outbound

This paper cites Germ: A generalist robotic model with mixture-of-experts for quadruped robot,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Germ: A generalist robotic model with mixture-of-experts for quadruped robot,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.777271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:45.590200Z digest=sha256:be9772bb46720d67883ef0e0de17239e0e1ef99f36d108c6afd8688127721838

Observation ea117410-28f8-4ea0-a23c-64f356e1d9d4 · outbound

This paper cites Octo: An open- source generalist robot policy,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Octo: An open- source generalist robot policy,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.766987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:45.667528Z digest=sha256:178bad1615d507d19790cd8732a53a0922a433953126963268ce7286c6c6eb93

Observation 80517fcf-175f-4148-b169-177e2975468e · outbound

This paper cites MoRE: Unlocking Scalability in Reinforcement Learning for Quadruped Vision-Language-Action Models.

RationalVLA: A Rational Vision-Language-Action Model with Dual System MoRE: Unlocking Scalability in Reinforcement Learning for Quadruped Vision-Language-Action Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:45.741915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:45.741915Z digest=sha256:bec92b1f8828ad7b524f41acbff887eb60ad5c5f36bc31b4673a59af120232e9

Observation f131b99b-d0ac-40ac-8297-7f1dd8473089 · outbound

This paper cites Accelerating vision-language-action model inte- grated with action chunking via parallel decoding,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Accelerating vision-language-action model inte- grated with action chunking via parallel decoding,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:45.805186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:45.805186Z digest=sha256:2f8e75c6c9c00d3e58618f15da98a0b066fdcaa1162b25eebcd536b1a1517c9c

Observation c4f4ca83-a28a-4d09-ada4-2c9cff80f355 · outbound

This paper cites 3d diffuser actor: Policy diffusion with 3d scene representations,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System 3d diffuser actor: Policy diffusion with 3d scene representations,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.757176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:45.882089Z digest=sha256:6eaa7ff02569847bca9efa7a24b1c84123bd7630b7855b5a0203722d799bb942

Observation 21c901ab-49cf-4d2e-9f55-8085eff1325c · outbound

This paper cites Calvin: A benchmark for language-conditioned policy learning for long-horizon robot manipulation tasks,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Calvin: A benchmark for language-conditioned policy learning for long-horizon robot manipulation tasks,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:45.970657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:45.970657Z digest=sha256:6a025b00a2298ef65dc35d3b08241715690429e9985c8d945694b25a3dcac805

Observation 3d798a09-016e-4b4a-badf-e51fd4ed9225 · outbound

This paper cites ChatVLA-2: Vision-Language-Action Model with Open-World Embodied Reasoning from Pretrained Knowledge.

RationalVLA: A Rational Vision-Language-Action Model with Dual System ChatVLA-2: Vision-Language-Action Model with Open-World Embodied Reasoning from Pretrained Knowledge

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:46.040172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:46.040172Z digest=sha256:68040ed3282a783110f4cd6eddf9d36216ab64bdb6b96be52104f779bfb3592b

Observation ca1709af-eec4-4714-a99f-1d3a1b4158ec · outbound

This paper cites Dynamic neural networks: A survey,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Dynamic neural networks: A survey,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.692542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:46.112873Z digest=sha256:7cc4edc1f59add0826d129b6f898a447084b772d9ef9363dec4a5f0d73926fe7

Observation c75630a4-3e9e-442c-ba4b-4c2fa5580d78 · outbound

This paper cites Gsva: Generalized segmentation via multimodal large language models,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Gsva: Generalized segmentation via multimodal large language models,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.510967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:46.188389Z digest=sha256:50a9d6c5d751229d5668d3666f9c7390bbe17b10b404e4627189135aa174246a

Observation c8133683-a472-459c-9c0b-c0c7c3506d3d · outbound

This paper cites A multimodal robust recognition method for grasping objects with robot flexible grippers,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System A multimodal robust recognition method for grasping objects with robot flexible grippers,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.380767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:46.242550Z digest=sha256:fdecd45bd3b9c403a960a52d5c7feb31012bf1a90d4f4a36b34adc83c839a3bd

Observation 8dea092c-2df3-4ec7-8a5e-43745a7c72a0 · outbound

This paper cites Learning generalizable vision-tactile robotic grasping strategy for deformable objects via transformer,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Learning generalizable vision-tactile robotic grasping strategy for deformable objects via transformer,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:54.187612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:46.320504Z digest=sha256:bab16c7f9bc56e00c77b3e4d9c7b7c7fabfea057ba7c6a36ac7d802b6cb95af9

Observation 2e35c844-d577-4243-837b-9dfdf08756db · outbound

This paper cites Dih-tele: Dexterous in-hand teleoperation framework for learning mul- tiobjects manipulation with tactile sensing,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Dih-tele: Dexterous in-hand teleoperation framework for learning mul- tiobjects manipulation with tactile sensing,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:53.968689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:46.408258Z digest=sha256:b2fbc8e3f927e76d2d7cf98aa0619113e4678e2bf942afda4cb1970d17ac0876

Observation c2a83e1e-9fc5-4800-a7dc-7957fb7ebd44 · outbound

This paper cites Efficient grasp detection network with gaussian-based grasp representation for robotic manipulation,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Efficient grasp detection network with gaussian-based grasp representation for robotic manipulation,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:46.486156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:46.486156Z digest=sha256:d9816538493e941a32427fac1b81ab28c2a9355592785c4835953e60009e6957

Observation 1d67e1e7-5647-4986-99a7-78a4812fb4b2 · outbound

This paper cites Language conditioned imitation learning over unstructured data,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Language conditioned imitation learning over unstructured data,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:53.685780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:46.561868Z digest=sha256:c009bd7794435212eb33131232168016c0fe6e13934efc400d53fc9c87d7e126

Observation ea9e6e9d-3174-467e-a56a-82cd197b0bf3 · outbound

This paper cites What matters in language conditioned robotic imitation learning over unstructured data,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System What matters in language conditioned robotic imitation learning over unstructured data,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:46.649628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:46.649628Z digest=sha256:4d21b1eed4f9601360f68766e03f6da927c83650877ba12c23d8f039bcc91ddf

Observation b4de731c-37d6-46ed-a747-53b8704be40b · outbound

This paper cites Diffusion policy: Visuomotor policy learning via ac- tion diffusion,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Diffusion policy: Visuomotor policy learning via ac- tion diffusion,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:46.728769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:46.728769Z digest=sha256:cad646297d2558abf3418c1bb00b0b56896e012e52d780d80d322089fcd75a76

Observation 286f726c-5063-499a-80de-a221ea9ee758 · outbound

This paper cites 3d diffusion policy,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System 3d diffusion policy,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:53.460891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:46.790181Z digest=sha256:c464bc6455fe2ab49835f417543c3917ab6222f168a1c94b011a858f0ace829c

Observation d3216e62-bf3c-4d88-a14c-33c0df62c9bc · outbound

This paper cites Consistency policy: Accelerated visuomotor policies via consistency distillation,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Consistency policy: Accelerated visuomotor policies via consistency distillation,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:53.259184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:46.844026Z digest=sha256:5e8ce3758bb71421fa5db2b4ff0786231a55eed3754946f8b90c518ec090675a

Observation b0c7a529-c8b4-4042-90cb-ee5bdb5e4d72 · outbound

This paper cites RT-trajectory: Robotic task generalization via hindsight trajectory sketches,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System RT-trajectory: Robotic task generalization via hindsight trajectory sketches,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:53.060753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:46.898184Z digest=sha256:bf8418cc210200aefc2783be2e4de7e4a5662ea9a9d04c933faa951dda390b49

Observation b8533a63-2a44-4ea7-8caf-241a80272b5d · outbound

This paper cites Sara-rt: Scaling up robotics transformers with self-adaptive robust attention,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Sara-rt: Scaling up robotics transformers with self-adaptive robust attention,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:52.862970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:46.957774Z digest=sha256:f1055642a33a713df20318d70557cd9f68392d10b81e9e27c62f54c2e8dd2c3b

Observation 50e6b248-cc1c-436b-81e9-660457f5e1dc · outbound

This paper cites Inner monologue: Em- bodied reasoning through planning with language models,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Inner monologue: Em- bodied reasoning through planning with language models,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:52.691441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:47.022196Z digest=sha256:20511abc32231e10441b86da04b86618cd7d871a2525b4e71185b4b02424ed5c

Observation f9abede6-c0b4-4513-b0cb-c0f4cf76bf20 · outbound

This paper cites Unleashing large-scale video generative pre-training for visual robot manipulation,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Unleashing large-scale video generative pre-training for visual robot manipulation,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:52.427299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:47.082933Z digest=sha256:0c3c103fbf8c268894e8752f0c9c91b77a2637a72ea9deb5ed9ecbf6cf87a9a7

Observation 2eda2f3a-330a-4827-826a-0b95e2ac8715 · outbound

This paper cites Video language planning,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Video language planning,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:52.151150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:47.145939Z digest=sha256:91960a9b563bce5fce005c23e0c76c4ad8c726aa6c596da7f9a1d65fff7414b4

Observation 9c3534bc-6442-43b5-a17d-b74bc742c8bd · outbound

This paper cites Robomamba: Efficient vision-language-action model for robotic reasoning and manipulation,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Robomamba: Efficient vision-language-action model for robotic reasoning and manipulation,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:52.039687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:47.193783Z digest=sha256:d075dd629a3d8c5270b0b58990046dde7817d2b8539f66e250fdd392d7a928ad

Observation 2a951cf5-e9f0-4ce4-bf9b-c545e4f26e67 · outbound

This paper cites Vlas: Vision-language-action model with speech instructions for customized robot manipulation,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Vlas: Vision-language-action model with speech instructions for customized robot manipulation,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:51.882889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:47.248717Z digest=sha256:652c80f743aa7ec0534817b8ada59d1b480ef206ed71fa226e5967e630e808d6

Observation cd995e3c-fb3a-43e7-a78a-355efaaefe90 · outbound

This paper cites Dual-arm robotic fabric manipulation with quasi-static and dynamic primitives for rapid garment flattening,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Dual-arm robotic fabric manipulation with quasi-static and dynamic primitives for rapid garment flattening,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:51.696294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:47.317109Z digest=sha256:b8ffb37ed485b978eab2153439b1bfa5b059f84b7e78c50121fa6b9be5473323

Observation cf663e19-9a37-4eee-b411-84fde9e53b8c · outbound

This paper cites From llms to actions: Latent codes as bridges in hierarchical robot control,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System From llms to actions: Latent codes as bridges in hierarchical robot control,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:51.511411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:47.373690Z digest=sha256:eced3ab1de4c39e8d54909d4e05ad083008565fe1ce936f02ed0a2fa110836d5

Observation 0c424af0-ad07-4135-855e-f574d34dd158 · outbound

This paper cites Hirt: Enhancing robotic control with hierarchical robot transformers,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Hirt: Enhancing robotic control with hierarchical robot transformers,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:51.318105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:47.438243Z digest=sha256:7e7695595d31b89ea2e48bf658aef94b07ef9ca60407e79012ab0e0add475bce

Observation de8c6ffc-a6fe-4607-b527-41a9ad8022d9 · outbound

This paper cites Towards synergistic, generalized, and efficient dual-system for robotic manipulation,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Towards synergistic, generalized, and efficient dual-system for robotic manipulation,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:51.113011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:47.485857Z digest=sha256:8f02f0bcc24f376058ff620df63e11b30625e47b0d9c190b1f22ed0ddeaf2655

Observation 095963b2-6f1c-4e10-bcb7-0c9684fb567d · outbound

This paper cites A Dual Process VLA: Efficient Robotic Manipulation Leveraging VLM.

RationalVLA: A Rational Vision-Language-Action Model with Dual System A Dual Process VLA: Efficient Robotic Manipulation Leveraging VLM

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:47.538512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:47.538512Z digest=sha256:48f8b8790172ce5704f3fa0a242f694a552c83a1feef6c544cce7c18130fa9cb

Observation 8132137a-43a1-49da-b5f8-88d0bfcb8855 · outbound

This paper cites Openvla: An open-source vision-language-action model,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Openvla: An open-source vision-language-action model,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:50.939829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:47.589923Z digest=sha256:29f6f25b8e3ccecde7306185b3b657c5a2c59c311d6e8b105d1967a5688e8e6d

Observation 6d5ba01f-60b7-4566-a5f4-f9693fc67b41 · outbound

This paper cites Gr00t n1: An open foundation model for generalist humanoid robots,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Gr00t n1: An open foundation model for generalist humanoid robots,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:50.721443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:47.631099Z digest=sha256:9ff28f9e96c2b27a267d564d6c6fbfdc5c2dc7bdc0e62138907189a8e9b2b0bd

Observation 7ac11666-e4db-4399-898c-10e4ab4007ec · outbound

This paper cites OpenHelix: A Short Survey, Empirical Analysis, and Open-Source Dual-System VLA Model for Robotic Manipulation.

RationalVLA: A Rational Vision-Language-Action Model with Dual System OpenHelix: A Short Survey, Empirical Analysis, and Open-Source Dual-System VLA Model for Robotic Manipulation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:47.684400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:47.684400Z digest=sha256:cf34d6f5b780ed551e0fb19dd9d2a74dd56fdb653fa78fe075991b3f9230f21e

Observation 55622c64-f47f-4ad1-8a5a-7fecc9d9c88b · outbound

This paper cites Hierarchical reinforcement learning with model guidance for mobile manipulation,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Hierarchical reinforcement learning with model guidance for mobile manipulation,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:50.469853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:47.745003Z digest=sha256:0bb76a33af49a5494c2cbbe1062570ff1fc4ebe8457cc4e94ce3836e4ff59846

Observation 797b14a9-7840-4dba-ace9-80fa9de8d5a5 · outbound

This paper cites Interactive imitation learning of bimanual movement primitives,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Interactive imitation learning of bimanual movement primitives,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:50.246959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:47.802533Z digest=sha256:d6dd5ae4085665082ef977f216e6d6fafa7d5368186ba02b3ae60e4a89097247

Observation 75f51d1f-f10b-4d7d-a507-1f7565cbca43 · outbound

This paper cites Navigating beyond in- structions: Vision-and-language navigation in obstructed environments,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Navigating beyond in- structions: Vision-and-language navigation in obstructed environments,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:50.037859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:47.872273Z digest=sha256:b1139489dc81199098aad608fd8b1c23d79c6b7e4cac02d3164c893b9ce58862

Observation dd672fc8-f90e-4b37-a440-925b000f4f7f · outbound

This paper cites BadNAVer: Exploring Jailbreak Attacks On Vision-and-Language Navigation.

RationalVLA: A Rational Vision-Language-Action Model with Dual System BadNAVer: Exploring Jailbreak Attacks On Vision-and-Language Navigation

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:47.924165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:47.924165Z digest=sha256:ca58999d6cc05471b2f3a5f23b11c46bffdc037caee270cf9458e5437a932083

Observation 23b0c067-e096-479b-8d13-fe55ba550797 · outbound

This paper cites Safety bounds in human robot interaction: A survey,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Safety bounds in human robot interaction: A survey,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:49.819340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:47.963015Z digest=sha256:0e395af8217d0026442b778efd941c6bd69696f5670520dbe383527504a4d17a

Observation 3d25a37a-ca7d-463e-953a-df8b123260e3 · outbound

This paper cites Gpt-4 technical report,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Gpt-4 technical report,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:48.002210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:48.002210Z digest=sha256:d48bec0aabfac917c137d32f2b77415411642d6f64fcb6c243dfd66f87547d17

Observation c5238311-00ae-44b5-a152-7a29b06dce08 · outbound

This paper cites Pybullet, a python module for physics simulation for games, robotics and machine learning,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Pybullet, a python module for physics simulation for games, robotics and machine learning,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:49.607146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:48.076903Z digest=sha256:056eb57b6937dc98fedff646962350450e0a8f106ea21ca85a6c676704ce19c1

Observation 2f4c3a85-1148-411e-a693-e5cf933c4902 · outbound

This paper cites LoRA: Low-rank adaptation of large language models,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System LoRA: Low-rank adaptation of large language models,

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:48.225606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:48.225606Z digest=sha256:5aedc15f7a14f513088722646f27d121581726ded3de1eb57021cc5aa20a22d0

Observation d08f8b2b-a536-4225-9c91-cd89043e73a5 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Learning transferable visual models from natural language supervision,

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:48.386974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:48.386974Z digest=sha256:19eaec7457116bd9258e556a70b549ca7d52205f9206449d6dad65b801d52d3b

Observation b6a9d7b3-7121-4f33-80c7-924bcbd4d6ea · outbound

This paper cites Investigating the catastrophic forgetting in multimodal large language models,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Investigating the catastrophic forgetting in multimodal large language models,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:49.387945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:48.482359Z digest=sha256:d9e22b8a36d9b80db9465058bdb231463b660536fbfa15db37b313453a6f0c02

Observation 7bfc2d50-168d-4edf-9806-fad476749dbf · outbound

This paper cites Allava: Harnessing gpt4v-synthesized data for lite vision-language models,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Allava: Harnessing gpt4v-synthesized data for lite vision-language models,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:49.134933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:48.559986Z digest=sha256:1cb16bd1592fdfe13b3294a72013c648ce3fb50b2a4425ea51e3c23d1b09a94e

Observation 87327f2a-f89b-43fb-814e-4764dd3287a2 · outbound

This paper cites Llama: Open and efficient foundation language models,.

RationalVLA: A Rational Vision-Language-Action Model with Dual System Llama: Open and efficient foundation language models,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:24:48.954179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:24:48.618223Z digest=sha256:6a0d58d3d508527833ee3205e802f523385e80e338932b6809e1362e7cf1ff3c

Pith citing papers

Observation aa00de08-9610-42d6-9bd4-d8f86cb31bc2 · inbound

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver cites this paper.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver RationalVLA: A Rational Vision-Language-Action Model with Dual System

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:46.126998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:46.126998Z digest=sha256:7606201e1555677c0d5f620f7a222aa1641686a63b30e519e164535be34375b7

Observation 393d89ca-2901-43e6-b271-35aaeadcf5f6 · inbound

Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey cites this paper.

Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey RationalVLA: A Rational Vision-Language-Action Model with Dual System

Reference 135

Resolution
verified exact
arxiv_id, observed 2026-05-17T20:28:16.364814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T20:28:15.818016Z digest=sha256:627c49a5a880efc5e70039b953862c90eaa0e83b4065303b6289e5010f0ba8c2

Observation 790f1a97-6d41-44c6-ad6f-6be6bc7ff1da · inbound

RoboMemArena: A Comprehensive and Challenging Robotic Memory Benchmark cites this paper.

RoboMemArena: A Comprehensive and Challenging Robotic Memory Benchmark RationalVLA: A Rational Vision-Language-Action Model with Dual System

Reference 33

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T06:56:26.744197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-12T03:50:18.706396Z digest=sha256:339629c6e3a29b2ba94ff0c917a75b6d16fa3056ccb0288a4e0a7098594bff26

Observation 88b85e26-b952-4e1d-b7d9-cadf72149d98 · inbound

Dual-Process Atomic Skill Learning: Decoupling Semantic Reasoning and Real-Time Control cites this paper.

Dual-Process Atomic Skill Learning: Decoupling Semantic Reasoning and Real-Time Control RationalVLA: A Rational Vision-Language-Action Model with Dual System

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-14T10:23:25.158961Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T10:23:25.158961Z digest=sha256:e89738d689dbe2092ac2c1aeae1f140eccf3eeb554bc883c6c14ecdde4095668

Observation 6512388d-82c6-4906-865d-8e13e2cf7f6e · inbound

Token-Wise Latent Streaming from Slow Reasoners to Fast Planners for Dynamic Vision Language Navigation cites this paper.

Token-Wise Latent Streaming from Slow Reasoners to Fast Planners for Dynamic Vision Language Navigation RationalVLA: A Rational Vision-Language-Action Model with Dual System

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-01T19:56:57.341989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T19:56:57.341989Z digest=sha256:e4aee4879322fe54d5d398af995ac7dc3c4869dfb18f4fdad64de18e25c1cf48