Pith. sign in

Paper Citation Record · LEDGER

Training Strategies for Efficient Embodied Reasoning

As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 22 inbound Pith citation observations for arXiv:2505.08243.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.08243 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 22 of 22 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T15:41:31.454469Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T13:37:06.859851Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 20ff8100-4ca7-47bb-9441-8c965fb14dc3 · inbound

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions cites this paper.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions Training Strategies for Efficient Embodied Reasoning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:01.209286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:01.209286Z digest=sha256:966c72fdfbbf9f3a73d1b18c0198bb6e99b7bc687945d97d8fd6b8468ffe3f87

Observation 451320f0-dd92-45fd-8e85-407168da5887 · inbound

Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey cites this paper.

Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey Training Strategies for Efficient Embodied Reasoning

Reference 95

Resolution
verified exact
arxiv_id, observed 2026-05-17T20:28:15.961055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-17T20:28:15.818016Z digest=sha256:16026d57d91705484c7059854026450f020cf69e2a7fb4766ae8fb50077a3832

Observation fb8a56cb-ec2d-4e5c-9c00-4f399dfe0038 · inbound

SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning cites this paper.

SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning Training Strategies for Efficient Embodied Reasoning

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-15T08:02:11.447034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-15T08:02:11.189795Z digest=sha256:01ff9a266d257bd23a063bb09fe05cbfd166632410408fe8f68fb2754fbab49e

Observation 81caf4b3-5bd4-4bee-bbad-5166d924e395 · inbound

Efficient Vision-Language-Action Models for Embodied Manipulation: A Systematic Survey cites this paper.

Efficient Vision-Language-Action Models for Embodied Manipulation: A Systematic Survey Training Strategies for Efficient Embodied Reasoning

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-04T09:08:16.019521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:08:16.019521Z digest=sha256:890e97ef4e1944fed871987f240ada8b875859f9cc8f7ad9772377dac2cacc8f

Observation 0398b2b6-653d-4d09-91cc-bf5076335373 · inbound

DeepThinkVLA: Enhancing Reasoning Capability of Vision-Language-Action Models cites this paper.

DeepThinkVLA: Enhancing Reasoning Capability of Vision-Language-Action Models Training Strategies for Efficient Embodied Reasoning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-18T03:10:48.972174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-05-18T03:09:09.713822Z digest=sha256:c6433a8b331daaec0bbab7c44bc06109ea938497fcf2b5743c29e20e241a1586

Observation 1ada9042-7d33-4288-813d-4d06bb32cfa0 · inbound

EVE: A Generator-Verifier System for Generative Policies cites this paper.

EVE: A Generator-Verifier System for Generative Policies Training Strategies for Efficient Embodied Reasoning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T14:10:34.393044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T14:10:34.393044Z digest=sha256:bb4ca46eb0a48d2c5fe9ec3f229c2bf81e6be5651b068d499144f189b3faa1ab

Observation 2bd83c1f-98b4-4efd-854a-a4c440649493 · inbound

Unified Embodied VLM Reasoning with Robotic Action via Autoregressive Discretized Pre-training cites this paper.

Unified Embodied VLM Reasoning with Robotic Action via Autoregressive Discretized Pre-training Training Strategies for Efficient Embodied Reasoning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-03T13:29:51.070384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:29:51.070384Z digest=sha256:af07608f0d989d9eea7a461a3dcf958db5f49703254406966d18806feb65c042

Observation 94e60e04-57d3-4790-97b1-c7dcb6b53174 · inbound

PoseVLA: Universal Pose Pretraining for Generalizable Vision-Language-Action Policies cites this paper.

PoseVLA: Universal Pose Pretraining for Generalizable Vision-Language-Action Policies Training Strategies for Efficient Embodied Reasoning

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-21T13:05:10.025747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-21T13:04:30.544504Z digest=sha256:90c62eb6c89e0d4ad8295221e2388809e3f35f1e227bcb04a45fc6fd4c34d153

Observation 695d346a-885e-4bbe-ab8f-f360d97c2d60 · inbound

PoseVLA: Universal Pose Pretraining for Generalizable Vision-Language-Action Policies cites this paper.

PoseVLA: Universal Pose Pretraining for Generalizable Vision-Language-Action Policies Training Strategies for Efficient Embodied Reasoning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T21:37:10.787322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:37:10.787322Z digest=sha256:0750b86e7bba7272aaf4a800fbe8461f0116124d3f18659dbb256fbf81993432

Observation 0f321d94-db77-45cb-9aec-8588946b1c8d · inbound

VP-VLA: Visual Prompting as an Interface for Vision-Language-Action Models cites this paper.

VP-VLA: Visual Prompting as an Interface for Vision-Language-Action Models Training Strategies for Efficient Embodied Reasoning

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:59:36.782735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-15T00:54:36.125258Z digest=sha256:a4ebbc2eab9810999364b227e9c2788be0751bfed21076f36b9e3b36c87a2d80

Observation cd3813ec-440b-42c3-87be-a2e62c9aabc0 · inbound

TRAP: Hijacking VLA CoT-Reasoning via Adversarial Patches cites this paper.

TRAP: Hijacking VLA CoT-Reasoning via Adversarial Patches Training Strategies for Efficient Embodied Reasoning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-13T19:50:50.950451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T19:50:50.950451Z digest=sha256:54b5b86b57d8c609bc65abba8fb3dcdca74e22df630483d18bc53ea3cff1fc9c

Observation 6cf40e9b-bdcf-4687-a81d-a531c9c364a8 · inbound

Environmental Understanding Vision-Language Model for Embodied Agent cites this paper.

Environmental Understanding Vision-Language Model for Embodied Agent Training Strategies for Efficient Embodied Reasoning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:26:11.056794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T03:36:52.685696Z digest=sha256:3709ffcd8053b58d203ded3feb42c1d0ce7b775e6bc29436232fab58408380c2

Observation e09b5153-3e2c-425c-9c4d-ded9a0310cbc · inbound

Beyond Thinking: Imagining in 360$^\circ$ for Humanoid Visual Search cites this paper.

Beyond Thinking: Imagining in 360$^\circ$ for Humanoid Visual Search Training Strategies for Efficient Embodied Reasoning

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T03:01:18.344630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-05-12T02:58:46.728868Z digest=sha256:a9ce24b74d7fb27374653eaaea7e466a048dbb876d37b7d049003330833c56cf

Observation 75ba3d46-16ff-40f2-a3ca-28d06292c94e · inbound

Revisiting Embodied Chain-of-Thought for Generalizable Robot Manipulation cites this paper.

Revisiting Embodied Chain-of-Thought for Generalizable Robot Manipulation Training Strategies for Efficient Embodied Reasoning

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-07-02T03:46:33.301446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-28T09:40:04.685274Z digest=sha256:418d5b5169a2a2f9de734ccf3185bfe3db840f87a27b6c382bdab53aeb6f31a7

Observation f2b8a074-1252-419e-bbfa-09b140029a78 · inbound

DIRECT: When and Where Should You Allocate Test-Time Compute in Embodied Planners? cites this paper.

DIRECT: When and Where Should You Allocate Test-Time Compute in Embodied Planners? Training Strategies for Efficient Embodied Reasoning

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-07-03T11:08:03.074252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-27T09:42:17.813126Z digest=sha256:c4312042bc7dbe95f33a2310654418328b9bbd70ea613e16f204132ddcc7c486

Observation faf1eda9-83a1-44b8-9392-2128b4d17f90 · inbound

Training Vision-Language-Action Models with Dense Embodied Chain-of-Thought Supervision cites this paper.

Training Vision-Language-Action Models with Dense Embodied Chain-of-Thought Supervision Training Strategies for Efficient Embodied Reasoning

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-06-30T15:44:48.962803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-30T05:10:51.005004Z digest=sha256:3651589769c97f1adddee53cbf7e2f3254fdf519668d8539aa2d571c00fcd537

Observation 8f5e6263-7b22-4862-b0d9-716889209cef · inbound

Training Vision-Language-Action Models with Dense Embodied Chain-of-Thought Supervision cites this paper.

Training Vision-Language-Action Models with Dense Embodied Chain-of-Thought Supervision Training Strategies for Efficient Embodied Reasoning

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:37:22.353552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-02T20:29:13.282030Z digest=sha256:1d5704c126ee081010831954c6b1e79f84708285f2c6433dd1905f12c682267a

Observation 9f03ab37-eb46-4b44-8256-c570f73bfebb · inbound

Do Vision-Language-Action Models Mean What They Say? On the Role of Faithfulness in Embodied Reasoning cites this paper.

Do Vision-Language-Action Models Mean What They Say? On the Role of Faithfulness in Embodied Reasoning Training Strategies for Efficient Embodied Reasoning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-11T15:16:38.426751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T15:16:38.426751Z digest=sha256:71a65b0f097ec530f6bad738ef4494ff43beba668699e897cf81b4627af21ed3

Observation a185804e-cf60-4653-b464-ed95f389c415 · inbound

APIVOT: Adaptive Planning with Interleaved Vision-Language Thoughts cites this paper.

APIVOT: Adaptive Planning with Interleaved Vision-Language Thoughts Training Strategies for Efficient Embodied Reasoning

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-07-10T13:37:06.861034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-10T13:33:17.574108Z digest=sha256:91ca4f612bc6efe0db67a36fdf9fb3ec550b70c3a78262c94dddef63c9dfa897

Observation 400c7946-e052-4d43-89aa-478c300bdf5e · inbound

Xiaomi-Robotics-1: Scaling Vision-Language-Action Models with over 100K Hours of Real-World Trajectories cites this paper.

Xiaomi-Robotics-1: Scaling Vision-Language-Action Models with over 100K Hours of Real-World Trajectories Training Strategies for Efficient Embodied Reasoning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T00:04:00.468016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:04:00.468016Z digest=sha256:b695ebc9fcf27a67fb0b50498fa0931929f951abcaa8ce54e286c141ef9e9f8b

Observation 284bca53-bfe9-40f6-b039-1ce943b45667 · inbound

Reasoning as a Double-Edged Sword: Architecture and Cross-Stage Robustness in Vision-Language-Action Models cites this paper.

Reasoning as a Double-Edged Sword: Architecture and Cross-Stage Robustness in Vision-Language-Action Models Training Strategies for Efficient Embodied Reasoning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T15:41:31.454469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:41:31.454469Z digest=sha256:0ff780203e0041fa4eec26bb06db7813dec08bcf73eef98e4a01b109f65e4897

Observation 6c9638e6-7219-48bb-b32e-acdb61da296a · inbound

RoboInter1.5: A Holistic Intermediate Representation Suite for Embodied World Modeling and Robotic Manipulation cites this paper.

RoboInter1.5: A Holistic Intermediate Representation Suite for Embodied World Modeling and Robotic Manipulation Training Strategies for Efficient Embodied Reasoning

Reference 126

Resolution
unresolved
no resolver link, observed 2026-08-01T14:39:46.327139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T14:39:46.327139Z digest=sha256:16dde9ef9936f020a780498b9ab78bbb4aacd0860af51999312a8d25b062f04c