Pith. sign in

Paper Citation Record · LEDGER

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use

As of 9 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 0 inbound Pith citation observations for arXiv:2608.05738.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.05738 v1

Coverage vector

measured 46 of 46 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T00:30:03.258938Z

measured 46 of 46 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

46 of 46 outbound references displayed

  • verified exact0
  • verified fuzzy9
  • unresolved37
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f7eb6eb6-0e15-47ef-8527-cdc713ce9b0c · outbound

This paper cites IEEE Transactions on Neural Networks and Learning Systems , year=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use IEEE Transactions on Neural Networks and Learning Systems , year=

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:30:04.570439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:30:03.006647Z digest=sha256:603761105a5bc1921fe49aa1e8e83f64a3ae47894276b63b85588dbdaddf6528

Observation 987ca5b0-57c6-471e-b0e4-7f1e53de1289 · outbound

This paper cites JoyAI-RA 0.1: A Foundation Model for Robotic Autonomy.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use JoyAI-RA 0.1: A Foundation Model for Robotic Autonomy

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.014525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.014525Z digest=sha256:e9e51be939c19cf3ee6b34edb7f345954fc2af4f90c58089c539a27a556a6588

Observation fa45f4ed-4fb9-453b-b8b6-afb419116e19 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use OpenVLA: An Open-Source Vision-Language-Action Model

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.020883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.020883Z digest=sha256:d61d58f30ff7c44d4a0c8cad9afdd020d87ad2530cd2b11adfa86beb56dd4e0d

Observation 7082172c-ce99-4768-a7f8-66ccfcb83440 · outbound

This paper cites Proceedings of the 60th annual meeting of the association for computational linguistics (volume 1: long papers) , pages=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Proceedings of the 60th annual meeting of the association for computational linguistics (volume 1: long papers) , pages=

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.028302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.028302Z digest=sha256:0d90e706e0b73e4870b78d292e5f8fd954e0bb7c5bdd63fd15d19499b22c8a5c

Observation 4ca46f22-ed91-4dab-a9d0-b77ea41e4422 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Advances in Neural Information Processing Systems , volume=

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.034197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.034197Z digest=sha256:c4b0edf1fcc5afe2b399df232c856f4dd42936094aadd07e8e42c89f41c21dd9

Observation e6bbec38-f956-48dc-b36b-dec14083be3a · outbound

This paper cites Do As I Can, Not As I Say: Grounding Language in Robotic Affordances.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Do As I Can, Not As I Say: Grounding Language in Robotic Affordances

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.040411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.040411Z digest=sha256:989d6dcc4a2f9d0ea5923f83cfb2add37eec92c3229a30dbe254ab9fb6d9759a

Observation 183a0138-3ca1-4870-b752-101aba51a2c4 · outbound

This paper cites 2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , pages=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use 2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , pages=

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.047791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.047791Z digest=sha256:183d157d94a10d5b926ea072ea4464ce586eb9da06ae9515449391734844f5c2

Observation 2ecdc0a2-7efb-4598-a93b-4def83e45495 · outbound

This paper cites VisualToolAgent (VisTA): A Reinforcement Learning Framework for Visual Tool Selection.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use VisualToolAgent (VisTA): A Reinforcement Learning Framework for Visual Tool Selection

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.053598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.053598Z digest=sha256:a86868e5db840618589951aec89f62539560936a6f096b706ba23186287274f8

Observation 939750d8-1c9e-4f39-83ce-3080f9eb4076 · outbound

This paper cites NeurIPS 2025 Workshop on Space in Vision, Language, and Embodied AI , year=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use NeurIPS 2025 Workshop on Space in Vision, Language, and Embodied AI , year=

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:30:04.516393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:30:03.059312Z digest=sha256:c5b7241b01214b2365e3249091a98111f73e799cf9797a800dd1572dffa4e23d

Observation 5812d0e6-efe0-49fa-8938-448723d2313a · outbound

This paper cites 9th Annual Conference on Robot Learning , year=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use 9th Annual Conference on Robot Learning , year=

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:30:04.498202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:30:03.064459Z digest=sha256:8c4e78bc7870053465e7b657bf202aaf0f97bb4edf90fd9b0fddb58b8d09e518

Observation fd67b1d6-3ee6-447b-ae0b-1e4aeee59680 · outbound

This paper cites MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.069585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.069585Z digest=sha256:e03fc21ee3474dd093522cf6ccfbde9790b3317343e3103533d732bc91ee2364

Observation d4782cff-fe70-40e2-962c-80cd24031efb · outbound

This paper cites European Conference on Computer Vision , pages=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use European Conference on Computer Vision , pages=

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:30:04.479702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:30:03.075063Z digest=sha256:9b65e30d0443b321e01f06b61f6c2c9e1e53d93b876e24b0ed7033f6cae91484

Observation 36f58bce-f1f6-4c18-adfe-381d103a2ba4 · outbound

This paper cites PaLM-E: An Embodied Multimodal Language Model.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use PaLM-E: An Embodied Multimodal Language Model

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.081177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.081177Z digest=sha256:18ef4d80e3d0f21b8825ff8e178822f69601d52b92ddcb35444a06b3ac118f92

Observation d78a3072-537f-4a59-9a36-aa028beb1cca · outbound

This paper cites VoxPoser: Composable 3D Value Maps for Robotic Manipulation with Language Models.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use VoxPoser: Composable 3D Value Maps for Robotic Manipulation with Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.086780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.086780Z digest=sha256:42c081d41696730111606e68c292513d6f53fffa3b8aa321b20bfce751498aca

Observation 4170b60c-f2a6-4a5a-93d4-35a4a86896e0 · outbound

This paper cites Inner Monologue: Embodied Reasoning through Planning with Language Models.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Inner Monologue: Embodied Reasoning through Planning with Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.092074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.092074Z digest=sha256:a85ee006525e91c25980fc9c0a0e439de2cc691717ad6c66e9e16d04896e99fa

Observation 2b43020b-bb89-4179-b01b-d18a17b3eb1f · outbound

This paper cites SpaceTools: Tool-Augmented Spatial Reasoning via Double Interactive RL.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use SpaceTools: Tool-Augmented Spatial Reasoning via Double Interactive RL

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.097660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.097660Z digest=sha256:8ea4527473ed22d349c8e0a0206fd9caf7ae6403aa0fcef3b89944e09bf694c8

Observation 1b5ba0d2-d480-4540-a2ae-53b7041dfae8 · outbound

This paper cites Towards Long-horizon Embodied Agents with Tool-Aligned Vision-Language-Action Models.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Towards Long-horizon Embodied Agents with Tool-Aligned Vision-Language-Action Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.103238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.103238Z digest=sha256:c2a10fc9c795357e8b6106ded30767ec7e404d1a12c9e8676a9ea575b463e44e

Observation bca905dc-e752-4eab-93fa-e195838f3fb6 · outbound

This paper cites Latent Reasoning VLA: Latent Thinking and Prediction for Vision-Language-Action Models.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Latent Reasoning VLA: Latent Thinking and Prediction for Vision-Language-Action Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.108171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.108171Z digest=sha256:147de7f66dd4d68cb34a687edf1cd1fecd16ea8342aa384b55dddfe70d139d07

Observation c23256b3-56ad-49d2-be2e-0c730d7a4f1a · outbound

This paper cites Proceedings of the AAAI Conference on Artificial Intelligence , volume=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Proceedings of the AAAI Conference on Artificial Intelligence , volume=

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:30:04.461262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:30:03.113661Z digest=sha256:7850134bdbaea54fed40f4533539f32be0a5a737866dc19d2bfb3e7123ba62ee

Observation aac46958-56d6-4430-a0de-e5471dff2a8d · outbound

This paper cites arXiv preprint arXiv:2601.11404 , year=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use arXiv preprint arXiv:2601.11404 , year=

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.118707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.118707Z digest=sha256:728791ffcafc0f7351d4eb318ecc93aa3232e546e143c4416c35bfb76e87bfbc

Observation b73c236e-f001-4b65-a4e7-d42c3af9ab61 · outbound

This paper cites arXiv preprint arXiv:2603.22280 , year=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use arXiv preprint arXiv:2603.22280 , year=

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.123624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.123624Z digest=sha256:3c1f057f62a2c20784f44bc0e115a59eb2ef544cd6deed4aac2917b126f9ed75

Observation 26b69c93-ab4b-4f41-87db-c949e7e991b6 · outbound

This paper cites arXiv preprint arXiv:2603.14523 , year=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use arXiv preprint arXiv:2603.14523 , year=

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.127765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.127765Z digest=sha256:a7c246f3d4fe5f2d7ea7b7607e782cc9106cc13d6268c396a1339c42fd648661

Observation f3b75b19-ac0b-4eb3-b5f6-da0aac8eba28 · outbound

This paper cites an unresolved cited work.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.132899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.132899Z digest=sha256:edb7a4bde4226e3f2e30382523fc333c601a55b9881e354ca900eb9b3a2caf3a

Observation 7c618703-eda2-49e8-b18e-84832367434d · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.137263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.137263Z digest=sha256:4388a9ef025bdf9c109f9e16d54316238de2c7e94368820756149bd3ac6b17e7

Observation d394e4ff-3465-491a-8843-6d3bfcb0ddf8 · outbound

This paper cites InternVLA-M1: A Spatially Guided Vision-Language-Action Framework for Generalist Robot Policy.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use InternVLA-M1: A Spatially Guided Vision-Language-Action Framework for Generalist Robot Policy

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.142726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.142726Z digest=sha256:68e76409ec08ad82c935337e8b41fa3e8ebc83bf153626bb6fbad33452207502

Observation b95ac65d-b420-4422-8ce7-55c7859b323a · outbound

This paper cites Forty-third International Conference on Machine Learning , year=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Forty-third International Conference on Machine Learning , year=

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:30:04.433143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:30:03.147823Z digest=sha256:b7ead09460df812aba1bc0211429b8b238e934a529e9a5f0296c9d0a94782a6e

Observation 37f111f6-d03b-47dc-8e04-8cffaddf19ad · outbound

This paper cites arXiv preprint arXiv:2602.10098 , year=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use arXiv preprint arXiv:2602.10098 , year=

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.153249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.153249Z digest=sha256:4dee804dd8a09884c36022fb5b7a4b92a5a67cc62f9725ed7fb380834952e131

Observation 0bb6acc2-4fcb-4a27-b008-8d0f96ce7529 · outbound

This paper cites Advances in neural information processing systems , volume=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Advances in neural information processing systems , volume=

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:30:04.417001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:30:03.158542Z digest=sha256:4dd0ebf9ee6d401bee79e9f7e4dcedd823df6cac8cdadc984639a426c11d7ec8

Observation 9605a5a1-b880-4e73-9d8b-a298bf1a501a · outbound

This paper cites CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.164482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.164482Z digest=sha256:1a8c2fff49268c1e5755b7e0b604292390bf3631274a754593d3f18cf25d5cd1

Observation 53eece55-de9e-4048-9e66-99fad28bd8ea · outbound

This paper cites ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.169903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.169903Z digest=sha256:bc931a9a61a65489e75e14f9d87eec10ace7b8a480fd9727695c8fd64e12a4bd

Observation 9bab54ad-2aea-4e55-8cc6-935a0fff07b2 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Advances in Neural Information Processing Systems , volume=

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:30:04.401760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:30:03.176272Z digest=sha256:03b98c6264bc07f77dfc8da245299a3e9d4b8baea03ee9f1095ba37e43e0a663

Observation d26d14f8-be1c-49df-a7d1-c0eb87eb3b33 · outbound

This paper cites arXiv preprint arXiv:2512.16793 , year=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use arXiv preprint arXiv:2512.16793 , year=

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.181620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.181620Z digest=sha256:0a913e3551280fed5f6585a95c8671af18fb793678d1b587c1c65006ebbba288

Observation d3eac13e-9ff4-4a6e-83a7-049686782f5b · outbound

This paper cites arXiv preprint arXiv:2601.14133 , year=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use arXiv preprint arXiv:2601.14133 , year=

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.187460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.187460Z digest=sha256:e8a5dbe3a735f7b50a8223ef4164cd7feee86c5c970e8374d9ea2c29b150b7dc

Observation b31fc78e-0ab8-413d-a379-dc434b8a0e3f · outbound

This paper cites F1: A Vision-Language-Action Model Bridging Understanding and Generation to Actions.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use F1: A Vision-Language-Action Model Bridging Understanding and Generation to Actions

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.192785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.192785Z digest=sha256:c5ddb225727acb85d53d966bb4a342430ed50a28f736b3807d3500ab868b4b2d

Observation 03e84617-c671-4265-959b-e2c0102bf50e · outbound

This paper cites SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.200006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.200006Z digest=sha256:3cdd5f2193b142100a969572d7ee9e12bc2c63f94f790f10fc836ae186c63a11

Observation cac90b40-68a7-4dc2-8389-5d728c22452b · outbound

This paper cites The International Journal of Robotics Research , volume=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use The International Journal of Robotics Research , volume=

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.205571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.205571Z digest=sha256:20e5cee35dcc8436805f871f86c0829f59c484a7c086f3c05f3f114741b67e03

Observation 10af046f-e2f2-46c3-9b49-32ac84ddabd1 · outbound

This paper cites AgiBot World Colosseo: A Large-scale Manipulation Platform for Scalable and Intelligent Embodied Systems.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use AgiBot World Colosseo: A Large-scale Manipulation Platform for Scalable and Intelligent Embodied Systems

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.210982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.210982Z digest=sha256:7381687d7fdfd2e1ea6bd264a193192bc363bb6fe1e5e16928705e88df6912b0

Observation fcabc654-5e06-4509-8f56-234e2db5dbec · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.216216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.216216Z digest=sha256:825141fa82f31da3e22ca1bee1d75d909d66fadf5a15df745f7f2f6da45d7930

Observation 2eae2f62-71b2-429d-8268-38737a1f2000 · outbound

This paper cites European conference on computer vision , pages=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use European conference on computer vision , pages=

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.221703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.221703Z digest=sha256:b8c80c2a527c024d81fbac18f07912730b2fac210d980f3966ec640ad960e679

Observation 30abe21b-8325-4def-a671-800650bead3f · outbound

This paper cites Evaluating Real-World Robot Manipulation Policies in Simulation.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Evaluating Real-World Robot Manipulation Policies in Simulation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.227603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.227603Z digest=sha256:b24e16ef95cdfd61f54cd4858ff38d8043db92bc38395b98f36dab16112be3c5

Observation 795120de-04e2-44e3-b023-8d31416b2d23 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.233497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.233497Z digest=sha256:bcb8acfe62f859d11e630ab861471b85a5895c927ea98401410c8e73ce8970ab

Observation d74d6ebf-ad62-4be2-b055-f1c89be2d433 · outbound

This paper cites GR00T N1: An Open Foundation Model for Generalist Humanoid Robots.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use GR00T N1: An Open Foundation Model for Generalist Humanoid Robots

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.239113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.239113Z digest=sha256:7ed7621c1c49fa3f02f9be9e3e02c7e3604d085afe2e35fd08c7ee5da89fa74a

Observation b0d84f58-b6ef-4187-afb3-b069aafdbe19 · outbound

This paper cites Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:30:04.347498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:30:03.244179Z digest=sha256:58689369df383077c8f2dd142287dca37b584b7ece38f4fd1f264b814ce5bbf7

Observation 4ac47c81-6e3d-4599-a3af-04d185bdae5f · outbound

This paper cites Advances in neural information processing systems , volume=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Advances in neural information processing systems , volume=

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.249206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.249206Z digest=sha256:03592aed44435e0504727282e63502043e56743dbfc7542435639678eb32cb29

Observation 973885bc-d72e-4e38-8ea7-39073732b0bd · outbound

This paper cites Advances in neural information processing systems , volume=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Advances in neural information processing systems , volume=

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.253931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.253931Z digest=sha256:02d1c167666690ef5410183a128b3786da3098880435a54bd4343145437ae550

Observation e9ec2099-f2a6-48dd-a2fa-c8b9eee4228f · outbound

This paper cites arXiv preprint arXiv:2602.01067 , year=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use arXiv preprint arXiv:2602.01067 , year=

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.258938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.258938Z digest=sha256:92c477b6fa39af1c30cde4adc759e58723c7e68e5fd91bc9f5d611d6348b9c4c

Pith citing papers

No inbound Pith citation observations are available.