Pith. sign in

Paper Citation Record · LEDGER

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use

As of 15 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 0 inbound Pith citation observations for arXiv:2608.05738.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.05738 v1

Coverage vector

measured 46 of 46 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T00:30:03.258938Z

measured 46 of 46 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

46 of 46 outbound references displayed

  • verified exact0
  • verified fuzzy9
  • unresolved37
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f7eb6eb6-0e15-47ef-8527-cdc713ce9b0c · outbound

This paper cites IEEE Transactions on Neural Networks and Learning Systems , year=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use IEEE Transactions on Neural Networks and Learning Systems , year=

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:30:04.570439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-08T00:30:03.006647Z digest=sha256:96d61897b46125baa22ff6fe32fdacee9b06c67027c3a79d338ace65c76b7487

Observation 987ca5b0-57c6-471e-b0e4-7f1e53de1289 · outbound

This paper cites JoyAI-RA 0.1: A Foundation Model for Robotic Autonomy.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use JoyAI-RA 0.1: A Foundation Model for Robotic Autonomy

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.014525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.014525Z digest=sha256:e42768f8f5c5006a90bd95a6a40bb7721a8b42078ea82f5b6dbeab409d120607

Observation fa45f4ed-4fb9-453b-b8b6-afb419116e19 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use OpenVLA: An Open-Source Vision-Language-Action Model

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.020883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.020883Z digest=sha256:603ef61910ccb8aa2f77ee184e09a24dac97ef5e9654503e3cb40a26be718030

Observation 7082172c-ce99-4768-a7f8-66ccfcb83440 · outbound

This paper cites Proceedings of the 60th annual meeting of the association for computational linguistics (volume 1: long papers) , pages=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Proceedings of the 60th annual meeting of the association for computational linguistics (volume 1: long papers) , pages=

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.028302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.028302Z digest=sha256:358d81bfd11381e6d75c1c0bc68b3d1114bb21b2425453a054782b2a60fdb8bd

Observation 4ca46f22-ed91-4dab-a9d0-b77ea41e4422 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Advances in Neural Information Processing Systems , volume=

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.034197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.034197Z digest=sha256:73ffd85acad26ff60c6ce15ae49c3057f5a393fb1a564b660906936fdfa0439f

Observation e6bbec38-f956-48dc-b36b-dec14083be3a · outbound

This paper cites Do As I Can, Not As I Say: Grounding Language in Robotic Affordances.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Do As I Can, Not As I Say: Grounding Language in Robotic Affordances

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.040411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.040411Z digest=sha256:8b4c722159e3dfd8148361bd5162b86797dcbdbade431f7bd847060717a2b06f

Observation 183a0138-3ca1-4870-b752-101aba51a2c4 · outbound

This paper cites 2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , pages=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use 2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , pages=

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.047791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.047791Z digest=sha256:fb9e1eb42f185328d4fa38590f3895110af422aefb9055ac173a8918b4252598

Observation 2ecdc0a2-7efb-4598-a93b-4def83e45495 · outbound

This paper cites VisualToolAgent (VisTA): A Reinforcement Learning Framework for Visual Tool Selection.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use VisualToolAgent (VisTA): A Reinforcement Learning Framework for Visual Tool Selection

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.053598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.053598Z digest=sha256:cdca55222b3e869de81de194f10487d71e27f1d4525562177561d35214b5ea84

Observation 939750d8-1c9e-4f39-83ce-3080f9eb4076 · outbound

This paper cites NeurIPS 2025 Workshop on Space in Vision, Language, and Embodied AI , year=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use NeurIPS 2025 Workshop on Space in Vision, Language, and Embodied AI , year=

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:30:04.516393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-08T00:30:03.059312Z digest=sha256:7263e968d814393d7fe36fc2d4ed1d4083e2f274760109871f7f659c2a927e3c

Observation 5812d0e6-efe0-49fa-8938-448723d2313a · outbound

This paper cites 9th Annual Conference on Robot Learning , year=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use 9th Annual Conference on Robot Learning , year=

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:30:04.498202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-08T00:30:03.064459Z digest=sha256:2d4ff19eb71beb347de7f261d604114668854f64927234da26759738fc2cbd11

Observation fd67b1d6-3ee6-447b-ae0b-1e4aeee59680 · outbound

This paper cites MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.069585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.069585Z digest=sha256:43a0b1abdf76a093af09dbe18031fb19034927ab3eac0e4b0cf4ba9e1699641f

Observation d4782cff-fe70-40e2-962c-80cd24031efb · outbound

This paper cites European Conference on Computer Vision , pages=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use European Conference on Computer Vision , pages=

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:30:04.479702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-08T00:30:03.075063Z digest=sha256:b54e40864178b94f2f709edcd390bd1e87639c8605201645ae9a511188ad2479

Observation 36f58bce-f1f6-4c18-adfe-381d103a2ba4 · outbound

This paper cites PaLM-E: An Embodied Multimodal Language Model.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use PaLM-E: An Embodied Multimodal Language Model

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.081177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.081177Z digest=sha256:19ed379791c680cb85d0bb4aaacaf90c007342d4e19fd7186fc4c066f46d6f59

Observation d78a3072-537f-4a59-9a36-aa028beb1cca · outbound

This paper cites VoxPoser: Composable 3D Value Maps for Robotic Manipulation with Language Models.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use VoxPoser: Composable 3D Value Maps for Robotic Manipulation with Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.086780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.086780Z digest=sha256:bba363638989c060c9d03e528da6ab3f3d2b36d8bbd83df83c88b04600802f57

Observation 4170b60c-f2a6-4a5a-93d4-35a4a86896e0 · outbound

This paper cites Inner Monologue: Embodied Reasoning through Planning with Language Models.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Inner Monologue: Embodied Reasoning through Planning with Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.092074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.092074Z digest=sha256:f9476ac4bc3c4376c54db9c954f1ad3cdb3045aa9447a453c0447e58aea244cb

Observation 2b43020b-bb89-4179-b01b-d18a17b3eb1f · outbound

This paper cites SpaceTools: Tool-Augmented Spatial Reasoning via Double Interactive RL.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use SpaceTools: Tool-Augmented Spatial Reasoning via Double Interactive RL

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.097660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.097660Z digest=sha256:a0f8ab12e09c169c568cceafed11cf7560c0defbba54155f3d67b83729994a3e

Observation 1b5ba0d2-d480-4540-a2ae-53b7041dfae8 · outbound

This paper cites Towards Long-horizon Embodied Agents with Tool-Aligned Vision-Language-Action Models.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Towards Long-horizon Embodied Agents with Tool-Aligned Vision-Language-Action Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.103238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.103238Z digest=sha256:2dee85dcd1f36dfe7c4429d34dc8355759138f9eb5d2eb0df2f85fab4750e6ad

Observation bca905dc-e752-4eab-93fa-e195838f3fb6 · outbound

This paper cites Latent Reasoning VLA: Latent Thinking and Prediction for Vision-Language-Action Models.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Latent Reasoning VLA: Latent Thinking and Prediction for Vision-Language-Action Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.108171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.108171Z digest=sha256:abf6c3ebe80c2b4dfe8dae80c229b244777ef848e9b687ea3ab4558f33b1e4f2

Observation c23256b3-56ad-49d2-be2e-0c730d7a4f1a · outbound

This paper cites Proceedings of the AAAI Conference on Artificial Intelligence , volume=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Proceedings of the AAAI Conference on Artificial Intelligence , volume=

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:30:04.461262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-08T00:30:03.113661Z digest=sha256:0c39cd15c813b75731d72a26f46e4f2c6828e05dc096bdc3c962bf63fa494855

Observation aac46958-56d6-4430-a0de-e5471dff2a8d · outbound

This paper cites arXiv preprint arXiv:2601.11404 , year=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use arXiv preprint arXiv:2601.11404 , year=

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.118707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.118707Z digest=sha256:a0f1f2dc419bcae7b1033f231df1594df9f3129dde6261e621d47b77de794ac4

Observation b73c236e-f001-4b65-a4e7-d42c3af9ab61 · outbound

This paper cites arXiv preprint arXiv:2603.22280 , year=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use arXiv preprint arXiv:2603.22280 , year=

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.123624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.123624Z digest=sha256:39fdd96035410c06033b28a366e8b5a95be30395eccf5170cad07777d212246d

Observation 26b69c93-ab4b-4f41-87db-c949e7e991b6 · outbound

This paper cites arXiv preprint arXiv:2603.14523 , year=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use arXiv preprint arXiv:2603.14523 , year=

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.127765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.127765Z digest=sha256:408f9837b5b6b36d35e57919c4e9895073d1274da01e8bd436ff3e3ccf038703

Observation f3b75b19-ac0b-4eb3-b5f6-da0aac8eba28 · outbound

This paper cites an unresolved cited work.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.132899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.132899Z digest=sha256:a5c5cfbb4441eb1472f5dd8ee0083bfc6a4e218fb750e5d9df4f9cc6eb560d52

Observation 7c618703-eda2-49e8-b18e-84832367434d · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.137263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.137263Z digest=sha256:4e54f5e756c75ea117f38d3ef33766ad9c2a1a39e3b7fbaff68f3934099736ba

Observation d394e4ff-3465-491a-8843-6d3bfcb0ddf8 · outbound

This paper cites InternVLA-M1: A Spatially Guided Vision-Language-Action Framework for Generalist Robot Policy.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use InternVLA-M1: A Spatially Guided Vision-Language-Action Framework for Generalist Robot Policy

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.142726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.142726Z digest=sha256:d5bfb19d81b60702288de2de5dab5e47e245542846de7dc46be00393746fa0c4

Observation b95ac65d-b420-4422-8ce7-55c7859b323a · outbound

This paper cites Forty-third International Conference on Machine Learning , year=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Forty-third International Conference on Machine Learning , year=

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:30:04.433143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-08T00:30:03.147823Z digest=sha256:027566cc48f907aa819dac320cf50a2bdbf62b2d2fa311d0b225cd5187abc376

Observation 37f111f6-d03b-47dc-8e04-8cffaddf19ad · outbound

This paper cites arXiv preprint arXiv:2602.10098 , year=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use arXiv preprint arXiv:2602.10098 , year=

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.153249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.153249Z digest=sha256:71cabe9570217e24e5e7a899c76237d86234e672e39e7729b4e73f6b00b5d2c8

Observation 0bb6acc2-4fcb-4a27-b008-8d0f96ce7529 · outbound

This paper cites Advances in neural information processing systems , volume=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Advances in neural information processing systems , volume=

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:30:04.417001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-08T00:30:03.158542Z digest=sha256:fa120a758998de3e8869690209d8aa6d593e111ab8547248d093c9f311369593

Observation 9605a5a1-b880-4e73-9d8b-a298bf1a501a · outbound

This paper cites CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.164482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.164482Z digest=sha256:e5bcaa282e867af095e804969f1d2156b5e4de3ec8e210ceffce569d10dba612

Observation 53eece55-de9e-4048-9e66-99fad28bd8ea · outbound

This paper cites ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.169903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.169903Z digest=sha256:6a9bfcadf82751b33ca1ae6aed974d79cdadd8aaa9a8bd0b5565d0ba9f5bae9e

Observation 9bab54ad-2aea-4e55-8cc6-935a0fff07b2 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Advances in Neural Information Processing Systems , volume=

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:30:04.401760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-08T00:30:03.176272Z digest=sha256:4258a0e6da5dd0e351852c7363b211797c2b92e969f3ff800e36ce339ac78221

Observation d26d14f8-be1c-49df-a7d1-c0eb87eb3b33 · outbound

This paper cites arXiv preprint arXiv:2512.16793 , year=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use arXiv preprint arXiv:2512.16793 , year=

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.181620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.181620Z digest=sha256:c6d60620c45ff995f49844bb4ae2704a5ba2c5766a55839f59284831e46d7201

Observation d3eac13e-9ff4-4a6e-83a7-049686782f5b · outbound

This paper cites arXiv preprint arXiv:2601.14133 , year=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use arXiv preprint arXiv:2601.14133 , year=

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.187460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.187460Z digest=sha256:a4eb09ef58953515450919f25acaf665115d64c1e8855289a0d8884a5bda4a7d

Observation b31fc78e-0ab8-413d-a379-dc434b8a0e3f · outbound

This paper cites F1: A Vision-Language-Action Model Bridging Understanding and Generation to Actions.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use F1: A Vision-Language-Action Model Bridging Understanding and Generation to Actions

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.192785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.192785Z digest=sha256:72a5277ad619dc3bede67d6345f3eb4e6821b2f988d4fcaca2ce7429a2d3f7b6

Observation 03e84617-c671-4265-959b-e2c0102bf50e · outbound

This paper cites SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.200006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.200006Z digest=sha256:ed29024ddeeec46849e8761e5b8593feeb80171d09ff60d48736fdf35176b272

Observation cac90b40-68a7-4dc2-8389-5d728c22452b · outbound

This paper cites The International Journal of Robotics Research , volume=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use The International Journal of Robotics Research , volume=

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.205571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.205571Z digest=sha256:3dc2db8f45948ed996f666bae7bce3a4de50d6e00bfbc10638530d40f909250f

Observation 10af046f-e2f2-46c3-9b49-32ac84ddabd1 · outbound

This paper cites AgiBot World Colosseo: A Large-scale Manipulation Platform for Scalable and Intelligent Embodied Systems.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use AgiBot World Colosseo: A Large-scale Manipulation Platform for Scalable and Intelligent Embodied Systems

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.210982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.210982Z digest=sha256:0006c2e571fbd9ec746c3ec42cdc62c73075495472ba4151968cd13392f9de3c

Observation fcabc654-5e06-4509-8f56-234e2db5dbec · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.216216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.216216Z digest=sha256:1357d6bec0fb5cdf173fd722b9d7fa591fa4d3db2b94e6c30aea06339303b04a

Observation 2eae2f62-71b2-429d-8268-38737a1f2000 · outbound

This paper cites European conference on computer vision , pages=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use European conference on computer vision , pages=

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.221703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.221703Z digest=sha256:3fcb85e8e450daa3186e54b426c969fbb6c57d1e77c4c1b06edae0f7639199da

Observation 30abe21b-8325-4def-a671-800650bead3f · outbound

This paper cites Evaluating Real-World Robot Manipulation Policies in Simulation.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Evaluating Real-World Robot Manipulation Policies in Simulation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.227603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.227603Z digest=sha256:692c4dbe2506f5bde4d456fb4d413c3896cdf7845ae3151614ff777790e82594

Observation 795120de-04e2-44e3-b023-8d31416b2d23 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.233497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.233497Z digest=sha256:708a04ae9f20e035c804a4bd7a8d5ed022a588c66ac6b9460b67ba5a1eac6b74

Observation d74d6ebf-ad62-4be2-b055-f1c89be2d433 · outbound

This paper cites GR00T N1: An Open Foundation Model for Generalist Humanoid Robots.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use GR00T N1: An Open Foundation Model for Generalist Humanoid Robots

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.239113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.239113Z digest=sha256:f32bbc7f1ccbe5bc4647465dca8775a54a78eeca3b596c387e27acef3d95c998

Observation b0d84f58-b6ef-4187-afb3-b069aafdbe19 · outbound

This paper cites Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:30:04.347498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-08T00:30:03.244179Z digest=sha256:482872831e5c1a8c87899ed4c6f0b5ebd13c8c6fed5d8729fa5fc035e8545ced

Observation 4ac47c81-6e3d-4599-a3af-04d185bdae5f · outbound

This paper cites Advances in neural information processing systems , volume=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Advances in neural information processing systems , volume=

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.249206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.249206Z digest=sha256:1858655f1d212f9d264d9ba69a4249f4e4b927a37da9047723ea3bdff03c15e9

Observation 973885bc-d72e-4e38-8ea7-39073732b0bd · outbound

This paper cites Advances in neural information processing systems , volume=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Advances in neural information processing systems , volume=

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.253931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.253931Z digest=sha256:0565bb7655dc9194810dbce03cea8ce5a70a72757b6d05e77a4698747ff184c1

Observation e9ec2099-f2a6-48dd-a2fa-c8b9eee4228f · outbound

This paper cites arXiv preprint arXiv:2602.01067 , year=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use arXiv preprint arXiv:2602.01067 , year=

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.258938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.258938Z digest=sha256:151d8cde4a718b4dc1befda1684d5966d508fc5959dfc6d349cc948eccda8075

Pith citing papers

No inbound Pith citation observations are available.