Pith. sign in

Paper Citation Record · LEDGER

Mobility VLA: Multimodal Instruction Navigation with Long-Context VLMs and Topological Graphs

As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 25 inbound Pith citation observations for arXiv:2407.07775.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.07775 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 25 of 25 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 25 of 25 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T16:09:09.491569Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T04:29:35.802946Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation dbab5946-d683-4b65-9bcc-d5e1fd584c9f · inbound

SayComply: Grounding Field Robotic Tasks in Operational Compliance through Retrieval-Based Language Models cites this paper.

SayComply: Grounding Field Robotic Tasks in Operational Compliance through Retrieval-Based Language Models Mobility VLA: Multimodal Instruction Navigation with Long-Context VLMs and Topological Graphs

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T18:42:53.105491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:42:53.105491Z digest=sha256:3c738c5044bee026739e3b001ecfa22778893279611df41e63a794860d70d572

Observation 58da1932-d45a-4aef-96fa-4b9dcf45d918 · inbound

Exploring the Adversarial Vulnerabilities of Vision-Language-Action Models in Robotics cites this paper.

Exploring the Adversarial Vulnerabilities of Vision-Language-Action Models in Robotics Mobility VLA: Multimodal Instruction Navigation with Long-Context VLMs and Topological Graphs

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:16.353817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:16.353817Z digest=sha256:e3a86db85efea9b510a38675ee2cac25c07479889fee14d7ead17b973833516b

Observation 377329ab-0abf-4a49-ae71-30697c46578d · inbound

Neural 4D Evolution under Large Topological Changes from 2D Images cites this paper.

Neural 4D Evolution under Large Topological Changes from 2D Images Mobility VLA: Multimodal Instruction Navigation with Long-Context VLMs and Topological Graphs

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T14:41:10.406746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:41:10.406746Z digest=sha256:956d8091d23cd6f3e0b8d78800955594752d0ca73ee6fc97fbc9a49b62be4c91

Observation 5051ee7d-f1c8-47ea-a058-f3c6afcdff71 · inbound

Collaborative Instance Object Navigation: Leveraging Uncertainty-Awareness to Minimize Human-Agent Dialogues cites this paper.

Collaborative Instance Object Navigation: Leveraging Uncertainty-Awareness to Minimize Human-Agent Dialogues Mobility VLA: Multimodal Instruction Navigation with Long-Context VLMs and Topological Graphs

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T04:35:32.629919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:35:32.629919Z digest=sha256:87c741af454e585ed18f0d4636af467c5143c69ce84936acb893e3809672c9c1

Observation 24f56ff0-e31c-4fa0-86c8-56cd84871e9f · inbound

V2PE: Improving Multimodal Long-Context Capability of Vision-Language Models with Variable Visual Position Encoding cites this paper.

V2PE: Improving Multimodal Long-Context Capability of Vision-Language Models with Variable Visual Position Encoding Mobility VLA: Multimodal Instruction Navigation with Long-Context VLMs and Topological Graphs

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T16:58:02.947998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:58:02.947998Z digest=sha256:afff7ac507ae0719c67142b5d597aa96000c3acdfc550f25b8e0094391e1c4f5

Observation f0d34255-e33e-49db-837d-38721ecf3295 · inbound

UniRS: Unifying Multi-temporal Remote Sensing Tasks through Vision Language Models cites this paper.

UniRS: Unifying Multi-temporal Remote Sensing Tasks through Vision Language Models Mobility VLA: Multimodal Instruction Navigation with Long-Context VLMs and Topological Graphs

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T23:18:21.196144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:18:21.196144Z digest=sha256:875f41d40053584fd645ce9d2f06b44b36ebb964e32ecf117ea27706ad634039

Observation a4ad4753-659c-4dfb-9d7b-f3adf0084e78 · inbound

ReFineVLA: Reasoning-Aware Teacher-Guided Transfer Fine-Tuning cites this paper.

ReFineVLA: Reasoning-Aware Teacher-Guided Transfer Fine-Tuning Mobility VLA: Multimodal Instruction Navigation with Long-Context VLMs and Topological Graphs

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:23:17.645992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:23:17.645992Z digest=sha256:0f8d8b5c04a4a0412df6c92c99aea4356503432e23e57a3a92f5fb0bcbf80180

Observation 8350be29-6b15-4fc8-91eb-b4463f5830ee · inbound

GraphPad: Inference-Time 3D Scene Graph Updates for Embodied Question Answering cites this paper.

GraphPad: Inference-Time 3D Scene Graph Updates for Embodied Question Answering Mobility VLA: Multimodal Instruction Navigation with Long-Context VLMs and Topological Graphs

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T11:55:47.755002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:55:47.755002Z digest=sha256:0a35a82019a9ca364a9dfea0bba38da52f3279ec7feee03d25cd483e4f3b10e6

Observation 5f28870d-87a9-47b5-aef3-fddf90adbc18 · inbound

Adversarial Attacks on Robotic Vision Language Action Models cites this paper.

Adversarial Attacks on Robotic Vision Language Action Models Mobility VLA: Multimodal Instruction Navigation with Long-Context VLMs and Topological Graphs

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T11:11:40.301692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:11:40.301692Z digest=sha256:e6768ed3b135bed4335253bc6d8e4b8ab4c7ea0b64a20562480a6f785c9bd1ca

Observation 55fdb703-89ad-4f8d-9e8f-1c51f6a336d8 · inbound

Block-wise Adaptive Caching for Accelerating Diffusion Policy cites this paper.

Block-wise Adaptive Caching for Accelerating Diffusion Policy Mobility VLA: Multimodal Instruction Navigation with Long-Context VLMs and Topological Graphs

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-19T09:37:14.003506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-19T09:36:09.790248Z digest=sha256:e7c7d2064d43537e614bcd2b6cb3813bd8c3672348d2db53a12297e1a3fc589a

Observation 48c744a6-e738-4f39-860f-39c49c376a46 · inbound

PixelNav: Towards Model-based Vision-Only Navigation with Topological Graphs cites this paper.

PixelNav: Towards Model-based Vision-Only Navigation with Topological Graphs Mobility VLA: Multimodal Instruction Navigation with Long-Context VLMs and Topological Graphs

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T13:14:24.028400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:14:24.028400Z digest=sha256:7af5ce1ea7a9be46b57832a0cd88a9b9a83e104b9d16f3965e67e0b64b3994c1

Observation 7129e8ec-9491-4168-a591-702f1a8d6c4a · inbound

Large Model Empowered Embodied AI: A Survey on Decision-Making and Embodied Learning cites this paper.

Large Model Empowered Embodied AI: A Survey on Decision-Making and Embodied Learning Mobility VLA: Multimodal Instruction Navigation with Long-Context VLMs and Topological Graphs

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:38.266231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:38.266231Z digest=sha256:b6c2fb899b21d2ffeb8e5b3adfc5fa768a74b3eeff384cb1bd5825a5d6dfeb77

Observation cdabf6ff-271d-431b-a380-6a3f1b806017 · inbound

CAST: Counterfactual Labels Improve Instruction Following in Vision-Language-Action Models cites this paper.

CAST: Counterfactual Labels Improve Instruction Following in Vision-Language-Action Models Mobility VLA: Multimodal Instruction Navigation with Long-Context VLMs and Topological Graphs

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T19:06:28.522220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:06:28.522220Z digest=sha256:6f544396939a569d16062a2b1a67e563f5f90edb85a81091fc6878238533aa8b

Observation 40206fad-ff53-4f74-ac6b-af18c1a46502 · inbound

TANGO: Traversability-Aware Navigation with Local Metric Control for Topological Goals cites this paper.

TANGO: Traversability-Aware Navigation with Local Metric Control for Topological Goals Mobility VLA: Multimodal Instruction Navigation with Long-Context VLMs and Topological Graphs

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-04T20:21:06.273621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:21:06.273621Z digest=sha256:06c344fde23918bc9dbe482b52e8930e8ff33ccb34e4c646fc1d62ede989bc26

Observation 3ce68ac1-087b-499d-8521-a083d36df4b1 · inbound

SocialNav-SUB: Benchmarking VLMs for Scene Understanding in Social Robot Navigation cites this paper.

SocialNav-SUB: Benchmarking VLMs for Scene Understanding in Social Robot Navigation Mobility VLA: Multimodal Instruction Navigation with Long-Context VLMs and Topological Graphs

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:09.491569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:09.491569Z digest=sha256:06fd02393a36797f300b968fd5c04ad8898d69be820bf5e56ecfdd1f877cefdf

Observation 53337fc5-e3d7-4525-bbde-5a132cb8dad7 · inbound

PLanAR: Planning-Language-Grounded Agentic Reasoning for Robot Manipulation cites this paper.

PLanAR: Planning-Language-Grounded Agentic Reasoning for Robot Manipulation Mobility VLA: Multimodal Instruction Navigation with Long-Context VLMs and Topological Graphs

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T05:38:47.862434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:38:47.862434Z digest=sha256:4681b55357acea19ff9b89872c284da08bd06b9fbdfc1da7d2816804cc17194a

Observation fcd62f20-31ea-499e-88b0-c225289c880e · inbound

AugVLA-3D: Depth-Driven Feature Augmentation for Vision-Language-Action Models cites this paper.

AugVLA-3D: Depth-Driven Feature Augmentation for Vision-Language-Action Models Mobility VLA: Multimodal Instruction Navigation with Long-Context VLMs and Topological Graphs

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-16T06:02:24.984071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-16T06:01:30.803128Z digest=sha256:4b2cbcc36443359cfc19f52d4625dfc26ac35cfb1250e46a738db33fe14cc1ae

Observation 4ac5f567-ed9d-4263-a439-31daf93ddfe7 · inbound

ReFineVLA: Multimodal Reasoning-Aware Generalist Robotic Policies via Teacher-Guided Fine-Tuning cites this paper.

ReFineVLA: Multimodal Reasoning-Aware Generalist Robotic Policies via Teacher-Guided Fine-Tuning Mobility VLA: Multimodal Instruction Navigation with Long-Context VLMs and Topological Graphs

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:48:48.404041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T05:06:38.517652Z digest=sha256:f15e0d60ad9eb1921ea640361e08a40fbe29985c951c539370d0de72156fd318

Observation ae1f6571-fa03-45f7-a9f2-0a36ddadcb6d · inbound

Explore Like Humans: Autonomous Exploration with Online SG-Memo Construction for Embodied Agents cites this paper.

Explore Like Humans: Autonomous Exploration with Online SG-Memo Construction for Embodied Agents Mobility VLA: Multimodal Instruction Navigation with Long-Context VLMs and Topological Graphs

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:41:01.504124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T03:20:35.854120Z digest=sha256:692e84cbe826d986642fa51572f1e3181d45ed09d3a6cfb7bbe6d5c2f018f5a9

Observation b936bc5a-ff0e-4026-92a9-ec9cf81b77c8 · inbound

AsyncShield: A Plug-and-Play Edge Adapter for Asynchronous Cloud-based VLA Navigation cites this paper.

AsyncShield: A Plug-and-Play Edge Adapter for Asynchronous Cloud-based VLA Navigation Mobility VLA: Multimodal Instruction Navigation with Long-Context VLMs and Topological Graphs

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:11:13.809789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-08T03:18:03.855477Z digest=sha256:4921c47b2b1eb3d2afd02317bc74a16e646945b37e2bdbbea1e972cb5c7559c3

Observation 2ea5df08-0ac6-4e8d-930a-515e476863d0 · inbound

Vesta: A Generalist Embodied Reasoning Model cites this paper.

Vesta: A Generalist Embodied Reasoning Model Mobility VLA: Multimodal Instruction Navigation with Long-Context VLMs and Topological Graphs

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-04T04:29:35.804604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-26T16:55:12.518255Z digest=sha256:65383954a78a8fe97edbeede9dcf2dcc8d77a95edd6499f92c364ede1a3f8a39

Observation e010aa7e-79b6-4b1e-9885-a936b56077a3 · inbound

Green for Go, Red for No: Visual Grounding via Semantic Segmentation for VLA Navigation Policies cites this paper.

Green for Go, Red for No: Visual Grounding via Semantic Segmentation for VLA Navigation Policies Mobility VLA: Multimodal Instruction Navigation with Long-Context VLMs and Topological Graphs

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-11T08:29:01.477958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T08:29:01.477958Z digest=sha256:1a29492e5735d9d854e9f805be531d9965396964ea7b651d83c606a0fff6e4c9

Observation 1b64d142-7b55-4f65-9f47-11f48f6b09ad · inbound

ABot-N1: Toward a General Visual Language Navigation Foundation Model cites this paper.

ABot-N1: Toward a General Visual Language Navigation Foundation Model Mobility VLA: Multimodal Instruction Navigation with Long-Context VLMs and Topological Graphs

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-14T12:10:21.115628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:10:21.115628Z digest=sha256:83a09e933c44087f1eec95c2772435ce9b5a24932c91f65800bb953476cd88d5

Observation b23e3ce2-1839-463e-808a-8e25af104d0d · inbound

ABot-N1: Toward a General Visual Language Navigation Foundation Model cites this paper.

ABot-N1: Toward a General Visual Language Navigation Foundation Model Mobility VLA: Multimodal Instruction Navigation with Long-Context VLMs and Topological Graphs

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T07:19:38.021123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:19:38.021123Z digest=sha256:966cb2527c80da24642a54b13cdbfa956a52bc1a438156701dd0a6fceba78e67

Observation f35ef524-a1b3-4784-83b3-6507948d0003 · inbound

Goal-oriented Navigation Instruction Generation with Tour Video Priors cites this paper.

Goal-oriented Navigation Instruction Generation with Tour Video Priors Mobility VLA: Multimodal Instruction Navigation with Long-Context VLMs and Topological Graphs

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-14T04:34:35.044831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:34:35.044831Z digest=sha256:73af62b87c396e63a678126bbb2e98d4709cf44ad3dc4c21ea35e9486561d3b6