Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T21:30:18.197698Z
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 100 of 122 outbound references and 85 inbound Pith citation observations for arXiv:2412.04453.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T21:30:18.197698Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T10:50:21.311984Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T09:09:43.511252Z
100 of 122 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation a363256c-9a04-4119-ab42-3c5db7d6dd56 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Vision-and- language navigation: Interpreting visually-grounded navigation instructions in real environments
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d84794ca-ccd1-47a0-a619-badf00fe81af · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Reinforced cross-modal match- ing and self-supervised imitation learning for vision- language navigation
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49ac0c45-089f-4424-9575-5f80791f90bd · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Learning to explore using active neural slam
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1ff295d-4ef9-48da-9f66-c083f0067529 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Object goal navigation using goal-oriented semantic exploration
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59b44ac5-0c8f-426c-94ee-b53caf157428 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Neural topological slam for visual navigation
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91f2abfc-1957-4d3f-b9b3-9d19712f5f05 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Habitat-web: Learning embodied object- search strategies from human demonstrations at scale
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 570268ab-f4bf-47d1-9f36-89bccd297f4b · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Rt-2: Vision-language-action models transfer web knowledge to robotic control
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0115472b-3e5d-4cc9-a46a-132900ce24fa · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Openvla: An open-source vision-language-action model
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a595661a-31d2-4e3a-9459-4973ba399154 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Open x-embodiment: Robotic learning datasets and rt-x mod- els
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 911cf3e0-e316-4d1c-975d-f41f01f0b0b1 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Spa- tialvlm: Endowing vision-language models with spatial reasoning capabilities
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58830bb4-a099-408d-839b-57f9c27ccb2c · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Spatialrgpt: Grounded spatial reasoning in vision- language models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05dc268e-3be0-4dc9-bf9c-9691b9a38488 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Navid: Video-based vlm plans the next step for vision-and-language navigation
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b737d52-472c-4fc1-b137-2d64785dd41a · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Vila: On pre-training for visual language models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65b3d0cb-92c5-4fb2-aafd-e039216f4add · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Vila-u: a unified foundation model integrating visual understanding and generation
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b722a745-26d8-4db6-8b7a-b9932ccb9d8a · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Vila 2: Vila augmented vila
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 378b2d1a-8f0d-4086-a726-c7c1049068d5 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Longvila: Scaling long- context visual language models for long videos
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a369e17e-c6ee-410e-954d-16bb275b8738 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation X-vila: Cross-modality alignment for large language model
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 559cbdd2-7d62-421e-b43d-bd57471b3cae · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Lita: Language instructed temporal-localization assistant
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4dba096-e6a8-4d49-840a-16ff475015d7 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Nvila: Efficient frontier visual language models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4881a756-aa64-4b9f-a570-54e1e1b9b507 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Visual instruction tuning
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40257328-7073-4b89-9ad0-2750db60f4db · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Coyo-700m: Image-text pair dataset
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9aab1108-f263-497c-9d0b-7e78489217d6 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Multimodal c4: An open, billion-scale corpus of images interleaved with text
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b52bf830-7b78-4975-9d3e-597987ec8b1a · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Improved baselines with visual instruction tuning
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76cb3944-cfb3-4372-a4f8-8f07b623a703 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Airbert: In-domain pretraining for vision-and-language navigation
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c99551f-8731-44eb-930b-e39e9e20490d · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Improving vision-and-language navigation with image-text pairs from the web
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 788944ab-8b84-4a78-b01a-50120c87446b · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Learning vision- and-language navigation from youtube videos
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dffa3edc-7555-4171-b21c-8f4f09ccac35 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Grounding image matching in 3d with mast3r
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3c6e179-c46f-49cf-93f9-51b690655d83 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Hello gpt-4o, 2024
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da16f356-3b44-4ff3-b4c3-c572eb74ddeb · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Beyond the nav-graph: Vision and language navigation in continuous environments
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e1f93db-0c9b-4786-91b9-f6412e9ae0d0 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Room-across-room: Multilingual vision-and-language navigation with dense spatiotem- poral grounding
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34249108-43b1-4f4b-9f12-db518486fb2b · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Learning to navigate unseen environments: Back translation with environmental dropout
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38a4a32d-5441-48b0-9719-2dfd89473c76 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Scanqa: 3d question answering for spatial scene understanding
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd9c79c8-a349-4d41-9a2b-9157de80c1be · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Sharegpt4v: Improving large multi-modal models with better captions
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1fd49c3f-bff9-4ea4-8392-56ee24df0c16 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Video-chatgpt: Towards de- tailed video understanding via large vision and language models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42287389-5a65-46e5-9e33-26d0b8e3f60a · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Extending regular expressions with context operators and parse extraction
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27a06b9c-6ffa-4631-a160-cd3fd2058a21 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Orbit: A unified simulation framework for interactive robot learning environments
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c00eb566-2ce2-43e5-bde4-a7809a56bc3c · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Proximal policy optimiza- tion algorithms
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 658ab226-eebe-4a8d-8709-0c0b27a14540 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Learning quadrupedal locomotion over challenging terrain
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a7520c2-85db-440e-8c16-b9913bcd0fe9 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Learn- ing robust perceptive locomotion for quadrupedal robots in the wild
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a52950d2-2359-4958-b3d9-4674ea5b8aa2 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Rapid locomotion via reinforcement learning
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4df7d7b8-2ae4-4d74-9ffe-e5fdf0dc7e83 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Learn- ing robust autonomous navigation and locomotion for wheeled-legged robots
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ed90a7e-ec21-4c97-9ff6-2e696f3c1e31 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Bridging the gap between learning in discrete and continuous environments for vision-and-language navi- gation
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47ea631a-dded-44f8-b46e-29e787c50265 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Waypoint models for instruction-guided navigation in continuous environ- ments
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b09310a8-0cba-4858-a773-4d435c63cebe · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Sim-2-sim transfer for vision-and-language navigation in continuous environ- ments
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation febb700a-a77f-4679-8494-70247f3dca8d · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Gridmm: Grid memory map for vision- and-language navigation
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 6fb11541-972e-441b-a928-69c4b5e022a6 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Learning navigational visual representations with se- mantic map supervision
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 43421e5f-c1da-4d46-a0be-5a344640446d · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Dreamwalker: Mental planning for contin- uous vision-language navigation
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation cf8ec2cd-05bb-41d7-b2ce-91fbc1171ff4 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation 1st place solutions for rxr-habitat vision-and-language nav- igation competition
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 5d1bef50-e143-4d6c-a505-b2388fd4d18c · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Etpnav: Evolv- ing topological planning for vision-language navigation in continuous environments
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 4b08ae4b-f300-4a04-a335-cc5e44873171 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Lookahead exploration with neural radiance representation for con- tinuous vision-language navigation
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation dc1d74ff-aded-4e15-9399-86acae984a74 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Bevbert: Multimodal map pre-training for language-guided navigation
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 52b2357d-cdce-48a5-a0f3-e42bebcb84e0 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Scaling data generation in vision-and-language naviga- tion
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 52f18842-3c32-4575-9805-d35b1afe0491 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Topological planning with transformers for vision-and-language navigation
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a9d66aaa-8382-4724-9e11-8ad16f82b235 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Language-aligned waypoint (law) supervision for vision-and-language navigation in continuous environments
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 10a49391-f1e2-452b-b672-65b3f8b8f1f1 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Cross-modal map learning for vision and language navigation
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation d1fc72cf-8589-4b90-969e-a40fc9d0ffe0 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Weakly- supervised multi-granularity map learning for vision- and-language navigation
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a9ad9588-d3b4-4ac1-b561-84df3c3886a4 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Affordances-oriented planning using foundation models for continuous vision-language navigation
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation aed9c626-93fa-4af7-a5b5-0efefccf9468 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Beyond the nav-graph: Vision- and-language navigation in continuous environments
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a35b5962-e278-41b2-8d07-08a71924346f · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Li, Gaowen Liu, Mingkui Tan, and Chuang Gan
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ec12e8a0-0f45-4e9b-a76f-938433fa0a4a · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Towards learning a generalist model for embodied navigation
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation fdaf41ca-a68b-48a6-93bc-1742da40723f · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Scene-llm: Extending language model for 3d visual understanding and reasoning
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation c9ab142d-3c1f-4452-8626-45a16f0c55e2 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation An embodied generalist agent in 3d world
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 8cda76cb-cfa7-4054-a680-331f0a34d6a6 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Deep modular co-attention networks for visual question answering
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a839c7b8-00c0-491e-827d-5c52b8cb752e · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation 3d-vista: Pre-trained transformer for 3d vision and text alignment
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation c8c102a5-fd58-40aa-8a5b-ba07e1aef51e · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation 3d- llm: Injecting the 3d world into large language models
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation bb6eca6d-821b-4875-9daf-0046ba95b65a · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Ll3da: Visual interactive instruction tuning for omni-3d understanding reasoning and planning
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 9772d1cf-c948-4466-a4c5-fd713197f4cb · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Chat-scene: Bridging 3d scene and large language mod- els with object identifiers
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b5bff175-0b86-4386-9dc5-725f84714d4b · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Deep whole-body control: Learning a unified policy for ma- nipulation and locomotion
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 79208f9c-74ad-4304-9292-b0b51c84219c · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Habitat: A platform for embodied ai research
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a2436d9b-a4ef-4d8e-b8b1-7a933ff0c338 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Robust speech recognition via large-scale weak supervision
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 817ab644-2f4d-40f9-b335-5cbaa4cb875f · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Obstacle avoidance and naviga- tion in the real world by a seeing robot rover
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 768bacb8-77ba-465d-bbd4-d2e010554ccc · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Sonar-based real-world mapping and navigation
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation c1177870-473c-4f9c-ad67-bb4c23cb61ac · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Robust monte carlo localization for mobile robots
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 7e2754ef-6ace-4c45-8185-bf5c0fa83369 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Navigating to objects in the real world
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 4ebb8a00-e351-40d0-8e52-bbc16aaa01d6 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Minerva: A second-generation museum tour-guide robot
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ed81421a-b222-4ec5-8499-d088410b203d · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Kinectfusion: Real-time dense surface mapping and tracking
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation e555f857-e0ad-417c-a19e-4d76eaf8dd69 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Monoslam: Real-time single camera slam
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation fee5065f-9262-4295-9d39-f15335605d15 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Visual-inertial navigation, mapping and localization: A scalable real- time causal approach
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 9fad0211-9fdc-4b78-99cf-13ab9e3ef97a · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Gated-attention architectures for task-oriented language grounding
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 074665ac-1205-4171-9e32-ecb935306093 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation End-to-end driving via conditional imitation learning
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 03f69ab1-da58-4f8c-a927-36739973faae · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Human-level control through deep reinforcement learning
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 694ac538-c33d-4070-b5e4-863dc42e8d3a · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Continuous control with deep reinforce- ment learning
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 88252727-b064-4087-b32e-9a5fb7e24972 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Reverie: Remote embodied visual referring expression in real indoor environments
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 810bbb99-44e4-4bcd-8c52-c29cfc621296 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Matterport3d: Learning from rgb-d data in indoor environments
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 28b42a25-66aa-4504-8c65-cb168d0fa6b9 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Speaker-follower models for vision-and- language navigation
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 8d86520f-65b8-444b-a154-fc8dd8d0b522 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Self-monitoring navigation agent via auxiliary progress estimation
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 24cac0a8-a7df-4165-a822-731150688f90 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Tactical rewind: Self-correction via backtracking in vision-and-language navigation
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation c594b6ea-60dc-4d2c-8d78-9d24bf5a1478 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Language and visual entity rela- tionship graph for agent navigation
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 7b1a9998-751d-47df-801c-00a4cc314965 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation History aware multimodal transformer for vision-and-language navigation
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation f3aaaf58-9ce7-4078-9c22-983ce7e00219 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Mapgpt: Map-guided prompting with adaptive path planning for vision-and- language navigation
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 1fba79c5-22a9-4a44-91b9-ccc40f8a6136 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Navgpt-2: Unleashing navigational reasoning capability for large vision-language models
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation eb2f9478-1a99-44ef-956a-d6ff04ee56a0 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Robust navigation with language pretraining and stochastic sampling
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ca0fbb2e-7ade-4a31-a7fc-cdca3270c246 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation A new path: Scaling vision-and-language navigation with synthetic instruc- tions and imitation learning
Reference 93
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 590399c6-d6f6-404a-8f8a-70744b063820 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Hierarchical cross-modal agent for robotics vision-and-language navigation
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation f96e3868-d9cf-4090-89e3-24f4062d0c67 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Octo: An open-source generalist robot policy
Reference 95
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ead327bc-0b54-4785-84a6-aa6763529d22 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Scaling cross-embodied learning: One policy for manipulation, navigation, lo- comotion and aviation
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 8213583d-38df-4f8b-93f6-56352140919d · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Pushing the limits of cross- embodiment learning for manipulation and navigation
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 61d3d485-36e4-4b27-a63d-8259bf2cd161 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Poliformer: Scaling on-policy rl with transformers results in mas- terful navigators
Reference 98
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 467acaa3-6b37-43ef-a20c-c5456742bbfc · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Gnm: A general navigation model to drive any robot
Reference 99
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation f0ab59dc-4a01-4c91-8f96-9b5d706f4b89 · outbound
NaVILA: Legged Robot Vision-Language-Action Model for Navigation Nomad: Goal masked diffusion policies for navigation and exploration
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation fac3d417-f91f-4a1e-a41c-8d39f05ae5cc · inbound
NVILA: Efficient Frontier Visual Language Models NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 87032ec4-0de6-42c4-8414-021e309234ab · inbound
doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 836504b2-6e5d-48a7-9dc3-fa941533a835 · inbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 64d3b1c8-7ccc-47a6-9849-95ad8e8a46b4 · inbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 5bef1725-593f-413c-94d9-7ccf30136d2b · inbound
MASR: Self-Reflective Reasoning through Multimodal Hierarchical Attention Focusing for Agent-based Video Understanding NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 257800ec-aff5-4796-b802-35990f8bda71 · inbound
RoboVerse: Towards a Unified Platform, Dataset and Benchmark for Scalable and Generalizable Robot Learning NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c7dc942-9294-4115-b044-85ea5dfe130a · inbound
SpatialReasoner: Towards Explicit and Generalizable 3D Spatial Reasoning NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 176cfe1f-fe6c-4c1b-8191-a8d663bae688 · inbound
A Survey of Robotic Navigation and Manipulation with Physics Simulators in the Era of Embodied AI NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e5caf3e-bffc-4b63-a42d-14a7215d9f47 · inbound
Omni-Perception: Omnidirectional Collision Avoidance for Legged Locomotion in Dynamic Environments NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c73ab88c-39f1-44a1-b111-8abc20fdbeba · inbound
Cross from Left to Right Brain: Adaptive Text Dreamer for Vision-and-Language Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59f4dc6f-7814-4d5a-8ac6-845a407fbd16 · inbound
Fast and Cost-effective Speculative Edge-Cloud Decoding with Early Exits NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2eef2a0e-9ef3-4ab0-ba8c-0c7f41be2c87 · inbound
TrackVLA: Embodied Visual Tracking in the Wild NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38816cc3-5d58-4c66-9782-0b54dcd0d5d0 · inbound
Visual Embodied Brain: Let Multimodal Large Language Models See, Think, and Control in Spaces NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 396dad35-108d-4f25-beea-2953d1be9b1d · inbound
LoHoVLA: A Unified Vision-Language-Action Model for Long-Horizon Embodied Tasks NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c65785dc-dc54-48cc-8a95-cbdfc092f495 · inbound
Real-Time Execution of Action Chunking Flow Policies NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation d427094e-e49b-4e57-8b62-ef3e1532db70 · inbound
OctoNav: Towards Generalist Embodied Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed17b4aa-87dc-430f-a22e-7824d1b5cc3f · inbound
Can Pretrained Vision-Language Embeddings Alone Guide Robot Navigation? NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7e8f828-9987-4e2f-887e-4330e5f0a59c · inbound
Hierarchical Vision-Language Planning for Multi-Step Humanoid Manipulation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80347b0f-17f0-42af-ba60-2eea9961e4d0 · inbound
A Survey on Vision-Language-Action Models: An Action Tokenization Perspective NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 153
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ce5d2b74-c720-4015-ae13-4a7af8fd3cc0 · inbound
LOVON: Legged Open-Vocabulary Object Navigator NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74ba35c4-3fda-4de6-8faf-1e5eeae78498 · inbound
Making VLMs More Robot-Friendly: Self-Critical Distillation of Low-Level Procedural Reasoning NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a6cdeb1-03f9-4203-8b30-d80f04486d72 · inbound
EmbRACE-3K: Embodied Reasoning and Action in Complex Environments NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ee95326-0488-4216-a12b-b8dff9a04446 · inbound
Shortcut Learning in Generalist Robot Policies: The Role of Dataset Diversity and Fragmentation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a42816a0-a620-41d5-9747-aa27b76a829a · inbound
SPARSE Data, Rich Results: Few-Shot Semi-Supervised Learning via Class-Conditioned Image Translation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ded1f365-b984-4087-a96d-a245e3a528f9 · inbound
CorrectNav: Self-Correction Flywheel Empowers Vision-Language-Action Navigation Model NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f291dc62-9fa7-4f61-b385-8632ae5e69f7 · inbound
SPG: Style-Prompting Guidance for Style-Specific Content Creation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7b4a824-620b-40e8-989f-60541a878e8f · inbound
CAST: Counterfactual Labels Improve Instruction Following in Vision-Language-Action Models NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 958c5fc4-594e-44d3-905f-748fd98e4e2c · inbound
From reactive to cognitive: brain-inspired spatial intelligence for embodied agents NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc10f7e0-86da-404d-8597-82dfcfea904a · inbound
Beyond Pixels: Introducing Geometric-Semantic World Priors for Video-based Embodied Models via Spatio-temporal Alignment NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 347d0f17-a116-4e37-9d7f-8a777b4109dd · inbound
Robix: A Unified Model for Robot Interaction, Reasoning and Planning NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6eb69b24-e8c1-4332-bb3e-2b5f293abe54 · inbound
Nav-R1: Reasoning and Navigation in Embodied Scenes NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64ed98ce-ebd6-4121-af1d-d60f1761d8ef · inbound
R2RGEN: Real-to-Real 3D Data Generation for Spatially Generalized Manipulation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 3fe1c212-e7fd-4c3c-8912-a7930b086f03 · inbound
Progress-Think: Semantic Progress Reasoning for Vision-Language Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 22394142-5144-4124-9272-e14f4f880d3f · inbound
AstraNav-World: World Model for Foresight Control and Consistency NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 9a86e709-e12f-4f98-b08b-8ebdb98f2a4a · inbound
VLN-Cache: Enabling Token Caching for VLN Models with Visual/Semantic Dynamics Awareness NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 5db8e094-517c-4759-b35a-d9a0eb429c3d · inbound
Structured Observation Language for Efficient and Generalizable Vision-Language Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b22d2ba4-de02-411a-9a1c-08d2a5beb5fc · inbound
Structured Observation Language for Efficient and Generalizable Vision-Language Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68842129-da80-4f90-be5e-25b2877a6721 · inbound
Learning Task-Invariant Properties via Dreamer: Enabling Efficient Policy Transfer for Quadruped Robots NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 36db350c-d7f4-4364-9154-3106a2bf5bec · inbound
HiRO-Nav: Hybrid ReasOning Enables Efficient Embodied Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a483cbf6-61b7-4079-acc5-4094ee500782 · inbound
Visually-grounded Humanoid Agents NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation f7431370-e85e-42e1-8961-779024c897df · inbound
Vision-and-Language Navigation for UAVs: Progress, Challenges, and a Research Roadmap NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 7b9d5fbe-234e-48c0-9b58-099b5820b2bb · inbound
CART: Context-Aware Terrain Adaptation using Temporal Sequence Selection for Legged Robots NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 6266283e-8f1e-4979-9cb4-72232016f271 · inbound
CART: Context-Aware Terrain Adaptation using Temporal Sequence Selection for Legged Robots NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84d45e28-7c51-45f1-8658-aac6ce8e678d · inbound
Think before Go: Hierarchical Reasoning for Image-goal Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 102
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation fe194dbe-af9d-46ea-968b-9dbfa45f548c · inbound
Dual-Anchoring: Addressing State Drift in Vision-Language Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 163b26c4-4daa-4f7b-950c-d0651a2e70c7 · inbound
SpaAct: Spatially-Activated Transition Learning with Curriculum Adaptation for Vision-Language Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 388dea21-e8b9-4e71-8a59-2c428a5b97e3 · inbound
Beyond Isolation: A Unified Benchmark for General-Purpose Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b01388b8-e8c8-4d07-b0bd-accfe5ee30ea · inbound
Terrain Consistent Reference-Guided RL for Humanoid Navigation Autonomy NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 55735769-cfc8-49e1-91ae-76d4454c88e3 · inbound
SEDualVLN: A Spatially-Enhanced Dual-System for Vision-Language Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 84154134-46ff-4693-a428-37c25129d403 · inbound
SEDualVLN: A Spatially-Enhanced Dual-System for Vision-Language Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 20405ac0-091c-4448-8806-d39c7ef3aef3 · inbound
GA-VLN: Geometry-Aware BEV Representation for Efficient Vision-Language Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 2120b067-32a6-4a06-8585-b872eaeb6435 · inbound
AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation deadbe25-e701-4ce6-8412-8e2732944916 · inbound
G-DRAGON: Geospatial Reasoning and Dynamic Planning for Retrieval-Augmented Outdoor Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 0ad83405-314b-48d1-861d-db91b4970d60 · inbound
Uni-LaViRA: Language-Vision-Robot Actions Translation for Unified Embodied Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 7d8c41c4-e31e-4da3-bf8b-14aaa25f3361 · inbound
POINav: Benchmarking and Enhancing Final-Meters Arrival in Real-World Vision-Language Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 1e1a8380-b7ab-4042-b09e-9aeb5418ec3e · inbound
PEACE: A Planner-Executor Agent with Constraint Enforcement for UAVs NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 2bd71fac-4c4b-4551-9a56-6e6f7fdb41bb · inbound
Hierarchical Semantic-Augmented Navigation: Optimal Transport and Graph-Driven Reasoning for Vision-Language Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b79fac24-9072-4f62-8a59-08199f454518 · inbound
Goal2Pixel: Grounding Goals to Pixels for Vision-Language Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation c09ffcf2-259f-4859-bf6e-bc2d6c216df2 · inbound
Ask When It Pays: Cost-Aware Open-Ended Interaction for Instance Goal Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation d7115002-5b6a-4bdb-83f5-cbf262b684e5 · inbound
Beyond Waypoints: A Trajectory-Centric Waypointing Paradigm for Vision-Language Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 17ad59cb-7c8d-4d41-9b1f-2347ee9a9ca5 · inbound
Act on What You See: Unlocking Safe Social Navigation in Vision-Language-Action Models NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 8eabb674-170a-4a53-8da8-8543c03f3941 · inbound
From Imitation to Alignment: Human-Preference Flow Policies for Long-Horizon Sidewalk Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 899df63f-9ade-4e5b-85f9-6121d359961f · inbound
VEGA: Learning Navigation VLAs from In-the-Wild Egocentric Video with Geometric Trajectory Supervision NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation cbada972-9842-43c1-b840-9bb5f6f2b4bc · inbound
VEGA: Learning Navigation VLAs from In-the-Wild Egocentric Video with Geometric Trajectory Supervision NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ccd85cd-0fa4-4d8d-9ba3-6250c061aa6c · inbound
Slow Brain, Fast Planner: Latency-Resilient VLM-Augmented Urban Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 091a2de6-e56a-4a62-a619-e2b06ba794e9 · inbound
Vesta: A Generalist Embodied Reasoning Model NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 38150bff-1b9d-4291-993b-8050aec8de85 · inbound
BIT-Nav: Brain-Inspired Trajectory Memory for Embodied Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 6f8c30ea-56fe-49c7-8600-cc3687652cac · inbound
FlowDec: Temporal Conditional Flow Decorruptor for Robust Continuous Vision-Language Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation c560a56e-eacc-4fe8-a460-cefc703ba6df · inbound
SpikeVLA: Vision-Language-Action Models with Spiking Neural Networks NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 613b6a72-ca78-421f-b77f-9dec93d39f73 · inbound
Vision-Language Models for Deployable Social Robot Navigation: Bridging Semantic Reasoning and Low-Level Control NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 73fef3cb-f37c-4a72-9c36-98d11587368b · inbound
FutureNav: Unified World-Action Modeling for Vision-and-Language Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation c132e310-40d5-4fa2-92b5-e7a178b3f5ab · inbound
Path-level Hindsight Instructions for Semantic Exploration in Vision-Language Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a3f816f0-a771-4f4e-9e6b-66e40e2249b1 · inbound
Exp2VLA: Enabling Vision-Language-Action for Drone Navigation from Expert Demonstrations NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60e3694c-c7e2-4522-8943-3ba9f774cccf · inbound
ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6e50527-e702-4417-aa23-4552559563ad · inbound
Green for Go, Red for No: Visual Grounding via Semantic Segmentation for VLA Navigation Policies NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4729f4d-613e-4a48-bbbf-f24dfdff1b23 · inbound
Anticipate Before Acting: Future-State-Conditioned Vision-Language Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4dcd7d93-6b4e-406e-a652-5bedd565ddcb · inbound
EA-Nav: Learning Safe Visual Navigation Policies with Embodiment Awareness NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e252989-0e19-47ea-a185-1296bbc615d6 · inbound
ACME: A Multi-Cultural, Multi-Embodiment Social-Navigation Dataset NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 292
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79160368-0112-47d0-b764-c75123e98fe3 · inbound
Zero-Shot Mission-Level Evaluation for Aerial MLLM Agents NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51a0af7a-2ead-4729-b06b-9357ef1cda00 · inbound
MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fda33824-fdbd-4099-8d6b-0184621082ef · inbound
Embodied Agents Take Control: Minimal-Interface Zero-Shot Agents Rival Industrial-Scale Policies in Vision-and-Language Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0ffe209-0598-4085-a7b1-ca7ea417fa2f · inbound
Embodied Agents Take Control: Minimal-Interface Zero-Shot Agents Rival Industrial-Scale Policies in Vision-and-Language Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73f86b6f-c688-413c-996f-c89124c25906 · inbound
WNM-3D: A World Navigation Model with 3D Scene Conditioning for Closed-Loop VLN NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11c936a4-78c7-458d-821f-0029b6b698df · inbound
Goal-oriented Navigation Instruction Generation with Tour Video Priors NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a860a33c-fd01-4431-890e-a63af14b9b69 · inbound
HumanoidVLN: A Physics-Grounded Simulator and Benchmark for Vision-Language Navigation Across Diverse Humanoid Embodiments NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.