Pith. sign in

REVIEW 1 cited by

ALT-Pilot: Autonomous navigation with Language augmented Topometric maps

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2310.02324 v1 pith:SFPQ6ST5 submitted 2023-10-03 cs.RO

classification cs.RO
keywords alt-pilotautonomousnavigationtopometricvehiclelandmarkslanguagemaps
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We present an autonomous navigation system that operates without assuming HD LiDAR maps of the environment. Our system, ALT-Pilot, relies only on publicly available road network information and a sparse (and noisy) set of crowdsourced language landmarks. With the help of onboard sensors and a language-augmented topometric map, ALT-Pilot autonomously pilots the vehicle to any destination on the road network. We achieve this by leveraging vision-language models pre-trained on web-scale data to identify potential landmarks in a scene, incorporating vision-language features into the recursive Bayesian state estimation stack to generate global (route) plans, and a reactive trajectory planner and controller operating in the vehicle frame. We implement and evaluate ALT-Pilot in simulation and on a real, full-scale autonomous vehicle and report improvements over state-of-the-art topometric navigation systems by a factor of 3x on localization accuracy and 5x on goal reachability

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Development of Vision-Language Model-based GNSS Spoofing Detection for Autonomous Vehicle Navigation

    cs.CV 2026-07 conditional novelty 5.0 of 10

    A three-stage fine-tuned vision-language model that compares camera/sensor-inferred maneuvers with GNSS-implied maneuvers detects wrong-turn, overshoot, and stop spoofing attacks at 94–95% F1 on an independent cross-r...

Pith tools