Pith. sign in

Paper Citation Record · LEDGER

BEVBert: Multimodal Map Pre-training for Language-guided Navigation

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2212.04385.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2212.04385 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:02:51.260957Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T11:26:54.066214Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ebe58894-a4b8-4119-b457-6a51c6b8dfba · inbound

NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation cites this paper.

NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-18T04:55:20.409420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T04:55:20.362512Z digest=sha256:d7cc0b1759381bf23cbcf34f397fde3ab157a340e03786cd8a3381a3a5a02da6

Observation 6c700b45-057a-4f36-b138-e8092009fdce · inbound

CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation cites this paper.

CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T15:02:51.260957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:02:51.260957Z digest=sha256:54a7c68cad73cf10364856792962cce784dceecaa588f315bb410ec5c501e69c

Observation 8515c56d-e73d-4b7b-be09-c17388a5d911 · inbound

Cross from Left to Right Brain: Adaptive Text Dreamer for Vision-and-Language Navigation cites this paper.

Cross from Left to Right Brain: Adaptive Text Dreamer for Vision-and-Language Navigation BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T13:52:03.792370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:52:03.792370Z digest=sha256:3f7670d93bd116e1d054d706c9d4097d2d8578d3347f6c8d4e53c2cc04e25997

Observation 6844dbe5-9679-49a8-bc10-f4558754dd79 · inbound

NavMorph: A Self-Evolving World Model for Vision-and-Language Navigation in Continuous Environments cites this paper.

NavMorph: A Self-Evolving World Model for Vision-and-Language Navigation in Continuous Environments BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T21:48:22.838916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:48:22.838916Z digest=sha256:b368f153695c3016c159e40ad2e0ce0270a9a30552ce35287da4f7c1cc3f1643

Observation 701d898e-fdc6-4ed5-9054-0438f6046e46 · inbound

Progress-Think: Semantic Progress Reasoning for Vision-Language Navigation cites this paper.

Progress-Think: Semantic Progress Reasoning for Vision-Language Navigation BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-17T21:05:15.982288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T21:02:38.013115Z digest=sha256:47898dc4050ff120f0c8265958d186bdd5228a38d241980462b66c7cd7caa021

Observation cb1db964-b710-47fe-b6d4-6105fbb5ae35 · inbound

HiMemVLN: Enhancing Reliability of Open-Source Zero-Shot Vision-and-Language Navigation with Hierarchical Memory System cites this paper.

HiMemVLN: Enhancing Reliability of Open-Source Zero-Shot Vision-and-Language Navigation with Hierarchical Memory System BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T18:10:13.897082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:10:13.897082Z digest=sha256:b0b8e620df2608cad6a2d83f00d932e7b203b427b425d3ece0dc5e54ac1cfb9a

Observation 7665012f-0536-40e8-8e37-b36cd17bb917 · inbound

LightZeroNav: Zero-Shot Vision Language Navigation in Continuous Environments Based on Lightweight VLMs cites this paper.

LightZeroNav: Zero-Shot Vision Language Navigation in Continuous Environments Based on Lightweight VLMs BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-21T10:44:07.845806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T10:41:47.297072Z digest=sha256:7eb069fd670b295918102da8ddc3aa99b5028236e4f16c0ab4ff7eb496dbe02e

Observation 0fade2b9-56f9-4a89-84ed-b91860c0a2d5 · inbound

Watching Movies Like a Human: Egocentric Emotion Understanding for Embodied Companions cites this paper.

Watching Movies Like a Human: Egocentric Emotion Understanding for Embodied Companions BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T08:53:04.305421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T08:49:33.107658Z digest=sha256:2d833327b167e2e7581ad611976785b9adecc144cee5044b3e3be3d35799e998

Observation 403e845d-de07-482b-894d-78bf9f1d6f2d · inbound

Dual-Anchoring: Addressing State Drift in Vision-Language Navigation cites this paper.

Dual-Anchoring: Addressing State Drift in Vision-Language Navigation BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 123

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T06:51:45.893511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T06:50:34.310831Z digest=sha256:b580da878601550b900eed96badb7700059971afe2cbdb9257374ab44d010cbd

Observation 6ff4f4a3-6243-4179-bf13-6ef065173d82 · inbound

LCGNav: Local Candidate-Aware Geometric Enhancement for General Topological Planning in Vision-Language Navigation cites this paper.

LCGNav: Local Candidate-Aware Geometric Enhancement for General Topological Planning in Vision-Language Navigation BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-12T02:06:14.900560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T02:04:57.627160Z digest=sha256:74c05f7c1dd333dd99baeeb683c0d45124e2eebe988f0b91dcaf98de4128ce4a

Observation 989c518d-6a76-4494-8fa5-8735a378b2e9 · inbound

GA-VLN: Geometry-Aware BEV Representation for Efficient Vision-Language Navigation cites this paper.

GA-VLN: Geometry-Aware BEV Representation for Efficient Vision-Language Navigation BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-22T06:46:10.657852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T06:45:35.920991Z digest=sha256:45ebda93257eba7e72f9b0b850e43fe7eab527daacfe63401e57c7df8b944356

Observation f4749ee8-2d9f-46e1-be8a-9d0851fd3236 · inbound

Large Depth Completion Model from Sparse Observations cites this paper.

Large Depth Completion Model from Sparse Observations BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:23:15.260029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T08:19:01.715788Z digest=sha256:35190fd8ef0ff8f9d6f441cd15ae79d7672bfac81fb06d4026ac491070406ffa

Observation 7edd71a2-9bbb-4ff3-809e-b820123e3123 · inbound

Ask When It Pays: Cost-Aware Open-Ended Interaction for Instance Goal Navigation cites this paper.

Ask When It Pays: Cost-Aware Open-Ended Interaction for Instance Goal Navigation BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-02T02:26:26.147588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T11:01:08.668612Z digest=sha256:efa0fc147bf155abd8e9a02ec1fbe1323798ee473620540b606958f28e7d0207

Observation 0561c1c0-5f30-4094-854e-45aaba081f13 · inbound

SpikeVLA: Vision-Language-Action Models with Spiking Neural Networks cites this paper.

SpikeVLA: Vision-Language-Action Models with Spiking Neural Networks BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-06-29T19:33:54.610574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T04:42:23.040915Z digest=sha256:5dbb573c7479a12367d4c43e8608c2febb55d1fd29318921b09ff663b0194c92

Observation 56c6f726-63d1-4d18-99e4-d6a3e14cc72d · inbound

DART-VLN: Test-Time Memory Decay and Anti-Loop Regularization for Discrete Vision-Language Navigation cites this paper.

DART-VLN: Test-Time Memory Decay and Anti-Loop Regularization for Discrete Vision-Language Navigation BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-02T11:26:54.068087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-02T11:17:26.529397Z digest=sha256:73719d85c269d5aaef78008de77a08821d18856c55dd782a0ece69bf1a10bfbf

Observation 0899cc28-37fb-4e2e-b629-ba12eccde1ed · inbound

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation cites this paper.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:19.022152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:19.022152Z digest=sha256:1d25bcff9002b7eaf18bf161c16e5812388f0da72ad5957ad98a474fb6a3d05e