Pith. sign in

Paper Citation Record · LEDGER

VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 29 inbound Pith citation observations for arXiv:2506.17221.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.17221 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 29 of 29 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T14:28:52.281460Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-05T11:41:02.489833Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 73c5e88f-7f8c-427d-a906-4abc614a6403 · inbound

Pseudo Depth Meets Gaussian: A Feed-forward RGB SLAM Baseline cites this paper.

Pseudo Depth Meets Gaussian: A Feed-forward RGB SLAM Baseline VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T23:56:04.195788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:56:04.195788Z digest=sha256:39d56af18036b26bf0ced3f2ddf4f018cd2a896d16b4565c3d165a621b9a98b9

Observation b85f0e97-f99f-4328-b6a6-3b51bc81b640 · inbound

Nav-R1: Reasoning and Navigation in Embodied Scenes cites this paper.

Nav-R1: Reasoning and Navigation in Embodied Scenes VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:42.089676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:42.089676Z digest=sha256:80c0df811d3f10d364739e8ecbca4e2293e0f076c95667a887c23f5ff8ac9860

Observation 5618983b-961c-4d25-8831-f4de5fb791b7 · inbound

DreamNav: A Trajectory-Based Imaginative Framework for Zero-Shot Vision-and-Language Navigation cites this paper.

DreamNav: A Trajectory-Based Imaginative Framework for Zero-Shot Vision-and-Language Navigation VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T17:00:45.020389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:00:45.020389Z digest=sha256:90e788212bcc1fa8ea71e9675eb7e64fbb6b4a6b57919cbefe4e41725b4e977a

Observation aa6ee4d9-00b3-47dd-9ceb-a3d06dbea990 · inbound

Reinforcement Learning Meets Large Language Models: A Survey of Advancements and Applications Across the LLM Lifecycle cites this paper.

Reinforcement Learning Meets Large Language Models: A Survey of Advancements and Applications Across the LLM Lifecycle VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 140

Resolution
unresolved
no resolver link, observed 2026-08-04T16:07:37.828525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:07:37.828525Z digest=sha256:6346ad93136368f232ef88b6279f9ab3257f5b14ea079ff52a2041c4575f108c

Observation 8f6d4493-0bc1-4a99-9994-6333f0c359cd · inbound

IRPO: Boosting Image Restoration via Post-training GRPO cites this paper.

IRPO: Boosting Image Restoration via Post-training GRPO VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-03T19:24:39.321499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T19:24:39.321499Z digest=sha256:fd8f7c68c410bfa7bfd40e0499cd11b2cd02b386823b8273b222ff0ff7307933

Observation 49bc9c5f-727f-4eb4-9bc8-98d0d5dd10a2 · inbound

Aerial Vision-Language Navigation with a Unified Framework for Spatial, Temporal and Embodied Reasoning cites this paper.

Aerial Vision-Language Navigation with a Unified Framework for Spatial, Temporal and Embodied Reasoning VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-16T23:58:42.589185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-16T23:58:40.992942Z digest=sha256:9a559d548d7e0cdb9af97a8d123018a2f76ac936dd26cf61123954ce0803ffd3

Observation 6ef87bd5-bd44-4240-8069-7b85f724ef06 · inbound

Token Warping Helps MLLMs Look from Nearby Viewpoints cites this paper.

Token Warping Helps MLLMs Look from Nearby Viewpoints VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-05-13T21:08:17.442855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-13T21:07:55.062113Z digest=sha256:36a9e31c71011acac9fd71a8d189bcd43d0eabf3156c792423c64cc30fb709c8

Observation 818345e2-d07f-4a43-a906-1cca8202e0c9 · inbound

Hypothesis Graph Refinement: Hypothesis-Driven Exploration with Cascade Error Correction for Embodied Navigation cites this paper.

Hypothesis Graph Refinement: Hypothesis-Driven Exploration with Cascade Error Correction for Embodied Navigation VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:08:00.907219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-13T17:05:57.205365Z digest=sha256:09dcc854e82a72d6b2e0db88fdce604436f970d3978efc7d712fc9985f0c0ac9

Observation fde93708-7efc-41f8-aaa4-f929867df094 · inbound

Enhancing MLLM Spatial Understanding via Active 3D Scene Exploration for Multi-Perspective Reasoning cites this paper.

Enhancing MLLM Spatial Understanding via Active 3D Scene Exploration for Multi-Perspective Reasoning VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:15:47.751455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T19:16:46.753641Z digest=sha256:f2c3d40080cf8655b4b0da13a2bd5cfd6dece846fe8ffc316dbf7918d0aedb78

Observation 53443418-d67e-4f2a-9708-e9072ca1faf8 · inbound

HiRO-Nav: Hybrid ReasOning Enables Efficient Embodied Navigation cites this paper.

HiRO-Nav: Hybrid ReasOning Enables Efficient Embodied Navigation VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:41:01.531030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T17:01:13.350219Z digest=sha256:6c38454ae3070617a3af8efec9b52e2a646fadc7054b84dd0a53429f358057d1

Observation 062ae992-5bef-4e9a-af06-a426058fbc1f · inbound

Think before Go: Hierarchical Reasoning for Image-goal Navigation cites this paper.

Think before Go: Hierarchical Reasoning for Image-goal Navigation VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-10T05:51:10.485653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-05-10T05:43:27.972164Z digest=sha256:54bda815213952108c1f17244f25be87f46e7b7c7e20ee5eb5607e2cfa0e4073

Observation d7e87bfa-b62d-4502-b272-c4e2f2ccce5b · inbound

Dual-Anchoring: Addressing State Drift in Vision-Language Navigation cites this paper.

Dual-Anchoring: Addressing State Drift in Vision-Language Navigation VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:51:45.828270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-05-10T06:50:34.310831Z digest=sha256:682e05d64b070cfd6a20b57809138c0dbc8a7a13378d1db31f01d0d86ca9a9de

Observation 9f612723-deee-4adf-a40e-1a7f51c2744e · inbound

Steadily moving semi-infinite fracture in plane poroelasticity cites this paper.

Steadily moving semi-infinite fracture in plane poroelasticity VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 77

Resolution
verified exact
local_arxiv, observed 2026-07-05T11:41:02.491147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-05T11:39:05.686584Z digest=sha256:3b4cfe15c7b6a30aff9c78fea451eff1eef0a6ef490964c91c6bd265a84af4d4

Observation c75ecfb5-fb62-42f1-a8d0-f4067bd19a30 · inbound

XEmbodied: A Foundation Model with Enhanced Geometric and Physical Cues for Large-Scale Embodied Environments cites this paper.

XEmbodied: A Foundation Model with Enhanced Geometric and Physical Cues for Large-Scale Embodied Environments VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 77

Resolution
verified exact
arxiv_id, observed 2026-05-10T05:51:10.296777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T05:46:36.865150Z digest=sha256:fa8b97f44b2625880871e00a36f84fbbaf2498c6f93bcd7f878574e0b0012f48

Observation e37d385e-ef1d-407e-acb5-948dbc1a8cc6 · inbound

SpaAct: Spatially-Activated Transition Learning with Curriculum Adaptation for Vision-Language Navigation cites this paper.

SpaAct: Spatially-Activated Transition Learning with Curriculum Adaptation for Vision-Language Navigation VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:36:31.243113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-07T05:05:33.606975Z digest=sha256:5c9ec7c8ff41ac7876103c6ddb93f35f9bfa171b8f3efc32cf5215f12f9ca651

Observation 79ce5e0b-fb83-47c5-bf81-7c5e0cf19292 · inbound

Beyond Thinking: Imagining in 360$^\circ$ for Humanoid Visual Search cites this paper.

Beyond Thinking: Imagining in 360$^\circ$ for Humanoid Visual Search VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 84

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:01:18.332757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-05-12T02:58:46.728868Z digest=sha256:898742d9c317d39bc5b3379943f6c3dc97770b12bd0a1b54e70cc12402f184ff

Observation f66da04d-f31d-4ccc-9688-5ff8f158570e · inbound

WorldVLN: Autoregressive World Action Model for Aerial Vision-Language Navigation cites this paper.

WorldVLN: Autoregressive World Action Model for Aerial Vision-Language Navigation VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-20T17:48:48.633274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-20T17:47:32.953903Z digest=sha256:ba477ffe3f96f7f64207bf1fa4f46ce07fccbd9536116ef98888f1942f17034a

Observation 0ce6c4d5-b991-4c1b-a393-9f1103944045 · inbound

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation cites this paper.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-22T04:46:04.568388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:1b6b7cafcdb9fd8668b330f98ed4cf0fc4294b3eea741773dadadccc16792b8e

Observation 8da6f638-a852-4621-a212-12dc15632643 · inbound

World Models as Group Actions cites this paper.

World Models as Group Actions VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-06-30T13:24:40.331396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-30T13:17:42.720825Z digest=sha256:70d51a42ce71bbb7a3721bc94dab48533c6d2aa3d5254b353709d37eebc63a1b

Observation ed9be44b-6172-4d41-b974-7b2b4d67a4ed · inbound

Bridging the 2D-3D Gap: A Hierarchical Semantic-Geometric Map for Vision Language Navigation cites this paper.

Bridging the 2D-3D Gap: A Hierarchical Semantic-Geometric Map for Vision Language Navigation VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:34:01.336932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-29T22:33:32.671550Z digest=sha256:9cd748754c63a9f00986aee767f746bab757bf9cb469f029068f9763046d84f9

Observation b0a0df41-f434-4da9-90ef-c03458f72283 · inbound

Reasmory: 3D Reconstruction as Explicit Memory for VLMs Spatial Reasoning cites this paper.

Reasmory: 3D Reconstruction as Explicit Memory for VLMs Spatial Reasoning VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-07-01T20:46:13.683917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-28T17:46:21.821764Z digest=sha256:46a2949946089cf61e7ea3ef16276b7dd849395bf39322c41900575025df1414

Observation 59072d98-3d15-47d9-b66f-d43b0980ef2a · inbound

Goal2Pixel: Grounding Goals to Pixels for Vision-Language Navigation cites this paper.

Goal2Pixel: Grounding Goals to Pixels for Vision-Language Navigation VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:16:15.779796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-28T15:39:13.803571Z digest=sha256:b0bb53401fc64c0aa85d2c1ed7d2949116a44b5d1a32213186c7d15a0cfdcca9

Observation 39828a1e-b0db-486a-b879-ad91aba9459b · inbound

Beyond Waypoints: A Trajectory-Centric Waypointing Paradigm for Vision-Language Navigation cites this paper.

Beyond Waypoints: A Trajectory-Centric Waypointing Paradigm for Vision-Language Navigation VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-07-02T19:07:17.765948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-27T21:40:58.329546Z digest=sha256:c59f76e85f181db962c6ad441a06f17044fb977b8366a05a1ba16ef835234789

Observation 9ee761ed-2ee7-4e27-8aab-274aa4eee27c · inbound

Watch, Remember, Reason: Human-View Video Understanding with MLLMs cites this paper.

Watch, Remember, Reason: Human-View Video Understanding with MLLMs VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 239

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:27:15.728048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-27T22:00:28.350003Z digest=sha256:dfd163af974e101b89a6051940eefb09a378cb3451e57137842d0a3c84528fa8

Observation afdb57f2-0228-4d53-b770-cbc3316ed17e · inbound

Path-level Hindsight Instructions for Semantic Exploration in Vision-Language Navigation cites this paper.

Path-level Hindsight Instructions for Semantic Exploration in Vision-Language Navigation VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-07-03T14:08:21.390625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-03T14:04:46.400336Z digest=sha256:d66cdfed340ef5288bc258a0f7f17edda8bfc6ba9aad55890b6b5f57f13814a2

Observation 43227267-f85c-427b-b2e8-587b41454996 · inbound

A Comprehensive Survey and Systematic Real-World Evaluation of Embodied Vision-and-Language Navigation cites this paper.

A Comprehensive Survey and Systematic Real-World Evaluation of Embodied Vision-and-Language Navigation VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-14T15:39:59.878169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T15:39:59.878169Z digest=sha256:0300f7ead77f7a586dc829bab8e01a786fb50a36c7f7ee7976ca35eab25915d7

Observation 86424cd5-e548-4337-b63d-c8f01723c1ad · inbound

Joint On-and-Off Policy Learning for Vision-and-Language Navigation cites this paper.

Joint On-and-Off Policy Learning for Vision-and-Language Navigation VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T05:13:54.793452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:13:54.793452Z digest=sha256:830ad91d0990f25723b522fcff7cd7cdcca3516d2bfb7460a729d613ce559904

Observation 7b9be73c-fe2f-4bf4-b991-db29888e2848 · inbound

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation cites this paper.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:15.484911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:15.484911Z digest=sha256:57bfd4eb2db6dd4c215b9e7fedaeee1ec5cee1262fa582a460d2f429ed1f194e

Observation 918d788f-3643-4bfe-b999-c2ffe3ed2f8f · inbound

RecoverFly: A Failure-Aware Reinforcement Learning Post-Training Framework for Aerial Vision-Language Navigation cites this paper.

RecoverFly: A Failure-Aware Reinforcement Learning Post-Training Framework for Aerial Vision-Language Navigation VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T14:28:52.281460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:28:52.281460Z digest=sha256:ef1d9c105191506d6ec7592ae368d66798ac1eeb9a89d4483909abd6cbae30d5