Pith. sign in

Paper Citation Record · LEDGER

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models

As of 10 August 2026, this Paper Citation Record lists 25 of 25 outbound references and 3 inbound Pith citation observations for arXiv:2507.13152.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.13152 v3

Coverage vector

measured 25 of 25 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T16:33:49.266555Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-02T11:17:26.529397Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T11:26:54.069284Z

Reference resolution

25 of 25 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5f9fe887-6591-4c1b-8e93-3448667cc371 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:46.720398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:46.720398Z digest=sha256:0911b1b89b5570ca9718458c602eaacfa9433440305eaf5a2a82db324392a5fa

Observation aaeb1df4-c2d2-46af-a4d1-50efa6cc7b47 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:51.984673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T16:33:46.816848Z digest=sha256:665d997bd84b8e30432333b78aec660a4edc7084ad6c83c02e82f622e2cba0bc

Observation 59ef6e08-e7a3-4da4-ad2b-08db9a86ccd9 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:51.820383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T16:33:46.913912Z digest=sha256:66ff71fe063d20f2ca8f0827608544eba76fd58f34a07e381878a237c91c1042

Observation 16dff492-3fd9-44d1-8154-15895b264f31 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:51.691448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T16:33:47.012239Z digest=sha256:ee3ebb50d623191b157707223e9f46aa72030fa8abb64f867c2d81d4308945af

Observation aac10cbf-0581-4bf0-ad0f-ebe39c0ec765 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:51.521250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T16:33:47.095163Z digest=sha256:49723806778c2c2e6d1d2960e3499067d5c91d9393562c32a825aac329392814

Observation 7fac222a-95ce-4707-8099-512751313678 · outbound

This paper cites LLM-Personalize: Aligning LLM Planners with Human Preferences via Reinforced Self-Training for Housekeeping Robots.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models LLM-Personalize: Aligning LLM Planners with Human Preferences via Reinforced Self-Training for Housekeeping Robots

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:47.179804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:47.179804Z digest=sha256:89fd94abd4dcfc599a32429a2467ea1dab7860b39095989348a058d942ac3923

Observation f5e63ab5-b490-41a1-94b7-fc045f30c9ce · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:47.314994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:47.314994Z digest=sha256:4e307091d30286b135c32d9bdb9085bcb71071f674675d51e7ebeac5d6ca30f5

Observation a7e134fc-fc35-4bac-a656-719b82f7c550 · outbound

This paper cites LLM-Based Agent Society Investigation: Collaboration and Confrontation in Avalon Gameplay.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models LLM-Based Agent Society Investigation: Collaboration and Confrontation in Avalon Gameplay

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:47.458300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:47.458300Z digest=sha256:41449eaec191c5c014a782ba6bac0017917912751096f06469a4ce00f6831175

Observation 39d8e93c-4499-46b4-896e-a073cfdf173b · outbound

This paper cites TINA: Think, Interaction, and Action Framework for Zero-Shot Vision Language Navigation.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models TINA: Think, Interaction, and Action Framework for Zero-Shot Vision Language Navigation

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-06T16:33:49.711508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T16:33:47.554701Z digest=sha256:67b8cf90ea94d02fa13faaa63c05c4c97e6c0b0f66f46c8d6c0672e34e20b1e7

Observation 63af0041-836b-439a-8e69-22cf5b2a4dd3 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:51.189888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T16:33:47.724791Z digest=sha256:73954e07b7f3caa5f5216ae8f149d7dd6fa95a83bc27e7bc37b2ca702cf01951

Observation 16615af1-33e5-49bc-adef-39c40a8ac3ad · outbound

This paper cites Vision-Language Navigation with Continual Learning.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Vision-Language Navigation with Continual Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:47.838131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:47.838131Z digest=sha256:a437a194b53112a71b1929e5dbfc289c4a198edf31b231cfc4f009b22deaff91

Observation 7e6d2c21-c56e-4111-b48f-fa9dcd3ade3e · outbound

This paper cites NavCoT: Boosting LLM-Based Vision-and-Language Navigation via Learning Disentangled Reasoning.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models NavCoT: Boosting LLM-Based Vision-and-Language Navigation via Learning Disentangled Reasoning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:47.914909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:47.914909Z digest=sha256:2a68704a84943705bcb3025cbba6d9aedb6e0d34499100fe36bc891d60f1d8a7

Observation 048da78d-3987-4aa0-8148-c009d3163823 · outbound

This paper cites L.; Wei, Z.; Han, M.; Xu, R.; Niu, M.; Han, J.; Lin, L.; Lu, C.; et al.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models L.; Wei, Z.; Han, M.; Xu, R.; Niu, M.; Han, J.; Lin, L.; Lu, C.; et al

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:48.006215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:48.006215Z digest=sha256:a9f7067094d813c770c933c3de5f1c7fa4c89e3b4a3a8f1361d1b129033185b5

Observation 4ba1cb70-9dcb-45ef-81ec-1e5e054dbd95 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:50.906011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T16:33:48.098155Z digest=sha256:ee69a64ea00d5800933b94eace92228a2612adfb616351c2df95bd07d0b5fa34

Observation 43b13e29-43f6-4ff4-910b-1bbe209cb286 · outbound

This paper cites Y.; Shen, C.; and Hengel, A.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Y.; Shen, C.; and Hengel, A

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:48.186292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:48.186292Z digest=sha256:c56511b337fc2c05964452707500a6470229707b650c51458735c68e3f82cacb

Observation c71e3b80-476e-4aec-9d5f-6ad3118963f3 · outbound

This paper cites Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:48.266792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:48.266792Z digest=sha256:b88902226eed9436d87749a75dd93a0080580cfb481710655e3d5bb5606a5fb5

Observation b69ad065-c245-40f4-898e-5aaf401c8e45 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:50.640028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T16:33:48.347861Z digest=sha256:413ad1b98aa440a212a5dbb897113048e843a5f9615bbd4ed320a38a1cc3de61

Observation 696b91a4-3a57-4ce2-9d1f-559f166aeefd · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:50.312498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T16:33:48.445760Z digest=sha256:ea3bb37db9d8ae47458a132541b27e79ab62bc3cbfc0a02e85108f779dc02857

Observation 95e06cbb-0741-48ae-90d7-971d35853b72 · outbound

This paper cites Learning to Navigate Unseen Environments: Back Translation with Environmental Dropout.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Learning to Navigate Unseen Environments: Back Translation with Environmental Dropout

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:48.525238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:48.525238Z digest=sha256:582c77848d9c8924c88cf891c6e77f6c1771022f466510ad17e27551c6037922

Observation b14dfcbb-3d1c-4a27-a073-fce6af454110 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:48.640454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:48.640454Z digest=sha256:948c5c8e44f3775a9d2271a9373d43a4634f77a5784b1566c360a03659ca3a86

Observation 5711c23d-ef90-4531-9079-ca52ce5795af · outbound

This paper cites MC-GPT: Empowering Vision-and-Language Navigation with Memory Map and Reasoning Chains.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models MC-GPT: Empowering Vision-and-Language Navigation with Memory Map and Reasoning Chains

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:48.794874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:48.794874Z digest=sha256:079bf6677ead5e0aafa8270437a58434125160891334b66ea88ce1c14100db22

Observation 0b0d1496-ebbc-45ad-9857-a97331296eca · outbound

This paper cites Agent-Pro: Learning to Evolve via Policy-Level Reflection and Optimization.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Agent-Pro: Learning to Evolve via Policy-Level Reflection and Optimization

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:48.946358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:48.946358Z digest=sha256:c1282ae17f60546f7b1f8441977b776e866a05f035011014bc71377fff5460e4

Observation c4274d8c-7f08-4095-850f-1fff80027225 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:49.998554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T16:33:49.068621Z digest=sha256:a740ef70426f3fffdc24459ac9d7c05bdfd902396c26f2aff0e83a819a1b0308

Observation 1f66ce62-3e29-43cf-8648-f9a4664093a6 · outbound

This paper cites , " * write output.state after.block = add.period write newline.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models , " * write output.state after.block = add.period write newline

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:49.166101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:49.166101Z digest=sha256:c0d4c8fd4da12ce893047902e6ecb643dd61b1ccc81b59bd7773569ebb866815

Observation 6a65c450-a3f3-44cb-8f92-1ecab44f107f · outbound

This paper cites write newline.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models write newline

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:49.266555Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:49.266555Z digest=sha256:926134e2e832fb239876607ba76599914cff05678c88eaf375de0181284065e0

Pith citing papers

Observation 60ab9d06-911e-4185-a101-8960541bc770 · inbound

Plan in Sandbox, Navigate in Open Worlds: Learning Physics-Grounded Abstracted Experience for Embodied Navigation cites this paper.

Plan in Sandbox, Navigate in Open Worlds: Learning Physics-Grounded Abstracted Experience for Embodied Navigation SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:11:27.308950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-12T03:36:24.941205Z digest=sha256:50dbacc7fa3d5ef69d3034b927725dbaffa5a61b7247c0f3f241e75beab45943

Observation 98762c74-a812-429d-b4dd-f485451161b8 · inbound

CLOSER-VLN: Closed-Loop Self-Verified Retrieval-Augmented Reasoning for Aerial Vision-Language Navigation cites this paper.

CLOSER-VLN: Closed-Loop Self-Verified Retrieval-Augmented Reasoning for Aerial Vision-Language Navigation SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models

Reference 26

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T15:25:47.562402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T01:35:04.288391Z digest=sha256:7377ecc45534f0c83eb956c19b51e81d1ea3ceeb23be7bb6bcf8abb7636c50b7

Observation f31a623b-0e8c-49cd-ab36-9ab66aeff250 · inbound

DART-VLN: Test-Time Memory Decay and Anti-Loop Regularization for Discrete Vision-Language Navigation cites this paper.

DART-VLN: Test-Time Memory Decay and Anti-Loop Regularization for Discrete Vision-Language Navigation SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T11:26:54.071949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-02T11:17:26.529397Z digest=sha256:6da69652074504bc5ad589d4d2534fb1869a08f71aee455c77ca08dd096b01a8