Pith. sign in

Paper Citation Record · LEDGER

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models

As of 8 August 2026, this Paper Citation Record lists 25 of 25 outbound references and 3 inbound Pith citation observations for arXiv:2507.13152.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.13152 v3

Coverage vector

measured 25 of 25 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T16:33:49.266555Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-02T11:17:26.529397Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T11:26:54.069284Z

Reference resolution

25 of 25 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5f9fe887-6591-4c1b-8e93-3448667cc371 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:46.720398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:46.720398Z digest=sha256:f053252a98d0cec0d30652b0ff8968a6007a672febf361238fc7fa3120b7ec36

Observation aaeb1df4-c2d2-46af-a4d1-50efa6cc7b47 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:51.984673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T16:33:46.816848Z digest=sha256:26a8b855627c36b134e8860c1d5fdf0e5c9b538f92f8dd39845e5d06b1d06256

Observation 59ef6e08-e7a3-4da4-ad2b-08db9a86ccd9 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:51.820383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T16:33:46.913912Z digest=sha256:04400653b7003adba6ae3de9941fad3b45c4e183ce4faace783a688bb39d37a0

Observation 16dff492-3fd9-44d1-8154-15895b264f31 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:51.691448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T16:33:47.012239Z digest=sha256:c03947181a4ab43903ecd4f64b661a39bbe75b4ccdc5adb60ba206ea90f8ea25

Observation aac10cbf-0581-4bf0-ad0f-ebe39c0ec765 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:51.521250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T16:33:47.095163Z digest=sha256:6778fa44b4e71bc275ec4030e4e9380af5078e135bf0a4ce652e74ca299f059f

Observation 7fac222a-95ce-4707-8099-512751313678 · outbound

This paper cites LLM-Personalize: Aligning LLM Planners with Human Preferences via Reinforced Self-Training for Housekeeping Robots.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models LLM-Personalize: Aligning LLM Planners with Human Preferences via Reinforced Self-Training for Housekeeping Robots

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:47.179804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:47.179804Z digest=sha256:067fad7fea4531dadea74ec82db49263c47f190c4113a510892a18a4be067966

Observation f5e63ab5-b490-41a1-94b7-fc045f30c9ce · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:47.314994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:47.314994Z digest=sha256:0b58606040da94de54ba1dfcac47788b2fd918fac687ccc4e3ccb0afc77b6efc

Observation a7e134fc-fc35-4bac-a656-719b82f7c550 · outbound

This paper cites LLM-Based Agent Society Investigation: Collaboration and Confrontation in Avalon Gameplay.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models LLM-Based Agent Society Investigation: Collaboration and Confrontation in Avalon Gameplay

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:47.458300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:47.458300Z digest=sha256:aee483b0652fcee044b12302e3d3768936c1e0fb619590bf86cc02148437ef86

Observation 39d8e93c-4499-46b4-896e-a073cfdf173b · outbound

This paper cites TINA: Think, Interaction, and Action Framework for Zero-Shot Vision Language Navigation.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models TINA: Think, Interaction, and Action Framework for Zero-Shot Vision Language Navigation

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-06T16:33:49.711508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T16:33:47.554701Z digest=sha256:91d51fc3e6ac0893d7ea5de69e15e2b122d829be34220b46d985db59d94883de

Observation 63af0041-836b-439a-8e69-22cf5b2a4dd3 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:51.189888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T16:33:47.724791Z digest=sha256:ba707fb2afd4848bd3febaf37b81e359facc2953a932d0e530837489544061c2

Observation 16615af1-33e5-49bc-adef-39c40a8ac3ad · outbound

This paper cites Vision-Language Navigation with Continual Learning.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Vision-Language Navigation with Continual Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:47.838131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:47.838131Z digest=sha256:287d74a283aa4a4b220623abbd9950bd3a1ba34a8abe5a4c313941c50920049b

Observation 7e6d2c21-c56e-4111-b48f-fa9dcd3ade3e · outbound

This paper cites NavCoT: Boosting LLM-Based Vision-and-Language Navigation via Learning Disentangled Reasoning.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models NavCoT: Boosting LLM-Based Vision-and-Language Navigation via Learning Disentangled Reasoning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:47.914909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:47.914909Z digest=sha256:bd3b936841d5044423662d996799219c11a93915d61777c60684fca72bd9ce1f

Observation 048da78d-3987-4aa0-8148-c009d3163823 · outbound

This paper cites L.; Wei, Z.; Han, M.; Xu, R.; Niu, M.; Han, J.; Lin, L.; Lu, C.; et al.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models L.; Wei, Z.; Han, M.; Xu, R.; Niu, M.; Han, J.; Lin, L.; Lu, C.; et al

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:48.006215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:48.006215Z digest=sha256:0a88cdbd24f3cedaa68d5f9c78e7466b51a0da27fa0858d92b7f938946e95c76

Observation 4ba1cb70-9dcb-45ef-81ec-1e5e054dbd95 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:50.906011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T16:33:48.098155Z digest=sha256:071636c840898fadc93d8faf105d0f1978308e970204b49f12cc6afa8b6d6dd4

Observation 43b13e29-43f6-4ff4-910b-1bbe209cb286 · outbound

This paper cites Y.; Shen, C.; and Hengel, A.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Y.; Shen, C.; and Hengel, A

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:48.186292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:48.186292Z digest=sha256:24f9acc7a9a4368f67bb5c681d955417f6cbe2e47cc60f888090c59dc90d3c14

Observation c71e3b80-476e-4aec-9d5f-6ad3118963f3 · outbound

This paper cites Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:48.266792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:48.266792Z digest=sha256:a7aef0daee67322c67acdb87536d70f4ea934163515e08d8e04cd5a04867cc82

Observation b69ad065-c245-40f4-898e-5aaf401c8e45 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:50.640028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T16:33:48.347861Z digest=sha256:61c8d55be398cdeb4b8d82c207588f585bb13fafea0589f74bf15653eace4859

Observation 696b91a4-3a57-4ce2-9d1f-559f166aeefd · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:50.312498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T16:33:48.445760Z digest=sha256:e16e61c5859cd262597479e2b1cf00b8bd77516e8aa44e56d3a306b0100edc6e

Observation 95e06cbb-0741-48ae-90d7-971d35853b72 · outbound

This paper cites Learning to Navigate Unseen Environments: Back Translation with Environmental Dropout.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Learning to Navigate Unseen Environments: Back Translation with Environmental Dropout

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:48.525238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:48.525238Z digest=sha256:1e9822544bc84145f14e2ae4d5455389a36e1e4e9a7f01f84a2e6a583aff7e02

Observation b14dfcbb-3d1c-4a27-a073-fce6af454110 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:48.640454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:48.640454Z digest=sha256:b7fac1519bfab55db1b81de4bd1d9388f1727e0b64e9a7bda1cf0110a7fc2543

Observation 5711c23d-ef90-4531-9079-ca52ce5795af · outbound

This paper cites MC-GPT: Empowering Vision-and-Language Navigation with Memory Map and Reasoning Chains.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models MC-GPT: Empowering Vision-and-Language Navigation with Memory Map and Reasoning Chains

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:48.794874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:48.794874Z digest=sha256:c084929e5bc0f3c2aa2cbb952e0296d5378e3af75ed0c2900ac8d29dae08fc1b

Observation 0b0d1496-ebbc-45ad-9857-a97331296eca · outbound

This paper cites Agent-Pro: Learning to Evolve via Policy-Level Reflection and Optimization.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Agent-Pro: Learning to Evolve via Policy-Level Reflection and Optimization

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:48.946358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:48.946358Z digest=sha256:8b95f511c975ea6e4410a4506cae20aaefa1e16c9c131fb7afac5f1076a9dca9

Observation c4274d8c-7f08-4095-850f-1fff80027225 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:49.998554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T16:33:49.068621Z digest=sha256:4832f47241a4da6535a6fe9805474188b9c4afcc80ea14bd31a20ee1a95f0034

Observation 1f66ce62-3e29-43cf-8648-f9a4664093a6 · outbound

This paper cites , " * write output.state after.block = add.period write newline.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models , " * write output.state after.block = add.period write newline

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:49.166101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:49.166101Z digest=sha256:496ddd8d849a09b978f6c0c01bd64365a6db98452df09b1b310e2c49f3c130bb

Observation 6a65c450-a3f3-44cb-8f92-1ecab44f107f · outbound

This paper cites write newline.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models write newline

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:49.266555Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:49.266555Z digest=sha256:7b8d4eb7c230b908aea005c061c4b452690a1b87adbc561f03d23426c0b2ca5d

Pith citing papers

Observation 60ab9d06-911e-4185-a101-8960541bc770 · inbound

Plan in Sandbox, Navigate in Open Worlds: Learning Physics-Grounded Abstracted Experience for Embodied Navigation cites this paper.

Plan in Sandbox, Navigate in Open Worlds: Learning Physics-Grounded Abstracted Experience for Embodied Navigation SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:11:27.308950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-12T03:36:24.941205Z digest=sha256:ab0e399ad4932e0de090ae3ac9300d86d55cffdd6391f27c3eb99cd1ad4f221b

Observation 98762c74-a812-429d-b4dd-f485451161b8 · inbound

CLOSER-VLN: Closed-Loop Self-Verified Retrieval-Augmented Reasoning for Aerial Vision-Language Navigation cites this paper.

CLOSER-VLN: Closed-Loop Self-Verified Retrieval-Augmented Reasoning for Aerial Vision-Language Navigation SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models

Reference 26

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T15:25:47.562402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T01:35:04.288391Z digest=sha256:96d194c044ff3ffb5fded2e4281e08b5429d5efe0c973dc6b99eb071edbf62ed

Observation f31a623b-0e8c-49cd-ab36-9ab66aeff250 · inbound

DART-VLN: Test-Time Memory Decay and Anti-Loop Regularization for Discrete Vision-Language Navigation cites this paper.

DART-VLN: Test-Time Memory Decay and Anti-Loop Regularization for Discrete Vision-Language Navigation SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T11:26:54.071949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-02T11:17:26.529397Z digest=sha256:ae99c4ab7affd35f428c039142784a8981a774f44f8c0b1d0bc79ebd8a7a9015