Pith. sign in

Paper Citation Record · LEDGER

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models

As of 20 August 2026, this Paper Citation Record lists 25 of 25 outbound references and 4 inbound Pith citation observations for arXiv:2507.13152.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.13152 v3

Coverage vector

measured 25 of 25 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T16:33:49.266555Z

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T15:40:14.058419Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T11:26:54.069284Z

Reference resolution

25 of 25 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5f9fe887-6591-4c1b-8e93-3448667cc371 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:46.720398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:46.720398Z digest=sha256:1cba92a0fc8ddc1ccfa9885fe068c74b3ee5f35680dd6c0e9110aa7c26ec787e

Observation aaeb1df4-c2d2-46af-a4d1-50efa6cc7b47 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:51.984673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T16:33:46.816848Z digest=sha256:f50a8539c485317c8ca3c0cb2a76c4f5015fbef70706f00f7595996136a8a46e

Observation 59ef6e08-e7a3-4da4-ad2b-08db9a86ccd9 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:51.820383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T16:33:46.913912Z digest=sha256:169a3582fd4effb2c2f04523fb6e0c5c27b9507b343dee2ef4a1b774ef7ad15a

Observation 16dff492-3fd9-44d1-8154-15895b264f31 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:51.691448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T16:33:47.012239Z digest=sha256:ef1e25bc636158f058224eb6dc17ec2cb04ae42958226e8d7b46c998134c8bec

Observation aac10cbf-0581-4bf0-ad0f-ebe39c0ec765 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:51.521250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T16:33:47.095163Z digest=sha256:4b4ee3add0cb7cb3798cf9eda8e184707d0226296638dbb8cab4338328fc35e8

Observation 7fac222a-95ce-4707-8099-512751313678 · outbound

This paper cites LLM-Personalize: Aligning LLM Planners with Human Preferences via Reinforced Self-Training for Housekeeping Robots.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models LLM-Personalize: Aligning LLM Planners with Human Preferences via Reinforced Self-Training for Housekeeping Robots

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:47.179804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:47.179804Z digest=sha256:8254dbced9c793644c2e5f597523b201872e62d286620535fe6d42dddc7fcc29

Observation f5e63ab5-b490-41a1-94b7-fc045f30c9ce · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:47.314994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:47.314994Z digest=sha256:7adbff0f75b7252590b4c1f2628197c248b4dcc14f43e4359ab8e79e65df0c62

Observation a7e134fc-fc35-4bac-a656-719b82f7c550 · outbound

This paper cites LLM-Based Agent Society Investigation: Collaboration and Confrontation in Avalon Gameplay.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models LLM-Based Agent Society Investigation: Collaboration and Confrontation in Avalon Gameplay

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:47.458300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:47.458300Z digest=sha256:18bc3203be597625fe05194bb817a05fcc5ad95637d6acc3c72ab3fc21d0621c

Observation 39d8e93c-4499-46b4-896e-a073cfdf173b · outbound

This paper cites TINA: Think, Interaction, and Action Framework for Zero-Shot Vision Language Navigation.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models TINA: Think, Interaction, and Action Framework for Zero-Shot Vision Language Navigation

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-06T16:33:49.711508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T16:33:47.554701Z digest=sha256:1cefb5f49e3efcdc33714089b03e2266961cf52381d22c329c57cd391027916c

Observation 63af0041-836b-439a-8e69-22cf5b2a4dd3 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:51.189888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T16:33:47.724791Z digest=sha256:f2b8897e808776f09e7ebccc00b559b083b21212fbefde63486a57f83d5989b5

Observation 16615af1-33e5-49bc-adef-39c40a8ac3ad · outbound

This paper cites Vision-Language Navigation with Continual Learning.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Vision-Language Navigation with Continual Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:47.838131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:47.838131Z digest=sha256:777f7a0de570326e2874919d5f773e0c2c5cd07edfc8b83ac4b60495c2982129

Observation 7e6d2c21-c56e-4111-b48f-fa9dcd3ade3e · outbound

This paper cites NavCoT: Boosting LLM-Based Vision-and-Language Navigation via Learning Disentangled Reasoning.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models NavCoT: Boosting LLM-Based Vision-and-Language Navigation via Learning Disentangled Reasoning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:47.914909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:47.914909Z digest=sha256:ee008b4f42ff3fd5ad82aa2598731219039de6636e7544b6a0b4564f4a827674

Observation 048da78d-3987-4aa0-8148-c009d3163823 · outbound

This paper cites L.; Wei, Z.; Han, M.; Xu, R.; Niu, M.; Han, J.; Lin, L.; Lu, C.; et al.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models L.; Wei, Z.; Han, M.; Xu, R.; Niu, M.; Han, J.; Lin, L.; Lu, C.; et al

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:48.006215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:48.006215Z digest=sha256:c8515d1ad1e284daf01ee93c23511d0334590fead936b60575890fa9af6b5366

Observation 4ba1cb70-9dcb-45ef-81ec-1e5e054dbd95 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:50.906011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T16:33:48.098155Z digest=sha256:9360df58a308ae07ba879f3a48b4ee3439a8c43ea5c9e5e82b243caad68e9b67

Observation 43b13e29-43f6-4ff4-910b-1bbe209cb286 · outbound

This paper cites Y.; Shen, C.; and Hengel, A.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Y.; Shen, C.; and Hengel, A

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:48.186292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:48.186292Z digest=sha256:dd1dcf6abdedaef1bd94a3885b4f6e11bc1957ab4acf4d2e65dedc6e2de64524

Observation c71e3b80-476e-4aec-9d5f-6ad3118963f3 · outbound

This paper cites Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:48.266792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:48.266792Z digest=sha256:9f1f0ede7df14ca72d175f6e5daaeefb87d2a862fdc22c8b01205c7a1cf875f4

Observation b69ad065-c245-40f4-898e-5aaf401c8e45 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:50.640028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T16:33:48.347861Z digest=sha256:f92b3619713830eb2cee6908d713a6f969cb71215e5b50e0f2555a53772edfb0

Observation 696b91a4-3a57-4ce2-9d1f-559f166aeefd · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:50.312498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T16:33:48.445760Z digest=sha256:ed30987f3f3409a27a86d3fe78c81ae64692f3834af35ec3b33862b06fd72e0e

Observation 95e06cbb-0741-48ae-90d7-971d35853b72 · outbound

This paper cites Learning to Navigate Unseen Environments: Back Translation with Environmental Dropout.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Learning to Navigate Unseen Environments: Back Translation with Environmental Dropout

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:48.525238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:48.525238Z digest=sha256:35340239a395bcbf8bd52584c5ae582b4657a432ea2a5f73ddee1c8de5104723

Observation b14dfcbb-3d1c-4a27-a073-fce6af454110 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:48.640454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:48.640454Z digest=sha256:fe525cf481f830d6ccbd9eb8d956d34b0957406e92ba6e451d21d6c5fdf1db6d

Observation 5711c23d-ef90-4531-9079-ca52ce5795af · outbound

This paper cites MC-GPT: Empowering Vision-and-Language Navigation with Memory Map and Reasoning Chains.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models MC-GPT: Empowering Vision-and-Language Navigation with Memory Map and Reasoning Chains

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:48.794874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:48.794874Z digest=sha256:ae35746d7db4455ba956abb17a6d4a3d37f6e8f1ecf2993d57649427ac7ac618

Observation 0b0d1496-ebbc-45ad-9857-a97331296eca · outbound

This paper cites Agent-Pro: Learning to Evolve via Policy-Level Reflection and Optimization.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Agent-Pro: Learning to Evolve via Policy-Level Reflection and Optimization

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:48.946358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:48.946358Z digest=sha256:d659cea22a5eeef1723b32d6d8d4c4c89545c1828b008429708ff20a3f17e6bd

Observation c4274d8c-7f08-4095-850f-1fff80027225 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:49.998554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T16:33:49.068621Z digest=sha256:f12674c315eee4de210ab290b128f0ed6b0846c0d732d54d7b8c2a36cf4673f1

Observation 1f66ce62-3e29-43cf-8648-f9a4664093a6 · outbound

This paper cites , " * write output.state after.block = add.period write newline.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models , " * write output.state after.block = add.period write newline

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:49.166101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:49.166101Z digest=sha256:2c818535c97ff0caf8db9721acd694a9bf30d028eca0832c55c69c9589f6345c

Observation 6a65c450-a3f3-44cb-8f92-1ecab44f107f · outbound

This paper cites write newline.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models write newline

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:49.266555Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:49.266555Z digest=sha256:a3ded522e42bca3e63d6047e6dafbc70197449a67ae9412a4d5d94eca91d6da8

Pith citing papers

Observation 60ab9d06-911e-4185-a101-8960541bc770 · inbound

Plan in Sandbox, Navigate in Open Worlds: Learning Physics-Grounded Abstracted Experience for Embodied Navigation cites this paper.

Plan in Sandbox, Navigate in Open Worlds: Learning Physics-Grounded Abstracted Experience for Embodied Navigation SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:11:27.308950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-12T03:36:24.941205Z digest=sha256:db648ed457b4338177d02886351dbd3cf9946c8ef0cb5d0adbcda5385225b517

Observation 98762c74-a812-429d-b4dd-f485451161b8 · inbound

CLOSER-VLN: Closed-Loop Self-Verified Retrieval-Augmented Reasoning for Aerial Vision-Language Navigation cites this paper.

CLOSER-VLN: Closed-Loop Self-Verified Retrieval-Augmented Reasoning for Aerial Vision-Language Navigation SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models

Reference 26

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T15:25:47.562402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-30T01:35:04.288391Z digest=sha256:2e4ff723529ba025ff64311121f207cbe28b8c5f15b7349bacd1cb13fb47dc88

Observation f31a623b-0e8c-49cd-ab36-9ab66aeff250 · inbound

DART-VLN: Test-Time Memory Decay and Anti-Loop Regularization for Discrete Vision-Language Navigation cites this paper.

DART-VLN: Test-Time Memory Decay and Anti-Loop Regularization for Discrete Vision-Language Navigation SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T11:26:54.071949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-02T11:17:26.529397Z digest=sha256:5b063c882d0428d4b517e5e6f5670b969c9dde67fced771b459a6d98b0fc18a7

Observation cd8f63ce-f7f3-47e8-846e-fe68a4e5cc08 · inbound

DART-VLN: Test-Time Memory Decay and Anti-Loop Regularization for Discrete Vision-Language Navigation cites this paper.

DART-VLN: Test-Time Memory Decay and Anti-Loop Regularization for Discrete Vision-Language Navigation SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T15:40:14.058419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:40:14.058419Z digest=sha256:1ae0e8a2dc3911688941b575d200c1c5895435a899b708ef4d8c7ae480395b52