Pith. sign in

Paper Citation Record · LEDGER

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models

As of 7 August 2026, this Paper Citation Record lists 25 of 25 outbound references and 3 inbound Pith citation observations for arXiv:2507.13152.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.13152 v3

Coverage vector

measured 25 of 25 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T16:33:49.266555Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-02T11:17:26.529397Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T11:26:54.069284Z

Reference resolution

25 of 25 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5f9fe887-6591-4c1b-8e93-3448667cc371 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:46.720398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:46.720398Z digest=sha256:f053252a98d0cec0d30652b0ff8968a6007a672febf361238fc7fa3120b7ec36

Observation aaeb1df4-c2d2-46af-a4d1-50efa6cc7b47 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:51.984673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T16:33:46.816848Z digest=sha256:d33cec5f243b5b4db218f51c2ea8dbf22f4d5d2800f354f545106bbd3890d0f6

Observation 59ef6e08-e7a3-4da4-ad2b-08db9a86ccd9 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:51.820383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T16:33:46.913912Z digest=sha256:c57e681dc0071bbe42546fa529f2a4893613e32e197dd00f7b26177a77ee3bb5

Observation 16dff492-3fd9-44d1-8154-15895b264f31 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:51.691448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T16:33:47.012239Z digest=sha256:f19b7808aa0ebe179ed59185312f10f369156cc7db4fe4f4bcb4699dcb1fb9f6

Observation aac10cbf-0581-4bf0-ad0f-ebe39c0ec765 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:51.521250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T16:33:47.095163Z digest=sha256:370500dce6974ed47c872a9a365783863dc1d928330ab74bff357e1f15c102e5

Observation 7fac222a-95ce-4707-8099-512751313678 · outbound

This paper cites LLM-Personalize: Aligning LLM Planners with Human Preferences via Reinforced Self-Training for Housekeeping Robots.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models LLM-Personalize: Aligning LLM Planners with Human Preferences via Reinforced Self-Training for Housekeeping Robots

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:47.179804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:47.179804Z digest=sha256:067fad7fea4531dadea74ec82db49263c47f190c4113a510892a18a4be067966

Observation f5e63ab5-b490-41a1-94b7-fc045f30c9ce · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:47.314994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:47.314994Z digest=sha256:0b58606040da94de54ba1dfcac47788b2fd918fac687ccc4e3ccb0afc77b6efc

Observation a7e134fc-fc35-4bac-a656-719b82f7c550 · outbound

This paper cites LLM-Based Agent Society Investigation: Collaboration and Confrontation in Avalon Gameplay.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models LLM-Based Agent Society Investigation: Collaboration and Confrontation in Avalon Gameplay

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:47.458300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:47.458300Z digest=sha256:aee483b0652fcee044b12302e3d3768936c1e0fb619590bf86cc02148437ef86

Observation 39d8e93c-4499-46b4-896e-a073cfdf173b · outbound

This paper cites TINA: Think, Interaction, and Action Framework for Zero-Shot Vision Language Navigation.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models TINA: Think, Interaction, and Action Framework for Zero-Shot Vision Language Navigation

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-06T16:33:49.711508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T16:33:47.554701Z digest=sha256:a7849bd362c24e28688aed76ee02df66febe8a4a5974fd3ab3b6abe6b661a62b

Observation 63af0041-836b-439a-8e69-22cf5b2a4dd3 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:51.189888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T16:33:47.724791Z digest=sha256:47d1ceda931118502604303f6a1a5103e92a7a527f92b6395d23ee1f8188aef6

Observation 16615af1-33e5-49bc-adef-39c40a8ac3ad · outbound

This paper cites Vision-Language Navigation with Continual Learning.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Vision-Language Navigation with Continual Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:47.838131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:47.838131Z digest=sha256:287d74a283aa4a4b220623abbd9950bd3a1ba34a8abe5a4c313941c50920049b

Observation 7e6d2c21-c56e-4111-b48f-fa9dcd3ade3e · outbound

This paper cites NavCoT: Boosting LLM-Based Vision-and-Language Navigation via Learning Disentangled Reasoning.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models NavCoT: Boosting LLM-Based Vision-and-Language Navigation via Learning Disentangled Reasoning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:47.914909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:47.914909Z digest=sha256:bd3b936841d5044423662d996799219c11a93915d61777c60684fca72bd9ce1f

Observation 048da78d-3987-4aa0-8148-c009d3163823 · outbound

This paper cites L.; Wei, Z.; Han, M.; Xu, R.; Niu, M.; Han, J.; Lin, L.; Lu, C.; et al.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models L.; Wei, Z.; Han, M.; Xu, R.; Niu, M.; Han, J.; Lin, L.; Lu, C.; et al

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:48.006215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:48.006215Z digest=sha256:0a88cdbd24f3cedaa68d5f9c78e7466b51a0da27fa0858d92b7f938946e95c76

Observation 4ba1cb70-9dcb-45ef-81ec-1e5e054dbd95 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:50.906011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T16:33:48.098155Z digest=sha256:d86c20f1d5c41c4ecd2dade7e2ed977c3632ea5d853c08624a765b177f0e7b57

Observation 43b13e29-43f6-4ff4-910b-1bbe209cb286 · outbound

This paper cites Y.; Shen, C.; and Hengel, A.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Y.; Shen, C.; and Hengel, A

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:48.186292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:48.186292Z digest=sha256:24f9acc7a9a4368f67bb5c681d955417f6cbe2e47cc60f888090c59dc90d3c14

Observation c71e3b80-476e-4aec-9d5f-6ad3118963f3 · outbound

This paper cites Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:48.266792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:48.266792Z digest=sha256:a7aef0daee67322c67acdb87536d70f4ea934163515e08d8e04cd5a04867cc82

Observation b69ad065-c245-40f4-898e-5aaf401c8e45 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:50.640028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T16:33:48.347861Z digest=sha256:52ae63a43517027acd56cc54519dc6d6b3d974a6778d3075a45331bf71527b66

Observation 696b91a4-3a57-4ce2-9d1f-559f166aeefd · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:50.312498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T16:33:48.445760Z digest=sha256:1c7fe86494bd8485f81dbe5eb842e00f4c67298e8879999e376481a7bfb41892

Observation 95e06cbb-0741-48ae-90d7-971d35853b72 · outbound

This paper cites Learning to Navigate Unseen Environments: Back Translation with Environmental Dropout.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Learning to Navigate Unseen Environments: Back Translation with Environmental Dropout

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:48.525238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:48.525238Z digest=sha256:1e9822544bc84145f14e2ae4d5455389a36e1e4e9a7f01f84a2e6a583aff7e02

Observation b14dfcbb-3d1c-4a27-a073-fce6af454110 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:48.640454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:48.640454Z digest=sha256:b7fac1519bfab55db1b81de4bd1d9388f1727e0b64e9a7bda1cf0110a7fc2543

Observation 5711c23d-ef90-4531-9079-ca52ce5795af · outbound

This paper cites MC-GPT: Empowering Vision-and-Language Navigation with Memory Map and Reasoning Chains.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models MC-GPT: Empowering Vision-and-Language Navigation with Memory Map and Reasoning Chains

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:48.794874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:48.794874Z digest=sha256:c084929e5bc0f3c2aa2cbb952e0296d5378e3af75ed0c2900ac8d29dae08fc1b

Observation 0b0d1496-ebbc-45ad-9857-a97331296eca · outbound

This paper cites Agent-Pro: Learning to Evolve via Policy-Level Reflection and Optimization.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Agent-Pro: Learning to Evolve via Policy-Level Reflection and Optimization

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:48.946358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:48.946358Z digest=sha256:8b95f511c975ea6e4410a4506cae20aaefa1e16c9c131fb7afac5f1076a9dca9

Observation c4274d8c-7f08-4095-850f-1fff80027225 · outbound

This paper cites an unresolved cited work.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:33:49.998554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T16:33:49.068621Z digest=sha256:1daeee7c0bf27ea7c2dce39c5b6db283c67871309135c3c1702f49409d100a08

Observation 1f66ce62-3e29-43cf-8648-f9a4664093a6 · outbound

This paper cites , " * write output.state after.block = add.period write newline.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models , " * write output.state after.block = add.period write newline

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:49.166101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:49.166101Z digest=sha256:496ddd8d849a09b978f6c0c01bd64365a6db98452df09b1b310e2c49f3c130bb

Observation 6a65c450-a3f3-44cb-8f92-1ecab44f107f · outbound

This paper cites write newline.

SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models write newline

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:49.266555Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:33:49.266555Z digest=sha256:7b8d4eb7c230b908aea005c061c4b452690a1b87adbc561f03d23426c0b2ca5d

Pith citing papers

Observation 60ab9d06-911e-4185-a101-8960541bc770 · inbound

Plan in Sandbox, Navigate in Open Worlds: Learning Physics-Grounded Abstracted Experience for Embodied Navigation cites this paper.

Plan in Sandbox, Navigate in Open Worlds: Learning Physics-Grounded Abstracted Experience for Embodied Navigation SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:11:27.308950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-12T03:36:24.941205Z digest=sha256:8b79b3aacc169d7693aed3f7c2586207554a75b0b5da9083131665e16eed82f6

Observation 98762c74-a812-429d-b4dd-f485451161b8 · inbound

CLOSER-VLN: Closed-Loop Self-Verified Retrieval-Augmented Reasoning for Aerial Vision-Language Navigation cites this paper.

CLOSER-VLN: Closed-Loop Self-Verified Retrieval-Augmented Reasoning for Aerial Vision-Language Navigation SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models

Reference 26

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T15:25:47.562402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T01:35:04.288391Z digest=sha256:26780eee2eba47ac97ffbf51326c8ec94f9f56c5bd9bc5621b034d1c05c920e3

Observation f31a623b-0e8c-49cd-ab36-9ab66aeff250 · inbound

DART-VLN: Test-Time Memory Decay and Anti-Loop Regularization for Discrete Vision-Language Navigation cites this paper.

DART-VLN: Test-Time Memory Decay and Anti-Loop Regularization for Discrete Vision-Language Navigation SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T11:26:54.071949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-02T11:17:26.529397Z digest=sha256:2a6e49e765648841dad2ac578f8a8ab3b62e6611e422c4272197964d18a586f5