Pith. sign in

Paper Citation Record · LEDGER

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model

As of 18 August 2026, this Paper Citation Record lists 87 of 87 outbound references and 3 inbound Pith citation observations for arXiv:2412.20903.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.20903 v4

Coverage vector

measured 87 of 87 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T23:12:19.269126Z

measured 90 of 90 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:43:47.468600Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T10:27:19.601310Z

Reference resolution

87 of 87 outbound references displayed

  • verified exact1
  • verified fuzzy39
  • unresolved47
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f52fc303-8b30-48d9-9058-601391939d1c · outbound

This paper cites GPT-4 Technical Report.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:18.822707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:18.822707Z digest=sha256:fd1f63261182e65fb5a50c078acbfee730e2f231d37ed4dab20dcfefb6df7d7f

Observation 7202eaf0-a26a-484f-8222-db5c74ce1942 · outbound

This paper cites Vita: An efficient video-to-text algorithm using vlm for rag-based video analysis system.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Vita: An efficient video-to-text algorithm using vlm for rag-based video analysis system

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:18.828631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:18.828631Z digest=sha256:31e285962426107688a555a60f9d5d32566be00c1a8c72c7e423b75cf4a30220

Observation 11a9bc38-acf2-4145-83cc-3a3e341bdaf5 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:18.834185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:18.834185Z digest=sha256:b9787b3995601413f8c0a2bd9d1cab5b8adebb2029ced5642a570c12bfaa4289

Observation 0ccfd3ce-1ff7-4a24-abb1-826d6c43ba69 · outbound

This paper cites An Introduction to Vision-Language Modeling.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model An Introduction to Vision-Language Modeling

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:18.839807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:18.839807Z digest=sha256:733e60c1960887a2d9a4d33ac7b487a303d0aaea8d4013d03901778bbcd12b35

Observation 06b2125e-953a-4853-bb03-830085277315 · outbound

This paper cites Visual challenges in the ev- eryday lives of blind people.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Visual challenges in the ev- eryday lives of blind people

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:18.845410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:18.845410Z digest=sha256:ce508713fe3b441a085bdaf203a7c1773e0f55379f0cb7cf71502d8aca34250a

Observation 9a9e43a5-f4e5-409f-b179-afeb400b7982 · outbound

This paper cites LLaVA-MoLE: Sparse Mixture of LoRA Experts for Mitigating Data Conflicts in Instruction Finetuning MLLMs.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model LLaVA-MoLE: Sparse Mixture of LoRA Experts for Mitigating Data Conflicts in Instruction Finetuning MLLMs

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:18.850612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:18.850612Z digest=sha256:4f5e7fdf78804102736f2e4921d268131af1f695c8a7e3c2b0f323f8f4343cb6

Observation 0001c851-8763-4979-b3c1-820ceeda2815 · outbound

This paper cites Vqask: a mul- timodal android gpt-based application to help blind users vi- sualize pictures.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Vqask: a mul- timodal android gpt-based application to help blind users vi- sualize pictures

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:18.856175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:18.856175Z digest=sha256:b3193b4dad1bd1e2351e1a8a6d24fbe2ff527d970798ec5ba9c161eeb3766d70

Observation cb97ee31-e4fb-4f8b-a4a4-10a0bad6c347 · outbound

This paper cites A Survey of Vision-Language Pre-Trained Models.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model A Survey of Vision-Language Pre-Trained Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:18.861631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:18.861631Z digest=sha256:a2030930790037d3772be46888222ffd4a310b1faf088eb4c5a5f3bba4807c24

Observation 586200d8-9149-4ea1-88bf-fa3a01093d21 · outbound

This paper cites The Llama 3 Herd of Models.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model The Llama 3 Herd of Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:18.867167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:18.867167Z digest=sha256:f5b995511f183e272d0e88044b2fa658eb9ba39ea598f9605d6303a280c1f802

Observation a4401e45-68f3-4709-8f05-a00c197ca5ed · outbound

This paper cites Retrieval-Augmented Generation for Large Language Models: A Survey.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Retrieval-Augmented Generation for Large Language Models: A Survey

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:18.873026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:18.873026Z digest=sha256:265a09d2e064866401ab214516b635c7374d3b086bcc8fa55a63fbc06efbc99b

Observation 6af24839-795c-4ce7-9201-21a9350a9e55 · outbound

This paper cites Use of an indoor navigation system by sighted and blind trav- elers: Performance similarities across visual status and age.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Use of an indoor navigation system by sighted and blind trav- elers: Performance similarities across visual status and age

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:18.878194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:18.878194Z digest=sha256:302aa94482cc5a4e13fccfe901e9403ff166b8433a1c2984c02cce2cf1c4604c

Observation 017c7324-1de4-405d-a529-f47cca0d25a5 · outbound

This paper cites Lvis: A dataset for large vocabulary instance segmentation.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Lvis: A dataset for large vocabulary instance segmentation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:18.883014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:18.883014Z digest=sha256:287cf6b1a2866bddd0e17570dea1c8436bd9b6441269b19bb8f52ece16a54498

Observation df14e819-ff0b-4714-80f8-d7086bb0599e · outbound

This paper cites Vizwiz grand challenge: Answering visual questions from blind people.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Vizwiz grand challenge: Answering visual questions from blind people

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:18.887823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:18.887823Z digest=sha256:de04eb644ada8e4ec8d9f4c8994a2f46d66599adbab46ef45cc2011c3dd672bb

Observation 2d3d64e1-afd2-4c35-bea6-aa88973b9ae5 · outbound

This paper cites Assistive technol- ogy for visually impaired and blind people.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Assistive technol- ogy for visually impaired and blind people

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:18.893207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:18.893207Z digest=sha256:8c133952b3b1b391d2ba80231d8e7c6b79ae19d3b0d4971818703dbdba9b9542

Observation 0a1dc93b-3d1e-4b5c-bd34-f77659eed05d · outbound

This paper cites GPT-4o System Card.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model GPT-4o System Card

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:18.898031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:18.898031Z digest=sha256:9a5214c0762269f22135da56b4d757d0a00f7c13310bc21ff7a75cb74b83f9e0

Observation 213d03a7-2377-45c0-8c65-8b667fc1325a · outbound

This paper cites A dataset for crucial object recognition in blind and low-vision individuals’ navi- gation.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model A dataset for crucial object recognition in blind and low-vision individuals’ navi- gation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:18.903088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:18.903088Z digest=sha256:c5d1c1e173865ff5935c5dd5abf5798956b353d0a08ba95fe9fe349b7b6f35a0

Observation 9b83c1b7-51b2-4478-8959-164e322aee14 · outbound

This paper cites Identifying Crucial Objects in Blind and Low-Vision Individuals' Navigation.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Identifying Crucial Objects in Blind and Low-Vision Individuals' Navigation

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-08-10T23:12:19.732215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:18.907994Z digest=sha256:fec1abedd5f16597d56a9f09a4b839e421fb08216deaaa3f45b1929477a5512b

Observation f1059897-ac53-4b8c-9ce5-f0a9b34651e7 · outbound

This paper cites ” i want to figure things out”: Supporting exploration in navigation for peo- ple with visual impairments.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model ” i want to figure things out”: Supporting exploration in navigation for peo- ple with visual impairments

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.794011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:18.913317Z digest=sha256:eb90edb61870ff00e0ee054a6477926ca12486fde43e55b847b8990743dca6ed

Observation 0926ee64-b7d7-4c1d-b1fe-c4cb63c78c17 · outbound

This paper cites Efficient multimodal large language models: A survey.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Efficient multimodal large language models: A survey

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:18.918415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:18.918415Z digest=sha256:9a77a636c9ad2b68c0c57b895f8fa1611b63217bd21d15d60401594060a9feb9

Observation b4aa0c02-e0f1-4e3a-a141-e0cde444e082 · outbound

This paper cites Segment any- thing.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Segment any- thing

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:18.923209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:18.923209Z digest=sha256:eedb397eb777c4eccb9d78a1423efa76bd1440a54d0cd26f1c499c6ded5cab68

Observation de37c228-7a18-40d0-942e-06c069b693a8 · outbound

This paper cites Chatgpt for visually impaired and blind.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Chatgpt for visually impaired and blind

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.767539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:18.928205Z digest=sha256:3ae699ce40e6dc9ff60dde811bd42be7c9b422fa193f0ac7c3603cd70ca0c126

Observation e3f89609-0b6d-4a67-80cc-804876014e47 · outbound

This paper cites Retrieval-augmented generation for knowledge-intensive nlp tasks.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Retrieval-augmented generation for knowledge-intensive nlp tasks

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.750199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:18.933093Z digest=sha256:821498c8e37509371ba81d4297b31f7c2da6014c0ff618b8954d37c5748c0faa

Observation e72f415e-5355-4c8b-b0cf-73fd43ccda11 · outbound

This paper cites Llava-med: Training a large language- and-vision assistant for biomedicine in one day.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Llava-med: Training a large language- and-vision assistant for biomedicine in one day

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.733274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:18.938312Z digest=sha256:dcb0ee5865d2fb8f399f725cf1e6a5ea276283383180a69a885af1c8ad888c9d

Observation f09d8a39-508e-4f99-add2-eaee27cc9d50 · outbound

This paper cites Rouge: A package for automatic evaluation of summaries.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Rouge: A package for automatic evaluation of summaries

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:18.943072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:18.943072Z digest=sha256:914aa757c3288df5d8ff34f69112948c8d91025b1e42d6aa4f983730bb0a673e

Observation 35f932c6-09d4-4a42-82b1-15414b745edb · outbound

This paper cites Visual instruction tuning, 2023.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Visual instruction tuning, 2023

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.706756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:18.948039Z digest=sha256:d0ddf677000cf347e9ddeecdded0dd22368dc53d44c425856995e8eef3cc6b21

Observation 44c019d7-51bd-48d4-ba9d-848a5613a953 · outbound

This paper cites Visual instruction tuning.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Visual instruction tuning

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.691188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:18.953733Z digest=sha256:984fa4a3ab8e3254ad32d4d80891f036b182cf3a64c09b5de721ee06037285be

Observation 525623e0-908c-41aa-98d1-331accf9a211 · outbound

This paper cites Open scene understanding: Grounded situation recognition meets segment anything for helping people with visual im- pairments.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Open scene understanding: Grounded situation recognition meets segment anything for helping people with visual im- pairments

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.674836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:18.958759Z digest=sha256:c1c43d1919aebc2b617341439c89b539735aa325211a1e3f6b963d572ff319c4

Observation 3cf7393c-d555-4072-bcdf-389a75837bcc · outbound

This paper cites A convnet for the 2020s.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model A convnet for the 2020s

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:18.963672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:18.963672Z digest=sha256:ee7eb6884a24f3605ed7ac488f5cae9accb09f3bc75322183424a8a5262b1979

Observation 99d9adcb-aa44-4a1c-b502-9e975a02b15a · outbound

This paper cites DeepSeek-VL: Towards Real-World Vision-Language Understanding.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model DeepSeek-VL: Towards Real-World Vision-Language Understanding

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:18.968985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:18.968985Z digest=sha256:db1931e73c4a07b0825168fd43af539252a2313d3b5cc0199e2b5c749a7848ec

Observation 1be9513b-3f96-4dfe-80ec-af8dee216dde · outbound

This paper cites Generating Contextually-Relevant Navigation Instructions for Blind and Low Vision People.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Generating Contextually-Relevant Navigation Instructions for Blind and Low Vision People

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:18.974180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:18.974180Z digest=sha256:3a4a9be465bd834d2ce223b4b79ed86184197f55860e8317781217661314a524

Observation 62f5f3a1-c6f9-4e3d-8fdc-93ca5a038127 · outbound

This paper cites Using tf-idf to determine word relevance in document queries.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Using tf-idf to determine word relevance in document queries

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.645451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:18.979233Z digest=sha256:80bb67a363dcdb2bfa8c25dd1d9db4fb08a8bd0d0efc18250ad09d155cc6d012

Observation cf099bac-4f3c-4aba-a4f8-a191003c10d5 · outbound

This paper cites Navigation systems for the blind and visually impaired: Past work, challenges, and open problems.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Navigation systems for the blind and visually impaired: Past work, challenges, and open problems

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.627120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:18.984141Z digest=sha256:1510511e32c8e78de9325e63fd2f29a915d03583d62b6d1b97e71f3fdc39b607

Observation ce947ee3-bd5f-47ed-af7a-6d3121b73e56 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:18.989071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:18.989071Z digest=sha256:a475cc6f86d672d42149763cc0c32a548b08989c506684eb5c0ee1e4dfd3cb3c

Observation 71317cfe-cb63-4ba1-a9dd-8a997b2b570d · outbound

This paper cites Math-LLaVA: Bootstrapping Mathematical Reasoning for Multimodal Large Language Models.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Math-LLaVA: Bootstrapping Mathematical Reasoning for Multimodal Large Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:18.994353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:18.994353Z digest=sha256:763bd40d39e50ada7a2825f31964f41970cea8e87fa914de3d1cbda69019056d

Observation 9772810b-3dbd-4730-a911-0711c1664b89 · outbound

This paper cites A dataset for the recognition of obstacles on blind sidewalk.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model A dataset for the recognition of obstacles on blind sidewalk

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.610340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.000847Z digest=sha256:83c492d3a33ab52f251e6e9d614f8d39a7af03fdbdefe10e4db3d22ed707f223

Observation f4c8b7c4-a347-471d-b77e-152bc9f1b93b · outbound

This paper cites Dynamic crosswalk scene understanding for the vi- sually impaired.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Dynamic crosswalk scene understanding for the vi- sually impaired

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.592475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.005636Z digest=sha256:d02c1177291669cabff6fe1c976b21c62d5e76a6ee46ac134f674db7a1fc12a8

Observation 969e7d32-e4a4-4869-a7b5-143c3411c930 · outbound

This paper cites DriveVLM: The Convergence of Autonomous Driving and Large Vision-Language Models.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model DriveVLM: The Convergence of Autonomous Driving and Large Vision-Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:19.010537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:19.010537Z digest=sha256:5dc4f2b10eb7ac009488b12071194e8e87445f053bfad27660e9b5d55a56c95a

Observation 9731851f-65c8-4063-a455-ef162f7bef95 · outbound

This paper cites VisionGPT: LLM-Assisted Real-Time Anomaly Detection for Safe Visual Navigation.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model VisionGPT: LLM-Assisted Real-Time Anomaly Detection for Safe Visual Navigation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:19.015582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:19.015582Z digest=sha256:85e48661f85f3b9c8186370389da60047d2b7e0ba65d6332c744350e6871f0f9

Observation bf1c821a-2402-4dc7-81b4-66bfa34878b7 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:19.020845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:19.020845Z digest=sha256:967f9c8be0b9ea68cca81bc267a7b3c606ae3fe85701a302f129ad87e203a162

Observation 25380c0b-0bf2-4c30-a9df-3ce8ef6ac537 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large lan- guage models.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Chain-of-thought prompting elicits reasoning in large lan- guage models

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.575753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.026210Z digest=sha256:3a6014062c0d601bbd08d335498503a6f8b6e1636002e5319448001476b3aad6

Observation 874f4fd7-562b-44d9-9b76-bf5cad97e02c · outbound

This paper cites A dataset for the visually impaired walk on the road.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model A dataset for the visually impaired walk on the road

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.558120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.031353Z digest=sha256:b894a454f05848c03c1591d9fd0e3b029946ef1b95bbb4edae9279fd6cb35fae

Observation 45d1efec-567f-427b-a136-c21455aae58d · outbound

This paper cites Emerging Practices for Large Multimodal Model (LMM) Assistance for People with Visual Impairments: Implications for Design.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Emerging Practices for Large Multimodal Model (LMM) Assistance for People with Visual Impairments: Implications for Design

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:19.036552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:19.036552Z digest=sha256:3ab42f4549d70f8393ce927618644b43509da315c0713f9a9efa7e0edb707579

Observation 16a92f8c-6dca-45b9-959a-7016c6f9c437 · outbound

This paper cites A survey of efficient fine- tuning methods for vision-language models—prompt and adapter.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model A survey of efficient fine- tuning methods for vision-language models—prompt and adapter

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.542181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.041730Z digest=sha256:5d8da364882fa3cd614524569f66638ee120bc5c75050e8cb688c76f5bb7544d

Observation 3127d5e9-485d-453c-a7bd-b11d37bd10f8 · outbound

This paper cites Qwen2 Technical Report.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Qwen2 Technical Report

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:19.046729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:19.046729Z digest=sha256:3a4686a9b2a6cd16186d03f9c50987d2e63219768576e82d81fa370c8bb58719

Observation 3f7b3f7b-10d2-4eea-ae66-1f8cb4f9b6b5 · outbound

This paper cites VIAssist: Adapting Multi-modal Large Language Models for Users with Visual Impairments.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model VIAssist: Adapting Multi-modal Large Language Models for Users with Visual Impairments

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:19.051938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:19.051938Z digest=sha256:6c45b6ee7372fd76765f956889aef9543fc389f0eca503ee2aa28ef48489d93c

Observation 619fc4b8-0eb3-4b5b-91a1-03417a34a3df · outbound

This paper cites Seeway: Vision- language assistive navigation for the visually impaired.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Seeway: Vision- language assistive navigation for the visually impaired

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.525355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.057234Z digest=sha256:e7a7884e063a26e69f1e73c16b44964339a3dabbff46ad05c87b752d8799e99c

Observation 68a87777-4a55-48e2-9ca2-812259c9f4dc · outbound

This paper cites MiniCPM-V: A GPT-4V Level MLLM on Your Phone.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model MiniCPM-V: A GPT-4V Level MLLM on Your Phone

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:19.062819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:19.062819Z digest=sha256:bd5a7c6ce079d4e53df00fd3d6157593d8d85f15ba27eb7e64fbccb17dc33b21

Observation 87fb6fc1-3a14-40d5-88bb-2e12276bf99d · outbound

This paper cites Yi: Open Foundation Models by 01.AI.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Yi: Open Foundation Models by 01.AI

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:19.068765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:19.068765Z digest=sha256:4efb9f1df0dbf49153a93900a40de3471ecf54c09d086a17498cb92faecae0e2

Observation e76e4dc2-ef84-4155-86ac-8c64ee39b27b · outbound

This paper cites Vision-language models for vision tasks: A survey.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Vision-language models for vision tasks: A survey

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:19.073764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:19.073764Z digest=sha256:56ffb4ac9b2e8f94dc22c150d0a7ee949c596b0b8160f0498e5e4c28ef6b32eb

Observation e60cf562-c603-4dd6-a3a9-d35626a63be9 · outbound

This paper cites Grfb-unet: A new multi-scale attention network with group receptive field block for tactile paving segmentation.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Grfb-unet: A new multi-scale attention network with group receptive field block for tactile paving segmentation

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.498924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.078879Z digest=sha256:6b0cc6f33bb234ab33c0b9cfb3a3fbf8220e76b6b41b292fe59c1a18a9542e19

Observation 7e87c874-6c10-4bef-94ff-43e2ef0e531f · outbound

This paper cites SparseVLM: Visual Token Sparsification for Efficient Vision-Language Model Inference.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model SparseVLM: Visual Token Sparsification for Efficient Vision-Language Model Inference

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:19.083793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:19.083793Z digest=sha256:d53775f7db22bf016a44defebac82d740f1821ead941a85ce8a196631d6baf58

Observation 5a7d3576-7308-4938-9d88-c5d7f3db478d · outbound

This paper cites VIALM: A Survey and Benchmark of Visually Impaired Assistance with Large Models.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model VIALM: A Survey and Benchmark of Visually Impaired Assistance with Large Models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:19.089009Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:19.089009Z digest=sha256:7193b9bc4d2c8f431f2f4d5a7b0b6a0bf3b46cae5b060916d81f9c61f78a1480

Observation afb450c2-2367-4618-873a-f6a6be4895de · outbound

This paper cites Judging llm-as-a-judge with mt-bench and chatbot arena.Advances in Neural Information Processing Systems, 36:46595–46623, 2023.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Judging llm-as-a-judge with mt-bench and chatbot arena.Advances in Neural Information Processing Systems, 36:46595–46623, 2023

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:19.094171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:19.094171Z digest=sha256:5f6622fd90d8a8c0c908ebe0d056ab3d1affb4a22ca6d1fd0f8dcde0c597b205

Observation b57bb9e5-23df-4485-9a44-c3fb7e64b069 · outbound

This paper cites Detecting twenty-thousand classes using image-level supervision.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Detecting twenty-thousand classes using image-level supervision

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.472133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.099057Z digest=sha256:6dddd51f3d95cc1232dfa2896186765f32f9df92a084684c20240c8cf2454954

Observation 066c069c-f4d9-4b05-9906-c9891d41afa8 · outbound

This paper cites LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-10T23:12:19.104157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:12:19.104157Z digest=sha256:196e48dd5af980e92f66aa259c55f8db193cd43b1720efce6ea60d62dec79cbc

Observation 966a674f-d632-4b28-839d-0d770faececa · outbound

This paper cites Data Collection.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Data Collection

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.456047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.109481Z digest=sha256:ac6302047b62b887e5bcc3e67c2effbb4d7dc8c1f305c17d07871cbd98fefd84

Observation 3b4dce1b-f7c7-4325-be3c-9a034494d533 · outbound

This paper cites Problem Formulation.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Problem Formulation

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.440398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.114364Z digest=sha256:357ead9c12c8da0451c8fb4a82ffa135c976caac46e60849c0c15051f6c13a39

Observation 208f177f-c8cf-4588-ad7f-d43ff41a5f15 · outbound

This paper cites Settings.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Settings

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.424697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.119440Z digest=sha256:3a06351fdde04717c15c6342cd327f82ae1f564ce93dc3f207555871ed775f8e

Observation ee4558f7-83a7-4b78-8b07-ab37183f39c9 · outbound

This paper cites Data Regional Distribution.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Data Regional Distribution

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.408918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.124301Z digest=sha256:9d28392903abed2c5c3d60ca60da4269b3814e6abf9135d11311b7ee1644debb

Observation 323e987c-e72f-4fb1-9a4d-d7b1db7a0116 · outbound

This paper cites All Prompts Used in Paper.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model All Prompts Used in Paper

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.393059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.129337Z digest=sha256:75eff353fc8ad08831743a96dc9f292deeffe4afc3aa4c2cdcc4d2a7d3737d62

Observation 2eacb0a7-3f6d-40dd-85dc-8a378aab14c9 · outbound

This paper cites Visualization of Hierarchical Reasoning.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Visualization of Hierarchical Reasoning

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.377270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.134086Z digest=sha256:38b24d137576e0611625e804372886166ae65fa6d91f8d68fe1fa2c53e76530c

Observation 172be62c-3674-413c-b5cb-dac19ceb2675 · outbound

This paper cites an unresolved cited work.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Unresolved cited work

Reference 62

Resolution
unresolved
raw_fallback, observed 2026-08-10T23:12:20.361647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.138813Z digest=sha256:daf827e61a182a23eac027754302c4984d1588e59f9b335e6a9b5716d49a8d2d

Observation bb27a21c-ccb4-4c38-94de-6bf77db81f66 · outbound

This paper cites Data Regional Distribution Table 8 shows the data distribution and corresponding du- ration in the W AD dataset.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Data Regional Distribution Table 8 shows the data distribution and corresponding du- ration in the W AD dataset

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.345168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.143806Z digest=sha256:a2dea4f63ca31517894b5dcf0c22bff449e35b0d2515c485363e53e7d8aec4b4

Observation 20ba17d5-aa53-435c-8643-3be40a57bafb · outbound

This paper cites an unresolved cited work.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Unresolved cited work

Reference 64

Resolution
unresolved
raw_fallback, observed 2026-08-10T23:12:20.329328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.149242Z digest=sha256:945b81936cee06a09e3ed97eb8f8c8870a4b651c6925d67d492bceb70d68edaf

Observation fa53ddef-f093-4a0a-b1a5-9926a7cde868 · outbound

This paper cites Visualization of Hierarchical Reasoning We have demonstrated the results of hierarchical reasoning using WalkVLM in Figure 18.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Visualization of Hierarchical Reasoning We have demonstrated the results of hierarchical reasoning using WalkVLM in Figure 18

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.314321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.154319Z digest=sha256:39b0df99f62b778c7fdb25c9e55bba2dc9f6b3f643cc8994abfec2648a26c6b7

Observation a06e5a22-9cee-4f47-a61a-b024effe1e81 · outbound

This paper cites The visualization results of the two models on the video stream can be viewed here, where WalkVLM is capable of generat- ing less temporal redundancy.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model The visualization results of the two models on the video stream can be viewed here, where WalkVLM is capable of generat- ing less temporal redundancy

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.297847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.159537Z digest=sha256:8142b61205d9a1d88795ecb62c44057fb99d4fb47dcbcca754c3809a1a2e19fa

Observation 37e46d0e-1bad-4d67-8f60-99c84c59ef01 · outbound

This paper cites an unresolved cited work.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Unresolved cited work

Reference 67

Resolution
unresolved
raw_fallback, observed 2026-08-10T23:12:20.282670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.164646Z digest=sha256:ccbea2871974f335c6d7b8d2d59e63a96638e34b82e839098f1fae27bd33262c

Observation af84912f-5a7a-433f-afe7-cad67856ffe1 · outbound

This paper cites 1. Location.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model 1. Location

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.267241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.169791Z digest=sha256:1075b732be69a35540d65e24563d425e6be96498ba928d83efeebc79ee10d32f

Observation 9860899e-e4ab-4aaf-8f45-213fcca8c135 · outbound

This paper cites an unresolved cited work.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Unresolved cited work

Reference 69

Resolution
unresolved
raw_fallback, observed 2026-08-10T23:12:20.250862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.175365Z digest=sha256:5a59e881834525ba561ff0ec482df203c8aee224a0ada4e45ce65aa3d0bd2a92

Observation 3ce6d488-b984-430c-9e2a-0b319af98de4 · outbound

This paper cites This entrance is marked by a structure with glass doors and an overhanging roof.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model This entrance is marked by a structure with glass doors and an overhanging roof

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.236242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.180203Z digest=sha256:38654017524d9d3ce93bd1fd617f7789d2ca4d8a82ebf086539a8439e6b4d287

Observation e8f45a7d-568f-4b97-9012-6880c5e70639 · outbound

This paper cites The pathway is paved and flanked by plantsers filled with lush green plantsings.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model The pathway is paved and flanked by plantsers filled with lush green plantsings

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.221192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.185566Z digest=sha256:8261f539a7c346cfc9e3767cb209eec6a3b12700f17c11c7f437f8299e73d8e6

Observation 79f39e30-9d6a-4402-a2eb-0c809a038457 · outbound

This paper cites Walk under this structure to shield yourself from potential sun exposure or weather elements.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Walk under this structure to shield yourself from potential sun exposure or weather elements

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.203675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.190374Z digest=sha256:6241810c646fb13aaa75485903d13823a7a2ca9cccc86b5aa626e1e9b5996e3b

Observation 283ea936-fbee-4e09-8d80-a9c6edec8857 · outbound

This paper cites an unresolved cited work.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Unresolved cited work

Reference 73

Resolution
unresolved
raw_fallback, observed 2026-08-10T23:12:20.187375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.197218Z digest=sha256:683e02acec67c8163af550550f89441250db4e489ac3cb08a03ea97afec4f5c5

Observation b44c0be8-ded3-4f12-9e74-5f24fcaf0fc6 · outbound

This paper cites The path is clear and unobstructed, with a slight curve to the left.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model The path is clear and unobstructed, with a slight curve to the left

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.171549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.203763Z digest=sha256:4919a78a3e913f88acd3a2b122a98c751d8b133cb10507214afb8e68ca59ed62

Observation 8bb3ad1d-60b2-4ad8-8fd1-f60c292bcf45 · outbound

This paper cites On your right, you will see a staircase and more greenery.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model On your right, you will see a staircase and more greenery

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.151684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.209235Z digest=sha256:7c8c14657659fe7f6ee605d4d216e9e7dc8063e179fbfae74b5125bcf69b46d5

Observation 7ed549b4-8e17-4cb5-8968-1a8d06ad997c · outbound

This paper cites Toastbox.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Toastbox

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.135501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.214742Z digest=sha256:8f1d5a88be7314c5c5e211639358c89ebf866772656e00807450251e773bc767

Observation 783f67ff-4669-4a45-bcf3-171f5ba4d485 · outbound

This paper cites an unresolved cited work.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Unresolved cited work

Reference 77

Resolution
unresolved
raw_fallback, observed 2026-08-10T23:12:20.119015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.219744Z digest=sha256:1ca5a2028e19b4932d587ebfddf67d9d0b9115c4b3ccb9e392db6bed7e187c50

Observation 4b1488c1-2a4e-4e83-9e3e-29deea4b7314 · outbound

This paper cites an unresolved cited work.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Unresolved cited work

Reference 78

Resolution
unresolved
raw_fallback, observed 2026-08-10T23:12:20.103333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.224589Z digest=sha256:fff041be863976487ecb4c24e7070204d4fa967de79f926d105e494d9afd25bc

Observation dadcf5f0-fde6-46ad-b493-2a554acda2e1 · outbound

This paper cites Whole Earth.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Whole Earth

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.087877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.229541Z digest=sha256:eb9ab53fd36ca25cb6d92d8494d1d88c44e2509dac23e3f193976eaaee98d501

Observation 6ce7f705-b099-4f8d-b1ac-eb9a9b4bbd76 · outbound

This paper cites an unresolved cited work.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Unresolved cited work

Reference 80

Resolution
unresolved
raw_fallback, observed 2026-08-10T23:12:20.072030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.234461Z digest=sha256:2607939727a6bda72baf86b4fa276270e3602fef7fcc03e5a86394506b77c57d

Observation 4d16c5ed-075e-486e-9292-9029ae59ca52 · outbound

This paper cites an unresolved cited work.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Unresolved cited work

Reference 81

Resolution
unresolved
raw_fallback, observed 2026-08-10T23:12:20.056663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.239320Z digest=sha256:354f270073448d0e8eab4f6f5d6541590022f0395e7770a32a16a813fefa9720

Observation ca414976-e03c-48f5-8ff6-4bb5db3de7b2 · outbound

This paper cites an unresolved cited work.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Unresolved cited work

Reference 82

Resolution
unresolved
raw_fallback, observed 2026-08-10T23:12:20.038850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.244331Z digest=sha256:d17db85f15d396d3b227e04547ba6dc28c04134e12c83789a92ab33f5ebe475d

Observation 9ec88863-9efe-4e47-bef3-af0badb6e14f · outbound

This paper cites By following these steps, you will be guided through the urban environment shown in the video, ensuring you stay within pedestrian pathways and avoid traffic lanes.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model By following these steps, you will be guided through the urban environment shown in the video, ensuring you stay within pedestrian pathways and avoid traffic lanes

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.020292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.249307Z digest=sha256:d00ad8df4b503fbb7bc95dbfa1da8bc7f6624309f6cdaeea250fb2173ffacc64

Observation 66c3703b-ff16-47fe-81d6-402951b2fcf0 · outbound

This paper cites Whole Earth.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Whole Earth

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:20.002662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.254009Z digest=sha256:53fb1c616bca9d9f2bb14133578e5e7deb4a8ff96144371a4e9f8ce8e7d76156

Observation 41adabdc-b30e-4d69-a260-18111e696262 · outbound

This paper cites Do not step onto the road.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Do not step onto the road

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:19.985365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.258853Z digest=sha256:e4d6fcb28a1923a243ebac3b8677a9e35f874c63c124f30d62421ccaa9dc2bca

Observation b6ee6707-7d2e-499c-bc81-7aef21a4e80e · outbound

This paper cites Ensure that you check for oncoming traffic from both directions.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Ensure that you check for oncoming traffic from both directions

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:19.969653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.264062Z digest=sha256:298f2c5a9b4e5d5f1186d9cf044dcb9f6c363e0ea6defe71ab053efed2f7b9d1

Observation b5c07023-2a10-4657-a6f6-c5acdff74cb7 · outbound

This paper cites Whole Earth.

WalkVLM:Aid Visually Impaired People Walking by Vision Language Model Whole Earth

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:12:19.953323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T23:12:19.269126Z digest=sha256:d46e1c4f5c4e8d26e8e5147aa4e0a313f66b192910fff392ac981f86d57e550f

Pith citing papers

Observation 1b62c9c2-7faa-47de-b423-b5ce211a1ec8 · inbound

Hallucination at a Glance: Controlled Visual Edits and Fine-Grained Multimodal Learning cites this paper.

Hallucination at a Glance: Controlled Visual Edits and Fine-Grained Multimodal Learning WalkVLM:Aid Visually Impaired People Walking by Vision Language Model

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T05:43:47.468600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:43:47.468600Z digest=sha256:9a634c55ac1b22cba7bc5550943a5887a2d0750c34cb37175c698fa5f8b17c50

Observation 87118920-cd2e-4364-bbde-4b910d574726 · inbound

A Synthetic-to-Real Dehazing Method based on Domain Unification cites this paper.

A Synthetic-to-Real Dehazing Method based on Domain Unification WalkVLM:Aid Visually Impaired People Walking by Vision Language Model

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-08-05T10:27:19.605209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T10:27:19.315551Z digest=sha256:2829987aa90e19c1adac28aa2505b28cbb3dc3a289b29aa47d94a9552028ceb8

Observation 275778dc-f220-43d0-9bd5-8ae77c0d927f · inbound

VIABench: A Comprehensive Video Benchmark Collected from Blind Individuals for Visual Impairment Assistance cites this paper.

VIABench: A Comprehensive Video Benchmark Collected from Blind Individuals for Visual Impairment Assistance WalkVLM:Aid Visually Impaired People Walking by Vision Language Model

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-02T01:31:33.936559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:31:33.936559Z digest=sha256:79db9cf30beb2573f7c5ef9ef5a10d121b136233d6c3e7a40ad10ea67bfe4c45