Pith. sign in

Paper Citation Record · LEDGER

A Large Vision-Language Model based Environment Perception System for Visually Impaired People

As of 18 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 0 inbound Pith citation observations for arXiv:2504.18027.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.18027 v1

Coverage vector

measured 57 of 57 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:31:00.288054Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

57 of 57 outbound references displayed

  • verified exact0
  • verified fuzzy40
  • unresolved17
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5238b3ec-0889-4319-9ec4-254728d96961 · outbound

This paper cites Ingold, The perception of the environment: Essays in livelihood, dwelling and skill.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Ingold, The perception of the environment: Essays in livelihood, dwelling and skill

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:01.025469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.070592Z digest=sha256:66661b0ea04f40d0693d5e3254290dbc1ab59a4bc3d4e29d2fb349b8f77795f6

Observation ca42d922-1ed1-460b-bb85-b2b5985e29fc · outbound

This paper cites Enabling independent navigation for visually impaired people through a wearable vision-based feedback system,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Enabling independent navigation for visually impaired people through a wearable vision-based feedback system,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:01.012861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.075930Z digest=sha256:fdfad289a99253baf5ad44196386ad0bb2b58db74da5d372e5f0d1dd5261722b

Observation 81703b36-5f97-486c-8024-ba29eec7ec0c · outbound

This paper cites Magnitude, temporal trends, and projections of the global prevalence of blindness and distance and near vision impairment: a systematic review and meta-analysis,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Magnitude, temporal trends, and projections of the global prevalence of blindness and distance and near vision impairment: a systematic review and meta-analysis,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:01.000856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.080180Z digest=sha256:07aac25f2c9be9709c9bdb09560d4b42ccd48b9d77170a4fcc861a7b80d3cc29

Observation 84177a55-74ac-4142-a187-32e566d797ca · outbound

This paper cites Objectively measured visual impairment and dementia prevalence in older adults in the us,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Objectively measured visual impairment and dementia prevalence in older adults in the us,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.989233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.083756Z digest=sha256:eaec0f472088d3ffe74fc01846d4733058bf5660b41de9746a9cc6ac5bfd6b3d

Observation 4239bd43-dd8e-4050-8236-e43af278c429 · outbound

This paper cites Virtual-blind-road following-based wearable navigation device for blind people,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Virtual-blind-road following-based wearable navigation device for blind people,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.976446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.087535Z digest=sha256:d882e3b794399886ff6bc12593fe66e62e6d890adc7e7cbd10f348d050dd359e

Observation 53098e2a-d256-4c56-90fd-bc8e9a837140 · outbound

This paper cites Embedded reading device for blind people: A user-centered design,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Embedded reading device for blind people: A user-centered design,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.962539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.091100Z digest=sha256:de80e2be80ebdbe24cc0cd77a35c550ec911639a843987f1f616ac41f9643803

Observation 0e357f4a-6a50-41a8-ab36-4d3cf8eb07d8 · outbound

This paper cites Hand-priming in object localization for assistive egocentric vision,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Hand-priming in object localization for assistive egocentric vision,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.949264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.095164Z digest=sha256:6ffce4fcdc43e90cca004adfe637b5fa243cbef6fb028d165ff4bec0cfdd0cb7

Observation 7586210f-3cae-425b-a8ab-3aad3216f4e5 · outbound

This paper cites Vizwiz-fewshot: Locating objects in images taken by people with visual impairments,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Vizwiz-fewshot: Locating objects in images taken by people with visual impairments,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.936168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.098659Z digest=sha256:cb8785698917093b5723282a80d3250ae518f0638ea20af7d24de9a085c4719c

Observation b8897f6e-bdb0-4ec7-922a-ecd80845d349 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:00.102041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:00.102041Z digest=sha256:894533b299fe038c6ca0a7c88e4afade48d4c5f5041e2bffacca424ddd995472

Observation 8362b807-6bb6-49bb-a64b-4545ea3c38fa · outbound

This paper cites Incorporating External Knowledge into Machine Reading for Generative Question Answering.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Incorporating External Knowledge into Machine Reading for Generative Question Answering

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:00.105795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:00.105795Z digest=sha256:75bb8083f70f61a7c622ad5f0ae3e484de151d364ce7a772af08285203b8ea83

Observation 7793f1ca-80db-4d2b-bdc9-aacadb2ff268 · outbound

This paper cites Check Your Facts and Try Again: Improving Large Language Models with External Knowledge and Automated Feedback.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Check Your Facts and Try Again: Improving Large Language Models with External Knowledge and Automated Feedback

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:00.109472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:00.109472Z digest=sha256:7b577318ac13efe95602abcdae0012644323535ffb9292f42040c1b985b3ccd5

Observation 6c32a6fa-7e82-4595-9138-59c04d5d0d6e · outbound

This paper cites Smart guiding glasses for visually impaired people in indoor environment,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Smart guiding glasses for visually impaired people in indoor environment,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.922921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.113421Z digest=sha256:fd8c4510ed30d2aba63bebdd13ccfa22e73d20f49b9eeeedaaac36e47bfa5deb

Observation 476705bc-afe4-4c72-85e4-c95e66a421d9 · outbound

This paper cites An enhanced obstacle avoidance method for the visually impaired using deformable grid,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People An enhanced obstacle avoidance method for the visually impaired using deformable grid,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.910049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.117676Z digest=sha256:35d8c0a8c5340f48adebfcbd6002869330e533640654d8d99c05a5c0b63b478a

Observation 4131347c-bc28-441d-8b39-f54f3886c515 · outbound

This paper cites Collision detection method using image segmentation for the visually impaired,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Collision detection method using image segmentation for the visually impaired,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.897088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.121663Z digest=sha256:500326c5e95ee51e23d49ada79f8d7cbbf954ae3ad5829c75283820cf83e6900

Observation 2f7d754c-8d43-45c7-aac0-9c60342808ac · outbound

This paper cites Tdraw: A computer-based tactile drawing tool for blind people,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Tdraw: A computer-based tactile drawing tool for blind people,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.884312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.125470Z digest=sha256:a803bf0065c4bd4cc576edf16079e906fad572b37887b571751d089a00ce113e

Observation 6c449d9b-0b3f-4423-bd65-fe56bd35dabe · outbound

This paper cites Usable gestures for blind people: Understanding preference and performance,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Usable gestures for blind people: Understanding preference and performance,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.872328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.129243Z digest=sha256:85dfcdff3b72b842a7d3a548e8d3a8bac7ffa2a8dd520f27c718110d57436adf

Observation a716678e-67de-4f5c-84de-4918c397b659 · outbound

This paper cites How teens with visual impairments take, edit, and share photos on social media,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People How teens with visual impairments take, edit, and share photos on social media,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.860378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.133075Z digest=sha256:fb7063bf8993e29704948c5c2d0af1c783ab66fcc9d2f7d1b904b1bf29b270ca

Observation 9f894f49-d49e-4c92-9fbd-8845beb8ba1e · outbound

This paper cites Flight: A low-cost reading and writing system for economically less-privileged visually-impaired people exploiting ink-based braille system,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Flight: A low-cost reading and writing system for economically less-privileged visually-impaired people exploiting ink-based braille system,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.847723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.136838Z digest=sha256:8234928ee3966b9dd2ad43b605e06ca87c7ccc6d7ac317acc9a602ebc9ff0018

Observation 5f6859ed-659e-41b2-af1e-c7fe1ff43ee5 · outbound

This paper cites Linespace: A sensemaking platform for the blind,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Linespace: A sensemaking platform for the blind,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.834865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.140734Z digest=sha256:82f352550100f794f2879152ef1044bf5983d0c45dde1b5bab31ecc08ef9383b

Observation 8ab3b6b7-2fae-428c-94a4-91c5f5189655 · outbound

This paper cites Tangible reels: Construction and exploration of tangible maps by visually impaired users,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Tangible reels: Construction and exploration of tangible maps by visually impaired users,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.821638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.144554Z digest=sha256:7c0495a006a92e4633656a21ec55783a7f812ed931ba2e5d2d69cfb58c1d8793

Observation 2745375e-50b7-476d-bfe5-2047884835fc · outbound

This paper cites Comparing computer- based drawing methods for blind people with real-time tactile feed- back,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Comparing computer- based drawing methods for blind people with real-time tactile feed- back,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.808129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.148569Z digest=sha256:5eaa0921b1937db4d79adfba4d8a3445324608b1ef0bc17331292a1f212cf888

Observation ee460e25-8e79-460d-9ac2-3ffa77a00f2f · outbound

This paper cites Accessible maps for the blind: comparing 3d printed models with tactile graphics,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Accessible maps for the blind: comparing 3d printed models with tactile graphics,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.795103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.152475Z digest=sha256:ce9645b1df5f3c99132ea0fb4f425b35c8af2bcc25d220541f299e04036ed77f

Observation 4f9b320c-d7ba-45bf-b420-4a433b9245b0 · outbound

This paper cites Taking into account sensory knowledge: The case of geo-techologies for children with visual impairments,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Taking into account sensory knowledge: The case of geo-techologies for children with visual impairments,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.782295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.156325Z digest=sha256:0e53252dffe9c8de14c04224718ed4e488d9c2870cf30fd292f2b1afe119e370

Observation 2c4c0c9f-7df3-4901-bf1b-846fbdbee5cc · outbound

This paper cites People with visual impairment training personal object recognizers: feasibility and challenges,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People People with visual impairment training personal object recognizers: feasibility and challenges,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.769894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.160306Z digest=sha256:6ef06a583ee722589227d2e70aedec203865d12a93fa8705b902678b3d8ad666

Observation 2d119cea-d298-4215-973f-999cd590083f · outbound

This paper cites Synthesizing stroke gestures across user populations: A case for users with visual im- pairments,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Synthesizing stroke gestures across user populations: A case for users with visual im- pairments,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.756993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.164179Z digest=sha256:978eec67662cc69c3e97e64d890432b93dbcc7f9454c98322be19a9045e15294

Observation 5a758cf3-9155-4672-9e47-5c8ec60709dd · outbound

This paper cites A face recognition application for people with visual impairments: Understanding use beyond the lab,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People A face recognition application for people with visual impairments: Understanding use beyond the lab,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.744620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.167995Z digest=sha256:15212bcda8ad2d764120b905b9c71cb44f6b2ece59eaa1d22aa0323ae9a895f7

Observation 96ec5311-e479-40ed-ba47-8d60998fa60b · outbound

This paper cites A multitask grocery assist system for the visually impaired: Smart glasses, gloves, and shopping carts provide auditory and tactile feedback,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People A multitask grocery assist system for the visually impaired: Smart glasses, gloves, and shopping carts provide auditory and tactile feedback,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.731875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.171921Z digest=sha256:f6064ae593aa95c8b1b011153d0bd5103862daabdc566f7ed3fa0ab216501eaa

Observation 6b9f7371-ff12-4baa-af1e-4f1a343d138a · outbound

This paper cites Hindsight: Enhancing spa- tial awareness by sonifying detected objects in real-time 360-degree video,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Hindsight: Enhancing spa- tial awareness by sonifying detected objects in real-time 360-degree video,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.719664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.175848Z digest=sha256:963a3521ff5216cbe4332041ee729ca37ef98e649a1358e6cb97e18fe355cc86

Observation 66c24848-801b-43a9-bba2-e0ddc698e84d · outbound

This paper cites You only look once: Unified, real-time object detection,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People You only look once: Unified, real-time object detection,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.706770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.180050Z digest=sha256:05e97d2b0a36036f8b81cba0c349f94fd760858a5535809249edab13ccb47029

Observation 782940ce-8c74-4448-bfb2-673331de9336 · outbound

This paper cites Ssd: Single shot multibox detector,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Ssd: Single shot multibox detector,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.692788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.183925Z digest=sha256:1edbbe565ac0c7adabd8c0c6819c5a125d2def46525c8d76879265834a2ad28b

Observation 2ee8d73a-286f-4d23-b07d-2e3cb947d673 · outbound

This paper cites Rich feature hierarchies for accurate object detection and semantic segmentation,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Rich feature hierarchies for accurate object detection and semantic segmentation,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.679459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.187686Z digest=sha256:b8e31c87fb83b6ec65e8cda1a0321c9e724fe5a7c5c967586c1925714c71c8da

Observation db88033f-d8b6-4980-86b5-a4533595940e · outbound

This paper cites Faster r-cnn: Towards real- time object detection with region proposal networks,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Faster r-cnn: Towards real- time object detection with region proposal networks,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.665125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.191599Z digest=sha256:b17a57800f9dfca579f02d96c681ce778c17b8f81e45198a55070ee81196c9ec

Observation a194552c-df25-496f-8c23-442091595cf6 · outbound

This paper cites Fully convolutional networks for semantic segmentation,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Fully convolutional networks for semantic segmentation,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.652065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.195421Z digest=sha256:24bcd46b7e5553796e57d890b0e4b10247aa6fbf02ee4fcdca8546de576c13d7

Observation cd9f58ff-1615-4cc4-a81e-1e7964239d26 · outbound

This paper cites Segnet: A deep convolutional encoder-decoder architecture for scene segmentation,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Segnet: A deep convolutional encoder-decoder architecture for scene segmentation,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.640256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.199320Z digest=sha256:767d157d29cd5590c4af50a8bd344e1a7a9d74b8ba712a3694df79c46f56a6dc

Observation a944ebca-3c19-4ccc-9772-cbc211c6bffc · outbound

This paper cites Deeplab: Semantic image segmentation with deep convolutional nets, atrous convolution, and fully connected crfs,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Deeplab: Semantic image segmentation with deep convolutional nets, atrous convolution, and fully connected crfs,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.628702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.203093Z digest=sha256:4646d5b1910c5bd11ee7e3752385d7b4b0f6bcc1e6d34e318694609a0e8d9251

Observation 9150d598-d671-489d-9a03-c988fa2185f2 · outbound

This paper cites Rethinking semantic segmentation from a sequence-to-sequence perspective with transformers,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Rethinking semantic segmentation from a sequence-to-sequence perspective with transformers,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.616863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.206866Z digest=sha256:96c8d5abb04ecd6b66882d62c3c9af89431a91d3c7fcdbf76984168a89a33650

Observation f05c32a9-1f5b-4acf-8032-411eef2c3f9c · outbound

This paper cites Swin transformer: Hierarchical vision transformer using shifted windows,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Swin transformer: Hierarchical vision transformer using shifted windows,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:00.210929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:00.210929Z digest=sha256:0aa072a76d13c748e24826bbcc6686a2120050ff3a67aefe4c2a732faf58803a

Observation f399b1a3-0bdb-41c3-b49f-62bc9b3321ba · outbound

This paper cites Attention is all you need,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Attention is all you need,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:00.214911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:00.214911Z digest=sha256:1aab207d8b34b1d349085373fef4663238a411813a816c2a32051ee777f7b0c9

Observation dc472b25-ca38-4440-bc89-769d0e8f45b2 · outbound

This paper cites Fusenet: In- corporating depth into semantic segmentation via fusion-based cnn architecture,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Fusenet: In- corporating depth into semantic segmentation via fusion-based cnn architecture,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.588242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.218978Z digest=sha256:b430d076b66704a9b0bc7a01491b5111740e9a0d36baf4ec6563436804e40339

Observation ff2e631d-caab-4573-9575-50ff1d525f6c · outbound

This paper cites RedNet: Residual Encoder-Decoder Network for indoor RGB-D Semantic Segmentation.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People RedNet: Residual Encoder-Decoder Network for indoor RGB-D Semantic Segmentation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:00.223131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:00.223131Z digest=sha256:946527802615d199ae8d42f757cdbd37d423e2eff93dc28ab4805dcc97c88266

Observation 740e6dea-7417-46aa-b70a-34ceda40eef0 · outbound

This paper cites Show and tell: A neural image caption generator,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Show and tell: A neural image caption generator,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.575223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.227421Z digest=sha256:c4e7bfe3d501ab50afae7e3f372b5d5111e7ac3e7c8ebdcaddfdc8c08b2a2758

Observation c8b8e99b-6775-45f8-9922-73d45b06afb5 · outbound

This paper cites Knowing when to look: Adaptive attention via a visual sentinel for image captioning,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Knowing when to look: Adaptive attention via a visual sentinel for image captioning,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.561847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.231212Z digest=sha256:7874986fd1d90a942b0f21ef1300aaac0501d0c1698e64cca4d2ebf69785debb

Observation b953f83a-e732-4730-9bd1-1cee6bfa1a0e · outbound

This paper cites Show, attend and tell: Neural image caption generation with visual attention,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Show, attend and tell: Neural image caption generation with visual attention,

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:00.235162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:00.235162Z digest=sha256:c8ae33d5ecb821a846ee54728278e0cb53e21f004def0e30e82b819039d67165

Observation cc6bd0fe-de3c-42c2-8f77-dbe17d3ac78a · outbound

This paper cites Dense captioning with joint inference and visual context,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Dense captioning with joint inference and visual context,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.540912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.239060Z digest=sha256:226dbf3922f50b073fd298dfbf251f0dfe5da1d1e3c108df2d03f8ef7c36e666

Observation 2964a558-64f0-4877-926c-80e27e40cccd · outbound

This paper cites Unifying vision-and-language tasks via text generation,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Unifying vision-and-language tasks via text generation,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:00.243006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:00.243006Z digest=sha256:bdb97ba6b93e31289d5b68f37290959a80d2b98079b4911fe2fceac54ab2702c

Observation a970cb18-308b-40d8-8047-b79c31476024 · outbound

This paper cites Multimodal few-shot learning with frozen language models,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Multimodal few-shot learning with frozen language models,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.518885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.246908Z digest=sha256:9cea2d97903d962a927c810abfb96861b5f02413f6a7d6c10ceb4c76cd731e65

Observation defa57c1-c932-44b8-b1d4-1991733ba014 · outbound

This paper cites SimVLM: Simple Visual Language Model Pretraining with Weak Supervision.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People SimVLM: Simple Visual Language Model Pretraining with Weak Supervision

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:00.250761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:00.250761Z digest=sha256:c7383d6ce69b38b06fffffcc683f1f183b7bba05814f3c7a56f47892426e327c

Observation 3618de6a-eedf-4b2d-b03a-ce85522cbad0 · outbound

This paper cites Visual instruction tuning,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Visual instruction tuning,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:00.254925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:00.254925Z digest=sha256:1fc50c6de06e156f7b81501ba2bb06160c0a14b918a00898d750f7dcfdec86c3

Observation 73657ceb-f69e-4b10-ac8f-332c7648fdbf · outbound

This paper cites BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:00.258268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:00.258268Z digest=sha256:5a71f898e5984d22fde4eeb0d14ca2924892007628a3d869199ab88ff735dd2f

Observation ced62b91-54c3-45b1-9e08-05e7e56ae140 · outbound

This paper cites Flamingo: a visual language model for few-shot learning,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Flamingo: a visual language model for few-shot learning,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.496661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.261705Z digest=sha256:2982fa04c56e17d1fada91a7329485a17f2ae5c913cdc5d49527b00312d0f8db

Observation f9486fcb-4b5a-4fea-b769-c90404a7b5a3 · outbound

This paper cites Detecting and Preventing Hallucinations in Large Vision Language Models.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Detecting and Preventing Hallucinations in Large Vision Language Models

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:00.264996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:00.264996Z digest=sha256:72c9ac0b277f15cbc64de5e67235787224944141ae2d5fb7fbd697d28a25f940

Observation 34c754b0-afdb-4e51-8015-78cbd4d6f1ad · outbound

This paper cites Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:00.268434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:00.268434Z digest=sha256:e3f11f892ef997bb4edeba2c72728ec82c349c96b2379d02f772a3c69dc27c2f

Observation c50934a6-876e-45e6-8396-b10a0d94a28d · outbound

This paper cites Evaluation and Enhancement of Semantic Grounding in Large Vision-Language Models.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Evaluation and Enhancement of Semantic Grounding in Large Vision-Language Models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:00.272098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:00.272098Z digest=sha256:2d768d18d7749328d4136f8a215c3a2870772cc52cee74cec436161eff23abe1

Observation bbbcd020-b0fb-4e83-bc27-793dbb5472e9 · outbound

This paper cites Woodpecker: Hallucination Correction for Multimodal Large Language Models.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Woodpecker: Hallucination Correction for Multimodal Large Language Models

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:00.276120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:00.276120Z digest=sha256:aa736369241794b902004b1f30181d6a569a04250198814a60b03339a7431b5a

Observation ea7e7403-23df-49dd-ada6-ca04bb3f694c · outbound

This paper cites Qwen-vl: A versatile vision-language model for under- standing, localization, text reading, and beyond,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Qwen-vl: A versatile vision-language model for under- standing, localization, text reading, and beyond,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.482844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T10:31:00.280139Z digest=sha256:a942171a1941f1eb6b4481c91f9a2607ed29f1968c2a1050edad03423216f9a8

Observation d2610f5d-784e-4134-b3e9-d6f8c114c06b · outbound

This paper cites Evaluating Object Hallucination in Large Vision-Language Models.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Evaluating Object Hallucination in Large Vision-Language Models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:00.283944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:00.283944Z digest=sha256:e231fb24f22d50899bd25a376eece29e5191a014ebe1051ea3fe07bcf6ed845b

Observation 41a2c272-c971-410a-a7a6-7bb40030c138 · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:00.288054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:00.288054Z digest=sha256:77226cef43530e81bc5a81402b98490a9c7571a9fdcc2eafdc741acb723ae7a2

Pith citing papers

No inbound Pith citation observations are available.