Pith. sign in

Paper Citation Record · LEDGER

BLaVe-CoT: Consistency-Aware Visual Question Answering for Blind and Low Vision Users

As of 16 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 1 inbound Pith citation observation for arXiv:2509.06010.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.06010 v1

Coverage vector

measured 21 of 21 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T04:43:20.831117Z

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T00:50:19.280108Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-12T00:50:19.564824Z

Reference resolution

21 of 21 outbound references displayed

  • verified exact1
  • verified fuzzy19
  • unresolved1
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a71d3b19-4d05-44c9-bd8a-678cf942a4cf · outbound

This paper cites Vision-language model-based polyformer for recognizing visual ques- tions with multiple answer groundings.

BLaVe-CoT: Consistency-Aware Visual Question Answering for Blind and Low Vision Users Vision-language model-based polyformer for recognizing visual ques- tions with multiple answer groundings

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T04:43:23.704356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T04:43:18.629579Z digest=sha256:a6b5a985d3abaed33c54860799941e57fe8f3ba68182561b0297b142d1e74f67

Observation 7969617d-6b83-4407-b0c4-90192c9a1b68 · outbound

This paper cites Remote assistance for blind users in daily life: A survey about be my eyes.

BLaVe-CoT: Consistency-Aware Visual Question Answering for Blind and Low Vision Users Remote assistance for blind users in daily life: A survey about be my eyes

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T04:43:23.597271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T04:43:18.722461Z digest=sha256:1f51e2c93e0c59e680469494af09e4aac8c9bbe72e2fa4c11a07fbc0569a6904

Observation 038400d4-9bfc-4000-91bf-9a075d73091f · outbound

This paper cites Vqa therapy: Exploring answer differences by visually grounding answers.

BLaVe-CoT: Consistency-Aware Visual Question Answering for Blind and Low Vision Users Vqa therapy: Exploring answer differences by visually grounding answers

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T04:43:23.481862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T04:43:18.812578Z digest=sha256:0f38a0919b66f5dd1b0a73aa9ba7691bef8dd2f155fb830b5892bd87ea2340d5

Observation 88fb4172-b046-4b0a-97de-69bb7b728e07 · outbound

This paper cites Refining pseudo labeling via multi- granularity confidence alignment for unsupervised cross domain object detection.IEEE Transactions on Image Processing, 2025.

BLaVe-CoT: Consistency-Aware Visual Question Answering for Blind and Low Vision Users Refining pseudo labeling via multi- granularity confidence alignment for unsupervised cross domain object detection.IEEE Transactions on Image Processing, 2025

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T04:43:23.318942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T04:43:18.934653Z digest=sha256:f7cd086cab12df043fdb126d8b4feae7a78efcc4310da85937b843d294993c23

Observation 17b92c3e-06ea-40ea-81ea-36dc0b1b2806 · outbound

This paper cites Vqask: a multimodal android gpt- based application to help blind users visualize pictures.

BLaVe-CoT: Consistency-Aware Visual Question Answering for Blind and Low Vision Users Vqask: a multimodal android gpt- based application to help blind users visualize pictures

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T04:43:23.155782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T04:43:19.079205Z digest=sha256:ab3258a9d746fb3536253217bed76647973bc8aa6b11a95f8b06f1ca1707b05c

Observation c35eb609-729a-4598-892a-f983135fa42f · outbound

This paper cites Towards understanding the use of mllm-enabled applications for visual interpretation by blind and low vision people.

BLaVe-CoT: Consistency-Aware Visual Question Answering for Blind and Low Vision Users Towards understanding the use of mllm-enabled applications for visual interpretation by blind and low vision people

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T04:43:23.013945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T04:43:19.198124Z digest=sha256:e1391efb77dbb51b1b6335a0d99f32f4402af191a433323fd8bd716b3307ca9a

Observation 61ae3016-6ceb-4779-84ac-d3683cd7abef · outbound

This paper cites Vizwiz grand challenge: Answering visual questions from blind people.

BLaVe-CoT: Consistency-Aware Visual Question Answering for Blind and Low Vision Users Vizwiz grand challenge: Answering visual questions from blind people

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T04:43:22.856931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T04:43:19.299838Z digest=sha256:93665636b2e024b9daf665e4fdac6d78b2f236784771de942e702f935be6121c

Observation 74eaca94-984d-4733-a645-be52a4ed3189 · outbound

This paper cites Lora: Low-rank adaptation of large language models.ICLR, 1(2):3, 2022.

BLaVe-CoT: Consistency-Aware Visual Question Answering for Blind and Low Vision Users Lora: Low-rank adaptation of large language models.ICLR, 1(2):3, 2022

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T04:43:19.427202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T04:43:19.427202Z digest=sha256:f5d42b55104853381ddd7ac4962df0f771beab1943f351b8ff959b5a0f593965

Observation 22e64f2a-a691-47bb-921b-5355012f4667 · outbound

This paper cites Consistency and uncertainty: Identifying unre- liable responses from black-box vision-language models for selective visual question answering.

BLaVe-CoT: Consistency-Aware Visual Question Answering for Blind and Low Vision Users Consistency and uncertainty: Identifying unre- liable responses from black-box vision-language models for selective visual question answering

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T04:43:22.658833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T04:43:19.511234Z digest=sha256:a7d80f3e48232f7a24d3193c107cfd29d1fe6ab05f18925f094235f84d9b30db

Observation 3e5d52a3-e44c-4d03-858f-be4c7f564a44 · outbound

This paper cites Dual-branch fusion with style modulation for cross-domain few-shot semantic segmentation.

BLaVe-CoT: Consistency-Aware Visual Question Answering for Blind and Low Vision Users Dual-branch fusion with style modulation for cross-domain few-shot semantic segmentation

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T04:43:22.540118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T04:43:19.621669Z digest=sha256:448acaa903ecbeaa0176fbe1ada967f3f20ee6a2bab72f8eeef556678b2bfa79

Observation 13112c59-87cd-4b19-bb27-628a24cc6a46 · outbound

This paper cites Natural language understanding and inference with mllm in visual question answering: A survey.ACM Computing Surveys, 57(8):1–36, 2025.

BLaVe-CoT: Consistency-Aware Visual Question Answering for Blind and Low Vision Users Natural language understanding and inference with mllm in visual question answering: A survey.ACM Computing Surveys, 57(8):1–36, 2025

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T04:43:22.407626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T04:43:19.745316Z digest=sha256:01720691ab964f90c29e78f83549f9b5265cf6c2df77f5103730b1e086d3a831

Observation 8a0b82b8-71f9-4ab7-9d16-3712a3cb19fa · outbound

This paper cites Blip-2: Boot- strapping language-image pre-training with frozen image encoders and large language models.

BLaVe-CoT: Consistency-Aware Visual Question Answering for Blind and Low Vision Users Blip-2: Boot- strapping language-image pre-training with frozen image encoders and large language models

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T04:43:22.242938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T04:43:19.846109Z digest=sha256:d4ea710367e30e0e0ce0e67af1690545897408cc727d70ae2f5928c2275d8219

Observation 59ab5756-76b5-4888-b221-fabfbf80582e · outbound

This paper cites Polyformer: Referring image segmentation as sequential polygon generation.

BLaVe-CoT: Consistency-Aware Visual Question Answering for Blind and Low Vision Users Polyformer: Referring image segmentation as sequential polygon generation

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T04:43:22.135019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T04:43:20.005306Z digest=sha256:7d7f465007ea5eef60f7a05e6e5f489bc6b98ef46f7be6ce1b44b1ee4f2bf643

Observation 4cc3c410-d558-428a-a2d3-d7f83cd29085 · outbound

This paper cites An astute assistive device for mobility and object recognition for visually impaired people.IEEE Transactions on Human-Machine Systems, 49(5):449–460, 2019.

BLaVe-CoT: Consistency-Aware Visual Question Answering for Blind and Low Vision Users An astute assistive device for mobility and object recognition for visually impaired people.IEEE Transactions on Human-Machine Systems, 49(5):449–460, 2019

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T04:43:22.046250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T04:43:20.210949Z digest=sha256:2b0f8b52638ca93ffd03c3f70d44cb2052ffc0157ae5da2b16fb44a16b117446

Observation 51efa02b-fb72-4e61-b604-ec675b7df3fa · outbound

This paper cites Dynamic conceptional con- trastive learning for generalized category discovery.

BLaVe-CoT: Consistency-Aware Visual Question Answering for Blind and Low Vision Users Dynamic conceptional con- trastive learning for generalized category discovery

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T04:43:21.954258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T04:43:20.357680Z digest=sha256:6f03409325410f6ad16540ca55f153b970305cff14e44afbc318de7e7c3e2e4f

Observation 73ebf983-0b0a-4f26-9844-b897620e2f23 · outbound

This paper cites Advances in few-shot action recognition: A comprehensive review.

BLaVe-CoT: Consistency-Aware Visual Question Answering for Blind and Low Vision Users Advances in few-shot action recognition: A comprehensive review

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T04:43:21.873693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T04:43:20.466310Z digest=sha256:7385c5ffbf26190a553d654236753fab8ec28d46c5d60b5eb439bca8a9b46571

Observation 2743f3ba-f3bf-4408-b218-19313375d57a · outbound

This paper cites DARE: Diverse Visual Question Answering with Robustness Evaluation.

BLaVe-CoT: Consistency-Aware Visual Question Answering for Blind and Low Vision Users DARE: Diverse Visual Question Answering with Robustness Evaluation

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-08-05T04:43:21.034136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T04:43:20.547086Z digest=sha256:96e9750f513fee5f38d3ae61e7411e405aef8551af7f3ec926b0917d806544d1

Observation 7cfad7f6-621b-40f6-80f5-cb3990f11da5 · outbound

This paper cites the smart vision glasses.

BLaVe-CoT: Consistency-Aware Visual Question Answering for Blind and Low Vision Users the smart vision glasses

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T04:43:21.753393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T04:43:20.632394Z digest=sha256:1a2e4cb1d7a35a1fb7f68c1c108f20bcf164ceb58bdc9d3e1ce9776306ad1785

Observation 55a58566-607d-42f3-8275-e293688f50dc · outbound

This paper cites A survey of 17 indoor travel assistance systems for blind and visually impaired people.IEEE Transactions on Human-Machine Systems, 52(1):134–148, 2021.

BLaVe-CoT: Consistency-Aware Visual Question Answering for Blind and Low Vision Users A survey of 17 indoor travel assistance systems for blind and visually impaired people.IEEE Transactions on Human-Machine Systems, 52(1):134–148, 2021

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T04:43:21.627116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T04:43:20.706463Z digest=sha256:d0ac8be51da5484a7c4d59bb2ef368e18aa95789b77c6d0e4e48ca7b0499e897

Observation 3793b5f3-9b50-4591-b94c-5f3bf8166f85 · outbound

This paper cites Minilm: Deep self-attention distillation for task-agnostic compression of pre-trained transformers.

BLaVe-CoT: Consistency-Aware Visual Question Answering for Blind and Low Vision Users Minilm: Deep self-attention distillation for task-agnostic compression of pre-trained transformers

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T04:43:21.422605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T04:43:20.778291Z digest=sha256:98c52ac86976d0f0d9cfc758b53555482ee1b6620986c7d462d1ccc002cce5b7

Observation c3c93b60-2266-480f-a13a-eabe623cbcae · outbound

This paper cites A survey on vqa: Datasets and approaches.

BLaVe-CoT: Consistency-Aware Visual Question Answering for Blind and Low Vision Users A survey on vqa: Datasets and approaches

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T04:43:21.238294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T04:43:20.831117Z digest=sha256:45107db256c02d3b3bf979f98c91680243328331bee427d1874f3ead1854f9c6

Pith citing papers

Observation ed1391d9-df16-421e-b77a-79a1b2965b96 · inbound

How Much Does It Cost to Answer My Question? Benchmarking Cloud VLM-based VQA Systems cites this paper.

How Much Does It Cost to Answer My Question? Benchmarking Cloud VLM-based VQA Systems BLaVe-CoT: Consistency-Aware Visual Question Answering for Blind and Low Vision Users

Reference 63

Resolution
verified exact
local_arxiv, observed 2026-08-12T00:50:19.569534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T00:50:19.280108Z digest=sha256:68d849c656fe7cf7270881784048a8f22bd1abfb4fdc7140d76274b251135dc2