Pith. sign in

Paper Citation Record · LEDGER

Overconfidence is Key: Verbalized Uncertainty Evaluation in Large Language and Vision-Language Models

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2405.02917.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.02917 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T04:54:19.511692Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T12:59:52.886459Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 4bd77b1f-166c-46a3-afd3-8d6de6b2468e · inbound

AdaSwitch: Adaptive Switching between Small and Large Agents for Effective Cloud-Local Collaborative Learning cites this paper.

AdaSwitch: Adaptive Switching between Small and Large Agents for Effective Cloud-Local Collaborative Learning Overconfidence is Key: Verbalized Uncertainty Evaluation in Large Language and Vision-Language Models

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-23T18:48:20.407450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-23T18:46:08.566035Z digest=sha256:4cc2f6cd4e655cc8c0359ff095ae05ab82810f2b338049691063a6ec32ccf57c

Observation e0621f0b-26a2-4cb7-ade2-cb1bacbcd63f · inbound

A Survey on Uncertainty Quantification of Large Language Models: Taxonomy, Open Research Challenges, and Future Directions cites this paper.

A Survey on Uncertainty Quantification of Large Language Models: Taxonomy, Open Research Challenges, and Future Directions Overconfidence is Key: Verbalized Uncertainty Evaluation in Large Language and Vision-Language Models

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-11T20:37:54.739559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:37:54.739559Z digest=sha256:b16b5b3e2c5bc7d740887af160a68dd8d3385c8c8ad1a4e36fd0ee95d848e7ac

Observation 9dd19933-507a-4be5-8b9a-bf316b60a3d1 · inbound

Pretraining with random noise for uncertainty calibration cites this paper.

Pretraining with random noise for uncertainty calibration Overconfidence is Key: Verbalized Uncertainty Evaluation in Large Language and Vision-Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T05:31:20.741505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:31:20.741505Z digest=sha256:b2ee94ec43b92a2831470be8de106adaf20308b9e11cf18360c08413d2c55057

Observation 528858d6-e646-40af-aa2d-851d24f2d46f · inbound

Calibrating Uncertainty Quantification of Multi-Modal LLMs using Grounding cites this paper.

Calibrating Uncertainty Quantification of Multi-Modal LLMs using Grounding Overconfidence is Key: Verbalized Uncertainty Evaluation in Large Language and Vision-Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T04:54:19.511692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:54:19.511692Z digest=sha256:f2a0a4b32124a275d139de92e06a2115382cefdf84f49b696ce6bb617173cdc5

Observation ecebaa37-a3f8-4971-bca7-bd7d50311e38 · inbound

Are vision language models robust to uncertain inputs? cites this paper.

Are vision language models robust to uncertain inputs? Overconfidence is Key: Verbalized Uncertainty Evaluation in Large Language and Vision-Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T20:52:56.308626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:52:56.308626Z digest=sha256:fc4419b36504791b67a59637c259ea501ce224b0b109d45270edcc81807e2198

Observation 7a216edf-40dd-4268-8e17-ed17375f73ad · inbound

Towards Harmonized Uncertainty Estimation for Large Language Models cites this paper.

Towards Harmonized Uncertainty Estimation for Large Language Models Overconfidence is Key: Verbalized Uncertainty Evaluation in Large Language and Vision-Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:22.373109Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:25:22.373109Z digest=sha256:acf98833312836b95e3b3a6f2a4edac3f16a9bb45b188f6bec13c76b54934485

Observation 87379cc9-e676-45fa-9c91-4facdfff5089 · inbound

From Calibration to Collaboration: LLM Uncertainty Quantification Should Be More Human-Centered cites this paper.

From Calibration to Collaboration: LLM Uncertainty Quantification Should Be More Human-Centered Overconfidence is Key: Verbalized Uncertainty Evaluation in Large Language and Vision-Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T05:38:25.518866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:38:25.518866Z digest=sha256:7be3451b45e9f2f81a8670ea8db554552c69aa1766f6d143c71e457ee60c5b28

Observation d321912e-b547-4446-952b-690dedede69e · inbound

Uncertainty Quantification for Motor Imagery BCI -- Machine Learning vs. Deep Learning cites this paper.

Uncertainty Quantification for Motor Imagery BCI -- Machine Learning vs. Deep Learning Overconfidence is Key: Verbalized Uncertainty Evaluation in Large Language and Vision-Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T18:43:11.825043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:43:11.825043Z digest=sha256:7f071cfcdd9cccd8331e45790cc9dbc7905db00222d9108fc780410412c01b88

Observation 324a01e3-eebf-4f9b-9774-fdcefc434fc9 · inbound

Post-Completion Learning for Language Models cites this paper.

Post-Completion Learning for Language Models Overconfidence is Key: Verbalized Uncertainty Evaluation in Large Language and Vision-Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T17:52:15.315509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:52:15.315509Z digest=sha256:3b9920abd823edccbca1d3cb1e298e55541c1bfa94d4a863a5b6d56cfea502f0

Observation 905043ce-46ae-454c-af0d-1afa74abdabb · inbound

Optimizing Active Learning in Vision-Language Models via Parameter-Efficient Uncertainty Calibration cites this paper.

Optimizing Active Learning in Vision-Language Models via Parameter-Efficient Uncertainty Calibration Overconfidence is Key: Verbalized Uncertainty Evaluation in Large Language and Vision-Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T12:46:01.828953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:46:01.828953Z digest=sha256:913ff800b191cac86eaab0a3e2173b68f0398ebe03612038bd635dcab91035d3

Observation fc1439dd-007b-46c1-8a7b-c9f07ade9911 · inbound

Visual-Language-Guided Task Planning for Horticultural Robots cites this paper.

Visual-Language-Guided Task Planning for Horticultural Robots Overconfidence is Key: Verbalized Uncertainty Evaluation in Large Language and Vision-Language Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-03T09:58:56.323219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T09:58:56.323219Z digest=sha256:968cb22bfbbb8e87365bb0b19be5b12da25ccf9c912a1af8fa1826d665e056ab

Observation 792083d3-559c-4f71-be52-c21469040129 · inbound

SIEVES: Selective Prediction Generalizes through Visual Evidence Scoring cites this paper.

SIEVES: Selective Prediction Generalizes through Visual Evidence Scoring Overconfidence is Key: Verbalized Uncertainty Evaluation in Large Language and Vision-Language Models

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:31:16.780201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-07T16:48:20.100672Z digest=sha256:e0f03828296edd11af17bcc0c3611fd8c6c809e1776575651985933a6f0f56fe

Observation a3889f1d-bfd0-4e65-818c-9db38f5bfe2e · inbound

SIEVES: Selective Prediction Generalizes through Visual Evidence Scoring cites this paper.

SIEVES: Selective Prediction Generalizes through Visual Evidence Scoring Overconfidence is Key: Verbalized Uncertainty Evaluation in Large Language and Vision-Language Models

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-15T06:19:50.029785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-15T06:15:55.968506Z digest=sha256:d021a3c42036d91efbc93aa89186267ec45897bf6c41e773a7ea14ef34c9ef56

Observation 326e2d1d-f887-487e-b372-90648d887af3 · inbound

MCMit: Hardware-Software Co-Design for Mid-Circuit Measurement Error Mitigation cites this paper.

MCMit: Hardware-Software Co-Design for Mid-Circuit Measurement Error Mitigation Overconfidence is Key: Verbalized Uncertainty Evaluation in Large Language and Vision-Language Models

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T08:45:35.090226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-07-01T08:36:37.641628Z digest=sha256:930c48643d753cc9cda15503753dd197034a13e2428eaa118588227b5f0fa4ca

Observation f857c91f-72bc-44d1-a9df-2616adab3c9f · inbound

Just how sure are you? Improving Verbalized Uncertainty Calibration in Medical VQA cites this paper.

Just how sure are you? Improving Verbalized Uncertainty Calibration in Medical VQA Overconfidence is Key: Verbalized Uncertainty Evaluation in Large Language and Vision-Language Models

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-07-04T12:59:52.887755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-26T05:35:35.054964Z digest=sha256:e5180c48ba9a99e7861755907b4e37f189e4ff2013b886a5032842f490ce6d0c

Observation 71ba7486-6762-478f-8831-a7c35ac78915 · inbound

Small Vision-Language Models Know When They Are Wrong But Cannot Say So: A Two-Model Study of Stated versus Internal Confidence Under Realistic Image Degradation cites this paper.

Small Vision-Language Models Know When They Are Wrong But Cannot Say So: A Two-Model Study of Stated versus Internal Confidence Under Realistic Image Degradation Overconfidence is Key: Verbalized Uncertainty Evaluation in Large Language and Vision-Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T06:03:16.970761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T06:03:16.970761Z digest=sha256:87e00f94fcfb589e0cf9a6d8c4d37220500d2536350dd133b05ae0dbcca86e18

Observation 15c8e9df-6aac-4cef-b6ce-eee9c2ab9832 · inbound

Bigger or Cheaper? Scale and Quantization Effects on Uncertainty Signals in Vision-Language Models Under Image Degradation cites this paper.

Bigger or Cheaper? Scale and Quantization Effects on Uncertainty Signals in Vision-Language Models Under Image Degradation Overconfidence is Key: Verbalized Uncertainty Evaluation in Large Language and Vision-Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-31T14:56:08.549295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T14:56:08.549295Z digest=sha256:0eda3e07c6901a44d1932a68227491a6812ad099c32bb5679440eb4982e50b17